GPU VulnDB

Database/AI/ML frameworks & serving

vLLM (prefix cache): Prefix-cache timing side channel leaks other tenants' prompts

CVE-2025-46570AI/ML frameworks & servingcurated

Impact

Prefix-cache timing side channel leaks other tenants' prompts

Who can reach it

Co-tenant issuing timed prompts against a shared serving instance

What to do

No clean fix while prefix caching is shared. Do not share a vLLM instance across tenants — the cache is a cross-tenant channel

References

This entry is curated: imported from vendor advisories with machine assistance, not yet individually verified. Confirm against your vendor's advisory before acting, and report anything wrong.