CVE-2026-105753
vLLM: Mirrored multimodal IPC caches desync after a rejected request โ a later request reusing the same media hash trips a receiver assertion in the engine core
CVSS Score
6.5
EPSS Score
0.0%
EPSS Percentile
0th
vLLM is an inference and serving engine for large language models. Prior to 0.28.0, the default mirrored multimodal LRU cache can commit a media hash in the frontend sender cache during multimodal rendering and before engine admission, while the engine receiver cache never receives the payload if that request is rejected. A later request reusing the same media hash causes MultiModalProcessorSenderCache to send no payload and MultiModalReceiverCache to reach an assertion with the message "Expected a cached item," producing a shared-service availability failure. This issue is fixed in version 0.28.0.
| CWE | CWE-617 |
| Vendor | vllm-project |
| Product | vllm |
| Published | Oct 5, 2026 |
Stay Ahead of the Next One
Get instant alerts for vllm-project vllm
Be the first to know when new medium vulnerabilities affecting vllm-project vllm are published โ delivered to Slack, Telegram or Discord.
Get Free Alerts โ
Free ยท No credit card ยท 60 sec setup
CVSS v3 Breakdown
CVSS:3.1/AV:N/AC:L/PR:L/UI:N/S:U/C:N/I:N/A:H Attack Vector
Network
Attack Complexity
Low
Privileges Required
Low
User Interaction
None
Scope
Unchanged
Confidentiality
None
Integrity
None
Availability
High
Affected Versions
vllm-project / vllm
< 0.28.0
References
github.com: https://github.com/vllm-project/vllm/security/advisories/GHSA-ph3r-5jfg-f84f github.com: https://github.com/vllm-project/vllm/pull/46747 github.com: https://github.com/vllm-project/vllm/pull/51897 github.com: https://github.com/vllm-project/vllm/commit/396204230423b7cc6798300926b8fa30190d26a9 github.com: https://github.com/vllm-project/vllm/releases/tag/v0.28.0