CVEFinder.io

CVE-2026-105753

đŸ”ļ medium
🔍 Scan for this CVE
Summary

vLLM is an inference and serving engine for large language models. Prior to 0.28.0, the default mirrored multimodal LRU cache can commit a media hash in the frontend sender cache during multimodal rendering and before engine admission, while the engine receiver cache never receives the payload if that request is rejected. A later request reusing the same media hash causes MultiModalProcessorSenderCache to send no payload and MultiModalReceiverCache to reach an assertion with the message "Expecte

Description

vLLM is an inference and serving engine for large language models. Prior to 0.28.0, the default mirrored multimodal LRU cache can commit a media hash in the frontend sender cache during multimodal rendering and before engine admission, while the engine receiver cache never receives the payload if that request is rejected. A later request reusing the same media hash causes MultiModalProcessorSenderCache to send no payload and MultiModalReceiverCache to reach an assertion with the message "Expected a cached item," producing a shared-service availability failure. This issue is fixed in version 0.28.0.

CVSS Score
6.5
Medium
EPSS Score
0.4
Exploit Probability
Published Date
2026-10-05
First Seen: 2026-10-08
📊 Relative Risk Intelligence

This CVE is Lower Risk - more severe than 46.5% of all 365,616 vulnerabilities in our database.

#195,441
Below average severity
Severity Percentile
đŸŽ¯ CISA SSVC Assessment Updated: Oct 6, 2026
🔍 Exploitation Status
None
No known exploits
âš™ī¸ Automatable
NO
Requires human interaction
đŸ’Ĩ Technical Impact
Partial
Limited system impact
SSVC data provided by CISA
Last Modified 2026-10-08
CVSS Vector 3.1 CVSS:3.1/AV:N/AC:L/PR:L/UI:N/S:U/C:N/I:N/A:H
CWE IDs (Weakness Types)

đŸ“Ļ Affected Products 1

🔗 References 5

🔗 Related CVEs 6

CVE ID Severity CVSS EPSS Summary Published
CVE-2026-105752 â„šī¸ low 3.1 0.2 vLLM is an inference and serving engine for large language models. Prior to 0.30.0, Harmony tool continuations submitted... 2026-10-05
CVE-2026-105754 đŸ”ļ medium 6.5 0.3 vLLM is an inference and serving engine for large language models. Prior to 0.30.0, the /inference/v1/generate endpoint ... 2026-10-05
CVE-2026-105755 đŸ”ļ medium 4.2 0.2 vLLM is an inference and serving engine for large language models. Prior to 0.30.0, flash late-interaction scoring at th... 2026-10-05
CVE-2026-105756 đŸ”ļ medium 6.5 0.3 vLLM is an inference and serving engine for large language models. Prior to 0.30.0, OpenAI-compatible request models acc... 2026-10-05
CVE-2026-105757 đŸ”ļ medium 6.5 0.3 vLLM is an inference and serving engine for large language models. Prior to 0.30.0, structured-output request failures c... 2026-10-05
CVE-2026-105758 đŸ”ļ medium 5.3 0.3 vLLM is an inference and serving engine for large language models. From 0.24.0 until 0.30.0, the Qwen2VLVideoBackend and... 2026-10-05
These CVEs affect the same products