CVE-2026-55514
MEDIUMvLLM denial of service via prompt embeds on M-RoPE models
Title source: cnaDescription
vLLM is a library for LLM inference and serving. From 0.12.0 to before 0.24.0, sending a pure prompt embeds payload in a /v1/completions request with a model using M-RoPE causes EngineCore to fail an assertion and fatally crash, shutting down the entire server application. Any remote user who is authorized to make a /v1/completions request can make such a request and induce a crash. This issue is fixed in version 0.24.0.
References (4)
Core 4
Core References
X_Refsource_Confirm x_refsource_confirm
https://github.com/vllm-project/vllm/security/advisories/GHSA-33cg-gxv8-3p8g
X_Refsource_Misc x_refsource_misc
https://github.com/vllm-project/vllm/pull/45252
X_Refsource_Misc x_refsource_misc
https://github.com/vllm-project/vllm/commit/470229c37efaf69c86e8bc97482b0b1ff7551c65
X_Refsource_Misc x_refsource_misc
https://github.com/vllm-project/vllm/releases/tag/v0.24.0
Scores
CVSS v3
6.5
EPSS
0.0037
EPSS Percentile
29.7%
Attack Vector
NETWORK
CVSS:3.1/AV:N/AC:L/PR:L/UI:N/S:U/C:N/I:N/A:H
CISA SSVC
Vulnrichment
Exploitation
none
Automatable
no
Technical Impact
partial
Details
CWE
CWE-617
Status
published
Products (3)
pypi/vllm
0.12.0 - 0.24.0PyPI
vllm/vllm
0.12.0 - 0.24.0
vllm-project/vllm
>= 0.12.0, < 0.24.0
Published
Jul 06, 2026
Tracked Since
Jul 07, 2026