Record summary

CVE-2026-73559 has a selected CVSS score of 6.5 (medium).

Description

vLLM is an inference and serving engine for large language models. From 0.19.0 until 0.26.0, the /v1/completions CompletionRequest.prompt field in vllm/entrypoints/openai/completion/protocol.py accepts an unbounded list[str] or list[list[int]], prompt_to_seq() in vllm/renderers/inputs/preprocess.py and OnlineRenderer.preprocess_completion() in vllm/renderers/online_renderer.py expand every element, and vllm/entrypoints/openai/completion/serving.py creates one engine generator and response slot per prompt, allowing an authenticated API client to exhaust CPU, memory, async scheduling capacity, engine request slots, and response buffering with one request. This issue is fixed in version 0.26.0.

Description source: CVE List

Affected products and versions

2
ProductSourceVersion rangeStatus
CVE List>= 0.19.0, < 0.26.0affected
GitHub Advisory0.19.0 to < 0.26.0 · Fixed in 0.26.0affected

References

6