{"id":"PYSEC-2026-2303","details":"vLLM is a library for LLM inference and serving. From 0.12.0 to before 0.24.0, sending a pure prompt embeds payload in a /v1/completions request with a model using M-RoPE causes EngineCore to fail an assertion and fatally crash, shutting down the entire server application. Any remote user who is authorized to make a /v1/completions request can make such a request and induce a crash. This issue is fixed in version 0.24.0.","aliases":["CVE-2026-55514","GHSA-33cg-gxv8-3p8g"],"modified":"2026-07-13T07:15:46.624479183Z","published":"2026-07-06T21:16:57.207Z","references":[{"type":"ADVISORY","url":"https://github.com/vllm-project/vllm/releases/tag/v0.24.0"},{"type":"ADVISORY","url":"https://github.com/vllm-project/vllm/security/advisories/GHSA-33cg-gxv8-3p8g"},{"type":"FIX","url":"https://github.com/vllm-project/vllm/commit/470229c37efaf69c86e8bc97482b0b1ff7551c65"},{"type":"FIX","url":"https://github.com/vllm-project/vllm/pull/45252"}],"affected":[{"package":{"name":"vllm","ecosystem":"PyPI","purl":"pkg:pypi/vllm"},"ranges":[{"type":"ECOSYSTEM","events":[{"introduced":"0.12.0"},{"fixed":"0.24.0"}]}],"versions":["0.12.0","0.13.0","0.14.0","0.14.1","0.15.0","0.15.1","0.16.0","0.17.0","0.17.1","0.18.0","0.18.1","0.19.0","0.19.1","0.20.0","0.20.1","0.20.2","0.21.0","0.22.0","0.22.1","0.23.0"],"ecosystem_specific":{},"database_specific":{"source":"https://github.com/pypa/advisory-database/blob/main/vulns/vllm/PYSEC-2026-2303.yaml"}}],"schema_version":"1.7.5","severity":[{"type":"CVSS_V3","score":"CVSS:3.1/AV:N/AC:L/PR:L/UI:N/S:U/C:N/I:N/A:H"}]}