{"id":"CVE-2026-100652","published":"2026-09-26T14:16:47.810","lastModified":"2026-09-29T21:35:07.960","description":"vLLM versions 0.22.0 through 0.23.0 fail to validate stop_token_ids against vocabulary bounds in Rust HTTP and gRPC frontends, allowing out-of-vocabulary token IDs to reach MinTokensLogitsProcessor. Attackers can submit requests with min_tokens greater than zero and out-of-vocabulary stop_token_ids to trigger CUDA tensor indexing failures that leave EngineCore in a fatal state requiring service restart.","cvssScore":5.9,"cvssSeverity":"MEDIUM","cvssVector":"CVSS:3.1/AV:N/AC:H/PR:N/UI:N/S:U/C:N/I:N/A:H","cwes":["CWE-20"],"vendors":[],"products":[],"references":[{"url":"https://github.com/vllm-project/vllm/security/advisories/GHSA-qff2-492f-9fm4","tags":[]},{"url":"https://www.vulncheck.com/advisories/vllm-0.22.0-through-0.23.0-denial-of-service-via-stop-token-ids","tags":[]}],"exploitRefs":[{"url":"https://github.com/vllm-project/vllm/security/advisories/GHSA-qff2-492f-9fm4","tags":[]}],"hasPoc":true,"ai":null}