CVE-2026-100651
Last modified
CVE-2026-100651 is a medium-severity vulnerability rated 6.5/10 on the CVSS scale. vLLM before 0.29.0 fails to enforce decoder prompt-length validation on the disaggregated serving endpoint /inference/v1/generate. When the request contains a 'features' (multimodal) payload, vllm/entrypoints/serve/disagg/serving.py builds a multimodal EngineInput directly from the caller-supplied token_ids, and GenerateRequest.token_ids (vllm/entrypoints/serve/disagg/protocol.py) is not checked against model_config.max_model_len.
Description
vLLM before 0.29.0 fails to enforce decoder prompt-length validation on the disaggregated serving endpoint /inference/v1/generate. When the request contains a 'features' (multimodal) payload, vllm/entrypoints/serve/disagg/serving.py builds a multimodal EngineInput directly from the caller-supplied token_ids, and GenerateRequest.token_ids (vllm/entrypoints/serve/disagg/protocol.py) is not checked against model_config.max_model_len. For multimodal processors that report skip_prompt_length_check=True (for example Nemotron Parse, Whisper, and FireRedLID), InputProcessor._validate_prompt_len() returns immediately for both encoder and decoder prompts, so an overlong prompt becomes an EngineCoreRequest and reaches the worker input-batch copy into a fixed max_model_len-wide NumPy row. A client able to reach the endpoint on an affected model configuration can therefore submit an overlong token_ids list to trigger a worker failure and denial of service. Fixed in 0.29.0.
Metrics
Weakness Enumeration
Affected Software
Source: CNA advisory (CVE.org). NVD analysis pending.
| Vendor | Product | Versions |
|---|---|---|
| vllm-project | vllm | < 0.29.0 |
References
Timeline
- Published
- Last Modified
- Status
- Received
Frequently Asked Questions
What is CVE-2026-100651?
How severe is CVE-2026-100651?
How do I fix CVE-2026-100651?
How Strix Helps
- How Strix found a critical auth bypass in etcdStrix autonomously discovered a critical authentication bypass in etcd, later designated CVE-2026-33413.
- Autonomous PentestingAI agents that find and validate exploitable vulnerabilities like this one across your applications.
- PR ReviewsPentest every pull request so vulnerable code is caught before it ships to production.
- AI Penetration TestingHow AI-driven penetration testing continuously covers your attack surface.
Related CVEs from 2026
- CVE-2026-100646SiYuan is a self-hosted personal knowledge management system…8.1
- CVE-2026-100647vLLM versions before 0.29.0 contain a denial-of-service vuln…5.3
- CVE-2026-100648vllm before 0.29.0 fails to enforce VLLM_MAX_AUDIO_CLIP_FILE…5.3
- CVE-2026-100649vLLM before 0.29.0 contains a resource-limit bypass vulnerab…3.7
- CVE-2026-10065A weakness has been identified in Shibby Tomato 1.28. This v…8.8
- CVE-2026-100650vLLM through 0.29.0 fetches and fully materializes remote or…6.5
- CVE-2026-100652vLLM versions 0.22.0 through 0.23.0 fail to validate stop_to…5.9
- CVE-2026-100653vLLM is an inference and serving engine for large language m…6.5
- CVE-2026-100654vLLM before 0.29.0 accepts user-controlled stop_token_ids on…6.5
- CVE-2026-100655Netty (io.netty:netty-codec-http) versions up to and includi…7.5
- CVE-2026-100656Netty (io.netty:netty-codec-http) contains an unbounded per-…7.5
- CVE-2026-100657Netty's STOMP codec (io.netty:netty-codec-stomp) contains a …7.5
Are you affected by CVE-2026-100651?
Run a free Strix scan to check your systems for this vulnerability.
Scan your code nowSource: NVD / NIST
