vLLM 在 0.29.0 版本之前,在解耦式服务(disaggregated serving)端点 /inference/v1/generate 上未能强制执行解码器提示词长度验证。当请求中包含“features”(多模态)负载时,vllm/entrypoints/serve/disagg/serving.py 会直接根据调用方提供的 token_ids 构建多模态 EngineInput,且未对 GenerateRequest.token_ids(定义于 vllm/entrypoints/serve/disag
Although we use advanced large model technology, its output may still contain inaccurate or outdated information.Shenlong tries to ensure data accuracy, but please verify and judge based on the actual situation.
| Vendor | Product | Affected Versions | CPE | Subscribe |
|---|---|---|---|---|
| vllm-project | vllm | 0 ~ 0.29.0 | - |
|
| # | POC Description | Source Link | Shenlong Link |
|---|
No public POC found.
Login to generate AI POC| CVE-2026-100654 | 6.5 MEDIUM | vLLM before 0.29.0 Denial of Service via out-of-range stop_token_ids |
| CVE-2026-100650 | 6.5 MEDIUM | vLLM before 0.29.0 Resource Exhaustion via Unbounded Media Materialization |
| CVE-2026-100653 | 6.5 MEDIUM | vLLM 0.22.1 before 0.28.0 Incomplete Artifact Pin Propagation |
| CVE-2026-100652 | 5.9 MEDIUM | vLLM 0.22.0 through 0.23.0 Denial of Service via stop_token_ids |
| CVE-2026-100647 | 5.3 MEDIUM | vLLM before 0.29.0 CPU Exhaustion via unbounded cache_salt |
| CVE-2026-100648 | 5.3 MEDIUM | vllm before 0.29.0 Uncontrolled Resource Consumption via Audio Decoding |
| CVE-2026-100649 | 3.7 LOW | vLLM before 0.29.0 Resource Limit Bypass via Sampler Subclass |
No comments yet