vLLM是vLLM团队开源的一个适用于 LLM 的高吞吐量和内存高效推理和服务引擎。 vLLM 0.24.0之前版本存在资源管理错误漏洞,该漏洞源于structured_outputs.regex API参数将用户提供的正则表达式字符串直接传递给语法编译器后端且无编译超时,导致嵌套量词模式可通过所有检查并引发指数级状态空间扩展,允许单个包含对抗性正则表达式的请求无限挂起推理工作线程造成拒绝服务。
| Vendor | Product | Version Range | Status |
|---|---|---|---|
| vllm-project | vllm | < 0.24.0 |
affected |
Although we use advanced large model technology, its output may still contain inaccurate or outdated information.Shenlong tries to ensure data accuracy, but please verify and judge based on the actual situation.
| Vendor | Product | Affected Versions | CPE | Subscribe |
|---|---|---|---|---|
| vllm-project | vllm | < 0.24.0 | - |
|
| # | POC Description | Source Link | Shenlong Link |
|---|
No public POC found.
Login to generate AI POC| CVE-2026-54234 | 7.5 HIGH | vLLM: Remote DoS in vLLM via Invalid Recovered Token Reinjection |
| CVE-2026-55646 | 6.5 MEDIUM | vLLM speech-to-text endpoints allocate full upload before enforcing the audio file-size li |
| CVE-2026-55514 | vLLM denial of service via prompt embeds on M-RoPE models |
No comments yet