vLLM 是一个用于大语言模型的推理和服务引擎。在 0.24.0 版本之前,/v1/chat/completions 接口在处理 input_audio 参数时,调用 AudioMediaIO.load_bytes 或 AudioMediaIO.load_file 方法时,未将 VLLM_MAX_AUDIO_DECODE_DURATION_S 参数传递给共享的音频解码器。因此,未认证的攻击者可以提交一个体积较小的压缩音频输入,使其在解码后展开为极其庞大的 float32 格式 PCM 内存分配,从而绕过 /v1/a
Although we use advanced large model technology, its output may still contain inaccurate or outdated information.Shenlong tries to ensure data accuracy, but please verify and judge based on the actual situation.
| Vendor | Product | Affected Versions | CPE | Subscribe |
|---|---|---|---|---|
| vllm-project | vllm | < 0.24.0 | - |
|
| # | POC Description | Source Link | Shenlong Link |
|---|
No public POC found.
Login to generate AI POC| CVE-2026-69147 | 6.5 MEDIUM | vLLM: Request-selected PyNvVideoCodec GPU decode bypasses static VRAM reservation |
| CVE-2026-92220 | 5.3 MEDIUM | vllm-project vLLM MoRIIO Acknowledgement moriio_connector.py MoRIIOWrapper._handle_release |
| CVE-2026-92365 | 4.3 MEDIUM | vllm-project vllm thinking_budget_state.py algorithmic complexity |
No comments yet