CVE-2026-90553
- EPSS 0.21%
- Veröffentlicht 12.09.2026 12:08:56
- Zuletzt bearbeitet 16.09.2026 17:31:40
vLLM before 0.28.0 contains a remote code execution vulnerability in the LlavaOnevision2 processor loader that ignores the trust_remote_code parameter when loading remote processor classes. Attackers can craft a malicious model with arbitrary code in...
CVE-2026-37237
- EPSS 0.53%
- Veröffentlicht 28.08.2026 00:00:00
- Zuletzt bearbeitet 29.09.2026 12:52:25
vLLM up to and including 0.17.0 allows remote attackers to cause a Denial of Service via memory exhaustion. The AsyncMediaIO.fetch_audio and AsyncMediaIO.fetch_image functions in multimodal/inputs.py fetch user-supplied media URLs using aiohttp and c...
CVE-2026-78684
- EPSS 0.33%
- Veröffentlicht 25.08.2026 11:33:22
- Zuletzt bearbeitet 29.09.2026 12:52:51
vLLM before 0.27.0 fails to properly classify DeepStream as a GPU backend and omits pixel-limit enforcement in its decode path. Unauthenticated attackers can activate DeepStream at request time to initialize the process-wide GPU decode pool and submi...
CVE-2026-73560
- EPSS 0.32%
- Veröffentlicht 17.08.2026 20:17:25
- Zuletzt bearbeitet 28.09.2026 18:41:05
vLLM is an inference and serving engine for large language models. Prior to 0.26.0, the MiMoV2OmniMultiModalProcessor in vllm/transformers_utils/processors/mimo_v2_omni.py passes attacker-controlled image and audio strings through _fetch_image, reque...
CVE-2026-71486
- EPSS 0.34%
- Veröffentlicht 17.08.2026 20:16:45
- Zuletzt bearbeitet 02.10.2026 19:30:34
vLLM is an inference and serving engine for large language models. Prior to 0.26.0, the /v1/completions/derender and /v1/chat/completions/derender endpoints accept caller-supplied GenerateResponse objects whose generate_responses, choices, token_ids,...
CVE-2026-73559
- EPSS 0.39%
- Veröffentlicht 13.08.2026 15:09:02
- Zuletzt bearbeitet 28.09.2026 18:41:23
vLLM is an inference and serving engine for large language models. From 0.19.0 until 0.26.0, the /v1/completions CompletionRequest.prompt field in vllm/entrypoints/openai/completion/protocol.py accepts an unbounded list[str] or list[list[int]], promp...
CVE-2026-73558
- EPSS 0.26%
- Veröffentlicht 13.08.2026 15:06:06
- Zuletzt bearbeitet 28.09.2026 18:41:49
vLLM is an inference and serving engine for large language models. Prior to 0.27.0, an integer overflow in blockIdx.x * 2 * d in activation_kernels.cu can cause act_and_mul_kernel to consume another batched user's input, allowing a request processed ...
CVE-2026-73556
- EPSS 0.32%
- Veröffentlicht 13.08.2026 14:56:52
- Zuletzt bearbeitet 28.09.2026 18:42:04
vLLM is an inference and serving engine for large language models. Prior to 0.26.0, the structured_outputs.regex parameter in vllm/v1/structured_output/backend_lm_format_enforcer.py is passed to lmformatenforcer.RegexParser without compile_regex_with...
CVE-2026-73555
- EPSS 0.26%
- Veröffentlicht 13.08.2026 14:50:03
- Zuletzt bearbeitet 02.10.2026 19:28:34
vLLM is an inference and serving engine for large language models. Prior to 0.26.0, the validation_exception_handler in vllm/entrypoints/openai/server_utils.py converts FastAPI RequestValidationError objects with str(exc), and sanitize_message in vll...
CVE-2026-55514
- EPSS 0.37%
- Veröffentlicht 06.07.2026 20:07:40
- Zuletzt bearbeitet 07.07.2026 19:02:37
vLLM is a library for LLM inference and serving. From 0.12.0 to before 0.24.0, sending a pure prompt embeds payload in a /v1/completions request with a model using M-RoPE causes EngineCore to fail an assertion and fatally crash, shutting down the ent...