Vllm

Vllm

85 Schwachstellen gefunden.

Hinweis: Diese Liste kann unvollständig sein. Daten werden ohne Gewähr im Ursprungsformat bereitgestellt.
  • EPSS 0.21%
  • Veröffentlicht 12.09.2026 12:08:56
  • Zuletzt bearbeitet 16.09.2026 17:31:40

vLLM before 0.28.0 contains a remote code execution vulnerability in the LlavaOnevision2 processor loader that ignores the trust_remote_code parameter when loading remote processor classes. Attackers can craft a malicious model with arbitrary code in...

  • EPSS 0.53%
  • Veröffentlicht 28.08.2026 00:00:00
  • Zuletzt bearbeitet 29.09.2026 12:52:25

vLLM up to and including 0.17.0 allows remote attackers to cause a Denial of Service via memory exhaustion. The AsyncMediaIO.fetch_audio and AsyncMediaIO.fetch_image functions in multimodal/inputs.py fetch user-supplied media URLs using aiohttp and c...

  • EPSS 0.33%
  • Veröffentlicht 25.08.2026 11:33:22
  • Zuletzt bearbeitet 29.09.2026 12:52:51

vLLM before 0.27.0 fails to properly classify DeepStream as a GPU backend and omits pixel-limit enforcement in its decode path. Unauthenticated attackers can activate DeepStream at request time to initialize the process-wide GPU decode pool and submi...

  • EPSS 0.32%
  • Veröffentlicht 17.08.2026 20:17:25
  • Zuletzt bearbeitet 28.09.2026 18:41:05

vLLM is an inference and serving engine for large language models. Prior to 0.26.0, the MiMoV2OmniMultiModalProcessor in vllm/transformers_utils/processors/mimo_v2_omni.py passes attacker-controlled image and audio strings through _fetch_image, reque...

Exploit
  • EPSS 0.34%
  • Veröffentlicht 17.08.2026 20:16:45
  • Zuletzt bearbeitet 02.10.2026 19:30:34

vLLM is an inference and serving engine for large language models. Prior to 0.26.0, the /v1/completions/derender and /v1/chat/completions/derender endpoints accept caller-supplied GenerateResponse objects whose generate_responses, choices, token_ids,...

Exploit
  • EPSS 0.39%
  • Veröffentlicht 13.08.2026 15:09:02
  • Zuletzt bearbeitet 28.09.2026 18:41:23

vLLM is an inference and serving engine for large language models. From 0.19.0 until 0.26.0, the /v1/completions CompletionRequest.prompt field in vllm/entrypoints/openai/completion/protocol.py accepts an unbounded list[str] or list[list[int]], promp...

Exploit
  • EPSS 0.26%
  • Veröffentlicht 13.08.2026 15:06:06
  • Zuletzt bearbeitet 28.09.2026 18:41:49

vLLM is an inference and serving engine for large language models. Prior to 0.27.0, an integer overflow in blockIdx.x * 2 * d in activation_kernels.cu can cause act_and_mul_kernel to consume another batched user's input, allowing a request processed ...

  • EPSS 0.32%
  • Veröffentlicht 13.08.2026 14:56:52
  • Zuletzt bearbeitet 28.09.2026 18:42:04

vLLM is an inference and serving engine for large language models. Prior to 0.26.0, the structured_outputs.regex parameter in vllm/v1/structured_output/backend_lm_format_enforcer.py is passed to lmformatenforcer.RegexParser without compile_regex_with...

  • EPSS 0.26%
  • Veröffentlicht 13.08.2026 14:50:03
  • Zuletzt bearbeitet 02.10.2026 19:28:34

vLLM is an inference and serving engine for large language models. Prior to 0.26.0, the validation_exception_handler in vllm/entrypoints/openai/server_utils.py converts FastAPI RequestValidationError objects with str(exc), and sanitize_message in vll...

  • EPSS 0.37%
  • Veröffentlicht 06.07.2026 20:07:40
  • Zuletzt bearbeitet 07.07.2026 19:02:37

vLLM is a library for LLM inference and serving. From 0.12.0 to before 0.24.0, sending a pure prompt embeds payload in a /v1/completions request with a model using M-RoPE causes EngineCore to fail an assertion and fatally crash, shutting down the ent...