Skip to content
vLLMGHSA-wcwg-c5fc-9vrc

vLLM is vulnerable to an Out-of-Memory (OOM) Denial of Service (DoS) attack due to unbounded frame count processing in the `VideoMediaIO.load_base64()` method

High7.5CVE-2026-5497 · Published Jun 11, 2026 · updated Aug 18, 2026

GitHub advisory

Affected versions

PackageAffectedFixed in
vllm
PyPI
>= 0.8.0, < 0.19.00.19.0
Details and references

vLLM versions 0.8.0 and later are vulnerable to an Out-of-Memory (OOM) Denial of Service (DoS) attack due to unbounded frame count processing in the `VideoMediaIO.load_base64()` method. When processing `video/jpeg` data URLs, the method splits the base64 data string on commas to extract individual JPEG frames without enforcing a frame count limit. An attacker can exploit this by crafting a single API request containing thousands of comma-separated base64-encoded JPEG frames in a data URL, causing the server to decode all frames into memory and crash due to excessive memory consumption. This vulnerability is reachable via the OpenAI-compatible chat completions API and does not require authentication.

CVSS 3.0
CVSS:3.0/AV:N/AC:L/PR:N/UI:N/S:U/C:N/I:N/A:H
Severity from
GitHub (reviewed advisory)
Weakness
CWE-400, CWE-770
Also known as
CVE-2026-5497, PYSEC-2026-2302

More vLLM advisories

All vLLM
DateAdvisory
Jun 10vLLM's Artifact Pin Decay allows pinned deployments to load unpinned code, weights, and processors
CVE-2026-47155Medium6.5fixed in 0.22.0
Jun 16vLLM: Security Check Bypass via assert Statement in Activation Function Loading Allows Arbitrary Code Execution
CVE-2026-41523High7.5fixed in 0.22.0
Jun 16vLLM: OpenAI auth bypass
CVE-2026-48746Critical9.1fixed in 0.22.0
Jun 17vLLM: temperature=NaN and temperature=Infinity bypass validation and propagate to GPU kernels
CVE-2026-54235Medium6.5fixed in 0.24.0
Jun 17vLLM: image EXIF Rotation & PNG tRNS Transparency Not Normalized, Causing Mismatch Between Model Input and Expectations
CVE-2026-12491Medium4.8fixed in 0.24.0
Jun 17vLLM: GGUF dequantize kernel int truncation exposes uninitialized GPU memory in multi-tenant serving
CVE-2026-53923Medium7.5fixed in 0.24.0

Critical advisories by email

Wednesdays: the week’s critical and high advisories in the AI and data stack, with the fixed versions. Only in weeks that have some.

Double opt-in. Unsubscribe any time.