Skip to content
vLLMGHSA-ggpf-24jw-3fcw

CVE-2025-24357 Malicious model remote code execution fix bypass with PyTorch < 2.6.0

Critical9.8Published Apr 23, 2025 · updated Aug 7, 2026

GitHub advisory

Affected versions

PackageAffectedFixed in
vllm
PyPI
< 0.8.00.8.0
Details and references

## Description https://github.com/vllm-project/vllm/security/advisories/GHSA-rh4j-5rhw-hr54 reported a vulnerability where loading a malicious model could result in code execution on the vllm host. The fix applied to specify `weights_only=True` to calls to `torch.load()` did not solve the problem prior to PyTorch 2.6.0. PyTorch has issued a new CVE about this problem: https://github.com/advisories/GHSA-53q9-r3pm-6pq6 This means that versions of vLLM using PyTorch before 2.6.0 are vulnerable to this problem. ## Background Knowledge When users install VLLM according to the official manual ![image](https://github.com/user-attachments/assets/d17e0bdb-26f2-46d6-adf6-0b17e5ddf5c7) But the version of PyTorch is specified in the requirements. txt file ![image](https://github.com/user-attachments/assets/94aad622-ad6d-4741-b772-c342727c58c7) So by default when the user install VLLM, it will install the PyTorch with version 2.5.1 ![image](https://github.com/user-attachments/assets/04ff31b0-aad1-490a-963d-00fda91da47b) In CVE-2025-24357, weights_only=True was used for patching, but we know this is not secure. Because we found that using Weights_only=True in pyTorch before 2.5.1 was unsafe Here, we use this interface to prove that it is not safe. ![image](https://github.com/user-attachments/assets/0d86efcd-2aad-42a2-8ac6-cc96b054c925) ## Fix update PyTorch version to 2.6.0 ## Credit This vulnerability was found By Ji'an Zhou and Li'shuo Song

CVSS 3.1
CVSS:3.1/AV:N/AC:L/PR:N/UI:N/S:U/C:H/I:H/A:H
Severity from
GitHub (reviewed advisory)
Weakness
CWE-1395

More vLLM advisories

All vLLM
DateAdvisory
Apr 292025Data exposure via ZeroMQ on multi-node vLLM deployment
CVE-2025-30202High7.5fixed in 0.8.5
Apr 292025vLLM Vulnerable to Remote Code Execution via Mooncake Integration
CVE-2025-32444Critical10.0fixed in 0.8.5
Apr 292025vLLM: Quadratic Time Complexity in Input Token Processing​ leads to denial of service
CVE-2025-46560Medium6.5fixed in 0.8.5
Apr 152025vLLM vulnerable to Denial of Service by abusing xgrammar cache
GHSA-hf3c-wxg2-49q9Medium6.5fixed in 0.8.4
May 62025Remote Code Execution Vulnerability in vLLM Multi-Node Cluster Configuration
CVE-2025-30165High8.0fixed in 0.10.0
May 202025vLLM Allows Remote Code Execution via PyNcclPipe Communication Service
CVE-2025-47277Critical9.8fixed in 0.8.5

Critical advisories by email

Wednesdays: the week’s critical and high advisories in the AI and data stack, with the fixed versions. Only in weeks that have some.

Double opt-in. Unsubscribe any time.