CVE-2026-24220
MediumCVSS 6.4Exploitation Probability (EPSS)
Low risk12th percentile - higher than 12% of all known CVEs
Summary
NVIDIA TensorRT-LLM contains a vulnerability in the visual gen server that allows unsafe deserialization via unauthorized zeroMQ deserialization. Successful exploitation could lead to code execution.
Risk Assessment
An attacker can remotely execute arbitrary code on the server, potentially leading to full system compromise and data breach.
Recommendation
Immediately update NVIDIA TensorRT-LLM to the latest patched version and restrict access to the visual gen server to trusted hosts only.
Other vulnerabilities in NVIDIA TensorRT-LLM
See all- CVE-2026-47475Medium
A vulnerability in the NVIDIA TensorRT-LLM library has been found in the OpenAI-compatible inference API. An attacker can trigger a reachable assertion in the sampler thread, leading to denial of service.
- CVE-2026-47473High
NVIDIA TensorRT-LLM contains a vulnerability that could allow an attacker to trigger a write-what-where condition. Successful exploitation may lead to data tampering, denial of service, and information disclosure.
- CVE-2026-47472High
A vulnerability in NVIDIA TensorRT-LLM's inter-process communication layer allows an attacker with local same-user access to trigger deserialization. Successful exploitation could lead to code execution, information disclosure, data tampering, and denial of service.
- CVE-2026-47471High
A vulnerability in NVIDIA TensorRT-LLM during tensor deserialization allows an attacker to trigger a heap-based buffer overflow. This could lead to information disclosure, data tampering, or denial of service.
- CVE-2026-47470Medium
NVIDIA TensorRT-LLM contains a vulnerability in the gRPC server chat API endpoint, allowing a local attacker to trigger CWE-20 (improper input validation). Successful exploitation could lead to denial of service (DoS).
- CVE-2026-24271Medium
NVIDIA TensorRT-LLM contains a vulnerability in the OpenAI-compatible inference API, where an attacker could cause allocation of GPU resources without limits or throttling. A successful exploit of this vulnerability might lead to denial of service.
- CVE-2026-24259Medium
NVIDIA TensorRT-LLM for Linux contains a vulnerability that could allow an attacker to bypass authentication for a critical function. Successful exploitation may lead to code execution, data tampering, and information disclosure.
- CVE-2026-24234Medium
A vulnerability in NVIDIA TensorRT-LLM for Linux exists in the multimodal media fetching functions, allowing a network-accessible attacker to perform server-side request forgery (SSRF). Successful exploitation could lead to denial of service and information disclosure.
- CVE-2026-24233High
NVIDIA TensorRT-LLM for Linux contains a vulnerability in the restricted unpickler used for model weight deserialization, allowing a local, unauthenticated attacker to deserialize untrusted data. Successful exploitation could lead to code execution, privilege escalation, data tampering, and information disclosure.
- CVE-2026-24229High
A vulnerability in NVIDIA TensorRT-LLM for Linux within the disaggregated orchestrator component allows an attacker to read, write, or delete internal cluster state by sending requests to the FastAPI server. Successful exploitation could lead to information disclosure, data tampering, and denial of service.
Original NVD description (English source)
NVIDIA TensorRT-LLM for any platform contains a vulnerability in visual gen server, where an attacker could cause an unsafe deserialization by unauthorized zeroMQ deserialization. A successful exploit of this vulnerability might lead to code execution.

