NVIDIA TensorRT-LLM vulnerabilities
12 known CVE vulnerabilities in NVIDIA TensorRT-LLM, translated and rated.
- CVE-2026-47475Medium
A vulnerability in the NVIDIA TensorRT-LLM library has been found in the OpenAI-compatible inference API. An attacker can trigger a reachable assertion in the sampler thread, leading to denial of service.
- CVE-2026-47473High
NVIDIA TensorRT-LLM contains a vulnerability that could allow an attacker to trigger a write-what-where condition. Successful exploitation may lead to data tampering, denial of service, and information disclosure.
- CVE-2026-47472High
A vulnerability in NVIDIA TensorRT-LLM's inter-process communication layer allows an attacker with local same-user access to trigger deserialization. Successful exploitation could lead to code execution, information disclosure, data tampering, and denial of service.
- CVE-2026-47471High
A vulnerability in NVIDIA TensorRT-LLM during tensor deserialization allows an attacker to trigger a heap-based buffer overflow. This could lead to information disclosure, data tampering, or denial of service.
- CVE-2026-47470Medium
NVIDIA TensorRT-LLM contains a vulnerability in the gRPC server chat API endpoint, allowing a local attacker to trigger CWE-20 (improper input validation). Successful exploitation could lead to denial of service (DoS).
- CVE-2026-24271Medium
NVIDIA TensorRT-LLM contains a vulnerability in the OpenAI-compatible inference API, where an attacker could cause allocation of GPU resources without limits or throttling. A successful exploit of this vulnerability might lead to denial of service.
- CVE-2026-24259Medium
NVIDIA TensorRT-LLM for Linux contains a vulnerability that could allow an attacker to bypass authentication for a critical function. Successful exploitation may lead to code execution, data tampering, and information disclosure.
- CVE-2026-24234Medium
A vulnerability in NVIDIA TensorRT-LLM for Linux exists in the multimodal media fetching functions, allowing a network-accessible attacker to perform server-side request forgery (SSRF). Successful exploitation could lead to denial of service and information disclosure.
- CVE-2026-24233High
NVIDIA TensorRT-LLM for Linux contains a vulnerability in the restricted unpickler used for model weight deserialization, allowing a local, unauthenticated attacker to deserialize untrusted data. Successful exploitation could lead to code execution, privilege escalation, data tampering, and information disclosure.
- CVE-2026-24229High
A vulnerability in NVIDIA TensorRT-LLM for Linux within the disaggregated orchestrator component allows an attacker to read, write, or delete internal cluster state by sending requests to the FastAPI server. Successful exploitation could lead to information disclosure, data tampering, and denial of service.
- CVE-2026-24226Medium
A vulnerability in NVIDIA TensorRT-LLM for Linux allows an attacker to cause improper control of code generation. Successful exploitation could lead to code execution, data tampering, and information disclosure.
- CVE-2026-24220Medium
NVIDIA TensorRT-LLM contains a vulnerability in the visual gen server that allows unsafe deserialization via unauthorized zeroMQ deserialization. Successful exploitation could lead to code execution.

