CVE-2026-24175
Triton Inference Server vulnerability analysis and mitigation

Overview

CVE-2026-24175 is a denial-of-service vulnerability in NVIDIA Triton Inference Server caused by an uncaught exception triggered by malformed request headers. An unauthenticated remote attacker can exploit this flaw to crash the server without any user interaction or special privileges. All versions of NVIDIA Triton Inference Server prior to r26.02 are affected. The vulnerability was disclosed on April 7, 2026, with a CVSS v3.1 base score of 7.5 (High), assigned by NVIDIA Corporation (Github Advisory, NVIDIA Advisory).

Technical details

The root cause is classified as CWE-248 (Uncaught Exception), where the Triton Inference Server fails to properly handle malformed HTTP request headers, allowing an unhandled exception to propagate and crash the server process. The attack vector is network-based with low complexity — an attacker simply sends a specially crafted malformed request header to the exposed server endpoint, requiring no authentication, privileges, or user interaction. No public proof-of-concept code or detailed technical write-up has been identified at this time (Github Advisory, NVIDIA Advisory).

Impact

Successful exploitation results in a server crash, causing a complete loss of availability for the Triton Inference Server and any AI/ML inference workloads it serves. There is no impact on confidentiality or data integrity, as the vulnerability is limited to a denial-of-service condition. In production environments, this could disrupt AI inference pipelines, model serving endpoints, and dependent applications, potentially causing significant operational downtime (Github Advisory, NVIDIA Advisory).

Exploitation steps

  1. Reconnaissance: Identify internet-facing or network-accessible NVIDIA Triton Inference Server instances running versions prior to r26.02 using network scanning tools (e.g., Shodan, Censys, or nmap targeting default Triton ports such as 8000/HTTP, 8001/gRPC, 8002/metrics).
  2. Craft malformed request: Construct an HTTP request with a deliberately malformed or invalid request header (e.g., invalid header field names, oversized values, or malformed encoding) targeting the Triton server's HTTP endpoint.
  3. Send the request: Transmit the crafted request to the Triton Inference Server's HTTP API endpoint. No authentication credentials are required.
  4. Trigger crash: The server fails to catch the resulting exception from processing the malformed header, causing the Triton Inference Server process to crash and become unavailable, achieving denial of service (Github Advisory, NVIDIA Advisory).

Indicators of compromise

  • Network: Unusual or repeated HTTP requests to Triton Inference Server endpoints (default ports 8000, 8001, 8002) containing malformed or anomalous header fields; high volume of requests from a single source IP.
  • Logs: Triton server logs showing uncaught exception stack traces or abrupt process termination entries; access logs with HTTP requests containing malformed headers immediately preceding server crashes.
  • Process: Unexpected termination or restart of the tritonserver process; monitoring alerts for process crashes or container/pod restarts in Kubernetes/Docker environments hosting Triton.

Mitigation and workarounds

NVIDIA has released a patch in Triton Inference Server version r26.02; all users should upgrade to this version or later as the primary remediation (NVIDIA Advisory). As a temporary workaround prior to patching, implement network access controls (firewalls, security groups, or API gateways) to restrict access to the Triton Inference Server from untrusted or external networks. Additionally, monitor server logs for unexpected crashes or abnormal request patterns to detect potential exploitation attempts.

Additional resources


SourceThis report was generated using AI

Related Triton Inference Server vulnerabilities:

CVE ID

Severity

Score

Technologies

Component name

CISA KEV exploit

Has fix

Published date

CVE-2026-47482HIGH7.5
  • Triton Inference Server logoTriton Inference Server
  • cpe:2.3:a:nvidia:triton_inference_server
NoNoJul 14, 2026
CVE-2026-47480HIGH7.5
  • Triton Inference Server logoTriton Inference Server
  • cpe:2.3:a:nvidia:triton_inference_server
NoNoJul 14, 2026
CVE-2026-47479HIGH7.5
  • Triton Inference Server logoTriton Inference Server
  • cpe:2.3:a:nvidia:triton_inference_server
NoNoJul 14, 2026
CVE-2026-47478HIGH7.5
  • Triton Inference Server logoTriton Inference Server
  • cpe:2.3:a:nvidia:triton_inference_server
NoNoJul 14, 2026
CVE-2026-47481MEDIUM6.5
  • Triton Inference Server logoTriton Inference Server
  • cpe:2.3:a:nvidia:triton_inference_server
NoNoJul 14, 2026

Free Vulnerability Assessment

Benchmark your Cloud Security Posture

Evaluate your cloud security practices across 9 security domains to benchmark your risk level and identify gaps in your defenses.

Request assessment

Get a personalized demo

Ready to see Wiz in action?

"Best User Experience I have ever seen, provides full visibility to cloud workloads."
David EstlickCISO
"Wiz provides a single pane of glass to see what is going on in our cloud environments."
Adam FletcherChief Security Officer
"We know that if Wiz identifies something as critical, it actually is."
Greg PoniatowskiHead of Threat and Vulnerability Management