Severity by source
AV:N/AC:L/PR:N/UI:N/S:U/C:N/I:N/A:H
Network-accessible inference endpoints require no authentication; impact is strictly availability loss with no confidentiality or integrity compromise possible from a DoS.
Primary rating from Vendor (nvidia).
CVSS VectorVendor: nvidia
Lifecycle Timeline
2DescriptionCVE.org
NVIDIA Triton Inference Server for Linux contains a vulnerability where an attacker could cause improper input validation. A successful exploit might lead to denial of service.
Articles & Coverage 1
AnalysisAI
Denial of service in NVIDIA Triton Inference Server for Linux exposes AI inference infrastructure to remote disruption through improper input validation. Unauthenticated network attackers can crash or render unresponsive the inference server by submitting malformed inputs, with CVSS 7.5 (AV:N/AC:L/PR:N/UI:N) confirming no authentication or user interaction is required. No public exploit code has been identified at time of analysis, and CISA KEV listing is absent, but the low attack complexity and zero-authentication requirement make this a meaningful availability risk for any Triton deployment reachable from untrusted networks.
Technical ContextAI
NVIDIA Triton Inference Server is an open-source model serving platform designed to deploy trained AI models at scale across CPU and GPU infrastructure on Linux. It exposes HTTP and gRPC endpoints for inference requests, making it a network-accessible service. The root cause is CWE-20 (Improper Input Validation): the server fails to adequately sanitize or bound-check incoming request data before processing, allowing a specially crafted payload to trigger abnormal server behavior. The CPE string cpe:2.3:a:nvidia:triton_inference_server:*:*:*:*:*:*:*:* with a wildcard version field indicates all currently tracked versions of the server on Linux are potentially within scope, though no upper or lower version bound is confirmed in the available data.
RemediationAI
Consult the NVIDIA Product Security advisory at https://github.com/NVIDIA/product-security/tree/main/2026/5865 for the authoritative patch guidance; no exact fixed version number was present in the provided intelligence data, so a specific upgrade target cannot be confirmed independently. Until a patch is applied, compensating controls should include restricting network access to Triton's HTTP (default port 8000) and gRPC (default port 8001) endpoints to trusted internal IP ranges using firewall rules or network policies, with the trade-off that this may block legitimate inference clients. Placing a reverse proxy or API gateway with input size limits and request rate limiting in front of Triton can reduce the attack surface without disabling the service, though it does not eliminate the underlying validation flaw. Monitoring inference endpoint error rates and process restart frequency can provide early warning of exploitation attempts.
More in Triton Inference Server
View allAuthentication bypass in NVIDIA Triton Inference Server allows unauthenticated remote attackers to reach protected funct
Path traversal in NVIDIA Triton Inference Server for Linux (CWE-22) lets remote attackers manipulate file paths to reach
Authentication bypass in NVIDIA Triton Inference Server allows remote unauthenticated attackers to circumvent access con
Integer overflow in the DALI backend of NVIDIA Triton Inference Server allows authenticated remote attackers to trigger
Out-of-bounds read in the DALI backend of NVIDIA Triton Inference Server allows authenticated remote attackers to trigge
NVIDIA Triton Inference Server for Linux contains a vulnerability where a user can set the logging location to an arbitr
NVIDIA Triton Inference Server for Linux contains a vulnerability in the tracing API, where a user can corrupt system fi
NVIDIA Triton Inference Server for Linux contains a vulnerability in shared memory APIs, where a user can cause an impro
Denial of service in NVIDIA Triton Inference Server for Linux allows remote unauthenticated attackers to exhaust server
Denial of service in NVIDIA Triton Inference Server for Linux (versions through 26.03) allows a remote attacker to crash
Unbounded resource allocation in NVIDIA Triton Inference Server for Linux allows remote, unauthenticated attackers to ex
Denial of service in NVIDIA Triton Inference Server for Linux allows remote unauthenticated attackers to trigger uncontr
Same weakness CWE-20 – Improper Input Validation
View allSame technique Denial Of Service
View allShare
External POC / Exploit Code
Leaving vuln.today
EUVD-2026-61110
GHSA-4hj7-x739-r9gh