Skip to main content

Triton Inference Server CVE-2026-47628

| EUVDEUVD-2026-61109 HIGH
Allocation of Resources Without Limits or Throttling (CWE-770)
2026-08-18 nvidia GHSA-j87r-h67h-c626
7.5
CVSS 3.1 · Vendor: nvidia
Share

Severity by source

Vendor (nvidia) PRIMARY
7.5 HIGH
AV:N/AC:L/PR:N/UI:N/S:U/C:N/I:N/A:H
vuln.today AI
7.5 HIGH

Network-reachable inference API, no authentication required, availability-only impact; metrics align with provided vector and CWE-770 characterization.

3.1 AV:N/AC:L/PR:N/UI:N/S:U/C:N/I:N/A:H
4.0 AV:N/AC:L/AT:N/PR:N/UI:N/VC:N/VI:N/VA:H/SC:N/SI:N/SA:N

Primary rating from Vendor (nvidia).

CVSS VectorVendor: nvidia

Attack Vector
Network
Attack Complexity
Low
Privileges Required
None
User Interaction
None
Scope
Unchanged
Confidentiality
None
Integrity
None
Availability
High

Lifecycle Timeline

2
Analysis Generated
Aug 18, 2026 - 18:52 vuln.today
CVE Published
Aug 18, 2026 - 18:23 cve.org
HIGH 7.5

DescriptionCVE.org

NVIDIA Triton Inference Server for Linux contains a vulnerability where an attacker could cause an allocation of resources without limits. A successful exploit might lead to denial of service.

AnalysisAI

Unbounded resource allocation in NVIDIA Triton Inference Server for Linux allows remote, unauthenticated attackers to exhaust server-side resources and cause denial of service against all tracked versions (CPE wildcard). The CVSS 7.5 High rating reflects a fully network-accessible attack path with no privileges or user interaction required, making the exposure surface broad for any Triton deployment reachable from untrusted networks. No public exploit code has been identified at time of analysis, and exploitation has not been confirmed by CISA KEV.

Technical ContextAI

NVIDIA Triton Inference Server is an open-source, production-grade ML inference serving platform for Linux, designed to host and serve AI models at scale via HTTP, gRPC, and metrics APIs. The vulnerability is rooted in CWE-770 (Allocation of Resources Without Limits or Throttling), meaning a code path in Triton's request handling, model-loading, or session management logic fails to enforce upper bounds on a finite resource - such as memory, file descriptors, thread pool slots, or request queue depth. An attacker who sends crafted inputs along this path can drive the server's resource consumption to exhaustion, starving legitimate inference clients. The CPE string cpe:2.3:a:nvidia:triton_inference_server:*:*:*:*:*:*:*:* applies a full version wildcard, indicating the flaw is not isolated to a specific release branch and potentially spans the entire product lineage on Linux.

RemediationAI

Consult the NVIDIA product security advisory at https://github.com/NVIDIA/product-security/tree/main/2026/5865 for authoritative patch guidance and fixed version information; no exact patched release version was confirmed in the available intelligence, so the advisory is the definitive source. As an immediate compensating control, restrict network access to Triton inference endpoints - default HTTP port 8000, gRPC port 8001, and metrics port 8002 - to trusted clients only using host-based firewall rules (iptables/nftables) or network segmentation, reducing the pool of potential attackers. Deploying a rate-limiting reverse proxy or API gateway (such as nginx with limit_req or Envoy with global rate limiting) in front of Triton can throttle the throughput of resource-exhausting requests; note this reduces exploitation opportunity but does not eliminate the underlying vulnerability. Monitor Triton process memory and file descriptor consumption via the metrics endpoint or system-level tooling and set alerting thresholds to detect anomalous consumption before full exhaustion occurs.

CVE-2026-24207 CRITICAL POC
9.8 May 20

Authentication bypass in NVIDIA Triton Inference Server allows unauthenticated remote attackers to reach protected funct

CVE-2026-47627 CRITICAL
9.8 Aug 18

Path traversal in NVIDIA Triton Inference Server for Linux (CWE-22) lets remote attackers manipulate file paths to reach

CVE-2026-24206 CRITICAL
9.8 May 20

Authentication bypass in NVIDIA Triton Inference Server allows remote unauthenticated attackers to circumvent access con

CVE-2026-24214 CRITICAL
9.8 May 20

Integer overflow in the DALI backend of NVIDIA Triton Inference Server allows authenticated remote attackers to trigger

CVE-2026-24213 CRITICAL
9.8 May 20

Out-of-bounds read in the DALI backend of NVIDIA Triton Inference Server allows authenticated remote attackers to trigge

CVE-2024-0087 HIGH
8.8 May 14

NVIDIA Triton Inference Server for Linux contains a vulnerability where a user can set the logging location to an arbitr

CVE-2024-0100 HIGH
8.1 May 14

NVIDIA Triton Inference Server for Linux contains a vulnerability in the tracing API, where a user can corrupt system fi

CVE-2024-0088 HIGH
8.1 May 14

NVIDIA Triton Inference Server for Linux contains a vulnerability in shared memory APIs, where a user can cause an impro

CVE-2026-24264 HIGH
7.5 Jul 01

Denial of service in NVIDIA Triton Inference Server for Linux allows remote unauthenticated attackers to exhaust server

CVE-2026-24266 HIGH
7.5 Jul 01

Denial of service in NVIDIA Triton Inference Server for Linux (versions through 26.03) allows a remote attacker to crash

CVE-2026-47629 HIGH
7.5 Aug 18

Denial of service in NVIDIA Triton Inference Server for Linux exposes AI inference infrastructure to remote disruption t

CVE-2026-47476 HIGH
7.5 Jul 14

Denial of service in NVIDIA Triton Inference Server for Linux allows remote unauthenticated attackers to trigger uncontr

Share

CVE-2026-47628 vulnerability details – vuln.today

This site uses cookies essential for authentication and security. No tracking or analytics cookies are used. Privacy Policy