Skip to main content

onnxruntime-ep-nv-tensorrt-rtx

NVIDIA TensorRT RTX Execution Provider plugin for ONNX Runtime.

Enables hardware-accelerated inference on NVIDIA RTX GPUs (Ampere / RTX 30xx and later) via the ORT Plugin EP ABI.


About NVIDIA TensorRT for RTX

NVIDIA® TensorRT™ for RTX (TensorRT-RTX) is an inference optimization library dedicated for deploying AI inference on NVIDIA GeForce RTX GPUs. It is a great choice for developers building applications that must run on Windows or Linux PCs, laptops, or workstations.

This package bundles the TensorRT-RTX runtime libraries alongside the ONNX Runtime EP plugin so that no separate TensorRT-RTX installation is required.

For more information about TensorRT-RTX, visit https://developer.nvidia.com/tensorrt-rtx.
Online documentation: https://docs.nvidia.com/deeplearning/tensorrt-rtx/latest/index.html
License agreement: https://docs.nvidia.com/deeplearning/tensorrt-rtx/latest/reference/sla.html


References


Requirements

  • NVIDIA RTX GPU (Ampere or later)
  • NVIDIA GPU driver with CUDA 12 support
  • pip install onnxruntime>=1.24

Installation

pip install onnxruntime>=1.24
pip install onnxruntime-ep-nv-tensorrt-rtx

Usage

import onnxruntime as ort
import onnxruntime_ep_nv_tensorrt_rtx as trt_ep

# Register the EP plugin
ort.register_execution_provider_library(trt_ep.get_ep_name(), trt_ep.get_library_path())

# List available devices
devices = [d for d in ort.get_ep_devices() if d.ep_name == trt_ep.get_ep_name()]
print(f"TensorRT RTX devices: {len(devices)}")

# Create session with EP
so = ort.SessionOptions()
so.add_provider_for_devices(devices, {})
sess = ort.InferenceSession("model.onnx", sess_options=so)

License

Apache 2.0. See LICENSE.

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distributions

No source distribution files available for this release.See tutorial on generating distribution archives.

Built Distributions

If you're not sure about the file name format, learn more about wheel file names.

onnxruntime_ep_nv_tensorrt_rtx_cu12-0.4.0-py3-none-win_amd64.whl (105.8 MB view details)

Uploaded Python 3Windows x86-64

onnxruntime_ep_nv_tensorrt_rtx_cu12-0.4.0-py3-none-manylinux_2_28_x86_64.whl (167.7 MB view details)

Uploaded Python 3manylinux: glibc 2.28+ x86-64

File details

Details for the file onnxruntime_ep_nv_tensorrt_rtx_cu12-0.4.0-py3-none-win_amd64.whl.

File metadata

File hashes

Hashes for onnxruntime_ep_nv_tensorrt_rtx_cu12-0.4.0-py3-none-win_amd64.whl
Algorithm Hash digest
SHA256 5d9dba6aafd8e0863f34d0e5f51ad7c7ae9a5d6a92a55fe45d9637d3a5f1e1a7
MD5 c0d625beb539ff217f95101a41ea541a
BLAKE2b-256 8be4f19fdbffcf8faf18d9f81973f81e3d378ef371f857b049a3f355535a015c

See more details on using hashes here.

File details

Details for the file onnxruntime_ep_nv_tensorrt_rtx_cu12-0.4.0-py3-none-manylinux_2_28_x86_64.whl.

File metadata

File hashes

Hashes for onnxruntime_ep_nv_tensorrt_rtx_cu12-0.4.0-py3-none-manylinux_2_28_x86_64.whl
Algorithm Hash digest
SHA256 221807b1797a6270f37dfd7981f66e25ca7483f3883bfe23ef2b6a4198fe2b10
MD5 021911f648b7f934f2ffc0ce17c36675
BLAKE2b-256 52a61405069d40e2d6d37b21031e5dc2cc8b8197fbb2e3d4e4f65b9155c57e4f

See more details on using hashes here.

Release history Release notifications | RSS feed

This release

0.4.0 This release

2 files

0.3.0

3 files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page