Skip to main content

Nvidia Model Optimizer: A unified library of SOTA model optimization techniques like quantization, pruning, Neural Architecture Search (NAS), distillation, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.

Project description

Checkout https://github.com/nvidia/Model-Optimizer for more information.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

nvidia_modelopt-0.46.0rc0.tar.gz (2.4 kB view details)

Uploaded Source

File details

Details for the file nvidia_modelopt-0.46.0rc0.tar.gz.

File metadata

  • Download URL: nvidia_modelopt-0.46.0rc0.tar.gz
  • Upload date:
  • Size: 2.4 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.14.6

File hashes

Hashes for nvidia_modelopt-0.46.0rc0.tar.gz
Algorithm Hash digest
SHA256 c006f329ecb0196ec28c18b3ceae22455972a370cd30dd9c40e178ebba47f5d0
MD5 46197773b9f4c2d4406bafde13a74c6e
BLAKE2b-256 44fd0cebe97f479c3fa3f8b20ff30a31c18e51708ccca720f3ff4d6977e3fd15

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page