OnnxSlim: A Toolkit to Help Optimize Large Onnx Model
Project description
OnnxSlim
OnnxSlim can help you slim your onnx model, with less operators, but same accuracy, better inference speed.
- 🚀 OnnxSlim is merged to mnn-llm, performance increased by 5%
- 🚀 Rank 1st in the AICAS 2024 LLM inference optimization challenge held by Arm and T-head
- 🚀 OnnxSlim is merged into ultralytics ❤️❤️❤️
- 🚀 OnnxSlim is merged into transformers.js 🤗🤗🤗
Installation
Using Prebuilt
pip install onnxslim
Install From Source
pip install git+https://github.com/inisis/OnnxSlim@main
Install From Local
git clone https://github.com/inisis/OnnxSlim && cd OnnxSlim/
pip install .
How to use
onnxslim your_onnx_model slimmed_onnx_model
For more usage, see onnxslim -h or refer to our examples
References
Contact
Discord: https://discord.gg/nRw2Fd3VUS QQ Group: 873569894
Project details
Release history Release notifications | RSS feed
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
onnxslim-0.1.34.tar.gz
(118.9 kB
view hashes)
Built Distribution
onnxslim-0.1.34-py3-none-any.whl
(140.3 kB
view hashes)
Close
Hashes for onnxslim-0.1.34-py3-none-any.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | 755cb13cfd7a0b47d33e5c1935db2bb7a6993e9098a33987c28794f9187cab17 |
|
MD5 | cc1d323474658cef6f2c9e2f378428cf |
|
BLAKE2b-256 | 2240a557ba55ac15bb31c1fcf654431116f37dd156dd80edcdc515f4d3748089 |