A Toolkit to Help Optimize Onnx Model
☆504Jun 2, 2026Updated last month
Alternatives and similar repositories for OnnxSlim
Users that are interested in OnnxSlim are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Large Language Model Onnx Inference Framework☆35Nov 25, 2025Updated 8 months ago
- A Toolkit to Help Optimize Large Onnx Model☆166Jul 2, 2026Updated 3 weeks ago
- ☆11Sep 30, 2019Updated 6 years ago
- caffe model to onnx☆33Nov 16, 2022Updated 3 years ago
- 修改的DFace代码,可以完整训练得到MTCNN模型☆15Jan 24, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆18Jan 12, 2022Updated 4 years ago
- A lightweight, production-ready C++ library for LLM tokenization, fully compatible with HuggingFace tokenizer.json.☆33Jan 4, 2026Updated 6 months ago
- ONNX Optimizer☆825Updated this week
- llm deploy project based onnx.☆49Oct 9, 2024Updated last year
- Simplify your onnx model☆4,373Updated this week
- llm-export can export llm model to onnx.☆355May 8, 2026Updated 2 months ago
- caffe to tensorrt☆17Jan 24, 2019Updated 7 years ago
- Model compression for ONNX☆103May 1, 2026Updated 2 months ago
- Cuda Version Image Processing API☆40Mar 17, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ONNX Script enables developers to naturally author ONNX functions and models using a subset of Python.☆448Updated this week
- Efficient in-memory representation for ONNX, in Python☆45Updated this week
- 一款简单易用和高性能的AI部署框架 | An Easy-to-Use and High-Performance AI Deployment Framework☆1,848Apr 25, 2026Updated 3 months ago
- Use safetensors with ONNX 🤗☆88Updated this week
- A lightweight, single-header C++11 Jinja2 template engine for LLM chat templates.☆20Mar 4, 2026Updated 4 months ago
- Everything in Torch Fx☆343Jun 7, 2024Updated 2 years ago
- A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative…☆3,306Updated this week
- A tool to modify ONNX models in a visualization fashion, based on Netron and Flask.☆1,624Jun 27, 2026Updated 3 weeks ago
- a ai infra framework for edge device base on nndeploy☆18Nov 27, 2025Updated 7 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Common utilities for ONNX converters☆303Dec 16, 2025Updated 7 months ago
- A tool convert TensorRT engine/plan to a fake onnx☆41Nov 22, 2022Updated 3 years ago
- Implementation of YOLOv9 QAT optimized for deployment on TensorRT platforms.☆139Apr 24, 2025Updated last year
- Native GStreamer plugins that integrate SAHI (Slicing Aided Hyper Inference) into NVIDIA DeepStream for real-time small object detection …☆30Jun 8, 2026Updated last month
- ☆126Dec 15, 2023Updated 2 years ago
- [AAAI2025] SUTrack: Towards Simple and Unified Single Object Tracking. Converter to onnx model file.☆16Apr 2, 2026Updated 3 months ago
- mnn asr demo.☆27Mar 24, 2025Updated last year
- StyleTTS2 + Vocos as a Decoder☆13Mar 24, 2025Updated last year
- onnxruntime-extensions: A specialized pre- and post- processing library for ONNX Runtime☆474Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Quantize yolov5 using pytorch_quantization.🚀🚀🚀☆15Oct 24, 2023Updated 2 years ago
- Multi-stream video inference with Ultralytics YOLO - Display multiple video streams in a grid layout with real-time object detection.☆16May 20, 2026Updated 2 months ago
- Generative AI extensions for onnxruntime☆1,089Updated this week
- Detect CPU features with single-file☆460May 22, 2026Updated 2 months ago
- ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator☆21,188Updated this week
- 🚀 Easier & Faster YOLO Deployment Toolkit for NVIDIA 🛠️☆1,859Mar 22, 2026Updated 4 months ago
- Feature-based knowledge distillation (CWD & MGD) for YOLOv9, built on the MIT-licensed YOLO repo☆16Apr 27, 2026Updated 2 months ago