Model Compression Toolkit (MCT) is an open source project for neural network model optimization under efficient, constrained hardware. This project provides researchers, developers, and engineers advanced quantization and compression tools for deploying state-of-the-art neural networks.
☆453Oct 5, 2026Updated this week
Alternatives and similar repositories for mct-model-optimization
Users that are interested in mct-model-optimization are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆22Updated this week
- Fully quantized Neural Networks for Audio Source Separation☆17Aug 11, 2024Updated 2 years ago
- ☆49Jul 28, 2020Updated 6 years ago
- QONNX: Arbitrary-Precision Quantized Neural Networks in ONNX☆194Sep 1, 2026Updated last month
- AIMET is a library that provides advanced quantization and compression techniques for trained neural network models.☆2,726Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- FakeQuantize with Learned Step Size(LSQ+) as Observer in PyTorch☆38Dec 18, 2021Updated 4 years ago
- ☆28Oct 21, 2020Updated 5 years ago
- Post-training sparsity-aware quantization☆34Feb 26, 2023Updated 3 years ago
- Improving Post Training Neural Quantization: Layer-wise Calibration and Integer Programming☆36Jun 29, 2023Updated 3 years ago
- Quantization library for PyTorch. Support low-precision and mixed-precision quantization, with hardware implementation through TVM.☆464May 15, 2023Updated 3 years ago
- A model compression and acceleration toolbox based on pytorch.☆333Jan 12, 2024Updated 2 years ago
- A curated collection of papers, benchmarks, surveys, and tools for model quantization, covering low-bit networks, LLMs, multimodal and ge…☆2,454Updated this week
- [IJCAI 2022] FQ-ViT: Post-Training Quantization for Fully Quantized Vision Transformer☆363Apr 11, 2023Updated 3 years ago
- Discover, configure, and deploy containerised software to Arm hardware over SSH☆50Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The official PyTorch implementation of the ICLR2022 paper, QDrop: Randomly Dropping Quantization for Extremely Low-bit Post-Training Quan…☆135Sep 23, 2025Updated last year
- The code for our paper "Neural Architecture Search as Program Transformation Exploration"☆17Apr 28, 2021Updated 5 years ago
- Model Quantization Benchmark☆882Apr 20, 2025Updated last year
- Neural Network Quantization With Fractional Bit-widths☆11Feb 19, 2021Updated 5 years ago
- The PyTorch implementation of Learned Step size Quantization (LSQ) in ICLR2020 (unofficial)☆139Nov 19, 2020Updated 5 years ago
- BitPack is a practical tool to efficiently save ultra-low precision/mixed-precision quantized models.☆58Feb 7, 2023Updated 3 years ago
- Pytorch implementation of BRECQ, ICLR 2021☆302Aug 1, 2021Updated 5 years ago
- Brevitas: neural network quantization in PyTorch☆1,585Updated this week
- Improving Post Training Neural Quantization: Layer-wise Calibration and Integer Programming☆98Jun 10, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆24Oct 7, 2021Updated 5 years ago
- [CVPR 2023] DepGraph: Towards Any Structural Pruning; LLMs, Vision Foundation Models, etc.☆3,359Sep 7, 2025Updated last year
- PyTorch implementation for the APoT quantization (ICLR 2020)☆287Dec 11, 2024Updated last year
- Code that accompanies the paper Bayesian Uncertainty for Gradient Aggregation in Multi-Task Learning - Accepted to ICML2024☆15May 8, 2025Updated last year
- Awesome Quantization Paper lists with Codes☆10Feb 24, 2021Updated 5 years ago
- HandLandmark Detection that can be performed only in onnxruntime. Pre-focusing by skeletal detection is not performed. This does not use …☆23Apr 30, 2024Updated 2 years ago
- Efficient Latent Image Restoration☆17Dec 28, 2025Updated 9 months ago
- ☆346Feb 12, 2026Updated 7 months ago
- ☆173Mar 9, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆45Jul 14, 2021Updated 5 years ago
- Fully Quantized Neural Networks For Speech Enhancement☆67Feb 15, 2024Updated 2 years ago
- Neural Architecture Search for Neural Network Libraries☆64Jul 24, 2026Updated 2 months ago
- Offline Quantization Tools for Deploy.☆143Dec 28, 2023Updated 2 years ago
- BSQ: Exploring Bit-Level Sparsity for Mixed-Precision Neural Network Quantization (ICLR 2021)☆41Jan 12, 2021Updated 5 years ago
- ☆82Jul 21, 2022Updated 4 years ago
- [CVPR'20] ZeroQ: A Novel Zero Shot Quantization Framework☆282Dec 8, 2023Updated 2 years ago