TinyNeuralNetwork is an efficient and easy-to-use deep learning model compression framework.
☆880Mar 3, 2026Updated 5 months ago
Alternatives and similar repositories for TinyNeuralNetwork
Users that are interested in TinyNeuralNetwork are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- AIMET is a library that provides advanced quantization and compression techniques for trained neural network models.☆2,678Updated this week
- Model Quantization Benchmark☆879Apr 20, 2025Updated last year
- PPL Quantization Tool (PPQ) is a powerful offline neural network quantization tool.☆1,815Mar 28, 2024Updated 2 years ago
- micronet, a model compression and deploy lib. compression: 1、quantization: quantization-aware-training(QAT), High-Bit(>2b)(DoReFa/Quantiz…☆2,266May 6, 2025Updated last year
- [ACCV2022 (Oral)] Efficient Hardware-aware Neural Architecture Search for Image Super-resolution on Mobile Devices☆18Oct 5, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR 2023] DepGraph: Towards Any Structural Pruning; LLMs, Vision Foundation Models, etc.☆3,341Sep 7, 2025Updated 11 months ago
- Simplify your onnx model☆4,384Updated this week
- A Simple framework for image restoration, it includes ECBSR, ELAN and other SOTAs.☆49Nov 13, 2022Updated 3 years ago
- A model compression and acceleration toolbox based on pytorch.☆331Jan 12, 2024Updated 2 years ago
- 针对pytorch模型的自动化模型结构分析和修改工具集,包含自动分析模型结构的模型压缩算法库☆260Apr 19, 2023Updated 3 years ago
- OpenMMLab Model Compression Toolbox and Benchmark.☆1,682Jun 11, 2024Updated 2 years ago
- Neural Network Compression Framework for enhanced OpenVINO™ inference☆1,189Updated this week
- [ICLR 2020] Once for All: Train One Network and Specialize it for Efficient Deployment☆1,953Dec 14, 2023Updated 2 years ago
- YOLOv5 🚀 in PyTorch > ONNX > CoreML > TFLite > UF2☆21Oct 15, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Offline Quantization Tools for Deploy.☆143Dec 28, 2023Updated 2 years ago
- MegCC是一个运行时超轻量,高效,移植简单的深度学习模型编译器☆483Oct 23, 2024Updated last year
- A primitive library for neural network☆1,367Nov 24, 2024Updated last year
- Personalized AEC☆19Nov 3, 2022Updated 3 years ago
- A tool for converting ONNX files to LiteRT/TFLite/TensorFlow, PyTorch native code (nn.Module), TorchScript (.pt), state_dict (.pt), Expor…☆987Aug 1, 2026Updated last week
- A tool to modify ONNX models in a visualization fashion, based on Netron and Flask.☆1,626Jun 27, 2026Updated last month
- edge-SR: Super-Resolution For The Masses☆61Jan 1, 2022Updated 4 years ago
- (ECCV'2020 Oral)EagleEye: Fast Sub-net Evaluation for Efficient Neural Network Pruning☆308Dec 8, 2022Updated 3 years ago
- Group Fisher Pruning for Practical Network Compression(ICML2021)☆164May 24, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A lite C++ AI toolkit: 100+ models with MNN, ORT and TRT, including Det, Seg, Stable-Diffusion, Face-Fusion.☆4,420Mar 19, 2026Updated 4 months ago
- RepVGG: Making VGG-style ConvNets Great Again☆3,477Feb 10, 2023Updated 3 years ago
- [NeurIPS 2020] MCUNet: Tiny Deep Learning on IoT Devices; [NeurIPS 2021] MCUNetV2: Memory-Efficient Patch-based Inference for Tiny Deep L…☆952Nov 27, 2024Updated last year
- ☆47Jun 6, 2021Updated 5 years ago
- A simple network quantization demo using pytorch from scratch.☆543Jun 18, 2023Updated 3 years ago
- NanoDet-Plus⚡Super fast and lightweight anchor-free object detection model. 🔥Only 980 KB(int8) / 1.8MB (fp16) and run 97FPS on cellphone…☆6,248Aug 8, 2024Updated 2 years ago
- Bolt is a deep learning library with high performance and heterogeneous flexibility.☆959Apr 11, 2025Updated last year
- [NeurIPS 2020] MCUNet: Tiny Deep Learning on IoT Devices; [NeurIPS 2021] MCUNetV2: Memory-Efficient Patch-based Inference for Tiny Deep L…☆709Mar 29, 2024Updated 2 years ago
- A DNN inference latency prediction toolkit for accurately modeling and predicting the latency on diverse edge devices.☆364Jul 30, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A list of papers, docs, codes about model quantization. This repo is aimed to provide the info for model quantization research, we are co…☆2,424Jul 10, 2026Updated last month
- ☆214Dec 4, 2023Updated 2 years ago
- PyTorch Neural Network eXchange☆712Aug 7, 2026Updated last week
- MNN: A blazing-fast, lightweight inference engine battle-tested by Alibaba, powering high-performance on-device LLMs and Edge AI.☆15,867Updated this week
- Brevitas: neural network quantization in PyTorch☆1,563Updated this week
- ☆33Nov 29, 2022Updated 3 years ago
- EasyQuant(EQ) is an efficient and simple post-training quantization method via effectively optimizing the scales of weights and activatio…☆407Nov 22, 2022Updated 3 years ago