TinyNeuralNetwork is an efficient and easy-to-use deep learning model compression framework.
☆879Mar 3, 2026Updated 4 months ago
Alternatives and similar repositories for TinyNeuralNetwork
Users that are interested in TinyNeuralNetwork are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- AIMET is a library that provides advanced quantization and compression techniques for trained neural network models.☆2,670Updated this week
- Model Quantization Benchmark☆874Apr 20, 2025Updated last year
- PPL Quantization Tool (PPQ) is a powerful offline neural network quantization tool.☆1,807Mar 28, 2024Updated 2 years ago
- micronet, a model compression and deploy lib. compression: 1、quantization: quantization-aware-training(QAT), High-Bit(>2b)(DoReFa/Quantiz…☆2,266May 6, 2025Updated last year
- [ACCV2022 (Oral)] Efficient Hardware-aware Neural Architecture Search for Image Super-resolution on Mobile Devices☆18Oct 5, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR 2023] DepGraph: Towards Any Structural Pruning; LLMs, Vision Foundation Models, etc.☆3,328Sep 7, 2025Updated 10 months ago
- Simplify your onnx model☆4,372Updated this week
- A Simple framework for image restoration, it includes ECBSR, ELAN and other SOTAs.☆49Nov 13, 2022Updated 3 years ago
- A model compression and acceleration toolbox based on pytorch.☆331Jan 12, 2024Updated 2 years ago
- 针对pytorch模型的自动化模型结构分析和修改工具集,包含自动分析模型结构的模型压缩算法库☆260Apr 19, 2023Updated 3 years ago
- OpenMMLab Model Compression Toolbox and Benchmark.☆1,677Jun 11, 2024Updated 2 years ago
- Neural Network Compression Framework for enhanced OpenVINO™ inference☆1,183Updated this week
- [ICLR 2020] Once for All: Train One Network and Specialize it for Efficient Deployment☆1,953Dec 14, 2023Updated 2 years ago
- YOLOv5 🚀 in PyTorch > ONNX > CoreML > TFLite > UF2☆21Oct 15, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Offline Quantization Tools for Deploy.☆143Dec 28, 2023Updated 2 years ago
- MegCC是一个运行时超轻量,高效,移植简单的深度学习模型编译器☆483Oct 23, 2024Updated last year
- A primitive library for neural network☆1,367Nov 24, 2024Updated last year
- Personalized AEC☆19Nov 3, 2022Updated 3 years ago
- A tool for converting ONNX files to LiteRT/TFLite/TensorFlow, PyTorch native code (nn.Module), TorchScript (.pt), state_dict (.pt), Expor…☆984Updated this week
- A tool to modify ONNX models in a visualization fashion, based on Netron and Flask.☆1,624Jun 27, 2026Updated 3 weeks ago
- edge-SR: Super-Resolution For The Masses☆61Jan 1, 2022Updated 4 years ago
- (ECCV'2020 Oral)EagleEye: Fast Sub-net Evaluation for Efficient Neural Network Pruning☆308Dec 8, 2022Updated 3 years ago
- Group Fisher Pruning for Practical Network Compression(ICML2021)☆163May 24, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 🛠A lite C++ AI toolkit: 100+ models with MNN, ORT and TRT, including Det, Seg, Stable-Diffusion, Face-Fusion, etc.🎉☆4,418Mar 19, 2026Updated 4 months ago
- RepVGG: Making VGG-style ConvNets Great Again☆3,479Feb 10, 2023Updated 3 years ago
- [NeurIPS 2020] MCUNet: Tiny Deep Learning on IoT Devices; [NeurIPS 2021] MCUNetV2: Memory-Efficient Patch-based Inference for Tiny Deep L…☆951Nov 27, 2024Updated last year
- ☆47Jun 6, 2021Updated 5 years ago
- A simple network quantization demo using pytorch from scratch.☆543Jun 18, 2023Updated 3 years ago
- NanoDet-Plus⚡Super fast and lightweight anchor-free object detection model. 🔥Only 980 KB(int8) / 1.8MB (fp16) and run 97FPS on cellphone…☆6,238Aug 8, 2024Updated last year
- Bolt is a deep learning library with high performance and heterogeneous flexibility.☆958Apr 11, 2025Updated last year
- [NeurIPS 2020] MCUNet: Tiny Deep Learning on IoT Devices; [NeurIPS 2021] MCUNetV2: Memory-Efficient Patch-based Inference for Tiny Deep L…☆704Mar 29, 2024Updated 2 years ago
- A list of papers, docs, codes about model quantization. This repo is aimed to provide the info for model quantization research, we are co…☆2,409Jul 10, 2026Updated 2 weeks ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A DNN inference latency prediction toolkit for accurately modeling and predicting the latency on diverse edge devices.☆364Jul 30, 2024Updated last year
- ☆214Dec 4, 2023Updated 2 years ago
- PyTorch Neural Network eXchange☆708Updated this week
- MNN: A blazing-fast, lightweight inference engine battle-tested by Alibaba, powering high-performance on-device LLMs and Edge AI.☆15,715Updated this week
- Brevitas: neural network quantization in PyTorch☆1,555Updated this week
- ☆33Nov 29, 2022Updated 3 years ago
- EasyQuant(EQ) is an efficient and simple post-training quantization method via effectively optimizing the scales of weights and activatio…☆407Nov 22, 2022Updated 3 years ago