PyTorch Quantization Aware Training Example
☆150May 18, 2024Updated 2 years ago
Alternatives and similar repositories for PyTorch-Quantization-Aware-Training
Users that are interested in PyTorch-Quantization-Aware-Training are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch Static Quantization Example☆41Apr 29, 2021Updated 5 years ago
- A simple network quantization demo using pytorch from scratch.☆543Jun 18, 2023Updated 3 years ago
- PyTorch Pruning Example☆53Dec 5, 2022Updated 3 years ago
- Manually implemented quantization-aware training☆22Oct 12, 2022Updated 3 years ago
- Brevitas: neural network quantization in PyTorch☆1,563Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Revisit Kernel Pruning with Lottery Regulated Grouped Convolutions. ICLR 2022☆11Nov 24, 2022Updated 3 years ago
- ☆16May 3, 2024Updated 2 years ago
- micronet, a model compression and deploy lib. compression: 1、quantization: quantization-aware-training(QAT), High-Bit(>2b)(DoReFa/Quantiz…☆2,265May 6, 2025Updated last year
- Delay estimation logic extracted from WebRTC☆18Jan 11, 2021Updated 5 years ago
- Nsight Compute In Docker☆13Dec 21, 2023Updated 2 years ago
- Team <skyb> solution for the AIM2020 mobile image signal processing challenge☆17Mar 15, 2021Updated 5 years ago
- Transformation process of a Python Pytorch GPU model into an optimized TensorRT C++ one.☆13Mar 8, 2021Updated 5 years ago
- 使用RV1126部署YOLOv5模型☆18May 23, 2024Updated 2 years ago
- ONNX Runtime Inference C++ Example☆262Apr 3, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Multiple-variance Volterra series Identification Tool☆16Sep 25, 2021Updated 4 years ago
- yolov5第四版☆15Oct 13, 2021Updated 4 years ago
- NanoDet for Jetson Nano☆11Sep 30, 2023Updated 2 years ago
- Implementation of Speculative Sampling as described in "Accelerating Large Language Model Decoding with Speculative Sampling" by Deepmind☆111Feb 29, 2024Updated 2 years ago
- A list of papers, docs, codes about model quantization. This repo is aimed to provide the info for model quantization research, we are co…☆2,423Jul 10, 2026Updated last month
- ☆213Nov 9, 2021Updated 4 years ago
- PyTorch implementation for the APoT quantization (ICLR 2020)☆288Dec 11, 2024Updated last year
- fast, lightweight dbscan implementation for peptide strings☆12Apr 29, 2020Updated 6 years ago
- Quantize,Pytorch,Vgg16,MobileNet☆44Jan 29, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Pytorch implementation of our paper accepted by ECCV2022 -- Dynamic Dual Trainable Bounds for Ultra-low Precision Super-Resolution Networ…☆30Sep 13, 2022Updated 3 years ago
- Diagonalwise Refactorization: An Efficient Training Method for Depthwise Convolutions (in PyTorch)☆21May 13, 2018Updated 8 years ago
- [CVPR 2023] DepGraph: Towards Any Structural Pruning; LLMs, Vision Foundation Models, etc.☆3,341Sep 7, 2025Updated 11 months ago
- AIMET is a library that provides advanced quantization and compression techniques for trained neural network models.☆2,675Updated this week
- Implementation of "SpEx: Multi-Scale Time Domain Speaker Extraction Network".☆37Jul 19, 2020Updated 6 years ago
- TensorRT-in-Action 是一个 GitHub 代码库,提供了使用 TensorRT 的代码示例,并有对应 Jupyter Notebook。☆15Jun 1, 2023Updated 3 years ago
- percepnet implemented using Keras, still need to be optimized and tuned.☆39Jul 23, 2021Updated 5 years ago
- Code repo for the paper "LLM-QAT Data-Free Quantization Aware Training for Large Language Models"☆328Mar 4, 2025Updated last year
- I'm going to use the Winograd’s minimal filtering algorithms to introduce a new class of fast algorithms for convolutional neural networks…☆12Mar 22, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Model Quantization Benchmark☆878Apr 20, 2025Updated last year
- Improving Post Training Neural Quantization: Layer-wise Calibration and Integer Programming☆36Jun 29, 2023Updated 3 years ago
- PyTorch implementation of 'Deep Compression: Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding' by …☆428Feb 27, 2020Updated 6 years ago
- A simple application of DTW Algorithm in isolate word speech recognition.☆17Mar 9, 2020Updated 6 years ago
- ☆12Feb 24, 2025Updated last year
- ipython notebooks for feature extraction and training of audio event classifier on ESC-50 dataset.☆10Mar 16, 2018Updated 8 years ago
- Attempts to prune yolo v3 tiny.☆10Dec 13, 2018Updated 7 years ago