☆11Jan 10, 2025Updated last year
Alternatives and similar repositories for quant_horizon
Users that are interested in quant_horizon are comparing it to the libraries listed below
Sorting:
- Summary of system papers/frameworks/codes/tools on training or serving large model☆57Dec 17, 2023Updated 2 years ago
- Offline Quantization Tools for Deploy.☆142Dec 28, 2023Updated 2 years ago
- ☆13Feb 16, 2022Updated 4 years ago
- NART = NART is not A RunTime, a deep learning inference framework.☆37Mar 2, 2023Updated 3 years ago
- [EMNLP 2024 & AAAI 2026] A powerful toolkit for compressing large models including LLMs, VLMs, and video generative models.☆688Mar 11, 2026Updated last week
- This is a repository of Binary General Matrix Multiply (BGEMM) by customized CUDA kernel. Thank FP6-LLM for the wheels!☆18Aug 30, 2024Updated last year
- An Tensorflow.keras implementation of Same, Same But Different - Recovering Neural Network Quantization Error Through Weight Factorizatio…☆10Dec 18, 2019Updated 6 years ago