wejoncy / QLLMView on GitHub
A general 2-8 bits quantization toolbox with GPTQ/AWQ/HQQ/VPTQ, and export to onnx/onnx-runtime easily.
184Apr 2, 2025Updated 11 months ago

Alternatives and similar repositories for QLLM

Users that are interested in QLLM are comparing it to the libraries listed below

Sorting:

Are these results useful?