wejoncy / QLLM

A general 2-8 bits quantization toolbox with GPTQ/AWQ/HQQ, and export to onnx/onnx-runtime easily.
Apache License 2.0
145 stars 14 forks source link

add assert message && ci upgrade torch 2.2.2 #124

Closed wejoncy closed 3 months ago