Introducting TextCLF Quant Factory
Hello All! I created a GitHub repo where you can quantize models from hugging to 4-bit using a quantization method I created called TextCLF TQ and use vllm for inference. TQ is a calibration-free quant method. Meaning you can quantize new models instantly without worrying about having calibration d…
Read the full story at r/huggingface ↗
Timeline · 1 report
- 2026-09-23 17:26 · r/huggingface
Introducting TextCLF Quant Factory