GPTQ vs AWQ for LLM Weight Quantization
AWQ edges out GPTQ on accuracy while both depend on optimized kernels for speed.
Vera Mbeki
Staff Writer
Vera Mbeki is a staff writer at Open Inference Review covering quantization methods. Based in Austin, Vera has written for Open Inference Review since 2017.
1 story · Austin
AWQ edges out GPTQ on accuracy while both depend on optimized kernels for speed.