LLM model quantization (compression) toolkit with HW acceleration support for Nvidia, AMD, Intel GPU and Intel/AMD/Apple CPU via HF, vLLM, and SGLang.
Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.
LLM model quantization (compression) toolkit with HW acceleration support for Nvidia, AMD, Intel GPU and Intel/AMD/Apple CPU via HF, vLLM, and SGLang.
LLM model quantization (compression) toolkit with HW acceleration support for Nvidia, AMD, Intel GPU and Intel/AMD/Apple CPU via HF, vLLM, and SGLang.
GPTQModel is an open-source project.
Yes. GPTQModel is free and open source — you can use, modify and self-host it.
Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.
[](https://opensourceai.tech/project/modelcloud-gptqmodel.html)