< cd ../gallery
[AI & ML]

GPTQModel

LLM model quantization (compression) toolkit with HW acceleration support for Nvidia, AMD, Intel GPU and Intel/AMD/Apple CPU via HF, vLLM, and SGLang.
ModelCloud
@ModelCloud
maintainer
★ 1.2k stars
go to website ↗github
gptqmodel — preview
GPTQModel repository preview
# README
LLM model quantization (compression) toolkit with HW acceleration support for Nvidia, AMD, Intel GPU and Intel/AMD/Apple CPU via HF, vLLM, and SGLang. Maintained by ModelCloud on GitHub, where it has earned 1,205 stars from the community.
It's actively developed around gptq, optimum, peft, and is a solid reference for anyone building with these tools.
# tags
# install
npm install GPTQModel
languages
Python85.171%
Cuda10.157%
C++4.535%
C0.076%
Shell0.061%
last commit1 week ago
license
more FastAPI repos
$ made-with-fastapi
rssllmmadewithwhat

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.