< cd ../gallery
[AI & ML]

llm-inspector

The htop for LLM inference see exactly where every GB of VRAM goes and get measured quantization savings.
helasaoudi
@helasaoudi
maintainer
★ 68 stars
github
llm-inspector — preview
llm-inspector repository preview
# README
The htop for LLM inference see exactly where every GB of VRAM goes and get measured quantization savings. Maintained by helasaoudi on GitHub, where it has earned 68 stars from the community.
It's actively developed around gpu-monitoring, htop, inference, and is a solid reference for anyone building with these tools.
# tags
# install
npm install llm-inspector
languages
Python100%
last commit3w ago
licenseMIT
latest versions
v0.6.0
more vLLM repos

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.