< cd ../gallery
[AI & ML]
llm-inspector
The htop for LLM inference see exactly where every GB of VRAM goes and get measured quantization savings.
@helasaoudi
maintainer
★ 68 stars
# README
The htop for LLM inference see exactly where every GB of VRAM goes and get measured quantization savings. Maintained by helasaoudi on GitHub, where it has earned 68 stars from the community.
It's actively developed around gpu-monitoring, htop, inference, and is a solid reference for anyone building with these tools.
# tags
# install
npm install llm-inspectormore vLLM repos