< cd ../gallery
[AI & ML]

llm-bench-rig

Dual-engine (llama.cpp + vLLM) LLM benchmarking pipeline for GGUF & safetensors on NVIDIA GPUs — speed, quality, live dashboard, publishable cards.
notwitcheer
@notwitcheer
maintainer
★ 23 stars
github
llm-bench-rig — preview
llm-bench-rig repository preview
# README
Dual-engine (llama.cpp + vLLM) LLM benchmarking pipeline for GGUF & safetensors on NVIDIA GPUs — speed, quality, live dashboard, publishable cards. Maintained by notwitcheer on GitHub, where it has earned 23 stars from the community.
It's actively developed around benchmarking, cuda, fastapi, and is a solid reference for anyone building with these tools.
# tags
# install
npm install llm-bench-rig
languages
Python89.259%
HTML5.609%
Shell5.132%
last commit2 weeks ago
license
more vLLM repos
$ made-with-vllm
rssllmmadewithwhat

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.