< cd ../gallery
[AI & ML]
fastassert
Dockerized LLM inference server with constrained output (JSON mode), built on top of vLLM and outlines. Faster, cheaper and without rate limits. Compare the quality and latency to your current LLM API provider.
@phospho-app
maintainer
★ 27 stars
# README
Dockerized LLM inference server with constrained output (JSON mode), built on top of vLLM and outlines. Faster, cheaper and without rate limits. Compare the quality and latency to your current LLM API provider. Maintained by phospho-app on GitHub, where it has earned 27 stars from the community.
It's actively developed around docker, llm, llm-inference, and is a solid reference for anyone building with these tools.
# tags
# install
npm install fastassertmore vLLM repos