< cd ../gallery
[AI & ML]

llm-inference

llm-inference is a platform for publishing and managing llm inference, providing a wide range of out-of-the-box features for model deployment, such as UI, RESTful API, auto-scaling, computing resource management, monitoring, and more.
OpenCSGs
@OpenCSGs
maintainer
★ 95 stars
github
llm-inference — preview
llm-inference repository preview
# README
llm-inference is a platform for publishing and managing llm inference, providing a wide range of out-of-the-box features for model deployment, such as UI, RESTful API, auto-scaling, computing resource management, monitoring, and more. Maintained by OpenCSGs on GitHub, where it has earned 95 stars from the community.
It's actively developed around deepspeed, llama-cpp, llm-inference, and is a solid reference for anyone building with these tools.
# tags
# install
npm install llm-inference
languages
Python100%
last commit2 years ago
licenseApache-2.0
more vLLM repos
$ made-with-vllm
rssllmmadewithwhat

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.