< cd ../gallery
[AI & ML]

ray_vllm_inference

A simple service that integrates vLLM with Ray Serve for fast and scalable LLM serving.
asprenger
@asprenger
maintainer
★ 79 stars
github
ray-vllm-inference — preview
ray_vllm_inference repository preview
# README
A simple service that integrates vLLM with Ray Serve for fast and scalable LLM serving. Maintained by asprenger on GitHub, where it has earned 79 stars from the community.
It's actively developed around inference, llm, llm-serving, and is a solid reference for anyone building with these tools.
# tags
# install
npm install ray_vllm_inference
languages
Python100%
last commit2 years ago
licenseApache-2.0
more vLLM repos
$ made-with-vllm
rssllmmadewithwhat

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.