< cd ../gallery
[AI & ML]

ray_vllm_inference

A simple service that integrates vLLM with Ray Serve for fast and scalable LLM serving.
asprenger
@asprenger
maintainer
★ 79 stars
github
ray-vllm-inference — preview
ray_vllm_inference repository preview
# README
A simple service that integrates vLLM with Ray Serve for fast and scalable LLM serving. Maintained by asprenger on GitHub, where it has earned 79 stars from the community.
It's actively developed around inference, llm, llm-serving, and is a solid reference for anyone building with these tools.
# tags
# install
npm install ray_vllm_inference
languages
last commit2y ago
licenseApache-2.0
latest versions
v0.2.0v0.1.0
more vLLM repos

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.