< cd ../gallery
[AI & ML]

pegaflow

High-performance KV cache storage for LLM inference — GPU offloading, SSD caching, and cross-node sharing via RDMA. Works with vLLM and SGLang.
novitalabs
@novitalabs
maintainer
★ 184 stars
github
pegaflow — preview
pegaflow repository preview
# README
High-performance KV cache storage for LLM inference — GPU offloading, SSD caching, and cross-node sharing via RDMA. Works with vLLM and SGLang. Maintained by novitalabs on GitHub, where it has earned 184 stars from the community.
It's actively developed around inference, kv-cache, llm, and is a solid reference for anyone building with these tools.
# tags
# install
npm install pegaflow
languages
Rust60%
last commit4w ago
licenseApache-2.0
more vLLM repos

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.