< cd ../gallery
[AI & ML]

pegaflow

High-performance KV cache storage for LLM inference — GPU offloading, SSD caching, and cross-node sharing via RDMA. Works with vLLM and SGLang.
novitalabs
@novitalabs
maintainer
★ 170 stars
github
pegaflow — preview
pegaflow repository preview
# README
High-performance KV cache storage for LLM inference — GPU offloading, SSD caching, and cross-node sharing via RDMA. Works with vLLM and SGLang. Maintained by novitalabs on GitHub, where it has earned 170 stars from the community.
It's actively developed around inference, kv-cache, llm, and is a solid reference for anyone building with these tools.
# tags
# install
npm install pegaflow
languages
Rust58.873%
Python40.584%
Shell0.499%
last commit1 week ago
licenseApache-2.0
more vLLM repos
$ made-with-vllm
rssllmmadewithwhat

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.