< cd ../gallery
[AI & ML]
pegaflow
High-performance KV cache storage for LLM inference — GPU offloading, SSD caching, and cross-node sharing via RDMA. Works with vLLM and SGLang.
@novitalabs
maintainer
★ 170 stars
# README
High-performance KV cache storage for LLM inference — GPU offloading, SSD caching, and cross-node sharing via RDMA. Works with vLLM and SGLang. Maintained by novitalabs on GitHub, where it has earned 170 stars from the community.
It's actively developed around inference, kv-cache, llm, and is a solid reference for anyone building with these tools.
# tags
# install
npm install pegaflowmore vLLM repos