< cd ../gallery
[AI & ML]
gfx906-fa-vllm
FlashAttention-style custom attention backend for vLLM on AMD MI50/MI60/Radeon VII (gfx906). Downstream fork of mixa3607/ML-gfx906 with replacement HIP kernels and a vllm.general_plugins entry point.
@nick413-bit
maintainer
★ 12 stars
# README
FlashAttention-style custom attention backend for vLLM on AMD MI50/MI60/Radeon VII (gfx906). Downstream fork of mixa3607/ML-gfx906 with replacement HIP kernels and a vllm.general_plugins entry point. Maintained by nick413-bit on GitHub, where it has earned 12 stars from the community.
It's actively developed around amd-gpu, flash-attention, gfx906, and is a solid reference for anyone building with these tools.
# tags
# install
npm install gfx906-fa-vllmmore vLLM repos