< cd ../gallery
[AI & ML]

gfx906-fa-vllm

FlashAttention-style custom attention backend for vLLM on AMD MI50/MI60/Radeon VII (gfx906). Downstream fork of mixa3607/ML-gfx906 with replacement HIP kernels and a vllm.general_plugins entry point.
nick413-bit
@nick413-bit
maintainer
★ 12 stars
go to website ↗github
gfx906-fa-vllm — preview
gfx906-fa-vllm repository preview
# README
FlashAttention-style custom attention backend for vLLM on AMD MI50/MI60/Radeon VII (gfx906). Downstream fork of mixa3607/ML-gfx906 with replacement HIP kernels and a vllm.general_plugins entry point. Maintained by nick413-bit on GitHub, where it has earned 12 stars from the community.
It's actively developed around amd-gpu, flash-attention, gfx906, and is a solid reference for anyone building with these tools.
# tags
# install
npm install gfx906-fa-vllm
languages
Python100%
last commit3 months ago
licenseApache-2.0
more vLLM repos
$ made-with-vllm
rssllmmadewithwhat

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.