< cd ../gallery
[AI & ML]
vllm-qwen
vLLM + Qwen3.6-27B (BF16) OpenAI-compatible inference server on AMD Strix Halo (Ryzen AI Max+ 395, gfx1151). Vision input, 256K context, /v1/responses with separated reasoning, via TheRock ROCm.
@hec-ovi
maintainer
★ 15 stars
# README
vLLM + Qwen3.6-27B (BF16) OpenAI-compatible inference server on AMD Strix Halo (Ryzen AI Max+ 395, gfx1151). Vision input, 256K context, /v1/responses with separated reasoning, via TheRock ROCm. Maintained by hec-ovi on GitHub, where it has earned 15 stars from the community.
It's actively developed around amd, docker, gfx1151, and is a solid reference for anyone building with these tools.
# tags
# install
npm install vllm-qwenmore vLLM repos