< cd ../gallery
[AI & ML]

vllm-qwen

vLLM + Qwen3.6-27B (BF16) OpenAI-compatible inference server on AMD Strix Halo (Ryzen AI Max+ 395, gfx1151). Vision input, 256K context, /v1/responses with separated reasoning, via TheRock ROCm.
hec-ovi
@hec-ovi
maintainer
★ 15 stars
github
vllm-qwen — preview
vllm-qwen repository preview
# README
vLLM + Qwen3.6-27B (BF16) OpenAI-compatible inference server on AMD Strix Halo (Ryzen AI Max+ 395, gfx1151). Vision input, 256K context, /v1/responses with separated reasoning, via TheRock ROCm. Maintained by hec-ovi on GitHub, where it has earned 15 stars from the community.
It's actively developed around amd, docker, gfx1151, and is a solid reference for anyone building with these tools.
# tags
# install
npm install vllm-qwen
languages
Python100%
last commit3 months ago
licenseUnlicense
more vLLM repos
$ made-with-vllm
rssllmmadewithwhat

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.