< cd ../gallery
[AI & ML]

dual-rtx-6000-blackwell-qwen3.6-27b-fp8

Optimized vLLM setup for Qwen3.6-27B-FP8 on dual RTX PRO 6000 Blackwell (192 GB GDDR7, no NVLink); config, benchmark sweep results, and custom chat template with thinking mode off by default.
theogravity
@theogravity
maintainer
★ 15 stars
github
dual-rtx-6000-blackwell-qwen3-6-27b-fp8 — preview
dual-rtx-6000-blackwell-qwen3.6-27b-fp8 repository preview
# README
Optimized vLLM setup for Qwen3.6-27B-FP8 on dual RTX PRO 6000 Blackwell (192 GB GDDR7, no NVLink); config, benchmark sweep results, and custom chat template with thinking mode off by default. Maintained by theogravity on GitHub, where it has earned 15 stars from the community.
It's actively developed around benchmark, blackwell, fp8, and is a solid reference for anyone building with these tools.
# tags
# install
npm install dual-rtx-6000-blackwell-qwen3.6-27b-fp8
languages
Shell100%
last commit2 months ago
licenseMIT
more vLLM repos
$ made-with-vllm
rssllmmadewithwhat

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.