< cd ../gallery
[AI & ML]

vLLM-Moet-GB10

Run DeepSeek-V4-Flash 304B on a single NVIDIA GB10 / DGX Spark - 2-bit MoE planes, FP4 quality recovery, speculative decoding
lrozewicz
@lrozewicz
maintainer
★ 11 stars
github
vllm-moet-gb10 — preview
vLLM-Moet-GB10 repository preview
# README
Run DeepSeek-V4-Flash 304B on a single NVIDIA GB10 / DGX Spark - 2-bit MoE planes, FP4 quality recovery, speculative decoding. Maintained by lrozewicz on GitHub, where it has earned 11 stars from the community.
It's actively developed around deepseek, dgx-spark, gb10, and is a solid reference for anyone building with these tools.
# tags
# install
npm install vLLM-Moet-GB10
languages
Sass89%
last commit3w ago
licenseApache-2.0
latest versions
v0.2.0v0.1.0
more vLLM repos

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.