< cd ../gallery
[AI & ML]
sndr_core_engine
SNDR Core Engine (Genesis) — vLLM runtime patch-overlay for Qwen3.6 + Gemma4 on consumer NVIDIA (Ampere sm_86, 2× A5000/3090). Qwen3.6-35B-A3B FP8 ~240 tok/s, 27B-int4 hybrid GDN+Mamba, Gemma4 26B/31B AWQ, 256K ctx. 321 patches: TurboQuant k8v4 KV, MTP/DFlash spec-decode, FULL cudagraph, hybrid GDN. vLLM pin dev424 + Control Center GUI.
@Sandermage
maintainer
★ 122 stars
# README
SNDR Core Engine (Genesis) — vLLM runtime patch-overlay for Qwen3.6 + Gemma4 on consumer NVIDIA (Ampere sm_86, 2× A5000/3090). Qwen3.6-35B-A3B FP8 ~240 tok/s, 27B-int4 hybrid GDN+Mamba, Gemma4 26B/31B AWQ, 256K ctx. 321 patches: TurboQuant k8v4 KV, MTP/DFlash spec-decode, FULL cudagraph, hybrid GDN. vLLM pin dev424 + Control Center GUI. Maintained by Sandermage on GitHub, where it has earned 122 stars from the community.
It's actively developed around awq, consumer-gpu, cuda, and is a solid reference for anyone building with these tools.
# tags
# install
npm install sndr_core_enginelanguages
Python90.449%
TypeScript6.31%
CSS1.848%
Shell0.906%
Jinja0.278%
Makefile0.158%
Dockerfile0.027%
JavaScript0.012%
HTML0.012%
last commit1 week ago
licenseApache-2.0
more vLLM repos