< cd ../gallery
[AI & ML]
local-llm-autotune
Zero-config local LLM optimization for Ollama, LM Studio, and Apple Silicon MLX. Reduces TTFT by 40%, wall time for local agents by 46%, and RAM usage by 3x.
@tanavc1
maintainer
★ 33 stars
# README
Zero-config local LLM optimization for Ollama, LM Studio, and Apple Silicon MLX. Reduces TTFT by 40%, wall time for local agents by 46%, and RAM usage by 3x. Maintained by tanavc1 on GitHub, where it has earned 33 stars from the community.
It's actively developed around apple-silicon, cli, inference-optimization, and is a solid reference for anyone building with these tools.
# tags
# install
npm install local-llm-autotunelatest versions
v1.0.0↗more FastAPI repos