< cd ../gallery
[AI & ML]

local-llm-autotune

Zero-config local LLM optimization for Ollama, LM Studio, and Apple Silicon MLX. Reduces TTFT by 40%, wall time for local agents by 46%, and RAM usage by 3x.
tanavc1
@tanavc1
maintainer
★ 31 stars
go to website ↗github
local-llm-autotune — preview
local-llm-autotune repository preview
# README
Zero-config local LLM optimization for Ollama, LM Studio, and Apple Silicon MLX. Reduces TTFT by 40%, wall time for local agents by 46%, and RAM usage by 3x. Maintained by tanavc1 on GitHub, where it has earned 31 stars from the community.
It's actively developed around apple-silicon, cli, inference-optimization, and is a solid reference for anyone building with these tools.
# tags
# install
npm install local-llm-autotune
languages
Python84.276%
HTML6.754%
PLpgSQL0.748%
Shell0.507%
Ruby0.306%
CSS0.039%
last commit3 weeks ago
licenseMIT
more FastAPI repos
$ made-with-fastapi
rssllmmadewithwhat

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.