Godwit
Product signal · 2026-08-09
What it does
Swift and Metal engine (zero dependencies) that runs 120-billion-parameter mixture-of-experts AI models on Apple Silicon Macs with limited RAM, such as a 16GB MacBook Air, by keeping a shared core in memory and streaming expert sub-networks from the SSD.
Who it’s for
Developers and AI enthusiasts who want to run large open-weight local LLMs on base Apple Silicon Macs without large unified memory.
Why it entered the Feed
Admitted from a r/SideProject post (score 75, 35 comments) demonstrating a working Apache-2.0 engine running a 120B MoE model on a 16GB MacBook Air at ~1.4 words per second.
Keyword suggestions
Godwitrun large AI models on low-RAM Maclocal LLM inference on Apple Silicon
Report in preparationLike and save this Feed item. Digmoro prioritizes Reports for items with more likes and saves.
