Godwit

Product signal · 2026-08-09

What it does

Swift and Metal engine (zero dependencies) that runs 120-billion-parameter mixture-of-experts AI models on Apple Silicon Macs with limited RAM, such as a 16GB MacBook Air, by keeping a shared core in memory and streaming expert sub-networks from the SSD.

Who it’s for

Developers and AI enthusiasts who want to run large open-weight local LLMs on base Apple Silicon Macs without large unified memory.

Why it entered the Feed

Admitted from a r/SideProject post (score 75, 35 comments) demonstrating a working Apache-2.0 engine running a 120B MoE model on a 16GB MacBook Air at ~1.4 words per second.

Keyword suggestions

Godwitrun large AI models on low-RAM Maclocal LLM inference on Apple Silicon
Report in preparationLike and save this Feed item. Digmoro prioritizes Reports for items with more likes and saves.