← Back to feed Fashion & Style

Show HN: Open-source engine running Gemma 4 26B in 2 GB RAM on any M-series Mac

Hacker News Best 29 July 2026 7h ago
Show HN: Open-source engine running Gemma 4 26B in 2 GB RAM on any M-series Mac
63
Relevance
3/25
Freshness
25/25
Authority
18/20
Brand Signal
11/15
Depth
6/15
Relevance Freshness Authority Brand Depth
Hi HN, I built a specialized inference engine for running 4-bit Gemma 4 26B-A4B-IT on any M-series Mac using about 2 GB of RAM. It is called TurboFieldfare and is written in Swift and Metal. I have always adored on-device AI. It feels like magic that you can run a powerful NN on your Mac or iPhone. So I wanted to push the limits a bit and run a model whose weights don’t fit in memory. The model’s
Read Full Article → Hacker News Best ↗