Show HN: Open-source engine running Gemma 4 26B in 2 GB RAM on any M-series Mac
Culture Index
Score Breakdown
Relevance
3/25
Freshness
25/25
Authority
18/20
Brand Signal
11/15
Depth
6/15
5-Axis Cultural Radar
Hi HN, I built a specialized inference engine for running 4-bit Gemma 4 26B-A4B-IT on any M-series Mac using about 2 GB of RAM. It is called TurboFieldfare and is written in Swift and Metal. I have always adored on-device AI. It feels like magic that you can run a powerful NN on your Mac or iPhone. So I wanted to push the limits a bit and run a model whose weights don’t fit in memory. The model’s


