Gemma 4 26B-A4B-IT
AssessPlatforms
A 26B-parameter instruction-tuned Gemma model variant used for local inference.
Why it's here
Placed in Assess: 1 article(s) of evidence from 1 source(s), led by open-source activity, with 1 in the last 30 days. Confidence 24%. Low accumulated evidence, so it defaults conservatively pending more signal.
Evidence (1)
- 7Hacker News·7/29/2026open_sourceOpen-source Swift engine runs Gemma 4 26B on 2 GB RAM on M-series Macs
A Hacker News poster introduced TurboFieldfare, an open-source inference engine written in Swift and Metal that can run the 4-bit Gemma 4 26B-A4B-IT model on M-series Macs using about 2 GB of RAM. The system streams routed experts from SSD while keeping shared weights and the KV cache in memory, and the author reports 5–6 tok/s on an M2 MacBook Air and 31–35 tok/s on an M5 MacBook Pro.