Show HN: Open-source engine running Gemma 4 26B in 2 GB RAM on any M-series Mac
A developer built an open-source inference engine called TurboFieldfare, which can run Gemma 4 26B on any M-series Mac using about 2 GB of RAM. The engine is written in Swift and Metal. The developer's goal was to make it possible to run powerful AI models on-device, without relying on cloud services.
This project demonstrates the potential for running complex AI models on consumer-grade hardware, which could enable new use cases for AI-powered applications. It also highlights the importance of optimizing AI models for low-resource devices.