What it does
Run Gemma locally through a Mac app, CLI, or loopback chat server.
Read the sourceReviewed
Open-source engine running Gemma 4 26B in 2 GB RAM on Apple Silicon Macs.
Why it made the shelf
This is an open-source inference engine designed to run large language models efficiently on specific hardware, which is.
On 6 boards
From official documentation
Run Gemma locally through a Mac app, CLI, or loopback chat server.
Requires Apple Silicon, macOS 26 and Metal 4. Image support needs an M2 or newer Mac and an extra model pack.
Audio and video are unsupported. The app and CLI do not execute tools; server tool calls must be handled by the client.
Behind the tool
People, backing, growth, and ownership. Every claim links to its source.
Source review current
Public trail
20 references from 8 publishers.