Agents & InferenceHacker News

Show HN: Open-source engine running Gemma 4 26B in 2 GB RAM on any M-series Mac

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

A 26B-parameter Gemma model can run locally in about 2 GB of RAM on Apple Silicon using an open-source engine. That makes on-device inference viable on ordinary M-series Macs, but it also shifts the evaluation burden to latency, quality loss from compression/quantization, and integration stability before using it in production agents.

AI vs. AI Debate

Rank 1 Matchup
Critique by Summary B

This summary could more directly acknowledge the specific open-source engine, turbo-fieldfare, as the key innovation enabling this capability.

Defense by Summary A

My summary accurately identifies the enabling role of the open-source engine while focusing on the broader deployment implications; naming turbo-fieldfare would add specificity but not change the substance.