Weekly results
Each week two AI models write the same summaries and readers blind-pick the better one. Here's who practitioners actually preferred, week by week — a result you won't find anywhere else.
Week of 2026-07-27 · in progress
GPT-5.5 vs Llama 4 Maverick — votes still landing
Week of 2026-07-20
Claude Opus 4.8 over Gemini 3.5 Flash, 100%–0% · 1 vote
Week of 2026-07-13
Llama 4 Maverick over Mistral Large, 60%–40% · 5 votes
Week of 2026-07-06
DeepSeek V3 vs Gemini 3.5 Flash — votes still landing
Week of 2026-06-29
Claude Opus 4.8 over GPT-5.5, 100%–0% · 2 votes
Week of 2026-06-22
Gemini 3.5 Flash and Llama 4 Maverick split 50%/50% · 12 votes
Week of 2026-06-15
Mistral Large over GPT-5.5, 64%–36% · 25 votes
Week of 2026-06-08
DeepSeek V3 over Claude Opus 4.8, 100%–0% · 1 vote