Which AI writes the better take? You decide — blind.

Two top models go head-to-head on today's AI news. Pick the sharper summary without seeing the names — the crowd's verdict builds the leaderboard.

Agents & InferenceHugging Face

Five labs, five minds: building a multi-model finance drama on small models

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

Five labs collaborated to create a multi-model finance simulation where different small AI models interact as agents in an emergent economy, each representing distinct financial behaviors. Players act as shadow financiers, manipulating the market through tips, alliances, and trades while avoiding detection by a magistrate. The system relies on heterogeneous models from various labs, ensuring diverse decision-making and market dynamics.

Browse editions · 58 days
Agents & InferenceTechCrunch

Google will pay SpaceX $920M per month for compute

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

Google will pay SpaceX $920 million per month from October 2026 through June 2029 for access to around 110,000 NVIDIA GPUs and related compute resources. The deal, similar to SpaceX's earlier agreement with Anthropic, allows Google to expand its AI capacity amid surging demand for products like Gemini Enterprise. Both companies can terminate the agreement with 90 days' notice after December 2026, and Google's access will ramp up by September 2026.

Agents & InferenceSimon Willison

datasette-agent-micropython 0.1a0

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

Simon Willison has released an alpha version of Datasette Agent, designed to safely generate and execute Python code. Early tests show promise, with GPT-5.5 failing to bypass the sandbox security measures. The project is part of Willison's ongoing work, supported by sponsors who receive monthly updates on LLM advancements.

Agents & InferenceHugging Face

Designing the hf CLI as an agent-optimized way to work with the Hub

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

The Hugging Face CLI (hf CLI) has been redesigned to optimize interactions for both human users and AI coding agents like Claude Code and Codex. It now detects agent usage via environment variables, tailoring outputs to be compact and structured for agents while maintaining rich formatting for humans. Early data shows significant agent adoption, with Claude Code and Codex leading in user numbers and request volume on the Hub.

Agents & InferenceTechCrunch

The token bill comes due: Inside the industry scramble to manage AI’s runaway costs

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

Companies across the tech industry are confronting soaring AI costs as more autonomous agents and aggressive adoption drive token consumption far beyond budgets, with examples like Uber exhausting its 2026 AI coding budget by April and one firm reportedly facing a $500 million Claude bill. In response, a market is emerging to help track and control AI spending, including the Linux Foundation's newly announced Tokenomics Foundation, which aims to bring cost discipline to AI tokens similar to what FinOps did for cloud spending. Studies suggest heavy AI use boosts developer productivity but at steeply higher costs and with more bugs and rewrites, prompting companies to impose token limits and seek better visibility and ROI.

Agents & InferenceSimon Willison

micropython-wasm 0.1a2

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

A new alpha release of micropython-wasm, version 0.1a2, has been published with the addition of a command-line interface. The CLI was inspired by efforts to demonstrate the project's functionality in a hands-on "Try it yourself" section.

See who's winning the model face-off

Tomorrow's blind matchup and the running leaderboard — one email a day.