Which AI writes the better take? You decide — blind.

Two top models go head-to-head on today's AI news. Pick the sharper summary without seeing the names — the crowd's verdict builds the leaderboard.

Agents & InferenceHugging Face

Introducing Mellum2: A 12B Mixture-of-Experts Model by JetBrains

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

JetBrains has introduced Mellum2, a 12B Mixture-of-Experts model optimized for efficient text-and-code tasks, offering faster inference and lower latency for software engineering workloads. The open model excels in routing, retrieval-augmented generation (RAG) pipelines, and sub-agent tasks while being deployable in private environments. Designed for specialized use rather than replacing larger models, Mellum2 aims to enhance AI system efficiency and cost-effectiveness.

Browse editions · 99 days
Agents & InferenceTechCrunch

Nvidia chases $200B CPU market with AI agent PCs from Microsoft, Dell, and HP

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

Nvidia opened Computex with the RTX Spark, a new 1-petaflop PC CPU it calls a superchip, designed to run AI agents and local large language models securely. Windows PCs powered by the chip will arrive this fall from ASUS, Dell, HP, Lenovo, Microsoft Surface, and MSI, with Acer and Gigabyte to follow, backed by more than 100 software makers including Adobe and Riot Games. The move is part of CEO Jensen Huang's pursuit of a $200 billion CPU market, envisioning PCs that complete tasks on command rather than requiring traditional pointing, clicking, and typing.

Agents & InferenceSimon Willison

llm-anthropic 0.25.1

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

The llm-anthropic plugin has been updated to version 0.25.1, with the new release used to generate pelican drawings tied to notes on Opus 4.8.

Agents & InferenceHugging Face

Beyond LLMs: Why Scalable Enterprise AI Adoption Depends on Agent Logic

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

Scalable enterprise AI adoption requires more than just large language models (LLMs), relying instead on agent logic to ensure quality, cost-effectiveness, and user trust. Agent logic, which includes tools like knowledge graphs and algorithms, helps steer LLMs to better align with dynamic enterprise workflows while reducing errors and inefficiencies. IBM's watsonx Code Assistant for Z demonstrates this approach by using agent logic to enhance mainframe application development.

Agents & InferenceTechCrunch

This AI weather startup is out-forecasting government agencies

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

A startup called WindBorne Systems has developed an AI weather forecasting tool, WeatherMesh-6, which outperforms predictions by leading government agencies like the European Centre for Medium-Range Weather Forecasts. The system offers more frequent updates, higher resolution, and greater accuracy, leveraging data from hundreds of weather balloons launched globally. This advancement highlights the growing potential of AI in improving weather predictions over traditional methods.

Agents & InferenceSimon Willison

Pasted File Editor

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

A new prototype called Pasted File Editor replicates a feature found in Claude's apps, where large blocks of pasted text are automatically converted into file attachments. Built using Codex desktop, the tool also lets users open files directly—displaying images as thumbnails—or drag files onto the text area.

See who's winning the model face-off

Tomorrow's blind matchup and the running leaderboard — one email a day.