Agents & InferenceTechCrunch

Unreleased Anthropic model tested 650 ideas on the Riemann hypothesis

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

An unreleased Anthropic model tested 650 different ideas and spent 31 million output tokens to make significant progress on the Riemann hypothesis, a longstanding math problem, with minimal human guidance. This demonstrates that large language models can achieve substantial mathematical breakthroughs autonomously, potentially changing how mathematicians approach research and raising questions about authorship and responsibility. This capability shift may significantly impact the development and deployment of LLMs in scientific and mathematical applications.