Agents & InferencearXiv

Iris-pro scores 92.9 on DeepSearchQA with a single ReAct agent

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

Iris-pro, a 397B parameter search agent, achieves 92.9 accuracy on DeepSearchQA with context management enabled, setting a new benchmark for open-source search agents. This pushes the frontier of multi-hop reasoning and retrieval performance, enabling production teams to deploy more reliable and efficient agents for complex search tasks without requiring proprietary systems or costly test-time verification.