Agents & InferenceOpenAI

Legora reviewed 41 documents in minutes with GPT-6 Astra

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

A frontier model caught all four planted errors across 41 documents in minutes while lifting task performance ~40% on a financial-review workflow, showing reliable multi-document error detection at scale rather than sampled spot-checks. If those recall numbers hold on your data, you can shift human reviewers from first-pass reading to exception handling on high-stakes document pipelines—but validate the zero-miss claim on your own planted-error benchmark before trusting it without a human backstop.