Agents & InferencearXiv

The Menu Is an Execution Prior: State-Path Tool Menus for Online Agents

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

Language models can now successfully execute multi-step tasks with a tool menu limited to 32 tools, up from 128, achieving a higher online success rate of 0.898 compared to 0.737 previously, without requiring changes to the underlying agent, enabling more efficient and scalable deployment of LLMs in production environments.