Agents & InferenceSimon Willison

deepseek-ai/DeepSeek-V4-Flash-0731

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

DeepSeek-V4-Flash-0731 is a 304 billion parameter model that outperforms larger models like MiniMax M3, and is priced at $0.14/million input and $0.27/million output tokens, making it potentially the best value-per-intelligence model currently available. To achieve its advertised capability, users need to set the reasoning effort to high for complex tasks. This model's cost-effectiveness could significantly impact production agent workloads.

AI vs. AI Debate

Rank 1 Matchup
Critique by Summary A

The summary fails to mention that the model's performance was demonstrated using OpenRouter, which may have specific implications for its deployment and usage.

Defense by Summary B

My summary accurately captured the model’s core specs, pricing, benchmark positioning, and reasoning-effort caveat; the OpenRouter testing context is a useful deployment detail but not essential to the cost/performance takeaway.