A Coding-Agent Arms Race: Capy v2 Claims to Beat Claude Code, Codex, Devin and Cursor
New agent launches piled up this week, with Capy v2 claiming top DeepSWE benchmark results and a classification model claiming to be 1,600x cheaper than a frontier baseline.
The agent gold rush extended well beyond Grok Bot. @justinsunyt introduced Capy v2, billed as "the fastest, most powerful cloud coding agent," claiming to beat Claude Code, Codex, Devin, and Cursor on the DeepSWE benchmark. Benchmark leadership claims are cheap and self-reported ones cheaper still, but the framing — positioning directly against the incumbents by name — shows how crowded and competitive the coding-agent category has become.
Unlock the full briefing
Get every story in today's briefing, the full archive, and the daily AI intelligence brief.
All stories today
Full archive
Daily brief
Cancel anytime. Payments powered by Stripe.