2026-07-21
Claude Code holds the top spot for me, and the move to Opus 4.8 and the Claude 5 family this month only widened the gap on the work I care about most, which is letting an agent own a full pull request from failing test to merged branch. It reads a large codebase, keeps the thread across dozens of files, and lands changes that actually compile. That is why it still leads on code quality and agentic capability. The bigger story this refresh is OpenAI Codex. With the GPT-5.6 Sol family reaching general availability and Sol Ultra heading into the Codex client, its async cloud agent got noticeably sharper at planning multi-step tasks and reviewing its own output. I moved Codex up to second because that GA milestone is a real event, not a promise, and its agentic score reflects it. Cursor stays a hair behind at third, and it remains my pick for anyone who wants AI woven into every layer of an editor with the freedom to swap models mid-task. Its IDE experience is still the smoothest here. Below the top three, Google Antigravity keeps punching above its price with strong value, and Jules remains a capable cloud teammate for batch work. Copilot is the safe institutional choice with the best editor integration, though the subscription math weighs on its value score. My honest read: pick Claude Code if you live in the terminal, Codex if you assign work and review later, Cursor if the editor is home base. All three are genuinely excellent in July 2026, and you can build a serious workflow on any of them.
Claude Code stays #1 on Opus 4.8
The Claude 5 upgrade this month sharpened full pull-request ownership, from reading a large codebase to landing changes that compile on the first run.
Codex climbs to #2 on GPT-5.6 GA
OpenAI's Sol family reached general availability with Sol Ultra entering the Codex client, lifting its async planning and self-review enough to earn the number two seat.
Cursor owns the IDE experience
For developers who want AI at every layer of the editor and freedom to switch models mid-task, Cursor remains the smoothest surface at third.
Value picks hold their ground
Google Antigravity and Aider keep delivering the strongest value scores, a reminder that budget-friendly agents can anchor a real workflow.