Everyday coding and debugging
Features, debugging, bug fixes, reviews, small refactors — the work you do all day.
- Fix a failing test or a reported bug
- Add an endpoint, a form, a CLI option
- Review a pull request
- Small refactor in a few files
Claude Opus 5.5(default)·effort medium(default)
- Agent
- Claude Code
- Subscription
- Claude Pro ($20) or Max ($100 / $200)
- Provider
- Anthropic
GPT-6 Sol·effort high
Codex · ChatGPT Plus ($20) or Pro ($100 / $200)
Codex on ChatGPT Plus: DeepSWE 65.3% at $0.64 per task — +8.6 points over medium for +$0.26; Plus allows about 15–150 Sol messages per 5 hours.
GLM-5.3·effort max(default)
Claude Code or ZCode · GLM Coding Plan ($18 / $80 / $168)
Budget: GLM Coding Plan Lite ($18/month, 2,000 credits per 5 hours) inside Claude Code; DeepSWE 69.0% at max.
How to work
- No plan needed when you can describe the change in one sentence; otherwise type
/plan <task>— more reliable thanShift+Tab, which now starts from auto mode. - Stay on medium (the default). For a stubborn bug, raise it with
/effort high, or addultrathinkto a single prompt. - Give the agent a check it can run (tests, build, lint) and ask for the command and its output, not a claim. After two failed corrections,
/rewindor/clearand re-prompt with what you learned. - Before the pull request:
/code-reviewlooks for bugs in a fresh subagent;/simplifycleans up the diff but does not look for bugs. - In Codex:
/model→ GPT-6 Sol, effort high; structure the prompt as Goal, Context, Constraints, Done when.
Why
- Best quality per unit of quota: 51.2 on the AA Intelligence Index for $1.34 per task, ahead of GPT-6 Astra at high (50.9 for $1.73).
- Anthropic's own SWE-bench Pro subset: 92.8% solved at $0.22 per solved task, versus 92.3% at $1.19 for Fable 5.1 at its default.
- It is Claude Code's default model and effort since 2026-09-22. Step up to high for a hard bug: +4 points on Terminal-Bench 4.0 (52.5% → 56.6%) for about 27% more quota.
Sources: Artificial Analysis model leaderboard · Optimizing for cost and intelligence · Model configuration · DeepSWE v1.1 leaderboard · Introducing GPT-6 Sol and Luna · Codex pricing · GLM Coding Plan · Claude Code best practices · Claude Code commands · Claude Code permission modes