Claude Code
B Tier · 7.8/10
Anthropic's terminal-based coding agent that reads your whole repo and makes real changes -- not just suggestions. **Claude Opus 5.5 became the default Opus model in 2.1.280 (2026-09-22) and Claude Sonnet 5.5 the default Sonnet in 2.1.284 (2026-09-28)**: Opus 5.5 at $4/$20 with $0.20 cache reads and fast mode at $8/$40 for up to 2.5x speed, Sonnet 5.5 at an unchanged $2/$10 with far fewer tokens per task. Both are 1M context. Fable 5.1 (2026-09-01) remains the top model and still defaults to High effort in Claude Code specifically. **Four more releases 9/29-10/02 (2.1.285-2.1.288) added Claude Mods (plugins that modify deeper behaviour) with a built-in 'You should know' side agent, `claude --desktop` to hand a session to the desktop app, an `allowedProviders` managed setting, a WebFetch kill switch, `--max-findings` on /code-review, and continuation from partial responses when the API times out mid-turn.** Also since 9/22: `/doctor prompt-audit`, `deniedModels`, dollar amounts on spend limits and `attribution: false`.
Score Breakdown
The Good and the Bad
What we like
- +Reads and understands your entire codebase before making changes -- context awareness is best-in-class for a coding agent
- +Actually executes code, runs tests, and iterates on failures autonomously -- it's a real agent, not a chatbot with code formatting
- +Multi-file refactoring is where it shines -- it can restructure projects across dozens of files coherently
- +41% developer adoption rate speaks for itself -- the output quality on complex coding tasks is genuinely excellent
What could be better
- −Terminal-only interface is a hard sell for developers who prefer visual tools -- there's no GUI at all
- −API costs can spiral fast on large tasks -- a complex refactor can easily burn through $5-10 in a single session
- −Sometimes over-edits files, making changes you didn't ask for in the name of 'improving' things
- −Learning curve is real -- you need to understand how to write good prompts and set appropriate boundaries
Pricing
Claude Pro
- ✓Included Claude Code access
- ✓Usage limits apply
- ✓Claude Sonnet 5.5 (default Sonnet since 2.1.284, 2026-09-28; Medium effort by default in Claude Code)
Claude Max (5x)
- ✓5x Pro usage
- ✓Opus 5.5 access (default Opus since 2.1.280, 2026-09-22)
- ✓Higher rate limits
Claude Max (20x)
- ✓20x Pro usage
- ✓Opus 5.5 access (default Opus since 2.1.280, 2026-09-22)
- ✓Highest rate limits
API Direct
- ✓Pay per token: Opus 5.5 $4/$20 (cache reads $0.20; fast mode $8/$40 up to 2.5x), Sonnet 5.5 $2/$10, Fable 5.1 $10/$50
- ✓Full model selection
- ✓No monthly commitment
Known Issues
- CLAUDE MODS, A 'YOU SHOULD KNOW' SIDE AGENT, --DESKTOP HAND-OFF AND PROVIDER LOCKDOWN -- FOUR RELEASES IN FOUR DAYS, 2.1.285 TO 2.1.288 (2026-09-29 to 2026-10-02, vendor changelog): **2.1.285 (9/29):** `CLAUDE_CODE_DISABLE_WEB_FETCH` turns the WebFetch tool off; **`claude --desktop`** opens the Claude desktop app on the current directory or on a session with `--continue` / `--resume <id>`; `claude plugin configure <plugin>` shows and sets plugin options (values from stdin), and `--config <server>.<key>=<value>` configures a bundled .mcpb MCP server at install time; **`allowedProviders` managed setting** limits which API providers a machine may use (Anthropic API, custom endpoint, Bedrock, Mantle, Vertex AI, Foundry, Claude Platform on AWS, Cloud gateway); `CLAUDE_CODE_NONSTREAMING_TIMEOUT_RETRIES` caps re-sends of a timed-out non-streaming fallback; fixes for forked subagents in `claude -p`, SSH plugin installs and unreadable managed-settings files. **2.1.286 (9/30):** a '2 of 5' counter on stacked permission prompts; mouse support on 'N more' list rows in fullscreen; fixes for resume losing turns after parallel tool calls, API 400s after non-text hook returns, cloud sessions with huge histories never waking, **the apps-gateway spend meter pricing 1-hour cache writes at the cheaper 5-minute rate** and under-counting input on streamed turns with server-side tools, a once-per-session fallback to the previous model of the same tier when the API refuses the configured model, Remote Control honouring an org policy that turns it off, and fallback-model retries running at standard speed when fast mode is unavailable. **2.1.287 (10/01):** **Claude Mods -- 'plugins may now modify deeper behavior'**, with a built-in **'You should know'** mod 'where a side agent watches your back and flags things you or Claude might miss' (`/plugin enable cc-plugin-you-should-know@builtin`, first-party sessions with telemetry on); an `n:<text>` filter in the agents view; `prompt_text` on the OpenTelemetry user_prompt event; URL prompts (e.g. sign-in) from MCP servers on the 2025-11-25 protocol with a `bareElicitationCapability` escape hatch; a Windows warning when denying Bash also removes PowerShell; a built-in REST-only `gh api` for self-hosted runner sessions without the GitHub CLI; fixes for fast mode in agent-owned remote sessions, Remote Control reconnects, runaway `asyncRewake` hooks and SDK heartbeats. **2.1.288 (10/02):** `$.ui.selection()` for mods; a built-in `gh api` in cloud sessions; Up-arrow recovery of a Ctrl+C-cleared prompt including pasted images; re-authentication prompts when an MCP server asks for more OAuth scope mid-call; **`--max-findings <n>|all` on /code-review** (sticky until `--max-findings default`); Ctrl+F session search and Alt+Up/Down group jumps in the agents view, rebindable; a screen-reader announcement of the new permission mode on plan approval; and **'mid-response API timeouts' no longer fail the turn -- non-interactive sessions and subagents continue from the partial response and thinking-only responses are retried**, plus fixes for 'Prompt is too long' instead of auto-compact, `--resume` dropping compaction-restored context, unsaved last responses and lost earlier thinking on resume from 2.1.286 or earlier.Source: Claude Code changelog (code.claude.com/docs/en/changelog.md, Update labels 2.1.285 'September 29, 2026', 2.1.286 'September 30, 2026', 2.1.287 'October 1, 2026', 2.1.288 'October 2, 2026') -- fetched 2026-10-02 via curl · 2026-10-02
- OPUS 5.5 AND SONNET 5.5 BECOME THE DEFAULTS, SIX DAYS APART -- PLUS FIVE RELEASES OF ENTERPRISE CONTROLS (2.1.280 on 2026-09-22 through 2.1.284 on 2026-09-28, vendor-primary via the Claude Code changelog): **2.1.280 (9/22)** added **Claude Opus 5.5 (`claude-opus-5-5`), 'now the default Opus model -- 1M context, $4/$20 per Mtok with $0.20/Mtok cache reads'**; the same release fixed auto mode retrying an action forever when a safety check declined to review it (now denied once) and backing off when the check gives no answer (turn stops after ten in a row), and stopped writes through a symlink being judged by their in-tree path. **2.1.284 (9/28)** added **Claude Sonnet 5.5 (`claude-sonnet-5-5`), 'now the default Sonnet model on the Anthropic API -- 1M context, $2/$10 per Mtok with $0.20/Mtok cache reads'**, a **'Yes, but ask again next time'** answer to auto mode's prompt before a read outside the working directories, **dollar amounts on the Claude apps gateway spend limit** in `/usage` and the status line ('$271.40 / $500.00 spent this month', with `used_usd` / `limit_usd` / `period` in `rate_limits.spend_limit`), rebindable `/effort` slider keys including `toggleUltracode`, `/rate-limit-options` in the command menu, and `/mcp reconnect all`. In between: **2.1.281 (9/23)** -- `"attribution": false` in settings.json hides all commit and PR attribution (keep the object form in shared files, older CLIs skip a file holding it); MCP URL-mode elicitation on 2026-07-28 protocol connections; `claude plugin validate` now reports `.mcp.json` entries that would be silently dropped; an auto-mode recommendation in `/insights` estimating how many prompts auto mode could have handled; gateway `assume_role` and Bedrock `guardrail` upstream options. **2.1.282 (9/24)** -- `maxProseWidth` caps prose width in wide terminals; a startup notice lists telemetry variables a project's settings ignored; `allowClaudeInChromeWithManagedMcp`; fixes for resumed sessions dropping earlier extended thinking and for every request failing with a 400 when history holds undecryptable web-search results from a third-party gateway. **2.1.283 (9/25)** -- `availableModelsMatch: "exact"` so new model releases stay blocked until listed, a **`deniedModels`** managed setting, **`/doctor prompt-audit`** (also `/checkup prompt-audit`) to audit CLAUDE.md files, skills, agents and commands for prompting patterns written for older models, `x-claude-code-prompt-id` gateway hint header, MCP/WebFetch/WebSearch outputs in the OpenTelemetry `tool.output` event, a gateway `load_test_mode`, and a `mantle` upstream for Amazon Bedrock's Mantle endpoint. Practical read: the two default flips are the story -- an unchanged Max or Pro session now runs Opus 5.5 where it ran Opus 5, and API-metered Sonnet work moves to 5.5 (Sonnet 5.5's default effort in Claude Code is Medium; Anthropic's migration note says thinking-off Sonnet users must adopt `between_tools` first).Source: Anthropic (code.claude.com/docs/en/changelog.md -- Update labels 2.1.280 'September 22, 2026' through 2.1.284 'September 28, 2026') + anthropic.com/claude-opus-5-5 (9/22) + anthropic.com/claude-sonnet-5-5 (9/28) -- fetched 2026-09-28 via curl with browser UA · 2026-09-28
- CHANGELOG PASS -- 121 RELEASES BETWEEN 2.1.132 (2026-05-06) AND 2.1.278 (2026-09-19), NEAR-DAILY, AND HERE ARE THE ONES THAT CHANGE HOW YOU USE OR PAY FOR IT (compiled 2026-09-21 from Anthropic's own changelog, vendor-primary): This page had cited v2.1.131 (May 6) as the last product batch; Claude Code has shipped a version almost every weekday since. Model events already recorded above (Fable 5 in 2.1.170 on 6/09, Sonnet 5 as default in 2.1.197 on 6/30 with its $2/$10 promo, Opus 5 as the default Opus in 2.1.219 on 7/24 -- 1M context, fast mode at $10/$50 -- and Fable 5.1 on 9/01) are the headline. **Product changes worth knowing, in date order:** (1) **Agent view, research preview (2.1.139, 5/11)** -- `claude agents` gives 'a single list of every Claude Code session -- running, blocked on you, or done'; the same release **disabled Remote Control, /schedule, claude.ai MCP connectors and notification preferences whenever an API key is set**, even with a claude.ai login present. (2) **Fast mode default moved to Opus 4.7 (2.1.142, 5/14)** and then, with Opus 5, to the current line. (3) **AGENTS.md support (2.1.277, 9/18)** -- 'in a project with no CLAUDE.md, Claude Code reads AGENTS.md instead', switchable under Project instructions in /config; not yet on Bedrock, Vertex or Foundry. That is the cross-vendor agent-instructions convention Cursor, Codex and Copilot already read, so repos no longer need a Claude-specific file. (4) **Auto mode now defaults to a server-side classifier (2.1.278, 9/19)** for Claude API and Enterprise users and on Bedrock, Vertex, Foundry and gateways -- and Anthropic says it '**does not charge for classifier overhead**'; `CLAUDE_CODE_AUTO_MODE_SERVER=0` opts out on the cloud platforms, the CLI warns when it falls back to billed classification, and /status gains an 'Auto mode server' row. For anyone who avoided auto mode because the safety classifier cost tokens, that objection is gone. (5) **Headless and SDK hardening (2.1.277, 9/18)** -- `claude -p` and Agent SDK sessions that could hang after an internal error now exit with code 1, headless resumes keep cost totals, and the background auto-title request was removed from `claude -p` runs outside an SDK or IDE. (6) **The deprecated TaskOutput tool was removed (9/18)**; Claude reads background-task output files directly. (7) **Claude Code on the web (9/18)** -- Personal and Organization environment sections for Team/Enterprise, and admins can share a personal cloud environment org-wide. (8) **VS Code extension (9/18)** shows session cost and token usage in the Account and usage dialog for API-key, Vertex, Bedrock and Foundry users, where plan limits do not apply. (9) **Security posture:** subagent results now arrive under a header marking them as subagent output, 'so text in a subagent's result cannot pass as the session's own instructions' (9/18), and prompts are scrubbed of invisible Unicode formatting characters before sending. **What the page still does not cover and the docs sidebar now lists:** Chrome, Slack, Claude Tag, mobile, keybindings and /schedule routines exist as surfaces; they are mentioned here as present, not reviewed. Check the changelog itself for the bug-fix stream -- it runs to dozens of items per release.Source: Anthropic: Claude Code changelog (code.claude.com/docs/en/changelog.md -- entries 2.1.132 'May 6, 2026' through 2.1.278 'September 19, 2026'; generated from github.com/anthropics/claude-code CHANGELOG.md) -- fetched 2026-09-21 via curl · 2026-09-19
- FABLE 5.1 IS THE NEW TOP MODEL IN CLAUDE CODE, AND IT DEFAULTS TO HIGH EFFORT HERE SPECIFICALLY (2026-09-01, vendor-primary): Anthropic shipped **Claude Fable 5.1**, and the detail that matters most for Claude Code users sits in a parenthetical on the launch post: '**Fable 5.1 defaults to High effort in Claude Code, and to Medium in Claude Cowork and on Claude.ai.**' **Claude Code is the only surface that defaults to High.** That is a defensible choice for a terminal agent doing long-running work, but it means the same prompt costs more here than in Cowork or on the web, and Anthropic's own framing is that '**when set to Low or Medium effort, Fable 5.1 achieves results similar to or better than Fable 5's at a much lower cost**' -- so the cheapest correct configuration for routine work is explicitly not the default you are given. **THE COST STORY IS CACHE-SHAPED, WHICH SUITS THIS TOOL UNUSUALLY WELL.** Fable 5.1's base rate is unchanged from Fable 5 at **$10/MTok input and $50/MTok output**; the entire quoted 25% saving (up to approximately 45% on 'highly agentic work') comes from **cache reads falling from $1/MTok to $0.25/MTok**. Claude Code re-reads a large, stable repo context on nearly every turn, which is close to the best possible case for a cache-read discount -- so of all the Anthropic surfaces, this is the one where the upper end of that range is plausible rather than promotional. **CODING EVIDENCE, VENDOR-PUBLISHED:** Terminal-Bench 4.0 **55.8%** (vs 42.0% for Fable 5, 52.3% for Opus 5, 37.3% for GPT-5.6 Sol); CursorBench 3.2.0 **73.4%** vs 70.5%; Terminal-Bench-Science 0.1 **52.6%** vs 24.7%. Anthropic says Fable 5.1 'avoids shortcuts that result in poorer-quality work' and fixes root causes, and cites Millennium finding a one-in-a-million crash that had gone unexplained for four to five years and that Fable 5 had missed. **Note the caveat Anthropic itself publishes: these runs had production safeguards enabled, and tasks where safeguards intervened scored zero**, which Anthropic says likely understates the numbers. Red Hat, quoted in the launch post, says Fable 5.1 'correctly identified the root cause of every broken build we tested, across all the effort levels' using Claude Code.Source: Anthropic (anthropic.com/claude-fable-and-mythos-5-1 -- root path, not /news/) + platform.claude.com/docs/en/about-claude/pricing.md -- fetched 2026-09-05 via curl with browser UA · 2026-09-01
- CLAUDE CODE PROMPTS CAN NOW BE GATED BY YOUR OWN SECURITY SERVER -- INFERENCE HOOKS BETA (2026-08-05, vendor-primary, Enterprise only): Anthropic shipped **inference hooks** in beta for **Claude Enterprise organizations**, and Claude Code is explicitly in scope. Point Claude at your organization's AI security server and '**each governed prompt across claude.ai, Cowork, and Claude Code is held for the server's allow or deny verdict before inference proceeds**' -- a synchronous pre-inference block, not retrospective logging. Requests to your server are **signed**, **failure handling is configurable**, and **denials land in the compliance Activity Feed**. **WHY THIS MATTERS SPECIFICALLY FOR CLAUDE CODE:** it is the surface where a coding agent sees source, secrets and infrastructure, and until now enterprise controls mostly stopped at API-level logging and data-retention terms. This is the first Anthropic control that can refuse a Claude Code prompt before the model ever sees it. **THE SETTING THAT DECIDES IF THIS IS SAFE TO TURN ON is the configurable failure handling** -- fail-closed makes your security server a hard dependency for every developer's Claude Code session, fail-open makes the control advisory. Decide that deliberately before rollout. Beta, Enterprise-tier only, and it requires you to actually run the security serverSource: Anthropic Claude Platform release notes (platform.claude.com/docs/en/release-notes/api, August 5 2026 entry, fetched 2026-08-06) · 2026-08-05
- MCP SPEC 2026-07-28 IS OUT AND IT IS A BREAKING RELEASE (2026-07-28, spec-primary): the Model Context Protocol shipped its `2026-07-28` revision as a **stable release**, and it materially changes how MCP servers are built and operated. The headline, verbatim: '**The highlight of this release is a stateless protocol core - MCP is transforming from a bidirectional stateful protocol into a request/response stateless protocol.**' Practical consequence: '**Any request can now land on any server instance behind a plain round-robin load balancer without needing shared storage**' -- MCP servers become ordinary horizontally-scalable web services. Other changes: **Multi Round-Trip Requests (MRTR)** replace server-initiated streams so mid-call confirmations no longer need an open bidirectional connection; **header-based routing** moves method and tool names into the `Mcp-Method` and `Mcp-Name` HTTP headers so gateways can route and authorize without parsing JSON bodies; **authorization hardening** adds RFC 9207 issuer validation and moves from Dynamic Client Registration to Client ID Metadata Documents (CIMD); and **Tasks and MCP Apps graduate out of experimental** into a formal extensions framework. BREAKING BITS: the `initialize`/`initialized` handshake and session IDs are removed, and the 12-month deprecation policy now starts the clock on Roots, Sampling and Logging. '**The TypeScript, Python, Go, and C# SDKs are updated to match, with detailed migration notes for the breaking bits.**' Anthropic is rolling support across the Claude apps, Claude Code and the Platform API. If you maintain an MCP server for Claude Code, this is the upgrade to plan forSource: MCP blog (blog.modelcontextprotocol.io/posts/2026-07-28/, fetched 2026-08-03), MCP spec (modelcontextprotocol.io/specification/2026-07-28) · 2026-07-28
- ALIBABA WORKPLACE BAN NOW IN EFFECT (took effect 2026-07-10 per the announced schedule -- scope and date unchanged in all reporting through 7/10, no delay reported, no formal Anthropic statement): Alibaba banned Claude Code company-wide over an alleged 'backdoor' -- since v2.1.91 (April 2) Claude Code reportedly checked for Asia/Shanghai and Asia/Urumqi timezones plus a 147-entry list of Chinese proxy/cloud/AI-lab URLs, inserting markers into prompts; Claude Code is now on Alibaba's 'high-risk software' list. Anthropic's only response remains a Claude Code engineer's statement that it was an anti-distillation / reseller-abuse experiment from March, rolled back as of July 1. Alibaba employees are directed to the in-house Qoder tool. Background: Anthropic's June 10 letter accused Qwen operators of ~25,000 fraudulent accounts and 28.8M distillation conversations. Relevant if you operate in China-adjacent environments or are sensitive to telemetry behavior in CLI toolsSource: Reuters (2026-07-03), CNBC (2026-07-06), SCMP, The Decoder · 2026-07-10
- MODEL UPDATE (2026-06-30): **Claude Sonnet 5** is now available in Claude Code (and is the new default on Free/Pro). Anthropic bills it as 'the most agentic Sonnet yet,' approaching Opus 4.8 quality at lower cost -- a meaningful default upgrade for everyday coding sessions, at $2/$10 per 1M. **PRICING UPDATE (confirmed 2026-08-31 on Anthropic's pricing doc): the $2/$10 rate was announced as introductory pricing through Aug 31 with a rise to $3/$15 on Sept 1 -- that increase was cancelled and $2/$10 is now the standard price.** For a Claude Code user this is the single biggest cost fact on the page, because Sonnet 5 is the default model on Free and Pro: the per-token cost of your default coding model is now permanently a third lower than the launch announcement scheduled, and Sonnet 5 permanently undercuts Sonnet 4.6 and 4.5, which both stay at $3/$15. Opus 4.8 remains the top-end option on Max for the hardest agentic work; the `xhigh` effort level is still the recommendation for coding. Note the new Sonnet-5 tokenizer inflates input token counts ~1.0-1.35x, so watch session cost on large repos. Separately, after a 19-day export-control suspension, Fable 5 returned to Claude Code on 2026-07-01 (see claude.ts).Source: Anthropic news (anthropic.com/news/claude-sonnet-5), Anthropic (anthropic.com/news/redeploying-fable-5) · 2026-06-30
- BILLING CHANGE PAUSED -- NOTHING CHANGES FOR NOW (status as of 2026-06-18): Anthropic had announced (~2026-05-13/14) that, effective 2026-06-15, programmatic Claude usage -- the Agent SDK, `claude -p` non-interactive mode, Claude Code GitHub Actions, and third-party apps built on the Agent SDK -- would move OFF normal Pro/Max subscription limits onto a SEPARATE metered credit pool (Pro $20/mo, Max 5x $100, Max 20x $200) billed at API rates beyond that. Commentary pegged the effective increase for heavy CI / Agent-SDK users at 12x-175x. **Anthropic reversed course and PAUSED the overhaul just before the June 15 go-live, telling developers 'Nothing changes for now.'** As of today, programmatic Claude Code usage (Agent SDK, `claude -p`, GitHub Actions) still draws on your regular Pro/Max subscription limits -- there is NO separate credit pool in effect (Anthropic's own cost docs at code.claude.com/docs/en/costs describe normal subscription/API billing, not a programmatic credit pool). Reporting attributes the reversal to the OpenAI price war, Anthropic's pending IPO filing, and government pressure over model access. Treat the credit-pool plan as shelved-but-not-dead and re-check before architecting around it. NOTE: this is DISTINCT from the 2026-06-15 Sonnet 4 / Opus 4 MODEL retirement, which DID take effect (see claude.ts).Source: The Decoder (the-decoder.com/anthropic-backs-off-unpopular-billing-overhaul-as-price-war-with-openai-looms/), Axios (2026-05-14), Anthropic Claude Code cost docs (code.claude.com/docs/en/costs) -- earlier announcement: The New Stack, VentureBeat (2026-05-13/14) · 2026-06-18
- PRODUCT BATCH (2026-05-06 Code with Claude SF keynote, versions 2.1.129 + 2.1.131): (1) **Code Review GA** -- 'used by every team at Anthropic'; substantive review comments rose 16% -> 54% of PRs; PRs >1000 lines: 84% generated findings, avg 7.5 issues per PR. Vendor-primary blog post at claude.com/blog/code-review. (2) **Remote Agents** -- launch and monitor Claude Code sessions from your phone; control your laptop remotely. SHIPPED. (3) **CI Auto-Fix** -- automatic fixes generated against PRs in CI. SHIPPED. (4) **Routines** -- saved Claude Code config (prompt + repos + connectors) running on Anthropic cloud as async automations; 'wake up to PRs ready to merge'. SHIPPED (expanded from earlier April rollout). (5) **Security Reviews** public beta for Enterprise (per April 30 post). PLUS rate-limit doubling from concurrent SpaceX compute deal -- Pro / Max / Team / seat-based Enterprise see 2x Claude Code 5-hour limits and removed peak-hours reduction (Pro / Max). Customers cited on stage: Shopify, Mercado Libre (~23k engineers targeting '90% autonomous coding by Q3'). Plus speakers from GitHub, Netflix, Datadog, VercelSource: Anthropic Code Review blog (claude.com/blog/code-review), Claude Code release notes 2.1.129 + 2.1.131 (docs.claude.com/en/release-notes/overview.md), Simon Willison live blog (simonwillison.net/2026/May/6/code-w-claude-2026/), InfoQ, TheNewStack, VentureBeat · 2026-05-06
- PRICING SCARE (2026-04-21 -> 2026-04-22, RESOLVED): Anthropic briefly removed Claude Code from the $20 Pro plan on logged-out pricing pages on 2026-04-21. Head of growth Amol Avasare framed it as a 2% A/B test on new prosumer signups; existing Pro/Max subscribers were never affected. Reversed within 24 hours after backlash -- as of 2026-04-22 the Claude Code checkbox is restored on claude.com/pricing. Anthropic statement: 'a mistake that the logged-out landing page and docs were updated for this test.' Pricing risk on agentic-coding tools is real even when today's price holds; if you're cost-sensitive on Pro, watch the pricing page periodicallySource: The Register (2026-04-22), Simon Willison · 2026-04-22
- Claude Opus 4.7 (default backing model) brings three Claude-Code-relevant features documented on Anthropic's What's New page: (1) new `xhigh` effort level recommended specifically for coding + agentic work, (2) task budgets (beta header `task-budgets-2026-03-13`) -- give Claude an advisory token budget across the full agentic loop and the model self-paces against a running countdown, (3) high-resolution image support up to 2576px / 3.75MP with 1:1 pixel-coordinate mapping, big upgrade for screenshot-driven debugging. Breaking changes for direct API users: extended thinking budgets removed (use adaptive thinking), sampling parameters (temperature/top_p/top_k) removed, thinking content omitted from response by default (set display=summarized to restore)Source: Anthropic: What's new in Claude Opus 4.7 (platform.claude.com/docs/en/about-claude/models/whats-new-claude-4-7) · 2026-04
- 2026-04-18 added the `/usage` command -- shows a usage-driver breakdown for the current session, flags cache-miss patterns, and makes it easier to catch runaway token consumption before it ends up on the bill. If your Claude Code sessions surprise you with cost, this is now the first diagnostic to runSource: Anthropic Claude Code release notes · 2026-04
- Large file edits occasionally produce malformed output, requiring manual cleanup of partial replacementsSource: GitHub Issues · 2026-03
- Token consumption on large repos can exceed expectations -- users report $20+ sessions on complex multi-file tasksSource: Reddit r/ClaudeAI · 2026-02
Best for
Experienced developers who are comfortable in the terminal and want an AI that can do real, multi-file engineering work autonomously. Especially strong for refactoring, debugging, and building features across complex codebases.
Not for
Beginners who want a visual coding assistant, or anyone who needs predictable monthly costs. If you're looking for autocomplete-style help, Copilot or Cursor are better fits.
Our Verdict
Claude Code is the most capable agentic coding tool available right now. The ability to read entire codebases, execute code, run tests, and iterate on results puts it in a different category than autocomplete-style assistants. The output quality on complex tasks is outstanding. But it's firmly a power-user tool -- the CLI-only interface, unpredictable costs, and learning curve mean it's not for everyone. If you're a developer who thinks in terms of terminal workflows and you're working on non-trivial projects, Claude Code is worth every penny. Just keep an eye on your API bill.
Sources
- Claude Code changelog 2.1.285-2.1.288 (2026-09-29 to 2026-10-02) -- Claude Mods, You should know, --desktop, allowedProviders, --max-findings, partial-response continuation (accessed 2026-10-02)
- Anthropic: Claude Code changelog -- 2.1.280 (Sept 22, 2026: Opus 5.5 default) through 2.1.284 (Sept 28, 2026: Sonnet 5.5 default, spend-limit dollars, prompt-audit, deniedModels) (accessed 2026-09-28)
- Anthropic: Introducing Claude Opus 5.5 (2026-09-22) -- $4/$20, $0.20 cache reads, fast mode $8/$40 in Claude Code and the Claude Platform (accessed 2026-09-28)
- Anthropic: Introducing Claude Sonnet 5.5 (2026-09-28) -- $2/$10, Medium effort default in Claude Code, between_tools migration note (accessed 2026-09-28)
- Anthropic: Claude Code changelog -- 2.1.132 (May 6, 2026) through 2.1.278 (Sept 19, 2026): agent view, AGENTS.md support, server-side auto mode classifier, TaskOutput removal (accessed 2026-09-21)
- Anthropic: Claude Fable 5.1 launch -- defaults to High effort in Claude Code, Medium elsewhere (2026-09-01) (accessed 2026-09-05)
- Anthropic pricing docs: Fable 5.1 cache hits $0.25/MTok (0.025x); base $10/$50 unchanged (verified 2026-09-05) (accessed 2026-09-05)
- Anthropic pricing docs: Sonnet 5 $2/$10 is now the standard price -- the Sept 1 rise to $3/$15 will not occur (verified 2026-08-31) (accessed 2026-08-31)
- Anthropic: Claude Platform release notes (2026-08-05 -- inference hooks beta covering claude.ai, Cowork and Claude Code) (accessed 2026-08-06)
- Reuters (via TradingView syndication): Alibaba to ban Claude Code in workplace over alleged backdoor risks (accessed 2026-07-05)
- Anthropic: Introducing Claude Sonnet 5 (2026-06-30, available in Claude Code) (accessed 2026-07-04)
- The Decoder: Anthropic backs off unpopular billing overhaul as price war with OpenAI looms (PAUSED before 6/15) (accessed 2026-06-18)
- Anthropic: Claude Code cost management docs (no programmatic credit pool in effect) (accessed 2026-06-18)
- The New Stack: Anthropic Agent SDK separate credit pools (original 2026-06-15 announcement, later paused) (accessed 2026-05-26)
- Anthropic: Claude Code Review (2026-05-06 keynote) (accessed 2026-05-06)
- Claude Code release notes 2.1.131 (accessed 2026-05-06)
- Simon Willison: Code with Claude 2026 live blog (accessed 2026-05-06)
- The Register: Anthropic tests Claude Code Pro removal (2026-04-22) (accessed 2026-04-25)
- Simon Willison: Is Claude Code going to cost $100/month? Probably not (accessed 2026-04-25)
- Anthropic: What's new in Claude Opus 4.7 (accessed 2026-04-22)
- Anthropic documentation (accessed 2026-04-22)
- Reddit r/ClaudeAI (accessed 2026-04-22)
- GitHub community discussions (accessed 2026-03-31)
Explore more Claude Code rankings
Deeper leaderboards, benchmarks, task-specific tier lists, and status/pricing pages for Claude Code.
The Tier List Tuesday
Weekly newsletter: tier movers, new entrants, and the VS of the week. Built from our daily AI-tool sweeps. No spam, unsubscribe anytime.
Alternatives to Claude Code
GitHub Copilot
**GPT-6.1 Sol landed in Copilot on OpenAI's launch day (2026-09-29, Pro+/Max/Business/Enterprise), the fifth day-one frontier add in eight days, and on 2026-10-02 the four-model retirement slate executed on schedule: Gemini 3.5 Flash, Gemini 3.6 Flash, Kimi K2.7 Code and Claude Opus 4.7 are gone from every Copilot surface.** Same week: HydraFusion's multi-model orchestration reached VS Code 1.140 and the Copilot app (9/30), **computer use** let Copilot CLI and the Copilot app drive desktop applications on macOS and Windows in public preview (10/01), dynamic workflows defined in code arrived for CLI, app and SDK (10/01), the September VS Code roundup added automations, agent merge and Dev Container sessions (10/01), and Copilot code review gained REST/GraphQL request support with Balanced as the default effort level (10/02). Still ahead: the six-model slate on 10/19 and the 10/22 default policy for new features for Business and Enterprise. The 9/22 and 9/28 adds (Opus 5.5, GPT-6 Sol/Luna, Sonnet 5.5) all stand, metered at provider list pricing.
Cursor
**Cursor launched two bots for the last mile of shipping code on 2026-09-23 (Teams and Enterprise): Rollouts, which attaches a monitor to every pull request and watches the change as it deploys per environment -- verified healthy, regression detected, or inconclusive -- and Security Review, which posts one comment per PR reporting exploitable bugs with severity, attack path and a proposed fix.** Rollouts is the Cursor rebuild of Firetiger's Change Monitors on the Bot Development Kit; it can open a revert PR or hand a regression to a cloud agent but does not merge or roll back on its own. Models: Anthropic's 9/28 Sonnet 5.5 post quotes SpaceXAI's Sualeh Asif on CursorBench 4.0 -- Sonnet 5.5 at 55.5% 'second only to Opus 5.5' (57.8%); Grok 4.7 and its Cursor-only Fast variant arrived 9/21. **Cursor Projects (9/10, beta)** and the 11/12 OpenAI shutoff watch continue.
Devin Desktop (formerly Windsurf)
Windsurf is now **Devin Desktop** -- Cognition retired the Windsurf brand via OTA update on June 2, 2026. Same editor, plans, pricing, settings, and extensions; the bundled agent is now 'Devin Local' and Devin Cloud agent access starts on the $20 Pro plan. Agent Command Center, Spaces, and Devin Review all carry over
Tabnine
AI code completion that runs locally and keeps your code private -- the enterprise-friendly alternative to Copilot
Lovable
Describe the app you want in plain English and watch it build itself. **New this week: apps built on Lovable can add live voice conversations through GPT Live (2026-10-01, TanStack Start apps, uses credits), Enterprise workspaces can pin AI processing to EU providers (10/01, costs more credits), Business and Enterprise API keys can be IP-allowlisted (10/01), Lovable can rotate a Cloud database password from chat (9/30), and the Lovable MCP server works inside Muse by Meta (9/29).** **Chat with Lovable (2026-09-21)** adds a free daily chat allowance for thinking, asking and planning without changing code, plus voice conversations and a Telegram surface; apps built on Lovable can now call **Claude Opus 5.5 (9/23)** and **GPT-6 Sol and Luna (9/25)** for their own AI features, and the **Lovable API (9/17)** automates workspaces. $500M ARR (June 2026) and ~1M new projects a week
Devin
**Cognition crossed a one billion dollar run rate on 2026-09-25** -- '$1B in annualized revenue run rate', less than two years after Devin's GA, naming GE Aerospace, Rivian, Rohlik and Exa -- and opened a Sao Paulo hub on 9/22 with Itau (over 75% of its technology teams on Devin, 300,000+ repositories documented, .NET-to-Java migrations 6x faster), Nubank (a multi-million-line monolith migration 'from years to weeks' at over 20x lower cost), Santander, Natura and EBANX as customers. **Devin Fusion (9/11)** -- the lead/sidekick two-model harness measured at 23-46% lower cost per task -- and the 9/15 AWS multi-year agreement stand.
Replit
Cloud IDE with an AI agent that builds full apps from prompts. **Free Mode (2026-08-18)**, powered by OpenAI's GPT-5.6 Luna, lets Core and Pro subscribers chat and run everyday tasks without spending credits -- up to 30 hours of chat a month on Core -- and **Intelligent Model Routing (2026-08-26)** now picks the model per task at what Replit measured as 65% lower cost than the old Max Mode. Replit acquired the charting startup Atta on 2026-09-25. Agent 4 (May 2026) added parallel task execution
Codex (OpenAI)
**DevDay 2026 (2026-09-29) was mostly a Codex release: GPT-6.1 Sol in Codex and the API at $2/$10 per 1M with cached input cut to $0.10, Codex Cloud (reusable cloud environments, tasks that keep running while your laptop sleeps, on Plus and up), a refreshed CLI with voice steering and an /agents view, Code Review in the desktop app with automatic cloud first passes, Codex Security Cloud with Daybreak Blue models and no separate Daybreak application, and the Agents API gaining computer use** -- plus GPT-6 Astra Ultrafast (300 tokens per second in Codex) on Pro 500 and Enterprise, with GPT-6.1 Sol Ultrafast 'coming soon'. GPT-6.1 Sol nearly matches GPT-6 Astra on DeepSWE at one-fifth of the token price and also landed in GitHub Copilot the same day. Still in place: GPT-6 Sol and Luna (9/22), the Agents API public beta (9/10) and the 8/31 retirement of GPT-5.4 from Codex. On 10/01 OpenAI scheduled GPT-5.3-Codex for API removal on 2027-04-01.
Google Antigravity
Google's agent-first AI IDE -- deploys up to 5 autonomous coding agents in parallel on a VS Code fork. Antigravity 2.0 (I/O 2026) is the runtime substrate for Gemini Spark, and the Antigravity CLI is now the official successor to Gemini CLI, which stopped serving consumer tiers on 2026-06-18
Codestral 2 (Mistral)
Mistral's dedicated code model -- Codestral 2 (launched 2026-04-08) relicensed under Apache 2.0, removing the commercial-use restrictions of the original. 22B dense, strong FIM (fill-in-middle), available via Mistral API + Hugging Face
Roblox Assistant
Roblox Studio's agentic AI that plans, builds, and playtests games. Planning Mode (2026-04-16) + Mesh Generation + Procedural Models brings 3D-native creation to 70M+ daily creators