AI Tool Tier List · Updated Daily

AI tools, ranked S to F.

Every tool tested, scored, and placed in its tier. We report the bugs, show the pricing traps, and tell you what actually works. No sponsored placements.

S
9.0+
A
8.0
B
7.0
C
6.0
D
5.0
F
<5
142
Tools ranked
23
Categories
10,011+
Comparisons
Daily
Updates

01 The Tier List

Top picks across 5 categories.

142 tools ranked
Last updated: Aug 6, 2026

03 Latest Reviews

Recently reviewed and updated.

See trending →
Claude (Anthropic) logo

Claude (Anthropic)

Anthropic's flagship LLM family. **Claude Opus 5 launched 2026-07-24** and is now the default model on Claude Max and the strongest model on Claude Pro -- same $5/$25 per 1M as Opus 4.8, but Anthropic says it lands within 0.5% of Fable 5 on CursorBench at half the cost. Sonnet 5 (June 30) stays the default on Free/Pro at $2/$10 per 1M (intro through Aug 31, then $3/$15), and Fable 5 -- back globally since July 1 after a 19-day export-control suspension -- remains the top of the range at $10/$50

A
8.5/10
Free tierFrom $0
Best writing quality of any LLM -- Opus ...1M token context window for enterprise A...
Updated 2026-08-06
Grok logo

Grok

SpaceXAI's irreverent chatbot with a direct line to X/Twitter -- and now Grok 4.5 (launched 2026-07-08), the frontier MoE model trained jointly with Cursor for coding, agentic tasks, and knowledge work at $2/$6 per 1M tokens. Grok 4.3 remains the value tier at $1.25/$2.50

B
7.5/10
Free tierFrom $0
Real-time access to X/Twitter data is ge...Grok 3 benchmarks are competitive with G...
Updated 2026-08-06
Kimi K3 (Moonshot) logo

Kimi K3 (Moonshot)

Moonshot's 2.8T-parameter Kimi K3 (launched 2026-07-16/17) is the largest open-weight model ever released -- 1M context, multimodal, $3/$15 per 1M via API, ranked best-available on Arena.AI at launch. WEIGHTS SHIPPED ~2026-07-26/27 on Hugging Face (2.8T total / 104B activated, safetensors) under a custom Kimi K3 License, not the Modified MIT of the K2 line

A
8.1/10
Free tierFrom $3 / $15
Frontier-tier performance -- Elo 1309 on...Beats Claude Opus 4.5 on several coding ...
Updated 2026-08-06
Runway logo

Runway

Runway Gen-4.5 (shipped 2025-12-01) -- #1 on Artificial Analysis text-to-video leaderboard at 1,247 Elo. **GWM-1 (General World Model family) announced May 2026** for Worlds / Avatars / Robotics, built on Gen-4.5. Gen-4.5 also gained **native audio generation + native audio editing** (May 2026). Gen-4 Turbo supports native 4K. The most capable professional AI video generator available in 2026

B
7.8/10
Free tierFrom $0
Gen-4.5 holds #1 on Artificial Analysis ...Gen-4 Turbo ships native 4K output (Gen-...
Updated 2026-08-06
GitHub Copilot logo

GitHub Copilot

AI code assistant that lives in your editor -- autocomplete on steroids, now with the broadest model picker of any coding tool. **Claude Opus 5 landed 2026-07-24** (Pro+, Max, Business, Enterprise) and **Grok 4.5 on 2026-07-28** (all five paid SKUs, up to 500K context, text and image input, low/medium/high reasoning effort). Both bill usage-based at provider list price, and both are off by default for Business/Enterprise until an admin enables the policy -- though that posture flips on **2026-08-26**, when new GA models covered by GitHub's data-retention agreement start auto-enabling for orgs. **GitHub Models (the separate free playground) was fully retired 2026-07-30.** Usage-based billing went live 2026-06-01 with AI Credits and token metering; code completions are still free; new signups for Student/Pro/Pro+/Max remain PAUSED

A
8.3/10
Free tierFrom $0
Inline code completions feel magical -- ...Works directly in VS Code, JetBrains, Ne...
Updated 2026-08-06
Claude Code logo

Claude Code

Anthropic's terminal-based coding agent that reads your whole repo and makes real changes -- not just suggestions. v2.1.131 (2026-05-06 Code with Claude conf) shipped Code Review GA + Remote Agents + CI Auto-Fix + Routines, plus 2x rate-limit increase from the SpaceX compute deal

B
7.8/10
From $20
Reads and understands your entire codeba...Actually executes code, runs tests, and ...
Updated 2026-08-06

04 Reviews you can actually trust

Every review is based on hands-on testing, cross-referenced user sentiment from G2, Reddit, and Capterra, and real pricing data. We report known bugs. We don't do paid placements.

No paid placements
Real bug reports
Updated daily