Kimi K3 (Moonshot)
Free tier available
- API (Kimi K3)$3 / $15/per 1M tokens (input/output)
- Self-hosted (Free -- K2.6/K2.7 line)$0
- API (Moonshot direct, K2.6)$0.60/per 1M input tokens

Kimi K3 (Moonshot)
Our pickPaperclip
Tier-list head-to-head. Paperclip takes the A-tier slot — here's the breakdown.
Spec sheet
| Tier | A-tier | A-tierwin |
| Overall score | 8.1 / 10 | 8.6 / 10win |
| Free tier | Yes | Yes |
| Starting price | $3 / $15 | $0 |
| Best for | Agentic coding workflows, tool-use agents, long-horizon repository work (1M context), and teams who want fr… | Operators running multiple agents who need real coordination -- an indie hacker running a content shop, a s… |
| Last reviewed | 2026-09-05 | 2026-05-13 |
Head-to-head
Rated 1-10 on the same rubric across all 130 tools we cover.
What you'll pay
Look past the headline number -- entry-tier limits drive most cost surprises.
Free tier available
Free tier available
Kimi K3 (2.8T, launched 2026-07-16) -- Arena.AI ranked it best-available at launch; vendor claims parity with Fable 5, third-party suites pending. Scores below are K2.6/K2.5-era baselines retained until K3 third-party runs publish benchmarks — Paperclip has no published benchmarks
| Benchmark | Description | Score |
|---|---|---|
| SWE-Bench Pro | 58.6% | |
| MMLU-Pro (K2.5 baseline) | 84.8% | |
| GPQA Diamond (K2.5 baseline) | 80.5% | |
| AIME 2025 (K2.5 baseline) | 91.2% | |
| LiveCodeBench (K2.5 baseline) | 74.1% |
The decision
Use-case anchors and category strengths, side by side.
Agentic coding workflows, tool-use agents, long-horizon repository work (1M context), and teams who want frontier-tier quality at a fraction of frontier pricing ($15/M output vs Fable 5's $50/M).
Visit Kimi K3 (Moonshot)Operators running multiple agents who need real coordination -- an indie hacker running a content shop, a small team testing autonomous-biz concepts, or anyone whose 'I'll just open another Claude Code tab' workflow has hit the wall. The org-chart framing is a huge upgrade if you have 5+ agents already.
Visit PaperclipBottom line
Paperclip edges out Kimi K3 (Moonshot) by 0.5 points (8.6 vs 8.1) -- a A-tier vs A-tier split that's narrow but real. Not a blowout; both belong on a shortlist. The score gap shows up most clearly in the categories that matter for Paperclip's strengths, so if those categories are your priority, the lead translates.
Pricing-wise, both tools have a free tier (Kimi K3 (Moonshot) starts $3 / $15, Paperclip starts $0), so you can test either without committing. Compare what each free tier actually unlocks -- usage caps, model access, and feature gates differ a lot more than the headline price suggests, especially as both vendors have tightened limits in 2026.
By use case: pick Kimi K3 (Moonshot) when agentic coding workflows, tool-use agents, long-horizon repository work (1m context), and teams who want frontier-tier quality at a fraction of frontier pricing ($15/m output vs fable 5's $50/m). Pick Paperclip when operators running multiple agents who need real coordination -- an indie hacker running a content shop, a small team testing autonomous-biz concepts, or anyone whose 'i'll just open another claude code tab' workflow has hit the wall. The two tools aren't fighting for the same person -- they're aiming at adjacent jobs that occasionally overlap. If you're squarely in Paperclip's lane, the tier-list ranking and the use-case fit point the same direction; if you're in Kimi K3 (Moonshot)'s lane, the score gap matters less than the fit.
Bottom line: Paperclip is the safer default for most readers, but Kimi K3 (Moonshot) is competitive enough that the tie-breaker is your specific workload, not the spec sheet.
Keep digging
Full Kimi K3 (Moonshot) review
Tier A · 8.1/10
Full Paperclip review
Tier A · 8.6/10
Kimi K3 (Moonshot) alternatives
Other tools in this lane
Paperclip alternatives
Other tools in this lane
Built from our daily AI-tool sweep, last touched September 5, 2026. Honest tier-list reviews — no affiliate-link pieces disguised as advice. See the rubric or how we review.