Grok
B Tier · 7.5/10
**Team Bots launched 2026-09-28 in public beta on Teams and Enterprise plans** -- shared Grok Bots built around a role, with team context (files, instructions, skills), plugins (Salesforce, Notion, GitHub), credentials for APIs without a plugin, and memories, where the bot is shared but each person's conversations and memories stay private; each Team Bot gets its own Slack handle. SpaceXAI runs one per major sales account, and its 9/22 case study says the merged SpaceXAI + Cursor support desk absorbed a 175% ticket increase on Grok Bot with no new hires at $0.20-$0.30 per resolution. **Grok 4.7 (9/21)** remains the flagship at $2/$6 per 1M with chart-only benchmarks; Grok Voice Transcribe 2.0 and the corrected $15/1M-character TTS price stand.
Score Breakdown
Benchmark Scores
Benchmarks for Grok 4.20 (baseline -- Grok 4.5 launched 2026-07-08; vendor-reported 4.5 scores in Known Issues pending third-party verification)
| Benchmark | Description | Score | |
|---|---|---|---|
| MMLU | Knowledge across 57 subjects | 88.5% | |
| GPQA Diamond | Graduate-level science questions | 85% | |
| HumanEval | Python code generation | 90% | |
| Humanity's Last Exam | Frontier difficulty questions | 50.7% |
Last updated: 2026-04-13
Personality & Tone
The irreverent contrarian
Tone: Casual, jokey, and willing to swear. Grok takes strong positions without hedging, leans into an edgy 'based' persona, and cracks jokes far more often than Claude, ChatGPT, or Gemini.
Quirks: Engages with topics other chatbots refuse, pulls live context from X so it reflects whatever is trending that hour, and will freely mock things -- including itself. In SuperGrok's multi-agent mode it can sound like several personalities arguing with each other.
The Good and the Bad
What we like
- +Real-time access to X/Twitter data is genuinely useful for tracking breaking news and trending topics
- +Grok 4.7 (2026-09-21) is a genuine frontier contender priced at $2/$6 per 1M -- a third of GPT-6 Astra input and a fifth of Fable 5.1 -- and it ships day-one inside Cursor, Grok Build and GitHub Copilot; xAI has held the price flat across 4.5, 4.6 and 4.7 while raising capability each time
- +The personality is refreshing if you're tired of overly cautious AI assistants -- it'll actually joke around
- +DeepSearch mode does solid multi-step research, pulling from web and X data simultaneously
What could be better
- −The snarky personality gets old fast when you're trying to get serious work done
- −Tied to the X ecosystem -- you need an X account, and the real-time data skews toward X's user base
- −SuperGrok at $30/mo is steep when Claude Pro and ChatGPT Plus are $20 with arguably better core models
- −Image generation is no longer the weak spot it was -- Imagine Image 2.0 (2026-08-07) added region editing, segmentation, background removal and 5-image multi-ref, and xAI places it #2 worldwide on both Arena image boards; the remaining gap vs ChatGPT/Gemini is analysis and document understanding, not generation
Pricing
Free
- ✓~10 prompts per 2 hours
- ✓Basic Grok access
- ✓Requires X account
X Premium
- ✓Higher query limits
- ✓Grok 4.20 access
- ✓Bundled X social features
X Premium+
- ✓Higher Grok 4.20 access
- ✓Ad-free X
- ✓Priority responses
SuperGrok
- ✓Full Grok 4.20 (4-agent multi-agent system)
- ✓DeepSearch mode
- ✓Highest rate limits
- ✓Think mode
- ✓$300/yr option (16% off)
SuperGrok Heavy
- ✓Grok 4 Heavy model
- ✓Highest priority
- ✓Multi-agent at scale
- ✓Note: Grok 4.3 beta-gating ended 2026-05-02
API (Grok 4.7, launched 2026-09-21)
- ✓500k context; cached input $0.50 per 1M
- ✓Long context (prompt at or above 200k tokens): $4 / $1 cached / $12 for every token in the request
- ✓Grok 4.7 Fast: 2x output speed at 2x price ($4 / $1 / $12; $6 / $1.50 / $18 above 200k) -- Cursor and Grok Build only, NOT on the public API, excluded from the Grok Build free tier
- ✓US regional endpoint (us.api.x.ai) bills 1.1x -- grok-4.7 and grok-4.6 only
- ✓Knowledge cutoff May 2026; Responses API always returns encrypted reasoning
- ✓Vendor benchmark charts only (CursorBench 4.0, DeepSWE v1.1, EEBench, AA Briefcase v1.1, Terminal-Bench 4.0, Harvey Legal, HealthBench Professional) -- no published table, numbers withheld here
API (Grok 4.5, launched 2026-07-08)
- ✓SpaceXAI + Cursor joint frontier MoE model
- ✓Faster variant $4/$18 via Cursor
- ✓Available in Grok Build, SpaceXAI console/API, Cursor (all plans)
- ✓NOT available in EU at launch
- ✓Vendor benchmarks: Terminal-Bench 2.1 83.3%, SWE-Bench Pro 64.7% (third-party verification pending)
API (Grok 4.3)
- ✓Production launch 2026-05-02 (~40% input / ~60% output price cut vs 4.20)
- ✓1M context window
- ✓Reasoning tokens billed at output rate
- ✓Native video input + PDF/PPT/spreadsheet output
- ✓Custom Voices voice cloning free on console (80+ presets, 28 languages)
- ✓Imagine Agent Mode (creative workflow agent, beta)
Known Issues
- TEAM BOTS -- SHARED GROK BOTS WITH PRIVATE CONVERSATIONS, PUBLIC BETA ON TEAMS AND ENTERPRISE (2026-09-28, vendor-primary): 'Today we're launching Team Bots, Grok Bots that work and learn alongside your team. Give one access to the files, apps, and expertise it needs, then share it so everyone can work from the same context.' A Team Bot is built around a role or workflow and combines four things: **Context** (files, instructions, skills -- brand guides, internal docs), **Plugins** (Salesforce, Notion, GitHub; connected per person or configured for the team), **Credentials** (secure access to third-party APIs without a plugin) and **Memories**. Privacy model: 'although the Bot is shared, each person's conversations with it remain private. The Bot keeps separate context and memories for each user while drawing on the skills shared across the team.' Slack: each Team Bot has its own handle and can be invited to a channel. SpaceXAI's own deployments: a Team Bot per major sales account that reviews company news, Gong calls, Notion and Slack nightly and posts a role-tailored morning briefing; engineering coordination; company data questions. Customer example: Harper (small-business insurance) built a lapsed-policy reinstatement bot in 24 hours and says it saved customers over $120,000 across hundreds of policies. **Availability: 'Team Bots is available today in public beta on Teams and Enterprise plans'**, with pre-built bots for sales, product management, marketing and data analytics; macOS download, contact sales. No price beyond plan membership is stated.Source: SpaceXAI (x.ai/news/team-bots, JSON-LD datePublished 2026-09-28) -- fetched 2026-09-28 via curl with browser UA · 2026-09-28
- GROK BOT RUNS THE COMBINED SPACEXAI + CURSOR SUPPORT DESK -- 175% MORE TICKETS, NO NEW HIRES, $0.20-$0.30 PER RESOLUTION (2026-09-22, vendor case study): 'When Cursor became part of SpaceXAI on August 14, our two customer support teams began coming together around a much broader product portfolio' while preparing to launch Grok Bot; the company used Grok Bot itself across the operation. Claims: 'our new combined team has seen a 175% increase in support tickets, but we have not had to hire any new people ... We might have hired 200 additional people otherwise'; against flat-rate AI support tools at '$1 to $4 per resolution', Grok Bot is billed on plan usage and 'with minor optimizations, we've been able to resolve tickets for as low as $0.20 to $0.30'. Method: connected Plain (ticketing) and Linear first, let it act as ticket owner with internal notes only and human approval on every write, added traces and evaluations to every run, then let it answer customers directly after a day of manual review; every ticket now gets a pre-investigation pass; known issues route to Linear, backend errors to Datadog, and it reproduces bugs with a video for engineering. This is vendor self-reporting on its own product, so treat the ratios as marketing-grade, but it is also the first concrete per-resolution cost figure xAI has published for Grok Bot.Source: SpaceXAI (x.ai/news/grok-bot-customer-support, JSON-LD datePublished 2026-09-22) -- fetched 2026-09-28 via curl with browser UA · 2026-09-22
- GROK 4.7 SHIPPED (2026-09-21, vendor-primary) -- A NEW BASE MODEL AT AN UNCHANGED PRICE, AND THE BENCHMARKS ARE PICTURES, NOT NUMBERS. SpaceXAI's framing: '**our most capable model for coding and knowledge work. It works longer on difficult tasks, checks its own work more carefully, and comes with our best-calibrated safeguards to date. Served at the same price and speed as Grok 4.6.**' Under the hood it is **a new, larger base model than Grok 4.6**, trained 'with a longer reinforcement learning run on a harder mix of tasks, weighted toward problems that take many hours to complete', and trained 'to natively understand the Grok Bot harness' -- so the model and the agent product are now co-designed. **PRICING, VERIFIED ON THE LIVE RATE CARD THE SAME DAY:** $2.00 input / $0.50 cached / $6.00 output per 1M for prompts under 200k tokens; **$4.00 / $1.00 / $12.00 once a prompt reaches 200k** (the long-context rate then applies to every token in the request); 500k context; knowledge cutoff May 2026. **Grok 4.7 Fast** is 'the same Grok 4.7 model served on faster infrastructure, at twice the standard token rates' -- $4 / $1 / $12 below 200k, $6 / $1.50 / $18 above -- and xAI is explicit that it is '**available only through Cursor and Grok Build; it is not available on the public xAI API, and Grok Build's free tier does not include it.**' A US regional endpoint (us.api.x.ai) bills 1.1x and currently serves only grok-4.7 and grok-4.6. **BENCHMARKS -- READ THIS BEFORE QUOTING ANY SCORE:** the post compares Grok 4.7 against Grok 4.6, GPT-5.6 Sol and Fable 5.1 on CursorBench 4.0, DeepSWE v1.1 (Grok's entry flagged as a high-effort run), EEBench, AA Briefcase v1.1, Terminal-Bench 4.0, Harvey Legal Agent Benchmark and HealthBench Professional, and against Fable 5.1 and GPT-6 Astra on GDPval, AA Briefcase and EEBench -- **but every figure is rendered inside charts and a styled table we could not extract, and xAI publishes no plain-text results table.** The only prose claims are that Grok 4.7 'is at the frontier in price-performance' on CursorBench 4.0 and that on GDPval and AA Briefcase it 'improves upon Grok 4.6 ... and performs comparably to other frontier models'. Per this site's rule, no numbers ship until a vendor table or a third-party leaderboard exists. **SAFETY IS THE ONE PLACE XAI GIVES NUMBERS:** 'the strongest model we've tested on refusals and jailbreak resistance', topping LatchBio's biosafety benchmark at **62.4%**, and on xAI's own HackerBench v0.3 allowing '**only 3.3% of risky dual-use prompts through**' -- with 'select cybersecurity partners' getting invite-only access to its red-team capabilities. That is the same gate-the-cyber-tier pattern OpenAI (Daybreak) and Google (Fairwind) adopted this month. **AVAILABILITY:** Cursor and Grok Build on day one, the Grok API, 'third-party coding harnesses, and model routers and cloud platforms' -- and **GitHub Copilot the same day (see next entry)**. Grok 4.6 remains on the rate card at identical prices, so there is no forced migration.Source: SpaceXAI (x.ai/news/grok-4-7, JSON-LD datePublished 2026-09-21); docs.x.ai/docs/pricing (rate table incl. long-context tier, Grok 4.7 Fast section, US regional endpoint) and docs.x.ai/docs/models ('Last updated: September 21, 2026'; 500k context; May 2026 cutoff); Cursor blog 'Introducing Grok 4.7' (cursor.com/blog/grok-4-7, a stub pointing to x.ai) -- all fetched 2026-09-21 via curl with browser UA · 2026-09-21
- GROK 4.7 REACHED GITHUB COPILOT ON LAUNCH DAY -- AND THIS TIME IT IS ON BY DEFAULT (2026-09-21, vendor-primary on GitHub's side): GitHub's changelog lists **Grok 4.7 for Copilot Pro, Pro+, Max, Business and Enterprise**, across VS Code, Visual Studio, Copilot CLI, the Copilot cloud agent, the Copilot app, JetBrains, Xcode and Eclipse, billed 'at provider list pricing under usage-based billing' with a 'gradual' rollout. **The enablement posture has flipped since Grok 4.6:** the 8/14 entry for 4.6 said the org policy was 'off by default'; the 4.7 entry says new models are **enabled automatically unless administrators disable them** under GitHub's global model policy (GA since 8/26). So for most Business/Enterprise orgs Grok 4.7 simply appears in the picker. **Zero-day placement is now xAI's standard playbook** -- 4.6 took two days to reach Copilot; 4.7 took zero -- and it lands the same week GitHub scheduled **Grok 4.5's removal from Copilot for 2026-10-19** (migration target Grok 4.6, per the 9/18 slate on github-copilot.ts). GitHub published no benchmark claims and no premium-request multiplier, describing the model only as 'xAI's latest reasoning model' for 'agentic coding and complex, multistep workflows'.Source: GitHub Changelog (github.blog/changelog/2026-09-21-grok-4-7-is-now-available-in-github-copilot, on-page 'September 21, 2026') -- fetched 2026-09-21; enablement contrast against github.blog/changelog/2026-08-14-grok-4-6-is-now-available-in-github-copilot (already on this page) · 2026-09-21
- GROK BUILD GAINS PERSISTENT MEMORY (2026-09-16, vendor-primary): Grok Build 'now carries conventions, decisions, and project facts from one session to the next'. Mechanics, per xAI: **capture runs in the background after every completed turn** and 'never blocks the session'; notes are **markdown files, one topic per subject**, kept in a per-project workspace scope plus a global scope for cross-project preferences; before starting related work Grok 'reads the topics that cover the area and applies them, including in sessions where the subject never comes up'; **instructions in the current conversation always take precedence** over a note. Two new commands: **/memory** (a read-only browser of every memory file, grouped by scope) and **/dream** (merges fresh observations into topic files; also runs periodically on its own). xAI states what is deliberately left out -- 'task state, tentative conclusions, secrets, and anything the repository or its docs already cover'. It applies to new sessions only (run /new or start a fresh session). **Why it matters for the comparison:** this is the persistent-project-memory layer that terminal coding agents have converged on, and Grok Build now has it alongside its 8-agent parallelism and plugin marketplace -- the remaining gap versus Claude Code and Cursor is ecosystem depth, not the feature list. Found via the daily pulse and confirmed against the vendor post; the 9/17 sweep's x.ai enumeration had missed it.Source: SpaceXAI (x.ai/news/grok-build-memory, JSON-LD datePublished 2026-09-16) -- fetched 2026-09-21 via curl with browser UA · 2026-09-16
- GROK BOT FOR ENTERPRISE -- AND THE FREE-USAGE OFFER IS BUNDLED WITH CURSOR, WHICH IS THE PART THAT MATTERS (2026-09-03, vendor-primary): SpaceXAI made **Grok Bot available for enterprises**, adding the governance layer the 8/11 early beta lacked: '**today's release adds access, network, and audit controls**' to let organizations 'govern Bots at scale', with each user's work running 'in its own secure and isolated environment'. **THE ACQUISITION IS NOW VISIBLE IN THE COMMERCIAL TERMS, NOT JUST THE ORG CHART:** '**Grok and Cursor Enterprise customers have free usage for the next two weeks, and can invite their whole organization, including people without an existing seat.**' **Cursor Enterprise customers are being given a SpaceXAI product for free, org-wide, with no seat requirement** -- the clearest instance yet of the consolidation being used as a distribution channel, and it lands in the same window that OpenAI is winding Cursor off its models. **This is a dated promotion: two weeks from 2026-09-03 puts expiry around 2026-09-17.** Treat 'free' as an acquisition subsidy with an end date, not a plan feature. **WHAT A BOT ACTUALLY IS, in xAI's framing:** 'a worker you create inside Grok Bot for a specific job', each running 'on its own computer in the cloud', usable via any app or website the way a person would. Training is by demonstration -- 'have it follow along once. It saves the routine, takes your corrections, and runs it on its own from then on' -- and working Bots can be handed to a colleague as a template. Bots can message each other and share context. xAI claims 'thousands of organizations' since launch, naming **Legora, Supermicro and Servi**. **PRICING FOR THE ENTERPRISE TIER IS STILL NOT PUBLISHED** -- the post routes to 'contact sales', so the standing gap on this page (Grok Bot standalone pricing) remains open after the free window.Source: SpaceXAI (x.ai/news/grok-bot-for-enterprise, JSON-LD datePublished 2026-09-03) -- fetched 2026-09-05 via curl with browser UA · 2026-09-03
- XAI PUBLISHES ITS OWN GROK BOT CASE STUDY -- A VENDOR-RUN INTERNAL TRIAL, NOT A CUSTOMER RESULT (2026-09-04, vendor-primary): SpaceXAI posted 'Setting Grok Bot loose on procurement', reporting that a Bot it named **Haggle Bot**, given access to '**vendor spend, contracts, and usage data**', '**found more than $100,000 in direct savings**' by surfacing unused SaaS seats, renewals that no longer matched usage, and recurring purchases that were never re-shopped. **READ THE PROVENANCE BEFORE THE NUMBER: this is xAI running its own product on its own procurement and publishing the result.** It is a plausible and specific use case, and the setup detail is genuinely informative -- 'you tell or show a Bot what you want it to do, give it access to the right tools, and let it execute without designing the workflow step by step' -- but $100,000 at an unstated company size, with no baseline, no methodology and no independent verification, is a marketing figure and should be cited as one. **It is included here because it is the clearest published description of what an enterprise Grok Bot deployment actually looks like in practice, not because the savings claim is verified.** The access it required -- full vendor spend, contracts and usage data -- is also the honest part of the story: this class of result needs broad financial-systems access, which is exactly what the 9/03 enterprise access, network and audit controls exist to govern.Source: SpaceXAI (x.ai/news/grok-bot-procurement, JSON-LD datePublished 2026-09-04) -- fetched 2026-09-05 via curl with browser UA · 2026-09-04
- GROK BOT GETS A FIRST-PARTY X INTEGRATION, AND xAI HANDS OUT X API CREDITS TO PULL YOU IN (2026-08-29, vendor-primary): xAI shipped an X connector for **Grok Bot**. The mechanics are unusually frictionless and that is the point: '**Connect your X account in Grok Bot and we'll create a developer account for you if you don't have one**' -- so the X developer-account step, historically the barrier to building anything against X, is provisioned automatically. '**Paid Grok Bot users get free X API credits to start.**' Once connected you can ask a Bot to '**search posts, read your timeline, check mentions, or pull together what's happening on X**'. **WHY THIS IS MORE THAN A CONNECTOR:** X API access has been expensive and rationed since 2023, and xAI is now giving it away as a bundled benefit of a Grok subscription. **That converts a paywalled data asset into a customer-acquisition subsidy for the agent product** -- and it is only possible because SpaceX owns both, following the consolidation that closed 2026-08-14. Read alongside the 8/26 change that pushed Grok Bot down to Cursor Pro and every paid Cursor tier: within two weeks xAI has both widened who gets Grok Bot and deepened what it can reach. **CAVEATS, AND xAI STATES THEM ITSELF:** '**This is the first version of this integration**', with no capability guarantees beyond the four read-oriented actions listed. **The credit grant is characterised only as 'free X API credits to start'** -- no volume, no duration, and no published rate for what happens after they run out, so do not build a production workflow on this without pricing the fallback. Access requires signing in with the X connector from inside Grok Bot.Source: xAI (x.ai/news/grok-bot-and-x, on-page 'Aug 29, 2026') -- fetched 2026-08-31 via curl with browser UA (x.ai 403s WebFetch) · 2026-08-29
- GROK 4.6 ON GOOGLE'S GEMINI ENTERPRISE AGENT PLATFORM, WITH A PUBLISHED RATE CARD (2026-08-21, vendor-primary): Grok 4.6 became available through **Model Garden** on the Gemini Enterprise Agent Platform. **Pricing, verbatim from xAI: $2.00 per 1M input, $0.50 per 1M cached input, $6.00 per 1M output.** Same 500k context and low/medium/high/xhigh reasoning efforts. **PRICE CORRECTION WORTH KEEPING: our daily triage reported cached input at $0.30 per 1M. The vendor page says $0.50. We ship $0.50** -- if an aggregator later shows $0.30, this note is why we did not follow it. Note also that xAI's post title says 'Vertex AI' in the URL slug while the page body consistently says **Gemini Enterprise Agent Platform**; Google renamed the surface, and the body text is the current name.Source: xAI (x.ai/news/grok-4-6-vertex-ai, on-page 'Aug 21, 2026') -- fetched 2026-08-28 via curl with browser UA · 2026-08-21
- GROK 4.6 ADDS MICROSOFT FOUNDRY -- THE FOURTH MAJOR CLOUD, AND ALL FOUR NOW CARRY IT (2026-08-26, vendor-primary): xAI put Grok 4.6 on **Microsoft Foundry**, restating **500k context and configurable reasoning efforts (low, medium, high, xhigh)**. xAI's pitch is procurement-shaped: Foundry gives orgs '**a single place to evaluate Grok 4.6 against other frontier models**', run workload-specific tests, deploy managed endpoints, and apply enterprise security and governance. **NO PRICING WAS PUBLISHED on the Foundry post** -- unlike the Gemini Enterprise Agent Platform post, which did carry a rate card. Do not assume the two are priced identically. **THE PATTERN IS NOW THE STORY: within roughly two weeks Grok 4.6 landed on GitHub Copilot, Amazon Bedrock, Google's Gemini Enterprise Agent Platform, and Microsoft Foundry.** For a model from a SpaceX subsidiary, distribution breadth -- not benchmark position -- is what changed this month. **This post was missed entirely by our daily triage; it was found by direct newsroom enumeration.**Source: xAI (x.ai/news/grok-4-6-microsoft-foundry, on-page 'Aug 26, 2026'); xAI links Microsoft's own 'Grok 4.6 comes to Microsoft Foundry Models' announcement -- fetched 2026-08-28 via curl with browser UA · 2026-08-26
- GROK BOT OPENS UP TO EVERY SUPERGROK AND CURSOR PAID TIER -- AND THE PULSE HAD THIS DATE WRONG BY FIVE DAYS (2026-08-26, vendor-primary): xAI expanded Grok Bot, launched in beta **2026-08-11**, from its narrow top-tier bundle to **all SuperGrok and Cursor paid plans**. **Full entitlement list, verbatim from the vendor page:** SuperGrok, SuperGrok Plus, SuperGrok Heavy, Cursor Pro, Cursor Pro+, Cursor Ultra, and Cursor Teams (Standard and Premium). **The commercially important line is the metering: 'Grok Bot comes with its own usage, separate from your Grok and Cursor plans, so anything you hand off to a Bot won't count against your existing usage.'** That is a genuine grant of new capacity, not a repackaging of existing quota -- and it is the detail that decides whether this is a real upgrade for an existing subscriber. **DATE CORRECTION, RECORDED SO A LATER SWEEP DOES NOT 'FIX' IT BACK: our daily triage reported this as 2026-08-21 on three consecutive days. The vendor page is dated Aug 26, 2026. Vendor-primary wins.** Earlier triage also described the new tiers as 'SuperGrok Plus and Cursor Pro+'; the page says **SuperGrok and Cursor Pro**, i.e. one tier lower in each ladder and therefore a wider expansion than reported.Source: xAI (x.ai/news/grok-bot-more-plans, on-page 'Aug 26, 2026') -- fetched 2026-08-28 via curl with browser UA · 2026-08-26
- GROK 4.6 IS GA ON AMAZON BEDROCK (2026-08-19, vendor-primary): xAI's flagship reached AWS's managed model service seven days after launch -- '**Grok 4.6 is now generally available on Amazon Bedrock**', available 'to all developers in supported AWS Regions'. **PRICING IS IDENTICAL TO xAI'S OWN API: $2 per 1M input tokens, $6 per 1M output tokens** -- no AWS resale premium, which is the fact most worth checking before assuming Bedrock costs more. **SPECS RESTATED ON THE BEDROCK PAGE:** 500K context window and **configurable reasoning effort across low, medium, high and xhigh**. **WHY THIS BELONGS IN THE SAME STORY AS THE LAUNCH:** Grok 4.6 has now landed on three distribution surfaces inside eight days -- **Cursor and Grok Build on day zero (8/12), GitHub Copilot (8/14), and Amazon Bedrock (8/19)** -- which is the fastest multi-surface rollout xAI has managed for any model. For buyers already standardised on Bedrock for procurement, data-residency or billing reasons, Grok 4.6 is now reachable without a separate xAI contract. **THE CAVEAT:** xAI says 'supported AWS Regions' without listing them, and the post gives no batch, provisioned-throughput or caching rates -- check the AWS region table and Bedrock's own pricing page before committing, because those are the dimensions where managed-service pricing usually diverges from a vendor's direct APISource: SpaceXAI (x.ai/news/grok-4-6-amazon-bedrock, dated Aug 19, 2026, fetched 2026-08-20 via curl -- x.ai 403s WebFetch) · 2026-08-19
- GROK 4.6 IS NOW IN GITHUB COPILOT -- TWO DAYS AFTER LAUNCH, ON EVERY PAID COPILOT SKU (2026-08-14, vendor-primary on both sides): xAI's newest coding model reached GitHub Copilot on **2026-08-14**, two days after its 8/12 launch. xAI's framing: Grok 4.6 is '**now live in GitHub Copilot for the millions of developers who work in VS Code and across GitHub every day**.' GitHub's changelog says it '**performed especially well on longer-horizon tasks requiring sustained reasoning and tool use**.' **Availability: Copilot Pro, Pro+, Max, Business, and Enterprise** -- the full paid list including base Pro. Surfaces: VS Code, Visual Studio, Copilot CLI, the Copilot cloud agent, the GitHub Copilot app, JetBrains, Xcode and Eclipse. **Org admins must still switch it on -- GitHub states 'the policy is off by default'** for Business and Enterprise; xAI's softer phrasing ('some businesses and enterprises will need to enable the model') understates that, so trust GitHub's wording. **PRICING CONFIRMED UNCHANGED AT THE SOURCE:** xAI's own post restates direct console pricing at **$2 per million input / $6 per million output** -- identical to Grok 4.5. So across the 4.5 -> 4.6 generation xAI has held price flat while raising capability, and has now placed the model in Cursor (day one, 8/12) and Copilot (day three, 8/14). **Strategic read:** xAI is no longer relying on its own surfaces for developer reach. Between the Anysphere/Cursor relationship and a same-week Copilot placement, Grok is now selectable inside the two highest-volume paid coding surfaces in the market, at a price well under Anthropic's and OpenAI's frontier tiers. The buyer-side caveat is unchanged from the Cursor entry: cheap frontier access inside someone else's IDE is still subject to that IDE's billing multipliers, and **GitHub published no premium-request multiplier for Grok 4.6**, so the effective cost inside Copilot is not yet knowable from vendor primary sourcesSource: SpaceXAI (x.ai/news/grok-4-6-github-copilot, on-page date Aug 14, 2026) and GitHub changelog (github.blog/changelog/2026-08-14-grok-4-6-is-now-available-in-github-copilot) -- both fetched 2026-08-17 · 2026-08-14
- GROK 4.6 SHIPPED (2026-08-12, vendor-primary) -- AND IT ARRIVED FIVE DAYS AFTER THE DATE WE RECORDED AS MISSED. Read this alongside the 7/16 entry below, which correctly stated on 8/10 that 4.6 had not shipped: the Musk-stated 8/7 target did pass without a release, but xAI then published **x.ai/news/grok-4-6, dated Aug 12, 2026**. xAI's framing: 4.6 'builds on Grok 4.5 with a particular focus on long-running agents and more ambitious interactive and visual work,' staying with tasks 'across many steps.' **PRICING: $2/M input and $6/M output** -- identical to Grok 4.5, so this is a capability bump at flat price, not a price move. **There is also a 'fast variant which is twice the price'** ($4/$12 by that arithmetic; xAI states the multiplier, not the absolute figures, so we do not print derived numbers as vendor facts). **VENDOR-PUBLISHED EVAL TABLE (Grok 4.6 High vs Grok 4.5 High):** AA Intelligence Index **61 (from 56)**, GDPVal-AA v2 **1753 (from 1526)**, CursorBench v3.2 **69.9% (from 66.7%)**, DeepSWE v1.1 **65.9% (from 54%)**, FrontierCode v1.1 Extended **61.3% (from 56.6%)**, APEX-Agents **57.5% (from 47.1%)**, Terminal-Bench v3.0 **26% (from 15.7%)**, APEX-SWE **56.4% (from 53.6%)**, AA-Briefcase **1577 (from 1313)**, Harvey LAB **15.8% (from 12.9%)**. **THE HEADLINE CLAIM, STATED PRECISELY:** xAI says 4.6 'matches GPT-5.6 Sol on the Artificial Analysis Intelligence Index' -- both at **61** in xAI's own table, with **Fable 5 Max at 62**. **BUT THE SAME TABLE SHOWS WHERE IT LOSES**, and this is the part marketing coverage will drop: on **DeepSWE v1.1 GPT-5.6 Sol scores 73% against Grok 4.6's 65.9%**, and on **Terminal-Bench v3.0 Sol scores 34.6% and Fable 5 Max 34.1% against Grok 4.6's 26%**. So 'matches Sol' is true of one composite index and not of the agentic-coding benchmarks underneath it. All competitor figures are xAI-selected from rivals' published system cards or leaderboards; third-party verification pending. **AVAILABILITY DAY ONE:** Cursor and Grok Build, the xAI API, and partners **OpenRouter, Vercel and Cloudflare**, with **2x included usage inside Grok Build and Cursor for the first week**. **METHOD NOTE FOR FUTURE SWEEPS: this is the resolution of a pending dated flip that had already been marked 'did not ship' once. A missed vendor date is not a cancelled product -- keep checking the slug index.**Source: SpaceXAI (x.ai/news/grok-4-6, on-page date Aug 12, 2026, fetched 2026-08-13 via curl -- x.ai 403s WebFetch); corroborated by Cursor's own post (cursor.com/blog/grok-4-6, dated Aug 12, 2026, fetched 2026-08-13) · 2026-08-12
- GROK BOT -- AN AGENT PRODUCT WITH ITS OWN COMPUTER, AND AN UNUSUAL CROSS-VENDOR ENTITLEMENT (2026-08-11, vendor-primary, EARLY BETA): xAI launched **Grok Bot**, described as 'your team of always-on agents' -- 'AI teammates you can give real work to.' The design claim that distinguishes it from chat-style agents: **bots get their own cloud computer**, sign into the tools you already use, and 'work across apps, inboxes, and more,' including **'platforms with no clean API or MCP'** -- i.e. it drives real UIs rather than requiring integrations. They 'finish jobs end to end, and only come back when something needs your approval,' retain conversation memory, and 'learn how you like things done.' xAI says it began as an internal prototype used for sales outbound, marketing campaigns, office operations and bug fixes. **AVAILABILITY IS THE NOTEWORTHY PART, QUOTED EXACTLY:** 'Grok Bot is in beta and available today for **SuperGrok Heavy, Cursor Ultra, and Cursor Teams Premium subscribers** on desktop and iOS,' with **enterprise access by waitlist**. That means an xAI product ships as an entitlement of a *Cursor* subscription -- which we verified is coherent rather than a typo: **Cursor's own pricing page lists Ultra and Teams Premium as real tiers and advertises 'Generous limits for Grok' on every paid plan, with Grok as a top-level nav item**. This follows SpaceX's Anysphere/Cursor acquisition and is the clearest sign yet that xAI treats Cursor's subscriber base as a first-party distribution channel. **Caveats to hold:** it is explicitly **early beta**, no pricing is published for standalone access, and no benchmark or reliability data accompanies the autonomy claims -- an agent with persistent credentials into your live SaaS tools is a materially larger blast radius than a chat assistant, and xAI publishes no security model for it in this postSource: SpaceXAI (x.ai/news/introducing-grok-bot, on-page date Aug 11, 2026, fetched 2026-08-13 via curl); tier names cross-checked against cursor.com/pricing (fetched 2026-08-13) · 2026-08-11
- IMAGINE IMAGE 2.0 IS GA -- PRECISE EDITING ARRIVES, AND GROK IS NOW #2 ON BOTH IMAGE ARENAS (2026-08-07, vendor-primary): xAI shipped **Imagine Image 2.0 as the new Quality Mode** on grok.com/imagine and the iOS and Android apps. The pitch is explicitly utilitarian -- 'make images you can use in real work' -- with the model planning typography and layout 'the way a designer would' so dense multi-part visuals hold together and small text stays sharp. **The substantive change is that editing is now first-class, not a re-roll.** Four tools: the **magic wand** edits only the region you point at and leaves the rest untouched; **segmentation** selects precise areas to change; **background removal** exports any subject on transparency; and **multi-ref editing accepts up to 5 input images in a single generation**, which removes a manual compositing step. **Smart resize** refills the frame across ten ratios (1:2, 9:16, 2:3, 3:4, 1:1, 4:3, 3:2, 16:9, 2:1). xAI also added **templates** -- prepackaged workflows for photo editing, product shots, headshots, icons, game assets, mascots, e-commerce and UGC photos, emoji and merch. **RANKING, QUOTED PRECISELY:** xAI says Image 2.0 'ranks second in the world in both text-to-image generation and image editing' on Arena overall Elo **as of Aug 7, 2026** -- note this is a vendor-stated placement citing the Arena leaderboards, third-party re-verification pending, and that **xAI models are listed on Arena under 'SpaceXAI'**, which is where to look if you go check it yourself. This directly retires the long-standing con on this page that Grok's image generation lags ChatGPT and Gemini -- on the vendor's own leaderboard reading it no longer does, though the con predates the Imagine line and is kept below as historical context for Grok 3-era usersSource: SpaceXAI (x.ai/news/grok-imagine-image-2, on-page date Aug 7, 2026, fetched 2026-08-10 via curl -- x.ai 403s WebFetch) · 2026-08-07
- IMAGINE VIDEO 1.5 GETS REFERENCES, TEXT-TO-VIDEO AND NATIVE 1080p (2026-07-31, vendor-primary -- this page had only ever recorded the 6/3 Imagine 1.5 *preview*, so the shipped feature set was several steps behind): xAI substantially expanded its video model. **Three things landed.** (1) **Multi-reference conditioning** -- pass reference images and 'each reference image locks one thing in place: a face, a product, a location', so you can keep a character and swap the scene, keep the scene and swap the character, or hold both and change only the action. **Up to seven references per generation.** (2) **Voice reference / voice consistency** -- supply a character image plus a voice reference and 'both hold: the same face and the same voice in every scene', which is the piece most competing video models still cannot do without a separate dubbing pass. (3) **Text-to-video and native 1080p** -- generation from a prompt with no starting image (xAI describes it as pairing their image generation with image-to-video), and 1080p output for both text-to-video and image-to-video, up from the 720p ceiling the 6/3 preview shipped with. **AVAILABILITY IS SPLIT AND WORTH READING CAREFULLY:** text-to-video and native 1080p are **generally available** on grok.com/imagine, iOS and Android. Image and voice references started **2026-07-31 in the US only, for SuperGrok Heavy and SuperGrok Plus**, on grok.com/imagine and iOS, with xAI saying they roll out to all tiers 'over the next few days' -- so by now that gate has probably widened, but the vendor post is the last first-party statement of scope. **In the API**, image references, text-to-video and native 1080p are live under the model id **`grok-imagine-video-1.5`**; **voice reference support is 'available on request'**, i.e. not self-serve. No pricing was published in the postSource: xAI/SpaceXAI (x.ai/news/grok-imagine-video-1-5-references, datePublished 2026-07-31T00:00:00Z, fetched 2026-08-06 via curl -- x.ai 403s WebFetch) · 2026-07-31
- PRODUCT PAIR (2026-07-15/16, vendor posts on x.ai/news): (1) **Grok Automations** (7/16) -- scheduled and autonomous recurring tasks in the Grok app (standing queries, monitoring, repeat jobs), SpaceXAI's answer to ChatGPT's Scheduled Tasks. (2) **Grok Build open-sourced** (7/15) -- the terminal coding agent's source is now public, relevant if you want to audit or extend the harness. Roadmap noise, clearly labeled: Musk indicated a **Grok 4.6** is in the pipeline (~7/17-18, no date, no specs -- aggregator-grade until a vendor post exists) and Grok 5's timeline has slid repeatedly. **GROK 4.6 UPDATE -- RESOLVED, SEE THE 2026-08-12 ENTRY ABOVE.** The history is worth keeping because it shows how these dates actually behave: the Musk-stated **8/7 target did pass with nothing shipped** (verified 8/10 by re-enumerating the full x.ai/news index -- no post, no slug), and then **Grok 4.6 launched on 2026-08-12**, five days late. So the 8/10 reading was accurate at the time and the product was not cancelled, merely slipped. Standing rule this produced: treat 'Grok X is out' as unsourced until a vendor post exists, but do **not** conclude from a missed date that a release is dead. EU AVAILABILITY STILL UNRESOLVED: the 'mid-July' EU promise for Grok 4.5 has conflicting reports (one aggregator claims EU access landed ~7/16; another says still unavailable) and x.ai's Grok 4.5 page wouldn't load for verification -- we are NOT flipping EU availability until vendor wording confirms; re-check next sweepSource: SpaceXAI (x.ai/news/grok-automations, x.ai/news/grok-build-open-source -- both listed on the x.ai news index, scraped in-session) · 2026-07-16
- GROK 4.5 LAUNCHED (2026-07-08 developer surfaces, public rollout reported 7/9): SpaceXAI shipped **Grok 4.5, 'our smartest model built for coding, agentic tasks, and knowledge work'** -- a mixture-of-experts frontier model **trained jointly with Cursor** (SpaceX closed its Anysphere/Cursor acquisition in June) on trillions of tokens of Cursor data. Musk's framing: 'Opus-class, but faster, more token-efficient and lower cost.' **API pricing: $2/M input + $6/M output** (a faster variant at $4/$18 is offered through Cursor). Day-one availability: **Grok Build, the SpaceXAI console/API, and Cursor on all plans**; press reports public access via grok.com and the X app from 7/9; **NOT available in the EU at launch**. Vendor-reported benchmarks (charts, third-party verification pending): **Terminal-Bench 2.1 83.3%, SWE-Bench Pro 64.7%, #1 on Harvey's Legal Agent Benchmark**, with standout token efficiency (~16K output tokens per SWE-Bench Pro task vs ~67K for Opus 4.8 max per launch charts); Artificial Analysis measured 91.3 tok/s on the API. CONSUMER ACCESS CONFIRMED (7/10 press): available to **X Premium and SuperGrok subscribers** in the Grok interface, and it's the **default model in Grok Build** (beta, SuperGrok + X Premium+). API also live on OpenRouter, Vercel, Cloudflare, Snowflake, Databricks. **INDEPENDENT BENCHMARKS PUBLISHED (Artificial Analysis, 7/9-10): Intelligence Index 54** -- AA's launch article places it 4th among frontier models behind Fable 5, GPT-5.5, and Opus 4.8 (the model page ranks it #8 of 188 counting all model variants); **GDPval-AA v2 Elo 1543**; 92.9 output tok/s measured; 500K context window per AA. AA calls the 16-point gen-over-gen jump xAI's largest ever. Press caveat: reviewers note a higher hallucination rate than frontier peers. Aggregator claims of a '1.5T-param V9 foundation' remain UNVERIFIED -- treat as rumorSource: SpaceXAI (x.ai/news/grok-4-5), Cursor blog (cursor.com/blog/grok-4-5), Artificial Analysis (artificialanalysis.ai/models/grok-4-5), Axios (2026-07-08) · 2026-07-10
- GROK BUILD PLUGIN MARKETPLACE (2026-06-11, vendor-primary): xAI launched a built-in **plugin marketplace for Grok Build** -- plugins bundle skills, slash commands, agents, hooks, MCP servers, and LSPs; installs are commit-SHA-pinned for supply-chain safety; the catalog is open to community submissions via PR. Launch partners: MongoDB, Vercel, Sentry, Chrome DevTools, Cloudflare. Mirrors the plugin/extension pattern Claude Code and Gemini-CLI-era tooling established -- Grok Build is maturing fast for a product still labeled betaSource: xAI news (x.ai/news/grok-plugin-marketplace), GitHub (github.com/xai-org/plugin-marketplace) · 2026-06-11
- JUNE CLUSTER (2026-06, all vendor-primary on x.ai/news): **Grok Imagine 1.5 Preview** (6/3) -- image-to-video generation up to 720p, available as an API preview. **Composer 2.5** (6/1) -- xAI's 'fast, SOTA model for long-running tasks,' now selectable in the Grok Build /models menu for SuperGrok and X Premium+ subscribers (NOT related to Cursor's Composer line despite the name). **Grok Build 0.1 on the API** (5/29) -- the coding-agent model behind Grok Build became directly callable via the xAI API: 256K context, always-on reasoning, text + image input. Grok Build itself ('Introducing Grok Build,' 5/25) is in early beta for ALL SuperGrok and X Premium+ subscribers -- broader than the original Heavy-tier-only gate. Also: Grok voice now powers Vapi (6/3) and Gopuff's 'Go' shopping agent (6/9). NOTE: 'Grok 5' / 'V9-Medium mid-June' claims remain aggregator-only with zero vendor signal -- not real until x.ai posts itSource: xAI news (x.ai/news/grok-imagine-1-5, x.ai/news/composer-2-5, x.ai/news/grok-build-0-1, x.ai/news/grok-build-cli), x.ai/build/changelog · 2026-06-03
- PRODUCT (2026-05-18): xAI shipped **Grok Skills** -- a persistent-memory Skills layer on Grok 4.3. Skills are user-defined named capabilities Grok carries across sessions on web / iOS / Android (recipe collection, code-review checklist, study-habits coach, etc.). Each Skill is a stored prompt + behavioral pattern Grok consults when invoked by name. Per-user storage; not shared across accounts. Differentiates Grok from ChatGPT Memory (passive recall) toward configurable named tools. Pairs with the 5/14 Grok Build CLI ship -- Skills are the consumer-facing persistent-state layer, Build CLI is the developer-facing one. Material in the 'agent goes where you go' competitive narrative alongside Codex on mobile (5/14) + Cursor Jira integration (5/19) + Devin Windows VMs (5/21).Source: xAI news (x.ai/news), xAI release notes (docs.x.ai/developers/release-notes) · 2026-05-18
- MODEL LINEUP CONSOLIDATION (2026-05-15, went live 12:00 PT): xAI auto-redirected **8 deprecated model slugs** to grok-4.3 (or grok-imagine-image-quality for the image model). Affected slugs: grok-4-1-fast-reasoning, grok-4-1-fast-non-reasoning, grok-4-fast-reasoning, grok-4-fast-non-reasoning, grok-4-0709, grok-code-fast-1, grok-3, grok-imagine-image-pro. All requests now silently bill at grok-4.3 rates ($1.25 input / $2.50 output per 1M tokens). Anyone with these slugs pinned in production or referenced inside a Copilot/Cursor/Codex multi-model selector now pays the new rate without any code change. Migration path: explicitly switch to grok-4.3 in your model selector and audit token-spend after 5/15 since the new rate may differ from what each deprecated slug was previously billed at. The grok-code-fast-1 slug retirement is the same event that took the model off GitHub Copilot's Chat/inline/agent surfaces on 5/15Source: xAI docs (docs.x.ai/developers/migration/may-15-retirement) · 2026-05-15
- PRODUCT (2026-05-14): xAI launched **Grok Build CLI** in early beta -- an agentic terminal-native CLI for coding, app development, and workflow automation. Spawns up to **8 concurrent agents** in parallel. Powered by Grok 4.3 beta with a 16-agent Heavy architecture and **2M token context window**. Vendor-primary launch posts at x.ai/news/grok-build-cli and x.ai/cli, plus Musk's public invitation to wider beta testers on X. **Access gate**: launched first to SuperGrok Heavy tier ($299/mo, intro offer $99/mo for 6 months) -- not yet available to standard Premium / SuperGrok subscribers. Positions Grok as a direct competitor to Claude Code, Codex CLI, and Cursor CLI for terminal-first agentic coding workflows. The 8-agent parallelism + 2M context is the differentiating feature -- single longest context window of any production coding CLI as of todaySource: xAI news (x.ai/news/grok-build-cli), xAI product page (x.ai/cli), Musk on X · 2026-05-14
- xAI joined SpaceX on 2026-02-02 -- SpaceX acquired xAI. Procurement, billing, and compliance workflows now route through SpaceX's vendor pipeline. For regulated industries (healthcare, finance, US government) this may require re-qualifying xAI as a vendor even if Grok itself was previously approvedSource: xAI announcement (x.ai/news/xai-joins-spacex), SpaceX updates · 2026-02
- Grok Speech (STT + TTS) APIs launched 2026-04-17 as separate products from the chatbot -- see /tools/grok-voice on this site. Built on the same stack Grok Voice uses. Not included in Premium/SuperGrok consumer tiers; billed separately at $0.10/hr STT batch and $4.20/1M char TTSSource: xAI Grok STT/TTS announcement · 2026-04
- Real-time X data can surface misinformation from viral posts without adequate fact-checkingSource: Reddit r/artificial · 2026-02
- Free tier rate limits are aggressive -- many users report hitting caps within a few queriesSource: X/Twitter user reports · 2026-03
- Grok 4.20's 4-agent system (Grok, Harper, Benjamin, Lucas) can take 30+ seconds for complex queries as agents debate internally. Grok 4.20 Beta 2 (landed ~2026-04-07) improved instruction-following, reduced hallucinations, better LaTeX and image search -- partially addresses the slowness and reliability complaints from early 4.20 feedbackSource: Reddit r/grok, IBTimes · 2026-04
- PRODUCTION LAUNCH (2026-05-02): Grok 4.3 went broadly available beyond the SuperGrok Heavy beta. New consumer + API features: **Custom Voices voice cloning suite** (clone voice from ~1 minute of speech in <2 minutes, two-stage passphrase + speaker-embedding consent gate, 80+ preset voices, 28 languages, free on console); **Imagine Agent Mode** (creative production workflow agent, beta); native video input + reasoning-by-default; native PDF / PowerPoint / spreadsheet output. **API pricing: $1.25 input / $2.50 output per 1M tokens** -- ~40% input cut + ~60% output cut vs Grok 4.20. 1M context window. Reasoning tokens billed at output rate. xAI's pattern is silent ship via grok.com model selector + console UI rather than vendor blog post -- vendor-primary verification through grok.com itself plus 4+ tier-1 press sources (VentureBeat, Winbuzzer, The Decoder, Phemex)Source: VentureBeat (venturebeat.com/technology/xai-launches-grok-4-3-at-an-aggressively-low-price-and-a-new-fast-powerful-voice-cloning-suite), Winbuzzer 2026-05-03, The Decoder, grok.com console · 2026-05-02
- Grok 4.3 Beta dropped 2026-04-17 as a SuperGrok Heavy exclusive ($300/mo tier). Elon Musk clarified on 2026-04-18 that the live checkpoint is ~0.5T params; the full 1T version is ~5 days from finishing training. Beta gating ENDED 2026-05-02 with broader rollout (see entry above)Source: PiunikaWeb, BuildFastWithAI, xAI release notes, Musk posts on X (2026-04-18) · 2026-04
Best for
People who live on X/Twitter and want an AI that can tap into that data in real-time. Also good for users who find mainstream chatbots too sanitized and want something with more personality.
Not for
Enterprise users who need reliable, consistent outputs. Also not the best pick if you don't use X -- the real-time data advantage disappears and you're left with a solid-but-not-best-in-class LLM.
Our Verdict
Grok has come a long way from being dismissed as Elon's pet project. As of Grok 4.7 (2026-09-21) xAI ships a frontier-class coding and knowledge-work model at $2/$6 per 1M -- far below OpenAI and Anthropic list prices -- and places it in Cursor, Grok Build and GitHub Copilot on launch day, which makes the API the strongest part of the story. The consumer side is less clear-cut: at $30/mo for SuperGrok you are paying a premium for personality and real-time X data, and xAI still publishes its benchmark comparisons as charts rather than tables, so independent verification lags every launch. If real-time X data or cheap frontier tokens matter to you, Grok is a serious pick. If not, Claude or ChatGPT still give you more polish for the money.
Sources
- SpaceXAI: Team Bots -- AI coworkers that learn from your team (2026-09-28) -- shared Grok Bots, private per-user memories, public beta on Teams and Enterprise (accessed 2026-09-28)
- SpaceXAI: How SpaceXAI is using Grok Bot to scale customer support (2026-09-22) -- 175% ticket growth, no new hires, $0.20-$0.30 per resolution (accessed 2026-09-28)
- SpaceXAI: Introducing Grok 4.7 (2026-09-21) -- same $2/$6 as 4.6, new base model, chart-only benchmarks, LatchBio 62.4%, HackerBench 3.3% (accessed 2026-09-21)
- xAI docs: Models and Pricing -- grok-4.7 $2/$0.50/$6, long-context $4/$1/$12 above 200k, Grok 4.7 Fast (Cursor + Grok Build only), US regional 1.1x (verified 2026-09-21) (accessed 2026-09-21)
- xAI docs: Models -- Grok 4.7 500k context, May 2026 knowledge cutoff, encrypted reasoning on Responses API (page 'Last updated: September 21, 2026') (accessed 2026-09-21)
- GitHub Changelog: Grok 4.7 is now available in GitHub Copilot (2026-09-21) -- all five paid SKUs, auto-enabled under global model policy (accessed 2026-09-21)
- Cursor blog: Introducing Grok 4.7 (2026-09-21, corroborates day-one Cursor availability) (accessed 2026-09-21)
- SpaceXAI: Memory in Grok Build (2026-09-16) -- background capture, /memory and /dream (accessed 2026-09-21)
- SpaceXAI: Grok Bot for Enterprise (2026-09-03) -- access/network/audit controls; free two weeks for Grok and Cursor Enterprise customers (accessed 2026-09-05)
- SpaceXAI: Setting Grok Bot loose on procurement (2026-09-04) -- vendor-run internal trial claiming $100k+ savings (accessed 2026-09-05)
- xAI: Grok Bot now works with X -- X connector, auto-provisioned developer account, free X API credits for paid users (2026-08-29) (accessed 2026-08-31)
- xAI: Grok 4.6 on Gemini Enterprise Agent Platform (2026-08-21) (accessed 2026-08-28)
- xAI: Grok 4.6 on Microsoft Foundry (2026-08-26) (accessed 2026-08-28)
- xAI: Grok Bot is now included with more plans (2026-08-26) (accessed 2026-08-28)
- SpaceXAI: Grok 4.6 on Amazon Bedrock -- GA, $2/$6 per 1M, 500K context (2026-08-19) (accessed 2026-08-20)
- SpaceXAI: Grok 4.6 in GitHub Copilot (2026-08-14) -- console pricing restated at $2/$6 per 1M (accessed 2026-08-17)
- GitHub changelog: Grok 4.6 is now available in GitHub Copilot (2026-08-14) -- all paid SKUs, policy off by default (accessed 2026-08-17)
- SpaceXAI: Introducing Grok 4.6 -- $2/$6, full eval table, Cursor + Grok Build day one (2026-08-12) (accessed 2026-08-13)
- Cursor blog: Grok 4.6 (corroborates launch date, pricing and 2x first-week usage) (accessed 2026-08-13)
- SpaceXAI: Introducing Grok Bot -- always-on agents, SuperGrok Heavy + Cursor Ultra/Teams Premium (2026-08-11) (accessed 2026-08-13)
- SpaceXAI: Imagine Image 2.0 -- precise editing, multi-ref, smart resize, #2 on both Arena image boards (2026-08-07) (accessed 2026-08-10)
- SpaceXAI: Imagine Video 1.5 with References -- multi-reference, voice consistency, text-to-video, native 1080p (2026-07-31) (accessed 2026-08-06)
- SpaceXAI: Grok Automations (2026-07-16) (accessed 2026-07-18)
- SpaceXAI: Grok Build open source (2026-07-15) (accessed 2026-07-18)
- SpaceXAI: Introducing Grok 4.5 (2026-07-08) (accessed 2026-07-09)
- Cursor blog: Introducing Grok 4.5 (joint training details) (accessed 2026-07-09)
- Axios: SpaceXAI launches new model, Grok 4.5 (accessed 2026-07-09)
- xAI May 15 model retirement docs (accessed 2026-05-19)
- VentureBeat: xAI launches Grok 4.3 with voice cloning (2026-05-02) (accessed 2026-05-05)
- Winbuzzer: xAI Grok 4.3 + Custom Voices (2026-05-03) (accessed 2026-05-05)
- xAI official site (accessed 2026-04-17)
- xAI Grok 4.20 announcement (accessed 2026-04-17)
- IBTimes: Grok 4.20 Beta 2 April 2026 (accessed 2026-04-17)
- BuildFastWithAI: Grok 4.3 Beta 2026-04-17 (accessed 2026-04-17)
- Artificial Analysis: Grok 4.20 (accessed 2026-04-17)
- Reddit r/grok, r/artificial (accessed 2026-04-17)
Explore more Grok rankings
Deeper leaderboards, benchmarks, task-specific tier lists, and status/pricing pages for Grok.
The Tier List Tuesday
Weekly newsletter: tier movers, new entrants, and the VS of the week. Built from our daily AI-tool sweeps. No spam, unsubscribe anytime.
Alternatives to Grok
Claude (Anthropic)
**Two Claude 5.5 models in six days: Claude Opus 5.5 (2026-09-22) and Claude Sonnet 5.5 (2026-09-28).** Opus 5.5 performs at Fable 5.1 level on most work and costs 40% less than Opus 5 to run -- $4/$20 per 1M (down from $5/$25) with cache reads cut to $0.20 (from $0.50), fast mode at $8/$40 up to 2.5x speed, output 30%+ faster, and five-hour usage limits raised on Pro, Max, Team and seat-based Enterprise with a saveable rate-limit reset. Sonnet 5.5 keeps Sonnet 5's $2/$10 but needs far fewer tokens (up to 30% cheaper per task), runs 30%+ faster and jumps Terminal-Bench 4.0 from 10.3% to 70.6%. Both are 1M context, Jun 2026 cutoff, on the Claude Platform, Bedrock, Vertex and Foundry; Haiku 5.5 follows in the coming weeks. Opus 5.5 launches with Fable-class cyber and bio safeguards (most cyber tasks fall back to Opus 4.8) and cannot run with thinking off.
Claude Mythos 5.1
Anthropic's trusted-access frontier model. **Mythos 5.1 launched 2026-09-01** alongside Fable 5.1, and Anthropic now states outright that they are **the same model with different safeguards** -- the 5.1-cycle gap is 60.9% vs 55.8% on Terminal-Bench 4.0, which Anthropic attributes to safeguard interventions rather than capability and expects to shrink. Originally launched June 9, 2026 alongside Claude Fable 5. Suspended June 12 by a US export-control order, then PARTIALLY RESTORED July 1, 2026 (US government lifted controls June 30): Mythos 5 is back for a set of US organizations with government approval, while Anthropic works to re-expand the broader Glasswing program. Public Fable 5 returned globally the same day. Gated to Project Glasswing orgs + select biology researchers.
Gemini (Google)
**Gemini 3.8 Flash TTS and 3.8 Flash-Lite TTS launched 2026-09-23 with voice design from a text prompt** -- describe a role, accent and character in more than 100 languages and get a new voice, replicate your own from a 30-second sample with consent verification, direct delivery line by line with tags like <laughs> and |mhm|, stage two-speaker scenes -- at $0.50 text in / $9.00 audio out per 1M for Flash TTS (about $0.00225 per 10 seconds) and $0.50 / $6.00 for Flash-Lite, intro rates that double on 2027-01-01. Google places Flash TTS #1 on Hume AI's Voice Design Benchmark (71.4) and #1/#2 on its Overall Quality Index. Same week: **Gemini 3.8 Live with Live Avatar (9/24)** puts a lip-synced, 97-language video persona on the Live model inside Gemini Enterprise; **Gemini Omni 1.1 Flash makes free 1080p video for any Google account in Google Vids (9/23)**; and a wave of Connected Apps (Airtable, Linear, monday.com, Adobe, Squarespace, Webflow, Experian, Peloton, SeatGeek) rolled into the Gemini app (9/23). Googlebook ships 10/04.
Muse Spark (Meta)
Meta's frontier model line from its Superintelligence Lab -- Muse Spark 1.1 (2026-07-09) adds substantially better coding, 1M-token multi-agent orchestration, and Meta's first paid developer API (Meta Model API, public preview)
GPT-Rosalind (OpenAI)
OpenAI's first domain-specific model -- life sciences, drug discovery, translational medicine. Launched 2026-04-16 as a Trusted Access research preview. Launch partners: Amgen, Moderna, Allen Institute, Thermo Fisher. Paired with a Life Sciences Codex plugin (50+ scientific tool integrations)
GPT-5.6-Cyber / GPT-5.4-Cyber (OpenAI)
**OpenAI is retiring GPT-5.4-Cyber (notice 2026-09-11, removed from the API 2026-10-01); its replacement is `gpt-5.6-cyber`**, an alias for OpenAI's most advanced purpose-trained cyber models, gated behind the Daybreak program and priced at $12.50/$75 per 1M (first published price for any OpenAI cyber model). Original page: OpenAI's defensive-cybersecurity variant of GPT-5.4, launched 2026-04-16. Lowered refusal boundary for security-research tasks and native binary reverse-engineering. Access gated via Trusted Access for Cyber (TAC) program -- thousands of verified defenders, hundreds of teams, no public pricing. On **2026-08-17 OpenAI published its first dedicated post on the Hugging Face incident**, conceding it 'underestimated the real-world cyber capabilities of our AI models' and confirming it now releases cyber capabilities only to trusted defenders. **On 2026-09-01 OpenAI confirmed GPT-6 Astra meets the Critical cyber threshold -- the first model it has ever designated at that level -- and on 2026-09-03 committed $1B to Daybreak for Frontline Defenders**
Microsoft MAI-Thinking-1
Microsoft's first in-house reasoning model -- launched 2026-06-02 at Build as the flagship of seven new MAI models. 35B-active / ~1T-total sparse Mixture-of-Experts, 256K context. AIME 2025 97.0%, matches leading models on SWE-Bench Pro, and beat Claude Sonnet 4.6 in human-preference testing. Available on Microsoft Foundry + OpenRouter / Fireworks / Baseten
Hunyuan 3 (Tencent Hy3)
Tencent's Hy3 reached GA 2026-07-06 (upgraded from the April preview) -- 295B total / 21B active MoE, 256K context, now Apache 2.0 open weights on HuggingFace + ModelScope with the EU/UK/South Korea restriction lifted. ~90% agent-task completion on Tencent's internal apps; API via Tencent Cloud TokenHub. Integrated into Yuanbao, WeChat, QQ
MiMo (Xiaomi)
Xiaomi's MiMo-V2.5 family launched 2026-04-22 -- Pro (1T total / 42B active MoE, 1M context, native vision+audio reasoning), Multimodal base, TTS (3 sub-models: base, VoiceDesign, VoiceClone), and ASR (open-source, English + Chinese + major dialects). Full voice pipeline for the agent era. Extra-charge 1M-context tier removed at launch