Grok icon

Grok 4.6

Grok 4.6

Run long-running coding, research, and knowledge-work agents on xAI's new 500K-context frontier model with low, medium, high, or xhigh reasoning Improve over Grok 4.5 on vendor-published agentic benchmarks, including AA Intelligence 61 versus 56 and DeepSWE 65.9% versus 54% Build through the xAI API, Grok Build, Cursor, OpenRouter, Vercel, and Cloudflare at $2 input, $0.50 cached input, and $6 output per 1M tokens

Reviewed by ToolWorthy Editors·updated today·Grok 4.6 released yesterday

Pricing:Free + from $2/per 1M input tokens
Categories:
Jump to section
No media available

More tools to compare

MakersClaw icon

MakersClaw

TypingMind icon

TypingMind

Doubao icon

Doubao

Z.ai icon

Z.ai

Odysseus icon

Odysseus

LumiChats Offline icon

LumiChats Offline

Pros & Cons

Pros

  • Clear benchmark lift over Grok 4.5 across every row xAI published
  • Better suited to long-running agents, visual project generation, and multi-step coding work
  • Adds xhigh reasoning effort for harder tasks
  • Available through the xAI API, Grok Build, Cursor, and major model gateways on launch day
  • Keeps the same headline $2 input / $6 output base pricing as Grok 4.5

Cons

  • Context window remains 500K, still smaller than Grok 4.3's 1M and Grok 4.1 Fast's 2M
  • Long-context pricing doubles at 200K prompt tokens
  • Benchmark data is vendor-published and should be validated on real internal tasks
  • Fable 5 and GPT-5.6 Sol still lead several coding-heavy or terminal-heavy benchmarks
  • Tool-heavy workflows can incur extra costs beyond token usage

Overview

Grok 4.6 launched on August 12, 2026 as xAI's new frontier model for coding, agentic tasks, and knowledge work. It builds on Grok 4.5 with a stronger focus on long-running agents, multi-step research, codebase work, interactive project building, and visual application generation. The model is available through the xAI API, Grok Build, Cursor, and partner gateways including OpenRouter, Vercel, and Cloudflare.

The headline tradeoff is continuity rather than a context expansion: Grok 4.6 keeps a 500K-token context window and $2 input / $6 output base pricing, but adds xhigh reasoning, stronger vendor-published benchmark results, and better behavior on extended agent trajectories. xAI positions it as matching GPT-5.6 Sol on the Artificial Analysis Intelligence Index while still competing closely with Anthropic's Fable 5 across agentic coding and knowledge-work benchmarks.

What's New

Better Long-Running Agents

xAI describes Grok 4.6 as focused on staying with complex tasks across many steps: researching unfamiliar topics, analyzing information, working through a codebase, and turning product ideas into working applications or artifacts. Compared with Grok 4.5, xAI says it saw stronger first passes on visual and interactive projects, plus more self-testing and verification on longer trajectories.

xhigh Reasoning Effort

The API model page lists reasoning modes as low, medium, high by default, and xhigh. That gives builders another quality-latency control for difficult agentic tasks. For production workflows, the practical migration point is to test Grok 4.6 under the same reasoning effort and prompt-cache configuration as Grok 4.5 before comparing cost or latency.

Stronger Vendor-Published Benchmarks

xAI's announcement reports Grok 4.6 High at 61 on the AA Intelligence Index versus 56 for Grok 4.5 High, 65.9% on DeepSWE v1.1 versus 54%, 61.3% on FrontierCode v1.1 Extended versus 56.6%, and 57.5% on APEX-Agents versus 47.1%. It also reports GDPVal-AA v2 at 1753 versus 1526 and AA-Briefcase at 1577 versus 1313. These are vendor-published figures, so teams should treat them as directional and validate on their own workloads.

Broader Launch-Day Access

Grok 4.6 is available through the xAI API, Grok Build, Cursor, OpenRouter, Vercel, and Cloudflare. xAI also says Grok Build and Cursor include 2x usage for the first week after launch, making those paths useful for evaluation before committing significant API spend.

Performance Benchmarks

Vendor-published results from xAI's launch post:

Benchmark Grok 4.6 High Grok 4.5 High GPT-5.6 Sol Max Fable 5 Max
AA Intelligence Index 61 56 61 62
GDPVal-AA v2 1753 1526 1728 1741
CursorBench v3.2 69.9% 66.7% 67.2% 70.5%
DeepSWE v1.1 65.9% 54% 73% 70%
FrontierCode v1.1 Extended 61.3% 56.6% 60.6% 63.6%
APEX-Agents 57.5% 47.1% 56.7% 59.2%
Terminal-Bench v3.0 26% 15.7% 34.6% 34.1%
APEX-SWE 56.4% 53.6% - 58.8%
AA-Briefcase 1577 1313 1502 1574
Harvey LAB (Vals) 15.8% 12.9% 2.5% 11.3%

The most useful comparison for Grok users is Grok 4.6 versus Grok 4.5: every published row improves. Against competitors, the picture is mixed. Grok 4.6 matches GPT-5.6 Sol on AA Intelligence, beats it on several knowledge-work rows, but trails on DeepSWE and Terminal-Bench. Fable 5 still leads many coding-heavy rows.

Migration Guide

Teams moving from Grok 4.5 should start with the model ID: use grok-4.6 in xAI API calls. The context window remains 500K tokens, so prompts that already fit Grok 4.5 should not need context restructuring. Pricing also starts at the same $2 input and $6 output per 1M tokens, but cached input is now listed at $0.50 per 1M tokens and long-context rates apply once prompts reach 200K tokens.

For cost-sensitive agents, set a prompt cache key so repeated context can hit cache reliably. For quality-sensitive agents, compare high and xhigh reasoning on a fixed evaluation set before switching defaults. For Cursor or Grok Build users, the first-week 2x included usage is a low-friction way to test the model on real coding and project-building tasks.

Pricing & Plans

Access Path Price Notes
xAI API, prompts below 200K tokens $2 input / $0.50 cached input / $6 output per 1M tokens Standard Grok 4.6 pricing
xAI API, prompts at or above 200K tokens $4 input / $1 cached input / $12 output per 1M tokens Long-context pricing applies to the request
Grok Build Included usage varies xAI offered 2x included usage for the first launch week
Cursor Included usage varies by plan Available in Cursor on launch day
Partner gateways Varies by provider xAI lists OpenRouter, Vercel, and Cloudflare

The model page lists a 500K context window, text and image input, text output, no text output limit, and support for Responses API and Chat Completions. Server-side tools such as web search, X search, and code execution add separate tool-invocation costs when used.

Best For

  • Engineering teams already using Cursor or Grok Build for coding agents and project generation
  • Developers who liked Grok 4.5's cost profile but need stronger long-running agent behavior
  • Product teams turning broad application ideas into first-pass working prototypes
  • Research and analysis workflows that need a model to sustain multi-step work across tools
  • Teams comparing lower-cost Grok API pricing against OpenAI and Anthropic frontier models

FAQ

How is Grok 4.6 different from Grok 4.5?

Grok 4.6 keeps the same 500K context size and headline $2/$6 API pricing, but improves long-running agent behavior, adds xhigh reasoning, and posts stronger vendor-published benchmark results across coding, agentic, and knowledge-work evaluations.

What is the API model name?

Use grok-4.6. xAI documents it for the Responses API and Chat Completions, with text and image input and text output.

How much does Grok 4.6 cost?

Below 200K prompt tokens, xAI lists $2 per 1M input tokens, $0.50 per 1M cached input tokens, and $6 per 1M output tokens. At or above 200K prompt tokens, rates are $4, $1, and $12 respectively.

Is Grok 4.6 better than Fable 5 or GPT-5.6 Sol?

It depends on the benchmark and workload. xAI reports Grok 4.6 matching GPT-5.6 Sol on AA Intelligence and beating it on several knowledge-work rows, while Fable 5 still leads many coding-heavy rows. Treat those numbers as starting points for internal testing.

Where can I use Grok 4.6?

xAI lists Grok 4.6 as available in the xAI API, Grok Build, Cursor, OpenRouter, Vercel, and Cloudflare. Availability and limits may vary by account, plan, and provider.

Version History

Grok 4.6

Current Version

Released on August 12, 2026

+What's new
3 updates
  • Run long-running coding, research, and knowledge-work agents on xAI's new 500K-context frontier model with low, medium, high, or xhigh reasoning
  • Improve over Grok 4.5 on vendor-published agentic benchmarks, including AA Intelligence 61 versus 56 and DeepSWE 65.9% versus 54%
  • Build through the xAI API, Grok Build, Cursor, OpenRouter, Vercel, and Cloudflare at $2 input, $0.50 cached input, and $6 output per 1M tokens

Grok 4.5

Released on July 8, 2026

View Update
+What's new
3 updates
  • Run coding, agentic, and knowledge-work tasks with xAI's flagship model, served at 80 TPS and using about 4.2× fewer output tokens than Opus 4.8 on SWE-Bench Pro
  • Use Grok 4.5 as the first jointly trained xAI + Cursor model, live in Cursor across desktop, web, iOS, CLI, and SDK from day one, alongside Grok Build and the xAI API
  • Access at $2 input / $6 output per million tokens (rising to $4/$12 above 200K-token prompts) — a smaller 500K context window than Grok 4.3's 1M

Grok 4.3

Released on April 30, 2026

View Update
+What's new
3 updates
  • Use Grok 4.3 for chat, coding, tool calling, and structured outputs with 1M context, configurable reasoning, and $1.25 input / $2.50 output below 200K-token prompts
  • Cut spend by roughly 40% on input and 60% on output versus Grok 4.20 — agentic and high-volume pipelines run materially cheaper while gaining +4 points on Artificial Analysis's Intelligence Index
  • Plan long-context workloads with Grok 4.3's tiered API rates: prompts at or above 200K tokens cost $2.50 input and $5 output per million tokens

Grok 4.20

Released on March 10, 2026

+What's new
3 updates
  • Run deep-research tasks with Grok 4.20 Multi-Agent, where multiple agents work in parallel to search, analyze, and synthesize complex findings
  • Choose reasoning or non-reasoning Grok 4.20 variants for fast, precise responses with image input, agentic tool calling, and a 1-million-token context window
  • Build research workflows through the xAI API with structured outputs, function calling, web and X search, and configurable reasoning effort

Grok 4.20

Released on March 10, 2026

View Update
+What's new
3 updates
  • Run deep-research tasks with Grok 4.20 Multi-Agent, where multiple agents work in parallel to search, analyze, and synthesize complex findings
  • Choose reasoning or non-reasoning Grok 4.20 variants for fast, precise responses with image input, agentic tool calling, and a 1-million-token context window
  • Build research workflows through the xAI API with structured outputs, function calling, web and X search, and configurable reasoning effort

Grok Imagine API

Released on January 28, 2026

+What's new
2 updates
  • Generate high-quality videos from text descriptions via the xAI Imagine API — state-of-the-art performance across quality, cost, and latency benchmarks for production video generation
  • Combine video and audio generation in a single unified API, enabling developers to build rich multimedia applications without managing separate video and audio providers

Grok Business & Enterprise

Released on December 30, 2025

+What's new
2 updates
  • Deploy Grok across your organization with team management, access controls, usage monitoring, unified billing, and the highest limits on xAI's strongest models
  • Protect sensitive business conversations with enhanced data privacy guarantees and compliance-ready infrastructure designed for corporate environments

Grok 4.1 Fast

Released on November 19, 2025

+What's new
3 updates
  • Build enterprise-grade agentic systems with a 2 million token context window — the largest in any Grok model — maintaining consistent tool-calling performance across long multi-turn conversations
  • Deploy production-ready customer support, finance, and research agents using the new Agent Tools API with native web search, X data, code execution, and MCP tool integrations
  • Get frontier tool-calling intelligence at reduced cost, with Grok 4.1 Fast cutting hallucination rates in half compared to Grok 4 Fast on information-seeking tasks

Grok 4.1

Released on November 17, 2025

+What's new
3 updates
  • Engage with a more perceptive assistant for emotional, creative, and collaborative work, preferred over the previous Grok model in 64.78% of blind comparisons
  • Get dramatically more accurate answers with hallucination rates reduced by 65% (from 12.09% to 4.22%), making research, fact-finding, and information-seeking tasks significantly more reliable
  • Create more natural stories and polished prose with Grok 4.1's strong Creative Writing v3 performance, while choosing Thinking or instant non-reasoning mode

Grok 4 Fast

Released on September 19, 2025

+What's new
2 updates
  • Get Grok 4-level intelligence at a fraction of the cost — optimized for high-volume tasks like content drafting, summarization, and rapid Q&A at scale
  • Run more API calls within budget while maintaining strong reasoning performance for agentic workflows and automated pipelines

Grok Code Fast 1

Released on August 28, 2025

+What's new
2 updates
  • Accelerate agentic coding workflows with grok-code-fast-1, a reasoning model built specifically for code generation, debugging, and multi-step software engineering tasks at low cost
  • Run more coding agents in parallel within budget — Grok Code Fast 1 maintains strong agentic performance for automated pipelines, PR reviews, and code refactoring at high throughput

Grok 4

Released on July 9, 2025

+What's new
3 updates
  • Solve the hardest real-world problems with native tool use and real-time web search — Grok 4 autonomously selects search queries and dives deep into X and the broader web without plugins
  • Achieve frontier-level intelligence across math, science, and coding tasks — xAI reports Grok 4 with tools reached 38.6% on Humanity's Last Exam (HLE), a challenging PhD-level benchmark
  • Access deeper research with the new SuperGrok Heavy tier, a multi-agent configuration that coordinates parallel reasoning threads for comprehensive answers

Grok 3

Released on February 19, 2025

+What's new
3 updates
  • Reason through complex problems with Think mode, inspect the model's reasoning trace, and reach 93.3% on AIME 2025 using xAI's highest test-time compute setting (cons@64)
  • Research any topic with DeepSearch, an agentic web and X search tool that synthesizes conflicting sources into clear, cited reports automatically
  • Process up to 1 million tokens of context, enabling analysis of large codebases, research papers, and lengthy documents in a single conversation

Grok for Everyone

Released on December 12, 2024

+What's new
2 updates
  • Use Grok for free on the X platform — access AI-powered chat, web search, and multilingual support without a Premium+ subscription
  • Get faster, more accurate responses with improved speed and enhanced multilingual support rolled out to all X users globally

Grok Image Gen

Released on December 9, 2024

+What's new
2 updates
  • Generate photorealistic and artistic images directly inside Grok chats using Aurora, xAI's new autoregressive image generation model built natively into the X platform
  • Create images from text descriptions without switching apps — Aurora delivers high-fidelity visuals across photography, illustration, and abstract styles

Grok 2

Released on August 13, 2024

+What's new
3 updates
  • Outperform previous Grok models in complex reasoning, coding, and math with Grok-2's significantly upgraded intelligence and expanded tool-use capabilities
  • Process images, charts, and documents alongside text queries with Grok-2's new vision capabilities — understand visual context and answer questions about any image
  • Access a faster, lighter Grok-2 mini variant that maintains strong performance for everyday tasks at reduced cost and latency

Grok 1.5 Vision Preview

Released on April 12, 2024

+What's new
2 updates
  • Explore xAI's first multimodal Grok preview, designed to understand documents, diagrams, charts, screenshots, photographs, and real-world spatial relationships
  • Compare Grok-1.5V on RealWorldQA and standard vision benchmarks, while noting that xAI announced it as a preview that would become available later

Grok 1.5

Released on March 28, 2024

+What's new
2 updates
  • Analyze books, codebases, and lengthy research papers within a single conversation using a 128,000-token context window — 16x larger than Grok 1
  • Solve harder coding and math challenges with significantly improved reasoning: MATH benchmark jumped from 23.9% to 50.6%, GSM8K from 62.9% to 90%

Grok 1 Open Source

Released on March 17, 2024

+What's new
2 updates
  • Download Grok-1's 314-billion-parameter Mixture-of-Experts weights and architecture under Apache 2.0, noting that the release is a raw base model rather than a chat model
  • Fine-tune or deploy the base weights on your own infrastructure for domain-specific research and applications, subject to the Apache 2.0 license terms

Grok 1

Released on November 3, 2023

+What's new
2 updates
  • Chat with an AI that answers almost anything, built on a 314-billion parameter Mixture-of-Experts model trained by xAI and integrated with real-time X platform data
  • Get timely answers on breaking news, trends, and live events through direct access to X's real-time data feed, Grok's defining advantage at launch

Top alternatives

Related categories

From the blog

View all →

Track Grok in ToolWorthy Weekly

Important tool updates, better alternatives, and selected AI signals in one weekly brief.

Weekly only. Unsubscribe anytime.