Grok 4.6 for Agentic Coding and Knowledge Work
发布时间:2026-09-18 | 浏览:1
A frontier model for long-running agents and ambitious interactive and visual work.
Introducing Grok 4.6
Built for long-running agents and more ambitious interactive and visual work.
Grok 4.6 Model Card
The official model card for Grok 4.6. It documents the model's capability benchmarks and safeguard evaluations.
Cursor partners with SpaceX on model training
Cursor is partnering with SpaceX to accelerate our model training efforts.
Built for coding and knowledge work alike
Long-horizon problems Long-horizon problems
Grok 4.6 stays with complex tasks across many steps. It uses tools, checks its work, adjusts its approach, and keeps moving toward a finished result.
Ambitious engineering Ambitious engineering
Grok 4.6 works across large codebases and extended engineering projects. It researches unfamiliar systems, edits across files, runs tests, and verifies the result.
In-depth knowledge work In-depth knowledge work
Beyond code, Grok 4.6 works across documents, spreadsheets, PDFs, and other professional artifacts. Give it a multi-step project that requires research, analysis, and synthesis.
Interactive and visual work Interactive and visual work
Grok 4.6 turns broad product ideas into working first versions. It can establish an application's structure and visual language, build the core interactions, and refine the result through feedback.
Built for more than software engineering
Grok 4.6 builds on Grok 4.5 with a focus on long-running agents and ambitious interactive and visual work. It handles projects that span research, analysis, implementation, and several rounds of refinement.
It matches GPT-5.6 Sol on the Artificial Analysis Intelligence Index, a composite of nine benchmarks. Cursor subscription plans for individuals and teams include significant usage of the model.
A strong foundation
We trained Grok 4.6 jointly with SpaceXAI. A longer supplemental training run used curated model-generated data for reasoning and advanced technical concepts, high-quality engineering data, and an improved optimizer and training recipe.
We then used Grok 4.5 to regenerate supervised fine-tuning trajectories across reasoning efforts, agent harnesses, STEM, software engineering, and knowledge work. Reinforcement learning covered general coding, knowledge work, kernel optimization, web development, computer-aided design, and other agentic environments.
Third-party model scores are the best self-reported or publicly available results.
(Above) Grok 4.6 results across agentic coding and knowledge work benchmarks.
(Left) Grok 4.6 results across agentic coding and knowledge work benchmarks.
Available everywhere you work
Grok 4.6 is available today in Cursor across desktop, web, iOS, CLI, and our SDK.
Manual to agentic coding, in one familiar editor.
Run agents in any terminal, script, or editor.
Spawn cloud agents on the go from your browser or phone.
Start agents from Slack, GitHub, Linear, JetBrains IDEs, and more.
What is Grok 4.6? ↓ ↑
Where can I use Grok 4.6? ↓ ↑
How much does Grok 4.6 cost? ↓ ↑
Try Grok 4.6 now.
A frontier model for long-running agents and ambitious interactive and visual work.
Introducing Grok 4.6
Built for long-running agents and more ambitious interactive and visual work.
Grok 4.6 Model Card
The official model card for Grok 4.6. It documents the model's capability benchmarks and safeguard evaluations.
Cursor partners with SpaceX on model training
Cursor is partnering with SpaceX to accelerate our model training efforts.
Built for coding and knowledge work alike
Long-horizon problems Long-horizon problems
Grok 4.6 stays with complex tasks across many steps. It uses tools, checks its work, adjusts its approach, and keeps moving toward a finished result.
Ambitious engineering Ambitious engineering
Grok 4.6 works across large codebases and extended engineering projects. It researches unfamiliar systems, edits across files, runs tests, and verifies the result.
In-depth knowledge work In-depth knowledge work
Beyond code, Grok 4.6 works across documents, spreadsheets, PDFs, and other professional artifacts. Give it a multi-step project that requires research, analysis, and synthesis.
Interactive and visual work Interactive and visual work
Grok 4.6 turns broad product ideas into working first versions. It can establish an application's structure and visual language, build the core interactions, and refine the result through feedback.
Built for more than software engineering
Grok 4.6 builds on Grok 4.5 with a focus on long-running agents and ambitious interactive and visual work. It handles projects that span research, analysis, implementation, and several rounds of refinement.
It matches GPT-5.6 Sol on the Artificial Analysis Intelligence Index, a composite of nine benchmarks. Cursor subscription plans for individuals and teams include significant usage of the model.
A strong foundation
We trained Grok 4.6 jointly with SpaceXAI. A longer supplemental training run used curated model-generated data for reasoning and advanced technical concepts, high-quality engineering data, and an improved optimizer and training recipe.
We then used Grok 4.5 to regenerate supervised fine-tuning trajectories across reasoning efforts, agent harnesses, STEM, software engineering, and knowledge work. Reinforcement learning covered general coding, knowledge work, kernel optimization, web development, computer-aided design, and other agentic environments.
Third-party model scores are the best self-reported or publicly available results.
(Above) Grok 4.6 results across agentic coding and knowledge work benchmarks.
(Left) Grok 4.6 results across agentic coding and knowledge work benchmarks.
Available everywhere you work
Grok 4.6 is available today in Cursor across desktop, web, iOS, CLI, and our SDK.
Manual to agentic coding, in one familiar editor.
Run agents in any terminal, script, or editor.
Spawn cloud agents on the go from your browser or phone.
Start agents from Slack, GitHub, Linear, JetBrains IDEs, and more.
What is Grok 4.6? ↓ ↑
Where can I use Grok 4.6? ↓ ↑
How much does Grok 4.6 cost? ↓ ↑
Try Grok 4.6 now.
(Above) Grok 4.6 scores 41.4% at extra high effort on CursorBench 4.0.
(Left) Grok 4.6 scores 41.4% at extra high effort on CursorBench 4.0.
(Above) Grok 4.6 scores 41.4% at extra high effort on CursorBench 4.0.
(Left) Grok 4.6 scores 41.4% at extra high effort on CursorBench 4.0.