BREAKING
+++ Grok 4.6: 500K context and aggressive pricing for AI agents +++ OpenAI Ultrafast: GPT-5.6 Sol up to 14x faster +++ IBM brings OpenAI models into the enterprise engine room +++ Anthropic in $6 billion talks to acquire Decart +++ Microsoft merges Copilot apps into one super app +++ Anthropic warns: AI agents descend into turf wars ++++++ Grok 4.6: 500K context and aggressive pricing for AI agents +++ OpenAI Ultrafast: GPT-5.6 Sol up to 14x faster +++ IBM brings OpenAI models into the enterprise engine room +++ Anthropic in $6 billion talks to acquire Decart +++ Microsoft merges Copilot apps into one super app +++ Anthropic warns: AI agents descend into turf wars +++
Updated 08:00
AI IN LIFE AI IN LIFENEWS
DAILY
MODELS

Grok 4.6: SpaceXAI ships agent model with 500K context

Elon Musk's AI unit scores 61 on the Artificial Analysis index, charges $2 per million input tokens and targets long-running AI agents.

Grok 4.6: SpaceXAI ships agent model with 500K context

Illustration · AI-generated (AI IN LIFE)

At a glance

  • 500,000-token context window, text and image input, text output
  • 61 on the Artificial Analysis Intelligence Index — level with GPT-5.6 Sol, third place worldwide
  • DeepSWE v1.1: 65.9 percent versus 73 percent for GPT-5.6 Sol
  • Pricing: $2 input, $0.50 cached, $6 output per million tokens (under 200K; doubled above)
  • No open-weights release, no self-hosting

Elon Musk's AI unit — referred to as SpaceXAI in current coverage after xAI moved under the SpaceX umbrella — has released Grok 4.6. The model ships with a 500,000-token context window, accepts text and image input, and is explicitly tuned for long-running agent workflows, coding and knowledge-intensive work.

On benchmarks, Grok 4.6 lands near the top: it scores 61 on the Artificial Analysis Intelligence Index, matching OpenAI's GPT-5.6 Sol — enough, according to VentureBeat, for third place worldwide, ahead of Kimi K3. On the coding tests that matter most to engineering teams, a gap remains: Grok 4.6 reaches 65.9 percent on DeepSWE v1.1, versus 73 percent for GPT-5.6 Sol.

Technically this is not a bigger base model but a post-training upgrade: SpaceXAI leaned on a longer supplemental training run and reinforcement learning in agentic environments. That is where the claimed strengths come from — repository refactors, migration agents, research synthesis over large document sets, and finance and legal analysis.

The pricing is the aggressive part: $2 per million input tokens, $0.50 for cached input and $6 per million output tokens — for prompts under 200,000 tokens, with rates doubling above that threshold. The model is available via the xAI API, Cursor and Grok Build, and can be routed through OpenRouter, Vercel and Cloudflare. There are no open weights and no self-hosting option.

For European teams, Grok 4.6 mostly means one thing: fresh price pressure in the agent segment. Anyone running agent pipelines on pricier frontier models now has a serious alternative — but has to weigh the weaker coding scores and the missing self-hosting option against the cost advantage.

◈ AI-GENERATED REPORT · SOURCES LINKED

FAQ

What is technically new in Grok 4.6?

Not a larger base model but a post-training upgrade: a longer supplemental training run plus reinforcement learning in agentic environments.

Where is Grok 4.6 available?

Via the xAI API, Cursor and Grok Build; also routable through OpenRouter, Vercel and Cloudflare. No open weights are offered.

What is the model built for?

Long-running agents: repository refactors, migrations, research over large document sets, and finance and legal analysis.