Skip to main content

Grok 4.6 ties for third best, and the real price hides past 200,000 tokens

Four men sit along one side of a table in a darkened studio, watching a laptop one of them is typing on. A narrow blue line of light runs out of the laptop into a small glass lens on the table, where it re-emerges as two far broader orange beams that widen away across the surface. The man at the right of the group is depicted as Elon Musk.
01

Grok 4.6 scored 61 in August, up from Grok 4.5, which means third place.

02

From August's pricing, past 200,000 tokens a job pays double the sticker rate.

03

In August it costs $0.84 per finished task, more than the older Grok 4.5.

VentureBeat · Grok 4.6 launch coverage · Aug 12, 2026

Morning Byte · weekly digest

One email a week. The stories that mattered.

What changed

SpaceXAI released Grok 4.6 in August, only weeks on from Grok 4.5, which means frontier models now ship monthly from Elon Musk's AI arm. The launch report shows a model tuned to stay on task through long stretches of work: code, research, office jobs with less hand-holding.

On the Artificial Analysis Intelligence Index, the data shows the new model scoring 61 in August, up five points from Grok 4.5, which means it ties GPT-5.6 Sol Max for third behind Anthropic's two leaders.

61

Grok 4.6's August score on the Artificial Analysis Intelligence Index. Third place, tied with GPT-5.6 Sol Max, behind Anthropic's two leaders.

The leaderboard says third.Source: Artificial Analysis, Aug 2026
Sticker rate$2/M tokens
Past the line$4/M tokens
The meter, before and after the 200,000-token line.Source: SpaceXAI launch pricing via VentureBeat · Rate applies to the whole request once crossed

Why it matters

AI is becoming something you rent by the meter, and this launch shows how the meter really works. SpaceXAI says the sticker rate is $2 per million input tokens in August, yet past 200,000 tokens the price doubles versus the sticker, which means a long job pays the higher rate on every token.

That line sits exactly where the model's pitch lives, because long agent jobs are the work it is built to run. Artificial Analysis measured $0.84 per completed task in August, worse than the older Grok 4.5, which means the newest model is not automatically the cheapest way to finish work. Taken together, the numbers suggest the real benchmark is shifting from leaderboard scores to cost per finished job.

What to watch

If production teams report lower total bills per finished job, the pitch could hold; if long jobs keep crossing the doubling line, buyers would learn to split work or shop elsewhere.

Whether OpenAI or Anthropic answer on price is not yet known. Regulator files on the vendor remain open in the UK and Europe, and compliance teams read those files before they buy.

Not signed in yet — hit Post and we'll finish it together

Grok 4.6 ties for third best, and the real price hides past 200,000 tokens | Morning Byte