Skip to content
wojciech.io
All insights
AI SystemsAIOpenAIAI Systems

Grok 4.6 costs the same as GPT-5.6 Terra on input and half on output. Its long-context line sits at 200K.

Grok 4.6 scores 44 to Terra's 42 at $2 / $6 against $2 / $12. It doubles every rate once a prompt reaches 200K tokens, and stops at 500K. Where each wins.

Wojciech Łuszczyński

Wojciech Łuszczyński

GTM Architect & Growth Operator · Now · 16 September 2026

TL;DR · Key insights

  • On Artificial Analysis's Intelligence Index v4.3, SpaceXAI's Grok 4.6 at high reasoning scores 44.41 and GPT-5.6 Terra at max effort scores 42.25. About two points to Grok.
  • Both cost $2 per million input tokens. Grok 4.6 charges $6 for output against Terra's $12, and Terra charges $0.20 for cached input against Grok's $0.50. Under 200K tokens, Terra is only cheaper when cached tokens outnumber output tokens by more than 20 to 1.
  • The long-context lines differ. Grok 4.6 doubles every rate for the whole request once a prompt reaches 200K tokens, and its window ends at 500K. Terra's surcharge starts above 272K and its window runs to 1,050,000.
  • Terra is faster, at 98.8 output tokens per second against 57.7. Grok is the cheaper default under 200K; Terra wins very cache-heavy loops, latency-sensitive work, and prompts between 200K and 272K or above 500K.

Grok 4.6 and GPT-5.6 Terra both charge $2 per million input tokens and sit two points apart on the index. Everything past those two numbers is built differently. Which one is cheaper for you depends on how much of your work is cached, how much of it is output, and how long your prompts get.

InfoGrok 4.6GPT-5.6 Terra
Input / output per 1M$2 / $6$2 / $12
Cached input per 1M$0.50$0.20
Long-context line200K prompt tokens272K input tokens
Price past the line, in / out$4 / $12$4 / $18
Context window500K1,050,000
Intelligence Index v4.344.41 (high)42.25 (max)
Output speed, tokens/s57.798.8
Blended price per 1M$1.35$1.74
Cost to run the full index$2,352$2,501

Prices, thresholds and context from SpaceXAI's model documentation and OpenAI's GPT-5.6 Terra model page. Index to two decimals from Artificial Analysis's v4.3 announcement; speed, blended price and run cost from its model pages. Read on 16 September 2026.

Two points apart, and close on total cost

Grok 4.6 at high reasoning scores 44.41 on Intelligence Index v4.3 and Terra at max effort scores 42.25. I would call that a small lead for Grok.

Running the full index cost Artificial Analysis $2,352 on Grok and $2,501 on Terra, about 6% apart. Averaged over a broad mix of tasks, the two cost about the same, but no single workload is an average, and on any one of them the price lists pull apart quickly.

Output-heavy work favours Grok

Both charge $2 per million input tokens. On output, Grok 4.6 charges $6 and Terra $12.

Work that writes a lot relative to what it reads is much cheaper on Grok. Long-form drafting and code generation are the obvious cases.

Terra needs a lot of cache to win

Cached input is where Terra fights back: $0.20 per million against Grok’s $0.50.

Input costs the same on both, so the bill turns on two differences. Grok charges $0.30 more per million cached tokens and $6 less per million output tokens. Terra only comes out cheaper when cached tokens outnumber output tokens by more than 20 to 1.

That line sits further out than it sounds. Artificial Analysis’s blended price assumes seven cached tokens for every output token, and at that mix Grok is still cheaper, $1.35 per million against $1.74.

Some agent loops do cross it. One that re-reads a 100K-token context every turn and writes a 2,000-token answer runs at 50 to 1. Per turn that comes to about 6.2 cents on Grok and 4.4 cents on Terra, before fresh input, which costs the same on both.

Reasoning pulls the ratio back down. OpenAI bills Terra’s reasoning tokens as output. xAI’s documentation says only that Grok’s are billed as part of total consumption, so check an invoice before relying on it. If they are priced as output too, the same loop thinking for 10,000 tokens a turn falls to about 8 to 1, and Grok is cheaper again.

Key takeaway

Under 200K tokens, work out one ratio before anything else on the price lists: cached tokens per output token, reasoning included. Below 20 to 1, Grok 4.6 is cheaper. Above it, Terra is.

The two long-context lines

Both vendors surcharge long prompts for the whole request. They draw the line in different places.

SpaceXAI’s documentation lists Grok 4.6 at its standard rates below 200K prompt tokens, and at $4 input, $1 cached and $12 output once the prompt reaches 200K, for every token in the request. The context window ends at 500K.

OpenAI’s Terra page puts its line at 272K input tokens, where the whole request moves to 2x input and 1.5x output: $4 / $18. The window runs to 1,050,000.

That creates four zones:

  • Under 200K: input is level at $2. Grok is cheaper on output, Terra on cache, with the 20-to-1 ratio deciding.
  • 200K to 272K: Grok has doubled to $4 / $12. Terra is still at $2 / $12. Terra is cheaper.
  • 272K to 500K: both surcharged, Grok at $4 / $12 and Terra at $4 / $18. Grok is cheaper on output.
  • Over 500K: only Terra can take the prompt.

The OpenAI side of this is in the 272K pricing cliff, and SpaceXAI’s version is the same idea with the cliff drawn lower.

Speed favours Terra

Terra generates 98.8 output tokens per second and Grok 4.6 57.7, so Terra is about 70% faster. In anything a person waits on, that difference lands on every response.

Which one I would use

The workModelWhy
Output is a real share of the bill, prompts under 200KGrok 4.6Half the output price, and two index points ahead.
Loops with over 20 cached tokens per output tokenGPT-5.6 TerraCached input at $0.20 against $0.50 outweighs the output price.
Prompts between 200K and 272K tokensGPT-5.6 TerraGrok has doubled every rate; Terra has not yet.
Prompts over 500K tokensGPT-5.6 TerraGrok 4.6's window ends at 500K.
Latency-sensitive interactive workGPT-5.6 Terra98.8 against 57.7 tokens per second.

Close on capability and total run cost. The cache ratio, prompt length and latency needs decide.

For Terra against the tiers people usually weigh it with, see Terra against Sonnet 5 and Fable 5.1 against Terra.

What would change my mind

A move of either long-context line. The 200K and 272K thresholds are pricing decisions, and they decide a whole band of workloads.

A Grok 4.6 speed improvement. Terra’s clearest advantage is being 70% faster.

Measured cost per task on a real agent loop. The 20-to-1 line is arithmetic on list prices, and the number of tokens each model actually spends on the same task can move it either way.

Questions people asked

Is Grok 4.6 better than GPT-5.6 Terra?

Slightly: 44.41 against 42.25 on the Intelligence Index v4.3. Terra is faster and has more than twice the context window.

Which is cheaper, Grok 4.6 or GPT-5.6 Terra?

Under 200K tokens, usually Grok 4.6: the same $2 input and half the output price. Terra is cheaper when cached tokens outnumber output by more than 20 to 1, and on prompts between 200K and 272K tokens.

How does Grok 4.6 price long prompts?

Below 200K prompt tokens it costs $2 / $0.50 / $6. At or above 200K, every token in the request costs $4 / $1 / $12, and the window ends at 500K. Terra’s surcharge starts above 272K.

Should I use Grok 4.6 or GPT-5.6 Terra?

Grok 4.6 for work under 200K tokens where output is a real share of the bill. Terra for loops with more than 20 cached tokens per output token and for latency-sensitive work. Terra also wins on prompts between 200K and 272K tokens, and on anything over 500K.

If you are matching a model’s price structure to your own token mix, the contact page is the fastest route to me.

About the author

Wojciech Łuszczyński

Wojciech Łuszczyński

GTM Architect and Growth Operator building AI-native revenue systems for B2B SaaS and technology companies. I connect positioning, SEO, content, paid acquisition, CRM, automation, analytics and AI workflows into practical growth infrastructure.

Newsletter

Get the next one first.

When I publish a new article on AI systems, GTM architecture, or growth operating models, you'll be the first to know.

Subscribe