Gemini 3.1 Pro Preview vs Grok 4.5: cost compared

Last verified

Grok 4.5 is 0% cheaper on input and 50% cheaper on output than Gemini 3.1 Pro Preview. Whether that makes it the right choice depends entirely on whether it does your task well enough — this page gives you the cost side of that trade, precisely.

Rates side by side

$ per million tokensGemini 3.1 Pro PreviewGrok 4.5
Input$2.00$2.00
Output$12.00$6.00
Cached input$0.20$0.30
Output : input ratio6.0×3.0×

What the difference is worth

Monthly cost for the same workload on each, without caching:

WorkloadGemini 3.1 Pro PreviewGrok 4.5Difference
Support chatbot
1,000 conversations/month · 2,000 in + 500 out each
$10.00$7.00$3.00 saved
RAG search
10,000 queries/month · 8,000 in + 400 out each
$208$184$24.00 saved
Bulk extraction
10,000 documents · 3,000 in + 300 out each
$96.00$78.00$18.00 saved
Coding agent
100 runs/month · 400,000 in + 20,000 out each
$104$92.00$12.00 saved

Which to use

Grok 4.5 if the task is well-specified and mechanical — extraction, classification, formatting, routine transformation. The saving is real and compounds with volume.

Gemini 3.1 Pro Preview if the task involves genuine reasoning, long-horizon planning, or work where a wrong answer is expensive to catch. Paying 1.0× more on input is trivial compared to the cost of shipping a bad result.

These are different providers, so switching means a different SDK and a different tokenizer. The same text is a different number of tokens on each, so the price difference above is approximate at the margin. Measure token counts on your real inputs before committing.

Don't skip caching

Cached input costs $0.20 on Gemini 3.1 Pro Preview and $0.30 on Grok 4.5. If your prompts share a stable prefix, caching on the more expensive model can beat switching to the cheaper one outright — worth checking before you migrate anything.

Sources

Google rates verified 2026-08-05 (source). xAI rates verified 2026-08-05 (source).