Gemini 3.1 Pro API Pricing Calculator
Compare what your prompts cost across every major LLM — with exact token counts for OpenAI models, right in your browser.
Prices verified 2026-07-24 · per 1M tokens · sources linked per provider
Input / 1M tokens
$2
Cached input / 1M
$0.2
Output / 1M tokens
$12
gemini-3.1-pro-preview>200k-token prompts: $4 in / $18 out.
At Gemini 3.1 Pro's rates ($2 per 1M input tokens, $12 per 1M output), a call with 1,000 input and 500 output tokens costs $0.008. At 10,000 requests a month, that workload runs $80.00. A heavier call — a 100,000-token prompt returning 2,000 tokens — costs $0.224.
Gemini 3.1 Pro is the priciest Google Gemini model at this workload, 26.7× the per-call cost of Gemini 2.5 Flash-Lite.
Cache-read input is billed at $0.2 per 1M tokens — 90% below the standard input rate, so a long system prompt reused across calls mostly bills at the discounted rate.
>200k-token prompts: $4 in / $18 out.
| Model | Cost / call | Cost / month | × best price |
|---|---|---|---|
| Gemini 2.5 Flash-Litecheapest | $0.0005 | $0.5 | — |
| Gemini 3.5 Flash-Lite | $0.0028 | $2.80 | 5.6× |
| Gemini 2.5 Flash | $0.0028 | $2.80 | 5.6× |
| Gemini 3.6 Flash | $0.009 | $9.00 | 18.0× |
| Gemini 3.5 Flash | $0.0105 | $10.50 | 21.0× |
| Gemini 2.5 Pro | $0.0112 | $11.25 | 22.5× |
| Gemini 3.1 Pro | $0.014 | $14.00 | 28.0× |
More Google Gemini models: Gemini 3.6 Flash · Gemini 3.5 Flash · Gemini 3.5 Flash-Lite · Gemini 2.5 Pro · Gemini 2.5 Flash · Gemini 2.5 Flash-Lite — or all Google Gemini pricing
Token counter
Token counts are exact for OpenAI models (tiktoken o200k_base, run in your browser). For other providers the count is an estimate — notably Claude Opus 4.7+/Sonnet 5 and similar newer models use a tokenizer that produces roughly 30% more tokens for the same text.
Related tools
Scaling AI content or ops? Elegant Atomics builds growth systems for B2B SaaS — Trevor’s company. Work with us →