Gemini 3.6 Flash API Pricing Calculator

Compare what your prompts cost across every major LLM — with exact token counts for OpenAI models, right in your browser.

Prices verified 2026-07-24 · per 1M tokens · sources linked per provider

Input / 1M tokens

$1.5

Cached input / 1M

$0.15

Output / 1M tokens

$7.5

API model ID: gemini-3.6-flash

At Gemini 3.6 Flash's rates ($1.5 per 1M input tokens, $7.5 per 1M output), a call with 1,000 input and 500 output tokens costs $0.00525. At 10,000 requests a month, that workload runs $52.50. A heavier call — a 100,000-token prompt returning 2,000 tokens — costs $0.165.

Within Google Gemini's lineup, Gemini 3.6 Flash runs 17.5× the per-call cost of Gemini 2.5 Flash-Lite and 1.5× less than Gemini 3.1 Pro at this workload.

Cache-read input is billed at $0.15 per 1M tokens — 90% below the standard input rate, so a long system prompt reused across calls mostly bills at the discounted rate.

ModelCost / callCost / month× best price
Gemini 2.5 Flash-Litecheapest$0.0005$0.5
Gemini 3.5 Flash-Lite$0.0028$2.805.6×
Gemini 2.5 Flash$0.0028$2.805.6×
Gemini 3.6 Flash$0.009$9.0018.0×
Gemini 3.5 Flash$0.0105$10.5021.0×
Gemini 2.5 Pro$0.0112$11.2522.5×
Gemini 3.1 Pro$0.014$14.0028.0×

More Google Gemini models: Gemini 3.5 Flash · Gemini 3.5 Flash-Lite · Gemini 3.1 Pro · Gemini 2.5 Pro · Gemini 2.5 Flash · Gemini 2.5 Flash-Lite — or all Google Gemini pricing

Token counter

0 tokens

Token counts are exact for OpenAI models (tiktoken o200k_base, run in your browser). For other providers the count is an estimate — notably Claude Opus 4.7+/Sonnet 5 and similar newer models use a tokenizer that produces roughly 30% more tokens for the same text.

Related tools

Scaling AI content or ops? Elegant Atomics builds growth systems for B2B SaaS — Trevor’s company. Work with us →