GPT-OSS 120B API Pricing Calculator
Compare what your prompts cost across every major LLM — with exact token counts for OpenAI models, right in your browser.
Prices verified 2026-07-24 · per 1M tokens · sources linked per provider
Input / 1M tokens
$0.15
Cached input / 1M
—
Output / 1M tokens
$0.6
openai/gpt-oss-120bAt GPT-OSS 120B's rates ($0.15 per 1M input tokens, $0.6 per 1M output), a call with 1,000 input and 500 output tokens costs $0.00045. At 10,000 requests a month, that workload runs $4.50. A heavier call — a 100,000-token prompt returning 2,000 tokens — costs $0.0162.
Within Groq's lineup, GPT-OSS 120B runs 5.0× the per-call cost of Llama 3.1 8B Instant and 2.2× less than Llama 3.3 70B Versatile at this workload.
GPT-OSS 120B's context window is 131,072 tokens.
| Model | Cost / call | Cost / month | × best price |
|---|---|---|---|
| Llama 3.1 8B Instantcheapest | $0.00013 | $0.13 | — |
| GPT-OSS 120B | $0.00075 | $0.75 | 5.8× |
| Llama 3.3 70B Versatile | $0.00138 | $1.38 | 10.6× |
More Groq models: Llama 3.3 70B Versatile · Llama 3.1 8B Instant — or all Groq pricing
Token counter
Token counts are exact for OpenAI models (tiktoken o200k_base, run in your browser). For other providers the count is an estimate — notably Claude Opus 4.7+/Sonnet 5 and similar newer models use a tokenizer that produces roughly 30% more tokens for the same text.
Related tools
Scaling AI content or ops? Elegant Atomics builds growth systems for B2B SaaS — Trevor’s company. Work with us →