Budget a chat completion
Plug in token counts before you ship a feature that calls GPT-5.6.
Sample
Input tokens: 3,200 Output tokens: 800
What you get
The tool multiplies by your selected per‑1M token rates so you can sanity‑check invoice size.
Estimate your OpenAI API spend before you commit. This calculator loads current GPT models (including the GPT-5.6 Sol / Terra / Luna family) and per-1M token rates from OpenRouter — so when OpenAI ships a new model, it can appear here without a site rebuild. Enter input/output tokens and request volume for an instant cost breakdown. Live model lists and $/1M rates load from OpenRouter’s public models API (https://openrouter.ai/api/v1/models), with TensorFeed and a local snapshot as fallbacks. New releases (for example Gemini 3.8 Flash or Claude Fable 5.1) appear automatically when OpenRouter lists them.
OpenAI · Local registry · as of 2026-09-02 · 3 models · refreshing…
Pricing data: OpenRouter models
| Model | Input | Output | Thinking |
|---|---|---|---|
| GPT-5.6 Sol | $2.00 | $10.00 | $10.00 |
| GPT-5.6 Terra | $2.00 | $12.00 | — |
| GPT-5.6 Luna | $0.20 | $1.20 | — |
Plug in token counts before you ship a feature that calls GPT-5.6.
Sample
Input tokens: 3,200 Output tokens: 800
What you get
The tool multiplies by your selected per‑1M token rates so you can sanity‑check invoice size.
Estimate costs for Claude Fable 5.1 and other current Anthropic models with live pricing.
Calculate costs for Gemini 3.8 Flash and other current Google models with live pricing.
Count tokens and estimate API costs across major LLMs — models update from OpenRouter live pricing.
Estimate GPU and cloud costs for self-hosting Llama models.
Compare RAG and fine-tuning costs to find the optimal approach for your project.