Estimate a Gemini 3.8 Flash call
Sample
Input: 10,000 tokens Output: 500 tokens
What you get
Adjust model tier to match Google AI Studio or Vertex pricing (Gemini 3.8 Flash and nearby tiers).
Google’s Gemini lineup changes quickly. This calculator loads active Gemini models — including Gemini 3.8 Flash and nearby Flash/Pro tiers — and rates from OpenRouter, and can estimate thinking/reasoning tokens when a model is marked for reasoning. Live model lists and $/1M rates load from OpenRouter’s public models API (https://openrouter.ai/api/v1/models), with TensorFeed and a local snapshot as fallbacks. New releases (for example Gemini 3.8 Flash or Claude Fable 5.1) appear automatically when OpenRouter lists them.
Google · Local registry · as of 2026-09-02 · 5 models · refreshing…
Pricing data: OpenRouter models
| Model | Input | Output | Thinking |
|---|---|---|---|
| Gemini 3.8 Flash | $0.75 | $3.75 | $3.75 |
| Gemini 3.7 Flash | $0.75 | $3.75 | $3.75 |
| Gemini 3.5 Flash | $1.50 | $9.00 | $9.00 |
| Gemini 3.1 Pro | $2.00 | $12.00 | $12.00 |
| Gemini 3.1 Flash Lite | $0.25 | $1.50 | $1.50 |
Sample
Input: 10,000 tokens Output: 500 tokens
What you get
Adjust model tier to match Google AI Studio or Vertex pricing (Gemini 3.8 Flash and nearby tiers).
Calculate API costs for current OpenAI GPT models with live OpenRouter pricing.
Estimate costs for Claude Fable 5.1 and other current Anthropic models with live pricing.
Count tokens and estimate API costs across major LLMs — models update from OpenRouter live pricing.
Estimate GPU and cloud costs for self-hosting Llama models.
Compare RAG and fine-tuning costs to find the optimal approach for your project.