Gemini 2.5 Flash API pricing
Gemini 2.5 Flash costs $0.30 per 1M input tokens and $2.50 per 1M output tokens on this gateway, with a 1M token context window. You reach it with the same key and the same OpenAI-compatible base URL as every other model in the catalogue.
Last updated: October 2026- Input / 1M tokens
- $0.30
- Output / 1M tokens
- $2.50
- Context window
- 1M
- Vendor
What does Gemini 2.5 Flash cost per request?
A typical request of 5,000 input tokens and 1,500 output tokens costs about $0.005 at list price. A thousand of those requests costs about $5.25.
| Workload | Tokens | Cost at list price |
|---|---|---|
| Single chat turn | 5K in / 1.5K out | $0.005 |
| 1,000 chat turns | 5M in / 1.5M out | $5.25 |
| Small production month | 10M in / 2M out | $8.00 |
| Busy production month | 100M in / 20M out | $80.00 |
Model usage is drawn from your balance at the upstream list price of each model. Figures above are indicative and exclude any plan fee; the console is the source of truth for live pricing.
Where Gemini 2.5 Flash fits
Gemini 2.5 Flash offers the 1M token context window of the Pro model at $0.30 per 1M input tokens, which makes long-input workloads affordable at volume.
Pick it when
- You are extracting fields from long documents at scale rather than reasoning over them.
- You want a large context window on a budget for classification, tagging or summarisation.
- You are prototyping a pipeline and want to see cost per thousand calls before switching to a stronger model.
What to watch
- Reasoning depth is lower than the Pro tier, so keep an escalation path for the cases it gets wrong.
- A 1M window at $0.30 per 1M tokens still costs real money at volume - measure before you assume the price is negligible.
Best for: Large-scale extraction, document summarisation and the first pass of any long-input pipeline.
How to call Gemini 2.5 Flash through the gateway
Point your client at https://api.quickrouter.homes/v1 with a QuickRouter key and pass gemini-2.5-flash as the model name. Nothing else in an OpenAI-style client changes.
curl https://api.quickrouter.homes/v1/chat/completions \
-H "Authorization: Bearer sk-qr-your-key" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-2.5-flash",
"messages": [{"role": "user", "content": "Hello"}]
}'The console lists the model ids your key can reach - use the id exactly as it appears there.
Frequently asked questions
How much does Gemini 2.5 Flash cost?+
List price on this gateway is $0.30 per 1M input tokens and $2.50 per 1M output tokens, with a 1M token context window.
Is Flash good enough for long-document work?+
For extraction and summarisation it is usually sufficient. For reasoning across the whole document, run the Pro model on the subset that Flash flags as hard.
Can I mix Flash and Pro in one pipeline?+
Yes. Both models are reached with the same key, so switching is a model-name change rather than a second integration.
One key covers every model in this guide
Create an account, pick a plan and copy an API key. Gateway plans start at $19 per month and model usage is billed at upstream list prices.