Google model prices

Gemini 2.5 Flash API pricing

Gemini 2.5 Flash costs $0.30 per 1M input tokens and $2.50 per 1M output tokens on this gateway, with a 1M token context window. You reach it with the same key and the same OpenAI-compatible base URL as every other model in the catalogue.

Last updated: October 2026
Input / 1M tokens
$0.30
Output / 1M tokens
$2.50
Context window
1M
Vendor
Google

What does Gemini 2.5 Flash cost per request?

A typical request of 5,000 input tokens and 1,500 output tokens costs about $0.005 at list price. A thousand of those requests costs about $5.25.

WorkloadTokensCost at list price
Single chat turn5K in / 1.5K out$0.005
1,000 chat turns5M in / 1.5M out$5.25
Small production month10M in / 2M out$8.00
Busy production month100M in / 20M out$80.00

Model usage is drawn from your balance at the upstream list price of each model. Figures above are indicative and exclude any plan fee; the console is the source of truth for live pricing.

Where Gemini 2.5 Flash fits

Gemini 2.5 Flash offers the 1M token context window of the Pro model at $0.30 per 1M input tokens, which makes long-input workloads affordable at volume.

Pick it when

  • You are extracting fields from long documents at scale rather than reasoning over them.
  • You want a large context window on a budget for classification, tagging or summarisation.
  • You are prototyping a pipeline and want to see cost per thousand calls before switching to a stronger model.

What to watch

  • Reasoning depth is lower than the Pro tier, so keep an escalation path for the cases it gets wrong.
  • A 1M window at $0.30 per 1M tokens still costs real money at volume - measure before you assume the price is negligible.

Best for: Large-scale extraction, document summarisation and the first pass of any long-input pipeline.

How to call Gemini 2.5 Flash through the gateway

Point your client at https://api.quickrouter.homes/v1 with a QuickRouter key and pass gemini-2.5-flash as the model name. Nothing else in an OpenAI-style client changes.

curl https://api.quickrouter.homes/v1/chat/completions \
  -H "Authorization: Bearer sk-qr-your-key" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-2.5-flash",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

The console lists the model ids your key can reach - use the id exactly as it appears there.

Frequently asked questions

How much does Gemini 2.5 Flash cost?+

List price on this gateway is $0.30 per 1M input tokens and $2.50 per 1M output tokens, with a 1M token context window.

Is Flash good enough for long-document work?+

For extraction and summarisation it is usually sufficient. For reasoning across the whole document, run the Pro model on the subset that Flash flags as hard.

Can I mix Flash and Pro in one pipeline?+

Yes. Both models are reached with the same key, so switching is a model-name change rather than a second integration.

One key covers every model in this guide

Create an account, pick a plan and copy an API key. Gateway plans start at $19 per month and model usage is billed at upstream list prices.

Gemini 2.5 Flash API price: $0.30 / 1M input tokens on QuickRouter