Gemini 2.5 Flash Lite API Pricing: Extreme Low-Cost High-Speed Inference
Comprehensive Gemini 2.5 Flash Lite API pricing analysis ($0.10/M input, $0.40/M output), sub-80ms first-token latency, classification efficiency, and upgrade comparisons.
How much does Gemini 2.5 Flash Lite cost per million tokens?
Gemini 2.5 Flash Lite costs $0.10 per million input tokens and $0.40 per million output tokens ($0.175/M blended at 3:1). Designed for ultra-high-volume micro-tasks, classification, and real-time voice latency. Verified 2026-09-08.
How much does Gemini 2.5 Flash Lite cost per 1,000 requests?
Computed from generated token pricing. Each row assumes the listed input and output tokens per request; this model has no measured verbosity factor, so the unadjusted output estimate is shown.
| Request shape | Input tokens | Output tokens | Cost / 1,000 requests |
|---|---|---|---|
| Short | 100 | 50 | $0.0300 |
| Medium | 1,000 | 500 | $0.3000 |
| Long | 4,000 | 2,000 | $1.2000 |
Formula: ((input price × input tokens) + (output price × output tokens × verbosity factor)) ÷ 1,000,000 × 1,000. Assumptions: short 100/50, medium 1,000/500, long 4,000/2,000 input/output tokens per request. Unadjusted — no measured verbosity factor is available.
Migration delta, task-ranking cross-check, and lifecycle ledger for Gemini 2.5 Flash Lite
Successor rate delta
| Rate | Gemini 2.5 Flash Lite | Gemini 3.5 Flash Lite | Delta ($ / %) |
|---|---|---|---|
| Input $/M | $0.10 | $0.30 | $0.20 (200.0%) |
| Output $/M | $0.40 | $2.50 | $2.10 (525.0%) |
Delta = successor rate − Gemini 2.5 Flash Lite rate, both taken from the current dated pricing registry. Cache, batch, and context-tier pricing are not substituted across models; each figure is model-specific or marked Unavailable.
Where Gemini 2.5 Flash Lite still ranks across every task page
| Task | Rank / status | Task-weighted $/M | Pick status |
|---|---|---|---|
| Coding | Rank 3 of 49 | $0.16 | Dated — excluded from picks |
| Structured Data Extraction | Rank 2 of 49 | $0.16 | Dated — excluded from picks |
| Writing & Content | Rank 2 of 49 | $0.26 | Dated — excluded from picks |
| Math & Reasoning | Not a qualifying candidate | Excluded by a hard requirement filter | — |
| Agents & Tool Use | Not a qualifying candidate | Excluded by a hard requirement filter | — |
| Long Documents & RAG | Rank 2 of 40 | $0.10 | Dated — excluded from picks |
| Summarization | Rank 2 of 49 | $0.10 | Dated — excluded from picks |
| Chatbots & Support | Rank 2 of 49 | $0.20 | Dated — excluded from picks |
| Translation | Rank 2 of 49 | $0.25 | Dated — excluded from picks |
| Image Understanding | Rank 1 of 37 | $0.16 | Dated — excluded from picks |
Rank is computed live from the same {price, speed, context, evidence} formula published on each /best-llm-for page; a dated model can still rank by price/speed/context but is excluded from the "overall pick" by policy.
Lifecycle and availability ledger
| Field | Recorded value | Note |
|---|---|---|
| Lifecycle status | legacy | Verified 2026-08-14 |
| Deprecation announced | 2026-06-01 | No inference beyond the dated record |
| Shutdown date | Unavailable | Null/unavailable is not a promise of indefinite availability |
| Successor | Gemini 3.5 Flash Lite | gemini-3-5-flash-lite priced in this registry |
| Context window | 1M tokens | Verified 2026-08-14 |
Verified 2026-04-06. "Unavailable" means no compatible dated evidence was found for that field; it is never treated as zero. Dated source · Lifecycle source
gemini-2-5-flash-liteGemini 2.5 Flash Lite API Pricing: Extreme Low-Cost High-Speed Inference
Gemini 2.5 Flash Lite costs $0.10 per million input tokens and $0.40 per million output tokens ($0.175/M blended at 3:1). Designed for ultra-high-volume micro-tasks, classification, and real-time voice latency. Verified 2026-09-08.
Gemini 2.5 Flash Lite delivers industry-leading cost efficiency at $0.175/M blended tokens.
| Scenario | Rendered Evidence & Bounds |
|---|---|
| Scenario 1 | Live voice assistant dialogue turn (400 in, 80 out): $0.000072 per turn |
| Scenario 2 | Real-time text autocomplete suggestion (150 in, 20 out): $0.000023 per keystroke completion |
| Scenario 3 | Support ticket intent routing (600 in, 40 out): $0.000076 per ticket |
| Scenario 4 | Document sentiment scoring (1K in, 50 out): $0.000120 per document |
| Scenario 5 | Automated data field normalization (500 in, 60 out): $0.000074 per record |
| Scenario 6 | Monthly 100M token classification fleet: $17.50 total API infrastructure spend |
Microscopic latency and rock-bottom token pricing make it the ideal engine for voice bots.
| Scenario | Rendered Evidence & Bounds |
|---|---|
| Scenario 1 | Sub-80ms TTFT enables natural, human-like voice conversation turn-taking |
| Scenario 2 | Streaming throughput exceeding 150 tps prevents UI buffering on client devices |
| Scenario 3 | Zero-latency prompt processing handles rapid burst traffic without queue buildup |
| Scenario 4 | High-volume webhook ingestion: processes 10,000 webhooks for less than $0.01 |
| Scenario 5 | Micro-memory footprint allows high parallel connection limits on cloud gateways |
| Scenario 6 | Optimal choice for automated phone bots, live chat routing, and realtime moderation |
Upgrading to Gemini 3.5 Flash Lite reduces token spend by 62.5% with higher precision.
| Scenario | Rendered Evidence & Bounds |
|---|---|
| Scenario 1 | Gemini 3.5 Flash Lite pricing ($0.0375/M in, $0.15/M out): 62.5% lower cost across all tiers |
| Scenario 2 | Gemini 3.5 Flash Lite improves multilingual accuracy and complex instruction following |
| Scenario 3 | Migrating 100M tokens/mo saves $10.94/mo ($6.56 vs $17.50) while boosting accuracy |
| Scenario 4 | Drop-in SDK compatibility: zero code modifications needed beyond updating model ID |
| Scenario 5 | Benchmark accuracy: 3.5 Flash Lite shows 5% higher intent classification precision |
| Scenario 6 | Recommended action: safe immediate migration to Gemini 3.5 Flash Lite for lowest total spend |
How fast is Gemini 2.5 Flash Lite?
How much does Gemini 2.5 Flash Lite cost at scale?
| Tokens / month | Est. cost (blended 3:1) |
|---|---|
| 100,000 | $0.02 |
| 1,000,000 | $0.18 |
| 10,000,000 | $1.75 |
| 100,000,000 | $17.50 |
How does Gemini 2.5 Flash Lite compare with other models?
What is Gemini 2.5 Flash Lite best for?
What should you explore next for Gemini 2.5 Flash Lite?
Which Gemini 2.5 Flash Lite head-to-head comparisons are available?
What are common questions about Gemini 2.5 Flash Lite?
Is Gemini 2.5 Flash Lite cheaper than Ministral 8B?
Gemini 2.5 Flash Lite costs $0.18/M blended tokens, Ministral 8B costs $0.15/M — Ministral 8B is cheaper.
How much does 1 million tokens cost with Gemini 2.5 Flash Lite?
At a 3:1 input:output ratio, 1 million blended tokens costs approximately $0.18. Pure input costs $0.10/M; pure output costs $0.40/M.
What does Gemini 2.5 Flash Lite cost at high volume?
At 100 million blended tokens a month, Gemini 2.5 Flash Lite costs approximately $17.50. See the cost-at-scale table below for other volumes.
