Skip to main content
← All Inference APIs

Gemini 2.0 Flash (Standard)

gemini
gemini-2.0-flash
Retired Aug 10, 2026

Provider: Google Gemini

Deprecated August 10, 2026.

Gemini 2.0 Flash is deprecated and was shut down June 1, 2026.

Sticker prices & cost-per-token

List rates as published by Google Gemini, before any volume or enterprise discounts.

Input cost

$0.10

Output cost

$0.40

Batch input / output

$0.05 / $0.20

Cached input

$0.01

Get this data free

Want to stay updated on price changes on this model? Pull this data free (no API key required) here:

curl "https://coolhandlabs.com/api/v2/inference_apis?q%5Bmodel_eq%5D=gemini-2.0-flash&q%5Bsource_api_eq%5D=gemini"

Full REST API docs →

Pricing History

Effective Price
Feb 23, 2026 – present Input cost: $0.10 / 1M, Output cost: $0.40 / 1M, Batch input cost: $0.05 / 1M, Batch output cost: $0.20 / 1M, Cached input cost: $0.01 / 1M, Cache creation cost: $0.00 / 1M, 5-minute cache write cost: $0.00 / 1M, 1-hour cache write cost: $0.00 / 1M, Reasoning output cost: $0.00 / 1M, Cache storage cost: $0.00 / 1M/hr

History

Last change detected August 10, 2026 (about 1 month ago) by an automated scan.

No history recorded yet.

Source

https://ai.google.dev/pricing

Comparable models

No comparable models identified yet.

Reliability

No reliability incidents recorded in the last 90 days.

Frequently asked questions

Track LLM pricing and deprecations automatically.

Start free