Skip to main content
← All Inference APIs

Gemini 3.1 Flash Lite Preview

vertex
gemini-3.1-flash-lite-preview
Retired Aug 13, 2026

Provider: google

Deprecated August 13, 2026.

Listed as deprecated on models.dev as of 2026-08-13.

Sticker prices & cost-per-token

List rates as published by google, before any volume or enterprise discounts.

Input cost

$0.25

Output cost

$1.50

Cached input

$0.02

Get this data free

Want to stay updated on price changes on this model? Pull this data free (no API key required) here:

curl "https://coolhandlabs.com/api/v2/inference_apis?q%5Bmodel_eq%5D=gemini-3.1-flash-lite-preview&q%5Bsource_api_eq%5D=vertex"

Full REST API docs →

Pricing History

Effective Price
Aug 16, 2026 – present Input cost: $0.25 / 1M, Output cost: $1.50 / 1M, Cached input cost: $0.02 / 1M, Cache creation cost: $0.00 / 1M, 5-minute cache write cost: $0.00 / 1M, 1-hour cache write cost: $0.00 / 1M, Reasoning output cost: $0.00 / 1M, Cache storage cost: $0.00 / 1M/hr

History

No automated scan changes recorded yet.

Date Changed by What changed
Aug 16, 2026 Unknown Initial record created

Source

https://models.dev/api.json (models.dev provider 'google-vertex', bulk import 2026-08-13)

Comparable models

No comparable models identified yet.

Reliability

No reliability incidents recorded in the last 90 days.

Frequently asked questions

Track LLM pricing and deprecations automatically.

Start free