LLMPrice.io

Gemini 2.0 Flash Lite API pricing 001

Published on-demand rates for Google's gemini/gemini-2.0-flash-lite-001, captured 2026-08-10 from public pricing data and committed to our archive. Context window 1.05M tokens; first observed in our archive 2026-02-16.

MeterUSD per 1M tokens
Input$0.075
Output$0.300
Cached input$0.019

Standard on-demand rates. Excludes batch endpoints and negotiated pricing; confirm on Google's pricing page before committing.

What it costs on a real workload

The four frozen profiles the LLMPrice indices price, per 1M tokens processed, at the rates above. Where no cache-read rate is published, the whole input bills at the standard rate.

ProfileIn / out mixCached shareCost
Retrieval800k / 200k70%$0.088
Chat500k / 500k30%$0.179
Content200k / 800k20%$0.253
Agent900k / 100k85%$0.054

Profile definitions are frozen and published on the methodology page. Compare every model on these profiles →

Observed rate history

No repricing observed since this model entered our archive on 2026-02-16. When a rate moves, the change appears here and on the sitewide ticker.

Priced near this one

The models closest to Gemini 2.0 Flash Lite on a retrieval workload, cheaper and more expensive alike. We publish no quality score, so this is what they cost and nothing more.

ModelProviderInputOutputvs Gemini 2.0 Flash Lite
DeepSeek V4 FlashDeepSeek$0.140$0.2803% more
Mistral Small 3.2Mistral$0.060$0.1805% less
GPT 5 NanoOpenAI$0.050$0.4007% more
Ministral 3 3BMistral$0.100$0.10013% more

Compared on the retrieval profile: 800k input, 200k output, 70% of input cached where the model offers a cache rate.

Price your own prompt on it Estimate a whole project Plain rate card Google's console →

No affiliate or referral arrangements: the console link carries no parameters and nothing on this page is paid placement. Gemini 2.0 Flash Lite is a trademark of its owner; LLMPrice.io is independent and unaffiliated.