Google route

gemini-2.5-flash-lite

Gemini 2.5 Flash Lite API pricing.

The cheapest Gemini route here: $0.10 input and $0.40 output per 1M tokens, 1M context. Use it for high-volume classification and extraction, not as a stand-in for 3.8 Flash.

Route status Available

#3

Cost rank / 47

budget

Cost band

Published route economicsUSD / 1M tokens

Input / 1M tokens

$0.10

Output / 1M tokens

$0.40

Context window

1M

Provider

Google

Workload calculator

Put the token rate into context.

These examples apply the published input and output rates directly. Actual cost follows the exact tokens your client sends and the model returns.

This route is the 3rd cheapest of 47 comparable text models and cheaper than 44 on the same key. Its context window ranks 2nd of 47.

Cost simulationPublished base rate

Short request

$0.0014

10K input + 1K output

Agent turn

$0.014

100K input + 10K output

Exact token billing

No request minimum

One model ID change

OpenAI compatible

Cost neighbourhood

Nearby routes, same workload.

Every comparison uses the same 100K-input / 10K-output turn. Changing route is a one-line model ID edit.

Price does not measure task quality. Check Gemini 2.5 Flash Lite benchmark evidence before optimizing on cost alone. For generation routes, see GPT Image 2.

First request

Route it in one call.

Use any OpenAI SDK or tool. Keep the request shape, point the base URL at RouterPlex, and set the model ID shown here.

Read the quickstart
curl /chat/completions
curl https://api.routerplex.com/v1/chat/completions \
-H "Authorization: Bearer $ROUTERPLEX_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-2.5-flash-lite",
"messages": [{"role": "user", "content": "Hello!"}]
}'

Route notes

Questions before production traffic.

Same endpoint, every route
01How much does the Gemini 2.5 Flash Lite API cost?+

Through RouterPlex, Gemini 2.5 Flash Lite costs $0.10 per 1M input tokens and $0.40 per 1M output tokens — the vendor list price with 0% markup, billed per token from a prepaid balance you top up from $5.

02What does one Gemini 2.5 Flash Lite agent turn cost?+

A 100K-token input with a 10K-token reply costs about $0.014 on Gemini 2.5 Flash Lite. That makes it the 3rd cheapest of 47 comparable text models on RouterPlex, roughly 30% more the cost of the cheapest option (DeepSeek V4 Flash).

03What is a cheaper alternative to Gemini 2.5 Flash Lite?+

MiMo V2.5 from Xiaomi is the closest cheaper model on RouterPlex — $0.014 per agent turn against $0.014 for Gemini 2.5 Flash Lite, at $0.11/1M input and $0.28/1M output. Both run on the same key and the same endpoint, so switching is a one-line model ID change.

04Is Gemini 2.5 Flash Lite cheaper than Hy3?+

Yes. Per agent turn, Gemini 2.5 Flash Lite costs about $0.014 against $0.028 for Hy3 — 50% less.

05What is the context window of Gemini 2.5 Flash Lite?+

Gemini 2.5 Flash Lite supports a 1M token context window on RouterPlex, the 2nd largest of the 47 models that publish one.

06Can I use Gemini 2.5 Flash Lite with the OpenAI SDK?+

Yes. RouterPlex serves Gemini 2.5 Flash Lite through an OpenAI-compatible endpoint at https://api.routerplex.com/v1 — point any OpenAI SDK or tool at that base URL with your RouterPlex key and set model to "gemini-2.5-flash-lite".

Deploy this route

Put Gemini 2.5 Flash Lite behind one key.

Top up from $5 with no subscription required.

Gemini 2.5 Flash Lite API Pricing · RouterPlex