google

Google: Gemini 2.5 Flash Lite

google/gemini-2.5-flash-lite

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

Prices updated: 7/28/2026

Input
20
credits / 1M tokens
Output
80
credits / 1M tokens
Context
1049k
tokens
Messages per $1 USD
≈ 1,700
typical message: 1k tokens in + 500 out

Prices in credits per million tokens. 1 credit = $0.01 USD. Final price, billed in MXN at the exchange rate on payment day + VAT (CFDI). Subject to change without notice.

Additional prices

Cache (read): 2.0 /1M
Cache (write): 17 /1M
Web search: 2,800,000 /1M
Image: 20 /1M

Spec sheet

Inputs
text, image, file, audio, video
Outputs
text
Max output tokens
65,535
Knowledge cutoff
2025-01-31
Tokenizer
Gemini
Capabilities
Text / LLM
Vision / OCR
Audio
Video
Tools
JSON
Reasoning

Use it in your code

One OpenAI-compatible endpoint. Switch models by changing one line.

curl https://api.deployalo.com/ai/gateway/v1/chat/completions \
  -H "Authorization: Bearer $DEPLOYALO_AI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-2.5-flash-lite","messages":[{"role":"user","content":"Hola"}]}'

Ready to use it?

Top up a credit pack, create your API key with a spend cap and start in minutes.