TokenRoute
Get started
Back to models

Google: Gemini 2.5 Flash Lite

google/gemini-2.5-flash-lite

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

Pricing

Input$0.0700 /M
Output$0.2800 /M
Reasoning$0.2800 /M
Cache read$0.0070 /M
Cache write$0.0583 /M
Image$0.00000 /image
Web search$0.0098 /req

Specs

Context length1M tokens
Max output65.5K tokens
Inputtext, image, file, audio, video
Outputtext

Quickstart

Point the OpenAI SDK at https://tokenroute.app/api/v1 and use a TokenRoute API key.

bash
curl https://tokenroute.app/api/v1/chat/completions \
  -H "Authorization: Bearer $TOKENROUTE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "google/gemini-2.5-flash-lite",
    "messages": [{ "role": "user", "content": "Hello!" }]
  }'

Supported parameters

include_reasoning
max_tokens
reasoning
response_format
seed
stop
structured_outputs
temperature
tool_choice
tools
top_p