TokenRoute
Get started
Back to models

Google: Gemini 3.1 Flash Lite Preview

google/gemini-3.1-flash-lite-preview

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

Pricing

Input$0.1750 /M
Output$1.05 /M
Reasoning$1.05 /M
Cache read$0.0175 /M
Cache write$0.0583 /M
Image$0.00000 /image
Web search$0.0098 /req

Specs

Context length1M tokens
Max output65.5K tokens
Inputtext, image, video, file, audio
Outputtext

Quickstart

Point the OpenAI SDK at https://tokenroute.app/api/v1 and use a TokenRoute API key.

bash
curl https://tokenroute.app/api/v1/chat/completions \
  -H "Authorization: Bearer $TOKENROUTE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "google/gemini-3.1-flash-lite-preview",
    "messages": [{ "role": "user", "content": "Hello!" }]
  }'

Supported parameters

include_reasoning
max_tokens
reasoning
reasoning_effort
response_format
seed
structured_outputs
temperature
tool_choice
tools
top_p