Google · live price
Gemini 3.1 Flash Lite API pricing
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
- Input
- $0.25
- per 1M tokens
- Output
- $1.50
- per 1M tokens
- Cached input
- $0.025
- per 1M tokens
- Context
- 1.0M
- 66K max output
List price from OpenRouter, which passes Google's price through without markup. Refreshed hourly. Released May 7, 2026. Model id google/gemini-3.1-flash-lite.
What Gemini 3.1 Flash Lite costs in practice
1,000 chat messages
$1.00
1,000 input + 500 output tokens each
1,000 RAG questions
$2.75
8,000 input + 500 output tokens each
100 coding-agent steps
$1.05
30,000 input + 2,000 output tokens each, no caching
Gemini 3.1 Flash Lite price history
Gemini 3.1 Flash Lite price by host
- Google$0.125 in · $0.75 out · 1.0M ctx · 99.9% uptime 24h
- Google AI Studio$0.125 in · $0.75 out · 1.0M ctx · 100.0% uptime 24h
Where to buy Google credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Gemini 3.1 Flash Lite cost calculator
Cheaper alternatives to Gemini 3.1 Flash Lite
| Model | Input / 1M | Output / 1M |
|---|---|---|
| OpenAI: GPT-6 Luna Pro | $0.1 | $0.5 |
| OpenAI: GPT-6 Luna | $0.1 | $0.5 |
| Qwen: Qwen3.8 Omni Flash | $0.15 | $0.47 |
| DeepSeek: DeepSeek V4.1 Flash | $0.03 | $0.5 |
| Qwen: Qwen3.8 Flash | $0.15 | $0.47 |
| Z.ai: GLM 5.3 Flash | $0.15 | $0.5 |
Other Google models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Gemini 3.5 Flash | $1.50 | $9 |
| Gemma 4 26B A4B | $0.09 | $0.3 |
| Gemma 4 31B | $0.09 | $0.34 |
| Nano Banana Pro (Gemini 3 Pro Image) | $2 | $12 |
| Nano Banana 2 (Gemini 3.1 Flash Image) | $0.5 | $3 |
| Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) | $0.25 | $1.50 |