Text generation / Google
gemini-3.1-flash-lite
Overview
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic workflows, simple data extraction, and applications where responsiveness and API cost are the primary constraints.
- Model type: Text generation
- Input: Text / Video / Image / Audio
- Output: Text
- Endpoints: chat
Tiered billing details
| Tier | Input | Output | Cache hit |
|---|---|---|---|
| Fallback tier | $0.0003/1K | $0.0015/1K | $0.0000/1K |
Billed bylen (input context token count). Coefficient is the $ / 1M tokens price. Default group price (other groups converted by discount).
Code example
Call with gemini-3.1-flash-lite :
curl https://www.starunion.net/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-3.1-flash-lite",
"messages": [{ "role": "user", "content": "Hello" }]
}'