Model information
Released
2026-05-07
Context window
1M
Max output
66K
Weights
Closed
Modalities
Input
AudioDocumentImageTextVideo
Output
Text
Identity
- Version
- gemini-3.1-flash-lite
- Knowledge cutoff
- 2025-01
- Aliases
- Gemini 3.1 Flash Lite
Capabilities
AttachmentsReasoningStructuredTemperatureTool call
Official price
input $0.25/M · output $1.5/M
cache read $0.025/M
Providers
| Provider | Provider model ID | Protocol | Context | Max output | Input | Output | Capabilities | Notes | Pricing | Streaming |
|---|---|---|---|---|---|---|---|---|---|---|
| Poe | google/gemini-3.1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool call | - | input $0.25/M · output $1.5/M | Yes | ||
| gemini-3.1-flash-lite | google_generate_content | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.025/M | Yes | |||
| NanoGPT | google/gemini-3.1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.025/M · cache write $0.083/M | Yes | ||
| NEAR AI Cloud | google/gemini-3.1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.025/M | Yes | ||
| LLM Gateway | gemini-3.1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.025/M · cache write $0.083/M | Yes | ||
| Kenari | gemini-3-1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0/M · output $0/M | Yes | ||
| Neon | gemini-3-1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.025/M | Yes | ||
| SAP AI Core | gemini-3.1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.025/M | Yes | ||
| ZenMux | google/gemini-3.1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.025/M | Yes | ||
| Abacus | gemini-3.1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.025/M | Yes | ||
| AIHubMix | gemini-3.1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.025/M · cache write $1/M | Yes | ||
| Vertex | gemini-3.1-flash-lite | custom | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.025/M | Yes | ||
| Kilo Gateway | google/gemini-3.1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.025/M · cache write $0.083/M | Yes | ||
| Merge Gateway | google/gemini-3.1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.025/M | Yes | ||
| Vercel AI Gateway | google/gemini-3.1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.25/M · output $1.5/M cache read $0.03/M | Yes | ||
| OpenRouter | google/gemini-3.1-flash-lite | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | Reasoning levels: high, medium, low, minimal, Default reasoning: minimal | input $0.25/M · output $1.5/M cache read $0.025/M · cache write $0.083/M | Yes |