Model information
Released
2025-06-17
Context window
1M
Max output
66K
Weights
Closed
Modalities
Input
AudioDocumentImageTextVideo
Output
Text
Identity
- Version
- gemini-2.5-flash
- Knowledge cutoff
- 2025-01
- Aliases
- Gemini 2.5 Flash
Capabilities
AttachmentsReasoningStructuredTemperatureTool call
Official price
input $0.3/M · output $2.5/M
cache read $0.03/M
Providers
| Provider | Provider model ID | Protocol | Context | Max output | Input | Output | Capabilities | Notes | Pricing | Streaming |
|---|---|---|---|---|---|---|---|---|---|---|
| AnyAPI | google/gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | - | Yes | ||
| Poe | google/gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool call | - | input $0.21/M · output $1.8/M cache read $0.021/M | Yes | ||
| gemini-2.5-flash | google_generate_content | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes | |||
| NanoGPT | gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes | ||
| FastRouter | google/gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callTemperature | - | input $0.3/M · output $2.5/M cache read $0.038/M | Yes | ||
| NEAR AI Cloud | google/gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes | ||
| LLM Gateway | gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes | ||
| Kenari | gemini-2-5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0/M · output $0/M | Yes | ||
| 302.AI | gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsTool callTemperature | - | input $0.3/M · output $2.5/M | Yes | ||
| Helicone | gemini-2.5-flash | openai_chat | 1M | 66K | ReasoningTool callTemperature | - | input $0.3/M · output $2.5/M cache read $0.075/M · cache write $0.3/M | Yes | ||
| Neon | gemini-2-5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes | ||
| FrogBot | gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callTemperature | - | input $0.3/M · output $2.5/M cache read $0.075/M | Yes | ||
| SAP AI Core | gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes | ||
| ZenMux | google/gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callTemperature | - | input $0.3/M · output $2.5/M cache read $0.07/M · cache write $1/M | Yes | ||
| Abacus | gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes | ||
| Auriko | gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes | ||
| QiHang | gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.09/M · output $0.71/M | Yes | ||
| AIHubMix | gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes | ||
| OrcaRouter | google/gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes | ||
| Vertex | gemini-2.5-flash | custom | 1M | 66K | AttachmentsReasoningTool callTemperature | - | input $0.3/M · output $2.5/M cache read $0.075/M · cache write $0.383/M | Yes | ||
| Requesty | google/gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callTemperature | - | input $0.3/M · output $2.5/M cache read $0.075/M · cache write $0.55/M | Yes | ||
| Kilo Gateway | google/gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M · cache write $0.083/M | Yes | ||
| Merge Gateway | google/gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes | ||
| Modelis | gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M | Yes | ||
| Vercel AI Gateway | google/gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes | ||
| OpenRouter | google/gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M · cache write $0.083/M | Yes | ||
| Qiniu | gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callTemperature | - | - | Yes | ||
| Jiekou.AI | gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callStructuredTemperature | - | input $0.27/M · output $2.25/M | Yes | ||
| Perplexity Agent | google/gemini-2.5-flash | openai_chat | 1M | 66K | AttachmentsReasoningTool callTemperature | - | input $0.3/M · output $2.5/M cache read $0.03/M | Yes |