Active Parameters
Undisclosed
Context Length
1.05M
Modality
Multimodal
Architecture
Mixture of Experts (MoE)
License
Proprietary
Release Date
25 Sept 2025
Knowledge Cutoff
Jan 2025
API Pricing (per 1M)
Input: $0.30 · Output: $2.50
Rank
#128
| Benchmark | Score | Rank |
|---|---|---|
Software Engineering | 0.29 | 37 |
Graduate-Level QA | 0.828 | 54 |
General Text | 1410 | 94 |
Intelligence Index | auto 0.15 Standard 0.10 | 177 219 |
Summarization Archived | 0.84 | 9 |
Coding Archived | 0.55 | 11 |
QA Assistant Archived | 0.722 | 20 |
Overall Rank
#128
Coding Rank
#136
Gemini 2.5 Flash is a high-throughput multimodal model by Google built to balance fast execution with verifiable chain-of-thought reasoning. It supports native multimodal inputs, function calling, and search grounding over a 1M context window.
Architecture specifications are undisclosed for proprietary models.
Google's advanced multimodal models with native understanding of text, images, audio, and video. Features massive context windows up to 2.1M tokens, max thinking modes for complex reasoning, and optimized variants for different performance/cost tradeoffs. Includes Pro, Flash, and Flash Lite variants with configurable thinking capabilities for transparent reasoning.
Assistant
Online