Active Parameters
Undisclosed
Context Length
1.05M
Modality
Multimodal
Architecture
Mixture of Experts (MoE)
License
Proprietary
Release Date
25 Sept 2025
Knowledge Cutoff
Jan 2025
API Pricing (per 1M)
Input: $0.30 · Output: $2.50
Rank
#116
| Benchmark | Score | Rank |
|---|---|---|
Software Engineering | 0.29 | 32 |
Graduate-Level QA | 0.828 | 50 |
General Text | 1410 | 94 |
Intelligence Index | auto 0.14 Standard 0.08 | 177 204 |
Summarization Archived | 0.84 | 9 |
Coding Archived | 0.55 | 11 |
QA Assistant Archived | 0.722 | 20 |
Overall Rank
#116
Coding Rank
#127
Gemini 2.5 Flash is a high-throughput multimodal model by Google built to balance fast execution with verifiable chain-of-thought reasoning. It supports native multimodal inputs, function calling, and search grounding over a 1M context window.
Architecture specifications are undisclosed for proprietary models.
Google's advanced multimodal models with native understanding of text, images, audio, and video. Features massive context windows up to 2.1M tokens, max thinking modes for complex reasoning, and optimized variants for different performance/cost tradeoffs. Includes Pro, Flash, and Flash Lite variants with configurable thinking capabilities for transparent reasoning.
APX AI
Online