Quick answer: Gemini 2.0 Flash (released February 5, 2025) is Google DeepMind's workhorse model for the agentic era: a 1M-token multimodal model at $0.10/$0.40 per million tokens that added native image generation, computer use (visual actions), and a live multimodal API to the Flash tier for the first time. It was the default model in Google AI Studio at launch.
Where Gemini 2.0 Flash leads
Where it lags
Best for: agentic computer use, real-time multimodal applications, production coding assistants, and workflows requiring long-context processing at moderate cost.
Gemini 2.0 Flash was released as Google's agent-era model. The defining additions vs Gemini 1.5 Flash were computer use (the model can interpret screenshots and output structured UI actions), native image generation (outputs images without a separate image model), and the Multimodal Live API (low-latency streaming interaction across text, audio, and video).
For Benchgen users, Gemini 2.0 Flash is notable as the first broadly available Flash-tier model capable of full computer use agent loops — tasks like navigating web UIs, filling forms, and interacting with desktop applications without requiring a separate vision-action model.
| Field | Value |
|---|---|
| Organization | Google DeepMind |
| API identifier | gemini-2.0-flash |
| Context window | 1,000,000 tokens |
| Max output | 8,192 tokens |
| License | Proprietary |
| Release date | February 5, 2025 |
| Modality | Multimodal (text, image, audio, video input; text + image output) |
| Special capabilities | Computer use, native image generation, Multimodal Live API |
| Input (per 1M tokens) | Output (per 1M tokens) | |
|---|---|---|
| Google AI | $0.10 | $0.40 |
Image input: $0.0258/image. Source: Google AI pricing.
| Benchmark | Score | Source | Date |
|---|---|---|---|
| MMLU | 83.2% | Google — Gemini 2.0 Flash | 2025-02 |
| HumanEval | 83.0% | Google — Gemini 2.0 Flash | 2025-02 |
| GPQA | 62.1% | Google — Gemini 2.0 Flash | 2025-02 |
Scores reported by Google. Not Benchgen measurements.
| Model | Context | Price (in/out per 1M) | Computer use |
|---|---|---|---|
| Gemini 2.0 Flash | 1M | $0.10 / $0.40 | Yes |
| Gemini 1.5 Flash | 1M | $0.075 / $0.30 | No |
| Gemini 3 Flash | 1M | TBD | Yes |
| GPT-4o (Nov 2024) | 128K | $2.50 / $10 | Limited |
Gemini 2.0 Flash adds computer use and image generation at 33% more than Gemini 1.5 Flash. GPT-4o costs 25× more per token at 8× less context. (Scores from respective announcements; not Benchgen measurements.)
Gemini 2.0 Flash is a large language model developed by Google.
This model isn’t on any benchmark leaderboard yet.