Gemini 2.0 Flash

by Google

Gemini 2.0 Flash is a fast multimodal Google model that takes text, images, video, audio, and PDF input. It is tuned for speed and efficiency, which makes it a fit for high-volume workloads and interactive applications where response time matters. The model supports tool use and structured output for programmatic control flow. With a 1M token context window, it handles long documents, extended conversation history, and complex multi-turn interactions, and it performs well on visual understanding, video analysis, and other multimodal reasoning tasks.

Key info

Input
Output
Features
Context window
1M
Max output
8K

Benchmarks

Independent benchmark scores — composite indices for reasoning, coding, and math, plus individual eval scores where available.

Global rank#353 of 596 LLMs
TierEfficient
Output speed0 tok/s
First token0.00s
Intelligence Index12.2
Math Index21.7
Reasoning & knowledge
MMLU-Pro
78%
GPQA Diamond
62%
Humanity's Last Exam
4%
Long-context reasoning
32%
Coding
LiveCodeBench
33%
SciCode
33%
Agentic & tool use
Terminal-Bench Hard
4%
τ²-Bench Telecom
30%
Math & instruction following
AIME 2025
22%
IFBench
40%

Available routes

No routes currently available — Gemini 2.0 Flash isn't routed through the Opper gateway right now. It may return.

Contact us about this model →

Available models from Google

Start building with 700+ models

One API key. Every major provider. Up and running in minutes.

Get startedView Documentation
Gemini 2.0 Flash by Google — not currently on Opper | Opper AI