Gemini 2.0 Flash API

High-speed AI model for efficient task execution
Output

How to use Gemini 2.0 Flash API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to .

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Gemini 2.0 Flash API Pricing

TypePrice
Input
Output

Gemini 2.0 Flash Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
Intelligence
8.9
Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026
Math
21.7
Composite score across standardised mathematics evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026

Gemini 2.0 Flash vs other models

ModelInputOutputContextBest for
$6.5 / 1M tokens
$39 / 1M tokens
1.05M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$0.65 / 1M tokens
$3.9 / 1M tokens
1.05M tokens
Reasoning + agents

Frequently asked questions

Yes, Gemini 2.0 Flash can stream responses as they are generated.

Yes, Gemini 2.0 Flash supports both function calling and structured outputs.

Gemini 2.0 Flash was built by Google.

Start building with Gemini 2.0 Flash

Get API Key
1000+ models, one API.