Gemini 2.0 Flash Thinking API

Advanced AI model with explicit reasoning capabilities
Output

How to use Gemini 2.0 Flash Thinking API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to .

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Gemini 2.0 Flash Thinking API Pricing

TypePrice
Input
Output

Gemini 2.0 Flash Thinking Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
Intelligence
8.9
Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026
Math
21.7
Composite score across standardised mathematics evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026

Gemini 2.0 Flash Thinking vs other models

ModelInputOutputContextBest for
$6.5 / 1M tokens
$39 / 1M tokens
1M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$1.95 / 1M tokens
$11.7 / 1M tokens
1M tokens
Reasoning + agents

Frequently asked questions

Yes, Gemini 2.0 Flash Thinking can stream responses as they are generated.

Yes, Gemini 2.0 Flash Thinking supports both function calling and structured outputs.

Gemini 2.0 Flash Thinking was built by Google.

Its listed capabilities are function calling, streaming, structured outputs, and reasoning; image generation is not included.

It is built for reasoning tasks and also supports function calling and structured output generation, making it useful for tasks needing logical steps and tool integration.

It is available through the v1 endpoint.

Start building with Gemini 2.0 Flash Thinking

Get API Key
1000+ models, one API.