Gemini 3.6 Flash API

Gemini 3.6 Flash is Google's most intelligent Flash model, balancing speed with frontier intelligence for strong performance on agentic, coding and multimodal tasks, with superior search and grounding.
Context
1.05M tokens
Input
$1.95 / 1M
Output
$9.75 / 1M
Released
Jul 21, 2026

How to use Gemini 3.6 Flash API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to google/gemini-3.6-flash.

import requests

r = requests.post(
   "https://api.aimlapi.com/v1/chat/completions",
   headers={"Authorization": "Bearer " + AIMLAPI_KEY},
   json={
     "model": "google/gemini-3.6-flash",
     "messages": [
       {
         "role": "user",
         "content": "Hello!"
       }
     ]
   },
)
print(r.json())

const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
 method: "POST",
 headers: {
   Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
   "Content-Type": "application/json",
 },
 body: JSON.stringify({
   "model": "google/gemini-3.6-flash",
   "messages": [
     {
       "role": "user",
       "content": "Hello!"
     }
   ]
 }),
});
console.log(await r.json());

curl -X POST https://api.aimlapi.com/v1/chat/completions \
 -H "Authorization: Bearer $AIMLAPI_KEY" \
 -H "Content-Type: application/json" \
 -d '{"model":"google/gemini-3.6-flash","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Gemini 3.6 Flash API Pricing

TypePrice
Input
$1.95 / 1M tokens
Output
$9.75 / 1M tokens

Google Search grounding billed separately at $0.0455 per call.

Gemini 3.6 Flash vs other models

ModelInputOutputContextBest for
$1.95 / 1M
$9.75 / 1M
1.05M tokens
Reasoning + agents
$2.6 / 1M
$13 / 1M
1M tokens
Balanced coding + agents
$3.9 / 1M
$19.5 / 1M
1M tokens
Long-context, multimodal & agentic workflows
$1.3 / 1M
$7.8 / 1M
1M tokens
Fast, high-volume tasks

Start building with Gemini 3.6 Flash

Get API Key
1000+ models, one API.