Gemini 3.5 Flash API

Gemini 3.5 Flash is a multimodal reasoning model from Google optimized for fast inference and agentic workflows. Supports text, image, audio and video understanding with large context window and strong coding capabilities.
Context
1.05M tokens
Input
$0.65 / 1M
Output
$3.9 / 1M

How to use Gemini 3.5 Flash API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to google/gemini-3-5-flash.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "google/gemini-3-5-flash",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "google/gemini-3-5-flash",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-3-5-flash","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Gemini 3.5 Flash API Pricing

TypePrice
Input
$0.65 / 1M tokens
Output
$3.9 / 1M tokens

Gemini 3.5 Flash vs other models

ModelInputOutputContextBest for
$0.65 / 1M
$3.9 / 1M
1.05M tokens
Reasoning + agents
$6.5 / 1M
$39 / 1M
1.05M tokens
Reasoning + agents
$2.6 / 1M
$13 / 1M
1M tokens
Balanced coding + agents
$3.9 / 1M
$19.5 / 1M
1M tokens
Long-context, multimodal & agentic workflows

Start building with Gemini 3.5 Flash

Get API Key
1000+ models, one API.