Qwen Turbo API

alibaba/qwen-turbo
Qwen Turbo: optimizes AI agent speed, integrates with RAG, large context window.
Context
1M tokens
Input
$0.065 / 1M tokens
Output
$0.26 / 1M tokens
Released
Sep 9, 2025

How to use Qwen Turbo API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to alibaba/qwen-turbo.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "alibaba/qwen-turbo",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "alibaba/qwen-turbo",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"alibaba/qwen-turbo","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Qwen Turbo API Pricing

TypePrice
Input
$0.065 / 1M tokens
Output
$0.26 / 1M tokens
Cached input
$0.013 / 1M tokens

Qwen Turbo Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
Intelligence
6.4
Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026

Qwen Turbo vs other models

ModelInputOutputContextBest for
Qwen Turbo
This page
$0.065 / 1M tokens
$0.26 / 1M tokens
1M tokens
Coding + agents
$6.5 / 1M tokens
$39 / 1M tokens
1.05M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$0.65 / 1M tokens
$3.9 / 1M tokens
1.05M tokens
Reasoning + agents

Frequently asked questions

Qwen Turbo has a 1,000,000 tokens context window and can return up to 16,384 tokens.

Qwen Turbo takes text as input and returns text.

Use alibaba/qwen-turbo as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.

Qwen Turbo became available on September 9, 2025.

Qwen Turbo is priced at input $0.065 / 1M tokens, output $0.26 / 1M tokens, cached input $0.013 / 1M tokens.

Yes, Qwen Turbo can stream responses as they are generated.

Qwen Turbo was built by Alibaba Cloud.

Qwen Turbo is accessed via the endpoint https://api.aimlapi.com/v1/chat/completions using standard chat completion requests.

Yes, Qwen Turbo includes reasoning as one of its capabilities.

Yes, Qwen Turbo supports function calling, tool use, and parallel tool calls.

Start building with Qwen Turbo

Get API Key
1000+ models, one API.