Qwen3-Next-80B-A3B Thinking API

Qwen/Qwen3-235B-A22B-Thinking-2507
Qwen3-Next-80B-A3B Thinking integrates seamlessly into modern AI workflows with flexible deployment options including serverless, on-demand dedicated, and reserved monthly instances.
Context
32K tokens
Input
$0.845 / 1M tokens
Output
$3.9 / 1M tokens
Released
Sep 30, 2025

How to use Qwen3-Next-80B-A3B Thinking API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to Qwen/Qwen3-235B-A22B-Thinking-2507.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "Qwen/Qwen3-235B-A22B-Thinking-2507",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "Qwen/Qwen3-235B-A22B-Thinking-2507",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"Qwen/Qwen3-235B-A22B-Thinking-2507","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Qwen3-Next-80B-A3B Thinking API Pricing

TypePrice
Input
$0.845 / 1M tokens
Output
$3.9 / 1M tokens

Qwen3-Next-80B-A3B Thinking Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
MMLU-Pro
82.7%
Multi-discipline knowledge + reasoning (harder MMLU)SourceJuly 12, 2026
LiveCodeBench
68.7%
Contamination-free competitive programming problemsSourceJuly 12, 2026
AIME 2025
87.8%
Competition mathematics (AIME), 2025SourceJuly 12, 2026
GPQA Diamond
77.2%
Google-proof graduate science questions (hardest subset)SourceJuly 12, 2026
Intelligence
11.2
Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026
Coding
17.4
Composite score across standardised coding evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026
Math
84.3
Composite score across standardised mathematics evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026

Qwen3-Next-80B-A3B Thinking vs other models

ModelInputOutputContextBest for
$0.845 / 1M tokens
$3.9 / 1M tokens
32K tokens
Coding + agents
$6.5 / 1M tokens
$39 / 1M tokens
1M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$1.95 / 1M tokens
$11.7 / 1M tokens
1M tokens
Reasoning + agents

Frequently asked questions

Qwen3-Next-80B-A3B Thinking has a 126,976 tokens context window and can return up to 81,920 tokens.

Qwen3-Next-80B-A3B Thinking takes text as input and returns text.

Use Qwen/Qwen3-235B-A22B-Thinking-2507 as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.

Qwen3-Next-80B-A3B Thinking became available on September 30, 2025.

Qwen3-Next-80B-A3B Thinking is priced at input $0.845 / 1M tokens, output $3.9 / 1M tokens.

Yes, Qwen3-Next-80B-A3B Thinking can stream responses as they are generated.

Qwen3-Next-80B-A3B Thinking was built by Alibaba Cloud.

Yes, it is described as a multilingual reasoning model with knowledge augmentation and creative capabilities.

Start building with Qwen3-Next-80B-A3B Thinking

Get API Key
1000+ models, one API.