Qwen3-Next-80B-A3B Instruct API

alibaba/qwen3-next-80b-a3b-instruct
Qwen3-Next-80B-A3B Instruct is a next-generation large language model that balances enormous parameter scale with sparse activation to deliver fast, cost-efficient, and scalable instruction-following capabilities.
Context
126K tokens
Input
$0.195 / 1M tokens
Output
$1.56 / 1M tokens
Released
Sep 30, 2025

How to use Qwen3-Next-80B-A3B Instruct API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to alibaba/qwen3-next-80b-a3b-instruct.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "alibaba/qwen3-next-80b-a3b-instruct",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "alibaba/qwen3-next-80b-a3b-instruct",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"alibaba/qwen3-next-80b-a3b-instruct","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Qwen3-Next-80B-A3B Instruct API Pricing

TypePrice
Input
$0.195 / 1M tokens
Output
$1.56 / 1M tokens

Qwen3-Next-80B-A3B Instruct Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
MMLU-Pro
80.6%
Multi-discipline knowledge + reasoning (harder MMLU)SourceJuly 12, 2026
LiveCodeBench
56.6%
Contamination-free competitive programming problemsSourceJuly 12, 2026
AIME 2025
69.5%
Competition mathematics (AIME), 2025SourceJuly 12, 2026
GPQA Diamond
72.9%
Google-proof graduate science questions (hardest subset)SourceJuly 12, 2026
Intelligence
9.6
Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026
Math
66.3
Composite score across standardised mathematics evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026

Qwen3-Next-80B-A3B Instruct vs other models

ModelInputOutputContextBest for
$0.195 / 1M tokens
$1.56 / 1M tokens
126K tokens
Coding + agents
$6.5 / 1M tokens
$39 / 1M tokens
1M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$1.95 / 1M tokens
$11.7 / 1M tokens
1M tokens
Reasoning + agents

Frequently asked questions

Qwen3-Next-80B-A3B Instruct has a 129,024 tokens context window and can return up to 16,384 tokens.

Qwen3-Next-80B-A3B Instruct takes text as input and returns text.

Use alibaba/qwen3-next-80b-a3b-instruct as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.

Qwen3-Next-80B-A3B Instruct became available on September 30, 2025.

Qwen3-Next-80B-A3B Instruct is priced at input $0.195 / 1M tokens, output $1.56 / 1M tokens.

Yes, Qwen3-Next-80B-A3B Instruct can stream responses as they are generated.

Qwen3-Next-80B-A3B Instruct was built by Alibaba Cloud.

It supports parallel tool calls, structured output, function calling, and web search.

Start building with Qwen3-Next-80B-A3B Instruct

Get API Key
1000+ models, one API.