Qwen3.6-35B-A3B API

alibaba/qwen3.6-35b-a3b
Alibaba's newest open-source mixture-of-experts model activates just 3 billion parameters at inference time, yet it goes toe-to-toe with dense models four to nine times its active size on agentic coding, reasoning, and multimodal tasks.
Context
256K tokens
Input
$0.4875 / 1M tokens
Output
$2.925 / 1M tokens
Released
Apr 23, 2026

How to use Qwen3.6-35B-A3B API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to alibaba/qwen3.6-35b-a3b.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "alibaba/qwen3.6-35b-a3b",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "alibaba/qwen3.6-35b-a3b",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"alibaba/qwen3.6-35b-a3b","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Qwen3.6-35B-A3B API Pricing

TypePrice
Input
$0.4875 / 1M tokens
Output
$2.925 / 1M tokens

Qwen3.6-35B-A3B Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
Terminal-Bench
51.5%
Autonomous shell/terminal task completionSourceJuly 12, 2026
GPQA Diamond
86%
Google-proof graduate science questions (hardest subset)SourceJuly 12, 2026
SWE-bench Verified
73.4%
Resolving verified real GitHub issuesSourceJuly 12, 2026
Intelligence
18.8
Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026
Coding
41.9
Composite score across standardised coding evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026

Qwen3.6-35B-A3B vs other models

ModelInputOutputContextBest for
$0.4875 / 1M tokens
$2.925 / 1M tokens
256K tokens
Reasoning + agents
$6.5 / 1M tokens
$39 / 1M tokens
1M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$1.95 / 1M tokens
$11.7 / 1M tokens
1M tokens
Reasoning + agents

Frequently asked questions

Qwen3.6-35B-A3B has a 256,000 tokens context window and can return up to 252,000 tokens.

Qwen3.6-35B-A3B takes image, text as input and returns text.

Use alibaba/qwen3.6-35b-a3b as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.

Qwen3.6-35B-A3B became available on April 23, 2026.

Qwen3.6-35B-A3B is priced at input $0.4875 / 1M tokens, output $2.925 / 1M tokens.

Yes, Qwen3.6-35B-A3B can stream responses as they are generated.

Yes, Qwen3.6-35B-A3B accepts image input alongside text.

Qwen3.6-35B-A3B was built by Alibaba Cloud.

Yes, it accepts image and text input and produces text output, supporting vision tasks.

Yes, reasoning is listed as one of its features and capabilities.

Yes, it supports tool use, including parallel tool calls and function calling.

Yes, web search is included among its supported features and capabilities.

Start building with Qwen3.6-35B-A3B

Get API Key
1000+ models, one API.