import requests
r = requests.post(
"https://api.aimlapi.com/v1/chat/completions",
headers={"Authorization": "Bearer " + AIMLAPI_KEY},
json={
"model": "minimax/minimax-m3",
"messages": [
{
"role": "user",
"content": "Hello!"
}
]
},
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", { method: "POST", headers: { Authorization: `Bearer ${process.env.AIMLAPI_KEY}`, "Content-Type": "application/json", }, body: JSON.stringify({ "model": "minimax/minimax-m3", "messages": [ { "role": "user", "content": "Hello!" } ] }), }); console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \ -H "Authorization: Bearer $AIMLAPI_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"minimax/minimax-m3","messages":[{"role":"user","content":"Hello!"}]}'
OpenAI-compatible — swap the base URL and it works with your existing SDK.
| Type | Price |
|---|---|
| Input | |
| Output | |
| Cached input |
| Benchmark | Score | What it measures | Source | Retrieved |
|---|---|---|---|---|
| SWE-bench Verified | 80.5% | Resolving verified real GitHub issues | Source | July 12, 2026 |
| Intelligence | 29.6 | Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial Analysis | Source | September 12, 2026 |
| Coding | 58.6 | Composite score across standardised coding evaluations, measured independently by Artificial Analysis | Source | September 12, 2026 |
| Model | Input | Output | Context | Best for |
|---|---|---|---|---|
MiniMax M3 This page | Coding + agents | |||
| Reasoning + agents | ||||
| Balanced coding + agents | ||||
| Long-context, multimodal & agentic workflows | ||||
| Reasoning + agents |
MiniMax M3 has a 524,288 tokens context window and can return up to 524,288 tokens.
MiniMax M3 takes image, text as input and returns text.
Use minimax/minimax-m3 as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.
MiniMax M3 became available on June 1, 2026.
MiniMax M3 is priced at input $0.39 / 1M tokens, output $1.56 / 1M tokens, cached input $0.078 / 1M tokens.
Yes, MiniMax M3 can stream responses as they are generated.
Yes, MiniMax M3 accepts image input alongside text.
MiniMax M3 was built by MiniMax.
Yes, it supports function calling along with parallel tool calls and structured outputs.