GLM 4.5V API

z-ai/glm-4.5v
GLM 4.5V on AIMLAPI.
Context
64K tokens
Input
$0.82524 / 1M tokens
Output
$2.47572 / 1M tokens
Released
Aug 11, 2025

How to use GLM 4.5V API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to z-ai/glm-4.5v.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "z-ai/glm-4.5v",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "z-ai/glm-4.5v",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"z-ai/glm-4.5v","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

GLM 4.5V API Pricing

TypePrice
Input
$0.82524 / 1M tokens
Output
$2.47572 / 1M tokens
Cached input
$0.151294 / 1M tokens

GLM 4.5V Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
OSWorld
35.8%
Computer-use across real desktop applicationsSourceJuly 12, 2026
MMMU
75.4%
College-level multimodal understanding + reasoningSourceJuly 12, 2026
Intelligence
7.6
Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026
Math
73
Composite score across standardised mathematics evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026

GLM 4.5V vs other models

ModelInputOutputContextBest for
GLM 4.5V
This page
$0.82524 / 1M tokens
$2.47572 / 1M tokens
64K tokens
Reasoning + agents
$6.5 / 1M tokens
$39 / 1M tokens
1M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$1.95 / 1M tokens
$11.7 / 1M tokens
1M tokens
Reasoning + agents

Frequently asked questions

GLM 4.5V has a 65,536 tokens context window and can return up to 16,384 tokens.

GLM 4.5V takes image, text as input and returns text.

Use z-ai/glm-4.5v as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.

GLM 4.5V is priced at input $0.82524 / 1M tokens, output $2.47572 / 1M tokens, cached input $0.151294 / 1M tokens.

Yes, GLM 4.5V can stream responses as they are generated.

Yes, GLM 4.5V accepts image input alongside text.

GLM 4.5V was built by Zhipu AI.

Start building with GLM 4.5V

Get API Key
1000+ models, one API.