scx.ai logo
All pricing

Z AI

GLM-5.2

GlobalChatReasoning

About model

Latest flagship model from Z AI, designed for long-horizon agentic tasks with strong reasoning and coding ability. 1M context

It takes text as input and returns text. It works with a context window of 1M tokens and can return up to 125k tokens in a single response.

Calls go through the SCX gateway on an OpenAI compatible endpoint, so moving an existing integration across means changing the base URL and the model name, and nothing else. The default tier allows 500 requests per minute, and higher limits are available on request.

Try GLM-5.2 today
Context window

1M

Tokens the model can attend to in a single request

LiveBench

73.2

Overall score, ranked 23 on the public leaderboard

Max output

125k

Tokens the model can return in one response

Model key capabilities

  • Reasoning

    Works through a problem step by step before answering, which lifts accuracy on multi-step maths, logic and planning tasks.

  • Tool calling

    Returns structured arguments for the tools you define, using the same schema as the OpenAI API, so existing agent frameworks work unchanged.

  • JSON mode

    Constrains output to valid JSON, so responses can be parsed without a repair step.

  • Streaming

    Streams tokens as they are generated, so an interface can show output before the response completes.

Benchmarks

Scores on LiveBench, a contamination-free benchmark refreshed every six months. Higher is better.

73.2overall
Ranked 23 of 39 models
Reasoning
78.6
Coding
79.7
Agentic Coding
51.8
Mathematics
89.8
Data Analysis
73.7
Language
76.2
Instruction Following
62.3

Source: LiveBench 2026-06-25 release.

Quick start

curl https://api.scx.ai/v1/chat/completions \
  -H "Authorization: Bearer $SCX_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "GLM-5.2",
    "messages": [{ "role": "user", "content": "Hello" }]
  }'
GLM-5.2 by Z AI: Pricing & Specs | SCX.ai