Z.ai
3 channels

glm-5.1

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work

Starting at
Input
$0.8163/ 1M
List price $0.8571
Output
$3.2653/ 1M
List price $3.4286
Open dashboard

Available from 3 channels

Each channel is compared at its lowest input-price tier; output breaks ties. A dash means /api/pricing does not provide that rate. Selecting a channel updates the endpoints and call example below.

jd/glm-5.1
Lowest input
Input
$0.8163 USD per 1M tokens
Output
$3.2653 USD per 1M tokens
Cache read
$0.1769 USD per 1M tokens
Cache write
- USD per 1M tokens
baidu/glm-5.1
Input
$1.3333 USD per 1M tokens
Output
$4.1905 USD per 1M tokens
Cache read
$0.2476 USD per 1M tokens
Cache write
- USD per 1M tokens
z-ai/glm-5.1
Input
$1.40 USD per 1M tokens
Output
$4.40 USD per 1M tokens
Cache read
$0.26 USD per 1M tokens
Cache write
- USD per 1M tokens

Endpoints

POST/v1/chat/completionsChat Completions

Call it

Using jd/glm-5.1

1from openai import OpenAI23client = OpenAI(4    base_url="https://www.realrelay.ai/v1",5    api_key="sk-***",6)78response = client.chat.completions.create(9    model="jd/glm-5.1",10    messages=[{"role": "user", "content": "Hello"}],11)1213print(response.choices[0].message.content)
1import OpenAI from "openai";23const client = new OpenAI({4  baseURL: "https://www.realrelay.ai/v1",5  apiKey: process.env.REALRELAY_API_KEY,6});78const response = await client.chat.completions.create({9  model: "jd/glm-5.1",10  messages: [{ role: "user", content: "Hello" }],11});1213console.log(response.choices[0].message.content);
1curl https://www.realrelay.ai/v1/chat/completions \2  -H "Authorization: Bearer $REALRELAY_API_KEY" \3  -H "Content-Type: application/json" \4  -d '{5    "model": "jd/glm-5.1",6    "messages": [{ "role": "user", "content": "Hello" }]7  }'

Prices and availability as published on Oct 4, 2026.

Capabilities

MCPPrompt cachingReasoningStructured outputTool use
Input
Text
Output
Text