SmartBrain API

Enterprise AI as Simple as Utilities

More Links
Model HubConsoleAPI KeysUsage QueryDocs
SmartBrain API. All rights reserved|Privacy Policy|Terms
SmartBrain API
SmartBrain API
  • Model Hub
  • Official Docs
  • Playground
G

z-ai/glm-5

Online Chat
智谱

Publish time

2026/2/12

Model Series

GLM

Input type

Output type

Input Price

¥4 / 1M tokens

Output Price

¥18 / 1M tokens

Cache Write Price

¥5 / 1M tokens

Cache Read Price

¥0.4 / 1M tokens

Context Window

128,000

Max Output Length

2,048

GLM-5 是面向 Coding 与 Agent 场景的新一代大模型,在复杂系统工程与长程任务中达到开源 SOTA,真实编程体验逼近 Claude Opus 级别;基于 744B 新基座、异步强化学习与稀疏注意力,实现从“写代码”到“写工程”的全面升级。

Providers for z-ai/glm-5

Zhinao API routes requests to the best-fit provider and automatically fails over to the one with highest availability.

Sort by

Performance for z-ai/glm-5

Compare different providers across Zhinao API

Throughput

52.25 tok/s

No data

TTFT

0.01 s

No data

Uptime for z-ai/glm-5

Uptime for z-ai/glm-5 across all providers

Sample code and API for z-ai/glm-5

Get API Key

Zhinao API normalizes requests and responses across providers for you

View full docs
This model may not support the OpenAI Chat Completions protocol, please note
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://api.360.cn/v1",
  apiKey: process.env.ZHINAO_API_KEY,
});

const response = await client.chat.completions.create({
  model: "z-ai/glm-5",
  messages: [
    { role: "user", content: "Hello, how are you?" }
  ],
  temperature: 0.7,
  max_tokens: 1000,
});

console.log(response.choices[0].message.content);