infery
← All models
GLM 5 Turbo

GLM 5 Turbo

glm-5-turbo

Chat & Textby Z.AI

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. 202,752 token context window, maximum output of 131,072 tokens. Includes independent benchmarks from Artificial Analysis.

chatstreamingtoolsjson mode

Details

Accepts
text
Context window
203K tokens

Pricing

Input
150 cr / 1M tokens
Output
500 cr / 1M tokens
Cached input
30 cr / 1M tokens

Prices in credits (1 credit = $0.01).

Data schema

approximate

Input

FieldTypeDescription
promptrequiredstringThe text prompt.

Output

FieldTypeDescription
textstringGenerated text (choices[0].message.content).