infery
← All models
GLM Flash Latest

GLM Flash Latest

glm-flash-latest

Chat & Textby Z.AI

This model always redirects to the latest model in the GLM Flash family. 1,310,720 token context window, maximum output of 131,072 tokens. Includes independent benchmarks from Artificial Analysis.

chatstreamingvisiontoolsjson mode

Details

Accepts
text
Context window
1M tokens

Pricing

Input
9.38 cr / 1M tokens
Output
31.25 cr / 1M tokens
Cached input
1.88 cr / 1M tokens

Prices in credits (1 credit = $0.01).

Data schema

approximate

Input

FieldTypeDescription
promptrequiredstringThe text prompt.

Output

FieldTypeDescription
textstringGenerated text (choices[0].message.content).