glm-5.3-flash
Text Generationby Z.ai
GLM-5.3-Flash is Z.ai's efficient open-weight model for coding and long-horizon agent tasks. Its hybrid sparse and linear attention keeps long-context behaviour accurate across a one-million-token window while reducing compute, and it supports function calling and adjustable reasoning effort.
Specifications
Parameters
Not disclosed
Context length
1,048,576 tokens
Task
Text Generation
Provider
Z.ai
Billing
Per token
API
OpenAI-compatible
Pricing
Input
$0.10
per 1M tokens
Output
$0.35
per 1M tokens