Long-context reasoning and coding model. The GLM release that expanded the context window from 200K to 1M tokens.
Z.ai models
Access 30 Z.ai models through the Oxyy unified API including GLM-5.2, GLM-5.3, GLM-5.3-Flash. Compare pricing, context windows and capabilities between Z.ai models.
Z.ai tokens processed on Oxyy· daily, UTC
Models 30
Z.ai's flagship model for complex software engineering and long-horizon agent tasks. Uses the same base model as GLM-5.2, with all gains driven by post-training.
The first native multimodal model in the GLM-5 series, delivering stronger intelligence than GLM-5.2 at very low cost.
Free tier text model from the GLM-4.7 generation.
Z.ai's new-generation flagship foundation model for agentic engineering, targeting complex system engineering and long-range agent tasks.
Leading open-source coding model with significant gains on long-horizon tasks. Predecessor to GLM-5.2.
Z.ai's video generation model.
Z.ai's CogView image generation model.
32-billion-parameter GLM-4 model with a 128K context window, flat-priced on input and output.
GLM-4.5 base text model. The release that introduced interleaved reasoning to the GLM line.
Lightweight, low-cost variant of GLM-4.5.
High-throughput variant of GLM-4.5-Air.
Free tier text model from the GLM-4.5 generation.
Previous-generation GLM vision-language model.
Highest-performance variant of GLM-4.5, and the most expensive model in Z.ai's published catalog.
Prior-generation GLM-4 text model.
Multimodal model for high-fidelity visual understanding and long-context reasoning across images, documents and mixed media. Handles complex page layouts and charts as visual input.
Free tier vision-language model.
High-throughput, low-cost vision variant of GLM-4.6V.
Previous flagship focused on task completion rather than single-point code generation, with interleaved, retained and round-level reasoning.
High-throughput, very low cost variant of GLM-4.7.
Higher-throughput FlashX variant of the GLM-5.3 generation.
Coding-specialized variant of GLM-5.
Speed-optimized variant of GLM-5.
Z.ai's automatic speech recognition model.
Text-to-image generation model, described by Z.ai as achieving open-source state of the art in complex scenarios.
Z.ai's OCR model for extracting text from images and documents.
Z.ai agent that generates slides and posters.
Z.ai's translation agent.
Z.ai agent that applies popular special-effects templates to video.

