
Z.ai
Multimodal AI models for image generation and video synthesis
Z.ai develops multimodal foundation models spanning image generation, video synthesis, and visual understanding. Its GLM-Image model combines autoregressive and diffusion architectures for high fidelity output with strong text rendering. As a Runware provider, Z.ai models are available through a single inference pipeline alongside other creators.
Models by Z.ai
GLM-5.2 is Z.ai's flagship language model for long-horizon coding, agentic engineering, and sustained multi-step execution. It is designed to keep large project context coherent over extended runs, with a 1M token context window, 128K max output, multiple thinking modes, function calling, structured output, context caching, streaming, and MCP support for tool-rich workflows.
GLM-5.1 is Z.ai’s flagship language model for agentic engineering, coding, reasoning, and tool-driven workflows. It supports a 200K token context window with up to 128K output tokens, deep thinking, function calling, structured output, and streaming tool calls, and is designed to stay effective over long multi-step sessions rather than only short-horizon tasks.

