GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
Specifications
- Provider
- Z.AI
- Type
- Vendor / proprietary
- Modality
- text+image+video->text
- Context window
- 202,752
- Released
- April 1, 2026
Capabilities
Input: imageInput: textInput: videoOutput: text
Strengths
include_reasoningmax_tokensreasoningresponse_formattemperaturetool_choicetoolstop_ktop_p
