GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
Specifications
- Provider
- Z.AI
- Type
- Open-source / open-weight
- Modality
- text+image->text
- Context window
- 65,536
- Knowledge cutoff
- 2024-12-31
- Released
- August 11, 2025
Capabilities
Input: textInput: imageOutput: text
Strengths
frequency_penaltyinclude_reasoningmax_tokenspresence_penaltyreasoningrepetition_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_p
