News
Zhipu has officially launched and open-sourced the GLM-4.6V series of multimodal large models.
Zhipu AI has officially launched and open-sourced its GLM-4.6V series of multimodal large models, including versions 106B and 9B. The models natively support autonomous tool invocation based on visual input, enabling them to handle complex tasks such as mixed text and image processing and image-based shopping. Their 128K long context window can understand documents up to 150 pages long or one hour of video content, and their capabilities in areas such as front-end code replication are improved.