ZhiPu's Open-Source Visual Model with 100 Billion Parameters Distinguishes Between McDonald's and KFC Fried Chicken
2025-08-12 / Read about 0 minute
Author:小编   

On August 11, ZhiPu unveiled its latest generation of visual understanding model, GLM-4.5V, and announced its open-source availability. GLM-4.5V builds upon ZhiPu's flagship text-based model, GLM-4.5-Air, continuing the technological trajectory established by GLM-4.1V-Thinking. Boasting a total of 106 billion parameters and 12 billion active parameters, GLM-4.5V achieves state-of-the-art (SOTA) performance among open-source models of its class across 41 public visual multimodal benchmarks. These benchmarks encompass a wide range of tasks, including image and video understanding, document analysis, and GUI Agent interactions. Furthermore, GLM-4.5V introduces an innovative thinking mode switch, empowering users to tailor the model's level of depth in processing according to their specific needs.

  • C114 Communication Network
  • Communication Home
7 X 24 Track global technological trends
Hot Topic