Loading...
Loading...
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
No benchmark data available.
Check your models and usage for free. See the relevant official changes, estimated cost, and a practical next step.