GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
| Type | Model |
| Section | Models & platforms / text+image->text |
| Pricing | paid (от $0.6/mo) |
| Platform | Self-hosted |
| Systems | api, python, self-hosted |
| Hosting | cloud |
| Site language | en |
| Vendor | Z-ai |
| Launched | 2025-08-11 |