GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...
| Type | Model |
| Section | Models & platforms / text+image+video->text |
| Pricing | paid (от $0.3/mo) |
| Platform | Self-hosted |
| Systems | api, python, self-hosted |
| Hosting | cloud |
| Site language | en |
| Vendor | Z-ai |
| Launched | 2025-12-08 |