GLM 4.5V
Z.ai's vision-language model for image and text understanding with a 65k-token context.
Pricing
per 1M tokensInput$0.6 /1M
Output$1.8 /1M
Cache read$0.11 /1M
Model Description
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
Get the full context.
Sign up to read complete case studies, access detailed metrics, and unlock all use cases.
Get the full context.
Sign up to read complete case studies, access detailed metrics, and unlock all use cases.