Loading...
Loading...
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
Best for
Price history
Local run command
# Refer to provider docs:
Presumptive Specs:
Typical GPU
12–24 GB VRAM
System RAM
16+ GB
Disk
20+ GB free
Runtime
Docker / local runtime
Community reliability estimate · not official
About this score: Community-estimated based on user reports and publicly available benchmark data (e.g. TruthfulQA). This is not an official score from the model provider. Scores may be inaccurate — always verify with the official leaderboard before making production decisions.
Input & output cost per 1M tokens over time
Price history is a Pro feature
Track pricing trends and catch price drops early.
Upgrade to Pro — $19/mo →Provider
Learn how to use Z.ai: GLM 4.5V
💡 Tip: Start with the sample prompts above to see how Z.ai: GLM 4.5V works best.
Proven prompts shared by the community for this model