Model Details
Where is GLM 4.5V a perfect fit?
- Complex image understanding and visual reasoning
- Video understanding and event recognition
- OCR and document extraction
- Long documents and research reports
- Charts, tables and diagrams
- Screenshot/UI understanding
- GUI agents and desktop automation
- Visual grounding and spatial reasoning
- Front-end/web coding from screenshots
- Multimodal agent workflows
GLM-4.5V is the vision counterpart to the GLM-4.5-Air generation: a relatively efficient 12B-active MoE model that adds full-spectrum visual reasoning, video understanding, grounding and GUI-agent capabilities rather than merely adding image input to a text model.
Quick Model Estimate
Your GLM 4.5V Cost Estimate
๐ฐ Total Cost
โ
for 1000 input + 1000 output tokens
Cost Breakdown
Prices are watched for changes and checked against the provider's own page.
Pricing
Benchmarks
Scores from standardized evaluations by Artificial Analysis and Design Arena. Higher is better โ the indices summarize overall ability, while the detailed scores break down performance on individual benchmarks.
Artificial Analysis
Higher is better ยท benchmarked by Artificial AnalysisGraduate-level questions in biology, chemistry & physics
Expert-level questions across many academic domains
Precise following of detailed instructions
Tool-using agent tasks in a telecom support setting
Unpublished physics research reasoning problems
Long-context reasoning across large inputs
Complex command-line and terminal workflows
Breadth of factual knowledge across domains
How reliably the model avoids fabricated answers
| Variant | GPQA Diamond | Humanity's Last Exam | IFBench | ฯยฒ-Bench Telecom | AA-LCR | Terminal-Bench Hard | CritPt | Omniscience Accuracy | Omniscience Non-Hallucination |
|---|---|---|---|---|---|---|---|---|---|
| GLM-4.5V (Non-reasoning) | 57.3% | 3.5% | 28.6% | 19.6% | 0% | 6.8% | 0% | 18.2% | 9.8% |
| GLM-4.5V (Reasoning) | 68.4% | 6.3% | 34.2% | 22.5% | 0% | 5.3% | 0% | 20.8% | 15.2% |
Other Models in the GLM 4.5 Family
FAQs about GLM 4.5V
How much does GLM 4.5V cost per 1M tokens?
GLM 4.5V costs $0.60 per million input tokens, and $1.80 per million output tokens.
What does a typical workload cost with GLM 4.5V?
1,000 requests of 2,000 input and 500 output tokens each โ 2,000,000 input and 500,000 output tokens in total โ costs $2.10 with GLM 4.5V at its lowest rates. Use the calculator on this page for your own volumes.
Where can I use GLM 4.5V?
GLM 4.5V is available through Z.ai, from $0.60 per million input tokens.
What is GLM 4.5V's context window?
GLM 4.5V has a context window of 65,536 tokens, and returns up to 16,384 tokens in a single response.
What input types does GLM 4.5V support?
GLM 4.5V accepts text, image and video input, and returns text.
What is GLM 4.5V's knowledge cutoff?
GLM 4.5V's training data runs to 31 December 2024. It has no built-in knowledge of events after that date.
