Model Details
Where is GLM 4.6V Flash a perfect fit?
- High-volume image understanding
- OCR and document extraction
- Charts, tables and diagrams
- Screenshot and UI understanding
- GUI agents
- Visual grounding and object localization
- Video understanding
- Multimodal RAG
- Local/private vision applications
- Low-latency multimodal agents
- Visual tool calling
Quick Model Estimate
Your GLM 4.6V Flash Cost Estimate
๐ฐ Total Cost
โ
for 1000 input + 1000 output tokens
Cost Breakdown
Prices are watched for changes and checked against the provider's own page.
Pricing
Other Models in the GLM 4.6 Family
FAQs about GLM 4.6V Flash
How much does GLM 4.6V Flash cost per 1M tokens?
GLM 4.6V Flash costs nothing for input tokens, and nothing for output tokens.
What does a typical workload cost with GLM 4.6V Flash?
1,000 requests of 2,000 input and 500 output tokens each โ 2,000,000 input and 500,000 output tokens in total โ costs nothing with GLM 4.6V Flash at its lowest rates. Use the calculator on this page for your own volumes.
Where can I use GLM 4.6V Flash?
GLM 4.6V Flash is available through Z.ai, free of charge for input tokens.
What is GLM 4.6V Flash's context window?
GLM 4.6V Flash has a context window of 128,000 tokens, and returns up to 32,000 tokens in a single response.
What input types does GLM 4.6V Flash support?
GLM 4.6V Flash accepts text, image, video and file input, and returns text.
