Model Details
Where is GLM 4.6V FlashX a perfect fit?
- High-volume image understanding
- Multimodal API applications
- OCR and document processing
- Screenshot and UI analysis
- Visual RAG
- Multimodal agents
- Visual grounding
- GUI-agent workflows
- Chart and table analysis
- Video understanding
- Applications where API cost and throughput matter more than using the full GLM-4.6V
GLM-4.6V-FlashX is best treated as the fast, inexpensive hosted serving variant of the GLM-4.6V multimodal family. Its defining characteristic is API economics, not a separately documented architecture or parameter count.
Quick Model Estimate
Your GLM 4.6V FlashX Cost Estimate
๐ฐ Total Cost
โ
for 1000 input + 1000 output tokens
Cost Breakdown
Prices are watched for changes and checked against the provider's own page.
Pricing
Other Models in the GLM 4.6 Family
FAQs about GLM 4.6V FlashX
How much does GLM 4.6V FlashX cost per 1M tokens?
GLM 4.6V FlashX costs $0.04 per million input tokens, and $0.40 per million output tokens.
What does a typical workload cost with GLM 4.6V FlashX?
1,000 requests of 2,000 input and 500 output tokens each โ 2,000,000 input and 500,000 output tokens in total โ costs $0.28 with GLM 4.6V FlashX at its lowest rates. Use the calculator on this page for your own volumes.
Where can I use GLM 4.6V FlashX?
GLM 4.6V FlashX is available through Z.ai, from $0.04 per million input tokens.
What is GLM 4.6V FlashX's context window?
GLM 4.6V FlashX has a context window of 128,000 tokens, and returns up to 32,000 tokens in a single response.
What input types does GLM 4.6V FlashX support?
GLM 4.6V FlashX accepts text, image, video and file input, and returns text.
