GLM 4.6V FlashX Pricing - Cost Calculator

GLM-4.6V-FlashX is Z.aiโ€™s cost-efficient multimodal vision model optimized for fast API inference, visual reasoning, document analysis, video understanding, and agentic applications.

Released 8 December 2025, GLM 4.6V FlashX is a video-generation model from Z.ai. It costs $0.04 per million input tokens through Z.ai. Text, image, video and file input goes in; text comes out. It reasons step by step.

Last updated Oct 1, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

๐Ÿ”’ We respect your privacy. Unsubscribe anytime.

Model Details

Released
Dec 8, 2025
Context Length
128,000
Max Output
32,000
Modalities
Text Image Video File → Text
Capabilities
Tool use Open weights

Where is GLM 4.6V FlashX a perfect fit?

GLM-4.6V-FlashX is perfect for below use-cases:
- High-volume image understanding
- Multimodal API applications
- OCR and document processing
- Screenshot and UI analysis
- Visual RAG
- Multimodal agents
- Visual grounding
- GUI-agent workflows
- Chart and table analysis
- Video understanding
- Applications where API cost and throughput matter more than using the full GLM-4.6V

GLM-4.6V-FlashX is best treated as the fast, inexpensive hosted serving variant of the GLM-4.6V multimodal family. Its defining characteristic is API economics, not a separately documented architecture or parameter count.

Quick Model Estimate

(USD 0.0400 per 1M tokens)
(USD 0.4000 per 1M tokens)

Your GLM 4.6V FlashX Cost Estimate

๐Ÿ’ฐ Total Cost

โ€”

for 1000 input + 1000 output tokens

๐Ÿ“ฅ Input (1000 ร— $0.040000) โ€”
๐Ÿ“ค Output (1000 ร— $0.400000) โ€”

Cost Breakdown

๐Ÿ“ฅ Input ๐Ÿ“ค Output

Prices are watched for changes and checked against the provider's own page.

Pricing

Provider โ†•
Modality โ†•
Service Tier โ†•
Input Price
(per 1M tokens)
โ†•
Output Price
(per 1M tokens)
โ†•
Cached Input
(per 1M tokens)
โ†•
Context Size โ†•
View
Z.ai LogoZ.aiTextStandard$0.0400$0.4000$0.0040128,000 tokensโ†’

FAQs about GLM 4.6V FlashX

How much does GLM 4.6V FlashX cost per 1M tokens?

GLM 4.6V FlashX costs $0.04 per million input tokens, and $0.40 per million output tokens.

What does a typical workload cost with GLM 4.6V FlashX?

1,000 requests of 2,000 input and 500 output tokens each โ€” 2,000,000 input and 500,000 output tokens in total โ€” costs $0.28 with GLM 4.6V FlashX at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use GLM 4.6V FlashX?

GLM 4.6V FlashX is available through Z.ai, from $0.04 per million input tokens.

What is GLM 4.6V FlashX's context window?

GLM 4.6V FlashX has a context window of 128,000 tokens, and returns up to 32,000 tokens in a single response.

What input types does GLM 4.6V FlashX support?

GLM 4.6V FlashX accepts text, image, video and file input, and returns text.