
GLM 4.6
智谱 AI (ChatGLM) • API Price Comparison
GLM 4.6 is a reasoning model by 智谱 AI (ChatGLM) with a 200K-token context window and a 128K-token maximum output. Across the 2 API channels tracked here, input pricing ranges from $0.29–$0.43 per 1M input tokens (USD-normalized); the cheapest is XycAi at ¥1.98 per 1M input tokens. Channels reachable from mainland China without a VPN include XycAi. Prices are collected daily from official provider pages and every tracked channel, then cross-checked by a data-accuracy audit.
Model performance benchmarks
Scores retain their original benchmark, version and metric. We do not combine them into a site-wide score; results from different versions are not directly comparable.
| Channel | Type | Input / 1M | Output / 1M | Rate Limit | China | vs Official | |
|---|---|---|---|---|---|---|---|
XycAi💰 Best Price | 转售商 | ¥1.98 | ¥7.91 | - | - | ||
OpenRouter5 variants | 聚合平台 | $0.43 | $1.75 | - | - | - | |
Show all price variants for this channelOpenRouter Auto / default routeDefault $0.43 / $1.75 per 1M default_route Venice | z-ai/glm-4.6 $0.43 / $1.75 per 1M upstream_provider DeepInfra | z-ai/glm-4.6 $0.50 / $2.00 per 1M upstream_provider Novita | z-ai/glm-4.6 $0.55 / $2.20 per 1M upstream_provider Z.AI | z-ai/glm-4.6 $0.60 / $2.20 per 1M upstream_provider | |||||||
| Usage Level | Tokens/Month | XycAi | OpenRouter |
|---|---|---|---|
| 🐣 Light | 0.1M in + 0.1M out | $0.09 | $0.13 |
| 📊 Medium | 1.0M in + 0.5M out | $0.86 | $1.31 |
| 🚀 Heavy | 10.0M in + 5.0M out | $8.60 | $13.05 |
| 🏢 Enterprise | 100.0M in + 50.0M out | $85.96 | $130.50 |
Frequently asked questions about glm-4.6
What is the cheapest API provider for GLM 4.6?
Of the channels tracked here, XycAi is the cheapest at ¥1.98 per 1M input tokens and ¥7.91 per 1M output tokens (about $0.29 per 1M input). Confirm on the provider's page; figures are updated daily.
Can I use GLM 4.6 from mainland China?
Yes — these channels are directly reachable from mainland China without a VPN: XycAi. Domestic channels generally support Alipay/WeChat Pay.
What is the context window of GLM 4.6?
GLM 4.6 offers a 200,000-token context window and a maximum output of about 128000 tokens.
How often are GLM 4.6 prices updated, and how reliable is the data?
Prices are scraped daily from official pricing pages and every tracked channel, with every significant change recorded in a price-history log. A daily read-only data audit cross-checks each channel against the model producer's published price and flags stale or inconsistent rows for correction.
Related pricing research