
glm-flash
智谱 AI (ChatGLM) • API Price Comparison
glm-flash is a large language model by 智谱 AI (ChatGLM) with a 1.3M-token context window. Across the 1 API channels tracked here, input pricing ranges from $0.08 per 1M input tokens; the cheapest is OpenRouter at $0.08 per 1M input tokens. None of the tracked channels are directly reachable from mainland China without a proxy or VPN. Prices are collected hourly from official provider pages and every tracked channel, then cross-checked by a daily data-accuracy audit.
| Channel | Type | Input / 1M | Output / 1M | Rate Limit | China | vs Official | |
|---|---|---|---|---|---|---|---|
OpenRouter💰 Best Price | 聚合平台 | $0.08 | $0.25 | - | - | - |
| Usage Level | Tokens/Month | OpenRouter |
|---|---|---|
| 🐣 Light | 0.1M in + 0.1M out | $0.02 |
| 📊 Medium | 1.0M in + 0.5M out | $0.20 |
| 🚀 Heavy | 10.0M in + 5.0M out | $2.00 |
| 🏢 Enterprise | 100.0M in + 50.0M out | $20.00 |
Frequently asked questions about glm-flash
What is the cheapest API provider for glm-flash?
Of the channels tracked here, OpenRouter is the cheapest at $0.08 per 1M input tokens and $0.25 per 1M output tokens (about $0.08 per 1M input). Confirm on the provider's page; figures are updated hourly.
Can I use glm-flash from mainland China?
None of the tracked channels are directly reachable from mainland China; reaching the official or international channels normally requires a proxy or VPN.
What is the context window of glm-flash?
glm-flash offers a 1,310,720-token context window.
How often are glm-flash prices updated, and how reliable is the data?
Prices are scraped hourly from official pricing pages and every tracked channel, with every significant change recorded in a price-history log. A daily read-only data audit cross-checks each channel against the model producer's published price and flags stale or inconsistent rows for correction.