NOT DEPRECATED
Listed by litellm, openrouter. No retirement date announced. Neither the provider's API nor its documentation has been checked for this model. Sources disagree on: input_price_per_mtok, max_output_tokens, output_price_per_mtok.
state=alive sunset=none days_remaining=n/a severity=low sources=litellm,openrouter evidence_tier=catalog confirmed_by_endpoint=false documented_by_provider=false replacement=none corroborated=yes divergent_fields=input_price_per_mtok,max_output_tokens,output_price_per_mtok checked=2026-09-30 15:40Z
zhipu/glm-4.6
Vendor: zhipu. Tracked by 2 of 5 sources.
What each source reports
| Source | ID | Context | Max out | In $/Mtok | Out $/Mtok | Sunset |
|---|---|---|---|---|---|---|
| litellm | novita/zai-org/glm-4.6 | 204,800 | 131,072 | $0.55 | $2.2 | — |
| openrouter | z-ai/glm-4.6 | 204,800 | 16,384 | $0.43 | $1.75 | — |
Change history
| Severity | Change | What | Detected |
|---|---|---|---|
| low | modified | together_ai/zai-org/GLM-4.6: context_tokens raised 200,000 to 202,752 | 2026-09-21 |
| high | modified | z-ai/glm-4.6: max_output_tokens cut 131,072 to 16,384 | 2026-09-19 |
| info | repriced | z-ai/glm-4.6: input_price_per_mtok $0.50/Mtok to $0.43/Mtok | 2026-09-19 |
| info | repriced | z-ai/glm-4.6: output_price_per_mtok $2.00/Mtok to $1.75/Mtok | 2026-09-19 |
| info | repriced | z-ai/glm-4.6: cache_read_price_per_mtok $0.10/Mtok to $0.08/Mtok | 2026-09-19 |
| low | modified | z-ai/glm-4.6: max_output_tokens raised 16,384 to 131,072 | 2026-09-19 |
| low | repriced | z-ai/glm-4.6: input_price_per_mtok $0.43/Mtok to $0.50/Mtok | 2026-09-19 |
| low | repriced | z-ai/glm-4.6: output_price_per_mtok $1.75/Mtok to $2.00/Mtok | 2026-09-19 |
| low | repriced | z-ai/glm-4.6: cache_read_price_per_mtok $0.08/Mtok to $0.10/Mtok | 2026-09-19 |
| high | modified | openrouter/z-ai/glm-4.6: max_output_tokens cut 131,000 to 16,384 | 2026-09-18 |
| low | modified | openrouter/z-ai/glm-4.6: context_tokens raised 202,800 to 204,800 | 2026-09-18 |
| low | modified | openrouter/z-ai/glm-4.6: capabilities gained response_schema | 2026-09-18 |
| info | repriced | openrouter/z-ai/glm-4.6: input_price_per_mtok $0.55/Mtok to $0.43/Mtok | 2026-09-13 |
| info | repriced | openrouter/z-ai/glm-4.6: output_price_per_mtok $2.20/Mtok to $1.75/Mtok | 2026-09-13 |
| info | repriced | openrouter/z-ai/glm-4.6: cache_read_price_per_mtok n/a to $0.08/Mtok | 2026-09-13 |
| high | modified | z-ai/glm-4.6: max_output_tokens cut 131,072 to 16,384 | 2026-09-09 |
| info | repriced | z-ai/glm-4.6: input_price_per_mtok $0.55/Mtok to $0.43/Mtok | 2026-09-09 |
| info | repriced | z-ai/glm-4.6: output_price_per_mtok $2.20/Mtok to $1.75/Mtok | 2026-09-09 |
| info | repriced | z-ai/glm-4.6: cache_read_price_per_mtok $0.11/Mtok to $0.08/Mtok | 2026-09-09 |
| medium | repriced | z-ai/glm-4.6: input_price_per_mtok $0.43/Mtok to $0.55/Mtok | 2026-09-08 |
| medium | repriced | z-ai/glm-4.6: output_price_per_mtok $1.75/Mtok to $2.20/Mtok | 2026-09-08 |
| medium | repriced | z-ai/glm-4.6: cache_read_price_per_mtok $0.08/Mtok to $0.11/Mtok | 2026-09-08 |
| low | modified | z-ai/glm-4.6: max_output_tokens raised 16,384 to 131,072 | 2026-09-08 |
| high | modified | z-ai/glm-4.6: max_output_tokens cut 131,072 to 16,384 | 2026-09-07 |
| info | repriced | z-ai/glm-4.6: input_price_per_mtok $0.55/Mtok to $0.43/Mtok | 2026-09-07 |
| info | repriced | z-ai/glm-4.6: output_price_per_mtok $2.20/Mtok to $1.75/Mtok | 2026-09-07 |
| info | repriced | z-ai/glm-4.6: cache_read_price_per_mtok $0.11/Mtok to $0.08/Mtok | 2026-09-07 |
| low | repriced | z-ai/glm-4.6: input_price_per_mtok $0.50/Mtok to $0.55/Mtok | 2026-09-06 |
| low | repriced | z-ai/glm-4.6: output_price_per_mtok $2.00/Mtok to $2.20/Mtok | 2026-09-06 |
| low | repriced | z-ai/glm-4.6: cache_read_price_per_mtok $0.10/Mtok to $0.11/Mtok | 2026-09-06 |
| medium | repriced | openrouter/z-ai/glm-4.6: input_price_per_mtok $0.40/Mtok to $0.55/Mtok | 2026-09-06 |
| medium | repriced | openrouter/z-ai/glm-4.6: output_price_per_mtok $1.75/Mtok to $2.20/Mtok | 2026-09-06 |
| info | repriced | z-ai/glm-4.6: input_price_per_mtok $0.55/Mtok to $0.50/Mtok | 2026-09-06 |
| info | repriced | z-ai/glm-4.6: output_price_per_mtok $2.20/Mtok to $2.00/Mtok | 2026-09-06 |
| info | repriced | z-ai/glm-4.6: cache_read_price_per_mtok $0.11/Mtok to $0.10/Mtok | 2026-09-06 |
| medium | repriced | z-ai/glm-4.6: input_price_per_mtok $0.43/Mtok to $0.55/Mtok | 2026-09-03 |
| medium | repriced | z-ai/glm-4.6: output_price_per_mtok $1.75/Mtok to $2.20/Mtok | 2026-09-03 |
| medium | repriced | z-ai/glm-4.6: cache_read_price_per_mtok $0.08/Mtok to $0.11/Mtok | 2026-09-03 |
| low | modified | z-ai/glm-4.6: max_output_tokens raised 16,384 to 131,072 | 2026-09-03 |
| low | added | new model deepinfra/zai-org/GLM-4.6 (202,752 ctx) | 2026-08-28 |