10% OFF your first GLM Coding subscription — invite token ROK78RJKNW · claim here →

Bonus inference, peak rates & token discounts

Every known z.ai bonus window, multiplier and limited-time discount — computed live in your browser with ticking countdowns, timezone-aware timelines and a Prometheus-compatible metrics endpoint. No stale screenshots, no guessing when to fire the heavy agent runs.

your local time

schedule time

Asia/Singapore · UTC+8

AI doc-sync starting…

live status · ai-synced hourly

Bonus windows

Toggle pills flip automatically as windows open and close. Countdowns tick every second; schedules are evaluated against Asia/Singapore time exactly as the official notices define them — and an AI bot re-verifies them against docs.z.ai every hour.

AI doc-sync starting…

z.ai

Peak hours surcharge

Time-variable rates apply during peak hours. Code off-peak and the same credits last twice as long.

status
syncing…
schedule
weekly Mon,Tue,Wed,Thu,Fri 14:00–18:00 Asia/Singapore
window tz
SGT (UTC+8)
  • GLM-5.33× quota at peak · 1× off-peak (legacy plans)
  • GLM-5.3-Flash1.2× quota at peak · 0.4× off-peak (legacy plans)
  • Credit plansstandard rate at peak · 50% credit rate off-peak
official docs

z.ai

GLM-5.3-Flash Usage Campaign

Nightly bonus window for GLM-5.3-Flash: unlimited usage in ZCode and doubled quota in other supported agents.

status
syncing…
schedule
2026-09-03..2026-09-20 daily 23:00–09:00 Asia/Singapore
event ends
2026-09-21 09:00 SGT
  • Via ZCodezero quota — unlimited GLM-5.3-Flash
  • Other agents2× plan quota for GLM-5.3-Flash
  • GLM-5.3standard rules apply (not included)
official docs

z.ai

Flash API −50% promo

Limited-time 50% discount on GLM-5.3-Flash API pricing — the list price is slashed across input, cached input and output tokens.

status
syncing…
schedule
until 2026-09-09 24:00 Asia/Singapore (always on)
event ends
2026-09-09 24:00 SGT
  • Input$0.15 → $0.075 per 1M tokens
  • Cached input$0.03 → $0.015 per 1M tokens
  • Output$0.50 → $0.25 per 1M tokens
official docs

planner

When to run what

The same 24 hours shown twice: once in schedule time (SGT), once in your local clock. Line up the bright segments and you never accidentally pay peak rates again.

Next 24 hours, visualised

peakflashflash-api
SGT (UTC+8)syncing…
00:0006:0012:0018:0024:00
your local time · UTC+00:00syncing…
00:0006:0012:0018:0024:00

Bars start at midnight in each timezone. The peak surcharge only runs Mon–Fri; the flash campaign paints the whole night. Wherever you are, the night bar is when the stack is deepest.

pay as you go

API token pricing

Prices per 1M tokens in USD. The GLM-5.3-Flash 50% promotion is live right now — input, cached input and output are all halved until 2026-09-09 24:00 SGT.

GLM-5.3

Flagship coding model

Input
$1.4
Cached input
$0.26
Output
$4.4

per 1M tokens · cached input storage free for a limited time

GLM-5.3-Flash

Fast · 50% off for a limited time

−50%
Input
$0.15$0.075
Cached input
$0.03$0.015
Output
$0.50$0.25

per 1M tokens · cached input storage free for a limited time

More API models

modelinput / cached input per 1Moutput per 1M
GLM-5.2$1.4 / $0.26$4.4
GLM-5.1$1.4 / $0.26$4.4
GLM-5$1.0 / $0.20$3.2
GLM-4.7$0.6 / $0.11$2.2
GLM-4.5-Air$0.2 / $0.03$1.1
GLM-4.7-FlashX$0.07 / $0.01$0.4
GLM-4.6V (vision)$0.3 / $0.05$0.9
GLM-4.7-FlashFreeFree

Full price list at docs.z.ai/guides/overview/pricing ↗

subscriptions

GLM Coding Plan tiers

Credits-based plans for Claude Code, Cline, OpenCode, ZCode and more. Off-peak usage bills at 50% of the standard credit rate — the same window the radar tracks above.

Lite

Daily driving for solo developers — quick fixes, codebase Q&A, completion-heavy work.

5-hour credits

2,000

weekly credits

10,000

  • GLM-5.3 + GLM-5.3-Flash
  • Vision, Web Search, Web Reader & Zread MCP
  • 50% credit rate off-peak
Get Lite — 10% off first order
popular

Pro

High-frequency feature work on real repositories — 6× the Lite allowance in both windows.

5-hour credits

12,000

weekly credits

60,000

  • 6× Lite credits
  • Same model + MCP access
  • Best value for daily repo work
Get Pro — 10% off first order

Max

All-day agentic coding on large codebases, long refactors and parallel agent fleets.

5-hour credits

28,000

weekly credits

140,000

  • 14× Lite credits
  • Up to ~4.4B Flash tokens / week at 98% cache hit
  • Built for autonomous agent runs
Get Max — 10% off first order

Credits refresh dynamically: the 5-hour bucket resets 5 hours after consumption, the weekly bucket resets every 7 days. Credit usage = (input × input multiplier + cached input × cached multiplier + output × output multiplier) ÷ 10,000 — and everything billed off-peak costs half. Source: GLM Coding Plan overview ↗

invite token · first order only

10% off your first GLM Coding subscription

Subscribe through this invite token and the discount is applied instantly at checkout — no manual activation. Valid for new users and existing accounts that have never had a paid subscription, on the first GLM Coding order only.

Claim 10% OFF
  • 10% is deducted from the first GLM Coding subscription order at checkout (Stripe minimum $0.50 applies).
  • One redemption per user (by phone / email); not stackable with other first-order promos.
  • Renewals, upgrades and follow-up orders bill at the standard price.
  • Rules: Invite Friends, Get Credits ↗