Subscription plan generations · not model generations
The plan everyone remembers as three dollars a month is gone. Here is exactly what replaced it, and what that costs you.
Z.ai’s has been rebuilt three times in twelve months. Two of those three generations can no longer be bought at any price, and the two that survive as “legacy” are treated very differently from one another — a distinction almost nobody makes, and the one that decides what you should do next.
All figures verified 2026-08-14 · 51 dimensions · 45 sources
- qualified“Legacy V1 was the most generous generation” 14254016
- confirmed“Legacy V1 is being phased out” 76
- corrected“The legacy plans are all going away” 6
- qualified“There are three generations” 614
Each of those verdicts is argued out in the section below, with the sources that decide it.
Entry price is the cheapest tier at the time. Select a generation to highlight it in every table on this page.
Orientation
If you have never heard of any of this
Four short answers, and then you can read the rest of the page like everyone else.
What a GLM Coding Plan actually is
Coding assistants like Claude Code, Cline and Roo Code talk to a language model over an API. Normally you pay that model’s vendor per request. A GLM Coding Plan is a flat monthly subscription that lets those same tools talk to Z.ai’s GLM models instead, at a fraction of the cost.
You change one setting — the address your coding tool points at — and your usage comes out of the subscription rather than off a card.
Why three generations is a real problem, not trivia
Each generation counts your usage in a different unit. The first two counted — one thing you ask is one prompt, no matter how much work it sets off. The current one counts , derived from .
That means the numbers you find in blog posts and comparison sites are frequently describing a plan that no longer exists, in a unit the current plan does not use. Several pricing sites we checked on 2026-08-14 still present the retired prompt figures as current.
Every row in the tables below says which generation it belongs to and how confident we are in it. That is the whole point of this page.
What the money means for a working day
The cheapest current tier is $18 a month, or $12.60 if you pay for a year. Z.ai’s own estimate is that it buys 43 to 87 million a week — the range depends on how much of your work lands in , when everything costs twice as much.
For most developers outside Asia that range’s top end is realistic, because peak is 06:00–10:00 UTC on weekdays and weekends are cheap all day.
How today’s offer compares to the high-water mark
In late 2025, $3 a month bought about 120 queries every five hours with no at all. Today’s $18 tier has a weekly ceiling and meters differently, so the two cannot be compared directly — but on price alone the entry point is six times what it was.
Z.ai used to advertise the plan as “3× usage, 1/7 cost”. It makes no equivalent claim now.
Every term on this page, defined
cache hit rate
The share of the text sent to the model that it has already seen and stored, and so charges less for.
Z.ai states 90.9% as the average for coding workloads, and uses that figure for its own token estimates. Cached input costs about a quarter of fresh input on GLM-5.3.
concurrency
How many requests you can have in flight at the same time — which decides whether you can run two projects, or a swarm of sub-agents, at once.
Z.ai has never published a number for this. It says only that the limits follow Max > Pro > Lite and are adjusted dynamically according to available capacity.
context window
How much text the model can hold in mind at once — your code, the conversation and everything it has read.
The current plan offers up to a million tokens on every tier, but you have to opt in by adding a suffix to the model name.
credit
A budget measured in text. Everything the model reads and writes is converted into credits by a published formula, with the model’s output costing far more than its input.
Credits = (input tokens × multiplier + cached input × multiplier + output × multiplier) ÷ 10,000. On GLM-5.3 the multipliers are 6.9, 1.7 and 24, so one output token costs about fourteen times what one cached input token costs.
GLM Coding Plan
A monthly subscription from Z.ai that lets you use its GLM models inside coding tools such as Claude Code, Cline or Roo Code, instead of paying per request.
You point your coding tool at a Z.ai endpoint instead of the model vendor’s own, and your usage comes out of the subscription rather than off a card. It is not a general API plan: using it from your own application is a terms violation.
grandfathered
Keeping the terms you originally signed up on after those terms stop being offered to new customers.
Here it is conditional in both directions: the first generation’s privilege required auto-renew to be switched on before a deadline, and Z.ai’s contract separately reserves the right to stop renewing anyone at all.
peak hours
A window when using the newest models costs more of your allowance. On this plan it is weekday afternoons in Singapore.
Monday to Friday, 14:00–18:00 Singapore time (UTC+8) — twenty hours a week. That is 06:00–10:00 UTC, so many developers outside Asia are almost never in it. Weekends are off-peak all day.
pro-rated
Splitting a charge by how much of the period you actually used.
On this plan the value that comes back from a pro-rated upgrade arrives as account credit, never as money returned to your card.
prompt quota
A budget measured in questions. One thing you ask counts as one prompt, no matter how much work the assistant does answering it.
Z.ai’s own definition: “One prompt refers to one query. Each prompt is estimated to invoke the model 15–20 times.” So a single prompt could cover a long chain of file reads, edits and test runs. The first two generations counted this way; the current one does not.
rate window
The stretch of time your allowance is measured over. Use it all inside the window and you wait for the window to refresh.
Every generation has used a five-hour window. On the current plan the refresh is rolling: each credit becomes available again five hours after you spent it, rather than everything resetting at one moment.
token
A chunk of text, roughly three-quarters of a word. Models read and write in tokens, and the current plan bills by them.
A medium source file might be two or three thousand tokens. A coding assistant re-sends much of the conversation on every turn, which is why cached input dominates the total.
weekly cap
A second, longer ceiling on top of the short one. You can be well inside your five-hour allowance and still be stopped for the rest of the week.
Introduced on 12 February 2026. The first generation had no weekly cap at all, which is the single biggest reason it is remembered as the generous one.
The claim under test
What people say about these plans, and whether it holds up
This page started from a widely-shared framing: that the first generation was the generous one, and that it is being phased out. We treated that as something to check rather than something to repeat.
“Legacy V1 was the most generous generation”Mostly true, with two real catches
True of V1 as it settled, and by a wide margin: about 120 user queries every five hours, no weekly ceiling at all, and new models arriving at no extra cost. One researcher measured that allowance at 40 million tokens every five hours — more, on the entry tier, than the current entry tier gets. Two catches, though. The famous $3 was a first-month price; the recurring price was $6, and Z.ai’s marketing went on quoting the $3 for months. And the famous 120 was not the launch figure either: the page captured on launch morning advertised 15 to 50, and was revised upward within hours. The generous V1 people remember is real, but it is a specific window, and it was never quite as cheap as the number everyone repeats.
“Legacy V1 is being phased out”True, and further along than “being”
Z.ai cancelled auto-renewal on every eligible V1 plan on 30 April 2026 and said legacy plans “will no longer remain as a supported subscription option going forward”. There is no renewal path. Every V1 subscription still alive today is running out its last paid term. It is not being phased out; it has been.
“The legacy plans are all going away”No — the two legacy cohorts are treated very differently
V2 is not on the same path. Z.ai’s July 2026 announcement explicitly lets V2 holders keep auto-renew switched on and upgrade tier or billing term. It offers no such right to V1. One legacy generation is terminal; the other can be renewed indefinitely, at least as of today. Anyone lumping “legacy” together is missing the single sharpest fact in the story.
Sources: 6
“There are three generations”Three is the right count, but only two are numbered
Z.ai names “Legacy Plan V1” and “Legacy Plan V2” in your account. The current one has no version number at all — the company just calls it the credits-based plan. “V3” is a community label. And V1 changed materially inside its own lifetime, so if you counted by terms rather than by Z.ai’s labels you could argue for four.
The spine
All three generations, side by side
51 dimensions across 3 generations. Sort it, filter it, hide what you do not care about, highlight one generation, or put two side by side and show only the change. Every cell carries a confidence badge and a numbered source.
| DimensionWhat is being compared | Legacy Plan V1Sold new: 1 September 2025 – 12 February 2026 | Legacy Plan V2Sold new: 12 February 2026 – 30 July 2026 | The credits planOn sale since 30 July 2026 | ChangeV2 → Credits |
|---|---|---|---|---|
| What Z.ai calls itThe label that appears in your account under “My plan”. | V1“Legacy Plan V1”Your account shows the words Legacy Plan V1. | V2“Legacy Plan V2”Your account shows the words Legacy Plan V2. | Credits“The new credits-based plan” — no version numberIt has no version number. Z.ai just calls it the credits plan.“V3” is a community shorthand. Z.ai numbers only the two retired generations. | |
| Can you buy it today?New subscriptions, not renewals. | V1No — withdrawn from sale, and auto-renew was force-cancelledNo. It is gone, and even people who had it were taken off automatic renewal. | V2No to new buyers — but existing holders may still renewYou cannot start one. If you already have one you can keep paying for it. | CreditsYesYes. This is the one on the pricing page. | |
| Sold to new customers duringWhen you could walk up and buy it. | V11 Sep 2025 → 12 Feb 2026From 1 September 2025 until 12 February 2026.Announced at 14:27 UTC on 1 September 2025 in a single post on X, with no launch blog post. The July 2025 date sometimes quoted belongs to the GLM-4.5 model, not to this subscription. | V212 Feb 2026 → 30 July 2026From 12 February 2026 until 30 July 2026. | Credits30 July 2026 → todayFrom 30 July 2026 until now. | |
| What ended the previous generationThe event that closed the door behind each generation. | V1— (the first generation)Nothing. It was the first. | V2A cutoff at 12 Feb 2026 (UTC+8): subscribe and enable auto-renew before it and you kept V1 terms for the life of the subscriptionA deadline on 12 February 2026. Get in before it with automatic renewal switched on and you kept the old, better terms. | CreditsThe 30 July 2026 announcement stopped selling the prompt plans and opened the credits planAn announcement on 30 July 2026 stopped selling the old plans and opened this one. | |
| Team versionTeam seats have always been metered on their own scheme. | V1Individual plans onlyIndividuals only.The July 2026 notice tags both legacy cohorts “individual plans only”. | V2Individual plans onlyIndividuals only. | CreditsStandard Seat $88/seat/mo, Premium Seat $188/seat/mo, minimum 2 seatsThere is a team version: $88 or $188 per person per month, minimum two people.Team seats meter in raw tokens, not credits, and are the only place overage billing exists. | |
| Lite — monthly list priceThe cheapest way in. | V1$3 for the first billing cycle, $6 recurringThree dollars for your first month, six a month after that.The famous $3 was a first-month offer, not the standing price. Z.ai’s own marketing kept saying “starting at just 3 USD per month” for months afterwards, which is how the number lodged itself in everyone’s memory. The recurring price was double that. | V2$10 / month, then $18 from around 19 April 2026Ten dollars a month at first, then eighteen.Two distinct price points inside one generation. The first-purchase discount disappeared at the same time, so a new buyer’s real cost went from $3 to $10 overnight. | Credits$18 / monthEighteen dollars a month. | |
| Pro — monthly list priceThe tier most working developers land on. | V1$15 for the first billing cycle, $30 recurringFifteen dollars for your first month, thirty after that.Same structure as Lite: the advertised figure was a first-month price and the recurring price was double it. | V2$30 / month, then $72 from around 19 April 2026Thirty dollars a month at first, then seventy-two.Pro was the one tier the February change left alone — Z.ai’s announcement named only Lite and Max. It then went up by 140% in April. | Credits$80 / monthEighty dollars a month.Also reported: $72 / month. Z.ai’s April 2026 migration notice printed $72 for Pro, and a third-party tracker still repeats it. That was the previous generation’s price. We read $80 as the struck-through list price on the live pricing page on 2026-08-14. 744 | |
| Max — monthly list priceThe heaviest individual tier. | V1Not offered at launch. Priced at twice Pro — so $30 first cycle, $60 recurring.There was no Max tier at first. When it came it was double the Pro price.Z.ai stated the two-times-Pro ratio officially; the dollar figures come from its archived checkout data. The tier appears in the documentation by 1 October 2025, a month after launch. | V2$80 / month, then $160 from around 19 April 2026Eighty dollars a month at first, then a hundred and sixty.A user watching it happen put it at $80 two weeks before 20 April and $160 on the day, which matches the archived checkout data to within a few days. | Credits$168 / monthA hundred and sixty-eight dollars a month.Also reported: $160 / month. $160 was the previous generation’s Max price and is still repeated by third-party trackers. 744 | |
| Discount for paying up frontWhat you save by committing to a quarter or a year. | V1No standing term discount. Annual was simply twelve months at list, with seasonal promotions on top.None as standard. There were seasonal offers instead.A Christmas 2025 promotion took 20% off annual on top of the first-purchase discount, which is how an annual Lite plan came to cost $28.80. | V2Quarterly −10%; annual −30%, cut to −20% at the April price riseTen percent off quarterly and thirty percent off yearly — later trimmed to twenty.Standing discounts appear for the first time in this generation. The annual discount was then reduced in the same move that raised the prices, so the annual cost rose by more than the headline monthly figure suggests. | CreditsQuarterly −20%, yearly −30%. Lite works out at $14.40 or $12.60 a month.Twenty percent off quarterly, thirty percent off yearly.Team seats get only 10%, not 30%. No page anywhere shows the total you actually pay up front — only the effective monthly rate. | |
| Priced in another currency?The mainland-China sibling sells the same thing in yuan. | V1Not establishedWe could not find out. | V2Not establishedWe could not find out. | CreditsYes: ¥118, ¥538 and ¥1,078 a month on bigmodel.cn, with identical allowancesYes. In China it is ¥118, ¥538 or ¥1,078 a month for exactly the same thing.At the yuan prices the tiers are not a straight currency conversion of the dollar ones. The Chinese pricing page also defaults to quarterly rather than yearly, so a careless comparison mixes two different billing terms. | |
| Cheapest real monthly costLite, on the longest commitment available. | V1$2.40 / month — $28.80 for a year, in the December 2025 promotionTwo dollars forty a month, if you caught the Christmas offer.Twenty percent off an annual Lite plan that already carried a first-purchase discount. This is the cheapest the GLM Coding Plan has ever been. | V2$7 / month early on ($84 a year), rising to $14.40 ($172.80 a year)Seven dollars a month at first if you paid yearly, later fourteen forty. | Credits$12.60 / month on yearly billingTwelve dollars sixty a month if you pay yearly.The effective monthly rate shown on the pricing page at the yearly −30% setting. | |
| Discounts on offerThings that reduce the sticker price. | V1Introductory pricing was itself the promotionThe low price was the offer. | V2A 50% migration discount for V1 holders, from 30 Apr 2026 until three months after their legacy term endedPeople coming off the first generation got half price for a while.It floats: it is always half of whatever the current discounted price is, not a frozen figure, and it could be used repeatedly inside the window. | Credits10% off a first order via a referral link; the old 50% migration discount still works hereTen percent off your first order through someone’s invite link.The referrer only gets paid after three successful invites, which is easy to miss. | |
| What gets countedThe unit your allowance is denominated in. | V1Prompts. One prompt = one user query, estimated to invoke the model 15–20 times.Your questions. One question counts as one, however much work it sets off behind the scenes. | V2Prompts, same definitionYour questions, counted the same way. | CreditsCredits = (input×mult + cached input×mult + output×mult) ÷ 10,000Tokens — the raw text going in and out, converted into credits by a formula.This is the single biggest change between generations. Under prompts, a runaway agent loop inside one query cost you one prompt. Under credits it costs whatever it burns. A captured response from Z.ai’s undocumented quota endpoint suggests the older plans were typed internally as token limits all along, so the prompt counts may always have been a presentation layer over the same thing. | |
| Five-hour allowance, LiteHow much you can use in one sitting on the cheapest tier. | V1~120 prompts — measured at about 40M tokensAbout 120 questions, which one researcher measured as 40 million tokens.The token figure is a community measurement, cross-checked against Z.ai’s own quota endpoint while that endpoint still returned real numbers. It is the only measurement anywhere that lets the generations be compared in one unit. | V2~80 prompts — about 27M tokens on the same measurementAbout 80 questions.A third less than before. The token equivalent is our arithmetic: the cut was exactly two thirds across all three tiers. | Credits2,000 credits ≈ 8.6M tokens at peak, 17.2M off-peak2,000 credits, which is a lot of text — twice as much outside busy hours.Derived by scaling Z.ai’s own weekly token estimate for Lite down to the five-hour allowance. Notice what this implies: measured against the first generation’s 40 million, the current entry tier’s sitting allowance is smaller. | |
| Five-hour allowance, ProSame question, one tier up. | V1~600 promptsAbout 600 questions. | V2~400 promptsAbout 400 questions. | Credits12,000 credits12,000 credits. | |
| Five-hour allowance, MaxSame question, top tier. | V1~2,400 promptsAbout 2,400 questions. | V2~1,600 promptsAbout 1,600 questions. | Credits28,000 credits28,000 credits. | |
| Weekly ceilingThe cap that decides whether a heavy week is possible at all. | V1NoneThere wasn’t one.Z.ai’s own shorthand for the first generation was “legacy plans (without weekly usage limits)”. | V2Lite ~400, Pro ~2,000, Max ~8,000 prompts per week — exactly five times the five-hour figureYes: about 400, 2,000 or 8,000 questions a week depending on tier.The weekly ceiling landed days after the grandfathering cutoff, not on it. One dated community record says it arrived at four times the five-hour allowance on 14 February and was loosened to five times on the 16th — so the very first version was 20% tighter than the figures that ended up in the documentation. | CreditsLite 10,000, Pro 60,000, Max 140,000 credits per weekYes: 10,000, 60,000 or 140,000 credits a week. | |
| Most you could use in a week, LiteThe ceiling if you never stopped. | V1~4,000 prompts ≈ 1,344M tokens — no weekly stop, only the five-hour oneAbout 4,000 questions, because only the five-hour limit could stop you.Our arithmetic: 40 million tokens per five-hour window, 33.6 windows in a week. No real person sustains that; it is the shape of the ceiling, not a forecast. | V2~400 prompts ≈ 133M tokens — the weekly cap binds firstAbout 400 questions. The weekly limit stops you long before anything else.And that is before the model multiplier: on the newest models a token counted two or three times, so the real ceiling was closer to 45–65M tokens of flagship use. | Credits10,000 credits ≈ 43–87M tokens10,000 credits, which Z.ai estimates at 43 to 87 million tokens.Z.ai’s own estimate, assuming all use is on its newest model at a 90.9% cache-hit rate. The range is peak versus off-peak. Read against the row above, the current entry tier lands close to the second generation’s effective ceiling and far below the first generation’s. | |
| Z.ai’s own value claimWhat the company said you were getting. | V1“About 3× the usage quota of the Claude Pro plan”, at “only ~1% of standard API pricing”Three times a Claude Pro allowance, for about one percent of normal API prices. | V2“Approximately 15–30× the monthly subscription fee” in API-equivalent valueFifteen to thirty times the fee, valued at API prices. | CreditsNo equivalent multiplier is claimedThey stopped making that claim.Absence of evidence: we found no comparable multiple on any current page. The credit table and token estimates replaced the marketing claim. | |
| Search and browse tool allowanceThe extra tools that sit alongside the model. | V1Not documented in the snapshots we recoveredWe could not find out. | V2Lite 100, Pro 1,000, Max 4,000 web search + reader calls per monthA monthly count of web searches: 100, 1,000 or 4,000 by tier.This detail survives only on the archived April 2026 page. | Credits1.2 credits per call, from the same pool as everything elseEach search costs 1.2 credits out of your normal allowance. | |
| Short windowThe one that bites during a long session. | V15 hoursFive hours. | V25 hoursFive hours. | Credits5 hours, refreshed rolling: each credit returns 5 hours after it was spentFive hours, but rolling — what you spend comes back five hours later, bit by bit.A rolling refresh, not a clock-aligned bucket. You are never waiting for a fixed reset time. | |
| Weekly resetWhen the weekly counter goes back to zero. | V1Not applicableNot applicableThere was no weekly counter to reset. | V27-day cycleEvery seven days. | Credits7 days from the moment you subscribed — not MondayEvery seven days, counted from when you bought it, not from Monday. | |
| Peak-hours penaltyUsage costs more in a narrow weekday window, Singapore time. | V1Not documented for this generationWe found no evidence of one.Peak coefficients first appear in the second generation’s documentation. Absence of evidence, not evidence of absence. | V2Newest models cost 3× at peak and 2× off-peakUsing the best model cost three times as much in busy hours, twice as much the rest of the time.Also reported: 3× peak / 1× off-peak. This is how Z.ai restated the same generation’s terms in July 2026, after it had ended. The pages subscribers actually read at the time said 2× off-peak, with 1× offered as a limited-time benefit that was extended repeatedly — “valid through the end of April”, then “through the end of June”. The two official texts do not agree, and the difference is a factor of two on most people’s usage. 620 | CreditsOff-peak costs 50% of the standard credit rateWorking outside the busy window is half price.Same window, reframed: the second generation anchored on off-peak and penalised peak; this one anchors on peak and rewards off-peak. The ratio also narrowed from 3:1 to 2:1. | |
| When peak actually isThis matters more than it sounds. | V1Not documentedWe could not find out. | V214:00–18:00 Singapore time, every dayAfternoons in Singapore, seven days a week.Also reported: Monday to Friday only. The weekdays-only qualifier appears in the July 2026 restatement, not in anything a subscriber read at the time. The contemporary pages say simply “peak hours are 14:00–18:00 (UTC+8)”, with no day restriction — so weekend afternoons appear to have been charged at peak rates during this generation. 6 | CreditsMon–Fri 14:00–18:00 SGT; all weekend is off-peakWeekday afternoons in Singapore. Weekends are cheap all day.Twenty hours a week are peak. That window is 06:00–10:00 UTC — so a developer in Europe or the Americas may almost never be in it. | |
| How many things at onceParallel sessions and sub-agents. | V1Published as numbers on launch day: 10 concurrent requests on Lite, 30 on ProTen at once on Lite, thirty on Pro — printed on the page.Those figures were on the pricing page on 1 September 2025 and were quietly removed within days. No later generation has published a number. | V2Numbers withdrawn. In practice one user observed a limit of one request in flight.No numbers any more. One user found he could only run one request at a time.Also reported: Officially: no number at all, only “Max > Pro > Lite”. Z.ai replaced the published figures with an ordering and a dynamic-adjustment clause. A Pro subscriber then documented reaching only 4% of his five-hour quota before hitting concurrency errors. One report is not a specification — but there is no specification to check it against, which is the point. 4130 | CreditsNo published number. Guidance only: Lite one project, Pro 1–2, Max 2+, with Max > Pro > Lite and higher limits off-peakNo hard number. Lite suits one project at a time, Pro one or two, Max more.Z.ai says the limits are adjusted dynamically according to available capacity, so there is no figure to hold them to. This may matter more than any allowance on this page: if throughput binds first, the ceilings never come into it. | |
| SpeedHow fast tokens actually arrive, which can bind before any quota does. | V1Not documentedWe could not find out. | V2Users reported congestion during peak and slower throughput than ClaudePeople said it slowed down in busy hours. | CreditsPro advertises “faster generation speeds”, Max “dedicated resources during peak times” — no figuresHigher tiers are advertised as faster. No numbers are given.Implies the cheapest tier may be speed-limited, but Z.ai publishes no throughput figures for any tier. | |
| Models you can callWhat the subscription actually gets you access to. | V1GLM-4.5 era modelsThe GLM-4.5 generation of models.The 2025 snapshot advertises GLM-4.5 in coding tools. Model line-ups shifted repeatedly during this plan’s life. | V2GLM-5.x and GLM-4.x models of the dayWhatever the current models were at the time. | CreditsGLM-5.3, GLM-5-Turbo, GLM-4.7 — and nothing elseThree models: GLM-5.3, GLM-5-Turbo and GLM-4.7.Requests for older models are silently routed to GLM-5.3. Calling anything outside the three produces a balance error. | |
| Do cheaper tiers get worse models?A question most people assume the answer to. | V1No. Every tier got every model, and new models arrived at no extra cost.No. Everyone got everything, including new models as they came out.GLM-4.5 became 4.6 became 4.7 with no price change. That is a real and under-remembered part of why this generation is remembered fondly. | V2Yes, for a while. Lite was excluded from GLM-5 at launch and calling it returned an error; resolved by May 2026.Yes, briefly. The cheapest tier could not use the newest model at first.The only period in which model access genuinely depended on tier. Users who had bought an annual Lite plan expecting the newest model were, in their own words, not pleased. | CreditsNo. All three tiers reach all three models.No. Every tier gets every model.Higher tiers get earlier access to new models and more capacity, not different models. | |
| Context windowHow much of your codebase the model can hold at once. | V1Not documented per tierWe could not find out. | V2Not documented per tierWe could not find out. | CreditsUp to 1,000,000 tokens, opt-in with a [1m] model suffix. Same for every tier.Up to a million tokens if you switch it on. The same on every tier.Context is a property of the model, not the tier. One Z.ai config page still shows an older model name for this setting. | |
| Search, browsing and visionThe tools bundled alongside the model. | V1Not documented in the snapshots we recoveredWe could not find out. | V2Web search and reader, on a monthly countWeb search, with a monthly allowance. | CreditsVision understanding, web search, web reader and Zread — on every tierImage understanding, web search, page reading and a repo reader. All tiers.The pricing page’s Pro card implies these are a Pro feature; two documentation pages say all plans get them. We follow the documentation. | |
| Where you may use itThe restriction newcomers trip over most. | V1Inside supported coding toolsInside approved coding apps. | V2Inside supported coding toolsInside approved coding apps. | CreditsOnly inside ~19 named tools. General API access from your own code is a terms violation.Only inside about nineteen named coding apps. You cannot use it as a normal API.This is not an API plan. Calling it from your own app, bot or website is explicitly prohibited and may get benefits restricted. | |
| Behaviour at the limitBlocked, throttled, queued, or billed. | V1Blocked until the five-hour window refreshedYou waited for the next five-hour window. | V2Blocked until the relevant window refreshedYou waited for the window to refresh. | CreditsHard block. No overage, no fallback to pay-per-call.You are stopped. Nothing is charged and nothing keeps running.Z.ai’s FAQ only mentions waiting for the five-hour cycle. If it is the weekly ceiling you hit, waiting five hours does nothing — you wait for the seven-day reset. | |
| Can you pay to keep going?An escape hatch, or the absence of one. | V1Not documentedWe could not find out. | V2Not documentedWe could not find out. | CreditsIndividual: no. Team seats: yes, opt-in overage at API list price less 10%.Not as an individual. Team accounts can switch on pay-as-you-go.The 10% overage discount is described as a limited-time offer with no stated end date. | |
| Does it eat your account balance?A common worry. | V1Not documentedWe could not find out. | V2Not documentedWe could not find out. | CreditsNo. Plan calls only ever use plan quota.No. It cannot quietly spend your balance. | |
| Auto-renewalWhether the plan renews itself. | V1Force-cancelled for all eligible holders on 30 April 2026Switched off for everyone on 30 April 2026.Applied to holders who had auto-renew enabled on that date. Z.ai never says what happened to holders who had already switched it off. | V2Still available: holders may keep auto-renew on and keep renewingStill on. You can keep renewing indefinitely. | CreditsOn by default at the end of each cycleRenews automatically unless you stop it.The legal terms describe renewal as something you opt into at checkout; the plan usage policy describes it as automatic. Two official pages, two framings. | |
| Is your price locked?The assumption most grandfathered subscribers are quietly making. | V1Terms held for the life of the subscription — but the subscription was endedYour terms were safe right up until the plan itself was retired. | V2Terms hold to the end of each billing cycleSafe until the end of the cycle you paid for. | CreditsNo. Renewal is charged at the price in force on the charge date.No. You pay whatever the price is on the day the card is charged.The contract says this outright, and adds that continuing to use the service after a price change counts as accepting it. There is no price-lock guarantee anywhere. | |
| Can Z.ai simply stop renewing you?The clause that made the first generation’s ending lawful. | V1Yes — and it didYes. That is exactly what happened. | V2Yes, the same clause appliesYes, the same clause covers it. | CreditsYes: Z.ai may “unilaterally discontinue providing automatic renewal services” as its operational strategy requires, with noticeYes. The contract lets them stop renewing anyone, with notice.This is the single most important sentence for anyone counting on a grandfathered price. | |
| Cancellation noticeHow long before your billing date you must act. | V1Same general termsThe same rules as now. | V2Same general termsThe same rules as now. | CreditsAt least 3 days before the billing date — on two Z.ai pagesThree days before your billing date, if you go by the safer of two answers.Also reported: 24 hours before the end of the current term. A genuine two-against-two split inside Z.ai’s own documentation. The usage policy and the pricing-page FAQ say at least three days; the docs FAQ and the legal subscription terms say twenty-four hours. Cancel three days out and you cannot be caught by either reading. 410 | |
| RefundsGetting money back. | V1NoneNone. | V2NoneNone. | CreditsNone, even for unused quotaNone, even if you never used it.Three official pages agree. The only carve-out is where law requires one. | |
| Upgrading mid-termWhat happens to the time you already paid for. | V1Cannot switch to a current plan before expiryYou cannot move early. You wait for it to run out. | V2May switch early only by upgrading a tier; the old plan ends immediately and its unused value goes toward the new priceYou can jump early only by moving up a tier. Doing so ends the old plan on the spot.A one-way door. Once a V2 holder upgrades early, V2 is gone. | CreditsCross-tier upgrade takes effect immediately with pro-rated value returned as account balance. Same-tier changes stack instead: monthly to annual gives you 13 months.Moving up a tier happens straight away. Moving from monthly to yearly adds on rather than replacing — you end up with thirteen months.The thirteen-month outcome surprises people. Pro-rated value comes back as account credit, never as cash. | |
| DowngradingMoving to a cheaper tier. | V1Not availableNot possible. | V2Queued to the end of the current termIt waits until your current term ends. | CreditsTakes effect after the current cycle ends. No proration.Starts after your current cycle ends. Nothing is refunded. | |
| Sharing the subscriptionOne plan, how many people. | V1Same general termsThe same rules as now. | V2Same general termsThe same rules as now. | CreditsLicensed to one named person. Sharing, renting, reselling or proxying is prohibited; enforcement runs rate limiting → freeze → ban after three violations.One person only. Sharing or reselling gets you throttled, frozen, then banned.The restriction is on multiple users, not multiple devices. One person on several machines is nowhere prohibited. | |
| Is your code used to train models?Worth knowing before you point it at a private repository. | V1Same general termsThe same rules as now. | V2Same general termsThe same rules as now. | CreditsFor individual users, yes: Z.ai reserves the right to process your prompts and the model’s outputs to develop and improve its models, and you consent to it. API and enterprise customers get the opposite default.Yes, if you are an individual. Your prompts and the answers can be used to improve their models. Business API customers are exempt by default.We found no opt-out documented for individuals — only an invitation to contact them. The Coding Plan is sold to individuals, so this is the relevant default. | |
| Liability capWhat you can recover if it goes wrong. | V1Same general termsThe same rules as now. | V2Same general termsThe same rules as now. | CreditsCapped at your spend in the most recent calendar monthWhatever you paid last month, and no more. | |
| Can a current holder keep renewing?The question that decides everything else. | V1No. Auto-renew was cancelled on 30 April 2026 and the renewal right is absent from the July 2026 notice.No. It ends when your paid time runs out.The July 2026 notice lists a renewal right for V2 and conspicuously does not list one for V1, which matches April’s force-cancellation. Z.ai never states the negative outright. | V2Yes. “If you currently have a legacy plan V2, you can enable auto-renew or upgrade your plan.”Yes. You can keep paying for it as long as you like. | CreditsYes — it is the plan on saleYes. It is the current plan. | |
| What forces you offThe specific events that end a grandfathered plan. | V1Expiry of the current paid term. Nothing else is needed — renewal was already removed.Time. When your paid period ends, so does the plan. | V2Letting it lapse, or upgrading a tier early — which ends the plan immediatelyStopping payment, or jumping up a tier early.Z.ai never states what happens if you lapse. Since the plan is “no longer sold to new users”, the strong reading is that a lapse is irreversible. That inference is ours. | CreditsNothing — it is the current planNothing yet. | |
| If you lapse, can you get it back?The irreversibility question. | V1NoNo. | V2Almost certainly not — but Z.ai never says so directlyProbably not, though they never spell it out.Inference from “previous plans are no longer sold to new users” plus a renewal right described as available only “while your plan is active”. No Z.ai sentence states the consequence of lapsing. | CreditsNot applicable — you can just subscribe againYou can just subscribe again. | |
| Can you pause it?For a holiday, a job change, a quiet quarter. | V1No pause feature is documentedNo. | V2No pause feature is documented — and cancelling is a one-way exitNo. And cancelling means you probably cannot come back.We searched the usage policy, FAQ and subscription terms. Cancel is the only lever, and it ends renewal. | CreditsNo pause feature is documentedNo, but you can resubscribe whenever. | |
| What you were given in exchangeMigration support offered when each generation closed. | V1Two complimentary months of the matching current tier, granted automatically on 30 April 2026 and added after the existing term; plus 50% off until three months after the legacy term ended, reusable within that window.Two free months of the equivalent new plan, plus half price for a while afterwards.Tier-matched and stacked after the existing term rather than replacing it. The 50% is a floating half of the current price, not a frozen figure. | V2Nothing beyond keeping the plan — which, since it stays renewable, is arguably the better dealNothing extra, but you get to keep the plan. | CreditsNot applicableNot applicable. | |
| Closest thing on sale todayWhere a holder of a closed generation would land. | V1Lite at $18/month for the price, or Pro at $80/month to get near the old headroom. Neither restores an uncapped week.Lite at $18 a month if you care about price; Pro at $80 if you care about headroom. Neither gives back the unlimited week. | V2The same tier of the credits plan. Z.ai says the generations “differ only in how usage is calculated”.The same tier of the current plan.Z.ai’s claim that only the calculation changed is a claim, not a measurement. Nobody has published a conversion between prompts and credits. | CreditsYou are already on itYou already have it. | |
| What your old money buys nowThe same monthly spend, priced at today’s list. | V1$3/month buys no plan today. The cheapest entry is $18, or $12.60 on yearly billing.Three dollars buys nothing now. The cheapest is $18 a month, or $12.60 if you pay yearly. | V2$18/month still buys Lite — the same nominal price, a different unit of account.Eighteen dollars still buys the entry plan. | Credits$18/month buys Lite; $80 Pro; $168 MaxEighteen for Lite, eighty for Pro, a hundred and sixty-eight for Max. | |
Nothing matches those filters.
Try a shorter search term, or reset the table.
Twelve months
Scrub through what happened
Drag the handle, or use the arrow keys. 14 dated moments from launch to today — three price rises, one quota cut, one grandfathering deadline and two sunsets.
Read the whole timeline as a list
- 1 Sep 2025The GLM Coding Plan launches, on X and nowhere elseZ.ai announces it at 14:27 UTC with a single post: one seventh the price of the Claude Code plans, three times the prompts. Two tiers — Lite at $3 for your first month, Pro at $15 — reverting to $6 and $30 afterwards. Access to GLM-4.5 and GLM-4.5-Air.There is no launch blog post; the announcement is the post. The July 2025 date sometimes quoted belongs to the GLM-4.5 model, not to this subscription.242521
- Same dayThe allowance is revised eightfold upward within hoursThe page captured at 11:20 UTC advertises about 15–50 queries per five hours on Lite and 80–240 on Pro, alongside hard concurrency figures — 10 requests at once on Lite, 30 on Pro. By 17:31 the same day it reads about 120 and 600.Both captures are of Z.ai’s own page, hours apart. The 120 figure is the one that stuck and the one that held until February 2026 — but the generous generation’s opening position lasted about six hours. The concurrency numbers were dropped within days and no later generation ever published one again.1424
- Oct 2025Newer models arrive at the same priceGLM-4.6, and later GLM-4.7, become available to existing subscribers without a price change. This is a genuine and often-forgotten part of why the first generation is remembered fondly.15
- Dec 2025A Max tier, a Christmas deal, and a sloganA third tier is documented at about 2,400 queries per five hours. The site banner reads “3× usage, 1/7 cost”, and the docs claim the plan delivers roughly three times a Claude allowance at about one percent of standard API pricing.Community reports from the same period describe a Black Friday offer of a full year of Pro for around $120 — about $10 a month.15
- 11 Feb 2026Z.ai announces price rises — and only price risesA post on X: first-purchase discounts removed, Lite and Max going up, existing subscribers keeping their pricing. A companion post adds, plainly, “To be upfront: compute is very tight.” Pro’s list price is untouched, exactly as announced.What the announcement does not mention is the thing that mattered more. The 33% cut to the five-hour allowance and the brand-new weekly ceiling were never announced anywhere — they arrived as an edit to a documentation page.2622
- 12 Feb 2026The grandfathering cutoffSubscribe and switch auto-renew on before this date, Singapore time, and your original quota holds for the life of the subscription with no weekly limit. Miss it and you are on the new terms.This is the boundary between the two legacy generations, in Z.ai’s own words. Note what the privilege was conditional on: auto-renew already being enabled. On the same day, Z.ai’s quota endpoint stopped returning token counts and began returning only percentages — which is why nobody has been able to measure the plan precisely since.1640
- 13–16 Feb 2026The allowance is cut and a weekly ceiling appearsLite drops from about 120 queries per five hours to 80, Pro from 600 to 400, Max from 2,400 to 1,600. Then a weekly cap lands on top: 400, 2,000 and 8,000 queries — exactly five times the five-hour figure.One dated community record says the weekly cap arrived on the 14th at four times the five-hour allowance and was loosened to five times on the 16th. Either way it landed after the grandfathering deadline had already passed, so nobody weighing that deadline knew what they were choosing between.172840
- 7 Mar 2026An apology, and fifteen days of creditAfter weeks of reliability complaints, Zhipu emails subscribers on the China side compensation equivalent to fifteen days of their subscription fee, with no expiry.Reposted and translated by users rather than published, and it is not clear whether international subscribers received the same. It corroborates the congestion complaints appearing on forums in the same period.40
- Early Apr 2026First price rise landsThe documentation now reads “starting at just 10 USD per month, with Pro plans from 30 USD per month”. With the first-month offer gone, a new buyer’s real entry cost has gone from $3 to $10.1922
- 12 Apr 2026Second price rise: everything roughly doubles againLite reaches $18, Pro $72, Max $160 — and the annual discount is cut from 30% to 20% in the same move, so an annual buyer is hit twice. No announcement was made at all.Z.ai’s own checkout data shows the same product identifiers carrying the new prices. A user watching it happen wrote that Max “was $80 two weeks ago, and now it’s $160”; another that a three-month Lite plan had gone from $27 to $48.23403233
- 21 Apr 2026The Legacy Plan Migration NoticeZ.ai announces that legacy plans — the ones without weekly limits — will be phased out and will not remain a supported option. It acknowledges, in its own words, that this lands on top of a recent price increase.7
- 30 Apr 2026V1's auto-renewal is switched off for everyoneExisting paid periods run to their end, but no V1 subscription can renew. In compensation, eligible holders get two complimentary months of the matching current tier and 50% off until three months after their legacy term ends.The 50% is a floating half of whatever the current discounted price is, and it can be used repeatedly inside the window — so a V1 holder could stack a long term at half price before the door closed.7
- 30 Jul 2026Everything switches to creditsThe unit of account changes from queries to tokens, converted into credits by a published formula. Both prompt-based generations stop being sold. Z.ai names them Legacy Plan V1 and Legacy Plan V2 for the first time.The asymmetry is set here: V2 holders may keep auto-renew on and upgrade; V1 holders may do neither and must wait for expiry.6
- 14 Aug 2026Where things stand as we checkedThe credits plan is on sale at $18, $80 and $168 a month. V2 remains renewable with no announced end date. Every remaining V1 subscription is counting down to its final billing date. GLM-5.3 became the flagship model on the plan today, with the credit multipliers unchanged.Everything on this page was verified on this date. Z.ai has changed plan terms three times in twelve months, so treat anything here as perishable — and note that most third-party comparisons you will find are describing a generation that no longer exists, in a unit the current plan does not use.13640
Burn it down
Push your week through all three generations at once
Describe how you actually work. We run that week against each generation’s ceilings and show you which ones stop you. This is the only place on the page where we bridge between and , and that bridge is an assumption you can move.
Your week, in numbers
- Queries a week
- —
- Queries in the heaviest five hours
- —
- Tokens a week
- —
- Credits a week on the current plan
- —
Which tier is being tested
Comparing like with like: the same tier name across all three generations. The first generation never published a Max price, so that cell stays empty rather than being invented.
Which model you lean on
This matters more than it looks. The first generation charged the same for everything; the second charged up to three times for its newest models; the current one charges less for the older one and halves everything off-peak.
Every assumption behind those bars, and why
Cross-generation comparison requires a bridge that Z.ai has never published. We have made ours visible and adjustable rather than burying it. If the bridge is wrong, the current generation’s bars move and the two legacy ones do not.
Calibrate yourself
How generous was it really?
Five figures. Guess each one before you see it. Most people are wrong in the same direction, and the gap between what you guessed and what is true is the actual lesson.
I’d rather just read the answers
- ?The plan launched in September 2025 advertising $3 a month. What did it actually cost from month two onward?Answer: $625211
- ?How many user queries could a Lite subscriber send every five hours by late 2025?Answer: 1201440
- ?After February 2026, how many queries per five hours did the same Lite tier get?Answer: 80176
- ?What share of the week counts as expensive “peak” time on the current plan?Answer: 12%36
- ?Between early and late April 2026, the Max tier went from $80 a month to what?Answer: $160327
For newcomers
Which tier, if any, fits you
Five questions about how you actually work. The answer is sometimes “none of them”, and we will say so.
For anyone holding a legacy plan
What you have, what you are about to lose, and what to do about it
Four questions. The answer is a reasoned verdict with the trade-offs stated, not a recommendation to buy anything. If holding is right, it will say hold. If leaving is right, it will say that too.
How much you have uncovered
The picture so far
Nothing here is required. Everything on this page is readable without earning a single one of these — they are a record of what you have actually looked at.
- OrientedRead what the plan actually isnot yet earned
- Table readerUsed a table controlnot yet earned
- Difference spotterCompared two generations directlynot yet earned
- Source checkerOpened a citationnot yet earned
- Jargon busterLooked up a definitionnot yet earned
- Time travellerScrubbed the whole timelinenot yet earned
- Burn testerRan a workload past a ceilingnot yet earned
- CalibratedFinished the generosity guessesnot yet earned
- MatchedCompleted the plan matchernot yet earned
- DecidedGot a personal assessmentnot yet earned
Take your result with you
Built from what you did on this page. Nothing is sent anywhere — this page makes no network requests at all.
What to do next
- Holding V1: find your renewal date. It is the date your plan stops existing. Decide before it, not after — there is no path back.
- Holding V2: you are the cohort that can still renew. Do not upgrade a tier casually; an early upgrade ends the legacy plan immediately and permanently.
- On the credits plan: check what share of your work is inside 06:00–10:00 UTC. Moving work out of it halves what it costs you.
- Considering it: confirm your coding tool is on the supported list first, and read the training clause if the code is not yours.
Then check the figures yourself. Every one of them has a number beside it that opens the source.
Show your working
Every source, in full
45 of them. Live Z.ai pages, dated archive snapshots of Z.ai pages that have since changed, and dated community reports where no official record survives.
- GLM Coding Plan pricing pagehttps://z.ai/subscribe OfficialZ.airead 2026-08-14The live pricing page, read in a real browser in all three billing-term states on 2026-08-14. It is client-rendered, so a plain fetch returns no prices at all. It also opens on the Yearly tab, which is why so many write-ups quote $12.60 as “the monthly price” — it is the effective monthly rate at the yearly −30% setting.
“$12.6/month $18/month … $56/month $80/month … $117.6/month $168/month”
- The same plan on bigmodel.cn, in Chinesehttps://docs.bigmodel.cn/cn/coding-plan/overview OfficialZhipu / bigmodel.cnread 2026-08-14Z.ai’s mainland-China sibling runs the identical credits system: the same three models, the same five-hour and weekly credit allowances, the same hard stop when they run out. Prices differ — ¥118, ¥538 and ¥1,078 a month — and its pricing page defaults to quarterly rather than yearly. A few details, such as an auto-renew ceiling and an uplift on legacy team plans, appear only in the Chinese notices.
“套餐类型 5 小时积分 每周积分 / Lite 套餐 2,000 10,000 / Pro 套餐 12,000 60,000 / Max 套餐 28,000 140,000”
- GLM Coding Plan — plan overview and credit tablehttps://docs.z.ai/devpack/overview OfficialZ.ai docsread 2026-08-14The primary specification page for the current credits-based plan: credit allowances, the credit formula, per-model multipliers, peak hours and the token estimates.
“Each plan is subject to both a 5-hour usage limit and a weekly usage limit.”
- GLM Coding Plan FAQhttps://docs.z.ai/devpack/faq OfficialZ.ai docsread 2026-08-14Covers what happens at the limit, upgrade and downgrade mechanics, cancellation and refunds.
“Once the quota is used up, you’ll need to wait until the next 5-hour cycle for it to refresh.”
- GLM Coding Plan usage policyhttps://docs.z.ai/devpack/usage-policy OfficialZ.ai docsread 2026-08-14Concurrency guidance, the renewal and cancellation rules, account-sharing prohibition and the enforcement ladder.
“Rate (concurrency) limits are tied to your plan tier … the general principle being Max > Pro > Lite.”
- Plan Update Announcement — the move to creditshttps://docs.z.ai/devpack/notice/usage-revision OfficialZ.ai docspublished 2026-07-30The launch document for the current generation, and the only page where Z.ai names its own grandfathered cohorts “Legacy Plan V1” and “Legacy Plan V2”. It also tabulates the superseded prompt quotas, which is why those figures are quotable at all.
“The new credits-based plan is now available. Previous plans are no longer sold to new users.”
- Legacy Plan Migration Notice — the V1 sunsethttps://docs.z.ai/devpack/transition OfficialZ.ai docspublished 2026-04-21The document that ended the first generation: auto-renew cancelled on 2026-04-30, plus the two complimentary months and the 50% migration discount offered in compensation.
“Legacy plans (without weekly usage limits) will be phased out and will no longer remain as a supported subscription option going forward.”
- Switching to the latest model / 1M contexthttps://docs.z.ai/devpack/latest-model OfficialZ.ai docsread 2026-08-14Model availability across tiers and how the one-million-token context is switched on.
“To enable GLM 1M context, add the [1m] suffix to the model name (e.g., glm-5.3[1m])”
- Supported tools and API endpointshttps://docs.z.ai/devpack/tool/others OfficialZ.ai docsread 2026-08-14The list of coding agents the subscription may legally be used inside, and the base URLs each needs.
“Incorrect endpoint configuration will result in inability to use GLM Coding Plan subscription quota.”
- Subscriptions, Fees, and Payment (legal terms)https://docs.z.ai/legal-agreement/subscription-terms OfficialZ.ai legalread 2026-08-14The binding contract. Contains the clause that pre-authorises exactly what happened to the first generation, and the clause that defeats any assumption of a locked-in price.
“the price for automatic renewal shall be the price … actually applied by the system on the date the charge is made, rather than the original price at the time of your initial subscription.”
- Terms of Usehttps://docs.z.ai/legal-agreement/terms-of-use OfficialZ.ai legalread 2026-08-14Defines User Content as your prompts and the model’s outputs, and sets a different default for individuals than for API and enterprise customers.
“For individual users, we reserve the right to process any User Content to improve our existing Services and/or to develop new products and services”
- Team Plan benefitshttps://docs.z.ai/devpack/teamplan OfficialZ.ai docsread 2026-08-14The separate team track, which meters differently again and is the only place overage billing exists.
“A minimum of 2 seats is required, with no upper limit on the number of seats”
- Credit campaign and referral ruleshttps://docs.z.ai/devpack/credit-campaign-rules OfficialZ.ai docslast updated 2026-03-15Referral discounts and their conditions. Note the page’s own date predates the July replatform.
“the final amount payable after applying all discounts and Credits must be at least 0.50 USD”
- Plan overview as it stood on 2025-09-09 (archived)https://web.archive.org/web/20250909015313/https://docs.z.ai/devpack/overview Archived officialInternet Archive / Z.ai docssnapshot 2025-09-09The clearest surviving statement of first-generation terms, on Z.ai’s own page, before it was rewritten. This is where the $3 entry price and the 120-prompt allowance come from.
“Lite Plan: Up to ~120 prompts every 5 hours — about 3× the usage quota of the Claude Pro plan.”
- Plan overview as it stood on 2025-12-23 (archived)https://web.archive.org/web/20251223043504/https://docs.z.ai/devpack/overview Archived officialInternet Archive / Z.ai docssnapshot 2025-12-23The first snapshot in which a Max tier appears alongside Lite and Pro in the first generation.
“Max Plan: Up to ~2400 prompts every 5 hours — about 3× the usage quota of the Claude Max (20x) plan.”
- Plan overview as it stood on 2026-03-10 (archived)https://web.archive.org/web/20260310070146/https://docs.z.ai/devpack/overview Archived officialInternet Archive / Z.ai docssnapshot 2026-03-10The grandfathering promise itself, in Z.ai’s own words. It fixes the generational boundary at 2026-02-12 and makes the privilege conditional on auto-renew already being switched on.
“For users who subscribed and enabled auto-renewal before February 12 (UTC+8), the original quota will remain in effect throughout the subscription validity period, and no weekly usage limits will apply.”
- Plan overview as it stood on 2026-02-13 (archived)https://web.archive.org/web/20260213210131/https://docs.z.ai/devpack/overview Archived officialInternet Archive / Z.ai docssnapshot 2026-02-13Captured the day after the cutoff. Shows the second generation’s lower five-hour ceiling already live.
“Lite Plan: Up to ~80 prompts every 5 hours”
- Plan overview as it stood on 2026-03-13 (archived)https://web.archive.org/web/20260313094957/https://docs.z.ai/devpack/overview Archived officialInternet Archive / Z.ai docssnapshot 2026-03-13The $3 entry price was still printed here a month after the quota change, which is why snapshot dates are upper bounds on when a change actually took effect rather than the change date itself.
“Starting at just 3 USD per month, with Pro plans from 15 USD per month”
- Plan overview as it stood on 2026-04-04 (archived)https://web.archive.org/web/20260404102848/https://docs.z.ai/devpack/overview Archived officialInternet Archive / Z.ai docssnapshot 2026-04-04Catches the first of the two 2026 price rises in flight.
“Starting at just 10 USD per month, with Pro plans from 30 USD per month”
- Plan overview as it stood on 2026-04-29 (archived)https://web.archive.org/web/20260429132114/https://docs.z.ai/devpack/overview Archived officialInternet Archive / Z.ai docssnapshot 2026-04-29Second-generation detail that no longer exists anywhere live: monthly counts for the search tools, and the off-peak coefficient as it was worded before the current page.
“Lite Plan: Include a total of 100 web searches and web readers per month”
- Z.ai’s own pricing endpoint, captured 2025-12-23https://web.archive.org/web/20251223123822id_/https://api.z.ai/api/biz/pay/batch-preview Archived officialInternet Archive / api.z.aisnapshot 2025-12-23The checkout API returned every plan’s list price, discounted price and campaign name as JSON, and the archive caught it. This is where the first generation’s real list prices come from — and it shows that the famous $3 was a first-purchase discount off a $6 list, not the standing price.
“Lite $6→$3, Pro $30→$15, Max $60→$30 monthly, all via “First Purchase Discount””
- Z.ai’s pricing endpoint, captured 2026-03-01, 03-27 and 04-05https://web.archive.org/web/20260327195211id_/https://api.z.ai/api/biz/pay/batch-preview Archived officialInternet Archive / api.z.aisnapshots 2026-03-01 to 2026-04-05Three captures agreeing on the second generation’s first pricing: Lite $10, Pro $30, Max $80 a month, with a standing 10% quarterly and 30% annual discount. Two of the three were captured from a session that was not subscribed, so these are the prices a new buyer saw.
“"campaignName":"-30% per Year"”
- Z.ai’s pricing endpoint, captured 2026-04-19https://web.archive.org/web/20260419181845id_/https://api.z.ai/api/biz/pay/batch-preview Archived officialInternet Archive / api.z.aisnapshot 2026-04-19The same product identifiers now carry the higher prices — Z.ai raised prices on existing items rather than creating new ones — and the annual discount has been cut from 30% to 20% in the same move.
“"monthlyOriginalAmount":18.00,"monthlyPayAmount":14.40 … "campaignName":"-20% per Year"”
- Z.ai announces the GLM Coding Plan on Xhttps://x.com/Zai_org/status/1962522757536887205 Official@Zai_org2025-09-01The launch, at 14:27 UTC on 1 September 2025. There was no launch blog post — this post is the announcement. The July 2025 date sometimes given belongs to the GLM-4.5 model, not to the subscription.
“Announcing GLM Coding Plan for Claude Code! … 1/7th the price of original Claude Code plans - 3x more prompts.”
- Cline’s write-up of the plan at launchhttps://cline.bot/blog CommunityCline2025-09A third-party integrator spelling out what the pricing page’s “1st Month Offer” badge meant: the famous $3 was a first-billing-cycle price and the recurring price was $6. We could not deep-link the individual post, so this points at the blog index — but the same fact is independently visible in Z.ai’s archived checkout data, which shows $6 and $30 as the list prices with a “First Purchase Discount” applied on top.
“these prices are for the first billing cycle after which the price will increase to $6/$30”
- Z.ai’s price-change announcement on Xhttps://x.com/Zai_org/status/2021656635668901985 Official@Zai_org2026-02-11The only announcement Z.ai made for the February change, and it covered price alone. A companion post said plainly: “To be upfront: compute is very tight.” The quota cut and the brand-new weekly ceiling were never announced anywhere — they arrived as silent edits to the documentation.
“Existing subscribers keep their current pricing.”
- Plan overview as it stood on 2026-02-06 (archived)https://web.archive.org/web/20260206230642/https://docs.z.ai/devpack/overview Archived officialInternet Archive / Z.ai docssnapshot 2026-02-06The last archived page showing first-generation terms. The word “weekly” appears nowhere on it. Together with the 13 February capture, it brackets the change to a single week.
“Lite Plan: Up to ~120 prompts every 5 hours”
- Plan overview as it stood on 2026-02-18 (archived)https://web.archive.org/web/20260218174544/https://docs.z.ai/devpack/overview Archived officialInternet Archive / Z.ai docssnapshot 2026-02-18The second generation’s specification in its settled form: the two-column quota table, the grandfathering sentence, and the peak multiplier as it was originally written — three times at peak and twice off-peak, not the one-times that later documentation claims.
“Its usage will be deducted at 3 × during peak hours and 2 × during off-peak hours.”
- Plan overview as it stood on 2026-05-15 (archived)https://web.archive.org/web/20260515135237id_/https://docs.z.ai/devpack/overview.md Archived officialInternet Archive / Z.ai docssnapshot 2026-05-15Late second-generation state. Shows the one-times off-peak rate as a temporary benefit with an expiry date, repeatedly extended — which is why the later claim that off-peak was simply one-times does not match what subscribers were told at the time.
“As a limited-time benefit, GLM-5.1 and GLM-5-Turbo will only consume 1× quota during off-peak hours, valid through the end of June.”
- The usage policy, first archived 2026-03-16https://web.archive.org/web/20260316185324id_/https://docs.z.ai/devpack/usage-policy Archived officialInternet Archive / Z.ai docssnapshot 2026-03-16This page did not exist during the first generation. Its arrival mid-second-generation is itself a finding: account sharing rules, non-coding-use restrictions and a three-strikes ban policy all appear for the first time here.
“it may be subject to risk control measures, including high-intensity throttling, account suspension, or permanent ban”
- Pricing-page FAQ content, captured 2026-04-23https://web.archive.org/web/20260423115543id_/https://api.z.ai/api/biz/operation/query?ids=1136 Archived officialInternet Archive / api.z.aisnapshot 2026-04-23The pricing page’s FAQ text, captured through the API that feeds it. Documents the one period when model access really was gated by tier, and the list of supported tools at the time.
“The Lite plan currently does not include GLM-5 quota… If you call GLM-5 under the plan endpoints, an error will be returned.”
- Hacker News: Max tier doubled in about two weekshttps://news.ycombinator.com/item?id=47838634 CommunityHacker News (user Kerrick)2026-04-20A dated user observation that brackets the second price rise more tightly than the archive can.
“It was $80 two weeks ago, and now it’s $160.”
- Hacker News: quarterly price moved $27 to $48, and a weekly cap hit on day onehttps://news.ycombinator.com/item?id=47855184 CommunityHacker News (user UncleOxidant)2026-04-21Two useful things in one comment: an independent price observation, and the only concrete report we found of a full week’s prompt allowance being burned inside a single day.
“I must’ve somehow hit their weekly limit on day one of the week.”
- Hacker News: 7 million tokens consumed 2% of a weekly quotahttps://news.ycombinator.com/item?id=47708846 CommunityHacker News (user recursivegirth)2026-04-09Implies roughly 350 million tokens a week on that tier, which brackets Z.ai’s own current estimate for Pro. A rare case of a user measurement and an official figure agreeing.
“7 million tokens has only gone through 2% of my weekly usage”
- Hacker News: $180 for a year, bought in Decemberhttps://news.ycombinator.com/item?id=48182676 CommunityHacker News (user rescbr)2026-05-18A first-generation annual price seen in the wild. The tier is not stated, so treat the mapping to Pro as inference rather than fact.
“Last December I paid $180 for an year of Z.ai’s coding plan.”
- Hacker News: off-peak is where the value sitshttps://news.ycombinator.com/item?id=48108363 CommunityHacker News (user brokegrammer)2026-05-12Confirms the off-peak discount was live and widely understood during the second generation.
“During off-peak hours, usage for GLM-5.1 uses only 1x of your quota”
- Hacker News: tighter limits after the price risehttps://news.ycombinator.com/item?id=48097115 CommunityHacker News (user aspectrr)2026-05-11Subjective, and the published quota numbers did not change in that window, so this most likely describes throughput and concurrency rather than quota.
“I like the GLM coding plan before they raised their prices, now their rate limits are more strict”
- Hacker News: good value, slower throughputhttps://news.ycombinator.com/item?id=48631479 CommunityHacker News (user nijave)2026-06-22Relevant because wall-clock speed, not quota, may be what actually limits a heavy day.
“z.ai seems a bit on the slower side for raw model tok/sec throughput.”
- Hacker News: plan abuse degraded service in early Aprilhttps://news.ycombinator.com/item?id=47641008 CommunityHacker News (user mariopt)2026-04-04The most commonly offered explanation for why caps and prices moved. Z.ai never gave a reason, and the token figure the poster cites is an unrelated marketplace total, so treat the causal story as plausible rather than established.
“Last week Z.ai coding plan was unusable due to a lot of people abusing the coding plan”
- jia.je — a dated knowledge base of Chinese coding planshttps://jia.je/kb/en/software/coding_plan.html CommunityJiegec (jia.je)read 2026-08-14An unusually rigorous community record: dated entries, stated methodology, explicit “tested” versus “speculated” labels, and figures for the current plan that match Z.ai’s own documentation exactly. It is the only source anywhere that measured what a prompt was worth in tokens, cross-checked against Z.ai’s own quota endpoint before that endpoint stopped returning numbers.
“speculated that GLM Coding Plan’s Lite plan usage limit is that the sum of all requests' input + output tokens does not exceed 40M per 5 hours (meaning each prompt corresponds to 40M/120=333K tokens)”
- GitHub issue: a Pro subscriber reaching 4% of quota before being rate-limitedhttps://github.com/anomalyco/opencode/issues/8618 CommunityGitHub2026-01-15The most consequential community finding on this page. Z.ai endorses sub-agents and says concurrency follows Max > Pro > Lite, but publishes no numbers; this subscriber observed a practical limit of one request in flight. If that is representative, fanning work out across sub-agents does not burn quota faster — it fails instead.
“I consistently hit 'Too much concurrency' errors and can only utilize approximately 4% of my 5-hour quota before being rate-limited.”
- GitHub issue: a captured payload from Z.ai’s undocumented quota endpointhttps://github.com/robinebers/openusage/issues/1104 CommunityGitHub2026-08-13A real captured response from an endpoint Z.ai does not document. Two things fall out of it: Z.ai does internally version the current plan as V3, and the previous generation was typed as a token limit — which suggests the published prompt counts were always a presentation layer over a token budget.
“New Z.ai coding plans (subscription entries report "version": "V3" …) return quota entries with "type": "CREDIT_LIMIT" instead of "TOKENS_LIMIT"”
- Reddit: a heavy Lite session measured in tokenshttps://www.reddit.com/r/ZaiGLM/comments/1r587il/glm_47_lite_plan_getting_frequent_5_hour_quota/ Communityr/ZaiGLMaround February 2026The closest thing to a citable heavy-coding hour anywhere. Its implied 200,000 to 250,000 tokens per prompt independently brackets the 333,000 measured elsewhere. Reddit shows only relative dates, so the date is approximate.
“even with 100+ prompts, I usually stayed around ~20–25M tokens and never hit the 5-hour rate limit”
- codingplan.org tracker entry for GLMhttps://codingplan.org/en/plans/glm Communitycodingplan.orgread 2026-08-14A third-party tracker whose weekly-credit figures match the official docs exactly, which is why it is quoted here as corroboration. It still names an older model as current, so it lags by weeks.
“quarterly 20% and yearly 30% discounts”
- Arithmetic performed for this page#methodology DerivedThis explainer2026-08-14A figure produced by combining cited numbers rather than one Z.ai published. The working is set out in the methodology note, and the inputs are cited beside it.
The small print
How this page was built, and where it will go wrong
What we did
We read every live Z.ai page describing the plan — the pricing page, the plan overview, the FAQ, the usage policy, the team plan page, the two migration announcements and the legal subscription terms. For the retired generations, whose terms no longer exist anywhere live, we worked from dated Internet Archive snapshots of Z.ai’s own documentation, which is server-rendered and therefore readable in the archive. The pricing page is not: from late September 2025 onward the archive captured only an empty shell, so archived prices come from the documentation rather than from the pricing page.
Where no record survives at all, we used dated community reports and labelled them as such. Where two sources disagree, we show both readings rather than picking one quietly.
What we are least sure about
- The exact day of the first price rise, which we could only bracket to a two-week window in March or early April 2026. The second rise is pinned to 12 April by a dated changelog and two corroborating user reports.
- The first generation’s Max tier price in dollars. Z.ai stated the ratio — twice the Pro price — officially, but the dollar figures come from archived checkout data rather than from a published price list.
- Whether a lapsed legacy plan can ever be repurchased. Z.ai never addresses it. We infer no from “previous plans are no longer sold to new users”, and we label that inference.
- Concurrency on any generation. It has never been published as a number.
- The bridge between prompts and credits in the simulator. That is our assumption, not Z.ai’s, and it is adjustable for exactly that reason.
Where sources contradict each other
- Cancellation notice. Two Z.ai pages say at least three days; two say twenty-four hours. A clean two-against-two split. Use three days and you cannot be wrong.
- What a prompt was worth. The English announcement estimates each legacy prompt at 15–20 model calls; the Chinese version of the same notice says 15–30.
- Auto-renewal. The plan usage policy describes it as automatic; the legal terms describe it as something you opt into at checkout.
- The off-peak coefficient on the second generation. An archived April 2026 page said 2× off-peak with a temporary 1× offer; the live page says a flat 1×.
- Whether V1 holders can renew. One summary line in the July 2026 announcement says holders of plans discontinued on 30 July can still renew — but the V1-specific section grants no renewal right, and V1 was discontinued in April, not July. Read with April’s force-cancellation, the summary line appears to be about V2. We flag it rather than resolve it.
What could change after 2026-08-14
Z.ai has changed plan terms three times in twelve months and its contract reserves the right to change prices, to amend the agreement on publication, and to “unilaterally discontinue providing automatic renewal services” as its operational strategy requires. Renewals are charged at the price in force on the day the card is charged, not the price you originally signed up at. Treat every figure on this page as true on the verification date and perishable after it.