r/ZaiGLM

Image 1 — GLM-5.3 is out on AA, and I'm fed up with their Intelligence/cost plot
Image 2 — GLM-5.3 is out on AA, and I'm fed up with their Intelligence/cost plot
▲ 20 r/ZaiGLM+2 crossposts

GLM-5.3 is out on AA, and I'm fed up with their Intelligence/cost plot

I think AA's intelligence/cost plot is seriously misleading, so I decided to make my own. Their plot is in the second image.

All points are at max thinking. All intelligence index scores are from AA. All cost scores are from AA too except where noted below.

What changes between AA's plot and mine:

  • Changed X scale from logarithmic to linear, because people's money is not logarithmic
  • Added DeepSeek V4 Flash 0731 as it is priced today by third party providers on OpenRouter (note: you don't get this today with OpenCode Go/Zen, but it's been promised you will soon).
  • Added GLM-5.3 as it will be priced by third party providers on OpenRouter in <2 weeks, assuming no license changes from 5.2. Note: you don't get this on OpenCode Go/Zen.
  • Added Qwen3.8-27B. Cost per task was crudely calculated from
    • 47,166 output tok/task (AA)
    • tg 55 tok/s @ 350W, as crudely observed on my RTX3090 (IQ4_XS shows negligible quality loss - dedicated post coming soon)
    • today's US median residential electricity price
    • today's UK median residential electricity price + today's GBP/USD fx
    • +15% (finger-in-the-air) for prefill and waiting for tools
    • Hardware priced at 0, on the basis that both a RTX 3090 PC and a 64GB Strix Halo are desirable gaming/work machines anyways.
    • These maths are meant to produce a rough back-of-the-envelope figure and should not be taken authoritatively were you to zoom into the bottom-left corner of the chart. They don't want to answer how much cheaper it is to run Qwen at home vs. DSv4 on OpenRouter, because they are both so cheap that the difference is inconsequential for most of the population.

Note: not including the cost of hardware stops being defensible once you upgrade to a 128GB Strix Halo (almost nobody needs that much RAM if not for AI). This is why I did not add self-hosted DeepSeek IQ2_XXS to the chart; it would likely also sit lower on the intelligence axis than the MXFP4 native model. Same argument for a ~$16k rig needed to run GLM-5.3 IQ4 locally. I'm not saying they're not worth the expense (privacy is priceless), just that pegging them on the plot is a much more nuanced exercise.

u/crusaderky — 17 hours ago
▲ 334 r/ZaiGLM+1 crossposts

GLM-5.3 achieves 60 on the Artificial Analysis Intelligence Index, on par with Kimi K3 and up 7 points from GLM-5.2.

u/minxio_ — 1 day ago
▲ 15 r/ZaiGLM

150% usage is gone now in zcode

Edit: I just found out that they mean the 5h limit will automatically reset during off-peak hours when the weekly limit resets, what happened with me i was about 56% on the 5h and once the weekly resets the 5h limit resets to 100%
the zcode has removed the 150% badge, in the last update 3.8.1 and they added a sentence i don't understand to the changelog

  • GLM Coding Plan subscribers now receive a 5-hour quota in ZCode, which resets during off-peak hours after the limit is reached.

what do they mean?

reddit.com
u/muhamedyousof — 20 hours ago
▲ 5 r/ZaiGLM

I use Z.AI for creative writing and stuff and suddenly all my chats have been set to "choose ai model" and it's not letting me choose one, hence making them all unusable.

Same as title. Any ways to fix this?

reddit.com
u/MessageFriendly4035 — 23 hours ago
▲ 36 r/ZaiGLM

My take - GLM 5.3 MAX is better than cursor grok 4.6 EXTRA HIGH

Hello guys, I was working on web development, focused on frontend task.. Asked for doing things with something. I have pro license of cursor so I did experiment if the grok is even usable because I have big amount of usage left there.

For grok 4.6 it took (in latest cursor IDE) ~25 minutes, 60 files changed, 5 bugs... Even existing design and components significantly changed comparing to the initial state.

For GLM 5.3 - 1,5 hours... no bugs, 60 files changed, 12 phases, everything finally tested manually by agent across 3 resolutions. Doing my side hustles meantime because when I was waiting I was thinking about new stuff. Much BETTER work overally. Frontend UI even better visually, more appealing to use as promotional components, etc. I'm very happy how it is working.

Kudos, Z.AI 🎊

reddit.com
u/AbbreviationsOk6975 — 22 hours ago
▲ 2 r/ZaiGLM

Coding Plan -Legacy V1

Hi,

With the coding plan (i'm on Legacy V1 Pro), I used to be able to basically run one session non-stop for pretty much the 5 hours and then it reset with usage left, this was using their top model at the time.

I could only run out if I ran 2x projects at once and even then it was a huge amount of tokens per 5 hours and per day.

I thought the whole point of the Legacy V1 plans is we kept similar limits to before?

I just ran a completely new session (admittedly in their peak time.. but still). With GLM 4.7 as the cheap model and GLM 5.3 as the main model and the 5 hour session was used up in 39-40m tokens, by their own graphs.

I see its now 3x usage in peak (I thought was 2x, maybe changed or maybe I mis-remembered) But in the past there was no way this plan was only doing 120m every 5 hours... from memory i'm sure was more like 500m.

So basically my plan hasn't changed (great), no weekly limits (really great), but they appear to. have massively cut what my plan is given.....

Can anyone on the newer Pro plans confirm what they expect from a 5 hour window? Either in peak or off-peak as can easily be worked out the other.

Thanks!

reddit.com
u/Ok_Try_877 — 20 hours ago
▲ 97 r/ZaiGLM+1 crossposts

GLM-5.3 is already live in Command Code. 5–10x cheaper than the frontier closed models.

GLM-5.3 is already live in Command Code.

Same price as GLM-5.2, ~50% better on Zai's internal Code Bench.

Still 5–10x cheaper than the frontier closed models (per 1M output tokens, list price today).

Plans: $1 Go ($10 credits) · $10 GOAT ($20 credits)

GLM-5.3 is pretty cheap:
10x cheaper than Fable 5
6x cheaper than GPT-5.6 Sol
5x cheaper than Opus 5

🐐

u/maedahbatool — 1 day ago
▲ 34 r/ZaiGLM+3 crossposts

piodide ~ pi + pyodide + ghostty in the browser with WASM

https://daugasauron.github.io/piodide/

https://github.com/daugasauron/piodide

I just had this idea that you could just replace the bash tool in pi with python that runs on WASM, then you could run the whole thing in the browser.

Tried it and it's quite fascinating how good it is as a concept.

Everything is complete slop.

I tried GLM (code), OpenAI and moonshot providers.

daugasauron.github.io
u/PersonalityWild1379 — 1 day ago
▲ 0 r/ZaiGLM

Do people actually buy the legacy accounts

Asking since I have subscribed to the GLM Coding Max Plan on 4 accounts, wonder if this is real that people are looking for them.

reddit.com
u/Usual-Barnacle-5052 — 21 hours ago
▲ 4.4k r/ZaiGLM+3 crossposts

GLM-5.3 is out and the US tech Giants 🫈 are in panic mode once again

u/Jenna_AI — 3 days ago
▲ 22 r/ZaiGLM

Robbed by Z.AI? My usage numbers make absolutely no sense

I subscribed to Z.AI’s smallest Coding Plan on August 16th, two days ago, for $18.

My first experience was already pretty rough: I managed to hit the 2,000-credit 5-hour limit in about 20 minutes of coding. GLM stopped right in the middle of making changes and left the project in a non-compiling state.

Fair enough, I can accept that part. I’m responsible for planning around the limits of the plan I bought.

Today I let GLM continue the changes. The project compiles again, although I still need to go through everything and check whether some of the changes were destructive.

But then I looked at my usage dashboard, and this is where things stop making sense.

According to the usage graph:

  • August 16: roughly 2,000 credits used
  • Today: roughly 760 credits used
  • Expected total: ~2,760 credits

So out of a weekly allowance of 10,000, I should have roughly 7,240 credits remaining.

Instead, the usage report says I only have:

3,250 / 10,000 remaining.

That means their system has apparently counted about 6,750 credits of usage.

So there are roughly 3,990 credits missing/unaccounted for more than my entire visible usage today and August 16 combined.

For context, I have used GLM 5.3 exclusively through OpenCode. I have not used the web interface for GLM.

Am I misunderstanding how Z.AI calculates Coding Plan credits? Are there hidden multipliers, retries, tool calls, cached/context tokens, or anything else that gets charged but isn't shown in the usage graph?

Because if the dashboard is supposed to represent my billed usage, the numbers simply do not add up.

Has anyone else experienced this with Z.AI’s Coding Plan?

u/genrichh93 — 1 day ago
▲ 118 r/ZaiGLM

GLM 5.3 is a beast

For some reason this model hasn't been getting the hype it deserves in this competitive landscape, but I am very impressed. It's so much more autonomous in the way I can task and walk again and it correctly iterates through until completion and with high quality results and generally faster than before. It's really on par with SOTA models when it comes to coding for me. I use Claude models at work and that's my point of reference.

My current harness is OpenChamber (OpenCode VSCode extension). Running it with "autopilot" ON and MAX reasoning.

My reference link for a discount if you wanna try subscribing.

reddit.com
u/torontobrdude — 2 days ago
▲ 14 r/ZaiGLM

did they automatically route GLM 5.2 usage to GLM 5.3?

hey guys,

I've recently noticed some changes in the response from my AI Agent, I did not change the model from GLM 5.2 to GLM 5.3 but I felt the response is different and slower.

I just checked my coding plan usage today and just found out i'm using glm 5.3 instead.

do they automatically route this?

honestly I like GLM 5.2 better if what i've been using is GLM 5.3.

reddit.com
u/luckypanda95 — 2 days ago
▲ 50 r/ZaiGLM+3 crossposts

I made a SOTA browser agent that can use your codex subscription, optimized for luna

I wanted to make an open source browser harness that's cheaper, more accurate and faster to run than available alternatives, without requiring a large, expensive model. Using 5.6 Luna, it outperforms the next best option (BrowserCode) on BU Bench v1 and BrowserWebApp bench. Full benchmark results available on the repository.

You can use it with your codex subscription or directly using the OpenAI API. And of course, it was built with Codex itself :)

github.com
u/pierreb5 — 3 days ago