cgorlla · 41 points · 19 comments · l’altro ieri · Open original
Comments
5 preview comments · loading full thread
Log in to use comments
Log in to h4cker, then connect Hacker News to publish comments.
NInijavel’altro ieri
Error messages matching Z.ai GLM I think are the simplest/most compelling. I had Opus 4.8 poke it and it came back with a couple different errors than the article mentions.
Matching the tokenizer is interesting tho
FRFrannkyl’altro ieri
I was using it side by side with GLM 5.3, and they were very, very similar. Also, the new zhipu 1 GW data center plus a flash(smaller?) model, can justify the 100T/day they said they were able to serve.
Pretty cool model, especially since it's not a nanny, if you want to unlock your own devices, like rooting an Android, it will happily help instead of flagging you.
Available for free via OpenRouter and Nous free tier. Also via OpenCode Go, but you have to pay a $5–$10 subscription. APIs are hammered now, so service is bumpy.
DAdangl’altro ieri
Recent and related:
Ox-Alpha Is GLM? - https://news.ycombinator.com/item?id=49422226 - Aug 2026 (65 comments)
A mysterious free AI model is impressing developers. Nobody knows who made it - https://news.ycombinator.com/item?id=49406289 - Aug 2026 (4 comments)
Ox Alpha - https://news.ycombinator.com/item?id=49381896 - Aug 2026 (202 comments)
JOjohndoughl’altro ieri
Another strong hint is that the uptime graph of GLM-5.3 by Z.ai is very similar to that of Ox Alpha:
https://openrouter.ai/stealth/ox-alpha#uptime
https://openrouter.ai/z-ai/glm-5.3#uptime
Screenshot of a recent blip: https://files.catbox.moe/haq90y.png
HYhypferl’altro ieri
Can someone explain why people care about that?
Both as in "Why is there a stealth launch like that in the first place?" but also "Why does it matter? Is it very good in something?"
Comments
5 preview comments · loading full threadLog in to h4cker, then connect Hacker News to publish comments.
Error messages matching Z.ai GLM I think are the simplest/most compelling. I had Opus 4.8 poke it and it came back with a couple different errors than the article mentions. Matching the tokenizer is interesting tho
I was using it side by side with GLM 5.3, and they were very, very similar. Also, the new zhipu 1 GW data center plus a flash(smaller?) model, can justify the 100T/day they said they were able to serve. Pretty cool model, especially since it's not a nanny, if you want to unlock your own devices, like rooting an Android, it will happily help instead of flagging you. Available for free via OpenRouter and Nous free tier. Also via OpenCode Go, but you have to pay a $5–$10 subscription. APIs are hammered now, so service is bumpy.
Recent and related: Ox-Alpha Is GLM? - https://news.ycombinator.com/item?id=49422226 - Aug 2026 (65 comments) A mysterious free AI model is impressing developers. Nobody knows who made it - https://news.ycombinator.com/item?id=49406289 - Aug 2026 (4 comments) Ox Alpha - https://news.ycombinator.com/item?id=49381896 - Aug 2026 (202 comments)
Another strong hint is that the uptime graph of GLM-5.3 by Z.ai is very similar to that of Ox Alpha: https://openrouter.ai/stealth/ox-alpha#uptime https://openrouter.ai/z-ai/glm-5.3#uptime Screenshot of a recent blip: https://files.catbox.moe/haq90y.png
Can someone explain why people care about that? Both as in "Why is there a stealth launch like that in the first place?" but also "Why does it matter? Is it very good in something?"