logoalt Hacker News

GLM-5.3: Frontier coding with emergent cyber capabilities

939 pointsby pellatoday at 5:19 AM475 commentsview on HN

Comments

Jacopos311today at 10:18 AM

This looks very interesting indeed!

aizktoday at 6:20 AM

The model releases just don't stop!

yogthostoday at 1:37 PM

I'm so glad I managed to get their subscription when it was on sale for 250 bucks a year back when it was 5.1. Back then it was just ok, but after 5.2, it's become my main workhorse. And 5.3 is looking fantastic.

cubefoxtoday at 7:10 AM

> Open Source: We will release the weights in two weeks after launch, once safety evaluation and hardening are complete.

What safety evaluation? What safety hardening? They already evaluated it and found it to be highly capable at exploiting security vulnerabilities. So we know it is not "safe", and they don't seem to plan to do anything against it. What could be more dangerous than hacking? Biological weapons research? I don't think Chinese labs are doing anything against this either.

show 2 replies
tw1984today at 5:39 AM

just imagine the world without these open weight models - we'd probably have to reverse mortgage our homes to pay for tokens to those trillion $ companies to have access to their models.

ofjcihentoday at 10:56 AM

The capabilities of open models approaching or meeting that of SOTAs is good in every way except for our short-sighted economic reliance on their success (in the US at least).

smurf9852today at 11:33 AM

" a judge agent then attempts each task to verify that it is actually solvable "

I understand you need to verify the goal is achievable. But if the judge agent has the same goal as the training agent (solve), and both are of the same model, then aren't the judge and the training agent doing the exact same thing? What is the point then? Can someone explain this to me.

show 1 reply
petesergeanttoday at 8:50 AM

Their own hardness (ZCode) seems to be a GUI, which doesn't work for me. They say they support other harnesses. However, it seems like I can inject the plan into other harnesses, like Claude Code[0]. Does anyone who's been using GLM models for a while have a strong feeling for if it does better in some harnesses than others, or should I just use my favourite harness?

0: https://docs.z.ai/devpack/tool/others

show 3 replies
bsenftnertoday at 11:05 AM

So, "cyber capabilities", whoa there horsey, what the fuck is that? Are we making up words or are you trying to court the black hat crowd?

show 1 reply
petesergeanttoday at 8:58 AM

[total rewrite: their subscription code is buggy. It takes a while for paid subscriptions to show up, and for upgrades to take effect. Original comment was whining about this]

show 1 reply
MrBuddyCasinotoday at 6:02 AM

An I the only one who was disappointed with GLM 5.2 after all the hype? It was thinking forever and sometime just stopped mid task.

steffi_olivertoday at 4:52 PM

[dead]

Satoshi_Brotoday at 1:41 PM

[flagged]

jocelynertoday at 7:26 AM

[dead]

zaj00ltoday at 2:48 PM

[dead]

libertas_quae_stoday at 9:37 AM

[dead]

bigwheeltoday at 11:28 AM

[dead]

felixlu2026today at 10:42 AM

[dead]

maxkim869today at 11:41 AM

[dead]

Culonavirustoday at 7:49 AM

[flagged]

show 2 replies