Remix.run Logo
Behaviorally fingerprinting Ox Alpha's provenance(ctgt.ai)
27 points by cgorlla 9 hours ago | 16 comments
nijave 4 hours ago | parent | next [-]

Error messages matching Z.ai GLM I think are the simplest/most compelling. I had Opus 4.8 poke it and it came back with a couple different errors than the article mentions.

Matching the tokenizer is interesting tho

randomblock1 an hour ago | parent [-]

I find the tokenizers most compelling. That's what the model is trained on, it's an immutable fact of the model and its architecture. You know for a fact that the model is at least related to other models that way. And if a tokenizer is unique / specific to one lab, like GLM's is, it's basically as good as it gets.

Comparatively, you can't be 100% sure that Z.ai isn't able to host some other lab's model (although in this case, the hosting errors still support the GLM theory).

Chu4eeno 14 minutes ago | parent [-]

No, that's not how anything works.

You can finetune an LLM to a new tokenizer by nudging just a few layers (even wildly different kinds of tokens), and there's nothing stopping a lab from using someone else's tokenizer.

dang 5 hours ago | parent | prev | next [-]

Recent and related:

Ox-Alpha Is GLM? - https://news.ycombinator.com/item?id=49422226 - Aug 2026 (65 comments)

A mysterious free AI model is impressing developers. Nobody knows who made it - https://news.ycombinator.com/item?id=49406289 - Aug 2026 (4 comments)

Ox Alpha - https://news.ycombinator.com/item?id=49381896 - Aug 2026 (202 comments)

johndough 4 hours ago | parent | prev | next [-]

Another strong hint is that the uptime graph of GLM-5.3 by Z.ai is very similar to that of Ox Alpha:

https://openrouter.ai/stealth/ox-alpha#uptime

https://openrouter.ai/z-ai/glm-5.3#uptime

Screenshot of a recent blip: https://files.catbox.moe/haq90y.png

hypfer 3 hours ago | parent | prev | next [-]

Can someone explain why people care about that?

Both as in "Why is there a stealth launch like that in the first place?" but also "Why does it matter? Is it very good in something?"

Philpax 3 hours ago | parent | next [-]

Stealth launch: builds hype, allows them to collect user preference data and see where the model fails. Why people care: it's free, decent, and people love a good mystery.

hypfer 3 hours ago | parent [-]

Aaah, free inference. Yeah that checks out and fits very well with weird Internet hype.

Okay, fair enough. Thanks!

handfuloflight 16 minutes ago | parent [-]

It's a good model ser.

chermi 2 hours ago | parent | prev | next [-]

I think there was some loud speculation is was Google. It's cool people can pretty definitively show it's not. But also maybe Google is being crazy sneaky and faking it! (Joking)

tcdent 3 hours ago | parent | prev | next [-]

I think the stranger thing is that people spend tens of hours doing analysis like this to hit an inevitably-expiring hype cycle that will give us a definitive answer shortly.

chermi 2 hours ago | parent | next [-]

Don't you think the forensics is fun?! This is exactly the sort of thing i would've expected hn to be broadly interesting to hn. It's a complicated technology and people are poking it and learning things

cgorlla 3 hours ago | parent | prev [-]

It's fun! Also it's an interesting commentary on where the AI industry is as a whole.

cgorlla 2 hours ago | parent | prev [-]

It's decent. How good depends on compute needs.

UncleOxidant 3 hours ago | parent | prev [-]

Speculation that this is GLM 5.3 Flash.

cgorlla 2 hours ago | parent [-]

Very likely, NYT confirmed it's releasing Friday.