| ▲ | Ox-Alpha Is GLM?(dejan.ai) |
| 46 points by jitbit 13 hours ago | 19 comments |
| |
|
| ▲ | gvkhna 5 minutes ago | parent | next [-] |
| If it’s not zhipu then why is it returning errors that zhipu does for other models? Who else would return the exact same errors even if they took a lot of core infra like tokenizer from z? |
|
| ▲ | mogili 5 minutes ago | parent | prev | next [-] |
| It's not a good model tbh, got a bunch of things wrong that Opus corrected in my codebase. |
|
| ▲ | jerrythegerbil 6 minutes ago | parent | prev | next [-] |
| As someone who uses NCD nearly every day, I have concerns about how it’s been used here. But while we’re “guessing”: Xiaomi MiMO |
|
| ▲ | volf_ 11 hours ago | parent | prev | next [-] |
| GLM 5.3 and all previous models don't have a vision encoder and can only accept text. Ox-Alpha can accept video and images, so unless Z-ai added a pretty good vision encoder for this model, I don't think so. My money is on Moonshot and this being Kimi K3.5. The measured tps and latency is in-line with K3's tps and latency from Moonshot. MiniMax M3.5 is also possible (but the MiniiMax provider is a lot more performant than the lab behind ox-alpha, so less likely). |
| |
| ▲ | minimaxir an hour ago | parent | next [-] | | The other tell from the provider angle is capacity. Whoever is hosting Ox Alpha has a lot of capacity which narrows down a lot of the Chinese companies. | |
| ▲ | nylonstrung 10 hours ago | parent | prev | next [-] | | It would be stranger to me that Kimi switched to GLM's tokenizer than that GLM added multimodal like Kimi and Deepseek both did recently | |
| ▲ | Bolwin 11 hours ago | parent | prev | next [-] | | Glm had made vision models in the past. Look up GLM 5v. The only question now is if it's 5.3v, 5.4/5.5 or a dedicated flash/vision model | | |
| ▲ | BoredomIsFun 4 minutes ago | parent | next [-] | | GLM made pretty decent for that time small 9b vision model, GLM-4.1. | |
| ▲ | volf_ 11 hours ago | parent | prev [-] | | Yeah. It could be. The Z.ai DC latency is still ~1.2s faster than whomever is serving this model. |
| |
| ▲ | Almondsetat 11 hours ago | parent | prev [-] | | DeepSeek literally just came out with the vision-enabled version of Flash v4 which was purely text based. Why would GLM not be able to do the same thing? | | |
|
|
| ▲ | xorgun 36 minutes ago | parent | prev | next [-] |
| Dont rule out ssi |
| |
|
| ▲ | ChrisArchitect 11 hours ago | parent | prev | next [-] |
| Related: Ox Alpha https://news.ycombinator.com/item?id=49381896 |
|
| ▲ | behnamoh 30 minutes ago | parent | prev [-] |
| You must have so much time on your hands to go to such great length to dox an anon model on the internet. What new piece of information am I supposed to learn from this passage? |
| |
| ▲ | minimaxir 29 minutes ago | parent [-] | | You cannot "dox" an AI model. Given the traction the model has received, it is extremely newsworthy to know who's developing and hosting it. | | |
| ▲ | behnamoh 27 minutes ago | parent [-] | | my question is: how does that affect a company's strategy? it's not like management is gonna switch models soon as a new shiny one drops. entire workflows depend on specific models working the way they do; you can't just swap out models. | | |
| ▲ | volf_ 2 minutes ago | parent | next [-] | | same business model as crack. best way to get people hooked is to make the first hits free. | |
| ▲ | minimaxir 25 minutes ago | parent | prev [-] | | If it's a really really good model, then yes, people will switch as long as the price is right. Ox Alpha is looking to be a really really good model to the point that it competes with Fable/Sol, and will likely beat them on price. |
|
|
|