Ox-Alpha Is GLM?
jitbit
57 points
32 comments
August 24, 2026
Related Discussions
Found 5 related stories in 53.1ms across 4,281 title embeddings via pgvector HNSW
- Ox Alpha mtokmak06 · 76 pts · August 20, 2026 · 70% similar
- GLM-5.3 Artificial Analysis Benchmarks apitman · 114 pts · August 18, 2026 · 54% similar
- Linear algebra done right the-mitr · 261 pts · August 17, 2026 · 52% similar
- GLM 5.2 is nearly as accurate as a human book keeper adamkurkiewicz · 196 pts · July 09, 2026 · 52% similar
- Oxide Joins Anthropic's Project Glasswing lwhsiao · 12 pts · July 28, 2026 · 51% similar
Discussion Highlights (9 comments)
volf_
GLM 5.3 and all previous models don't have a vision encoder and can only accept text. Ox-Alpha can accept video and images, so unless Z-ai added a pretty good vision encoder for this model, I don't think so. My money is on Moonshot and this being Kimi K3.5. The measured tps and latency is in-line with K3's tps and latency from Moonshot. MiniMax M3.5 is also possible (but the MiniiMax provider is a lot more performant than the lab behind ox-alpha, so less likely).
ChrisArchitect
Related: Ox Alpha https://news.ycombinator.com/item?id=49381896
xorgun
Dont rule out ssi
behnamoh
You must have so much time on your hands to go to such great length to dox an anon model on the internet. What new piece of information am I supposed to learn from this passage?
jerrythegerbil
As someone who uses NCD nearly every day, I have concerns about how it’s been used here. But while we’re “guessing”: Xiaomi MiMO
gvkhna
If it’s not zhipu then why is it returning errors that zhipu does for other models? Who else would return the exact same errors even if they took a lot of core infra like tokenizer from z?
mogili
It's not a good model tbh, got a bunch of things wrong that Opus corrected in my codebase.
petesergeant
I think within 12 months we’re going to see a frontier (inc open models) that’s so good at almost all human-directed tasks that which model you use just won’t matter. Only differences that remain will be in deep research or very long-range tasks.
tadkar
I wonder if the NCD metric says something about distillation too. Would you expect that a model that has been distilled/seen traces from other models would have a smaller NCD? It would be really interesting to see if this holds up and provides evidence of distillation or certainly evidence of model outputs being used in the training mix.