Ox-Alpha Is GLM?

(dejan.ai)

61 points | by jitbit 14 hours ago ago

26 comments

  • gvkhna an hour ago ago

    If it’s not zhipu then why is it returning errors that zhipu does for other models? Who else would return the exact same errors even if they took a lot of core infra like tokenizer from z?

    • e9 7 minutes ago ago

      Someone could've trained model on top of GLM. Same way Cognition trained their SWE model on top of Kimi and Cursor did same with their Composer model.

      • gvkhna 3 minutes ago ago

        While possible the amount of variation in serving infrastructure is unlikely to land with actually giving the exact same errors zhipu does.

        It feels like glm flash, and there was a report zhipu had secured a huge new cluster suggesting they have the capacity. My guess anyway.

  • tadkar 43 minutes ago ago

    I wonder if the NCD metric says something about distillation too. Would you expect that a model that has been distilled/seen traces from other models would have a smaller NCD? It would be really interesting to see if this holds up and provides evidence of distillation or certainly evidence of model outputs being used in the training mix.

  • jerrythegerbil an hour ago ago

    As someone who uses NCD nearly every day, I have concerns about how it’s been used here.

    But while we’re “guessing”: Xiaomi MiMO

    • walrus01 32 minutes ago ago

      Have also seen people guess it's a next version of Longcat, but I also think that's unlikely

  • mogili an hour ago ago

    It's not a good model tbh, got a bunch of things wrong that Opus corrected in my codebase.

    • petesergeant an hour ago ago

      Yet to find a model that cross-model review doesn’t find a bunch of things wrong with. I’m running simultaneous review with whichever of Grok4.6/GLM5.3/Fable/Sol didn’t write it, and each model tends to find items the others didn’t.

      • epolanski a minute ago ago

        If your changes are non trivial even the same model will loop over and over with the feedback.

      • mogili 20 minutes ago ago

        Wasn't just a review, it failed the task I gave and Opus completed the task

  • volf_ 12 hours ago ago

    GLM 5.3 and all previous models don't have a vision encoder and can only accept text. Ox-Alpha can accept video and images, so unless Z-ai added a pretty good vision encoder for this model, I don't think so.

    My money is on Moonshot and this being Kimi K3.5. The measured tps and latency is in-line with K3's tps and latency from Moonshot.

    MiniMax M3.5 is also possible (but the MiniiMax provider is a lot more performant than the lab behind ox-alpha, so less likely).

    • nylonstrung 11 hours ago ago

      It would be stranger to me that Kimi switched to GLM's tokenizer than that GLM added multimodal like Kimi and Deepseek both did recently

    • minimaxir 2 hours ago ago

      The other tell from the provider angle is capacity. Whoever is hosting Ox Alpha has a lot of capacity which narrows down a lot of the Chinese companies.

    • Bolwin 12 hours ago ago

      Glm had made vision models in the past. Look up GLM 5v.

      The only question now is if it's 5.3v, 5.4/5.5 or a dedicated flash/vision model

      • BoredomIsFun an hour ago ago

        GLM made pretty decent for that time small 9b vision model, GLM-4.1.

      • volf_ 12 hours ago ago

        Yeah. It could be. The Z.ai DC latency is still ~1.2s faster than whomever is serving this model.

    • Almondsetat 12 hours ago ago

      DeepSeek literally just came out with the vision-enabled version of Flash v4 which was purely text based. Why would GLM not be able to do the same thing?

      • volf_ 12 hours ago ago

        It's possible

  • petesergeant an hour ago ago

    I think within 12 months we’re going to see a frontier (inc open models) that’s so good at almost all human-directed tasks that which model you use just won’t matter. Only differences that remain will be in deep research or very long-range tasks.

    • stingraycharles an hour ago ago

      People were saying this last year, and they’ll be saying the exact same thing next year. The goalpost keeps moving.

      • Tepix 28 minutes ago ago

        It‘s already happening, people are using cheaper models because they are good enough

      • petesergeant 13 minutes ago ago

        Someone else having been too early on a prediction has little bearing on my prediction.

  • xorgun 2 hours ago ago

    Dont rule out ssi

    • nullbio an hour ago ago

      That would be insanely disappointing.

  • ChrisArchitect 12 hours ago ago