Ox-Alpha Is GLM?

(dejan.ai)

47 points | by jitbit 12 hours ago

8 comments

  • petesergeant 4 minutes ago

    I think within 12 months we’re going to see a frontier (inc open models) that’s so good at almost all human-directed tasks that which model you use just won’t matter. Only differences that remain will be in deep research or very long-range tasks.

    • stingraycharles 3 minutes ago

      People were saying this last year, and they’ll be saying the exact same thing next year. The goalpost keeps moving.

    • gvkhna 18 minutes ago

      If it’s not zhipu then why is it returning errors that zhipu does for other models? Who else would return the exact same errors even if they took a lot of core infra like tokenizer from z?

      • jerrythegerbil 19 minutes ago

        As someone who uses NCD nearly every day, I have concerns about how it’s been used here.

        But while we’re “guessing”: Xiaomi MiMO

        • mogili 17 minutes ago

          It's not a good model tbh, got a bunch of things wrong that Opus corrected in my codebase.

          • petesergeant 8 minutes ago

            Yet to find a model that cross-model review doesn’t find a bunch of things wrong with. I’m running simultaneous review with whichever of Grok4.6/GLM5.3/Fable/Sol didn’t write it, and each model tends to find items the others didn’t.

          • volf_ 11 hours ago

            GLM 5.3 and all previous models don't have a vision encoder and can only accept text. Ox-Alpha can accept video and images, so unless Z-ai added a pretty good vision encoder for this model, I don't think so.

            My money is on Moonshot and this being Kimi K3.5. The measured tps and latency is in-line with K3's tps and latency from Moonshot.

            MiniMax M3.5 is also possible (but the MiniiMax provider is a lot more performant than the lab behind ox-alpha, so less likely).

            • minimaxir 1 hour ago

              The other tell from the provider angle is capacity. Whoever is hosting Ox Alpha has a lot of capacity which narrows down a lot of the Chinese companies.

              • nylonstrung 9 hours ago

                It would be stranger to me that Kimi switched to GLM's tokenizer than that GLM added multimodal like Kimi and Deepseek both did recently

                • Bolwin 11 hours ago

                  Glm had made vision models in the past. Look up GLM 5v.

                  The only question now is if it's 5.3v, 5.4/5.5 or a dedicated flash/vision model

                  • BoredomIsFun 17 minutes ago

                    GLM made pretty decent for that time small 9b vision model, GLM-4.1.

                    • volf_ 10 hours ago

                      Yeah. It could be. The Z.ai DC latency is still ~1.2s faster than whomever is serving this model.

                    • Almondsetat 11 hours ago

                      DeepSeek literally just came out with the vision-enabled version of Flash v4 which was purely text based. Why would GLM not be able to do the same thing?

                  • xorgun 48 minutes ago

                    Dont rule out ssi

                    • nullbio 19 minutes ago

                      That would be insanely disappointing.

                    • ChrisArchitect 11 hours ago
                      • behnamoh 43 minutes ago

                        You must have so much time on your hands to go to such great length to dox an anon model on the internet. What new piece of information am I supposed to learn from this passage?

                        • JSR_FDED 6 minutes ago

                          You can learn from the variety of techniques they used to come to this conclusion.

                          • minimaxir 41 minutes ago

                            You cannot "dox" an AI model.

                            Given the traction the model has received, it is extremely newsworthy to know who's developing and hosting it.

                            • behnamoh 40 minutes ago

                              my question is: how does that affect a company's strategy? it's not like management is gonna switch models soon as a new shiny one drops. entire workflows depend on specific models working the way they do; you can't just swap out models.

                              • minimaxir 37 minutes ago

                                If it's a really really good model, then yes, people will switch as long as the price is right. Ox Alpha is looking to be a really really good model to the point that it competes with Fable/Sol, and will likely beat them on price.

                                • volf_ 14 minutes ago

                                  same business model as crack. best way to get people hooked is to make the first hits free.