Finally confirmation that Haiku was not forgotten and will be coming soon, althouhg I find it quite interesting they skipped 5 and directly skip to 5.5 with all models, including Sonnet which is not super old. I suspect they found something breaking that allows to release this. Recently they struggled with keeping up a 50 % weekly limit increase and now they're putting out 30-40% faster and cheaper models even faster, with much more better benchmarks, a limt reset command and five hour limit increase. It seems more like the opposite and as if they never struggled, thus, I very much believe they found something very effective and new.
IT's Artificial Analysis Index is the same as Kimi K3, which is about 4.6x bigger, and GLM 5.3, which is about 1.25x bigger. Pricing is $1/$2.70 i/o. Openweights on October 15.
Most of these models, both open and closed, are now so over tuned to agentic and coding tasks that they no longer work well for general purpose. Kimi K3 is an exception to that – maybe you need that larger size todo well on a broader range of tasks.
It's quite unfortunate, a lot of people or organisations I am interested in - AI for example - only post on X, and perhaps it's only a matter of time until services like bird.makeup also get taken down. Also, accessibility wise it's sad too, X's web interface is not the best out there and without login you cannot do really much.
If AI really is democratizing technology, it won’t be long before we get back to everyone having a personal website.
I remember blowing everyone’s mind when i wrote a php script to scrape my bands myspace to keep our website up to date with show details. It was like a magic trick.
That sort of power is in the hands of the masses now, they just don’t know it yet.
> The company on Wednesday confirmed speculation that the Ox Alpha model is a new iteration of its GLM series and said it will release the weights for it tonight, in response to queries by Bloomberg News.
Seems legit.
It's really hard to know how good it is. So much hype around it.
Nobody (with the probable exception of Anthropic given their work on character training) really trains models on their identity and Claude is the only AI persona that's well-defined so if you put yourself into the AI's shoes it's a pretty reasonable guess that it might be Claude. I've had basically every open model claim it's Claude when the topic comes up.
Anthropic changed their character training policy many times, and the name is probably separate from it anyway (besides the bits from the constitution etc). Name training usually comes last in all models, if at all, and it's pretty shallow. Certain Claude models say they are Qwen or Deepseek when asked in Chinese, for example.
Yeah, Claude is actually surprisingly unsure of his identity considering that their most recent publication on their constitutional AI training literally had graphs demonstrating how certain properties differed based on whether they were phrased as questions about "Claude" versus "You", but it's actually that which makes me fairly confident that they're probably doing _something_ to try and close that gap.
More generally, their current approach to constitutional AI pretty much only makes sense if they believe that they can first teach the model what the Claude character is like and also teach the model that the persona responding is Claude, so I figure that has to be part of the pipeline even if they're not very good at it.
reply