OpenAI’s amazing — but vastly oversold — new model Astra
Posted by champagnepapi 20 hours ago
Comments
Comment by keeganpoppen 17 hours ago
Comment by malshe 15 hours ago
Comment by axus 15 hours ago
It wouldn't hurt their argument to observe how well LLMs do at reading and writing, but it's probably painful for them to admit.
Comment by kelseyfrog 11 hours ago
Comment by igor47 19 hours ago
Comment by SpicyLemonZest 18 hours ago
Comment by perching_aix 19 hours ago
You really don't need to break open the fallacy dictionary to see why those tweets are phony, or to telegraph Astra as just an incremental [0] improvement that's even better tuned for math than what came before it. It's the obvious direction of development.
[0] There's a mathematician guy I follow on YouTube who keeps taking LLMs for a spin, and the primary failure mode seems to be persisting. It's not dissimilar to any other field; the models are chatterboxes, and keep going off about stuff that doesn't matter, while quickly jumping over things that do. They're also comparatively slow. According to another mathematician's review of the 250 page paper OAI put out of those 10 breakthroughs, the former persists with Astra.
I wonder if Astra can run on those Cerebras wafers. An order of magnitude faster inference would at least make the iteration process quicker. But then they were announced for Sol too, and they're nowhere to be found. The 2.5x fast mode is nice, but it's a far cry from the 750 tok/sec suggested with Cerebras.
Comment by dude250711 19 hours ago
Comment by eec33 17 hours ago
Comment by semiquaver 18 hours ago