logoalt Hacker News

aerhardttoday at 3:54 PM3 repliesview on HN

I really enjoy the balance of speed and accuracy of Astra. I can definitely see it become my driving model for most tasks, technical and non-technical.

However, I don't see it as such a massive leap compared to Fable or Sol. As ever, there's a mismatch between the benchmarks and my daily experience of the models.

What do you all think about Astra now that it's been out for a few weeks?


Replies

rfgplktoday at 4:12 PM

> What do you all think about Astra now that it's been out for a few weeks?

Best model put out so far by any of the frontier labs. Way better than Anthropics models, especially in actual text generation. Claudes fodder heavy text is ridiculous.

> However, I don't see it as such a massive leap compared to Fable or Sol.

It's hard to quantify these things without burning tons of tokens. But Fable has been a huge disappointment for me with the sole exception of graphics (UI/GPU shaders). It burns an obscene amount of tokens and barely produces output better than Opus 5.

Edit because I forgot to mention that Fable is the only modern model that seems to splat out random Chinese or Arabic glyphs. And 5.1 does it more than 5

kbrannigantoday at 4:14 PM

Such a massive leap at averaging possible use cases from previous data collected.

My guess is : collect all the prompt and their satisfaction score. group them by similarity . For each group pretrain the next model on that . Get these results ready.

Next model generation feed them back those answers.

mythrwytoday at 3:57 PM

Extremely capable and one shots large tasks from somewhat vague descriptions. Not AGI, not even close, that is complete nonsense. Just my opinion.

show 1 reply