I can't help but feel that every 1% improvement over a benchmark from new releases of flagship LLMs isn't really that impressive anymore. To some extent it feels like LLM improvements is slowing down significantly despite increasing resources being spent on them.
The people promising AGI just 6 months ago are now more and more quiet about the possibility and rightfully so. LLMs alone are most likely not the path to AGI.
My personal experience diaagrees.
The Opus moment in November felt very different but despite that, a handful things did not woork well with Opus in November and I tried them last week and they now just work.
In parallel GPT-6 and Fable 5.1 show significant skill improvements in 3D modeling and Astra shows significant progress in computer use.
Do we live in a parallel universe? I mean it, this week was crazy from a progress point of view.
No, the people promising AGI 6 months ago are now saying it's here. (and I agree)