logoalt Hacker News

howunfortunatetoday at 3:39 AM1 replyview on HN

Yes, catastrophic forgetting is absolutely one of the problems that needs to be solved to enable something like this.

My broader point is just that there's nothing inherent to the structure of LLMs that stops them from updating their weights and continuously learning from environmental feedback in the way humans do, and there's already solid templates for how they could push even further in that direction.

But as an assessment of the current state, I agree with you, LLMs lag humans severely in ability to self-update.


Replies

imtringuedtoday at 7:22 AM

>My broader point is just that there's nothing inherent to the structure of LLMs that stops them from updating their weights and continuously learning from environmental feedback in the way humans do, and there's already solid templates for how they could push even further in that direction.

"LLM" is a branded model as a product. Of course it could be anything, as long as it fulfills the product category.

But we live in reality, we can only look at what models are out there and we see that they don't do any of those things and yet we're supposed to act as if these models already do.