logoalt Hacker News

rybosworldtoday at 2:43 PM5 repliesview on HN

Unless I'm misunderstanding, calling this RSI seems misleading?

This looks like an optimization of current training methods, and a good one, but not "RSI" in the sense of a system that can perpetually improve itself forever.


Replies

EthanHeilmantoday at 3:30 PM

Does RSI actually mean anything specific anymore? RSI, AGI, at this point seem like buzzwords. Sure AGI has definition that are measurable, say "better than 95% of humans on 95% of intellectual tasks" but if we used that definition we already have AGI and almost no one thinks we have achieved AGI. We use AIs to train AIs which we use to train AIs, why is that not RSI? How much human intervention means that is not RSI?

show 2 replies
jephstoday at 3:36 PM

Yeah, this is absolutely not what anyone reasonable is thinking about when they say recursive self-improvement.

I'd say it's much closer to the concept of continual learning, but I'm only a few pages deep and haven't groqued it fully yet.

addagtoday at 4:14 PM

Agreed, what I understand from RSI would be models creating new models, or at least upgrading their own weights/architecture. It does not seem to be the case here.

xidong_wutoday at 6:52 PM

This paper optimizes a controller/policy which will be used to agent itself in the next round

show 1 reply
IAmGraydontoday at 5:17 PM

According to industry leaders, we currently have AGI and RSI in the last month or so. Of course, we've seen zero evidence of any of this and have to take their word for it.