logoalt Hacker News

xdavidliutoday at 1:45 AM1 replyview on HN

that is if it even a human commenter at all


Replies

linkjuice4alltoday at 2:38 AM

State-sponsored psyop meta comments aside, the models obviously continue to get better, but there is still a lot of 'guard railing' required to keep even the latest models completely on-task. The chess example is interesting because it's clearly a well-studied and established domain so the rules, strategies, and whatever else is in the training data should make yield excellent results; but clearly there is some behavior in these systems that's difficult to engineer out.