>You could ask an LLM what 1+1 is, and the number of times it says "3" is so small that it makes no sense to worry about it...
I think the disturbing fact is that you can take a frontier model with all the intelligence of humanity, and make it say 1 + 1 = 3, by specifically training for it...
A human with that much knowledge will refuse that attempt. There in lies the difference..
What’s your point though, really? “You can train a model to say things that are objectively wrong”? You can do the same with a human.
> A human with that much knowledge will refuse that attempt.
Well, that's not true.
https://www.youtube.com/playlist?list=PLO3a3Ax6Yh6bbtKuxfYBP...