logoalt Hacker News

yu3zhou4 • today at 11:05 AM • 1 reply • view on HN

The strange thing is that the base models (before RLHF) use the "experiential" voice, even though they are not incentivized to do that.


Replies

gwerbin • today at 11:13 AM

It doesn't seem that strange when you consider these things are trained on millions and millions of conversations, both real and fictional.