logoalt Hacker News

notfromhere • 08/18/2026 • 0 replies • view on HN

LLMs seem to be trained to work very well against a goal, especially one it can verify against. I guess because it can easily know if it passed or failed, va other tasks where good/bad output is subjective