logoalt Hacker News

_heimdallyesterday at 10:50 PM1 replyview on HN

> We are placing stricter requirements on alignment

This is comical. Its impossible to align a black box and that's precisely what LLMs are. It also seems impossible to align recursive text prediction algorithms, which LLMs are.

How exactly do they gate on alignment today, and how can they tighten it? Is it purely gates based on input/output pairs to check whether they're happy enough with responses regardless of how and why the response was actually chosen?


Replies

willmarchtoday at 12:16 AM

Aren't humans black boxes? Aren't humans prediction algorithms?

How do we align humans?

show 2 replies