Great summary. The fact that the auto encoding task is not grounded in thoughts, and their initial t...

mxwsn • yesterday at 11:39 PM • 0 replies • view on HN

Great summary. The fact that the auto encoding task is not grounded in thoughts, and their initial training on guessed internal thoughts, raise serious concerns on faithfulness. Feels like they might get better results by just training a supervised model on activations and "internal thoughts" measured by some different behavioral way.

alt Hacker News