logoalt Hacker News

red75primetoday at 12:47 AM0 repliesview on HN

> they don't fundamentally "get it"

There's no clear decision criteria for this. Do trick questions demonstrate that most people don't "get it"? And, well, older model saying dumb things doesn't establish a general principle that LLMs don't "get it" in general.

> The sample efficiency is just crazy low

Autoregressive pretraining requires huge amount of data to go from a blank state to a somewhat functional model. Fine-tuning, LORA, reinforcement learning of foundation models and in-context learning are much more sample efficient.

> Chinese room

...creates a wrong intuition that by cranking a Leibniz's mill you are somehow responsible for whether it understands something or not.