Three years ago, people were saying "LLMs just generate plausible-sounding text, they don't understand the notion of truth so they can't do verifiable work like math proofs."
Right, which was true at the time. So hundreds of billions of dollars have been poured into making LLMs better at these tasks via pretraining, RL, RLHF, post training, etc. again all with something verifiable in the loop. In order to improve the thing in the loop, the loop itself needs to be verifiable.
There have only been a few thousand wars, and they’re all different and all different in the world in which they occurred. The dimensionality is absurd, which is not a problem for LLMs if there’s enough data, but in this case there isn’t.
You understand that those are different kinds of "truth", right?
> they don't understand the notion of truth so they can't do verifiable work like math proofs."
a) No one ever said that.
b) Your comment shows a lack of understanding of the notion of truth.
> they don't understand the notion of truth so they can't do verifiable work like math proofs.
Nobody that understands automated proof checking was claiming that.
They still can't. But very smart humans constructed ways to use the monkeys with typewriters (with a statistical advantage) to find correct answers to problems where they already knew how to verify the answer.