> No human being is going to genuinely suggest to glue cheese onto a pizza to stop it sliding off.
That mistake was made more than two years ago by whatever experimental version of Google's AI overview (which has strict performance requirement- i.e. needs to answer within a split-second) was up at the time. In other words, it was a very small and primitive system specialised in spitting answers as quickly as possible without a second thought. I hope you realise that basing your assessment of what LLMs can or can't do on that example is not much better than suggesting to put glue on pizza. A mistake that, if we were to adopt your reasoning, would in turn set a hard limit to the analytic skills of all humanity.
Nice try with the gaslighting. Google's AI "summaries" get things wrong all the time even now.
It's a clear illustration that these systems are not oracles. No matter how much they've improved since then, we should assume they are still capable of making hilariously big mistakes. Only now the obvious mistakes are fixed and the big mistakes will be a lot more subtle and a lot harder to catch.