It's a clear illustration that these systems are not oracles. No matter how much they've improved since then, we should assume they are still capable of making hilariously big mistakes. Only now the obvious mistakes are fixed and the big mistakes will be a lot more subtle and a lot harder to catch.
> It's a clear illustration that these systems are not oracles.
Of course they are not oracles. Oracles belong to magic, these things are real.
I really can't understand how people seem to expect "intelligence" to be deterministic and infallible and uniform across domains when all examples we have of it (that is, us) are anything but.