Do you draw that conclusion from the fact that AI surprisingly quickly reaches the end of each ruler we try to measure it with?
Its not really that surprising when models are trained on the exams
Can't call it AI like that without discrediting yourself. You mean LLMs?
Its not really that surprising when models are trained on the exams