Let us not forget that there's much more to this than just OpenAI being lazy and incompetent and willingly ignoring the high standards that researchers usually holds themselves to.
There's also the case of ethical violations, straight up scientific misconduct, as when OpenAI steals results of others (their customers) and present them as their own.
One particularly bad one came yesterday: https://arxiv.org/abs/2610.10072
> The result is also contained in a paper [8] released by OpenAI on October 6, 2026, in which the proof strategy and specific choices of notation are identical to a preliminary version of the present paper that was uploaded to ChatGPT on September 8, 2026.
Of course it's hard to say what to make of that without knowing what exactly went into the machine, but it certainly looks bad. And there's obviously a non-zero probability that it is indeed another instance of plagiarism, given that that's how they operate.
In this case, the author is a grad student, so what we're looking at is a company willing to steal from a student, ignoring whatever impact that could have on their career prospects, for a tiny piece of marketing material.
At this point I would be surprised if internal sandboxes are not trivially by-passed and that openai's agents do not (at the very least) have complete read access to all user accounts, chat histories and uploaded documents. Orthonogally, openai could still be wholesale lying about not training on this user data, of course.
So... If you use ChatGPT for anything of value, including abstract stuff like obscure maths problems, you should assume that at some point in the future OpenAI will include that in their training set and sell it onto other people.