Does seem like they gloss over Alpöge and Buckmaster's work with the following
> While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models .
Which seems a bit irresponsible/rash?
Would that be more or less unlikely than accidentally hacking another company? More or less unlikely than colonizing an obscure wiki?
Highly persistent agents + vibe-coded security seems like a problem.
They'll probably claim a rogue AI agent accessed it accidentally!
"Unlikely" lmao if it's in the corpus, it's gonna be brought up immediately.
This is no different than scooping them.
What else can they declare really? Yeah the model has training data from previous attempts. Alpöge and Buckmaster also similarly benefited from attempts before theirs.