Their models, all of them. This has been shown repeatedly when models can return, verbatim, copyrighted material when asked. New York Times is suing over this.
It would be like if I sold you a pizza, but when you opened the box it contained all the source code for the latest GTA. The pizza wasn't copyrighted by anyone, but the what was inside the box was.
Please provide me an example prompt to do so or some algorithm to extract from the weights for an open source model. I will accept any copyrighted work, any model.
A xerox machine can produce verbatim copyrighted works when asked as well. That doesn’t make distributing the xerox machine the same as distributing the copyrighted works.
Strong IP advocates have argued for years that devices that can be used to infringe copyright are themselves infringement of copyright. So far that hasn’t held up to court analysis provided that device can be and is also used for non-infringing purposes. Given that so far the courts have found that training an AI model is sufficiently transformative to qualify as fair use, it doesn’t seem likely that distributing a model counts as distributing copyrighted material.