Correct. If they were doing this with Gutenberg/Smithsonian/some library so there's both a public archive and training the language model, it wouldn't have the ick factor.