If we ignore the massive amount of books being destroyed, the outages and increased usage bills for once free websites (some resulting in closure), the increased difficulty to access public information such as Reddit and Twitter, and lastly ignore the amount of conversations being taken away from the public in favor of LLMs, I would agree.
Most of these things are bad. None of them is robbery, though. Some are of questionable legality, but most are not.
People need to face reality. The old Internet was fragile and couldn't have lasted long. It was dying slowly due to closed social media and LLMs are accelerating that death.
I'd rather it die quickly and be replaced by something more robust than decay slowly. I was sick of how it was before LLMs even came among.
Except books are already destroyed in physical form every single day. My local library has a section near the front for free and/or extremely cheap books ($1) they are desperately trying to get rid of. My impression is that if they are untaken/unsold they will end up in a dumpster.
[Edit] Here's an article from 2003 talking about how Random House destroyed up to 25,000 books a day: https://www.theguardian.com/books/2002/mar/19/fiction.stephe...
> the massive amount of books being destroyed
Surely that's not about AI? Or you mean the small fraction of that that's result of "destructive format shift" process[0], which itself is a consequence of copyright regulation that would otherwise prevent anyone from accessing these works?
> the outages and increased usage bills for once free websites (some resulting in closure)
That's assumed to be AI companies for some reason, even if they have no real incentive to do that, while the usual business underbelly of people scraping web for whatever reasons (which now may include some AI upstart wannabies too, to be fair) is forgotten about. Not helping is people confusing AI agents acting as user-agents and doing one-off fetches with "AI scrappers".
> the increased difficulty to access public information such as Reddit and Twitter
It was never public, and they started locking down before AI, when both platforms (as well as all other social media platforms) run out of VC subsidsies and realized they need to start monetizing; first step they did was to wage a war on third-party clients and API users. AI came later.
> ignore the amount of conversations being taken away from the public in favor of LLMs
You mean human agency? Like, humans deciding it's better for them to ask a machine instead of posting questions online? I can understand that, it's usually much better experience (particularly on sites like StackOverflow - they dug their own grave here, and they know it; there have been memes about this way before LLMs were a thing).
You're also not considering the amount of questions answered that would not have been asked otherwise. I for one don't ask many questions on-line, so anything I ask LLMs that they solve for me, is a question that would've remained unanswered for me otherwise. Not everyone is gregarious online, many people are more self-reliant and only answer questions, but solve their own problems without asking for help (it's probably not optimal thing to do, but that's another topic).
--
[0] - Digitizing and retaining physical copy is clear infringement, digitizing but destroying the physical original can be argued to be fair use.
Eh, StackOverflow deserved to go.
Your talking about a bunch of books that are out of print and sat unsold for years in a warehouse somewhere. The rights holders could print another run tomorrow if they wanted, or better yet digitize them but they don’t. Where’s your ire for the rights holders who sit on these books and don’t do anything with the ?