I'm a bit skeptical of arguments of the form, "You should not use LLMs without disclosure because LLMs at bad at writing."
My reasoning is: If LLMs get better at writing—which I think is extremely likely—will you switch positions and say that now using LLMs without disclosure is A-OK?
Surely some people are willing to bite that bullet and say yes. But for most people, my guess is that the answer will remain no. Thus, I tend to think that the "real" reason most of us don't like it when people use LLMs to write without disclosure is that it's misleading: It's a sort of a claim that certain thoughts can be attributed to a human being when in fact they can't.
(The em dashes in this message were rendered using keyboard shortcuts.)
I don't think anyone needs to disclose the tools they used to write something. Nobody ever disclosed Word's grammar suggestions or spelling corrections, or the fact that they used some autocomplete tool to produce code, so we should they now need to start disclosing this tool. Besides, in it's current state, it discloses itself to anyone paying attention.
I do think people own the output they produce regardless of the tools they use, and if they want to put crappy writing out there with their name on it, that's on them. Even if they do disclose that it was written by an LLM, they're still 100% responsible for the content.
I suppose using LLMs to write is like not wiping your ass, and using LLMs that are bad at writing is like not wiping your ass and wearing a T-shirt that says "Wiping is for losers."
I might have smelled something suspicious before, but knowing for sure is worse.
> If LLMs get better at writing—which I think is extremely likely—will you switch positions and say that now using LLMs without disclosure is A-OK?
If LLM writing improves to the point where they can infer the business (or other real world) context and serves the functional purpose of informing others as opposed to being an intellectually lazy piece being produced only for the purpose of being produced, then I’d be fine. It’s likely the author would need to spend some effort on the said piece of writing, regardless of how good the LLMs get good at writing.
People shouldn’t pass off largely unedited AI writing as their own, but I’d much rather live in a world where that writing would be of high quality in both form and content, than in the world we live now. In that sense, lack of disclosure is the lesser evil.
Some people type 100x more words than the final article just digging into a topic and turning it on all sides. The final article being generated tells you nothing of the size of the effort going in.
LLMs write better than most journalists. It also writes better than majority people in the world when it comes to English as majority are really bad in writing.
And of course, LLMs write as you prompt.
Personally, I'm happy to read broken English from non English people or good English from English people. But I'd rather read LLM writing than read typical long form perfect English journalist crap.
I wonder why people think posting undisclosed LLM writing is any better than posting some other person’s writing uncited. Even if the original author is okay with it, it’s still misleading and dishonest. Or maybe all these people would be fine having someone else write their content for them if it were free and easy? A social media is a reputation economy, but reputation doesn’t work if gaining karma requires no effort.
I agree. LLM writing is atrocious, but people also want LLM writing to be atrocious in many cases, because they don’t like LLMs.
The worst of this is with image/video AI models, where the results today, although still very imperfect, some people will still pretend like it’s awful and the worst thing they’ve ever seen. They refuse to admit the technology is at all impressive or making progress because they don’t like the technology.
I don’t like AI generated images and video either, but I can regrettably admit that the technology has gotten remarkably better over time.
Personally, I'd be fine with LLM writing if it wasn't so unpleasant to read
I think, there's a systematic flaw in RLHF: constructs reinforced as effective or sophisticated style will always suffer overuse and significant overexposure. Which will be also very obvious when used out of the intended context, it's copied from. There will be always a "smell", something that causes us to turn our backs to these texts, and possible the supposed authors, as well.
There's also the problem that certain types of phrasing are perceived as effective, because they mark a pivotal point in the progress of the text and are, as such, used sparingly, but are now becoming everyday templates that incorporate whatever is available in the context. There is no way this passes the Turing test of a competent reader.
And there's yet another issue: in social research, there has been the concept of semantic position, indicated by deviation from the mean (or median). If you don't deviate from the mean, your semantic position is zero. There's simply no expression. In this sense, next token prediction really amounts to a desemantificiation of the context. There's really no sense in uttering any of these productions, no plausible motivation, other than for the purpose of raising you hand to be seen.
Agreed. LLMs for writing is like intellectual catfishing. You are presenting yourself one way in written form.. but that's not who you will ever be once someone has to actually have a conversation with you. And that's wrong.
> If LLMs get better at writing—which I think is extremely likely—will you switch positions and say that now using LLMs without disclosure is A-OK?
I would say it is certainly significantly more ok. There are two reasons reading AI-generated text sucks now:
1. It's usually low value and not trustworth - I could have just asked the AI myself.
2. The prose style is horrible to read.
If we eliminate the second reason then it's definitely an improvement. (Although on the other hand the terrible prose can be quite a helpful indication that you're wasting your time reading slop, so maybe we shouldn't complain about it!)
This misunderstands what it means to be better at writing. Root cause here is that writing is how humans communicate ideas. An LLM written thing isn’t doing that, it’s something equivalent to copy/pasting a Wikipedia article. Even if the writing gets better, it will still be necessarily empty when expanding anything that isn’t completely specified.
LLMs seem already to be pretty good at translating, where you already have something fully written and are changing the language. It’s when they get rough ideas and fill in the gaps you get the empty prose they are known for.