In a study published in the journal Judgment and Decision Making, 1,682 adults were asked to read one of six short stories, three of which were written by humans and three by ChatGPT. Each AI story had a similar theme to one of the human-authored works.
The team told participants whether their story was written by a human or AI, but this information was not always correct. The researchers then asked participants to rate how absorbing and engaging they found the story, and its quality.
Participants who read an AI-generated story rated it as more absorbing and of higher quality than those who read a story written by a human. However, participants gave higher ratings to stories they had been told were written by people.
Dr Deena Skolnick Weisberg, a senior author of the research from Villanova University in Pennsylvania, said: “AI systems can already generate short stories that are seen as being at least as good as – if not better than – human-written stories. We should update our views of AI’s abilities accordingly.”
However, Weisberg, who is herself a creative writer, said that did not mean writing should be left to AI, noting that novels produced by tech were probably going to be different from those written by humans.
“We may need to make room for AI-generated novels, and for AI/human co-written novels, but that doesn’t mean that there’s no longer space for us to appreciate the process of human creativity,” she said.
Look, man, I have been in writing groups and classes with people of all ages and skill levels. I have read some truly horrendous stories. Is the infinite word predictive machine going to write a better story than the worst, least skillful writer? Easily. Is it going to produce a better story than some of the more skillful writers? Absolutely not.
Worth noting also, I’m not saying someone who is noted for selling a lot of books, or even someone in an MFA program for writing. I am saying, LLMs are unlikely to produce a better work than me, and I am just some guy who is pretty skilled and practiced at writing. Maybe the one or two most skilled in a group of hobbyists who meet and share feedback and try to get better.
AI is always going to shoot for the absolute middle in terms of quality, and can definitionally produce nothing new or innovative.
- 1 day
This is like comparing landscape paintings to photographs of the same landscape from the same perspective and asking random people to judge which one they think is better. This is an insane comparison.
- eleijeep@piefed.socialEnglish1 day
Here are the stories they used, if you want to decide for yourself. Although it won’t be a blind test since in these documents they tell you which is which:
- 1 day
Having read the compared stories now, I have to say I’m deeply offended at the results. The AI stories are quicker to read because they are chains of cliches which state and restate the theme of the story (as stated in the prompt) outright instead of demonstrating it through the charged and personal perspectives of the characters. The LLM generated text is meaningless word sausage formatted to be easier to read for people who don’t read much. I have never read something more cynical, and I have read plenty of trash by authors who hate their audience and consider them all to be idiots.
- 1 day
You nailed it. For me, the human-written stories were interesting because it’s not immediately apparent where they’re going. The prose evokes images and emotion, and I want to put the pieces together to identify the theme.
The AI stories blurt out the theme almost immediately, and then my attention wanes because there’s nothing left in the text except for a base retelling of events.
- 19 hours
Most people are dumb and are functionally illiterate. Of course the incredibly dumbed down text will appeal to them more.
- XLE@piefed.socialEnglish1 day
I like how story pair #1’s AI just took the prompt and basically kept repeating “yeah, them’s fish” over and over without any attempt at significance
- 2 days
Reading the journal paper this article referenced, I was a little disappointed that there weren’t more variables accounted for regarding the participants and the texts. With the participants, factors I see commonly accounted for are levels of education, socioeconomic indicators (say income), or how many books the participants have read in the past month or something similar.
Regarding the text, some form of analysis to then account for or attempt to equalise the AI-generated texts so that they’re similar to the author-written ones would have been nice - like average sentence lengths or average syllables per word.
However, this is still a data point (and a scary one) that it can be very hard to distinguish AI-generated content from human-generated ones…
- 15 hours
I was a little disappointed that there weren’t more variables accounted for regarding the participants
I think it’s even worse than that. quoting from the study’s abstract:
In Study 1, participants (1,682 adults recruited from Prolific)
…
In Studies 2 and 3, participants (905 adults recruited from Prolific)
hadn’t heard of Profilic before, they describe themselves as:
so the common theme is that 100% of the study participants have opted-in to this “help evaluate AI as a paid side-hustle” thing. that is going to introduce a huge selection bias that is definitely not representative of the population as a whole.
maegul (he/they)@lemmy.mlEnglish
2 daysTo be honest, this moment was always coming or is yet to come for anti-AI folks.
The real issue isn’t whether it’s good or bad. The reason it’s such a hot thing is because generally it’s good enough to get people interested. And that’s very unlikely to go away.
Too interested? Sure.
But machines that are better than people at something people value about themselves. That’s where we’re heading if we’re not there already.
Which unlocks a massive question: What kind of world do you want to live in and how much are you willing to do to maintain it? Just about everything else is no longer a given. Not our value or importance, not our position at the top of the food chain, not our democracy or wealth, or the continuous growth of capitalism, or even our continued existence.
We’re basically at the beginning of an AI apocalypse movie. You can’t presume that you know the ending. You have to figure out how you want it to end. And not just for you, but your children and grandchildren children. Cuz it’s all at stake now.
- XLE@piefed.socialEnglish1 day
I think you’re taking the doomer narrative a bit too seriously, especially with that final paragraph. The prominent pushers of said narrative usually have something they want to sell, or a real-world issue to distract you from. Some (Yudkowsky) from their cults, some (Altman, Amodei) to enable monopoly or drive profits. Short stories do not an AI apocalypse make.
Sources need to be vetted and held to account when they are wrong or make unfalsifiable statements.
- 24 hours
I don’t think it’s a doomer comment at all, I think the last paragraph is trying to tell you without saying the word “datacenter” to go look up the addresses of data centers.
- 1 day
I read The machine stops short story and I really hope we don’t head that direction, but it seems we are.
maegul (he/they)@lemmy.mlEnglish
1 dayWhat strikes me about the present moment is that it feels like particularly bad timing.
I’m not so sure that technological developments follow a well defined path and like to think about what alternative timelines would look like, especially with humanity being better prepared for them.
For instance, I’ve figured for a long time that the internet was a few decades earlier than it should have been for our social and political development. Probably TV too.
AI just feels 100s of years too early. Like we’re only just getting to grips with post-capitalism and climate change as ideas let alone problems we can solve.
- 24 hours
I think this is actually advantageous to us in a way, because clearly AI is something of a system shock to many people right now, whereas in a hundred years it may have not been nearly as apparent as the danger it is (and I don’t mean LLMs, I mean AI in general, especially for surveillance and weapons).
maegul (he/they)@lemmy.mlEnglish
21 hoursYea I hear. We’ll see I guess.
Personally, I’m seeing much more falling in line instincts than anything else and fear that any “shock” and reaction is confined to bubbles without any solid culture and infrastructure to break out into any sort of effective movement.
- 21 hours
Luckily, revolutions have never required the majority of a populace to actively participate in to succeed, just a majority going “meh, sure”.
But yes, I am personally also very worried that we’re reaching a point with technology that state surveillance and violence will both be too easily applied and too pervasive for the revolution to overcome.
Handles@leminal.spaceEnglish
1 dayQuite agree, but there is the other POV that “AI” has emerged just at the right time to distract us from doing anything about climate change and rampant late stage capitalism… That is of course the perspective of capitalist entities that really need to wrench another few decades of profit out of us before we die in sundry environmental disasters.
See also, “why it would be really convenient if we could upload our minds to a metaverse while the world burns and billionaires escape to Mars”.
P03 Locke@lemmy.dbzer0.comEnglish
1 dayThere’s only two choices:
- Learn how it works, adapt it to your lifestyle, use it as a weapon against the rich fucks that try to use it against you.
- Continue to act like luddites, bury your head in the sand, and pretend that it will go away. Which it won’t.
That’s it. Those are the choices. Most are sucking in that copeium for the second option.
Why are those two the only options that come to mind? You can’t imagine working against the rich fucks without using autocomplete?
Nice try fbi.
But the implication that a chatbot is needed to knock rich fucks down a peg is silly to me. Orcas didn’t need Claude to attack yachts, the ocean didn’t need Gemini to crush that dipshit and his sub, saint Luigi didn’t need grok when he was being a referee for my child’s soccer game
If a token predictor has removed your creative spark, then you’ve already lost. That’s unfortunate, but assuming everyone else has lost would make you wrong.
- TehPers@beehaw.orgEnglish17 hours
What’s yours? Pay Anthropic $2k/mo to generate fake C&D letters and send them out to random businesses?
Unless you have a plan that doesn’t rely on the service-based models, then I don’t see where you’re going with this. Sure, you can use self-hosted models, assuming you’re fine paying for the GPUs (Jensen Hwang? Lisa Su? Lip-Bu Tan if you’re feeling special?) and power to run them of course. But you’ll be “behind” the big cloud-based models, endlessly chasing after them.
Or, hear me out, don’t use them and do it all yourself and you won’t have to pay these companies and make their execs richer.
maegul (he/they)@lemmy.mlEnglish
22 hoursI don’t entirely disagree. I’d add that whatever other options there are, they involve skills, values and perspectives that are unpracticed and atrophied in modern western culture (and likely globally) to the point that that is its own problem.
- Onomatopoeia@lemmy.cafeEnglish1 day
AI text can be very hard or nearly impossible to detect from human generated, at times, depending in the model used.
It’s rather surprising.
- 17 hours
People are also way worse at distinguishing them than they think they are. And it’s the same story with images.
- MagicShel@lemmy.zipEnglish2 days
I would love to know the prompts and parameters of their short stories, but no. I generate stories all the time as a sort of self-roleplay since I don’t have a tabletop troupe to play with. The stories are terrible and get worse with length. I only tolerate it because roleplaying already has a lot of bad writing and retconning, and I’m not looking for depth or a coherent plot.
P03 Locke@lemmy.dbzer0.comEnglish
1 dayI think a lot of it depends on the model and the amount of context it has. LLMs usually get worse with time because they are running out of context, or they don’t have a good summarization of the story at large. For story-writing, it’s best to have a full outline planned and set up as its own Markdown doc, along with critiques on things you think don’t make sense or just don’t like as a plot device. Then, start a new session with each section, keeping it in a Markdown doc as it goes. Some sort of character doc is probably a good idea, too. Give it as much reference material as possible.
It’s just like planning features for code. You don’t just say “gimme a story like blah”. You have to set it up like writers write movies.
- BlameThePeacock@lemmy.caEnglish2 days
People down voting a study. Hilarious.
Feels over reals. We’re fucking doomed.
A study isn’t valid simply because it’s a study, and headlines telling you what to think of their results without at least a vague sense of the numbers are propagandistic by nature.
Rated better quality? By who and how much? Pretty basic details to include.
From what the guardian reports on the details, this study looks like garbage. From the posted stories, the results are definitely garbage, and whether that’s the fault of the researcher or humanity will need a closer inspection.
I’ll tell you right now, coming to such a broad conclusion because some American sophomore students liked three AI stories better than three selected human written stories sounds pretty fucking stupid.
- BlameThePeacock@lemmy.caEnglish1 day
Are you kidding me? This was a University Prof, doing a proper study, published in the Cambridge University Press, and involving 1700 participants.
Did you not even bother to look up the source before jumping down my throat?
This is even further evidence of why we’re fucking doomed. Your feelings about this headline override any reasonability you have to the point where you won’t even look at the source.







