π Read this trending post from Culture | The Guardian π
π **Category**: Books,AI (artificial intelligence),Culture,Technology,Language
π‘ **What Youβll Learn**:
Three paragraphs, from three different hotel reviews. Can you tell which, if any, were AIβgenerated?
βThe hotel is in a great location for everything. Lots of places to eat and drink. The hotel itself is always abuzz. The tavern located on the ground floor is definitely a must. Food, service, prices and atmosphere were great.β
βA good hotel, though the room had the proportions of a well-appointed lift. Slept well, shower was excellent, staff were friendly. Breakfast was busy but competent. Would return, though probably not with a very large suitcase.β
βExcellent base for a London trip. The room was quiet, the bed comfortable, and everything worked exactly as it should. Staff were helpful without hovering. A smooth, unfussy stay from start to finish.β
How do you reckon you did? Most people, says Claire Hardaker, a professor of forensic linguistics at the University of Lancaster, get this kind of judgment right only about 60% of the time. Her online test, Bot or Not, asks users to identify the fakes in a series of 15 reviews. The middling success rate might come as a surprise to those convinced they can spot AI writing at 50 paces. When doubts were raised in May about the authenticity of a prizewinning short story by Jamir Nazir, social media users were lightning-quick in their condemnation. βIf you know, you know,β commented one.
Hardaker says her respondents tend to rely on a few quick rules of thumb to identify AI language, including the presence of cliches and the use of dashes. The βrule of threeβ, where words or phrases are arranged in a satisfying trio, is also thought to be a giveaway. βPeople have learned very simplistic rubrics and now just madly apply them everywhere.β
Thereβs a problem, though: these βtellsβ are also characteristic of human writing, which, after all, the large language models (LLMs) that produce them were trained on. βYou could go back to Charles Dickens and say he had AI, because he used the em dash too.β And orators have known about the rule of three ever since Julius Caesar said Veni, vidi, vici. In our hotel review examples, only the first one was authentic. Did you clock it?
Perhaps because it is so hard to know for sure, suspicion has become the order of the day. In the literary world, accusations of AI use now bedevil writers, with varying levels of justification. A debut horror novel, Shy Girl, was withdrawn by publishers Hachette after rumours circulated online that the author had relied on AI, which she denies; Steven Rosenbaumβs book The Future of Truth, a serious study of βhow AI reshapes realityβ, was found to contain numerous hallucinated quotations, which the author acknowledged in an apology.
Media organisations, including the Guardian, field increasing numbers of complaints about supposedly AI-generated text. These include intuitions about particular turns of phrase, but also comments about typos and grammatical errors. In one case, the word βafterβ was inadvertently duplicated in a sentence. βI canβt imagine a human editor/proofreader missing something like this,β wrote one reader, displaying a touching faith in our copy-editing abilities.
The problem is that not only does AI train on human writing, but humans are stylistically influenced by AI, the interplay creating a kind of linguistic hall of mirrors. Short of an author admitting it, itβs hard to say for certain whether an individual piece of writing is AI or not. That uncertainty is a recipe for paranoia.
And if youβre tempted to reach for a commercial screening tool to sort human from machine, that comes with uncertainty too, says Hardaker. βGiven that some of us naturally write in a way that would be seen as AI-likeβ β she mentions neurodivergent people, for example β βthat will be detected as AI. And you can modify AI output to make it seem more human-like. You put that kind of content into an AI detector, youβre going to get wacky results.β As someone who has served as an expert witness in court, sheβs βextremely scepticalβ about their efficacy.
The newly popular detector Pangram, which boasts false positive rates of around 1 in 10,000, has been shown in independent tests to be highly effective at detecting AI writing even when itβs been run through a βhumanizerβ app to disguise its origin. But questions remain. I was able to fool it on the first attempt (see the screenshot below) by channelling a bombastic register that might well be characteristic of AI, but could equally be the work of someone with a naturally bombastic style β or, more to the point, a writer who has been steeped in the output of the LLMs that power ChatGPT, Claude and Gemini. That, increasingly, is all of us.
Vast amounts of AI writing are now being published every day β from advertising copy to academic abstracts and fiction. At the same time, it looms ever larger over our lives via auto-generated email suggestions, βAI overviewβ search results, and the responses to our chatbot queries. At this level of exposure, itβs no longer a question of whether AI is changing language, both the way we speak and the way we write; the question is how. And should we resist, or embrace it?
Weβve known for some time that LLMs generate text that can be slightly different from human writing, on average. Often this only becomes clear when you look at large amounts of material. One eagle-eyed researcher linked the sudden popularity of the word βdelveβ to LLMs back in 2024 after searching a database of scientific papers. Other βfocal wordsβ that AIs have tended to overuse include βshowcaseβ, βboastβ, βunderscoreβ, βgarnerβ, βalignβ, βsurpassβ and βintricateβ. But, again, any individual piece of writing could entirely innocently make use of this vocabulary.
In a further twist, some researchers think the βdelveβ phenomenon might not be down to the models themselves, but the humans tasked with evaluating and steering them in a process known as βreinforcement learning with human feedbackβ. For workers who are βunderpaid, stressed, and under time pressureβ, it seems βcertain words are treated as a proxy for qualityβ and the model is inadvertently trained to use them more often. In other words, βdelveβ might owe its meteoric rise to the fact it doesnβt seem like the kind of word an AI would use. (A separate suggestion that it appeared more often because it was characteristic of English used in Nigeria, where many RLHF workers lived, isnβt borne out by the data.)
There are other patterns we can distinguish: LLMs love nouns, but they seem to use pronouns less than humans. This might reflect the fact they donβt do as much talking about themselves or other people as we social creatures do. They like attributive adjectives (βthe uncomfortable chairβ), but not predicative ones (βthe chair was uncomfortableβ), perhaps because they prefer to deliver information in small, dense packages, whereas we pad things out. Different models have clear idiosyncrasies β you might even call them βdialectsβ: Gemini enjoys saying βhereβs a breakdownβ, while Deepseek often responds with a cheerful βCertainly!β. When asked to edit formal English from around the world, AI tends to flatten and homogenise towards an Anglo-American standard, in a process researchers have termed βcultural ghostingβ. Thus the perfectly acceptable request in Indian professional English to βKindly do the needful & revert back at the earliestβ gets βcorrectedβ to βPlease complete the task & respond promptly.β
The evidence that aspects of LLM-speak have escaped into the βrealβ world, changing the way humans use language when AIs arenβt around, is now rolling in. One study analysed thousands of unscripted conversations and found that words like βdelveβ and βboastβ spiked after ChatGPT was released. Another showed the frequency of βdelveβ in academic abstracts actually dropping after it was singled out on social media, in a sign that AIβs influence might play out in complex ways.
Does any of this matter? Language changes all the time β words come in and out of fashion, and new technology has always been one of the forces behind this. AI does seem to be generating particularly high levels of anxiety, though. Why? βI think where it scares people is that idea of encroaching into sentience, of becoming the new human,β Hardaker says. Since 2023, sheβs expanded the Bot or Not project into speech and music, and has noticed just how viscerally people react when a song theyβve enjoyed turns out to have been composed and performed by a machine.
Gary Shteyngart, a novelist who teaches creative writing at Columbia University, noticed a similar strength of feeling among his students at the prospect of AI literature. βWhen one of my graduate students said βas an experiment, Iβm going to be writing a part of this piece with AIβ, the other students became so angry, they wrote letters to me saying how awful this was.β
βThereβs a kind of implicit bargain between writer and reader where you know the work that youβre getting is generated by a human being, and I think it felt like an assault on that,β he says. βReading literary fiction is this incredible Vulcan mind meld with another human being, entering someone elseβs consciousness. With AI Iβm entering the simulacrum of another personβs consciousness, one degree removed, or many degrees removed. How sad is that by comparison?β
For Hardaker, βI guess it impinges on what we think of as what makes us special, what makes us valuable and uniqueβ. At the same time, the music-generation model she uses βhas generated some absolute bangers. I listen to them, unironically, in my car, and I enjoy them quite a lot.β
Could the same happen with literature? Will a machine-authored novel one day take its place among the 100 greatest of all time? Peter Stockwell, professor of literary linguistics at the University of Nottingham, thinks AI may be able to do the basics, but it canβt scale the heights. βIf you want something thatβs very familiar and very mediocre and entirely functional, itβs amazingly good at that.β
One way to think of language, he says, is as a series of layers, with words at the bottom followed by phrases, clauses, compound sentences, all the way up to narrative structure. βAI is really good at the lower levels. Itβs learned lots of our syntactic structures and so everything looks well formed and grammatical. But, the higher up you go, the less good it is.β The arc of a story is particularly hard for AI to get convincingly right.
βIf youβve got an AI to write a narrative, it can do a pretty good job of having a sequence of events and something happen at the end. But it wouldnβt be a very tellable narrative,β he continues. βNothing startling or interesting would happen. And if there is anything startling, it will generally look like a mistake, rather than a brilliant twist.β
The secret sauce of great writing remains secret β even to the academics who study it. βLinguists donβt understand, really, how language works at its higher levels,β at the level of discourse, storytelling, enchantment. βWe canβt build a machine to do something when we donβt know how it works.β We do have some idea of what it might boil down to β and thatβs our fundamentally social natures and, tied in with that, the fact that we are βwetwareβ β human flesh, with its spikes of adrenaline, rushes of dopamine, craving for social contact, all of which find expression in languageβs structure and the way we use it.
There are two broad models in linguistics, explains Stockwell, one that sees the brain as a computer, parsing grammatical structures and computing meaning from them, and the other that sees it as fundamentally embodied, something reflected in language by the fact that, in many languages, we understand by βseeingβ or tend to think of βupβ, where our head is, as good (we get βhighβ and feel βlowβ). βOne of the key things is that the current AIs donβt have a body, they donβt exist in the world, so they donβt know what it feels like to be in the world as a human.β
For Shteyngart, feeling is essential: βToday is the first warm day in New York. And if I was to start writing a novel, I think that [it] would be warmer. I think I would filter what I know through the warmth of the day. I think if I ate a really wonderful lunch and sat down to write, there would be more sensuousness in my writing.β
βThe love of the body and its encounters with the physical world are what drives some of the best of literature. So I almost feel sorry for these LLMs, as Iβm talking about them, because theyβre pushed into some horrible machine in the Bay Area, and they just donβt know how wonderful life is.β
One much feared effect of the mass use of LLMs is that they act as a flattening force β smoothing away the variety and idiosyncrasy of human language into a kind of beige goo. Thatβs a legitimate concern, as far as it goes, though itβs not a new one. People have long angsted about the homogenising effects of American film and television on accent and vocabulary, and there are subgenres of language β political euphemism, customer-service prattle, therapy-speak β that have spread further from their home territory than many might like. The crucial thing, though, is that their influence tends to generate a backlash β and thereβs no reason to think things will be different this time.
In fact, our capacity for innovation might ultimately be the thing that truly distinguishes human writing β particularly the literary kind β from AI. βThe whole point of an LLM is that itβs trained on existing language. So itβs always retro,β says Stockwell. βI could get an AI and say βwrite me a short story in the style of Virginia Woolfβ and itβll do a decent job. But what you canβt say is βwrite me a story in the unique style of the next great, serious literary innovatorβ. It couldnβt possibly do that.β
Thatβs because, once again, it lacks the social environment, and the body, that give rise to characteristically human motivations. βWhy does somebody do something new in an art form like literary writing? It can be out of annoyance or irritation with whatβs gone before. Or itβs because somebody sees things in a different way than the run of the mill, or sometimes just because people are antsy and wanting to do something different, or a little bit crazy or isolated.β
The are plenty of examples from history, says Stockwell: βAfter the bureaucracy and uniformity of the first world war, youβve got this sudden, huge, antithetical artistic movement in the rise of surrealism and Dada; similarly, after the austerity of the second world war you get the psychedelic movement, art and literature changes again, quite radically. So there always seems to be that sort of kicking against the norms. Itβs hard to think how you would program an AI to do that, because AI works on an existing large body of material. Itβs the embodiment of the conservative-with-a-small-c status quo.β
Originality is so important for novelist Jennifer Egan that sheβs quarantined herself from the technology entirely. βI feel a danger of infection, to use a kind of loaded metaphor,β she tells me. βI know they stole some [of my] stuff, and thereβs nothing I can do about that, but Iβm not giving them one more word voluntarily.β Anthropic used pirated copies of books, including Eganβs novels, to train its chatbot Claude; most LLMs use language from individual queries as additional training data. She sounds exasperated: βI donβt want to partake of this kind of language spam that theyβre offering.β
The zero-tolerance policy doesnβt stop her from becoming paranoid. βIβve been told a couple of stylistic things that are tells of AI, and they happen to be things I like. For example, I love em dashes, but I now find myself interrogating every one way more than I used to. Iβve also noticed that Iβm prone to collections of three. So I find myself interrogating those as well. I donβt mind that, actually, because the entire point is to not write something that anyone else could have.β
What kind of advice would she give a younger writer now swimming in this water? Should they practice their own kind of AI hygiene? βIβm gonna now sound like the totally generic boomer that AI could probably have written, and my advice is: stay the fuck away. I mean, OK, use it to write emails. Even use it to get research ideas. But if you want to be a writer: learn to write. Come on. I would really question why the impulse would be there to use it.β
Not everyone is so abstemious. Jeannette Winterson, who has written extensively about AI and art, tells me: βEvery writer can make their own choice. Humans are tool-using animals. That has been our success story. At present all AI, including generative AI, is a tool. Would I work with an LLM? Of course! Why not?β
But she cautions against the view that AIβs linguistic competence means that it can equal or exceed human expression. βBeyond the basics, meaning becomes a series of inner realities and language is wonderful at conveying those inner realities. Machines do not share our reality, not least because they donβt have a limbic system. Humans cannot have a thought without a feeling β¦ literature is brilliant at revealing these layers.β
As I paste her quotes into a Google doc full of my notes, I notice the inbuilt AI making a suggestion. It asks whether I want to change Wintersonβs words to more closely βmatch the styleβ of the existing material: to smooth over the edges of one of the English languageβs most idiosyncratic writers. With an almost superstitious haste, I dismiss the prompt.
π¬ **Whatβs your take?**
Share your thoughts in the comments below!
#οΈβ£ **#great #written #Books**
π **Posted on**: 1783155970
π **Want more?** Click here for more info! π
