{"id":18862,"date":"2026-08-09T05:02:47","date_gmt":"2026-08-09T05:02:47","guid":{"rendered":"https:\/\/futureknowledge.in\/?p=18862"},"modified":"2026-08-09T05:02:47","modified_gmt":"2026-08-09T05:02:47","slug":"how-to-spot-if-something-has-been-written-by-ai","status":"publish","type":"post","link":"https:\/\/futureknowledge.in\/?p=18862","title":{"rendered":"How to spot if something has been written by AI"},"content":{"rendered":"<p>You have reached your maximum number of saved items.<\/p>\n<p>Ghost writer is haunting the English language. The linguistic spectre can turn its hand to prose, poetry, journalese and corporate jargon. It is frightfully versatile: you can get it to mimic Shakespeare\u2019s sonnets or a schlocky beach read; Ernest Hemingway\u2019s taut prose or the office-printer manual. It is frightfully fast, churning out thousands of words a minute. (Hemingway rarely produced as many in a day, and required much more booze.) Wordsmiths are spooked.<\/p>\n<p>AI writing is everywhere. It is in your inbox and on your LinkedIn feed. It is all over the internet, drafting more than a third of new websites by one count. Large language models (LLMs) are helping students write essays and probably helping scientists write papers. Some allege AI-generated prose won the Commonwealth Short Story prize this year, with judges praising its \u201cquiet authority\u201d. (The Commonwealth Foundation denied the claim.)<\/p>\n<p>LLMs have stylistic quirks. They are thought to maximise the use of long em-dashes\u2014and the use of words like \u201cmaximise\u201d. They like to \u201cdeep dive\u201d (and, better yet, \u201cdelve\u201d) into the \u201crich tapestry\u201d of the world. AI writing is not about a single word or phrase, but a rich tapestry of things.<\/p>\n<p>Spotting AI texts can be tricky. This is in part because you need evidence beyond a few words or dashes: claiming that a text is by an LLM because it uses the word \u201cdelve\u201d is like claiming one is by Jane Austen because it uses \u201cimprudence\u201d. Bots also write in slightly different ways. There is no single style of AI writing, explains Karolina Rudnicka, a linguist at the University of Gdansk in Poland, just as there is no single style of human writing. Writers have idiosyncrasies\u2014Emily Dickinson, for instance, loved em-dashes\u2014and bots may do, too.<\/p>\n<p>But there are a few ways to identify LLM-generated text. One is to use detection algorithms that are trained to spot the texture of human or AI prose. Pangram, a leading firm, claims to have 99.98 per cent accuracy. (It has partnered with Substack, a blogging platform, on such a tool.) Detectors, however, are black-box algorithms that can give false positives. They do not give reasons for why they reach their conclusions.<\/p>\n<p>Researchers have also tried scouring texts for suspicious words or comparing papers from before and after LLMs were made available to the public. But these approaches have drawbacks too, not least because it is hard to disentangle AI quirks from other language trends.<\/p>\n<p>You can discover AI\u2019s hallmarks by comparing the writing of man and machine. To do this you need a baseline that is distinctive and familiar. The Economist turned to prose that we\u2019re sure is human and that readers will recognise: our own. We designed a study to ask top LLMs\u2014OpenAI\u2019s ChatGPT, Anthropic\u2019s Claude, Google\u2019s Gemini and xAI\u2019s Grok\u2014to write versions of our articles without consulting the web. (As a prompt, we gave them the AI-generated summaries that we have experimentally added to some of our articles.)<\/p>\n<p>This gave us a corpus of human and AI creations and we compared them across 55,940 sentences and 1.2 million words. To make sure we were detecting AI quirks rather than our own, we also checked the AI texts against journalism from CNN, the New York Times and the Washington Post. Excerpts from hit novels published between 1950 and 2022 offered another test.<\/p>\n<p>Our findings are surprising. AI prose is distinguishable by word and punctuation choice as well as sentence and paragraph structure. But its hallmarks are not what you might expect, partly because its writing style has changed with software updates. That does not mean that LLMs are great writers: their prose lacks lucidity and elegance and is often formulaic. So those aspiring to be impressive (human) storytellers should avoid the following peculiarities in their own prose.<\/p>\n<p>First, consider words. The vocabulary that bots overuse has changed: they no longer \u201cdelve\u201d and there are not as many \u201ctapestries\u201d. Instead they offer a significant number of polysyllables like \u201csignificant\u201d, \u201cincreasingly\u201d and \u201cconsequences\u201d. They use more rare words (\u201cinterdependence\u201d, \u201creindustrialisation\u201d) and scientific lingo (\u201cparameter\u201d, \u201cmethodology\u201d) than humans, and are fond of nominalisations (making nouns from verbs, such as \u201cexpansion\u201d from \u201cexpand\u201d). All the LLMs in our study use such words, but particularly Gemini and Claude.<\/p>\n<p>Much of this language could be described as what George Orwell called \u201cpretentious diction\u201d. He railed against writers who \u201cdress up simple statements\u201d with complicated words and jargon to sound clever. Such pontificating penmen, Orwell observed, also believe that \u201cLatin or Greek words are grander than Saxon ones\u201d. (Bots agree: more Latinate suffixes crop up in their writing than in human texts.)<\/p>\n<p>Then look at punctuation. Many believe LLMs stuff their prose with em-dashes, but that is not true after the most recent updates. Today only Claude uses more em-dashes than human writers, with ChatGPT using markedly fewer than any other writer in our study. Humans rejoice\u2014and start using dashes again.<\/p>\n<p>A better way to spot AI-generated writing would be to look for texts without much punctuation at all. LLMs are very Joycean about it: they use fewer commas and semicolons than humans (and hardly any parentheses). They use less punctuation in part because they write longer sentences\u2014\u201cand\u201d is their most overused word\u2014and in part because they do not quote experts.<\/p>\n<p>Finally, study the sentence. Bots\u2019 sentences tend to be long; paragraphs are rarely interrupted with short, punchy statements. How dull. When LLMs want to make their sentences more lively, they often reach for a rhetorical device. Their favourites include: \u201cnot X but Y\u201d, \u201cnot only but also\u201d and the \u201crule of three\u201d. (Grouping ideas in threes makes them more engaging, as we did just then.) ChatGPT and Claude use more of these constructions per 1,000 sentences than other LLMs and humans.<\/p>\n<p><em>Source: <a href='https:\/\/www.theage.com.au\/technology\/how-to-spot-if-something-has-been-written-by-ai-20260807-p60m9d.html?ref=rss&#038;utm_medium=rss&#038;utm_source=rss_technology' target='_blank'>Read the original article on www.theage.com.au<\/a><\/em><\/p>\n","protected":false},"excerpt":{"rendered":"<p>You have reached your maximum number of saved items. Ghost writer is haunting the English language. The linguistic spectre can turn its hand to prose, poetry, journalese and corporate jargon. It is frightfully versatile: you can get it to mimic Shakespeare\u2019s sonnets or a schlocky beach read; Ernest Hemingway\u2019s taut prose or the office-printer manual. [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":18863,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[4,36,3],"tags":[10,28,34],"class_list":["post-18862","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-important","category-share-suggestions","category-technology","tag-impact-googl","tag-signal-buy","tag-stage-stage-2"],"_links":{"self":[{"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/posts\/18862","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=18862"}],"version-history":[{"count":0,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/posts\/18862\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/media\/18863"}],"wp:attachment":[{"href":"https:\/\/futureknowledge.in\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=18862"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=18862"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=18862"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}