I recently read a good article on the different kinds of truth a language model operates with by Luke Mann, Liam Magee and Vanicka Arora, Truth Machines: Synthesizing Veracity in AI Language Models, but despite its lovely typology of truths (consensus, correspondence, coherence and pragmatic) it doesn’t help me with this example of a ChatGPT untruth from last May.

So someone wrote an op-ed for a Norwegian newspaper and quoted a much loved former prime minister, Kåre Willoch, as having said “Russia annexed Crimea because they were afraid that Ukraine might become part of Russia.” Today another op-ed was published by Eirik Sakariassen who said he had been surprised by the quote so had looked for it but couldn’t find it anywhere. When he asked the author of the original op-ed he said he had got the quote from ChatGPT.

Argh. Eg reagerte også på det Willoch-sitatet då eg las innlegget til Martens, men kunne ikkje førestilla meg at det var funne på av ChatGPT.

magnusbe (@magnusbe.bsky.social) 2025-05-30T08:53:40.915Z

OK, so far this is a typical when-will-people-stop-thinking-everything-they-generate-with-ChatGPT is true thing. Stuff like this happens every day. I mean the same day we also heard that RFK published a report where several of the sources were made up. A few days earlier several big newspapers had published that summer reading list of non-existent books. The month before Tromsø politicians voted to permanently shut down seven schools based on a report that turned out to be AI-generated and cite made-up sources. Etc.

But the next thing that happened was different. People started posting that Kåre Willoch could have said that.

I dette sitatet – fra intervjuboka med Torbjørn Røe Isaksen – går han langt i den retningen, jeg ville tro det er herfra misforståelsen stammer. www.nb.no/items/c4feef… (krever innlogging)

Anders Heger (@andersheger.bsky.social) 2025-05-30T09:18:31.627Z

And sure, Willoch sort of discusses the general topic but he certainly doesn’t say exactly what the first op-ed quoted him as saying.

But does it matter? I think it does. Yes, humans also make stuff up and misremember things. But LLMs do this systematically. That’s a major difference, and it means we need to systematically distrust LLM-generated claims.


Discover more from Jill Walker Rettberg

Subscribe to get the latest posts sent to your email.

Leave A Comment

Recommended Posts

AI STORIES

AI-generated stories have longer endings than human stories

My colleague Jessica Witte has just shared a preprint where she compared the emotional arcs of the stories we generated using gpt-4o-mini to those of human-authored (pre-2022) stories from the subreddit r/WritingPrompts, conveniently gathered in this dataset. She found a distinct difference in the endings of the LLM-generated stories: both […]

How to peer review a paper in 2026

Dorothy Bishop, a psychologist, wrote a useful list of what reviewers need to look out now that so many papers are bad science that looks good thanks to LLMs. Read her whole blog post, she explains it well, but here’s a brief version because I’m pretty sure we’ll be needing […]

AI STORIES

AI shimmer and sparkle

I have this hunch that sparkles and glow and shimmer are somehow a point in the latent spaces of LLMs that have more connections than you would expect. Perhaps their connotation to magic and to the unknown matches some of the mystique of genAI? Or perhaps these words are used […]

Don’t do a systematic review if you’re in the humanities

This paper is a great example of why you probably shouldn’t use a systematic literature review for a theoretical and conceptual research question like “How does artificial intelligence affect the perception of authenticity and aura in art?” However, if you’re looking for an annotated list of 48 recent articles about […]

Screenshot of a paragraph from a New York Times article published May 12, 2026. Text reads: "The price of tomatoes -tart bursts of flavor in salads and sandwiches — surged nearly 40 percent in April from a year ago on a combination of bad weather, high tariffs and climbing transportation costs."
AI STORIES

Genre glitches and unexpected promotional phrases as a sign of AI writing

A genre glitch is a characteristic of LLM-assisted writing where the text suddenly switches genre, typically inserting a short promotional phrase full of sensory details into an informational text. Genre glitches occur when a word in the generated text is heavily associated with a genre or context that is markedly […]