80% of the articles published in AI & Society seem to be AI slop these days. OK, maybe that’s an exaggeration, but there are a lot. I hate wading through the slop to find the good articles they also publish. After seeing two “Curmudgeons’ Corner” opinion pieces yesterday that were very obviously AI-generated I fired off a Curmudgeons’ Corner of my own to the editors. It’s titled “AI & Society needs to stop publishing AI slop opinion pieces”. I hope they publish it – it’s certainly curmudgeonly, which is what they ask for. If they don’t I’ll send it somewhere else and publish it here.

I have an article of my own forthcoming in AI & Society, so I’m a little disappointed that the quality of the journal seems to be nosediving so rapidly. My article is precisely about AI slop. Here’s a link to the not-yet-resolving DOI in case you read this in the future and it’s already been published.

So I skimmed one of today’s freshly published LLM-generated articles on AI & Society, “Designing for rhythm: tempo-setting infrastructures and the temporal conditions of digital life” by Robert Atkinson, a professor of computer science at Arizona State University. The title made me suspicious: it has that dense rhythm that LLMs aim for in an academic article, but then again, that’s how we’re trained to write. The title of my forthcoming article might not be that much better, though I like to think it’s more concrete: “LLMs are metaphor machines: orbital argumentation, misaligned citations, and scientific fabrication in AI-generated writing”. It has a similar rhythm, unfortunately. Atkinson’s title has that evocative vagueness to it that LLMs love and, I suspect, that computer scientists think is the way humanities and social science research should be written. I ran it through Pangram, and unsurprisingly, Pangram says it’s 87% AI. That seems quite a generous assessment.

Anyway, I kept reading, and quickly saw exactly the same LLM-produced scientific fabrication in Atkinson’s article as I analyse in my forthcoming paper. Atkinson doesn’t cite well-known research on his topic, but instead cites an irrelevant paper by the author of one of the books that should have been cited. These are misaligned citations, and I argue that they’re a systematic error made by LLMs.

Here’s the abstract of my paper about this phenomenon:

Screenshot of title and abstract of forthcoming paper.
I sent back the corrected proofs of my article about misaligned citations and scientific fraud a few days ago.

Let me explain. Atkinson’s topic is digital temporality, and he starts off by saying there’s not much research on this, which surprised me, because I have been hearing about research on the topic for decades and it’s not even something I’ve directly done research on myself.

I read STS legend Judy Wajcman’s book Pressed for Time: The Acceleration of Life in Digital Capitalism a few years ago, and know many other scholars are publishing on the topic. So this seems like an odd claim to make. I did a quick Google Search to check. I’m in Greece right now on a writing retreat, so that’s why my Google Scholar interface is nice and Greek today.

OK, that’s just the first few hits. Let’s see if Atkinson cites Rob Kitchin’s book, the top hit, Digital Timescapes: Technology, Temporality and Society. Nope. But look, he cites a different article by Kitchin from 2017 about something not related to temporality. And wow, the sentence the reference is supposed to support is a great example of an LLM-generated not X but Y sentence too, just for good measure. With three references to articles that are all about digital media but not really specifically about the very vague claim made in the sentence.

OK, so the citations here are drive-by citations, performative decorations that make a text look like scholarship without actually doing the work of scholarship. The point of citing sources is to engage in a dialogue with previous researchers, to build on or challenge their work. Drive-by citations don’t do that. They’re performative rather than actually doing the work of research. And often, especially in LLM-generated texts, they’re completely irrelevant to the claim being made.

But this isn’t just an irrelevant citation. The reason the LLM picks Kitchin 2017 is that it should have cited Kitchin 2023, the book Kitchin wrote specifically about the topic of the whole article.

LLMs don’t pick references logically. They pick them based on statistical pattern recognition, they pick them based on proximity between tokens in the text in the context window (in this case the text Atkinson had already generated at this point of writing the article, or if he pasted the whole article into a tool like Grammarly and asked it to add references, the full text of the article) and tokens in the training data.

You could imagine a chain of connections, though that’s not quite what the multidimensional semantic space of an LLM really looks like. I think this might be a useful way for humans to analyse what’s going on, though.

digital temporality – Kitchin 2023 – Kitchin 2017 – “They organize the rhythm through which information, communication, and social activity become available (Bucher 2018; Gillespie 2014; Kitchin 2017).”

My forthcoming article analyses equivalent examples in another clearly LLM-generated or at least heavily LLM-assisted paper, and I propose a theoretical argument for why this happens, building on John Gallagher’s concept of orbital argumentation, Anne Sigrid Refsum’s concept floating motifs and the narratological concept of unnarratability.

Atkinson’s article doesn’t cite Wajcman or the authors of other top Google hits for publications about digital temporality. This orbital argumentation where the LLM can’t directly name the originator of a term (as in the example I discuss in my forthcoming article) or the key scholars in a field, but instead pulls them in in the wrong way doesn’t happen every time. But I think every single irrelevant or half-relevant reference provided by an LLM (like Grammarly) is actually a misaligned citation, a citation generated through networks of similarity that are in the model but not direclty visible to us in the generated output. That means that AI-literacy requires a reader to evaluate those hidden connections – we can no longer assume there is a logical reason for a reference. We can’t even assume it’s sloppy scholarship (though it’s AI slop). It’s systematically wrong, and a symptom of something else having been left out. In this case, what’s left out is the extensive scholarship on the topic that is not mentioned in the article.

I have a lot more to say about this. I think it’s extremely important to understand. Please read my article when it’s out and tell me what you think! Oh, and if you see more examples of misaligned citations like this, please let me know in the comments!


Discover more from Jill Walker Rettberg

Subscribe to get the latest posts sent to your email.

1 Comment

  1. […] Babies Were Aborted to Make the MMR Vaccine Misaligned citations and LLM-generated scientific fraud Behold the ‘glueball,’ a strange new form of matter A Pragmatist-Artifactualist view of […]

Leave A Comment

Recommended Posts

AI STORIES

AI-generated stories have longer endings than human stories

My colleague Jessica Witte has just shared a preprint where she compared the emotional arcs of the stories we generated using gpt-4o-mini to those of human-authored (pre-2022) stories from the subreddit r/WritingPrompts, conveniently gathered in this dataset. She found a distinct difference in the endings of the LLM-generated stories: both […]

How to peer review a paper in 2026

Dorothy Bishop, a psychologist, wrote a useful list of what reviewers need to look out now that so many papers are bad science that looks good thanks to LLMs. Read her whole blog post, she explains it well, but here’s a brief version because I’m pretty sure we’ll be needing […]

“So what if it was ChatGPT? It *could* have been true!”

I recently read a good article on the different kinds of truth a language model operates with by Luke Mann, Liam Magee and Vanicka Arora, Truth Machines: Synthesizing Veracity in AI Language Models, but despite its lovely typology of truths (consensus, correspondence, coherence and pragmatic) it doesn’t help me with […]

AI STORIES

AI shimmer and sparkle

I have this hunch that sparkles and glow and shimmer are somehow a point in the latent spaces of LLMs that have more connections than you would expect. Perhaps their connotation to magic and to the unknown matches some of the mystique of genAI? Or perhaps these words are used […]

Don’t do a systematic review if you’re in the humanities

This paper is a great example of why you probably shouldn’t use a systematic literature review for a theoretical and conceptual research question like “How does artificial intelligence affect the perception of authenticity and aura in art?” However, if you’re looking for an annotated list of 48 recent articles about […]

Screenshot of a paragraph from a New York Times article published May 12, 2026. Text reads: "The price of tomatoes -tart bursts of flavor in salads and sandwiches — surged nearly 40 percent in April from a year ago on a combination of bad weather, high tariffs and climbing transportation costs."
AI STORIES

Genre glitches and unexpected promotional phrases as a sign of AI writing

A genre glitch is a characteristic of LLM-assisted writing where the text suddenly switches genre, typically inserting a short promotional phrase full of sensory details into an informational text. Genre glitches occur when a word in the generated text is heavily associated with a genre or context that is markedly […]