I gave a talk about AI today for a group of knowledge workers in the public sector – people who are asked to write reports summarising research on a specific topic, and who coordinate funding schemes and advise on policy – and I asked them how they use LLMs in their work. Based on their free-text responses on Mentimeter it looks like there are four types of things they have found LLMs actually useful for:

  1. Translation and copy-editing – checking the comma rules was one example, another was revising a text to be easier to read
  2. Summarising long texts
  3. Structuring information, comparing data (but this was also in the “didn’t work” category)

I also asked for examples of things they had tried to use LLMs for that hadn’t worked:

  1. Literature searches and finding factual information
  2. Analysing a topic or texts, reflective analysis, complex topics, public consultations (høring, an institutional response to a policy proposal),
  3. Summarising large or complex documents
  4. Excel
  5. Making a seating chart
  6. Maintaining ambiguity or multiple perspectives, e.g. in an article with two argumments
  7. Anything related to specialised expertise that is not heavily discussed online

These are slightly different types of tasks to the ones I wrote about yesterday, in the benchmarking test where LLMs failed at 70% of office tasks. My main advice in the talk today was to think carefully about whether or not a language model is actually the right tool for the task at hand – and to consider what is lost by not doing it yourself. Sometimes speed and a good-enough product is the most important thing. Other times the friction of working with a process is important in itself because that is where the thinking happens. And then there are tasks that are just better suited to other tools.


Discover more from Jill Walker Rettberg

Subscribe to get the latest posts sent to your email.

Leave A Comment

Recommended Posts

AI STORIES

AI-generated stories have longer endings than human stories

My colleague Jessica Witte has just shared a preprint where she compared the emotional arcs of the stories we generated using gpt-4o-mini to those of human-authored (pre-2022) stories from the subreddit r/WritingPrompts, conveniently gathered in this dataset. She found a distinct difference in the endings of the LLM-generated stories: both […]

How to peer review a paper in 2026

Dorothy Bishop, a psychologist, wrote a useful list of what reviewers need to look out now that so many papers are bad science that looks good thanks to LLMs. Read her whole blog post, she explains it well, but here’s a brief version because I’m pretty sure we’ll be needing […]

“So what if it was ChatGPT? It *could* have been true!”

I recently read a good article on the different kinds of truth a language model operates with by Luke Mann, Liam Magee and Vanicka Arora, Truth Machines: Synthesizing Veracity in AI Language Models, but despite its lovely typology of truths (consensus, correspondence, coherence and pragmatic) it doesn’t help me with […]

AI STORIES

AI shimmer and sparkle

I have this hunch that sparkles and glow and shimmer are somehow a point in the latent spaces of LLMs that have more connections than you would expect. Perhaps their connotation to magic and to the unknown matches some of the mystique of genAI? Or perhaps these words are used […]

Don’t do a systematic review if you’re in the humanities

This paper is a great example of why you probably shouldn’t use a systematic literature review for a theoretical and conceptual research question like “How does artificial intelligence affect the perception of authenticity and aura in art?” However, if you’re looking for an annotated list of 48 recent articles about […]