Theory of mind: why the best stories are about who knows what
You know the innkeeper poisoned the wine. Your companion does not. She raises the cup, and for one sentence the whole story is the distance between what you know and what she believes.
What theory of mind is
Psychologists test it with a marble. Sally puts the marble in her basket and leaves the room. Anne moves it to her box. Sally comes back. Where will she look?
To answer the basket, you have to hold two versions of the world at once: where the marble is, and where Sally thinks it is. Most children manage it from about the age of four. Younger children tend to answer with where the marble really is, because for them there is only one world, the true one.
Fiction is a machine for knowledge gaps
Stories are built out of that double bookkeeping. The literary scholar Lisa Zunshine argues that we enjoy fiction because it keeps inviting us to guess at other minds, which our brains do all the time anyway (Why We Read Fiction, 2006). Look at what the oldest devices of storytelling are made of:
- Dramatic irony. The audience knows something a character does not. In Romeo and Juliet we know Juliet is only asleep; Romeo does not, and acts on what he believes.
- Suspense. The audience knows the danger is coming, and the characters do not.
- Mystery. Nobody knows yet, and the reader finds out alongside the detective.
- The lie and the betrayal. Someone acts on a false belief that someone else planted, or on a loyalty that was never real.
- Minds inside minds. She suspects that he knows that she lied. Readers follow several of these layers without noticing they are doing it.
Some researchers went further and claimed that reading literary fiction trains this ability. A 2013 study in Science reported exactly that. When the same authors ran three preregistered replications, the results were mixed: one succeeded, and two failed without settling the question (Kidd and Castano, 2019). In two of those experiments, readers also judged characters in popular fiction more predictable and more stereotyped than characters in literary fiction. Whether or not fiction trains theory of mind, it runs on it.
Why an AI narrator gets it wrong by default
A language model writing a story sees everything in its context at once: the villain's plan, the heroine's secret, the name of the stranger you have not met yet. Nothing in that text says who was in the room when something was said. So its default is an omniscient cast, where every character knows what the prompt knows.
I ask the woman at the gate for directions to the old mill.
"Of course, Elowen," she says. "Though I'd have thought you knew the way to your father's mill."
Nobody has told her your name, or who your father was.
In our own playtests this month we met all three classic leaks: a name the hero had never heard, a hidden truth written plainly into the prose, and a twist given away before its scene.
This is not because models cannot reason about beliefs. Given a classic false-belief puzzle, recent models often solve it: a 2024 study found GPT-4 solving 75% of its tasks, about as well as six-year-old children (Kosinski, PNAS). Others showed that small changes to the same puzzles can make models fail (Ullman, 2023), and the debate is still open. In a story, the problem comes before any reasoning. The narrator is handed everything, and nothing separates what each character has actually witnessed.
What it opens up for a player
When a story keeps track of who knows what, a kind of play appears that an omniscient cast makes impossible:
- The bluff. You lie to a guard, and the lie holds, because he has no way to know better.
- The secret. What you hide stays hidden until you, or the story, reveal it.
- The investigation. Clues are found, not recited. Nobody hands you the answer because the narrator happens to know it.
- The misunderstanding. Someone acts on a false belief, and the consequences are real.
- The betrayal. An ally knew all along, and there is a scene where you realise it.
- The surprise. The story knows something you don't, and keeps it.
What we are building, in broad strokes
Nostyss keeps track of who has met whom and who has learned what, and writes each scene from inside what your character knows. A secret stays out of the prose until the story reveals it. Someone you have not met is described rather than named, until the story introduces them.
This is recent work, and it is not finished. It matters to us more than any single feature, because it is the difference between a story that remembers and a story with people in it.
What still breaks
- Inference. A sharp character should be able to guess. We would rather she guessed than simply knew, and drawing that line is hard.
- News off the page. Rumours travel between scenes, and a story has to decide what reached whom.
- The second order. What she believes you know is harder to keep straight than what she knows.
- The suggestions. The choices we propose can still name someone you have not met yet.
None of these is fixed by a bigger model. They are fixed, when they are, by the story keeping better books.
Why this is the fun part
A story is not only what happens. It is who knows it, when, and what they do about it. Give the player a mind among other minds, and the oldest pleasures of fiction come back: the held breath, the lie that works, the reveal that lands.
To check this on any AI story app, the fourth of our five memory tests does exactly that: does a character know something they never learned?
Questions
What is theory of mind in simple terms?
The ability to understand that other people have their own beliefs, knowledge and intentions, which can differ from yours and from the truth. It is what lets you predict that someone will look for their keys where they left them, not where you moved them while they were out.
What is an example of theory of mind in a story?
Dramatic irony. In Romeo and Juliet, the audience knows Juliet has only taken a potion that makes her seem dead; Romeo does not, and acts on what he believes. The scene works because the audience holds both minds at once.
Why do AI characters know things they shouldn't?
Because the model writing them sees the whole story at once, including secrets and names the characters never learned. Unless the story keeps track of what each character has actually witnessed, every character draws on the same pool of knowledge, and secrets leak.
Can AI have theory of mind?
Recent language models solve many classic false-belief tests: a 2024 study found GPT-4 solving 75% of its tasks, about as well as six-year-olds. Other researchers showed that small changes to the tests can make models fail, and the question is still debated. In a story, the bigger problem is structural: the narrator is told everything.
What is dramatic irony?
Dramatic irony is when the audience knows something a character does not. The tension comes from watching the character act on a belief we know is wrong.
Sources
- Premack & Woodruff, Does the chimpanzee have a theory of mind? · Behavioral and Brain Sciences, 1978
- Wimmer & Perner, Beliefs about beliefs · Cognition, 1983
- Baron-Cohen, Leslie & Frith, Does the autistic child have a theory of mind? · Cognition, 1985
- Sally–Anne test · Wikipedia
- Zunshine, Why We Read Fiction: Theory of Mind and the Novel · Ohio State University Press, 2006
- Kidd & Castano, Reading Literary Fiction Improves Theory of Mind · Science, 2013
- Kidd & Castano, Reading Literary Fiction and Theory of Mind: Three Preregistered Replications · Social Psychological and Personality Science, 2019
- Kosinski, Evaluating large language models in theory of mind tasks · PNAS, 2024
- Ullman, Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks · arXiv, 2023
See whether it actually remembers
Ten worlds you can start without an account. Play a few chapters, contradict something you established earlier, and watch what the world does about it. That test is the honest way to judge any of this.
Choose a world