Show someone a few hundred pictures, then test them later, and their memory is almost eerily good — people can recognize images they saw once, out of thousands, well above chance. Do the same with words and the recall is markedly worse. This gap has a name, the picture-superiority effect, and the leading explanation is quietly useful for anyone designing a first screen.
The idea, from Allan Paivio, is called dual-coding. A word gets encoded one way — verbally, as a symbol. A concrete picture gets encoded twice: once as an image and once as the word it evokes. Two routes into memory, two chances to retrieve it later. That redundancy is why a diagram sticks when its caption evaporates, and why you can recall the look of a room years after forgetting anything anyone said in it.
Now put that next to the screen your users see most and understand least: the empty state, the first run, the moment before the product has done anything. That screen almost always defaults to words. A headline, a subhead, three bullet points explaining what will happen once you get going. It’s the honest instinct — explain the thing — but it’s fighting the brain’s filing system. Prose is the single-coded channel. It’s the one people forget.
The design translation is almost too simple: in the places where you most need someone to get it fast and remember it, show the finished state instead of describing it. An empty task list that displays a faint, realistic example of a filled-in list. A blank canvas that shows a ghosted version of what a good first result looks like. A pricing page that pictures the outcome, not just lists the features that produce it. You’re not decorating the screen — you’re giving the concept a second channel so it lands and stays.
But there’s a real catch, and it’s the difference between using this effect and cargo-culting it. The picture superiority only shows up for images that carry the meaning. A concrete, relevant picture of the thing itself codes twice. A stock photo of smiling people at a laptop codes as… a stock photo — it’s a separate item competing for the same attention. Mayer’s research on “seductive details” found that vivid-but-irrelevant images can actually lower comprehension, because they pull working memory toward the wrong thing. So the rule isn’t “add an image.” It’s “make the image be the point.” If you could delete the picture and lose none of the meaning, it was decoration, and decoration doesn’t get the memory bonus — it just spends attention.
The other honest boundary: this is about concreteness. Picture superiority is strongest for things you can actually depict. Abstract ideas — “privacy,” “flexibility,” “trust” — don’t have a natural image, and forcing a metaphor onto them often adds a puzzle rather than a shortcut. For those, words may genuinely be clearer, or you find the one concrete instance that stands in for the abstraction (a lock you control, a setting you can flip). The skill is knowing which of your concepts have a true picture and which are pretending.
Put together, it’s a small reallocation with an outsized payoff. The screens where you’re most tempted to over-explain are usually the screens with the least earned attention — a first run, a blank state, a moment of “wait, what is this.” That’s exactly where a single honest picture of the destination does the work three paragraphs can’t, and does it on a channel the brain won’t drop on the way out the door.
Liked this? Get the next one in Working Theory.
Going weekly in August (it's in beta now). One genuinely interesting read on building, the brain, and the science most people missed.