Why do I lose the thread when someone talks for more than two minutes?
Because listening is real-time processing, and your working memory holds about four chunks. A two-minute explanation contains dozens. The thread drops when a new idea can't connect to one you're still holding, and the person explaining can't see which link you missed. That is overflow, not an attention-span failure. Rescue it with one targeted question, not "can you repeat that?"
Where does the thread actually go?
It runs out of space. Your focus of attention holds about four chunks at once, and closer to three when you cannot rehearse. A two-minute explanation arrives at roughly 150 words per minute, which is dozens of ideas, and each one needs a buffer slot while the next arrives. The chain breaks when every slot is taken.
Cognitive psychologist Nelson Cowan's review of working-memory research established the four-chunk limit, and Chen and Cowan's follow-up put the core verbal limit closer to three when rehearsal is blocked. At about 150 words per minute, a two-minute explanation is roughly 300 words: dozens of separate ideas, links, names, and numbers. Each needs a buffer slot while being integrated with the next, and the moment idea N+1 arrives with every slot taken, the chain breaks. You didn't lose the thread. The buffer overflowed, and that is a capacity fact, not a verdict on your listening.
Why do unfamiliar topics make it worse?
Because chunking needs prior knowledge. When you already understand a domain, related details compress into a single chunk: "the Q3 revenue call" is one item, not thirty. With an unfamiliar topic, nothing compresses, so every idea burns a whole slot and the buffer fills in seconds.
This is also why working-memory capacity predicts listening comprehension in the Daneman and Carpenter research: the same buffer has to store and process at once, and processing unfamiliar material consumes the storage. It is not that you can't listen to experts explain their field. It is that their field arrives un-chunked, and there is nowhere to put it.
Why does the explainer just repeat the same thing, louder?
The curse of knowledge. In the classic 1989 study by Camerer, Loewenstein, and Weber, people who knew more systematically overestimated how much others knew, and even trading markets only cut that bias by about half. The explainer built the explanation on assumptions you were never told, then re-sends the same layer because they can't see which link you are missing.
The original study, published in the Journal of Political Economy, is the classic source for this effect. You know the loop: "it's hot! At! the bottom!" same words, more volume, as if loudness were the missing ingredient. That is not a listening failure on your side. It is an explaining failure on theirs, and knowing that changes what you do next.
Is my attention span actually the problem?
No, and the research that says attention dies on a timer does not hold up. Wilson and Korn's 2007 review found no evidence for the famous claim that students can only focus for 10 to 15 minutes, and Bunce and colleagues' 2010 clicker data show attention waxing and waning across a lecture rather than sliding steadily downhill.
The Wilson and Korn review in Teaching of Psychology is the go-to debunk of the ten-minute claim. Stuart and Rutherford's classic study actually found concentration peaking 10 to 15 minutes in, then falling. What does decay fast is vigilance during passive, monotonous formats: the Southampton lecture research found the vigilance decrement kicking in within minutes of one-way delivery, while varied presentation kept it away. So the two-minute cliff is not a magic timer. It is overflow plus monotony plus a format with no rewind, and all three are fixable. None of them are you.
How do I get the thread back mid-explanation?
Four moves, in order. First, pre-load before the conversation: skim the doc, check the agenda, spend thirty seconds building chunks to hang ideas on. Second, at the first sign of overflow, ask one targeted question: "Which part connects X to Y?" That pinpoints the exact link you missed and turns a monologue into a dialogue.
"Can you repeat that?" re-sends the same stream; a targeted question forces the explainer to find the missing bridge. Third, run a two-minute paraphrase checkpoint: restate the last idea in your own words, briefly. It forces the storage-and-processing loop to do real work, and lets the explainer correct you while the cost is still one sentence. Fourth, take one written anchor note per explanation. One line. That is your external slot, and it frees the buffer for the next link.
The same reflex works when you are the explainer. People who actually know what they are doing tend to explain less, because they have already been through the messy part and can compress. Structure beats volume: check understanding, then add one layer at a time.
Every lost thread is a decision you will re-litigate later, a doc you will reread, a question you will ask twice. That is the real cost, and it is why the paraphrase-and-chunk reflex is worth training like a skill. It is a comprehension habit, and it is exactly what Absorb's daily listening practice builds: short, real-world listening and analysis workouts that make you actively process what you hear instead of letting it wash past. If you want to train the reflex before your next long explanation, see Absorb on the App Store.