There is a narrow statement at the heart of this April 2 YouTube video about audiobooks and reading that is plainly true: using the eyes to decode printed symbols is not physically identical to using the ears to receive spoken language. Light and sound are different forms of energy. The eyes and ears contain different receptors, and the signals initially travel through different neural pathways. If that were the complete argument, there would be little to fight over. The video goes much further. It calls equivalence an enormous lie, portrays audiobook listeners as soft people seeking a “participation badge,” treats listening as passive content consumption, questions whether a person using double speed receives the work with full fidelity, and ends by ordering listeners to sit in a chair and actually read. That moves the discussion away from sensory biology and into moral judgment. My conclusion is that the video identifies real differences between modalities and raises a fair warning about divided attention, but then overstates the neuroscience, confuses medium with behaviour, and mistakes one preferred ritual of engaging with art for the only respectful one.
I am not defending the idea that every listening situation produces perfect comprehension. Someone playing an audiobook while watching Netflix, replying to a group chat, and trying to follow another conversation will probably miss things. I also would not pretend that a printed book and its narrated performance create precisely the same artistic experience. What I reject is the leap from those reasonable observations to the claim that audiobook listeners are usually taking a shortcut for their ego. That leap ignores how attention works, how differently human brains regulate stimulation, how neurodevelopmental conditions affect task initiation and sustained focus, and why a second low-demand activity can sometimes support concentration rather than destroy it. In HowlStrom terms, the video correctly notices that the eyes and ears are different gates through the castle wall, then declares that only one gate leads into the real library. The science does not support that conclusion.
What the Video Is Actually Arguing
The video begins with ridicule, but its serious argument has four layers. First, it accepts that reading and listening may recruit highly similar brain networks once the sensory signal has been converted into language and meaning. Second, it argues that controlled laboratory comparisons miss ordinary behaviour because listeners often drive, exercise, or do chores at the same time. Third, it defines visual reading as inherently active because the reader decodes symbols, chooses the pace, imagines voices, and constructs scenes, while the audiobook listener supposedly hands those choices to the narrator and becomes a spectator. Fourth, it changes from a scientific argument to an aesthetic and moral one: literature is art or “soul food,” so optimising it through multitasking or double speed allegedly reduces it to disposable content. The final command to reclaim attention by sitting still with a printed book follows from this moral definition rather than from the neuroscience presented earlier.
That is a stronger argument than simply saying that audiobooks are lazy, because it contains several truths. Reading print requires learned visual decoding. A self-paced reader can stop, reread, inspect spelling, study a difficult paragraph, and control rhythm without operating an app. A narrator supplies pronunciation, prosody, timing, accents, and character interpretation that a silent reader would otherwise create internally. Dividing attention between two demanding tasks can reduce memory and comprehension. Playback speed can eventually exceed a listener’s processing capacity. None of those points is controversial. The trouble is the repeated use of categorical words such as active, passive, art, content, reading, and respect as if each had one scientifically fixed meaning. They do not. The conclusion depends on sliding between definitions while comparing the most attentive version of print reading with an exaggerated worst-case version of audiobook listening.
How the Eyes Turn Marks Into Language
Visual reading begins as biology, not literature. Light reflected from the page is focused onto the retina. Photoreceptors convert that light into electrical activity, which passes through retinal circuits and along the optic nerve toward the brain. The United States National Eye Institute gives the basic pathway from light to retinal signal to brain. From there, visual information is relayed and processed through networks that detect edges, shapes, orientation, position, and patterns. Reading is not a smooth camera scan. The eyes alternate between rapid movements called saccades and short fixations during which useful visual information is acquired. Fluent readers recognise familiar letter combinations and words rapidly, but the system was not born specifically for novels. Literacy trains existing visual and language circuitry to cooperate.
A region of left ventral occipitotemporal cortex is commonly called the visual word form area. Research following children as they learn to read shows that word-selective responses in this area develop with literacy, rather than appearing as a complete book-reading module at birth. A longitudinal neuroimaging study on the emergence of the visual word form and related work on word selectivity and reading skill describe this learned specialisation. The visual system must identify the orthographic form, connect it with sounds and known words, access meaning, maintain information across sentences, make inferences, update a mental model of events, and connect new material with prior knowledge. That visual decoding stage is real and valuable. Audiobooks do not exercise it, so listening alone cannot teach every skill involved in literacy, including spelling, punctuation recognition, or fluent decoding of print.
This is the point where the video’s narrow distinction is strongest. If the question is, “Did this person practise decoding written symbols with the eyes?” an audiobook listener did not. If a school is measuring visual reading fluency, an audiobook cannot simply replace the test. If someone needs to learn how a name is spelled, inspect a diagram, compare footnotes, study source citations, or move repeatedly between dense technical passages, print may be better. However, those statements answer a question about the route and particular skills. They do not settle whether a listener understood the story, learned its ideas, formed images, followed character motives, felt its emotional force, or can discuss the book intelligently. The video treats the extra visual-decoding work as though cognitive effort itself proves a morally superior encounter. Difficulty can build a skill, but extra difficulty is not automatically extra understanding.
How the Ears Turn Pressure Waves Into Language
Audiobook listening also begins as biology. Sound waves move the eardrum and the small bones of the middle ear, creating waves in the fluid of the cochlea. Hair cells convert that mechanical movement into electrical signals. The auditory nerve carries those signals through brainstem and thalamic relays toward auditory cortex. The National Institute on Deafness and Other Communication Disorders explains this conversion, while an NCBI overview of the auditory pathway describes the ascending neural route in more detail. Auditory processing extracts changing frequencies, timing, intensity, phonemes, word boundaries, rhythm, and voice information from a stream that disappears as it arrives.
Listening to speech is therefore not a state in which sound simply washes through an open ear. The brain must separate speech from background noise, identify sounds despite variation between speakers, predict likely words, resolve ambiguous meanings, follow syntax, retain earlier information in working memory, make inferences, and build a situation model. Unlike print, spoken language is transient. A silent reader can leave a difficult sentence physically present while thinking. A listener must retain it, pause, or rewind. That can make listening harder for some material and some people. At the same time, narration adds information that ordinary print does not encode completely: emphasis, emotion, tempo, irony, pronunciation, breath, hesitation, and relationships expressed through vocal performance. Neither pathway is empty. They distribute effort differently.
Different Entrances, Strongly Shared Meaning Networks
The video’s account of neuroscience is close enough to sound decisive, but the phrase “exact same networks” is too absolute. Early visual and auditory processing are necessarily different, and written-word recognition adds modality-specific operations. Higher-level language and semantic processing overlap strongly. In a 2019 Journal of Neuroscience study by Fatma Deniz and colleagues, nine adults listened to and read matched natural stories while undergoing functional MRI. Semantic tuning across much of the cortex was highly correlated between modalities, and models trained using one modality predicted responses in the other. That supports a substantially modality-independent representation of meaning after the sensory route has delivered language to the wider comprehension system.
The result is important, but it has limits the video does not discuss. Nine participants are enough for intensive within-person brain mapping, not enough to represent every age, disability, language background, attention profile, or listening habit. Functional MRI also measures blood-oxygen changes as an indirect marker of neural activity. Similar activation does not mean every computation, subjective experience, or learning outcome is identical. Even the researchers distinguish early sensory pathways from the semantic system. The accurate summary is that reading and listening start differently, retain some modality-specific processes, and converge strongly when the brain represents narrative meaning. That is more interesting than either slogan. The doors differ, but much of the work performed once the story is inside the mind is shared.
What Comprehension Research Actually Finds
Brain activation alone cannot tell us what somebody remembers or understands, so behavioural research matters. A 2022 meta-analysis by Virginia Clinton-Lisell combined 46 studies involving 4,687 participants. The overall difference between reading and listening comprehension was not statistically reliable, with an effect size of g = 0.07 and p = .23. Print showed a small advantage when reading was self-paced, g = 0.13 with p = .049, which makes sense because readers can vary their speed and look back without pressing controls. That finding does not prove universal equivalence, but it directly contradicts the claim that serious comprehension belongs categorically to print.
A separate randomised experiment by Beth Rogowsky, Barbara Calhoun, and Paula Tallal assigned 91 adults to a digital audiobook, an e-text, or both together. Their immediate comprehension and two-week retention results showed no significant modality effect for the nonfiction material used. This is one study with one text and a particular adult sample, so it should not become another slogan. Dense mathematics, a complicated reference work, poetry whose typography matters, and a character-driven fantasy novel are not interchangeable tasks. Still, when a critic says laboratory work only finds similarity because scientists ignore the individual listener, that criticism overlooks how research actually handles comprehension aptitude, age, self-pacing, text type, prior knowledge, and other moderators. Science has not forgotten the person. It simply separates variables so claims can be tested.
The honest conclusion is conditional. Print has unique advantages for orthography, navigation, annotation, visual structure, and effortless rereading. Audio has unique advantages for pronunciation, prosody, performance, accessibility, and hands-free continuity. For many ordinary narrative and expository texts, average comprehension can be similar. Individual differences may be larger than the average difference between media. A modality cannot guarantee understanding because comprehension is an interaction among the person, the material, the purpose, the environment, and the way attention is managed.
Reading Is Not Automatically Active, and Listening Is Not Automatically Passive
The video calls visual reading categorically active and content consumption categorically passive. Psychology does not support that clean division. A person’s eyes can continue crossing lines while the mind has left the book. Researchers call this mindless reading. In an eye-tracking study, Erik Reichle and colleagues found that eye movements continued during mind wandering, although their pattern changed. A later meta-analysis and broader literature estimate that mind wandering occurs regularly during reading and harms comprehension. Holding a printed book proves that visual decoding is happening at some level. It does not prove deep attention, emotional respect, or successful construction of meaning.
Listening can also drift. A study comparing ways of encountering the same material found more mind wandering in its listening condition than in reading silently or aloud, and reading aloud produced the least mind wandering. That is fair evidence for a possible auditory disadvantage in some settings. It still does not make listening inherently passive. It shows that attention depends partly on task structure. Topic interest, motivation, working-memory capacity, fatigue, and personal concerns all affect mind wandering during print reading too. Research by Unsworth and McMillan found that interest, motivation, working memory, prior experience, and off-task thought jointly help explain reading comprehension. The book does not cast a spell of attention merely because it is made of paper.
An attentive audiobook listener predicts, interprets, imagines, evaluates, remembers, and notices contradictions. A passive print reader can decode every word and retain almost nothing. An attentive film viewer can analyse framing, performance, symbolism, and narrative structure. A distracted museum visitor can stand in front of a painting while thinking about dinner. Active and passive describe modes of engagement, not permanent moral properties attached to media. The video’s strongest mistake is treating a behavioural variable as a format label.
The Netflix Example Is a Real Problem but a Bad Comparison
The video asks us to imagine an audiobook playing while the listener watches Netflix, does laundry, and answers a group chat. That person is not merely multitasking. They are combining multiple streams that compete for language processing, working memory, visual attention, decision-making, and response selection. Comprehension probably will suffer. The same would happen if a print reader stopped every sentence to answer messages while television dialogue competed in the room. This example proves that several demanding tasks can interfere with one another. It does not isolate the effect of audiobook format.
All secondary activities are not cognitively equivalent. Folding familiar clothes, washing dishes, walking a known route, navigating dense city traffic, writing a message, and watching a drama place radically different demands on the brain. Two tasks interfere most when their combined needs exceed available resources or require overlapping operations at the same time. A routine motor activity may need little conscious planning once practised. A group chat requires reading language, deciding what it means, composing a response, and switching goals. Netflix adds another narrative with speech and images. Calling all of this “doing something else” hides the variable that matters.
Driving deserves special caution because safety outranks book completion. A 2025 high-fidelity simulator study of dual-task costs while listening and driving found that both younger and older adults showed more variable lane position during the most demanding city conditions when listening was added. Older adults also lost more listening accuracy under difficult conditions. The study used connected speech in noise rather than a novel, and its authors did not claim that the measured changes automatically meant unsafe driving. It nevertheless shows why a listener should pause during heavy traffic, unfamiliar roads, bad weather, or any moment when the driving task becomes complex. It also shows why an easy rural drive cannot be treated as cognitively identical to a difficult city drive.
For Some Minds, a Second Activity Is an Anchor
There is another possibility missing from the video. For some people, especially those with restless or under-stimulated attention, a low-demand physical activity can help maintain the level of arousal needed to stay with spoken language. The hands are occupied, the body receives movement, and the narrative gives the mind a stable target. That is not the same as claiming that humans possess two unlimited attention systems or that every kind of multitasking helps. It means the effect of a secondary activity can depend on whether it competes with the audiobook or regulates the listener.
This distinction fits my own experience. My mind is rarely quiet. If I force myself into a chair with nothing except a printed book, part of my attention may fight the text, search for stimulation, or create its own internal noise. When I listen while doing a familiar practical task, the physical activity can occupy the restless layer while the story holds the language-processing layer. The result can be more focus, not less. The correct test is not whether my body remained still. It is whether I followed the characters, understood the world, retained earlier events, noticed the writing, and could form an independent judgment afterward.
ADHD Is an Attention-Regulation Problem, Not a Refusal to Try
The video’s command to stop being soft and force the brain to focus treats attention mainly as discipline. Discipline matters, but ADHD is a neurodevelopmental condition, not a personality flaw. The United States National Institute of Mental Health describes ADHD through persistent patterns of inattention, hyperactivity, and impulsivity that affect functioning. Adults may struggle with task initiation, organisation, time management, sustained focus, distractibility, restlessness, and finishing large tasks. ADHD does not mean a person is incapable of attention. It often means attention is difficult to allocate and regulate consistently, especially when a task is slow, repetitive, weakly rewarding, or requires effort without immediate feedback.
Research on mind wandering adds useful detail. A review of ADHD from a mind-wandering perspective discusses difficulties regulating internally generated thought and possible interference from the brain’s default mode network during tasks that require external attention. A later study on context regulation of mind wandering in adults with ADHD found that sustained-attention demands were especially important. This does not mean every wandering thought comes from one network, and brain-network language should not be turned into a simple dopamine meme. It does mean that “just concentrate” can describe the goal without explaining why reaching it costs one person far more effort than another.
Movement may also have different effects in ADHD. In research with children, greater physical activity during demanding working-memory or cognitive-control tasks has sometimes been associated with better performance. A study asking whether ADHD-related hyperactivity can be compensatory rather than purely impairing and a separate trial-by-trial study of movement and cognitive control support that possibility under specific conditions. These are mainly child studies and cannot prove that washing dishes improves every adult’s audiobook comprehension. They do, however, make the video’s still-chair prescription scientifically unjustified as a universal rule. Stillness is not a direct measurement of attention.
For me, beginning a printed book often requires force before I reach the first chapter. Even when I want the story, finishing can take about a month. That pace is too slow for how I naturally live with stories, and the effort needed to restart after each break becomes another frozen wall. Audiobooks remove the visual-decoding and sitting-still demands while preserving the language, characters, conflicts, and ideas. I can begin listening while walking, travelling, cooking, cleaning, or handling a routine task. My body remains occupied and my mind has less empty space in which to run away from the story. Calling this a shortcut misses the destination. A bridge is not cheating because another traveller prefers to climb down into the gorge.
This does not mean audiobooks are automatically better for everyone with ADHD. Some people lose spoken words quickly, find narration too slow, become distracted by the environment, or need the visible page to return after a lapse. Some benefit from reading and listening together. Others need silence. ADHD is heterogeneous, and the same person may need different strategies for fiction, study material, or a difficult unfamiliar subject. The reasonable standard is self-observation: pause and rewind when attention breaks, reduce competing language, choose a playback speed that sustains understanding, and use the format that produces real engagement rather than the format that looks virtuous from outside.
Autism Does Not Point Toward One Correct Medium
Autism makes the universal claim even less defensible because autistic reading and listening profiles vary enormously. Some autistic people are highly drawn to written language and become unusually strong decoders. Hyperlexia is associated with early or advanced word recognition, but strong decoding can coexist with weaker comprehension. A systematic review of hyperlexia and a meta-analysis of reading comprehension across the autism spectrum show why sounding out words, reading quickly, and building meaning must not be treated as one ability.
Auditory experience is equally varied. Reviews of auditory sensory differences in autism describe both heightened and reduced sensitivity, along with wide variation in how sound is processed and tolerated. A narrator’s voice may create a clear, predictable stream for one autistic listener and become exhausting for another. Written text can help because it is stable, silent, controllable, and available for repeated inspection. Audio can help because it supplies prosody, reduces decoding demands, and allows the eyes or hands to avoid overload. Some people find reading while listening useful because the two channels support each other. There is no scientifically honest sentence beginning with “autistic people learn best by” that ends in one universal format.
My friend Shilka demonstrates the opposite preference from mine without needing to represent every autistic or neurodivergent person. He reads books quickly and chooses English books because he does not like Danish as much and notices too many translation problems. For him, original English text gives direct control over wording and lets him move at his own fast pace. I can respect that without pretending his brain should be my instruction manual. He may cross a written forest quickly while I become stuck between the first trees. I can travel farther through audio. The presence of two very different readers in the same friendship should make us cautious about declaring one route the proof of strength and the other a symptom of softness.
Language, Translation, and Why the Original Format Matters Differently
Shilka’s preference also exposes another variable the video ignores: the language and edition. A poor translation can flatten humour, distort idioms, create awkward sentences, or move the reader farther from the author’s original style. An English original may therefore be more satisfying than a Danish translation for a reader who is comfortable in both languages. An audiobook introduces yet another interpretive layer because the narrator decides pronunciation, emphasis, pacing, and character sound. Those changes can improve a work, reveal emotion, or occasionally pull the listener away from how the prose might sound internally.
None of this creates a universal ranking. A silent reader using a translation is not automatically closer to the author than a listener hearing the original language. A listener using an unabridged English audiobook may receive the author’s actual words while also receiving the narrator’s performance. A reader sees spelling and typography that a listener cannot hear. A listener hears vocal information the printed page can only suggest. The right comparison must identify the exact editions, purpose, and person rather than treating “book” and “audiobook” as two featureless blocks.
Double Speed Is a Processing Choice, Not a Moral Confession
The video treats double-speed listening as evidence that the person is optimising away the soul of the work. Speed can affect comprehension and emotional rhythm, but 2x on a player is not a direct measure of mental haste. Narrators have different natural rates, pauses, and delivery styles. Listeners also adapt with practice. Marc Brysbaert’s review and meta-analysis of 190 reading-rate studies, covering more than 18,000 participants, estimated average silent English reading at about 238 words per minute for nonfiction and 260 for fiction, with broad individual ranges. Silent readers already move through language at very different speeds without being accused of disrespecting art.
Research on accelerated recorded material is mixed and highly dependent on rate, complexity, age, familiarity, and the outcome being tested. A 2024 series of experiments on playback speed, modality, and distraction found that university students could preserve substantial immediate learning at speeds up to 2x, with no statistically significant speed effect in its audio-only experiment even at 2.5x, although scores numerically declined and the authors advised caution above 2x. Those were short educational clips, not forty-hour fantasy novels or emotionally paced performances. Older work on time-compressed speech found greater losses as compression, age, and speech complexity increased. The correct conclusion is not that speed never matters or that 2x destroys meaning. It is that the threshold differs across people and material.
I normally listen at around 2x because slower narration often leaves enough unused mental space for my attention to drift. Faster delivery can keep the stream dense enough to hold me. I still slow down, rewind, or stop when a scene, narrator, accent, system explanation, or emotional moment requires it. After years of listening to roughly 140 audiobooks annually, I have experience with the format and with the genres I choose. That does not grant me superhuman comprehension, but it matters. Expertise and familiarity affect how efficiently new information connects with existing knowledge. Judging my attention from the number on the speed control is like judging how deeply Shilka read a book from how quickly he turned the pages.
Does the Narrator Turn the Listener Into a Spectator?
The video is right that narration transfers some artistic control. A silent reader chooses pace moment by moment, imagines every voice, and supplies internal prosody. An audiobook narrator or cast makes those choices audible. A performance can establish character identity before the text names the speaker, turn a dry joke into a successful one, make grief feel immediate, or impose a voice the listener dislikes. That difference is large enough that I review story and performance separately. A brilliant narrator can strengthen flawed prose, while an inconsistent performance can damage a strong novel.
What does not follow is that the listener becomes a “mere member of the audience.” Every art form has interpreters and audiences. Theatre does not stop being Shakespeare because actors choose the voices. Music does not become mindless because performers choose tempo. An audiobook listener still constructs places that are never visually shown, tracks motives, predicts consequences, detects foreshadowing, compares new facts with earlier chapters, judges prose, and responds emotionally. The narrator directs the vocal performance, not the listener’s entire imagination. The app also permits pausing, rewinding, bookmarking, changing speed, and replaying a chapter. Audio gives up some control and gains a performed layer. That is a trade, not a downgrade from participant to empty seat.
The video also describes audiobooks as films for the ears with different voice actors and extensive sound design. That fits full-cast productions and audio dramas, but not the whole medium. Many unabridged audiobooks use one narrator, no music, and no sound effects beyond introductory or closing material. Peter Bradshaw’s Mercy, the example mentioned in the video, really was released as an audio-first Audible production with a cast and Dolby Atmos. However, Bradshaw had already published traditional novels, so this one audio-only release does not mean his entire career is exclusively audio. The fact is interesting because it proves audio can be the intended original art form, not merely a substitute for print.
The 37.5 Percent Statistic Is Real but Misdescribed
The video’s National Literacy Trust figure is based on a real report, but calling 37.5% a conversion rate says more than the study can establish. In 2024, the Trust asked more than 37,000 people aged 8-18 who listened to audiobooks and podcasts in their free time about their attitudes. Its published findings say that 37.5% reported that listening to audio had sparked their interest in reading books. The category included audiobooks and podcasts, the result was self-reported, and the design does not prove that audio alone caused a later behaviour. A true conversion rate would require a defined starting group, a defined outcome, and evidence that participants actually changed behaviour across time.
Even with those limitations, the figure matters. The same report found that 48.4% said audio helped them understand a story or subject, 52.9% said it made them use their imagination more than video, and 52% said it helped them relax or feel better during stress or anxiety. These are perceptions, not laboratory measurements, but they contradict the claim that audio is generally a lesser stream washing over passive minds. The video gives the statistic genuine credit as a gateway, then fails to let that evidence affect its final moral command. If a format brings a young person from avoidance into stories, supports wellbeing, and sometimes leads toward print, telling that person the gateway does not count can damage the very engagement literacy organisations are trying to build.
The Gender Statistic Does Not Explain Motivation
The claim that 33% of men and 24% of women listen to audiobooks weekly appears in the January 2026 Guardian article that also contains the National Literacy Trust statistic and the Peter Bradshaw example. The matching sequence of facts makes that article the likely factual spine behind the video. The percentage can describe a reported gender difference in one survey population. It cannot explain why the difference exists.
The video’s suggestion that men simply like optimisation and life hacks is a cultural guess. To establish it, researchers would need to measure motivations, control for age, work patterns, commuting, genre, platform use, disability, education, and other factors, then determine whether optimisation explained the difference. Men may be attracted by efficiency, but the percentage alone cannot show that. It also cannot tell us whether those listeners are deeply engaged, distracted, using audio for accessibility, or replacing other media. A demographic gap is a question requiring explanation, not an answer wearing a number.
Art Does Not Become Content When It Enters Through the Ears
The video draws a moral border among art, data, and content. It is correct that literature deserves more than mindless scrolling, and I agree that the attention economy encourages fragmentation. Notifications, short-form feeds, autoplay, and constant task switching can make sustained engagement harder. Print can be a valuable refuge because it can remove screens and lock attention onto one stable object. Anyone who enjoys that ritual should protect it.
The mistake is treating paper as the ritual’s only possible home. Spoken storytelling existed long before mass literacy. Poetry, myth, history, and communal memory have survived through voices and attentive ears. A performance of Anna Karenina, Justine, or Don Quixote does not cease to be art because the audience cannot see ink. Respect is expressed through the quality of attention, thought, feeling, and later reflection, not through which sensory organ received the words. A person can skim a classic for status and remember almost nothing. Another can listen, rewind difficult scenes, think about the characters for days, and discuss the work with care. The first used visual reading. The second may have had the deeper artistic encounter.
Audio can even resist the attention economy. A ten, twenty, or fifty-hour audiobook asks the listener to stay with one world across days or weeks. For someone who would otherwise scroll during a commute or repetitive chore, an audiobook can replace fragmented novelty with sustained narrative. It can also turn necessary work into the only reliable space where a restless mind can remain inside a long story. That is not automatically optimisation against the soul. Sometimes it is the structure that lets the soul reach the work at all.
The Participation-Badge Argument Tries to Read Minds
Some people undoubtedly chase totals, social-media status, reading challenges, or the appearance of being well read. That can happen with audiobooks. It can also happen through skimming, choosing very short print books, abandoning comprehension for speed, copying opinions, or displaying untouched hardbacks. The medium does not reveal the motive. Saying that a listener probably wants a badge turns an imaginable behaviour into an accusation about strangers.
This framing also creates a psychological trap. If the listener defends the value of audio, that defence is treated as proof of insecurity. If the listener accepts a lower-status label, the gatekeeper wins. The discussion stops being about comprehension and becomes a test of whether the listener will submit to someone else’s cultural hierarchy. Words such as soft, shortcut, ego, and disservice are identity attacks wrapped around a definitional dispute. They invite shame and reactance rather than honest examination.
I do track the books I finish, but a number cannot replace the experience. I buy audiobooks, listen closely enough to evaluate story and narration separately, write detailed reviews, return to series after long gaps, and compare systems, relationships, writing, voices, and production across hundreds of titles. I usually say that I listened when the format matters because that is precise. That precision is not an admission that I failed to encounter the book. My ears did not decode ink, but my mind still had to meet the characters, carry the world, remember the rules, and decide what the work deserved.
What “Reading” Should Mean Depends on the Question
The entire fight becomes clearer when we stop demanding one definition for every purpose.
- If the question concerns visual decoding, spelling, punctuation, eye movement, or print fluency, listening is not the same activity.
- If the question concerns understanding the language and ideas of an unabridged book, listening and reading often produce similar outcomes, although the person and context matter.
- If the question concerns the artistic experience, the formats differ because narration adds performance while print gives the reader more direct control over pace and imagined voice.
- If the question concerns accessibility and sustained engagement, the better medium is the one that lets the individual reach and retain the work.
- If the question concerns a casual statement such as “I read that book,” audiobook use is widely understood shorthand, while “I listened to the audiobook” is more precise when format matters.
Science can compare brain activity, comprehension, retention, attention, and skill development. It cannot decree the only socially acceptable use of the verb “read.” Language communities decide how words broaden, and context resolves most ambiguity. A person discussing typography should specify print. A reviewer discussing narration should specify audio. A friend asking whether someone knows the story usually cares about whether the book was experienced, not which sensory pathway delivered it. Turning every casual sentence into a purity test creates heat without improving understanding.
A Better Challenge Than “Sit in a Damn Chair”
The video ends by asking listeners to reclaim their attention. That goal is worth keeping after removing the shame and the single-format command. A better challenge would be to learn which conditions produce genuine attention for you. Try print without notifications. Try audio while sitting still. Try audio during a familiar walk or chore. Notice when you rewind, what you remember the next day, whether you can explain motives and arguments, and whether speed improves focus or merely shortens exposure. Pause during demanding traffic. Do not combine a book with another stream of language and then blame the format when both dissolve.
For ADHD and autistic people, the experiment must remain individual. Movement may anchor one mind and overload another. Narration may organise language for one listener and create sensory stress for another. Print may offer calm control, or it may turn initiation into a wall so high that the story never begins. Reading and listening together may help some people connect spelling, sound, and meaning. No result makes the person weak. It identifies the conditions under which that nervous system can engage.
Final Judgment: Different Medium, Real Reading Experience
The video earns credit for refusing a lazy slogan. Eye-reading and audiobook listening are not neurologically identical from entrance to exit. Print uniquely exercises orthographic decoding and provides effortless self-pacing, spatial navigation, annotation, and rereading. Audiobooks uniquely provide performed speech, pronunciation, prosody, character voices, and access while the eyes and hands are engaged elsewhere. Multitasking can reduce comprehension, especially when the secondary task competes for language, working memory, or safety-critical attention. Speed can exceed a person’s capacity. Those warnings are factual.
The larger verdict does not survive scrutiny. Higher-level semantic processing overlaps strongly. Across dozens of studies, average comprehension is often similar, with a small print advantage under some self-paced conditions rather than a categorical gulf. Print readers mind-wander. Audio listeners think actively. Low-demand movement can sometimes support attention, particularly for minds that struggle with stillness and sustained under-stimulation. ADHD and autism increase the need for individual fit, not obedience to one ritual. The children statistic is real but not a causal conversion rate. The gender statistic does not prove an optimisation motive. Double speed does not reveal disrespect. A narrator changes the art but does not erase the listener’s imagination.
Shilka and I can enter the same library by different gates. He can race through English pages and avoid Danish translations that weaken his experience. I can spend a month fighting one printed book, or I can use audio to keep my always-running mind occupied, focused, and moving through stories throughout the year. Neither route makes the other fraudulent. I will keep saying “I listened” when the format matters, but I will not accept that my hundreds of completed audiobooks are empty background noise or a badge pinned to borrowed fur. In the frozen halls of HowlStrom, I judge the tracks by where they lead and what the traveller carried home, not by whether they were made with paws that somebody else approves of.
Selected Research and Sources
- The YouTube video discussed in this article
- Deniz et al. (2019), semantic representations during listening and reading
- Clinton-Lisell (2022), meta-analysis of reading and listening comprehension
- Rogowsky, Calhoun, and Tallal (2016), reading, listening, and dual modality
- National Literacy Trust, Children and Young People’s Listening in 2024
- Reichle et al. (2010), eye movements during mindless reading
- National Institute of Mental Health, ADHD overview
- Bozhilova et al. (2018), ADHD and mind wandering
- Sarver et al. (2015), movement as possible compensatory behaviour in ADHD
- Ostrolenk et al. (2017), hyperlexia and autism
- Goncalves and Monteiro (2023), auditory sensory differences in autism
- Chen et al. (2024), playback speed, modality, distraction, and comprehension
- Bak et al. (2025), dual-task costs of listening while driving






Leave a Reply