From “Couch Potatoes” to Reading Kant: Why Short Video and AI Are Not the Enemies of Deep Thought, but the Key to the Door
Responding to warnings in The New York Times and The Economist about AI summaries and 'cognitive shallowing': from television-era 'couch potato' panics to how YouTube, TikTok, and AI gave me the scaffolding to read Kant's Critique of Pure Reason—why staring at raw text does not equal understanding, and what really determines whether we keep questioning after opening the door.
Conversational walkthrough of the core argument
Recently, a wave of essays in The New York Times and The Economist has sounded the alarm over artificial intelligence and the decline of reading. Neuroscientists and humanities scholars warn that as professionals and students grow accustomed to using AI to compress thick books into three-minute bullet points, the neural circuits responsible for complex empathy and long-chain reasoning are beginning to atrophy. In response, several Western universities have launched a counter-cultural classroom experiment: confiscating all digital screens and requiring students to spend an entire semester grinding through eight-hundred-page classics like The Brothers Karamazov word by word without assistance.
At first glance, this appeal sounds noble and beyond reproach—in an impatient era, who could possibly argue against deep reading? Yet the longer I sit with this conclusion, the less it holds up. Beneath its moral earnestness lies an old intellectual illusion that has repeated itself for more than half a century: mistaking the physical length and medium of a text for the capacity to understand a complex world.
We have watched this exact movie several times before. When television sets first entered family living rooms, elders and cultural critics warned that an entire generation would turn into "couch potatoes" who would never read or think again. Decades later, when video platforms like YouTube and TikTok emerged, public commentary replayed the exact same script, declaring that human attention had shattered beyond repair and no one would ever grasp serious ideas again. Today, now that large language models can distill the backbone of a book in three minutes, the same moral panic has simply found a new target.
Yet the real world never unfolded the way those prophecies claimed. If we examine honestly how human beings actually come to understand complex things, this popular narrative that "new tools make us shallow" is wrong on two fundamental premises—and when people in real life do stop after memorizing two superficial slogans, artificial intelligence is the last thing we should blame.
I. Piercing the First Myth: The Capacity to Understand a Complex World Never Depended on the Length of the Medium
The first flaw in the "cognitive shallowing" thesis is its assumption that television, short-form video, or AI summaries inherently strip people of the ability to grasp complexity. That judgment confuses the container of information with the structural depth of the thought inside it.
A five-hundred-page printed book can be repetitive, muddled, and intellectually hollow; conversely, a great film on a television screen, a finely crafted drama series, or a serious documentary can lay bare the ambiguities of human nature, institutional dilemmas, and historical causality with extraordinary depth. In the same way, countless creators on YouTube and TikTok today use incisive visual storytelling to unpack intricate scientific, historical, and social problems. No one can credibly claim that an idea loses its depth simply because it travels through moving images rather than ink on paper.
Even more importantly, we need to acknowledge a truth that traditional critics rarely admit: when a short video or an AI synthesis succeeds in explaining a complex story or underlying mechanism clearly within a few minutes, its contribution to understanding a complex world can actually be clearer and sharper than a meandering long video or a bloated long-form text.
The reason is straightforward. To explain a tangled, multi-layered problem clearly in three minutes, a creator or model must strip away decorative filler and circular throat-clearing, pulling the central tension and the load-bearing causal chain directly into view. From the standpoint of information transmission, during those few minutes of focused synthesis, the information density and cognitive intensity reaching our minds do not drop—in a very real sense, they increase.
Clarity has never been the enemy of depth. For centuries, people often equated "spending thirty painful hours getting through an opaque book" with profundity because they confused the linguistic friction of the medium with the depth of the idea. When a high-density video or an AI synthesis helps us see the load-bearing blueprint of a complex edifice in five minutes, what it eliminates is the time wasted getting lost in the outer corridors—freeing our mental energy for the actual work of weighing the core problem.
II. Piercing the Second Myth: The Illusion That “Staring at Text Equals Understanding,” and How Video and AI Helped Me Read Kant
The second error in the conventional critique is even more pervasive, though rarely stated out loud: people talk as if simply sitting down and staring at a long text means you have understood it—and as if forcing yourself to read long texts automatically builds the capacity to understand a complex world.
In reality, staring at raw text very often means understanding nothing at all. And if you cannot even understand what the text is saying, how could grinding through long passages possibly build your capacity to understand a complex world?
Let me use myself as a concrete example. Over the past few years, I have spent a great deal of time reading Immanuel Kant’s Critique of Pure Reason and using his three-tier architecture of Sensibility, Understanding, and Reason to examine the epistemological boundaries of the AI era. Yet I know one thing about myself with complete certainty: without video platforms—without YouTube, without TikTok, and without AI—if I had tried to pick up the Critique of Pure Reason cold, with no prior contact with Kant, I could never have understood or finished a single edition of that book.
Why? Because Kant wrote in the dense, eighteenth-century academic apparatus of German idealism. When a reader without years of formal training in a philosophy department opens page one of the original text, every individual word may be familiar, yet sentences packed with "transcendental aesthetic" and "the synthetic unity of apperception" form an impenetrable stone wall. Locking your phone in a drawer and sitting in agony before those eight hundred pages for ten hours will not magically wire "deep reading circuits" into your brain—it will simply leave you lost and defeated on page fifteen.
As I wrote in an earlier essay on Kant’s epistemology, genuine human comprehension requires climbing step by step from Sensibility (vivid intuition) to Understanding (conceptual connection). When an intelligent reader approaches an unfamiliar masterpiece cold, what blocks them is the absence of an entry ramp across the wall of specialized jargon: they first need a sensory, intuitive grasp of what problem Kant was actually fighting over, and then a conceptual translation that maps his eighteenth-century terminology onto modern logic.
That is precisely what video platforms and artificial intelligence built for me. Before I ever wrestled with Kant’s raw text, creators on YouTube and short-video platforms gave me the intuitive entry point—using vivid everyday metaphors and visual diagrams to show why Kant was trying to rescue both scientific law and human freedom. Then, once I opened the book myself and hit passages where the original syntax refused to yield, AI served as an unwearying conceptual translator, unpacking two-hundred-year-old arguments and connecting them to legal and analytical structures I already knew.
It was precisely because of video platforms and AI that I gained the capacity to enter, read, and genuinely understand that complex text in the first place.
Once we look at this lived process honestly, the elitist nostalgia for "unassisted raw-text reading" loses its halo. Written text is a static storage medium, not a magical altar that confers wisdom on anyone who suffers before it. Dropping a traveler without a map into an eight-hundred-page fog does not cultivate intellectual stamina; turning on the searchlight of high-density video and AI translation is what finally makes the classic accessible to a wider humanity.
III. The Deeper Question: Once Handed the “Key to the Door,” Why Do Some People Stop After Two Slogans?
At this point, a thoughtful reader—or an educator worried about the next generation—will naturally raise a deeper question: "Even if video and AI gave you the bridge to read Kant, what about the broader reality around us? Now that AI can summarize any book in three minutes, don't most people simply memorize two catchy slogans and stop right there? Isn't that worth worrying about?"
That question deserves to be explored seriously. In everyday life, plenty of people really do watch a three-minute video or read an AI summary, pick up two polished phrases, and close the page assuming they have mastered the subject. Yet if we examine why people stop at the doorway, we must not blame AI—because blaming the tool misjudges both the historical baseline and the true human cause. Unpacking this requires separating three layers:
1. The Wrong Baseline: In the Past, People Were Not “Reading Deeply”—They Couldn’t Even Get Through the Door
When critics lament that "people today only remember a three-minute summary of Kant or Dostoevsky," they are measuring the present against an imaginary past in which ordinary people spent their evenings reading eight-hundred-page philosophical treatises from cover to cover.
That past never existed. Before video platforms and AI, for works as demanding as Kant’s, the issue for ninety-nine percent of humanity was not whether they read deeply or shallowly—it was that they could not get through the door at all. Back then, the vast majority of people either walked past those towering books in complete intimidation, or—when forced by a school curriculum—memorized two dry textbook sentences just to survive an exam.
What video platforms and AI have done is hand every ordinary person the key to the door for the very first time. Even if ten thousand people take that key, unlock the door, spend three minutes looking at the architectural layout of the room, and nine thousand of them walk away remembering only the basic outline, that clear mental sketch is already a vast leap forward from being locked outside the wall. And even more importantly, the remaining one thousand people—who never would have opened Kant in the old world—catch sight of the landscape inside those three minutes, feel a spark of genuine fascination, and walk into the depths.
2. This Is a Human Problem, Not an AI Problem: The Old “Standard-Answer Habit” vs. Having a Real Question
Once the key is in everyone’s hand, why do some people stop at two slogans while others keep walking deeper? Whether a person keeps questioning into the depths is a personal issue, not a problem caused by the age of AI. If we ask what is truly to blame when someone stops at two slogans, the culprit is our long-conditioned "exam and standard-answer habit," paired with the absence of a real personal question.
Decades of standardized testing and bureaucratic evaluation have trained many people to treat reading as an exercise in "finding a standard answer to turn in" or "collecting two impressive terms to repeat at a dinner table." For someone operating under that habit, even in the era of paper books, they never read deeply—they skimmed the table of contents, read book reviews, or lined their shelves with unread hardcovers for display. AI did not make them shallow; a three-minute summary simply gave them the two slogans they were looking for more efficiently, allowing them to check the box and move on.
Conversely, what drives someone—as happened in my own encounter with Kant—to keep pushing past the three-minute video and the initial AI summary into the original text? It is never external discipline; it is carrying a real, unresolved question of one’s own.
When I turned to Kant, I was not looking for two decorative quotes. I was wrestling with questions that pressed on my own work and thinking every day: Where does statistical machine inference end and human judgment begin? Why do equally intelligent people across the Pacific look at the exact same world and arrive at opposite conclusions? When you carry a genuine perplexity inside you, two slogans from a summary can never satisfy you. Instead, that clear three-minute overview acts like a match struck in a dark room—immediately prompting you to ask the next question: How did Kant actually prove this step? If we test his premise against modern large models, where does it hold and where does it break?
The dividing line between stopping at the door and walking inside has never been the medium in your hand; it is whether you are shopping for pre-packaged conclusions or trying to solve a real question in your own life.
3. Confiscating Screens to Force Raw-Text Drudgery Is Essentially “Confiscating the Key”
Seen in this light, the movement in some universities to confiscate digital tools and force unassisted raw-text reading is a deeply flawed prescription. Because professors see some students take the key, snap a quick picture at the doorway, and walk away, they decide to confiscate everyone’s key and padlock the heavy iron door once again.
Wrestling with moral ambiguity and complex human motives inside a great work is productive cognitive friction; staring helplessly at archaic jargon without a map is unproductive friction—the friction of a locked gate. Confiscating video and AI does not turn passive test-takers into deep thinkers; it merely demolishes the entry ramp and locks ordinary learners back outside the wall.
IV. Stepping Back Further: The Exaggeration of “Genetic Decline,” Old-World Distractions, and the “Fitness Center” of the Mind
Of course, some critics push this anxiety even further onto biological ground. They argue that because children today grow up immersed in short-form videos and AI summaries, the younger generation will not only lack mental training, but may even be altered or degraded "genetically" by this new media environment. Frankly, I think this claim is wildly over-exaggerated.
First, the notion that a media format could "genetically change" human beings vastly underestimates human biology and the timescale of evolution. Genetics does not rewrite a human being in twenty or thirty years; biological evolution operates across tens of thousands of years. In fact, mass literacy and printed books have existed for only two or three centuries—a fraction of a blink in evolutionary time. Our underlying biological hardware is far more resilient than a few decades of screens.
Second, when we talk about the pull of short-form videos or AI summaries, we cannot pretend that in the "old times" no other addictive attractions existed. It is a romantic illusion to imagine that before smartphones, everyone spent their leisure hours exclusively reading dense philosophical classics. Every era in history had its own powerful attractions and addictions competing for human attention—from taverns, gambling houses, and street theater to serialized pulp fiction, radio dramas, and video arcades. The temptation of easy distraction has always been part of the human condition; short video is simply today’s version, not a sudden mutation of our species.
If there is one small piece of ground I would grant to the critics, it is this: just as modern life gave rise to fitness centers for the body, we should use our abundant resources to deliberately train our minds.
Think about what happened when automobiles, elevators, and modern appliances replaced the daily physical labor of walking miles and hauling heavy loads by hand. Human genes did not collapse, and society did not smash its cars and elevators to force people back into manual drudgery. Instead, we recognized that because everyday convenience required less physical exertion by default, we needed dedicated spaces to train our muscles—so we built fitness centers, backed by better nutrition and training science than any previous era possessed.
The exact same logic applies to the mind. Even if convenient tools mean daily information gathering requires less raw drudgery, we today have far richer resources to train our mental muscles than any generation before us. That is the one room I would fairly grant to the critics: we should build and use "fitness centers" for our minds. Other than that, the apocalyptic narrative that new tools are destroying human thought is completely false.
V. Closing Note: Don’t Confiscate the Key—Build “Fitness Centers” for the Mind and Keep Questioning
Looking back across the evolution from television to the internet, short video, and generative AI, every generation of tools has lowered the barrier blocking ordinary people from humanity’s most complex ideas. There is no reason to feel a shred of intellectual guilt for using instruments that deliver higher clarity and density in less time.
In an age when AI can summarize any book in three minutes, the wisest posture is neither nostalgic hostility toward new media nor lazy satisfaction with two machine-generated slogans. It is to gladly accept the key that video and AI place in our hands, and treat the world beyond that open door as a fitness center for training our minds.
In everyday practice, that comes down to two simple disciplines. First, whenever we stand before an intimidating frontier—whether Kant’s philosophy, a new technical architecture, or an unfamiliar institutional system—we should use short videos, long-form lectures, and AI syntheses without hesitation to build a high-density map of the terrain, turning an unscalable wall into an open doorway.
Second, the moment a three-minute synthesis gives us that first flash of clarity, we should treat it not as a full stop that concludes our thinking, but as the starting line of genuine inquiry. Bring your own lived experience and real perplexities to the table, and ask the second and third questions: Why did the author reach this conclusion? What tension did the summary smooth over? And does this framework still hold up against the reality I have walked with my own two feet?
The capacity to understand a complex world was never a medal awarded for enduring opaque pages. It belongs to those who use the best tools of their time to unlock the door—and who keep the curiosity and courage to step into the mental gym and keep asking questions all the way into the depths.
Discussion
CONTINUE EXPLORING
Related Essays in This Domain
Isomorphism and Agnosticism: Why We Keep Looking for a Human Mind Inside the Machine
Perhaps we suspect AI is conscious simply because it exhibits a structural isomorphism with conscious human beings. Across seven distinct levels—from output and behavior to causality and subjective experience—isomorphism is both our sharpest tool for approaching the unknowable and our deepest source of misrecognition.
How to Put Legal Work in the Loop: Building the Verification Harness for Legal AI
If coding transformed faster than law because software has compilers and unit tests, how do we build a verification loop and test harness for legal work without asking busy lawyers to grade AI homework?
Never Let Disliking a Leader Sabotage Your Own Work: Piercing the Zero-Sum Myth in Global Business
From an enterprise perspective, we must never view public officials through personal likes and dislikes—and we must outgrow both the two-way risk misjudgments across the Pacific and the extreme zero-sum myth that doing business in the other country merely strengthens a rival.
If you liked this:
My newsletter has more "signal → action" content.
Leave your email, and I'll send you new signals first.