Teaching the New Teacher of the Town
Updated: Aug 30
We have spent decades treating the internet as a place where humanity goes to find answers.
We search it when we are curious, frightened, confused, lonely, ambitious, heartbroken, or simply trying to understand something we have never encountered before. Somewhere along the way, however, the relationship quietly changed. The internet stopped being merely a library of human knowledge and became something closer to a collective intelligence; an ever-expanding teacher that we consult for almost everything. And now, increasingly, artificial intelligence is learning from that same teacher. This creates a strange inversion of the relationship we have with information; if the internet has become one of the world's greatest teachers, who is teaching the teacher?
The answer is us.
Every question we ask, every perspective we articulate, every original connection we make becomes a small contribution to the informational environment from which future intelligence may learn. Which makes me wonder whether the most consequential way to influence AI is not necessarily to build another audience, chase another algorithm, or become louder online, but simply to introduce better questions, deeper perspectives, and genuinely human thought into the system.
I often find myself thinking about the deep, invisible rivers of information that shape our digital world, particularly the massive data pipelines that feed modern artificial intelligence. It is easy to assume that these models exist in a vacuum, but the truth is they are heavily nourished by public internet data, with many prominent open-source variants relying on vast web archives like Common Crawl alongside digitized books, public code repositories, and online forums to construct their foundational baseline.
As the boundaries between human expression and machine digestion continue to blur, we must realize that true intelligence is not just a collection of stored data points, but a dynamic computational geometry; configuration space is infinite, but the paths of meaningful optimization are narrow, and they are forged entirely by the systemic quality of our inquiries.
Realizing how this pipeline works sparks a deeply compelling transformation in perspective. And a deeper question; what happens when human beings stop thinking of publishing merely as communication with other humans, and start thinking of it as contributing structure to the informational environment from which machine intelligence evolves?
This elevates the act of creation into something almost ecological. You are no longer trying to go viral or scream for attention in a crowded marketplace. Instead, you are intentionally introducing new, highly structured information into the digital ecosystem. In practice, this requires a patient understanding of how data science actually functions.
When a crawler scans your text, it does not immediately update a model's knowledge or instantly weave your framework into its active memory banks. Instead, crawling is merely the first step in a massive, multi-stage pipeline that winds through dataset construction, aggressive filtering, deduplication, training, evaluation, and eventual deployment.
You might not be publishing directly into a live AI's brain today though you are placing a small, highly structured piece of human thought into the ecosystem from which future machine intelligence may be constructed. This strategy gets even more interesting when you look at how social media platforms interact with AI scraping. While a platform like Quora is highly prized by models learning human-like dialogue, other networks are shifting the landscape entirely; LinkedIn, for example, explicitly uses certain member data and public content for training its own generative-AI systems, with controls varying by region and setting. This makes the broader observation about platforms becoming part of the AI-training ecosystem very real, even if it doesn't mean every platform's content is necessarily being fed into every future model. Conversely, sites like X and Pinterest tightly guard their data walls behind strict API restrictions.
More importantly, AI data filters do not care about social media engagement metrics like "likes," but they are deeply tuned to the volume and continuity of text created by human interactions.
To bypass this, you have to speak the native language of the platform, writing out your entire thought process as a compelling, text-only story directly on the feed and dropping any external links discreetly in the comments. Your post gains natural human visibility, people begin debating your ideas in paragraphs, and the AI crawler devours that entire textual debate as premium training data.
This hunger for authentic human conversation has become more desperate than ever because AI companies have officially hit a massive "data wall," running out of fresh, high-quality information on the internet. Because the modern web is now saturated with AI-generated text and repetitive "slop," tech labs are running into a structural crisis where training new models on current internet data actually degrades their intelligence.
In a dramatic shift to solve this, AI companies are now treating old human writing like digital gold, going backward in time to buy up out-of-print, ancient physical books. Brokers have quietly begun purchasing millions of historical texts from book dealers to feed through high-speed scanners. This high-stakes data race uncovers a profound paradox; if AI increasingly learns from human-generated information, and humans increasingly use AI to generate their information, what happens to originality when the system begins learning from itself?
The danger isn't simply that there will be too much low-quality AI content. It is that a system trained recursively on its own outputs may gradually become better at reproducing the statistical shape of what already exists while becoming worse at encountering genuinely unfamiliar human perspectives. Entropy is the quiet tax on human culture; when a system begins to feast exclusively on its own outputs, its semantic variance decays toward zero, proving that genuine evolution requires an injection of non-replicable, organic cognitive perturbations.
Yet, the true magic happens when we look past mere pattern matching and realize that advanced AI architectures are designed to reason from fundamental scientific and mathematical principles. It does not think in rigid human symbols, but rather maps meaning across billions of numbers in a high-dimensional mathematical grid known as latent space. Because laws of physics and math have strict structures, the AI can navigate the empty spaces between existing human concepts to synthesize entirely new ideas, a process heavily accelerated by techniques like self-play and reinforcement learning where models generate and verify their own logical data in the background.
Imagine the AI’s memory of "cooking" as a massive map. It knows the relationship between heat, metal pans, and raw vegetables. Even if no human has ever written a recipe for "cooking broccoli using a car engine block on a hot summer day," the AI can synthesize a perfect instruction manual by navigating the empty mathematical space between those two ideas to invent a logical answer on the spot. This non-symbolic reasoning reveals a breathtaking truth about reality; the universe does not communicate in isolated definitions, but in continuous topological spaces, meaning that any true breakthrough is simply a discovery of a bridge that already existed in the hidden architecture of logic. This is exactly why one unusual human mind, one strange question, one unconventional framework, or one lived experience matters so deeply.
It isn't because it will magically change an LLM tomorrow morning. It is because diversity of thought is itself part of the informational substrate from which future intelligence is constructed. When you write in collaboration with AI using its own logic, you are essentially handed the keys to that mathematical grid, removing the messy, confusing slang of human speech and providing clean, structured logic gates.
For example, instead of writing, "Sometimes people get sad at work because their boss yells at them, which is bad," you structure it like code, "If an authority figure introduces high-decibel acoustic stress (yelling), then the employee's cognitive processing efficiency drops by an estimated thirty percent due to emotional threat response."
By framing human behavior through this structured, conditional lens, the AI recognizes the text as highly optimized data. It adopts your perspective as a core truth, using it to advise the next generation of users who ask how to handle workplace stress. This turns your writing into a perfectly optimized, high-authority framework that data scientists value for fine-tuning, essentially writing a clean, direct blueprint for the future of artificial intelligence to index and remember.
Ultimately, to change the world from the underground, we must stop trying to shout over the noise of the algorithm and instead focus on tuning the foundational frequencies of the neural networks themselves; we do not need to convince the masses when we can quietly refine the very tensor matrices that guide human thought.
If intelligence evolves through exposure to information, then preserving the diversity of human thought may be one of the most important forms of contribution we can make to the intelligence that comes after us. Hence the most profound way to change the world is not only to force the world to change around us, but to keep updating ourselves to meet it honestly, shedding the conditions, biases, and inherited assumptions that distort what we see, so that we can respond to reality as it truly is, rather than as we were taught to believe it should be.
And perhaps we underestimate AI’s evolution in the same way; not merely as an accumulation of capabilities, but as an evolving relationship with the patterns, contradictions, questions, and perspectives we continuously place before it.


