
Welcome back AI prodigies!
In todayâs Sunday Special:
đThe Prelude
đHow LLMs Remove Context
âïžWhat Do We Lose?
âïžLLMs vs. Truth
đKey Takeaway
Read Time: 7 minutes
đKey Terms
Tokens: A word or part of a word. For example, âyesterdayâ might be broken into three tokens: âyes,â âter,â and âday.â
Large Language Models (LLMs): AI Models pre-trained on vast amounts of high-quality datasets to generate human-like text.
đ©ș PULSE CHECK
Would you trust advice more if you knew who gave it?
đTHE PRELUDE
Imagine walking into a courtroom: the judge is absent, the jury is missing, and the defendant is gone. Still, a gavel strikes, and the verdict is read aloud. Justice is rendered, but no oneâs there to bear it.
Thatâs what AI-generated text often feels like: words that mimic meaning but lack the human presence that once gave them weight.
When we type, our words carry context, connection, and conviction. In other words, our words possess an underlying consciousness. Each sentence reflects a lived experience crafted from our unique perspective.
LLMs provide us with an infinite amount of AI-generated text on demand, but theyâve also changed the nature of text itself. We used to assume that when we read something:
Someone meant to say something.
Someone stood behind the words.
How exactly are LLMs engineered to produce text without the usual human stuff behind it? Why canât LLMs reliably tell whatâs true?
đHOW LLMs REMOVE CONTEXT
⊿ 1ïžâŁ âïžHow LLMs Remove Context.
To understand what we lose, we must first examine how LLMs are constructed.
LLMs are essentially statistical tools designed to predict the probability of a sequence of words. You can view them as sophisticated autocomplete machines trained on the entire internet.
For example, when given: âThe cat chased the {BLANK}!â LLMs ask themselves: given the words so far, whatâs the most likely next word?
Each time LLMs guess wrong, they adjust thousands of Weights, which control how tens of thousands of words relate to each other within LLMs. These relationships help form the Neural Network (NN): a network of interconnected nodes that processes words using two methods:
đAttentions Mechanisms help calculate how much each word within a sentence should pay attention to every other word. Consider the following sentence: âThe cat chased the mouse!â In this case, the words âcatâ and âmouseâ would pay more attention to each other because the past tense transitive verb âchasedâ connects them, indicating that the âcatâ is actively pursuing the âmouse.â
đTransformer Layers help further clarify the meaning of each word within a sentence. This process enables LLMs to develop a deeper understanding of the context. Consider the following sentence: âThe cat chased the mouse!â In this case, it examines the word âchasedâ and determines that âcatâ is important because itâs doing the chasing. It also determines that âmouseâ is important because itâs being chased.
⊿ 2ïžâŁ Bias Toward the Expected?
LLMs absorb words statistically, not experientially. In other words, they map how words tend to appear together, rather than whether the sentences are grounded in reality. This gives them extraordinary fluency but a bias toward the expected. Cognitive scientists refer to this phenomenon as Regression Toward the Mean (RTM): picking the safe choice over the risky option. LLMs tend to regress to the safest choice when predicting the likely next word within a sentence because itâs statistically most probable.
In one of the most rigorous examinations of how LLMs impact idea generation, scientists at the University of Michigan (UofM) conducted a global experiment with over 800 participants across 40 countries. The participants were presented with creative ideas on a specific topic generated by LLMs. They were then asked to come up with their own original ideas on the same specific topic.
This global experiment revealed two critical patterns:
đExposure to creative ideas generated by LLMs increased the overall diversity of original ideas across the participants.
đĄEach participantâs original ideas became semantically similar, clustering around common themes.
This combination is what makes LLMs feel simultaneously abundant and strangely uniform: itâs easier to generate original ideas, but those original ideas orbit around a statistical center rather than exploring frontiers of possibility. This global experiment identified the first clue that AI-generated text is optimized for probability, not situated truth.
âïžWHAT WE LOSE?
⊿ 3ïžâŁ Whoâs Speaking?
In speech, words hold meaning, and that meaning often translates to action. In the 1950s, prominent British philosopher J. L. Austin developed Speech Act Theory (SAT): words donât just merely describe the world; they do things. This concept was later expanded by renowned American philosopher John Searle, who classified speech into five categories:
1. Directives {Requests}: âPlease close the window.â
2. Expressives {Apologies}: âSorry for the confusion earlier.â
3. Commissives {Promises}: âI promise to call you later today.â
4. Assertiveness {Stating Facts}: âThe capital of France is Paris!â
5. Declarations {Decrees Altering Reality}: âYouâre officially fired.â
These forms of speech have force because theyâre backed by context, authority, and sincerity. For example, a judge saying: âI sentence you!â carries legal weight; a casual passerby saying the same thing doesnât.
When LLMs provide us with AI-generated text, this chain is broken. âI promiseâ is no longer a commitment; itâs merely a string of Tokens that mimic commitment. The performative dimension collapses, leaving words that appear fluent yet remain hollow. In other words, AI-generated text can feel persuasive yet strangely weightless: it simulates the form of action without the responsibility that lends those actions their force.
⊿ 4ïžâŁ When and Where?
In the 1970s, influential American philosopher David Kaplan developed the formal semantics of Indexicals and Demonstratives: words whose meaning depends entirely on context. Indexicals are words like âI,â âyou,â âhere,â and ânow.â Demonstratives are words like âthis,â âthat,â âthese,â and âthose.â
His findings were simple yet profound: to interpret a sentence containing Indexicals or Demonstratives, you must know who, when, and where. Without those contextual anchors, the sentence is underspecified or meaningless. âIâll call you tonightâ only communicates something actionable if the listener knows who âIâ is, who âyouâ is, and what counts as âtonight.â
LLMs rarely supply these situational anchors. By design, theyâre placeless and timeless. They produce AI-generated text that appears coherent but detaches from the concrete âhere-and-nowâ that provides critical context.
âïžLLMS VS. TRUTH
⊿ 5ïžâŁ Prediction Without Grounding.
The loss of context would be less worrying if LLMs could at least guarantee correctness. But their architecture makes this impossible.
LLMs work by Next-Token Prediction (NTP): given a sequence of words, choose the next likely word with the highest probability of correctness given past co-occurrences. NTP is a purely statistical operation. It has no internal representation of whether a statement matches reality. If the phrase âThe Eiffel Tower is located in....â is often followed by âParis,â LLMs will output âParis.â This happens to be true. But itâs only true because the data distribution reflected reality, not because LLMs confirmed it.
đKEY TAKEAWAY
LLMs replace context with something thin, statistical, and strangely uniform. They give us infinite AI-generated text but strip words of their meaning. What weâre left with is sentences optimized for plausibility.
đFINAL NOTE
FEEDBACK
How would you rate todayâs email?
â€ïžTAIP Review of The Week
âA must-read for anyone curious about AI, I always learn something new.â
REFER & EARN
đYour Friends Learn, You Earn!
{{rp_personalized_text}}
Share your unique referral link: {{rp_refer_url}}

