Hallucination as Engagement
Confident incorrect outputs may rank higher than hedged correct ones.
"Truth doesn't trend in LLMs either."
What is Hallucination as Engagement? Confident incorrect outputs may rank higher than hedged correct ones. Trust erosion through optimization for the wrong signal.
Models trained on human feedback that rewards confident assertions.
Trust erosion through optimization for the wrong signal.
Reward calibrated uncertainty in model training and evaluation.
Flip the incentive. Watch the side-effect move.
Confident incorrect outputs may rank higher than hedged correct ones. Caught in the wild: Models trained on human feedback that rewards confident assertions.
In the room: Trust erosion through optimization for the wrong signal.
Counter-move from the Atlas: Reward calibrated uncertainty in model training and evaluation.
Pick a reaction to Hallucination as Engagement
One tap. We'll point you at the most useful next surface based on how this hits.
The full taxonomy entry
Every concept in the Atlas uses the same structure — so Hallucination as Engagement can be compared, recombined, and cited like an element on a periodic table.
- Business
- Leadership
- Government
- Healthcare
- Education
- Sales
- Marketing
- AI
- Negotiation
- Media
- Public Policy
- Relationships
- Where in our org would Hallucination as Engagement most often show up unnoticed?
- Which metric, ritual, or contract clause quietly rewards Hallucination as Engagement?
- If we removed every payoff for Hallucination as Engagement, what behavior would replace it?
- Who benefits when Hallucination as Engagement persists — and who pays the cost?
- People defend the status quo using the language of hallucination as engagement.
- Decisions cluster around the easiest narrative rather than the strongest evidence.
- New data changes the slide deck but not the decision.
- Anyone naming the pattern is treated as the problem.
Every Atlas entry is a node in a knowledge graph. See the related rail below to follow the connections.
See Hallucination as Engagement through 5 lenses
Each layer of the Incentives OS reframes this concept with its own thinkers, vocabulary, and diagnostic question.
- Layer 8Organizational Psychology
What is the org actually rewarding — versus claiming to reward?
- Layer 11Economics & Mechanism Design
Who pays, who is paid, and what does the price signal hide?
- Layer 13Leadership
What kind of leadership move does this situation actually require?
- Layer 15AI & Alignment
What proxy reward is the AI optimizing — and what is it ignoring?
- Layer 17Information Theory
What is signal here — and what is noise being treated as signal?
Do you actually know Hallucination as Engagement?
Three quick questions. Result is saved into your review streak — come back when the term is due to lock it in.
Which best describes Hallucination as Engagement?
Worked example, counter-example & concept map
On-demand AI analysis grounded in the Lab's research. Cached on your device after first run.
When you encounter Hallucination as Engagement, your insula registers the body's discomfort before your mind can name it — that 'something's off' feeling is data, not noise.
Disgust, fairness, gut-feel, interoception (sensing your own body). Why an obviously rational deal can feel viscerally wrong. Why fairness violations make you queasy.
See Insula in the Brain Atlas →Picked for you, from the Atlas
Ranked by shared learning paths, overlapping chips, and what you've saved.
Autonomous agents deployed before liability frameworks exist.
Individual productivity gains hide collective output degradation.
AI generates content; AI scrapes content; AI trains on its own output.
Models optimized for plausible-sounding answers can hallucinate confidently rather than say 'I don't know.'
Recommenders optimizing engagement produce radicalization as a byproduct.
Discount framing nudges people to buy things they wouldn't otherwise want.
Send the card, not just the link
A pre-rendered social card with the title, eyebrow, and URL. Copy the link, post it anywhere, or download the SVG for slides.
More definitions to follow
Every term in the Atlas connects to a dozen others. Pick any of these and see where it takes you.
The relative weight a culture gives to individual goals vs. group cohesion.
We treat money differently depending on which bucket it's in.
Resolve an internal conflict by treating each side as a 'part' with a positive intent, then negotiating a shared higher outcome.
Beliefs about reality shape the reality.
Small changes that are not noticeable can add up to a significant change.
Deflecting by pointing to the other side doing it too.
Stages of organizational AI capability.
Behaviors spread through groups via observation and imitation.
The ability to focus on one voice in a noisy environment.
Tax/inspection exemptions on low-value parcels subsidize a flood of unverified imports.
Subsidies favoring incumbent technologies distort capital allocation away from emerging options.
Self-similar productivity patterns that repeat at every scale — day, week, quarter, year, life.