Scientific Definition
Models optimized for plausible-sounding answers can hallucinate confidently rather than say 'I don't know.'
Plain-English Definition
Models optimized for plausible-sounding answers can hallucinate confidently rather than say 'I don't know.'
Feynman Explanation
Confidence pays. Calibration doesn't.
Core Principle
Models optimized for plausible-sounding answers can hallucinate confidently rather than say 'I don't know.'
Mechanisms
Pending editorial review.
Models optimized for plausible-sounding answers can hallucinate confidently rather than say 'I don't know.'
Pending editorial review.
Pending editorial review.
Training-reward design produces deployment behavior.
Pending editorial review.
Pending editorial review.
Inputs (Triggers)
Pending editorial review.
Outputs (Behaviors)
Pending editorial review.
Behavioral Signature
Confidence pays. Calibration doesn't.
Examples
- Public LLM hallucinations under RLHF pressure for helpfulness.
- Training-reward design produces deployment behavior.
Pending editorial review.
Famous Experiments
Pending editorial review.
Design Principles
- Calibration-weighted RLHF. Uncertainty-aware UX.
Measurement Approaches
Pending editorial review.
Evidence
Pending editorial review.
Pending editorial review.
The Perverse Incentive Lens™
How this behavior is exploited — and how to redesign around it.
- Calibration-weighted RLHF. Uncertainty-aware UX.
Pending editorial review.
Pending editorial review.
Interactive Mini Network
Click any neighbor to re-center the graph and follow the threads of connection.
Knowledge Graph Neighbors
Auto-linked to the rest of the Human Behavior Taxonomy by family, domain, dimension, and shared keywords.
Autonomous agents deployed before liability frameworks exist.
Individual productivity gains hide collective output degradation.
AI generates content; AI scrapes content; AI trains on its own output.
Confident incorrect outputs may rank higher than hedged correct ones.
Recommenders optimizing engagement produce radicalization as a byproduct.
Discount framing nudges people to buy things they wouldn't otherwise want.
Productivity targets compress visits, raising misdiagnosis and burnout.
A federal mandate intended to lower drug costs for the poor became a profit engine for hospitals and contract pharmacies.
Earn-outs designed to retain founders often demotivate the team they bought.
Funnels rewarded for new logos under-invest in retention and lifetime value.
Free products monetize attention, structurally aligning incentives against user time well spent.
Tenure-track jobs replaced by low-paid adjuncts, lowering cost and quality.