Scientific Definition
Why did it produce this? vs. How does it work?
Plain-English Definition
Why did it produce this? vs. How does it work?
Feynman Explanation
Different questions. Different methods. Both useful.
Core Principle
Why did it produce this? vs. How does it work?
Mechanisms
Pending editorial review.
Pending editorial review.
Pending editorial review.
Pending editorial review.
Pending editorial review.
Pending editorial review.
Pending editorial review.
Inputs (Triggers)
Pending editorial review.
Outputs (Behaviors)
Pending editorial review.
Behavioral Signature
Different questions. Different methods. Both useful.
Examples
- SHAP values vs. mechanistic interpretability.
- Compliance, user trust, debugging.
Pending editorial review.
Famous Experiments
Pending editorial review.
Design Principles
- Match the tool to the question.
Measurement Approaches
Pending editorial review.
Evidence
Pending editorial review.
Pending editorial review.
The Perverse Incentive Lens™
How this behavior is exploited — and how to redesign around it.
- Match the tool to the question.
Pending editorial review.
Pending editorial review.
Interactive Mini Network
Click any neighbor to re-center the graph and follow the threads of connection.
Knowledge Graph Neighbors
Auto-linked to the rest of the Human Behavior Taxonomy by family, domain, dimension, and shared keywords.
Written rules about how AI may be used internally.
Logging of AI inputs, outputs, and decisions.
Inventory of models, data, tools, and dependencies in an AI system.
Tracing AI components for risk and compliance.
Cross-functional governance body for AI decisions.
Agents deployed before anyone owns the consequences.
Cryptographic tracking of content origin.
Tracking where training and inference data came from.
Comprehensive AI regulation in the EU.
Human review at critical AI decision points.
Human oversight without per-decision review.
Documentation of model purpose, performance, limitations, and risks.