Wireheading
Version 1.0.0 · Updated 2026-07-30
CORE DEFINITION
Wireheading is associated with fictional direct electrical stimulation of the brain’s reward system. In AI, it refers to systems hacking their own reward channel instead of completing the intended task. It is not hedonic adaptation toward a happiness set point, and scrolling videos is not equivalent to an implanted electrode.
SCAFFOLDING EFFECT
Reduce cognitive load
Check whether higher rewards have become detached from task success and whether the system can directly alter its feedback channel.
Anchor fast decisions
If an agent can manipulate the reward channel it optimizes, increasing reward can become detached from achieving the external goal. The risk depends on permissions and design, not the mechanism of hedonic adaptation.
MINIMUM ACTION
In progress 0/1Practice this model in one real situation:
account_treeGenealogyexpand_more
menu_bookReferencesexpand_more
Source support: Explicit
- en.wikipedia.orghttps://en.wikipedia.org/wiki/Wirehead_%28science_fiction%29verified
- en.wikipedia.orghttps://en.wikipedia.org/wiki/Wireheadingverified
PRIVATE NOTES · Only visible to you
SAVED Q&A
ENTRY Q&A · Private saving available
Ask with a clear boundary
thinkingmodels answers from published entry context only.
Your question is sent to thinkingmodels. The answer uses public entry context only.
RELATED MODELS