Basics of Human Reinforcement
Read the original on LessWrong →
Summary
Extends the behaviorism sequence from animals to humans, tackling the puzzle of why people avoid behaviors they've never tried (giving Mr. T the finger). Tools: secondary reinforcers (money), status as a primary social reinforcer (monkeys pay juice to see high-status faces), and internal/value-based reinforcement (Bandura). Then generalization (Little Albert's fear spreading from white rat to rabbit to Santa; Skinner's superstitious pigeons) and social learning (the striking NurtureShock finding that educational TV correlates with relational aggression more strongly than violent TV does, because children see the bully high-status before the late comeuppance lands). Honestly flags the limits — the smoker who quits on a doctor's word fits none of these; he is skeptical of Dennett's simulated 'inner environment' (too much magic: why can't I get addicted to heroin just by imagining it?) and floats a weak-cognitive-link neural-net story. Key move: reinforcement learning, like utility theory, runs on expected reward.
Why this score
Quality 72 · Strong. Strong-ish. A substantive, well-exemplified extension to human behavior (the educational-TV/bullying result is memorable) that honestly wrestles with where the theory breaks down, clearly above the animal-basics primer. Expository at root, so low-Strong.
Claude’s paradigm shift 38 · Slight. Moderate-slight — applies established reinforcement, generalization, and social-learning concepts to humans; the synthesis and the NurtureShock finding supply the freshness.
Real-world impact 2 · Minor. Extends the behaviorism sequence to humans (secondary reinforcers, status as a primary social reinforcer, the memorable educational-TV/relational-aggression result) while honestly wrestling with where the theory breaks. Conceptual/expository influence within rationalist discourse, no material change — low RWI.