Scott Alexander, curated
← Back to curation

Basics of Human Reinforcement

Quality
72
Strong
Claude Shift
38
Slight
RWI
2
of 10

Summary

Extends the behaviorism sequence from animals to humans, tackling the puzzle of why people avoid behaviors they've never tried (giving Mr. T the finger). Tools: secondary reinforcers (money), status as a primary social reinforcer (monkeys pay juice to see high-status faces), and internal/value-based reinforcement (Bandura). Then generalization (Little Albert's fear spreading from white rat to rabbit to Santa; Skinner's superstitious pigeons) and social learning (the striking NurtureShock finding that educational TV correlates with relational aggression more strongly than violent TV does, because children see the bully high-status before the late comeuppance lands). Honestly flags the limits — the smoker who quits on a doctor's word fits none of these; he is skeptical of Dennett's simulated 'inner environment' (too much magic: why can't I get addicted to heroin just by imagining it?) and floats a weak-cognitive-link neural-net story. Key move: reinforcement learning, like utility theory, runs on expected reward.

Why this score

Quality 72 · Strong. Strong-ish. A substantive, well-exemplified extension to human behavior (the educational-TV/bullying result is memorable) that honestly wrestles with where the theory breaks down, clearly above the animal-basics primer. Expository at root, so low-Strong.

Claude’s paradigm shift 38 · Slight. Moderate-slight — applies established reinforcement, generalization, and social-learning concepts to humans; the synthesis and the NurtureShock finding supply the freshness.

Real-world impact 2 · Minor. Extends the behaviorism sequence to humans (secondary reinforcers, status as a primary social reinforcer, the memorable educational-TV/relational-aggression result) while honestly wrestling with where the theory breaks. Conceptual/expository influence within rationalist discourse, no material change — low RWI.