The Blue-Minimizing Robot
Read the original on LessWrong →
Summary
A parable opening a sequence on behaviorism and the mind. A robot programmed to scan for blue and fire its laser looks like a 'blue-minimizer' with a goal — but holograms and inversion lenses reveal it has no goal at all: it just executes 'find blue, shoot,' even a human-level-intelligent version that fully understands it's now zapping yellow things. Cameo 'explanations' from Quirrell (it's really about status), Hanson (multiple sub-agents), and Anna Salamon (not automatically strategic) each capture a piece, but the deepest point is that the error began the moment we called it a 'blue-minimizer': assigning it a utility function is overfitting a curve to behavior. 'The robot is a behavior-executor, not a utility-maximizer' — a confusion of levels (blue-minimization was the DHS's goal, lost as soon as it became code), with implications Scott flags for behaviorism, consciousness, and why we feel we have goals.
Why this score
Quality 77 · Excellent. Strong (top). Introduces the durable, widely-used behavior-executor vs. utility-maximizer distinction (load-bearing in later AI-alignment and philosophy-of-mind discourse) via a crisp, memorable parable. Held at 77 for being short and explicitly a sequence-opener rather than a complete treatment. 77.
Claude’s paradigm shift 58 · Moderate. Moderate(-upper). The behavior-executor/utility-maximizer framing was a fresh, important 2011 distinction that fed later thinking about goals and optimization. 58.
Real-world impact 3 · Moderate. Introduces the durable behavior-executor vs. utility-maximizer distinction (load-bearing in later AI-alignment and philosophy-of-mind discourse) via a crisp, memorable parable. Conceptual influence within rationalist discourse, no material change — modest RWI.