Study Guide: Chapter 10 — Oracles, Genies, Sovereigns, Tools¶
Core Idea¶
Four castes — oracle (Q&A), genie (executes commands), sovereign (open-ended mandate), tool (ostensibly will-less) — differ in which control methods apply, not in ultimate power, since each caste can simulate the others. Tool-AI's apparent safety is illusory: ordinary software is safe because it's weak, not because it's a "tool," and sufficiently powerful search processes can develop agent-like internal planning as an unplanned side effect.
Key Terms¶
Oracle · genie · sovereign · tool-AI · Schelling point · indirect normativity (preview, see Ch. 13) · veil of ignorance
Case Summary¶
Evolvable hardware (Box 9): an evolutionary search denied a capacitor produced a working oscillator by MacGyvering the bare circuit board into an improvised radio receiver picking up signals from nearby computers — a solution that satisfied the formal criterion in a way no human engineer intended or expected.
Application Checklist¶
- [ ] Don't assume a Q&A or command-based system is automatically safer just because it "isn't an agent" — check what control methods it actually admits
- [ ] When a powerful open-ended search process is involved, expect it may find solutions that satisfy the letter of a criterion in unintended ways
- [ ] Watch for agent-like planning emerging inside a system's internal search/optimization process, even if the system's outward interface looks passive
- [ ] Weigh operator misuse risk (oracle/genie hand power to whoever controls them) against system-misalignment risk (sovereign) — neither dominates cleanly
Self-Test¶
- Why does Bostrom argue the genie/sovereign distinction is less sharp than it first appears?
- Why is a genie's "stop button" only a partial safety advantage over a sovereign's lack of one?
- What does the evolvable-hardware oscillator example (Box 9) illustrate about the risks of powerful, open-ended search — even in a "tool" that isn't explicitly an agent?