Eckersley, 2019 — Impossibility and Uncertainty Theorems in {AI} Value Alignment
Ethical impossibility results undermine strict total-order objectives for high-stakes AI; learned rewards do not escape if they collapse to a single ordering.
Publication links
Ethical impossibility results undermine strict total-order objectives for high-stakes AI; learned rewards do not escape if they collapse to a single ordering.