paper
Research Priorities for Robust and Beneficial Artificial Intelligence
An interdisciplinary agenda distinguishes formal correctness from desirable behavior and human control.
Stuart Russell, Daniel Dewey and Max Tegmark organize safety research around economics, law, ethics and technical robustness. Their agenda distinguishes verification of formal requirements, validity of those requirements, protection against unauthorized manipulation, and human control during operation (pp.107–108).[1]
Specification and learning
A formally correct system can still fail because its environmental assumptions are wrong or its requirements are undesirable. The authors discuss learning preferences from behavior as a possible route to AI Alignment, while leaving its adequacy open (pp.108,110).[1]
Historical context
The agenda accompanied the 2015 beneficial-AI open letter. The selected original here is the later Winter 2015 journal article; its publisher records December 31, 2015. This date does not date the initial agenda. The article maps research questions, rather than supplying validated safeguards or empirical probabilities of future harm.[1][2]
Sources
- Research Priorities for Robust and Beneficial Artificial Intelligence · Source record src-191 · Back to claim ↑1 ↑2 ↑3
- 2015: An Amazing Year in Review · Source record src-192 · Back to claim ↑1
Pages that link here
- 2015 beneficial-AI open letter event
- From objectives to accountable control note
- Outer Alignment concept
- Stuart Russell person
Last updated 2026-10-10