|
You are here |
www.alignmentforum.org | ||
| | | | |
www.greaterwrong.com
|
|
| | | | | Paul Christiano and "MIRI" have disagreed on an important research question for a long time: should we focus research on aligning "messy" AGI (e.g. one found through gradient descent or brute force search) with human values, or on developing "principled" AGI (based on theories similar to Bayesian probability theory)? I'm going to present my current model of this disagreement and additional thoughts about it. I put "MIRI" in quotes because MIRI is an organization composed of people who have differing views. I'm going to use the term "MIRI view" to refer to some combination of the views of Eliezer, Benya, and Nate. I think these three researchers have quite similar views, such that it is appropriate in some contexts to attribute a view to all of them collectiv... | |
| | | | |
www.lesswrong.com
|
|
| | | | | Nate Soares reviews a dozen plans and proposals for making AI go well. He finds that almost none of them grapple with what he considers the core prob... | |
| | | | |
www.lesswrong.com
|
|
| | | | | This is a public adaptation of a document I wrote for an internal Anthropic audience about a month ago. Thanks to (in alphabetical order) Joshua Bats... | |
| | | | |
www.greaterwrong.com
|
|
| | | |||