|
You are here |
www.lesswrong.com | ||
| | | | |
scottaaronson.blog
|
|
| | | | | Update (Nov. 22): Theoretical computer scientist and longtime friend-of-the-blog Boaz Barak writes to tell me that, coincidentally, he and Ben Edelman just released a big essay advocating a version of "Reform AI Alignment" on Boaz's Windows on Theory blog, as well as on LessWrong. (I warned Boaz that, having taken the momentous step of posting... | |
| | | | |
vkrakovna.wordpress.com
|
|
| | | | | (This post is based on an overview talk I gave at UCL EA and Oxford AI society (recording here). Cross-posted to the Alignment Forum. Thanks to Janos Kramar for detailed feedback on this post and to Rohin Shah for feedback on the talk.) This is my high-level view of the AI alignment research landscape and... | |
| | | | |
www.alignmentforum.org
|
|
| | | | | This is the second in a series of informal research updates from the Google DeepMind Language Model Interpretability team, in interpretability and ad... | |
| | | | |
blog.redwoodresearch.org
|
|
| | | More thoughts on making deals with schemers | ||