|
You are here |
www.lesswrong.com | ||
| | | | |
www.greaterwrong.com
|
|
| | | | | Eric DrexlerCentre for the Governance of AIUniversity of Oxford This document argues for "open agencies" - not opaque, unitary agents - as the appropriate model for applying future AI capabilities to consequential tasks that call for combining human guidance with delegation of planning and implementation to AI systems. This prospect reframes and can help to tame a wide range of classic AI safety challenges, leveraging alignment techniques in a relatively fault-tolerant context. | |
| | | | |
www.alignmentforum.org
|
|
| | | | | This piece gives an overview of the alignment problem and makes the case for AI alignment research. It is crafted both to be broadly accessible to th... | |
| | | | |
amatria.in
|
|
| | | | | 2024 has been an intense year for AI. While some argue that we haven't made much progress, I beg to differ. It is true that many of the research advances from 2023 have still not made it to mainstream applications. But, that doesn't mean that research is not making progress all around! | |
| | | | |
www.alignmentforum.org
|
|
| | | AI researchers warn that advanced machine learning systems may develop their own internal goals that don't match what we intended. This "mesa-optimiz... | ||