|
You are here |
www.lesswrong.com | ||
| | | | |
www.greaterwrong.com
|
|
| | | | | TL;DR:Strong problem-solving systems can be built from AI systems that play diverse roles, LLMs can readily play diverse roles in role architectures, and AI systems based on role architectures can be practical, safe, and effective in undertaking complex and consequential tasks. This article explores the practicalities and challenges of aligning large language models (LLMs[1]) to play central roles in performing tasks safely and effectively. It highlights the potential value of Open Agency and related role architectures in aligning AI for general applications while mitigating risks. | |
| | | | |
scottaaronson.blog
|
|
| | | | | Two weeks ago, I gave a lecture setting out my current thoughts on AI safety, halfway through my year at OpenAI. I was asked to speak by UT Austin's Effective Altruist club. You can watch the lecture on YouTube here (I recommend 2x speed). The timing turned out to be weird, coming immediately after the... | |
| | | | |
joecarlsmith.com
|
|
| | | | | A high-level picture of how we might get from here to safe superintelligence. | |
| | | | |
www.alignmentforum.org
|
|
| | | In this post I'm going to describe my basic justification for working on RLHF in 2017-2020, which I still stand behind. I'll discuss various argument... | ||