/explore

Click through on any links that interest you or select the planets on the right to continue exploring the Outer Web.
You are here

www.lesswrong.com
| | www.greaterwrong.com
3.6 parsecs away

Travel
| | TL;DR:Strong problem-solving systems can be built from AI systems that play diverse roles, LLMs can readily play diverse roles in role architectures, and AI systems based on role architectures can be practical, safe, and effective in undertaking complex and consequential tasks. This article explores the practicalities and challenges of aligning large language models (LLMs[1]) to play central roles in performing tasks safely and effectively. It highlights the potential value of Open Agency and related role architectures in aligning AI for general applications while mitigating risks.
| | www.alignmentforum.org
1.9 parsecs away

Travel
| | Preface The following text is my submission for theAI Safety Public Materials contest. In it, I try to lay out the importance of AI Safety Research...
| | distill.pub
3.5 parsecs away

Travel
| | If we want to train AI to do what humans want, we need to study humans.
| | magnusvinding.com
25.3 parsecs away

Travel
| The following is a point-by-point critique of Lukas Gloor's essay Altruists Should Prioritize Artificial Intelligence. My hope is that this critique will serve to make it clear - to Lukas, myself, and others - where and why I disagree with this line of argument, and thereby hopefully also bring some relevant considerations to the table...