Modeling Risks From Learned Optimization
📖 Description
*This post, which deals with how risks from learned optimization and inner alignment can be understood, is part 5 in our* [*sequence on Modeling Transformative AI Risk*](https://www.alignmentforum.org/s/aERZoriyHfCqvWkzg)*. We are building a model to understand debates around existential risks from advanced AI. The model is made with* [*Analytica*](https://en.wikipedia.org/wiki/Analytica_(software)) *software, and consists of nodes (representing key hypotheses and cruxes) and edges (representing the relationships between these cruxes), with final output corresponding to the likelihood of various potential failure scenarios. You can read more about the motivation for our project and how the model works in the* [*Introduction post*](https://www.alignmentforum.org/posts/qnA6paRwMky3Q6ktk/modelling-transformative-ai-risks-mtair-project-introduction)*. The previous post in the sequence,* [*Takeoff Speeds and Discontinuities*](https://www.alignmentforum.org/posts/pGXR2ynhe5bBCCNqn/takeoff...
📊 Game Impacts
| Variable | Change | Condition |
|---|---|---|
| Research | +10 | Always |
| Vibey Doom | +5 | Always |
| Ethics Risk | -5 | Always |
💭 Reactions
"Critical insights for the field"
"Discussed in AI safety community"
🔗 Sources
🤝 Found an Issue?
This event data is sourced from the pdoom-data repository. If you notice errors or want to suggest improvements:
GitHub Issue (Preferred) 📧 Email (No GitHub)