Chain-of-Thought Unfaithfulness Research
📖 Description
Research shows AI reasoning steps often don't represent actual decision-making process, undermining interpretability and monitoring approaches
📊 Game Impacts
| Variable | Change | Condition |
|---|---|---|
| Research | +25 | Always |
| Papers | +20 | Always |
| Ethics Risk | +30 | Always |
| Technical Debt | +25 | Always |
| Vibey Doom | +20 | Always |
💭 Reactions
🔬 Safety Researcher Reaction:
⚠️ Placeholder - Needs Real Quote
"'We can't trust what the model claims to be thinking'"
"'We can't trust what the model claims to be thinking'"
📰 Media Reaction:
⚠️ Placeholder - Needs Real Quote
"AI's 'reasoning' may not reflect actual decision process"
💡 Found a Real Quote? Suggest it here
"AI's 'reasoning' may not reflect actual decision process"
🔗 Sources
🤝 Found an Issue?
This event data is sourced from the pdoom-data repository. If you notice errors or want to suggest improvements:
GitHub Issue (Preferred) 📧 Email (No GitHub)