Main takeaways
AI safety
- Near-term and long-term concerns both matter
- Concrete problems: safe exploration, side effects, reward hacking
- Uncertainty is a reason to act now
Alignment
- Getting AI to do what we actually want
- Hard because of specification, Goodhart’s law and value disagreement
- Current approaches: RLHF, Constitutional AI
- Major open problems remain
Future trajectories
- Genuine uncertainty about where AI is heading
- Expert disagreement is real
- Plan for multiple futures
What we can do
- Technical safety research
- Informed citizens who scrutinise AI rules
- Individual choices
- Neither panic nor complacency
The future of AI is not determined. Choices about alignment, oversight and governance will shape it

















