What is AI Safety?
From Wikipedia, "AI safety is an interdisciplinary field focused on preventing accidents, misuse, or other harmful consequences arising from artificial intelligence systems. It encompasses AI alignment (which aims to ensure AI systems behave as intended), monitoring AI systems for risks, and enhancing their robustness. The field is particularly concerned with existential risks posed by advanced AI models."
If you'd like to learn more, the following resources are a good starting point
- The Future of AI. Self-paced, 45-minute course by BlueDot Impact
- Risks from power-seeking AI systems. Survey article by 80,000 Hours. 90 minute read.
- International AI Safety Report 2026. Report led by Turning Award winner Yoshua Bengio published by the UK's AI Security Institute.