Area of inquiry
How do we keep AI systems safe and aligned?
One of the field's big questions. It gathers the themes below — all questions the research asks, not subjects it is about. (For subjects, browse Topics.)
Themes within this area 5
Each theme is a narrower question. Follow one down toward its lines of inquiry.
- How do alignment methods inadvertently affect model behavior and generalization?
- How can humans maintain effective control over AI systems?
- How do adversarial attacks exploit vulnerabilities in AI safety monitoring?
- How do architectural choices affect system safety and transparency?
- How does AI reshape human understanding and epistemic accountability?