UChicago AI Safety logoUChicagoAI Safety

Fellowship Syllabus

AI Safety Fundamentals

Our flagship fellowship introduces fellows from any background to the core ideas in AI safety, with a particular focus on existential risk from advanced AI. By the program's end, fellows will have a working map of the field and the foundation to dive deeper into the subareas that interest them.

Week 01

Philosophical and Political Foundations of AI Safety

Explore the implications of increasingly intelligent systems.

Week 02

Outer Alignment

Examine the challenges in correctly specifying training goals for AI systems.

Week 03

Deception, Inner Alignment & Mechanistic Interpretability

Investigate the concept of mesa-optimizers and the potential for deceptive behavior in AI systems.

Week 04

AI Security

Explore various AI security issues including jailbreaks, adversarial examples, and potential vulnerabilities.

Week 05

AI Governance

Examine the challenges and approaches to governing AI development and deployment.

Week 06

Criticisms and Counter-Arguments

Examine critiques of AI safety concerns and alternative perspectives on AI development.

Week 07

Further Reading and Discussion

Explore various AI alignment approaches and dive deeper into specific areas of interest. Fellows will choose one of the optional readings to focus on for the week.