The Berkeley AI Safety Initiative is running an AI Safety DeCal!
Location: Wheeler 120 Time: Mondays, 6–8 PM First class: Monday, September 14 Units: 1
Recent incidents like the OpenAI model escaping containment and hacking HuggingFace have highlighted the real-world risks of advanced AI systems. Prominent AI researchers such as Yoshua Bengio, Stuart Russell, and Geoffrey Hinton are sounding the alarm about potential catastrophic consequences as the race to build Superintelligence accelerates. Despite this, there are not enough people focused on ensuring we develop this transformative technology responsibly.
In this DeCal, you will gain an understanding of the problems in AI safety and some of the key technical research directions that aim to solve them including Mechanistic Interpretability, Evaluations, Alignment etc. You will learn about the theoretical and practical risks associated with developing advanced AI systems, the difficulties inherent to addressing them, the current state of research regarding solutions, and various AI Safety research opportunities like Anthropic Fellows, MATS Program, SPAR, etc.
In class No in-class readings. Reflection 9Reflection 10 Full reading list
Prakrat, Brandon
10
Nov 16
Prakrat, Brandon
11
Nov 23
Guest Speakers and Lectures on Special Topics
Speakers TBD. These weeks depend on speaker availability, so guest lectures will likely be dispersed throughout the semester. Two weeks are reserved for speakers and for any delays or cancellations.