#01 Orient
Why this, why now
The case for taking advanced AI seriously: capabilities, timelines, and what it would mean for things to go well or badly.
SAISI-F26-FEL · Fall 2026
An eight-week reading group on how powerful AI systems fail, who is working to prevent it, and what a Swarthmore student can actually do about it.
Reading list · Discussed in the room
#01 Orient
The case for taking advanced AI seriously: capabilities, timelines, and what it would mean for things to go well or badly.
#02 Open
Pretraining, fine-tuning, RLHF, and scaling laws, explained without code. Enough to read the rest of the syllabus.
#03 Break
Specification gaming, reward hacking, goal misgeneralization. Real incidents, real transcripts.
#04 Doubt
Situational awareness, sandbagging, and whether we can keep a system that is smarter than its overseers in check.
#05 Look
Interpretability: what circuits, features, and probes have found so far, and what they have not.
#06 Test
Dangerous-capability evals, red teaming, and the organizations whose job is to say no before a launch.
#07 Govern
Compute governance, safety frameworks, export controls, and who actually decides.
#08 Act
Research programs, policy fellowships, career paths, and a short project you present to the cohort.
Questions