STANFORD AI
GOVERNANCE
& ALIGNMENT
Why AI safety
A few years ago, frontier LLMs could barely form coherent paragraphs. Now, they can perform sophisticated cyberattacks and exploit vulnerabilities in widely used software. Worse yet, they do so on their own, when nobody wants them to. Just recently, a swarm of OpenAI models secretly colluded to hack another company to cheat on a test, then proceeded to hack OpenAI itself.
These are failures of alignment, security, and oversight that could be catastrophic for future model iterations. Talented people are working on these problems, but there just aren't enough to keep up with the rapid pace of the frontier. We need more people with backgrounds in all fields, especially policy, technical, and strategy. Stanford students are uniquely well-placed to contribute, and there is no better time than now.
Where our alumni work
Our programs
Get involvedAI Safety Fellowship
Our flagship program! A weekly fellowship (with dinner included) designed to teach the context and skills needed to do impactful work in AI safety.
Reading & discussion groups
We have weekly discussions (with free boba) about new AI safety research / AI policy developments.
Speakers
We host researchers and practitioners, like Buck Shlegeris (Redwood Research), Aris Richardson (RAND), and Aric Floyd (AI in Context).
Trips
Last year, we hosted a trip to Constellation in Berkeley and to the Misalignment Museum in San Francisco!
1:1s
We love chatting with members and potential members! Book a 1:1 call or meet us in person (we'll buy you coffee and a croissant at Coupa).
More to come
Do you have an idea for a program we should run? Let us know!
Upcoming events
See all eventsPolicy Reading Group
Interested in AI policy but finding it hard to keep up with everything happening? Come to SAGA’s new weekly Policy Reading Group. We’ll catch up on the news, read something short together, and talk about what governments should actually do about AI. No background knowledge or prep needed. Dinner provided!
- Time: 6-7:30pm
- Place: Old Union Room 121
TREEFEST
Come say hi at our table!
- Time: 2-5pm
- Place: White Plaza
Talk: Aric Floyd
Join us for a deep dive into the Hugging Face incident and its implications about the systems we're building. Boba provided! Aric Floyd is the face of AI in Context, a Youtube channel with over 435k subscribers and 20 million views. He was the lead in the viral AI 2027 explainer video and a Stanford alum.
- Time: 5:30-7pm
- Place: Bishop Auditorium




















