Skip to content
SAGA

STANFORD AI
GOVERNANCE
& ALIGNMENT

Why AI safety

A few years ago, frontier LLMs could barely form coherent paragraphs. Now, they can perform sophisticated cyberattacks and exploit vulnerabilities in widely used software. Worse yet, they do so on their own, when nobody wants them to. Just recently, a swarm of OpenAI models secretly colluded to hack another company to cheat on a test, then proceeded to hack OpenAI itself.

These are failures of alignment, security, and oversight that could be catastrophic for future model iterations. Talented people are working on these problems, but there just aren't enough to keep up with the rapid pace of the frontier. We need more people with backgrounds in all fields, especially policy, technical, and strategy. Stanford students are uniquely well-placed to contribute, and there is no better time than now.

Where our alumni work

  • AI Safety Fundamentals
  • OpenAI
  • Google DeepMind
  • RAND Corporation
  • METR
  • Center for AI Safety
  • Existential Risk Alliance
  • Supervised Program for Alignment Research
  • Stanford Center for AI Safety
  • ARC Theory
  • Goodfire
  • Krueger AI Safety Lab
  • Anthropic
  • Stanford HAI
  • U.S. House of Representatives
  • Centre for the Governance of AI
  • Center on Long-Term Risk
  • ML Alignment & Theory Scholars
  • Stanford NLP Group
  • Center for Research on Foundation Models
  • Stanford Existential Risks Initiative

Our programs

Get involved

AI Safety Fellowship

Our flagship program! A weekly fellowship (with dinner included) designed to teach the context and skills needed to do impactful work in AI safety.

Reading & discussion groups

We have weekly discussions (with free boba) about new AI safety research / AI policy developments.

Speakers

We host researchers and practitioners, like Buck Shlegeris (Redwood Research), Aris Richardson (RAND), and Aric Floyd (AI in Context).

Trips

Last year, we hosted a trip to Constellation in Berkeley and to the Misalignment Museum in San Francisco!

1:1s

We love chatting with members and potential members! Book a 1:1 call or meet us in person (we'll buy you coffee and a croissant at Coupa).

More to come

Do you have an idea for a program we should run? Let us know!

Upcoming events

See all events

Policy Reading Group

Interested in AI policy but finding it hard to keep up with everything happening? Come to SAGA’s new weekly Policy Reading Group. We’ll catch up on the news, read something short together, and talk about what governments should actually do about AI. No background knowledge or prep needed. Dinner provided!

  • Time: 6-7:30pm
  • Place: Old Union Room 121
RSVP

TREEFEST

Come say hi at our table!

  • Time: 2-5pm
  • Place: White Plaza

Talk: Aric Floyd

Join us for a deep dive into the Hugging Face incident and its implications about the systems we're building. Boba provided! Aric Floyd is the face of AI in Context, a Youtube channel with over 435k subscribers and 20 million views. He was the lead in the viral AI 2027 explainer video and a Stanford alum.

  • Time: 5:30-7pm
  • Place: Bishop Auditorium
RSVP