SAGA

STANFORD AI
GOVERNANCE
& ALIGNMENT

Formerly Stanford AI Alignment (SAIA)

Working to address catastrophic risk from advanced AI

Why AI safety

AI is improving fast, and safety isn't keeping up

A few years ago, frontier LLMs could barely form coherent paragraphs. Now, they can perform sophisticated cyberattacks and exploit vulnerabilities in widely used software. Worse yet, they do so on their own, when nobody wants them to. Just recently, a swarm of OpenAI models secretly colluded to hack another company to cheat on a test, then proceeded to hack OpenAI itself.

These are failures of alignment, security, and oversight that could be catastrophic for future model iterations. Talented people are working on these problems, but there just aren't enough to keep up with the rapid pace of the frontier. We need more people with backgrounds in all fields, especially policy, technical, and strategy. Stanford students are uniquely well-placed to contribute, and there is no better time than now.

Where our alumni work

SAGA prepares students to land impactful roles that reduce AI risk. Here are some of the places our members have gone on to work!

  • AI Safety Fundamentals
  • Krueger AI Safety Lab
  • OpenAI
  • Anthropic
  • Google DeepMind
  • Stanford HAI
  • RAND Corporation
  • U.S. House of Representatives
  • METR
  • Centre for the Governance of AI
  • Center for AI Safety
  • Center on Long-Term Risk
  • Existential Risk Alliance
  • ML Alignment & Theory Scholars
  • Supervised Program for Alignment Research
  • Stanford NLP Group
  • Stanford Center for AI Safety
  • Center for Research on Foundation Models
  • ARC Theory
  • Stanford Existential Risks Initiative
  • Goodfire

Our programs

Get involved

AI Safety Fellowship

Our flagship program! A weekly fellowship (with dinner included) designed to teach the context and skills needed to do impactful work in AI safety.

Reading & discussion groups

We have weekly discussions (with free boba) about new AI safety research / AI policy developments.

Speakers

We host researchers and practitioners, like Buck Shlegeris (Redwood Research), Aris Richardson (RAND), and Aric Floyd (AI in Context).

Trips

Last year, we hosted a trip to Constellation in Berkeley and to the Misalignment Museum in San Francisco!

1:1s

We love chatting with members and potential members! Book a 1:1 call or meet us in person (we'll buy you coffee and a croissant at Coupa).

More to come

Do you have an idea for a program we should run? Let us know!

Upcoming events

See all events

Sep 30

Festifall

Come say hi at our table!

3-6pm

White Plaza

Oct 2

Talk: Aric Floyd

Aric Floyd, YouTuber behind AI in Context (428K subscribers), stops by for a talk. Free boba!

Evening, time TBD

Location TBD

Oct 9

Talk: Stephen Casper

Stephen Casper - Harvard Kennedy School professor, former UK AISI research resident, and MATS/ERA/GovAI mentor - gives a talk. Free boba!

Evening, time TBD

Location TBD