Resources

What we send to someone who asks where to start. Almost all of it is in English, which is where the field publishes.

Five readings to start with

Videos

  • The reference video library on alignment, control and the problems that are still open.

  • AI In Context

    80,000 Hours

    Long documentaries on where AI is heading. The first one, on the AI 2027 scenario, passed ten million views.

  • Animations adapting classic essays on AI, rationality and existential risk.

  • Short animated explainers, each on one concrete risk or proposal.

  • Doom Debates

    Liron Shapira

    Debates and interviews where both positions on the risk are argued face to face.

Podcasts

  • AXRP

    Daniel Filan

    The reference podcast on technical alignment: long interviews with the people doing the research.

  • Conversations with researchers, founders and policy people about what to do with your own career.

  • Weekly, with people from the frontier labs. Useful for keeping up with what comes out.

Books

  • The Alignment Problem

    Brian Christian

    An accessible account of why it is hard to align a system with what we actually want.

  • Human Compatible

    Stuart Russell

    The case, from one of the authors of the classic AI textbook, for rebuilding the field around uncertainty.

  • Superintelligence

    Nick Bostrom

    The book that brought the discussion of general AI into public debate. It is from 2014 and it shows, but it fixed the vocabulary.

Courses

  • AGI Strategy

    BlueDot Impact

    Twenty-five hours in groups of eight, taking a threat apart to see where it is worth intervening.

  • Thirty hours across six units. Each one leaves something finished: a brief, a map of power, a roadmap.

  • Technical AI Safety

    BlueDot Impact

    Thirty hours of alignment, interpretability, evaluations and control. It opens the door to the project sprint.

  • AGI Safety Course

    Google DeepMind

    Seventy-five recorded minutes on the two ways an objective goes wrong.

  • AI Safety, Ethics and Society

    Center for AI Safety

    Eight chapters, from catastrophic risks to governance. Free as a book, an audiobook and a course.

  • ARENA

    In London, for people who already code

    Four or five weeks in person in London, writing code: transformers, reinforcement learning and evaluations.

If you want to see the whole field

aisafety.com is an open directory with hundreds of organisations, programmes and projects, and with who funds each one.

They go further with company

Go through them with someone else

The reading group discusses one every fortnight, and the WhatsApp group is where you ask what does not add up.