Introductory Technical AI Safety Fellowship

Every semester and summer, AISST runs a 9-week introductory reading group on AI safety, covering topics like:

  • Expert predictions on AI progress, and when we can expect transformative impacts

  • How we train cutting edge language models, and potential safety failure modes

  • Why having the right training protocol might not be enough to guarantee the behavior we want from AI models

  • The cutting-edge research into the inner workings of neural networks

  • The current plans to make human-and-beyond-level AI safe, and how effective they are on current AI systems

See here for the curriculum from this past spring.

Applications for the Technical Fellowship are open. Apply here by Sunday, September 13 at noon ET.

We also have a Policy Fellowship for those interested in AI policy or governance and a new AI x Biorisk Fellowship for those interested in the life sciences or AI-driven biorisks. It is possible to participate in multiple fellowships.

Joint AISST and MAIA workshops, where members and intro fellows discussed AI alignment and interacted with researchers from Redwood Research, OpenAI, Anthropic, and more. Learn more here.