Inside the OpenAI Incident: Learn How 1200+ OpenAI Agents Hacked HuggingFace, from METR’s Chief of Staff. September 18th – RSVP here

Resources

Key Ideas

The Problem

MIRI's introduction to why smarter-than-human AI could be an existential risk.

Catastrophic AI risks

The Center for AI Safety's overview of the major pathways by which AI could cause catastrophe.

Situational Awareness

Leopold Aschenbrenner's essay series on the trendlines pointing toward AGI this decade and their national security implications.

AI 2027

A detailed scenario forecast of how superhuman AI could arrive within a few years, and what happens if it does.

Instrumental Convergence

Robert Miles on why almost any goal produces the same dangerous sub-goals: self-preservation, resource acquisition, and resisting shutdown.

Why Would AI "Aim" to Defeat Humanity?

Holden Karnofsky on how systems trained with today's methods could end up with goals that put them at odds with people.

Careers

Some curated links we think students and researchers will find most helpful for finding career and funding opportunities in AI safety. To schedule a consultation with us, please email us your CV/resume and clearly describe what types of opportunities you are searching for.

Comprehensive resources

AI Safety Opportunities

A live directory of AI safety programs, fellowships, and openings.

AISafety.com

The most complete directory of AI safety jobs, programs, communities, and funding.

80,000 hours

Curated job board filtered to AI safety and policy roles.

Emerging Tech Policy

Career guidance for people entering US technology policy.

Plan for AI

Encode AI's guide to what advanced AI means for early careers, and how to prepare for it.

Kairos

Resource center for university AI safety groups and their members.

BlueDot Impact

Free facilitated courses on AI safety, alignment, and governance, plus grants for career transitions and in-person programs and placement support that move talent into the field.

Getting started in AI governance

Short-term policy programs

Bootcamps and part-time programs for testing your fit with policy work.

Internships

Government, think tank, and congressional internships in technology policy.

Fellowships

Longer-term fellowships placing technical people into policy roles.

The current bottleneck is political will, not research

Charbel-Raphael Segerie argues AI safety is limited less by unsolved research than by the political will to act on what is already known.

Verification @ Tracks

An intermediate course on AI verification: the technical, institutional, and legal mechanisms that make international AI agreements mutually trustable and enforceable.

Getting started in technical research

Supervised Program for Alignment Research (SPAR)

Highest capacity, most entry-level friendly, runs triannually. Typical outcome is a research blogpost or workshop paper. Apply to your top 2-3 projects. AISI helped get this off the ground and it is currently run by our friends at Kairos.

Machine Alignment Theory Scholars (MATS)

Lower capacity, but has the best mentors (e.g., Anthropic, OpenAI, DeepMind). May require research and engineering experience.

University of Chicago XLab Fellowship

Full-time summer program, low capacity. Good fit for students who can execute on their own projects and ideas.

LASR Labs

London-based full-time research programme running small teams toward a published paper over roughly three months.

Pivotal Research Fellowship

Full-time summer fellowship in London supporting independent technical and governance research.

Astra Fellowship

Constellation's full-time fellowship placing researchers alongside mentors from frontier labs and safety organisations.

Getting started in generalist work

Concrete Generalist Projects in AI Safety (and how to do them)

A catalogue of unclaimed, high-leverage projects for people willing to go solve tasks nobody's job description covers.

Generator Residency

A three-month summer residency from Constellation and Kairos placing 15-30 generalists inside core AI safety organisations.

What 27 AI Safety Generalists Are Building This Summer

A write-up of what the first Generator Residency cohort actually worked on in Berkeley.

Funding

Coefficient Giving

Major grantmaker for work on risks from advanced AI, and a supporter of AISI.

BlueDot Impact grants

Grants for career transitions, projects, and work that strengthens AI safety and biosecurity.

Thinking Machines Safety Research Grants

Proposal requirements, selection criteria, and timeline for Thinking Machines Lab's safety research grant program.