Organizations

A list of organizations conducting or supporting mathematically relevant AI safety research—from theory and interpretability to evaluations and control.

Iliad

Research and field-building for AI safety as a theoretical and experimental science. Iliad incubates research bets and runs training programs and conferences.

  • Research incubation
  • Training
  • Field-building

METR

Develops scientific evaluations of frontier-model autonomy and capabilities that could contribute to catastrophic risk.

  • Frontier evaluations
  • Autonomy
  • Risk measurement

Redwood Research

Works on strategic deception, AI control, threat assessment, and mitigations for risks from advanced AI systems.

  • AI control
  • Deception
  • Threat assessment

Resolution

A nonprofit scaling a portfolio of theoretical and empirical alignment research, with an emphasis on automation and higher-confidence approaches.

  • Alignment research
  • Automation
  • Research support

Timaeus (now part of Resolution)

Applies singular learning theory to interpretability and alignment, including developmental interpretability and susceptibility-based spectroscopy.

  • Singular learning theory
  • Interpretability
  • Learning dynamics

UK AI Security Institute

Government technical research on frontier evaluations, risk mitigations, alignment, control, and AI security.

  • Evaluations
  • Mitigations
  • AI security