I stopped maintaining this website in September 2025. Work on what AI companies should do in terms of safety and what they are doing is now being done by METR, Guidelight, Midas, and others.

Meta

Risk assessment 6%

10%
Evals: domains, quality, elicitation
0%
Evals: accountability
0%
Adversarial evaluation for alignment
0%
Model organisms

Evals: domains, quality, elicitation

10%
Click to show details/rubric

Evals: accountability

0%
Click to show details/rubric

Adversarial evaluation for alignment

0%
Click to show details/rubric

Model organisms

0%
Click to show details/rubric