I stopped maintaining this website in September 2025. Work on what AI companies should do in terms of safety and what they are doing is now being done by METR, Guidelight, Midas, and others.

DeepSeek

Risk assessment 0%

0%
Evals: domains, quality, elicitation
0%
Evals: accountability
0%
Adversarial evaluation for alignment
0%
Model organisms

Evals: domains, quality, elicitation

0%
Click to show details/rubric

Evals: accountability

0%
Click to show details/rubric

Adversarial evaluation for alignment

0%
Click to show details/rubric

Model organisms

0%
Click to show details/rubric