Tool or datasetJul 2026For the curious
SaferAI · tracker.safer-ai.org
Rates frontier AI companies' published safety frameworks against established risk-management practice; even the top-rated companies, Anthropic and OpenAI, score only about a third.
Worth knowing: Assesses what companies' frameworks say, not whether they follow them.
Tool or dataset2026Technical
METR · metr.org
METR's index of the safety frameworks published by frontier AI companies, including Anthropic, OpenAI, Google DeepMind, Meta, xAI, Microsoft and Amazon, for comparing what each has committed to.
Worth knowing: The frameworks are written by the companies themselves.
Tool or datasetApr 30, 2024For the curious
Zach Stein-Perlman · AI Lab Watch · ailabwatch.org
A scorecard rating frontier AI companies' safety practices, from risk assessment and security to safety research and planning, with pages on their commitments and integrity incidents.
Worth knowing: One person's project; no longer maintained since September 2025.
Tool or datasetFor the curious
AISafety.com
Directory of the AI safety field: courses, training programmes, communities, events, jobs and funding, for people who want to get involved.
Worth knowing: Framed around preventing human extinction from AI.
Tool or datasetFor the curious
The Collective Intelligence Project · Weval · weval.org
Open platform where experts and communities write tests for AI models and publish the results, including checks on mental-health crisis responses and sycophancy.
Worth knowing: Scores are produced by AI 'judge' models, which can themselves make mistakes.