NeoLabsAn index of research-first AI labs
Home / Topics / Safety & interpretability
Research area

AI safety and interpretability labs

Mechanistic interpretability, oversight and non-agentic AI, often as a public-benefit company or nonprofit. We track 4 labs in this area, which have raised $0M in disclosed rounds between them; the most valuable is Goodfire, at $1.25B.

4labs
$0Mdisclosed raised
$0Mreported valuations
LabTopicLocationFounders fromFoundedLast roundValuationRaisedStatus
GoodfireMechanistic interpretability as a product (Ember API), with applications in biology and he…site · in Safety & interpretability San Francisco, US ex-Google DeepMind 2024 · $1.25Best. · API
IrregularA frontier AI security lab; its evaluations appear in OpenAI and Anthropic system cards.site · x · in Safety & interpretability Tel Aviv, IL ex-IBM, ex-Google DeepMind 2023 · $450Mest. · Partnerships
LawZeroNon-agentic 'Scientist AI'; a nonprofit backed by $300M from Canada and Germany (Sept 2026…site · in Safety & interpretability Montreal, CA ex-Mila 2025 · · · Research only
TransluceScalable oversight and open interpretability infrastructure; a nonprofit.site · in Safety & interpretability San Francisco, US ex-UC Berkeley, ex-MIT 2024 · · · Research only

News