About Goodfire
Goodfire is an AI interpretability research lab. We build interpretability agents and infrastructure to understand, monitor, and align AI models.
We believe that AI is the most consequential technology of our time. Yet AI models are trained by trial and error today, with limited understanding of what drives their behavior. Our goal is to advance the research and technology needed to fully understand and align AI so humanity can trust superintelligence.
Goodfire is a public benefit corporation founded by researchers who helped pioneer the field of interpretability at OpenAI and Google DeepMind. We are headquartered in San Francisco and have raised over $200M from leading investors including Menlo, Lightspeed, and B Capital.
We are a team of researchers, engineers and builders shaping the frontier of AI
Our team includes founding members of interpretability efforts at Google DeepMind and OpenAI, professors on leave, and engineers who have built and deployed large-scale ML systems at organizations like OpenAI, Google, and Palantir.
Many of us helped pioneer core research directions in interpretability—from discovering sparse, human-meaningful neural network features using sparse autoencoders, to automated feature interpretation, to extracting knowledge from superhuman models.





Contact us
Interested in partnering with Goodfire?