Since 2007
19 years of continuous editorial operation. Listings audited and re-audited as the directory ages — no stale 2008 records masquerading as live.
Some links on DirJournal are affiliate links. We only recommend verified businesses that clear our 12-point editorial audit standard.
Independent. Human-Curated. Established 2007.
AI safety research organizations and alignment labs working on AI alignment, interpretability, evaluation, and existential risk research focused on ensuring advanced AI systems remain safe and beneficial.
AI Safety Research Labs currently lists 8 entities pending editorial audit.
Verified status arrives after a 12 point audit covering entity ownership, citation consistency across answer engines, schema validity, and operating history.
Ranking weights editorial tier and entity health score across the full listed set. No listings here carry the verification badge yet — they appear in order of editorial signal. Last updated .
Berkeley, United States
Redwood Research is an AI safety research organization focused on technical alignment research and AI control, founded and headquartered in Berkeley, California. Founded by Nate Thomas, Buck Shlegeris, and Bill Zito, Redwood Research conducts empirical alignment research focused on AI control techniques designed to safely deploy AI systems even when alignment is uncertain, mechanistic interpretability research, model organisms research, and AI safety evaluations. The organization publishes influential research on AI control protocols, scalable oversight, and adversarial robustness, partners with major AI labs and the broader AI safety community, and contributes to setting research agendas for ensuring frontier AI systems remain controllable and safe through their deployment lifecycle.
The Machine Intelligence Research Institute (MIRI) is one of the worlds longest-established AI safety research organizations focused on existential risks from advanced AI, founded in 2000 and headquartered in Berkeley, California. Founded by Eliezer Yudkowsky as the Singularity Institute and renamed to MIRI, the organization conducts foundational research on the alignment problem, AI risk theory, and existential risk from artificial general intelligence. MIRI publishes influential books and papers including Eliezer Yudkowskys foundational AI safety writings, Nate Soares research, and runs the MIRI Communications program advocating for international policy responses to existential AI risks. The organization is widely credited with founding the modern AI safety field.
METR (Model Evaluation and Threat Research) is a leading AI safety research organization specializing in dangerous capabilities evaluations of frontier AI models, founded and headquartered in Berkeley, California. Originally launched as the Alignment Research Center Evaluations team led by Beth Barnes, METR partners with major AI labs including OpenAI, Anthropic, and Google DeepMind to conduct pre-deployment evaluations measuring autonomous task completion, agent capabilities, biorisk, cyberrisk, and AI research and development capabilities. METR publishes the influential Time Horizon benchmark tracking how long autonomous AI tasks AI agents can reliably complete, providing critical empirical data informing AI policy and frontier model deployment decisions globally.
FAR AI (Fund for Alignment Research) is a non-profit AI safety research organization, founded and headquartered in Berkeley, California. Founded by Adam Gleave, FAR AI conducts technical AI safety research, fiscal sponsorship of AI safety projects, AI safety field building, and operates the FAR AI labs research program studying neural network robustness, adversarial attacks on AI systems, scalable oversight, and emerging risks from frontier AI models. The organization publishes peer-reviewed safety research, hosts AI safety workshops including the influential AI Risk Forum, supports independent researchers through grants, and contributes empirical safety research to the broader AI safety community building knowledge essential for safer AI development globally.
The Alignment Research Center (ARC) is a non-profit AI alignment research organization founded by former OpenAI alignment lead Paul Christiano and headquartered in Berkeley, California. ARC conducts theoretical and applied research on AI alignment with the mission of ensuring that future advanced AI systems are aligned with human interests. The organization originally housed the influential ARC Evaluations team (which spun off as METR ) and continues research on heuristic arguments, mechanistic anomaly detection, eliciting latent knowledge, and theoretical alignment foundations. Paul Christianos research at ARC including the original ARC Theory work has been highly influential in shaping modern alignment research methodology and theoretical foundations.
Why Trust DirJournal
19 years of continuous editorial operation. Listings audited and re-audited as the directory ages — no stale 2008 records masquerading as live.
Every entry clears a 12-point editorial audit before publication. No pay-to-play ranking — listing position reflects merit, not ad spend.
Schema-validated for ChatGPT, Gemini, and Perplexity. Machine-readable trust signals that AI engines parse and cite with confidence.
Join 8 listed AI Safety Research Labs providers. One payment, lifetime authority.
Secure Your $249.95 Permanent ListingList Your Business
Join 8 listed AI Safety Research Labs providers