
معرفی
Fazl Barez is a Senior Research Fellow at the University of Oxford leading research on Technical AI Safety and Governance. He is also affiliated with Cambridge's CSER, NTU's Digital Trust Centre, Edinburgh's Informatics, and is a member of ELLIS. Previously, he was a researcher at Amazon and Huawei, and Co-director and Head of Research at Apart Research. He currently serves as an advisor to Martian and has worked with Anthropic's Alignment team (2024-2025).
- University of Oxford: Senior Research Fellow
- Cambridge CSER: Affiliate
- NTU Digital Trust Centre: Affiliate
- Edinburgh Informatics: Affiliate
- ELLIS: Member
- Anthropic: Alignment Team Collaborator (2024-2025)
- Martian: Advisor
Dr. Barez's research focuses on ensuring AI systems remain safe, interpretable, and beneficial as they grow in capability. His work spans four interconnected areas: Interpretability (developing methods to reveal how AI models process information internally), Safety and Alignment (creating tools to detect and address deceptive behaviors), Technical Governance (translating technical insights into governance frameworks), and Societal Impact (examining broader implications of AI on society). His research is funded by OpenAI, Anthropic, Schmidt Sciences, Future of Life Institute, and NVIDIA.
The trends in Dr. Barez's publications show a consistent focus on making AI systems more transparent and safer. His recent work explores mechanistic interpretability techniques like sparse autoencoders, investigates how language models relearn removed concepts, examines machine unlearning for safety applications, and develops frameworks for value alignment measurement. His publications appear in top venues including NeurIPS, ICML, ICLR, ACL, and EMNLP, reflecting his significant contributions to both theoretical and practical aspects of AI safety.
- Future of Humanity Institute PhD Affiliate (2022-2024)
- EPSRC PhD Student Scholarship (2019-2023)
- MSc Scholarship (2017-2018)
- BA (Hons) Sports Performance Scholarship (2013-2017)
Dr. Barez has mentored numerous students who have gone on to prominent positions at organizations like Microsoft Research, DeepMind, and Martian. His research is generously funded by major AI organizations including OpenAI, Anthropic, Schmidt Sciences, Future of Life Institute, and NVIDIA. He has served as an Area Chair for ACL 2025 and on program committees for major conferences including ECAI 2024. His work has practical impact, with algorithms like N2G adopted by OpenAI to evaluate sparse autoencoders for interpretability.
Dr. Barez leads research at the intersection of technical AI safety and governance. His work connects with multiple research groups including the UK AI Security Institute, Alan Turing Institute, and various university centers. He has co-organized workshops such as the first Mechanistic Interpretability workshop at ICML 2024 and actively collaborates with researchers across the AI safety ecosystem. His research bridges the gap between theoretical safety research and practical implementation in real-world AI systems.
Fazl Barez در سایتهای دیگر
جستوجوهای مرتبط
شاید اینها هم برایتان مناسب باشند
Mario GiulianelliSwiss Federal Institute of Technology in Lausanne · دانشیار
Thomas FelHarvard University · پژوهشگر ارشد- AAlexei EfrosUniversity of Quebec · استاد
Michael HahnSwiss Federal Institute of Technology in Lausanne · استاد
Yu MengUniversity of Virginia · استادیار- SSandro PezzellePompeu Fabra University · استادیار