معرفی
Stuart Russell is a Professor at the University of California, Berkeley, affiliated with the Department of Electrical Engineering and Computer Sciences in the College of Engineering. His research spans Artificial Intelligence, Reinforcement Learning, and AI Ethics, focusing on safety, alignment, and theoretical foundations.
- Research Interests:
- Provably beneficial AI systems
- Human-compatible AI
- Reinforcement learning theory
- Language model interpretability
- AI safety and control mechanisms
- Recent Work Trends:
- 2025-2024: Safety frameworks (off-switch games, reversal curse analysis)
- 2024: Reward learning with partial observability, program synthesis via diffusion
- 2023: Societal risk taxonomy, human feedback coordination
Collaborations include leading researchers like Anca D. Dragan (UC Berkeley), Yoshua Bengio (Mila), and Michael I. Jordan (UC Berkeley), with publications in venues such as NeurIPS, ICLR, and ICML.
۰مقاله ثبتشده




