Amir-massoud Farahmand is an Associate Professor at the Polytechnique Montréal (Department of Computer and Software Engineering) and a Status-Only Associate Professor at the University of Toronto (Department of Computer Science). He is also a Core Academic Member at Mila (Quebec AI Institute). His research focuses on computational and statistical mechanisms for designing efficient reinforcement learning (RL) agents and adaptive algorithms. Dr. Farahmand's research spans reinforcement learning, optimal transport, adversarial robustness, and model-based methods. He has extensively studied regularization in RL, distributional approaches, and algorithm design for stability and convergence. His textbook Lecture Notes on Reinforcement Learning (2021) emphasizes mathematical intuition over algorithmic collections. Recent publications highlight trends in high-update-ratio RL, distributional equivalence, and self-prediction for task understanding. He is actively involved in teaching, having previously instructed courses on machine learning, neural networks, and RL at the University of Toronto. Scientific Awards : Ontario Early Researcher Award (2024) for Accelerated Reinforcement Learning Algorithms Dr. Farahmand has mentored numerous students, including his first PhD graduate Yangchen Pan (now at Oxford) and MSc students like Allen Bao (AMD) and Farnam Mansouri (University of Waterloo). He is currently recruiting graduate students at Polytechnique Montréal and Mila for 2025 admissions.










