Shivaram Kalyanakrishnan is an Associate Professor at the Department of Computer Science and Engineering , Indian Institute of Technology Bombay , specialising in Artificial Intelligence and Machine Learning . His research spans sequential decision making , multiagent learning , multi-armed bandits , and humanoid robotics , with applications in robot soccer , computer games , and online advertising . He teaches advanced courses like CS 747: Foundations of Intelligent and Learning Agents and CS 748: Advances in Intelligent and Learning Agents , focusing on end-to-end system design and theoretical analysis. His scientific awards include the Best Student Paper Award at RoboCup International Symposium 2006 and nomination for Best Student Paper Award at AAMAS 2007 . His work on reinforcement learning and policy iteration has been published in leading venues such as IJCAI , ICML , and COLT , with recent contributions to railway scheduling and bandit algorithms. While no explicit list of advisees is provided, his research projects and publications suggest mentorship of students in collaborative efforts. Contact : shivaram@cse.iitb.ac.in .
Kevin Jamieson is an Associate Professor at the Paul G. Allen School of Computer Science & Engineering and an Adjunct Professor in the Department of Statistics at the University of Washington . His academic journey includes a B.S. (2009) , M.S. (2010) , and Ph.D. (2015) in electrical engineering from the University of Washington, Columbia University, and University of Wisconsin–Madison respectively. He completed a postdoc at UC Berkeley's AMP Lab before joining UW in 2017. Ph.D., Electrical Engineering, University of Wisconsin–Madison (2015) M.S., Electrical Engineering, Columbia University (2010) B.S., Electrical Engineering, University of Washington (2009) Jamieson's research lies at the intersection of interactive machine learning , active learning , and sequential decision making . His work focuses on: Adaptive sampling strategies in multi-armed bandits and reinforcement learning (RL) Developing instance-dependent optimal algorithms that adapt to problem difficulty Applications in robotics , human perception studies , and hyperparameter optimization Representation learning for large models and experimental design frameworks His 15 most recent publications (2025-2022) demonstrate expertise in bandit theory , contextual RL , and game-theoretic learning . Notable trends include sample-efficient optimization , adaptive A/B testing , and sim-to-real transfer in robotics. Jamieson has received: NSF CAREER award for foundational contributions Amazon Faculty Research award for innovation in learning systems He actively recruits graduate students and postdocs , emphasizing collaboration in areas like: Multi-agent RL and strategic actor learning Empirical process suprema and adaptive sampling theory Applications in robotics , large language model finetuning , and biomedical data analysis Jamieson leads the Washington AI Lab (WAIL) and develops open-source learning systems like the NEXT framework for real-world adaptive data collection. He serves as co-PI for the Institute for the Foundations of Data Science (IFDS) and co-organizes the Distinguished Seminar in Optimization & Data .
Benjamin Van Roy is a Professor at Stanford University since 1998, affiliated with the Departments of Electrical Engineering and Management Science and Engineering, and the Institute for Computational and Mathematical Engineering. He leads the Efficient Agent Team at Google DeepMind and previously held leadership roles at Morgan Stanley, Unica, and Enuvis. He holds SB, SM, and PhD degrees in Computer Science and Electrical Engineering from MIT, advised by John Tsitsiklis. His research focuses on reinforcement learning, alignment, and information theory, with contributions to machine learning foundations, optimization, and finance. He has authored over 150 publications, including influential works on Thompson Sampling, approximate dynamic programming, and exploration strategies. His honors include INFORMS and IEEE Fellowships and the INFORMS Lanchester Prize. Van Roy advises doctoral students across academia and industry, with graduates at top institutions and companies like Meta, Tesla, and Citadel. He teaches courses on reinforcement learning, stochastic control, and optimization. His open-source projects include Epistemic Neural Networks and the Neural Testbed for evaluating machine learning models. Key contributions include foundational work in reinforcement learning theory, scalable methods for recommendation systems, and applications in finance and resource allocation. His research bridges theoretical insights with practical applications, emphasizing alignment and safety of AI systems.
Cicek Cavdar is an Associate Professor at the School of Electrical Engineering and Computer Science (EECS) at KTH Royal Institute of Technology , Sweden. She leads the Intelligent Network Systems research group and specializes in Telecommunication Networks , with a focus on Beyond 5G/6G Mobile Networks , Energy Efficiency , and AI-Assisted Network Management . PhD in Computer Science (2009) from University of California, Davis and Istanbul Technical University Her research spans Cell-Free Massive MIMO , Reconfigurable Intelligent Surfaces (RIS) , UAV Communication Systems , and Green Network Technologies . She actively contributes to 6G Network Architecture and Non-Terrestrial Networks , including satellite and aerial systems. Recent publications highlight AI-driven network optimization for handover management, energy-aware resource allocation , and multi-agent reinforcement learning in complex communication environments. She teaches advanced courses in Communication Systems , Machine Learning , and Software Engineering at KTH.
Rochester Institute of Technology (RIT)United States
Daniel Krutz is an Associate Professor in the Department of Software Engineering at Rochester Institute of Technology (RIT) , with a secondary appointment as Courtesy Associate Professor at the University of Florida's Department of Computer & Information Science & Engineering. He directs the AWARE Lab , focusing on Autonomous Systems, Warfare Applications, and Software Engineering , supported by $6M+ in federal grants from NSF, NSA, DOD, and AFRL. His research spans AI/ML, Self-Adaptive Systems, Decision Support Systems, and Accessibility in Computing Education . He holds a PhD from Nova Southeastern University and has taught courses including Software Engineering Freshman Seminar and Capstone Research Project . He is an NSF CAREER Award recipient (2022) and former AFRL Fellow (2018). Educations: BS, St. John Fisher College MS, Rochester Institute of Technology PhD, Nova Southeastern University Labs: AWARE Lab (Autonomy, Warfare, and Engineering) Awards: NSF CAREER Award (2022) AFRL Research Faculty Fellowship (2018) Teaching: Undergraduate/Graduate courses in software engineering, accessibility, and AI/ML Research Interests: Focus on experiential learning in computing education, neuroevolutionary algorithms, and AI ethics. Recent projects address empathy-building interventions for inclusive software development and context-aware decision support systems . Grants & Funding: Over $6M secured since 2018 from NSF, DOD, and NSA for projects in AI/ML, cybersecurity, and adaptive systems. Current work explores AI-driven stock trading frameworks and accessibility education modules.
Bo Dai is an Assistant Professor at the School of Computational Science and Engineering, Georgia Institute of Technology, and a Staff Research Scientist at Google DeepMind. His research focuses on Agent AI, Generative Models, and Representation Learning, aiming to create decision-making agents through world modeling. He holds a Ph.D. from Georgia Tech (2013–2018) and previously worked at Google Brain. Dai has authored numerous influential papers in top conferences like NeurIPS, ICML, and ICLR, and received the AISTATS Best Paper Award (2016). Education: Ph.D., School of Computational Science and Engineering, Georgia Tech (2013–2018) Research Interests: Reinforcement Learning and Representation Learning for decision-making agents Generative Models and their integration with Agent AI Provable and scalable algorithms for real-world applications His work bridges theory and practice, emphasizing spectral representations and provable guarantees in complex systems. Key Article Trends: His recent work emphasizes scalable spectral methods for multi-agent systems, diffusion policies, and representation-based techniques in reinforcement learning. He also explores LLM alignment and sim-to-real transfer learning. Awards: AISTATS Best Paper Award (2016) NeurIPS Workshop Best Paper (2017) Ross Fellowship (2011–2012) Advising & Grants: Supervises 6 current Ph.D. and M.S. students. Active in organizing workshops on reinforcement learning and serves as an Area Chair for top conferences. Labs & Software: Leads development of Representation-based Reinforcement Learning and Repr-Control toolboxes for nonlinear control and stochastic systems.
Osman Yağan is a Research Professor in the Department of Electrical and Computer Engineering at Carnegie Mellon University (CMU), with affiliate faculty status in the School of Computer Science. He is also a core member of CyLab Security and Privacy Institute. Prior to joining CMU in 2013, he was a Postdoctoral Research Fellow at CyLab. He holds a Ph.D. in Electrical and Computer Engineering from the University of Maryland (2011) and a B.S. from Middle East Technical University (2007). His research focuses on modeling, analysis, and optimization of computing systems, leveraging applied probability, network science, data science, and machine learning. Key areas include multi-armed bandits, resilient machine learning, contagion processes in networks, and cybersecurity. Research Interests include Machine Learning, Data Science, Network Science, Cybersecurity, and Robustness in Cyber-Physical Systems. He has advised numerous students, including current PhD candidates Yurun Tian, Orkun İrsoy, and Ishank Juneja, as well as notable alumni such as Mansi Sood (now at MIT) and Jun Zhao (Assistant Professor at Nanyang Technological University). Key Awards include the CIT Dean's Early Career Fellowship, IBM Academic Award, and Best Paper Awards at ICC 2021, IPSN 2022, and ASONAM 2023. His work spans theoretical contributions (e.g., contagion models in multi-layer networks) and applied research (e.g., mitigating cascading failures in power systems). He leads or co-leads grants from ONR, NSF, and ARO, focusing on resilient machine learning, network robustness, and pandemic modeling.
Dai Zhongxiang is an Assistant Professor and Presidential Young Fellow at the School of Data Science, The Chinese University of Hong Kong, Shenzhen (CUHKSZ), where he joined in August 2024. Previously, he was a Postdoctoral Associate at MIT's Laboratory for Information and Decision Systems (January-June 2024) and a Postdoctoral Fellow at the National University of Singapore's Department of Computer Science (April 2021-December 2023). He completed his Ph.D. in Artificial Intelligence at NUS under the supervision of Bryan Kian Hsiang Low and Patrick Jaillet. Dr. Dai's research focuses on the intersection of theoretical and practical AI, with particular emphasis on large language models (LLMs) and optimization techniques. His work spans both theoretical foundations of multi-armed bandits and Bayesian optimization, as well as practical applications in LLM inference, including prompt optimization, in-context learning, personalization of LLMs, LLM-based agents, and scaling up test-time computation of LLMs. His research approach often bridges theoretical principles with real-world applications, particularly in AI4Science problems. His recent publications demonstrate a clear trend toward advancing LLM capabilities through optimization techniques, with increasing focus on practical deployment challenges. The research spans both theoretical contributions to optimization theory and applied work on enhancing LLM performance in real-world scenarios. His work on dueling bandits, neural bandits, and zeroth-order optimization has been consistently published in top-tier venues including NeurIPS, ICML, ICLR, and ACL. Presidential Young Fellow, CUHKSZ (2024) Dean's Graduate Research Excellence Award, NUS (2021) Research Achievement Award × 2, NUS (2019 & 2020) Singapore-MIT Alliance Graduate Fellowship (2017) Dr. Dai actively mentors multiple Ph.D. students and research assistants, with several of his students' papers accepted to top conferences. His research has received significant attention, with invitations to serve as Area Chair for NeurIPS 2025 and ICLR 2025, reflecting his growing influence in the machine learning community. His work bridges theoretical machine learning with practical applications in large-scale AI systems.
Fei Fang is an Associate Professor in the Software and Societal Systems Department at Carnegie Mellon University (CMU) , where she explores the intersection of artificial intelligence and multi-agent systems . Her work integrates machine learning with game theory to address challenges in security , sustainability , and mobility , aligning with the AI for Social Good mission. Ph.D. in Computer Science, University of Southern California (2016) B.Eng. in Electronic Engineering, Tsinghua University (2011) Recent research focuses on reinforcement learning , large language models (LLMs) , and human-AI collaboration . Her team’s work has been recognized with 15+ awards across prestigious venues like IAAI, AAAI, and IJCAI. Notable accolades include the 2023 Allen Newell Award , 2022 Sloan Fellowship , and NSF CAREER Award (2021) . She actively contributes to educational initiatives , including teaching "Demystifying AI for Everyone" at CMU, and has sought part-time teaching assistants for course development. Her research spans 15+ domains , including AI ethics , cyber defense , traffic optimization , and public health .
I-Hong Hou is a Professor in the Department of Electrical and Computer Engineering at Texas A&M University, part of the College of Engineering. He holds a B.S. in Electrical Engineering from National Taiwan University (2004), and M.S./Ph.D. in Computer Science from the University of Illinois, Urbana-Champaign (2008/2011). His research focuses on wireless networks, cloud/edge computing, and machine learning with notable contributions to real-time systems and network optimization. Education : B.S., Electrical Engineering, National Taiwan University, 2004 M.S., Computer Science, University of Illinois at Urbana-Champaign, 2008 Ph.D., Computer Science, University of Illinois at Urbana-Champaign, 2011 Research Highlights : Hou’s work emphasizes Age of Information (AoI) , distributed learning, and scheduling algorithms for edge computing. He has pioneered frameworks integrating machine learning with network protocols, such as deep reinforcement learning for restless bandits and second-order optimization for wireless systems. His methods address real-time communication challenges in multi-hop networks and dynamic environments. Awards : Best Paper Awards at ACM MobiHoc (2017, 2020) Best Student Paper, WiOpt 2017 C.W. Gear Outstanding Graduate Student Award, UIUC Advising & Grants : Advised PhD student Siqi Fan (graduated 2024). His research has been supported by grants exploring edge-cloud reconfiguration, real-time video delivery, and neural Whittle index networks. Recent work includes optimizing freshness of information in multi-user systems and developing threshold-optimal policies for complex decision-making. Labs/Teams : Leads the Computer Engineering and Systems Group (CESG) at Texas A&M, collaborating on projects blending networking, machine learning, and distributed systems.
Danica Kragic is a Professor of Computer Science at the School of Electrical Engineering and Computer Science at the Royal Institute of Technology (KTH) in Stockholm, Sweden. She serves as the Director of the Centre for Autonomous Systems and leads the Robotics, Perception and Learning Lab at KTH. Her research focuses on advancing robotics capabilities through computer vision and machine learning approaches. MSc in Mechanical Engineering from the Technical University of Rijeka, Croatia (1995) PhD in Computer Science from KTH (2001) Professor Kragic's research primarily centers on robotics, computer vision, and machine learning, with particular emphasis on robotic manipulation, grasp planning, and human-robot interaction. Her work bridges theoretical foundations with practical applications, exploring how robots can understand and interact with objects in complex environments. She investigates how visual and tactile sensing can be integrated to improve robotic perception and manipulation capabilities, with applications ranging from industrial automation to assistive robotics. Her recent publications demonstrate a strong focus on advanced grasp planning techniques, tactile sensing for manipulation, and mathematical representations for robotic control. Kragic's research shows increasing integration of machine learning approaches with traditional robotics frameworks, particularly in the areas of grasp synthesis, object recognition, and human-robot collaboration. Her work spans theoretical contributions in mathematical representations of grasps to practical implementations of robotic systems capable of adapting to novel objects and situations. 2007 IEEE Robotics and Automation Society Early Academic Career Award IEEE Fellow ERC Starting Grant (2012) Member of The Royal Swedish Academy of Sciences Member of The Royal Swedish Academy of Engineering Sciences Honorary Doctorate from Lappeenranta University of Technology Professor Kragic's research has been supported by major funding bodies including the EU, Knut and Alice Wallenberg Foundation, Swedish Foundation for Strategic Research, and Swedish Research Council. While specific student names aren't listed in the provided information, her publication record suggests extensive mentorship of PhD students and postdoctoral researchers in robotics and computer vision. Her lab, the Robotics, Perception and Learning Lab, serves as a hub for interdisciplinary research connecting computer science, engineering, and cognitive science perspectives on robotic systems. As Director of the Centre for Autonomous Systems at KTH, Kragic oversees a major research initiative focused on advancing autonomous technologies. Her Robotics, Perception and Learning Lab brings together researchers working on visual perception, machine learning, and robotic manipulation, with particular emphasis on developing systems that can understand and interact with objects in unstructured environments. The lab's work spans theoretical foundations of robotic manipulation to practical implementations of systems capable of learning from experience.
Shipra Agrawal is an Associate Professor at the Department of Industrial Engineering and Operations Research, Columbia University, with affiliations to the Data Science Institute and the Department of Computer Science. Her research bridges optimization and machine learning, focusing on decision-making in uncertain environments. PhD in Computer Science from Stanford University (2011) Researcher at Microsoft Research India (2011–2015) Her work addresses online optimization , reinforcement learning , and game theory , aiming to develop algorithms that balance exploration and exploitation for long-term goals. Applications include internet advertising , revenue management , and resource allocation . Recent publications examine dynamic pricing models, regret bounds in reinforcement learning, and convex knapsack optimization. Her research has been supported by NSF CAREER , Google Faculty Research , and Amazon Research Awards . NSF CAREER Award CMMI-1846792 (2019) Google Faculty Research Award (2017) Amazon Research Award (2017) She has advised PhD students who now hold positions at institutions like Google DeepMind, Amazon, and Facebook. Agrawal serves as an associate editor for Management Science , INFORMS Journal on Optimization , and Journal of Machine Learning Research , and co-chaired major conferences such as COLT 2024 and AISTATS 2025.
Theo Damoulas is a Professor of Machine Learning at the University of Warwick with a joint appointment in the Department of Computer Science and Statistics. He is a Turing AI Fellow (2021-2026) through UK Research and Innovation, an ELLIS member, and a Visiting Professor at New York University's Center for Urban Science and Progress (CUSP). He founded and leads the Warwick Machine Learning Group and has directed major projects at The Alan Turing Institute including Project Odysseus and the London Air Quality project. Education includes: PhD in Probabilistic Multiple Kernel Learning (University of Glasgow, 2009) MSc in Informatics (Distinction, University of Edinburgh, 2004) MEng in Mechanical Engineering (1st Class, University of Manchester, 2003) His research focuses on probabilistic machine learning and Bayesian statistics, emphasizing the integration of structural priors, spatiotemporal dependencies, physical laws, and causal relationships. Key applications include Digital Twins, urban science, and computational sustainability. His work advances robust and scalable inference methodologies for complex real-world systems. Publications demonstrate strong emphasis on Bayesian methods, spatiotemporal modeling, and uncertainty quantification, with applications spanning battery modeling, urban mobility, federated learning, and causal inference. Recent work shows increased focus on physics-informed models, federated learning frameworks, and causal abstraction techniques. Major scientific awards: Turing AI Acceleration Fellowship (2021-2026) Best Paper Awards (Wilkes 2024, AISTATS 2022, IEEE ICMLA 2010) ACM SIGMOD Most Reproducible Paper (2017) Dissertation Award (Classification Society 2012) Teaching Excellence nominations (Warwick 2015-2017) He actively advises PhD students and secured significant grants including the £multi-million Turing AI Fellowship. Current doctoral researchers investigate federated learning, causal inference, and spatiotemporal modeling. He leads the Warwick Machine Learning Group, a cross-departmental team developing foundational ML methods for scientific and societal challenges.
Manjesh Kumar Hanawal is an Associate Professor at the Industrial Engineering and Operations Research (IEOR) center of IIT Bombay , India. His academic journey includes a Ph.D. from University of Avignon/INRIA (2013), M.Sc (Engg) from IISc Bangalore (2009), and B.E. from NIT Bhopal (2004). Pre-Ph.D. work: Scientist-B at DRDO's CAIR Postdoctoral: Boston University (2013-2015) Appointed as first Professor-In-Charge of TCA2I center Research focuses on Machine Learning algorithms for limited feedback environments, Communication Networks resource allocation, and Cybersecurity threat detection. Publications span top venues like IEEE Transactions, NeurIPS, INFOCOM, and AISTATS. Recent work trends include: Bandit algorithms for distributed learning in heterogeneous networks Contextual information integration in sequential selection Energy efficiency optimization in wireless sensor networks Net neutrality violation detection frameworks Anti-jamming countermeasures in cognitive networks Scientific recognition includes the SERB Early Career Research Award (2019-2022) for machine learning applications in wireless networks. Advisees include Ph.D. awardee Arun Verma and Best Masters Thesis Awardee Sayan Chatterjee.
Sharan Vaswani is an Assistant Professor in the School of Computing Science at Simon Fraser University (SFU). His research focuses on designing algorithms for sequential decision-making under uncertainty, stochastic optimization, and their interplay with machine learning generalization. He holds a PhD from the University of British Columbia (2019) and postdoctoral experiences at the University of Alberta and Mila. His academic journey includes MSc (UBC, 2015) and BTech (BITS Pilani, 2012) degrees. Education: PhD (UBC, 2019), MSc (UBC, 2015), BTech (BITS Pilani, 2012) Postdoctoral Work: University of Alberta (2020-2021), Mila (2019-2020) Teaching includes courses on Probability and Computing (CMPT 210), Optimization for Machine Learning (CMPT 409/981), and Theoretical Foundations of Reinforcement Learning (CMPT 419/983). His research group focuses on developing scalable optimization algorithms with theoretical guarantees. He advises multiple PhD and MSc students, contributing to areas like constrained MDPs, adaptive learning rates, and reinforcement learning theory. Research Highlights: Contributions to bandit algorithms, stochastic gradient methods, and reinforcement learning theory. Notable work includes global convergence analysis of policy gradients and variance-reduced optimization frameworks.