Shivaram Kalyanakrishnan is an Associate Professor at the Department of Computer Science and Engineering , Indian Institute of Technology Bombay , specialising in Artificial Intelligence and Machine Learning . His research spans sequential decision making , multiagent learning , multi-armed bandits , and humanoid robotics , with applications in robot soccer , computer games , and online advertising . He teaches advanced courses like CS 747: Foundations of Intelligent and Learning Agents and CS 748: Advances in Intelligent and Learning Agents , focusing on end-to-end system design and theoretical analysis. His scientific awards include the Best Student Paper Award at RoboCup International Symposium 2006 and nomination for Best Student Paper Award at AAMAS 2007 . His work on reinforcement learning and policy iteration has been published in leading venues such as IJCAI , ICML , and COLT , with recent contributions to railway scheduling and bandit algorithms. While no explicit list of advisees is provided, his research projects and publications suggest mentorship of students in collaborative efforts. Contact : shivaram@cse.iitb.ac.in .
Stephen Licht is an Associate Professor of Ocean Engineering and Graduate Director at the University of Rhode Island's College of Engineering, where he directs the Robotics Laboratory for Complex Underwater Environments (R-CUE). His research focuses on developing maritime robots capable of operating in dynamic and unpredictable environments through biologically inspired propulsion, distributed pressure sensing, model-based optimal control, and compliant underwater manipulation technologies. Ph.D. in Oceanographic and Mechanical Engineering from MIT/WHOI Joint Program (2008) B.S. in Mechanical Engineering from Yale University (1998) Former Senior Research Scientist at iRobot and Senior Robotics Engineer at Vecna Robotics Current Research Affiliate with MIT Department of Mechanical Engineering Former Visiting Faculty at Libera Università di Bolzano (2019-2020) Dr. Licht's research spans marine robotics with emphasis on biologically inspired propulsion systems that provide high authority and bandwidth thrust, nonlinear attitude control for maneuvering in dynamic conditions, compliant underwater manipulation technologies, and unmanned aerial monitoring of coastal structures. His work bridges mechanical engineering principles with oceanographic applications to create more capable underwater robotic systems that can operate in complex marine environments. His recent publications demonstrate a strong trend toward soft robotics applications for deep-sea exploration, with particular focus on jamming grippers and neutrally buoyant manipulation systems. The research also shows increasing integration of additive manufacturing techniques for field-deployable solutions and computational methods for autonomous systems operating in challenging marine environments. His work spans fundamental control theory, mechanical design, and practical field applications. Dr. Licht has secured significant research funding as both Principal Investigator and Co-Principal Investigator from major organizations including the Office of Naval Research, NOAA, NSF, and various university collaborations. His grants focus on advancing unmanned underwater vehicle technology, soft robotics for deep-sea applications, and coastal monitoring systems. Active mentor to numerous graduate and undergraduate students in Ocean Engineering Successful track record of student placements at organizations including Jaia Robotics, Scripps Institution of Oceanography, FORSSEA Robotics, and government research labs Collaborates with researchers at MIT, WHOI, University of Connecticut, University of Maine, and international institutions Licht leads the R-CUE lab which develops innovative solutions for underwater robotics challenges, with particular expertise in biomimetic propulsion, soft robotics for deep-sea applications, and autonomous systems for environmental monitoring. The lab maintains strong industry connections with OceanGate Inc. and FabNewport, and engages with local educational institutions through outreach programs with Roger Williams Middle School.
Kevin Jamieson is an Associate Professor at the Paul G. Allen School of Computer Science & Engineering and an Adjunct Professor in the Department of Statistics at the University of Washington . His academic journey includes a B.S. (2009) , M.S. (2010) , and Ph.D. (2015) in electrical engineering from the University of Washington, Columbia University, and University of Wisconsin–Madison respectively. He completed a postdoc at UC Berkeley's AMP Lab before joining UW in 2017. Ph.D., Electrical Engineering, University of Wisconsin–Madison (2015) M.S., Electrical Engineering, Columbia University (2010) B.S., Electrical Engineering, University of Washington (2009) Jamieson's research lies at the intersection of interactive machine learning , active learning , and sequential decision making . His work focuses on: Adaptive sampling strategies in multi-armed bandits and reinforcement learning (RL) Developing instance-dependent optimal algorithms that adapt to problem difficulty Applications in robotics , human perception studies , and hyperparameter optimization Representation learning for large models and experimental design frameworks His 15 most recent publications (2025-2022) demonstrate expertise in bandit theory , contextual RL , and game-theoretic learning . Notable trends include sample-efficient optimization , adaptive A/B testing , and sim-to-real transfer in robotics. Jamieson has received: NSF CAREER award for foundational contributions Amazon Faculty Research award for innovation in learning systems He actively recruits graduate students and postdocs , emphasizing collaboration in areas like: Multi-agent RL and strategic actor learning Empirical process suprema and adaptive sampling theory Applications in robotics , large language model finetuning , and biomedical data analysis Jamieson leads the Washington AI Lab (WAIL) and develops open-source learning systems like the NEXT framework for real-world adaptive data collection. He serves as co-PI for the Institute for the Foundations of Data Science (IFDS) and co-organizes the Distinguished Seminar in Optimization & Data .
Shimon Whiteson is Professor of Computer Science at the University of Oxford, leading the Whiteson Research Lab focused on reinforcement learning, multi-agent systems, and deep learning. His research develops algorithms for efficient learning in complex environments. Current work explores meta-reinforcement learning frameworks that enable agents to rapidly adapt to new tasks, with applications in autonomous driving simulation and robotics. Recent innovations include novel methods for offline reinforcement learning, multi-agent coordination, and morphology-aware control. Publications demonstrate advances in: Meta-RL algorithm design for few-shot adaptation Multi-agent reinforcement learning environments and benchmarks Imitation learning in autonomous driving Bayesian methods for sample-efficient learning Research outputs include widely used benchmarks and tools including JaxMARL for accelerated multi-agent RL research. Current doctoral supervision focuses on temporal abstraction in RL, multi-agent coordination, and reinforcement learning theory.
Sewoong Oh is a Professor at the Paul G. Allen School of Computer Science & Engineering at the University of Washington, where he has been faculty since 2019. His research focuses on the foundations of machine learning with particular emphasis on differential privacy, secure and robust machine learning, and federated learning. Prior to joining UW, he was at the Department of Industrial and Enterprise Systems Engineering at the University of Illinois at Urbana-Champaign from 2012-2019. He is affiliated with multiple NSF AI Institutes including ACTION (Agent-based Cyber Threat Intelligence and Operation), AI-EDGE (Future Edge Networks and Distributed Intelligence), and IFML (Foundations of Machine Learning). Oh received his PhD in Electrical Engineering from Stanford University in 2011 under Andrea Montanari, followed by postdoctoral work at MIT's Laboratory for Information and Decision Systems under Devavrat Shah. His educational background spans top institutions in theoretical computer science and electrical engineering. His research spans the critical intersection of machine learning security, privacy, and robustness. Oh's work addresses fundamental challenges in making AI systems reliable and trustworthy, with particular focus on defending against backdoor attacks, developing privacy-preserving algorithms, and creating efficient tokenization methods for language models. His recent SuperBPE work demonstrates how moving beyond traditional subword tokenization can significantly improve language model efficiency and performance. His research consistently bridges theoretical foundations with practical implementations, as evidenced by numerous open-source repositories containing production-ready code. Oh's publications reveal a consistent trajectory toward addressing security and privacy concerns in increasingly complex machine learning systems, with recent work focusing on language model tokenization, data-centric AI, and federated learning frameworks. His research shows strong interdisciplinary connections between theoretical computer science, statistics, and practical machine learning systems. ACM SIGMETRICS best paper award (2015) NSF CAREER award (2016) ACM SIGMETRICS rising star award (2017) GOOGLE Faculty Research Awards (2017, 2020) 2024 ICML Best Paper Award (for work by advisee Jon Hayase) Professor Oh maintains an active research group with numerous PhD students, postdocs, and undergraduate researchers. His students have secured prestigious positions at leading tech companies including Amazon, Google, and Snap, as well as academic appointments at institutions like University of Wisconsin-Madison and Shanghai Jiao Tong University. He has received substantial research funding through his participation in multiple NSF AI Institutes and industry research awards from Google. His lab maintains several active GitHub repositories implementing cutting-edge research in backdoor defense, robust statistics, and privacy-preserving machine learning.
Hao Su is an Associate Professor in the Department of Computer Science and Engineering at University of California, San Diego . He serves as Chairman & CTO of Hillbot Inc , and leads the SU Lab which focuses on building autonomous systems that learn actively in physical environments. His affiliations include the Institute for Learning-enabled Optimization at Scale , Artificial Intelligence Group , Contextual Robotics Institute , Halicioğlu Data Science Institute , and Center for Visual Computing . As a researcher in Computer Vision, Robotics, and Neural Geometry , he has made significant contributions to 3D foundation models, reward-free world models, diffusion policy frameworks, and GPU-accelerated simulation environments. His 2024-2025 publications include advancements in hand-eye calibration, dynamic mesh reconstruction, and multi-stage robotic manipulation. His scientific awards include: Frontiers of Science Award (2025) TPAMI Young Research Award (2025) NSF CAREER Award (2023) ACM SIGGRAPH Best Doctorate Thesis Honorable Mention (2019) He has served as Program Chair for CVPR 2025 and Area Chair for ICLR 2022 and NeurIPS 2023 , while previously serving as Publication Chair for 3DV 2016 and Program Committee for SIGGRAPH Asia Workshops .
Nadia Figueroa is the Shalini and Rajeev Misra Presidential Assistant Professor in the Mechanical Engineering and Applied Mechanics (MEAM) Department at the University of Pennsylvania. She holds secondary appointments in Computer and Information Science (CIS) and Electrical and Systems Engineering (ESE), and is a core faculty member at the GRASP Lab. Prior to Penn, she was a Postdoctoral Associate at MIT’s CSAIL and earned her PhD in Robotics from EPFL under Prof. Aude Billard. Her research focuses on physical and perceptual adaptive intelligence for robots, enabling fluid collaboration with humans in dynamic environments. Key applications include robot learning from demonstration , human-robot co-manipulation , safe navigation in human-centric spaces , and rehabilitation robotics . Her work integrates machine learning control theory artificial intelligence biomechanics psychology with guarantees of stability, safety, and robustness . Recent publications highlight advancements in reactive collision avoidance dynamical system learning intent estimation EEG-driven assistive control origami-based reconfigurable robots across platforms like autonomous vehicles and humanoid robots. She has authored a 2022 textbook on dynamical systems for robot control and received the Presidential Assistant Professorship at Penn.
Colin Raffel , currently an Associate Professor at the University of Toronto and Associate Research Director at the Vector Institute , is a leading researcher in machine learning and natural language processing . His career spans roles at Hugging Face (Faculty Researcher), Google Brain (Senior Research Scientist), and UNC Chapel Hill (Assistant Professor). Education: PhD in Electrical Engineering (Columbia), MA in Music/Science (Stanford), BA in Mathematics (Oberlin) Key affiliations: Google Brain (2016-2020), Hugging Face (2021-present), Vector Institute (2023-present) His research focuses on language model development , attention mechanisms , efficient machine learning , and music information retrieval . Recent work explores model merging , parameter-efficient fine-tuning , and data-constrained language models . Teaching : Has instructed courses at University of Toronto and UNC Chapel Hill on Neural Networks , Deep Learning , and Information Theory . Academic service includes organizing ICLR workshops and serving as Senior Area Chair for NeurIPS and EMNLP . Notable awards : NSF CAREER (2022), Caspar Bowden Award (2023), NeurIPS Outstanding Paper (2023) Key contributions : Core developer of WT5 , Git-Theta , and mir_eval software
Lerrel Pinto is an Assistant Professor of Computer Science at the Courant Institute of Mathematical Sciences at New York University (NYU), where he leads the General-purpose Robotics and AI Lab (GRAIL) as part of the CILVR research group. His work bridges the gap between theoretical machine learning and practical robotics applications, with a focus on enabling robots to generalize and adapt in real-world environments. Dr. Pinto received his undergraduate degree from IIT Guwahati, followed by a PhD from the Robotics Institute at Carnegie Mellon University (CMU). He then completed a postdoctoral fellowship at the University of California, Berkeley before joining NYU as faculty. His research program centers on robot learning and decision making, with several key thrusts that demonstrate his innovative approach to robotics. Pinto's work emphasizes large-scale learning techniques that leverage both extensive data and sophisticated model architectures. A significant portion of his research focuses on representation learning for sensory data, particularly developing methods that enable robots to make sense of visual, tactile, and auditory inputs. His lab has made notable contributions to reinforcement learning algorithms that allow robots to adapt to new scenarios with minimal retraining. Pinto also champions open-source robotics , developing affordable robot platforms that democratize access to robotics research. Analysis of Pinto's recent publications reveals a strong trend toward multimodal perception in robotics, integrating visual, tactile, and auditory information to create more robust robot systems. His work increasingly focuses on zero-shot and few-shot learning capabilities, enabling robots to handle novel situations without extensive retraining. There's also a clear progression toward general-purpose robotics , moving away from task-specific solutions toward more flexible systems that can handle diverse real-world challenges. Dr. Pinto's scientific contributions have been recognized with several prestigious awards: Sloan Research Fellowship (2025) NSF CAREER Award (2024) RAL Early Career Award (2024) Best Student Paper Award at ICRA (2016) Outstanding Paper Award at MFM-EAI workshop at ICML (2024) Best Paper Award at NGSM workshop at ICML (2024) Best Student Paper Award at RSS (2023) As an advisor, Pinto has mentored numerous students who have gone on to impactful careers in both academia and industry. His former PhD student Denis Yarats co-founded Perplexity.AI, while Mahi Shafiullah became a postdoc at UC Berkeley and Meta AI. Many of his Masters students have pursued PhDs at top institutions like CMU, MIT, and Stanford, or joined leading robotics companies including 1X, Fauna Robotics, and NVIDIA. Pinto's lab has secured significant research funding, including the NSF CAREER award and likely other grants supporting his robotics research program. The General-purpose Robotics and AI Lab (GRAIL) that Pinto leads brings together a diverse team of researchers working on cutting-edge robotics challenges. The lab maintains strong collaborations with industry partners and other academic institutions, facilitating technology transfer and real-world impact. GRAIL's research spans multiple robotics platforms and focuses on developing algorithms that enable robots to learn from diverse experiences and generalize across environments.
Roberto Martin-Martin is an Assistant Professor of Computer Science at the University of Texas at Austin, where he leads the Robot Interactive Intelligence (RobIN) Lab. His research bridges robotics, computer vision, and machine learning to enable robots to operate autonomously in human-centric environments like homes and offices. Previously, he was a Postdoctoral Scholar at the Stanford Vision and Learning Lab working with Fei-Fei Li and Silvio Savarese, and an AI Researcher at Salesforce AI. Education: Ph.D. and M.Sc. in Robotics from Technische Universität Berlin (TUB), advised by Professor Oliver Brock B.Sc. from Universidad Politécnica de Madrid Dr. Martin-Martin's research focuses on developing AI algorithms that combine reinforcement learning and imitation learning with advanced planning and control to address core challenges in robot perception. His work spans mobile and whole-body manipulation, dexterous and contact-rich interactions, and long-horizon tasks in unstructured environments. He takes inspiration from human cognition through psychology and cognitive science to develop solutions for skills ranging from simple pick-and-place operations to complex tasks like cooking and furniture assembly. His recent publications demonstrate a strong trend toward enabling robots to learn from human demonstrations, particularly through video, and to safely adapt these demonstrations to their own morphology. There's significant emphasis on mobile manipulation, bimanual tasks, and developing hardware that supports robust robot learning through trial and error. His work shows increasing integration of large language models and vision-language models to enhance robot understanding and task execution. Scientific Awards: RSS Pioneer (2020) Winner of Amazon Picking Challenge (2015) RSS Best Systems Paper Award (2016) ICRA Best Paper Award IROS Best Mechanism Award Amazon Faculty Award AAAI Young Faculty IJCAI Early Faculty Nominated for Best Paper at IROS (2014, 2017) Dr. Martin-Martin advises PhD students including Arpit Bahety, who is working on mobile manipulation and learning. He serves as Chair of the IEEE Technical Committee on Mobile Manipulation and is a co-founder of QueerInRobotics. His research is supported by industry partnerships and academic funding sources that enable his lab to develop both hardware and software innovations in robotics. He directs the Robot Interactive Intelligence (RobIN) Lab at UT Austin, which takes a holistic approach to robot intelligence, developing both the hardware (like the BaRiFlex gripper) and software frameworks necessary for robots to learn from interaction. The lab's research addresses the full pipeline from perception to action, with particular emphasis on learning from human demonstrations, safe exploration, and adapting to novel objects and environments.
Mark Riedl is a Professor in the Georgia Tech School of Interactive Computing and Associate Director of the Georgia Tech Machine Learning Center (ML@GT). His research focuses on human-centered artificial intelligence, emphasizing the development of AI technologies that naturally interact with humans. Key areas include story understanding/generation, computational creativity, explainable AI, and ensuring AI safety. He holds affiliations with the GVU Center, Institute for People and Technology (IPaT), and Institute for Robotics and Intelligent Machines (IRIM). His work is supported by NSF, DARPA, ONR, and industry partners like Google and Meta. Notable awards include the DARPA Young Faculty Award and NSF CAREER Award, plus three Pulitzer Prizes (collaborative with Roko M. Bask). Riedl's recent projects include STORY2GAME (AI-driven game design) and ethical AI frameworks addressing transparency and accountability. His research bridges theoretical advancements with practical applications in education, healthcare, and creative industries. Research interests span AI ethics, narrative systems, and AI's societal impact. He explores how AI can be made more transparent through explainable mechanisms while maintaining creativity and safety. Collaborations with Roko Bask on futuristic culinary trends have produced influential works. His labs and teams focus on interdisciplinary approaches, combining computer science with social sciences to shape responsible AI development. Key grants and projects include NSF-funded initiatives on AI in education and DARPA-supported work on AI safety. His contributions to explainable AI challenge traditional XAI paradigms, advocating for human-centered approaches that prioritize user understanding and ethical implications. Current efforts emphasize adapting LLMs for world modeling and enhancing RL agents with causal reasoning capabilities.
Anca Dragan is an Associate Professor in the Department of Electrical Engineering and Computer Sciences at the University of California, Berkeley, where she runs the InterACT Lab focused on algorithms for human-AI and human-robot interaction. Currently on leave from Berkeley, she leads AI Safety and Alignment at Google DeepMind, overseeing safety for Gemini models and preparing for future advancements. Dragan has been a co-PI of the Center for Human-Compatible AI and served on the steering committee for the Berkeley AI Research (BAIR) Lab. B.Sc. in Computer Science from Jacobs University Bremen, Germany Ph.D. from Carnegie Mellon University Dr. Dragan's research focuses on enabling AI agents to work effectively with, around, and in support of people. Her work bridges robotics, machine learning, and game theory to create systems that better understand human preferences and coordinate with users. Key areas include AI alignment (ensuring AI does what people actually want), learning reward functions from diverse human feedback forms, and developing algorithms for human-AI collaboration across domains like autonomous vehicles, brain-machine interfaces, and recommender systems. Her research emphasizes maintaining uncertainty about human preferences and accounting for the plurality of human values. Dr. Dragan's recent publications reveal a strong focus on addressing fundamental challenges in AI safety and alignment. Her work spans theoretical foundations of reward learning, practical implementations for human-AI coordination, and critical examinations of limitations in current approaches. There's a clear trajectory toward more robust, safe, and value-aligned AI systems that can handle complex human preferences while avoiding both present-day harms and potential catastrophic risks. IEEE RAS Early Academic Career Award in Robotics and Automation (2021) McEntyre Award for Excellence in Teaching (2020) PECASE (Presidential Early Career Award for Science and Engineering) (2019) Sloan Fellowship (2018) NSF CAREER Award (2017) Okawa Foundation Award (2017) MIT Tech Review 35 Innovators Under 35 (2017) Multiple best paper awards at top robotics and AI conferences Dr. Dragan has mentored numerous successful students who have gone on to faculty positions at MIT, Stanford, CMU, and Princeton, as well as industry roles at DeepMind, Waymo, and Meta. Her advising philosophy emphasizes both technical rigor and consideration of broader societal impacts. She has secured significant research funding including NSF CAREER, ONR Young Investigator, and Okawa Foundation awards, supporting work on human-AI interaction and alignment. Dragan has also consulted for Waymo for six years, helping develop roadmaps for deploying increasingly learning-based safety-critical systems. Dr. Dragan leads the InterACT Lab at UC Berkeley, which has produced influential work on Cooperative Inverse Reinforcement Learning, Inverse Reward Design, and other foundational concepts in human-AI interaction. The lab's research has significantly shaped the field of AI alignment, with applications spanning autonomous vehicles that coordinate with human drivers, brain-machine interfaces that adapt to user needs, and language models that better understand human preferences. Current work focuses on scaling safety approaches as AI capabilities advance, ensuring alignment keeps pace with technological progress.
Fei Fang is an Associate Professor in the Software and Societal Systems Department at Carnegie Mellon University (CMU) , where she explores the intersection of artificial intelligence and multi-agent systems . Her work integrates machine learning with game theory to address challenges in security , sustainability , and mobility , aligning with the AI for Social Good mission. Ph.D. in Computer Science, University of Southern California (2016) B.Eng. in Electronic Engineering, Tsinghua University (2011) Recent research focuses on reinforcement learning , large language models (LLMs) , and human-AI collaboration . Her team’s work has been recognized with 15+ awards across prestigious venues like IAAI, AAAI, and IJCAI. Notable accolades include the 2023 Allen Newell Award , 2022 Sloan Fellowship , and NSF CAREER Award (2021) . She actively contributes to educational initiatives , including teaching "Demystifying AI for Everyone" at CMU, and has sought part-time teaching assistants for course development. Her research spans 15+ domains , including AI ethics , cyber defense , traffic optimization , and public health .
Sebastian Scherer is an Associate Research Professor at the Robotics Institute (RI), Carnegie Mellon University (CMU), where he leads cutting-edge research in autonomous aerial systems and robotics. His work focuses on enabling unmanned rotorcraft to operate safely and efficiently in cluttered, low-altitude, and extreme environments. Education: Ph.D. in Robotics, Carnegie Mellon University (2010) MS in Robotics, Carnegie Mellon University (2007) BS in Computer Science (Minor in Robotics), Carnegie Mellon University (2004) His research interests span robotics, artificial intelligence, autonomous navigation, obstacle avoidance, SLAM, visual-inertial odometry, energy infrastructure, and public policy . He has made seminal contributions to UAV autonomy, including the first obstacle avoidance for micro aerial vehicles in natural environments (2008) and the first automatic landing zone detection and landing on a full-size helicopter (2010). His recent publications (2023–2025) demonstrate a strong focus on resilient autonomy, multi-robot exploration, foundation models for robotics, and large-scale dataset development. His team has released key datasets like TartanGround , BETTY , and SubT-MRS , and simulation tools like Pegasus Simulator , indicating a systems-level approach to advancing real-world autonomy. The research trends emphasize self-supervised learning, robust perception, risk-aware planning, and multi-modal fusion for off-road and urban environments. Scientific Awards: Popular Science Best of What's New 2010 Award AIAA@Infotech Best Paper Runner-up Award (2010) Siebel Scholar Dr. Scherer has advised numerous students and leads a vibrant research group focused on high-impact robotics applications. He has secured significant grants related to UAV autonomy, energy infrastructure, and urban air mobility. His lab develops experimental infrastructure such as AIrTonomy for testing next-generation autonomous aerial vehicles. He is actively involved in advancing SLAM and localization in extreme environments, notably through participation in the DARPA Subterranean Challenge. His team develops large-scale datasets and benchmarking frameworks to push the boundaries of robustness and generalization in mobile robotics.
Karthik R. Narasimhan is a Professor at Princeton University's School of Engineering and Applied Science in the Department of Computer Science. Previously, he earned his PhD from MIT under Regina Barzilay and served as a visiting research scientist at OpenAI during 2017-18. His research focuses on the intersection of language and decision-making, building autonomous agents that learn from both experience and human knowledge. His research spans multiple high-impact areas including language agents (Text-DQN, CALM, ReAct, Tree of Thoughts), reinforcement learning (h-DQN, Multi-Objective RL), and AI safety (Toxicity in ChatGPT, DataMUX). He has developed critical datasets and benchmarks such as WebShop, InterCode, SWE-bench, and SILG that have become standard evaluation tools in the field. Current work emphasizes agent capabilities, software engineering automation, and multimodal interaction. His publication trends show strong focus on practical agent deployment (SWE-agent, Tree of Thoughts), safety evaluation (Probing AI Safety), and efficiency improvements (DataMUX). Recent work increasingly addresses real-world challenges in software engineering, security, and human-AI collaboration through rigorous benchmarking. Co-author of foundational GPT (2018) paper Key developer of Text-DQN (2015), CALM (2020), ReAct (2022), Tree of Thoughts (2023) Creator of influential benchmarks: WebShop (2022), SWE-bench (2023), InterCode (2023) He actively advises students through Princeton's computer science program, with research supported by multiple grants focused on autonomous agent development and language-based decision systems. His GitHub repositories (nlp-datasets, text-world-player) demonstrate strong community engagement in open-source research tools. Current projects include advancing language agent capabilities through Reflexion (2023) and Tree of Thoughts (2023) frameworks while addressing critical safety and efficiency challenges.