Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Michael Kaess is an Associate Professor at the Robotics Institute, Carnegie Mellon University (CMU), within the School of Computer Science. He leads the Robot Perception Lab (RPL) and contributes to the Field Robotics Center (FRC) and Computer Vision Group (CV). His research focuses on efficient perception algorithms for mobile robots, particularly in 3D mapping, SLAM, and sensor fusion using vision, LiDAR, inertial, and sonar data. Kaess holds a PhD in Computer Science from Georgia Tech and was a postdoc at MIT's Marine Robotics Lab. Education: Georgia Institute of Technology, PhD in Computer Science (2008) MIT, Postdoctoral Associate (2008–2010) Research Interests: Kaess develops algorithms for robust and efficient inference in robotics, emphasizing factor graphs and linear algebra. His work spans underwater robotics, aerial systems, tactile SLAM, and multi-sensor integration. Key areas include SLAM with planes/lines, imaging sonar reconstruction, and neural field methods for LiDAR-visual fusion. Publications: Over 145 papers, including work on EDPLVO (visual odometry), HoloOcean (underwater simulation), and neural radiance fields with LiDAR. Recent trends focus on robust incremental smoothing, acoustic-optical fusion, and real-time volumetric mapping. Awards: Recognized with the RSS Test of Time Award (2020), Outstanding Associate Editor (2022), and paper awards at ICRA/ICRA. Active in conference organization (IROS/ICRA program committees). Advising & Grants: Supervises 10+ current PhD/MSc students, with past advisees contributing to CoRL/ICRA work. Manages grants in perception, autonomy, and marine robotics. Teaches courses like Robot Localization and Mapping (16-833). Labs/Teams: Directs RPL, collaborates with FRC on field robotics. Develops open-source tools like GTSAM (GNU Toolkit for Smoothing and Mapping).
Katerina Fragkiadaki is the JPMorgan Chase Associate Professor of Computer Science in the Machine Learning Department at Carnegie Mellon University. She works at the intersection of Artificial Intelligence, Computer Vision, Machine Learning, Language Understanding, and Robotics. PhD from GRASP Lab, University of Pennsylvania Postdoctoral researcher at UC Berkeley (with Jitendra Malik) and Google Research Recipient of NSF CAREER, DARPA Young Investigator, Amazon, Google, Sony, UPMC, and AFOSR awards Organizer of CoRL 2023 Workshop on Generalist Robots ICLR 2024 Program Chair, multiple area chair roles Her research group focuses on developing machines that autonomously improve world models through human-environment interactions, with specific emphasis on: Representation learning and video understanding 2D/3D unified vision-language models Generative simulation and reinforcement learning Real2Sim/Sim2Real robot learning Continual learning and spatial common sense 3D scene reconstruction and dynamics Recent publications highlight advancements in: 3D mesh generation with compositional transformers Unified 2D/3D perception frameworks Physics-aware generative models Diffusion-based robotic manipulation policies Embodied agents with memory prompting Awards include: 2024: DARPA Young Investigator Award 2023: Amazon Faculty Award 2022: Sony Faculty Research Award 2021: UPMC Faculty Research Award 2020: NSF CAREER Award 2019: Google Faculty Award Key collaborations span institutions including UC Berkeley, Google Research, Stanford, MIT, and University of Tsukuba. Her work bridges theoretical innovation with practical applications in: Autonomous robot manipulation 4D world modeling Language-grounded perception Visual dynamics prediction Embodied program synthesis Physics-based simulation engines
Olga Sorkine-Hornung is a Professor of Computer Science at ETH Zurich and head of the Institute of Visual Computing. She leads the Interactive Geometry Lab, focusing on theoretical and practical advancements in digital content creation, geometry processing, and shape modeling. Current position: ETH Zurich, Department of Computer Science Previous roles: Courant Institute (NYU), Technical University of Berlin Education: BSc and PhD from Tel Aviv University, postdoc at TU Berlin Her research spans shape representation, digital fabrication, computer animation, and fundamental geometry processing. Key contributions include Laplacian surface editing, as-rigid-as-possible deformation, and generalized winding numbers. She works on applications in VR, AR, and autonomous systems. Awards include Test of Time Awards (2024), ACM Fellow (2020), ERC Consolidator Grant (2020), and EUROGRAPHICS Young Researcher Award (2008). She has supervised numerous students and co-developed software libraries like libigl and Instant Meshes . Co-chair roles for SIGGRAPH, Eurographics, and Pacific Graphics Editorial board member for ACM Transactions on Graphics and other journals Keynote speaker at VMV, CVPR, and SIAM conferences Her work bridges mathematical rigor with practical implementation, advancing computer graphics and geometry processing through intuitive algorithms that maintain surface detail while enabling efficient computation.
Alexei A. Efros is the Howard Friesen Professor in the EECS Department at the University of California, Berkeley, and a core member of the Berkeley Artificial Intelligence Research (BAIR) Lab. Previously, he spent a decade at Carnegie Mellon University’s Robotics Institute. His research focuses on data-driven computer vision, self-supervised learning, computational photography, and generative models. He has pioneered advancements in visual representation learning, including seminal work on neural radiance fields and generative adversarial networks. Education Background: Efros holds a PhD in Computer Science from MIT, though specific details of his academic journey are not explicitly provided in the text. His career includes postdoctoral research at the University of Oxford with Andrew Zisserman and collaborative work with Team WILLOW at INRIA Paris. Research Interests: Efros explores how vast uncurated visual data can be leveraged for understanding and synthesizing the visual world. Key areas include self-supervised learning, generative models, and applications in robotics and art. His lab has contributed influential techniques such as Style Transfer, GAN-based image synthesis, and neural scene representation learning. Recent work emphasizes real-time adaptation (Test-Time Training), 3D perception models, and ethical AI implications of generative systems. Publications: Over 150+ publications span topics like Generative Adversarial Networks (GANs), unsupervised learning, and visual-linguistic models. Notable works include Unpaired Image-to-Image Translation (CUT/GAU), Style Transfer , and Swapping Autoencoder . His research has significant industry impact, with techniques adopted in Adobe’s software and generative AI applications. Grants & Collaborations: Efros has secured major funding from NSF, DARPA, and industry partnerships (e.g., Adobe, NVIDIA). He co-leads projects on scalable vision models, ethical AI, and real-world perception systems. Current collaborations include work with MIT, NYU, and INRIA Paris. Labs & Teams: Leads the BAIR Vision Group at Berkeley, fostering interdisciplinary research between computer vision, graphics, and robotics. The group emphasizes Slow Science principles, prioritizing deep exploration over rapid publication.
Sarita V Adve is the Richard T. Cheng Professor of Computer Science at the University of Illinois at Urbana-Champaign, where she conducts research spanning hardware, programming languages, operating systems, and applications with a focus on domain-specific systems. Her work bridges theoretical foundations and practical implementations, particularly in extended reality and heterogeneous computing. Her educational background includes a Ph.D. and M.S. in Computer Science from the University of Wisconsin-Madison (1993, 1989) and a B.Tech in Electrical Engineering from the Indian Institute of Technology Bombay (1987). Prior to joining Illinois, she served on the faculty at Rice University from 1993 to 1999. Adve's research centers on generalizable and scalable specialization for domain-specific systems, with current emphasis on extended reality (XR) systems including virtual, augmented, and mixed reality. She chairs the ILLIXR consortium to democratize XR research and developed the first fully open-source XR system (ILLIXR). Her foundational contributions include memory consistency models for C++ and Java programming languages, the Spandex coherence framework for heterogeneous systems, and software-driven approaches for hardware reliability. Her work spans hardware reliability (SWAT and RAMP projects), power management (GRACE system), and instruction-level parallelism. Recent publications reveal a strong focus on energy-efficient XR systems, hardware-software co-design for AI workloads, and resilience analysis. Her team explores rendering offload, visual-inertial odometry optimization, and compositional error injection frameworks, often targeting tradeoffs between energy, latency, and accuracy in mobile and edge environments. Fellow of the American Academy of Arts and Sciences Fellow of the ACM and IEEE ACM/IEEE-CS Ken Kennedy Award Anita Borg Institute Woman of Vision in Innovation Award ACM SIGARCH Maurice Wilkes Award Alfred P. Sloan Research Fellowship UIUC University Scholar University of Illinois Campus Award for Excellence in Graduate Student Mentoring Adve actively mentors students and has received multiple teaching awards. She co-founded the CARES movement to address discrimination in CS research events and chairs CS@Illinois CARES. Her service includes leadership roles in ACM SIGARCH (2015-2019), DARPA/ISAT study group, ACM Council, and Computing Research Association. She has secured significant funding including DARPA initiatives and Google Faculty Research Awards. She leads the ILLIXR consortium and has established collaborative research programs such as the $8.3M DARPA Joint University Microelectronics Program. Her lab focuses on open-source XR development, heterogeneous system architectures, and reliability-aware designs, with strong industry and government partnerships.
Deva Ramanan is a Professor at the Robotics Institute of Carnegie Melllon University, where he leads research in computer vision and machine learning. His work focuses on modeling human visual perception, leveraging large-scale visual data, and developing systems for 3D understanding, neural rendering, and autonomous systems. He advises a large group of PhD students and has mentored numerous postdoctoral researchers now in leading roles across industry and academia. His research interests include computer vision, machine learning, human perception modeling, 3D scene understanding, neural rendering, autonomous driving, video understanding, and multimodal foundation models. These areas reflect his focus on both foundational models and their application to real-world problems in robotics and AI. The recent publications highlight a strong trend toward multimodal and 3D-aware models, with increasing use of diffusion models, neural fields, and large vision-language systems. Key themes include scene flow, 3D reconstruction from monocular video, autonomous driving perception, and robust evaluation of vision-language models. There is a clear emphasis on both methodological innovation and practical deployment in dynamic environments. Marr Prize, Honorable Mention (ICCV 2021) Best Paper, Honorable Mention (ECCV 2020) Best Paper Finalist (WACV 2024) Best Paper Award (WACV 2016) Best Industrial Paper, Honorable Mention (BMVC 2017) Marr Prize winner (ICCV 2009) Deva Ramanan has advised numerous PhD and master’s students, many of whom are now at top institutions and companies including Apple, Meta, Google, Nvidia, OpenAI, and Princeton. He has received substantial funding from IARPA, DARPA, NSF, Intel, Google, and Facebook for projects in video analytics, dispersed computing, visual cloud systems, and multi-task recognition. His group has developed influential datasets and benchmarks used widely in the community. He leads a vibrant research lab focused on advancing computer vision through deep learning and multimodal integration. His team works on core challenges in perception, including 3D reconstruction, motion modeling, object detection, and scene understanding, with applications in robotics and autonomous systems.
Sebastian Scherer is an Associate Research Professor at the Robotics Institute (RI), Carnegie Mellon University (CMU), where he leads cutting-edge research in autonomous aerial systems and robotics. His work focuses on enabling unmanned rotorcraft to operate safely and efficiently in cluttered, low-altitude, and extreme environments. Education: Ph.D. in Robotics, Carnegie Mellon University (2010) MS in Robotics, Carnegie Mellon University (2007) BS in Computer Science (Minor in Robotics), Carnegie Mellon University (2004) His research interests span robotics, artificial intelligence, autonomous navigation, obstacle avoidance, SLAM, visual-inertial odometry, energy infrastructure, and public policy . He has made seminal contributions to UAV autonomy, including the first obstacle avoidance for micro aerial vehicles in natural environments (2008) and the first automatic landing zone detection and landing on a full-size helicopter (2010). His recent publications (2023–2025) demonstrate a strong focus on resilient autonomy, multi-robot exploration, foundation models for robotics, and large-scale dataset development. His team has released key datasets like TartanGround , BETTY , and SubT-MRS , and simulation tools like Pegasus Simulator , indicating a systems-level approach to advancing real-world autonomy. The research trends emphasize self-supervised learning, robust perception, risk-aware planning, and multi-modal fusion for off-road and urban environments. Scientific Awards: Popular Science Best of What's New 2010 Award AIAA@Infotech Best Paper Runner-up Award (2010) Siebel Scholar Dr. Scherer has advised numerous students and leads a vibrant research group focused on high-impact robotics applications. He has secured significant grants related to UAV autonomy, energy infrastructure, and urban air mobility. His lab develops experimental infrastructure such as AIrTonomy for testing next-generation autonomous aerial vehicles. He is actively involved in advancing SLAM and localization in extreme environments, notably through participation in the DARPA Subterranean Challenge. His team develops large-scale datasets and benchmarking frameworks to push the boundaries of robustness and generalization in mobile robotics.
Rakesh Kumar is a Professor and John Bardeen Faculty Scholar in the Electrical and Computer Engineering Department at the University of Illinois at Urbana-Champaign. His work focuses on computer architecture, system-level design automation, and low-power computing. PhD in Computer Engineering from University of California, San Diego BS in Electrical Engineering from IIT Kharagpur His research spans all layers of the computing stack, with key contributions to flexible computer systems , waferscale computing , error-resilient architectures , and approximate computing . He has pioneered work on voltage-reliability tradeoffs and peak power management techniques. Recent publications highlight trends in space microdatacenters , printed microprocessors , and neural graph accelerators . His work on plastic chips was recognized as one of the three biggest semiconductor headlines of 2022 by IEEE Spectrum. IEEE Fellow (2024) ISCA Influential Paper Award MICRO Test-of-Time Award ICCAD Ten Year Retrospective Most Influential Paper Award Best Paper Awards at CASES, SELSE, HPCA He has received teaching accolades including the Stanley H. Pierce Faculty Award and Ronald W. Pratt Outstanding Teaching Award . His research group explores hardware-software co-design for emerging applications in AI, IoT, and sustainable computing.
Karthik R. Narasimhan is a Professor at Princeton University's School of Engineering and Applied Science in the Department of Computer Science. Previously, he earned his PhD from MIT under Regina Barzilay and served as a visiting research scientist at OpenAI during 2017-18. His research focuses on the intersection of language and decision-making, building autonomous agents that learn from both experience and human knowledge. His research spans multiple high-impact areas including language agents (Text-DQN, CALM, ReAct, Tree of Thoughts), reinforcement learning (h-DQN, Multi-Objective RL), and AI safety (Toxicity in ChatGPT, DataMUX). He has developed critical datasets and benchmarks such as WebShop, InterCode, SWE-bench, and SILG that have become standard evaluation tools in the field. Current work emphasizes agent capabilities, software engineering automation, and multimodal interaction. His publication trends show strong focus on practical agent deployment (SWE-agent, Tree of Thoughts), safety evaluation (Probing AI Safety), and efficiency improvements (DataMUX). Recent work increasingly addresses real-world challenges in software engineering, security, and human-AI collaboration through rigorous benchmarking. Co-author of foundational GPT (2018) paper Key developer of Text-DQN (2015), CALM (2020), ReAct (2022), Tree of Thoughts (2023) Creator of influential benchmarks: WebShop (2022), SWE-bench (2023), InterCode (2023) He actively advises students through Princeton's computer science program, with research supported by multiple grants focused on autonomous agent development and language-based decision systems. His GitHub repositories (nlp-datasets, text-world-player) demonstrate strong community engagement in open-source research tools. Current projects include advancing language agent capabilities through Reflexion (2023) and Tree of Thoughts (2023) frameworks while addressing critical safety and efficiency challenges.
Professor Hanumant Singh leads the Electrical and Computer Engineering department at Northeastern University, with a joint appointment in Mechanical and Industrial Engineering , and serves as Program Director for the Master of Science in Robotics. He earned his Ph.D. from MIT/WHOI Joint Program in 1995 and has conducted over 60 expeditions globally, focusing on marine geology, polar studies, and coral reef ecology. His research emphasizes field robotics , including SLAM, underwater manipulation, and imaging in extreme environments. He developed the Seabed AUV and Jetyak ASV , widely used in scientific research. His labs include the Field Robotics Lab and the Institute for Experiential Robotics . Research Interests: Machine Learning for Fisheries SLAM with dynamic objects Underwater imaging and manipulation Autonomous surface and aerial systems Polar and marine robotics Awards: ICRA Best Student Paper Award, IEEE Oceanic Engineering Society Distinguished Faculty Award (2025), Lifetime Achievement Award (2022), and IEEE Fellow status. His work has been featured in Nature Geoscience , Polar Biology , and media outlets like WGBH. Students & Collaborations: Advises students like Srinidhi Pattala (MS Robotics) and Dennis Giaya (PhD Computer Engineering). Collaborates with institutions on projects such as Antarctic sea ice thickness estimation and deep-sea submersible missions.
Dr Andrew Rhead is a Senior Lecturer in the Department of Mechanical Engineering at the University of Bath, specializing in aerospace composites and damage tolerance analysis. His research focuses on impact damage detection, failure mechanism modeling, and Non-Destructive Evaluation (NDE) techniques for composite structures. MSci in Mathematical Sciences (Dynamical Systems) - University of Bristol (2006) PhD in Composite Damage Tolerance - University of Bath (2009) His work develops computationally efficient analytical models for compression after impact (CAI) strength prediction in composite laminates, surpassing traditional finite element methods. Key projects include hydrogen storage systems for aircraft, cryogenic composite testing, and steered fiber manufacturing optimization. Active in 10 projects including ASPIRE and HyFIVE Collaborates with Airbus, GKN Aerospace, and EPSRC Research trends show emphasis on sustainable aviation materials, structural battery integration, and advanced testing methodologies. Current affiliations include the Institute for Mathematical Innovation (IMI) and Centre for Integrated Materials, Processes & Structures (IMPS).
Patrick Slade is an Assistant Professor of Bioengineering at Harvard University's School of Engineering and Applied Sciences (SEAS). His lab, the Slade Lab, focuses on developing assistive devices to enhance mobility through the integration of biomechanics, robotics, and human-centered artificial intelligence. Key research areas include exoskeletons, prosthetics, wearable sensors for health tracking, and navigation aids for visually impaired individuals. Research Interests: The lab emphasizes translating research into practical solutions, such as personalized exoskeletons and robotic systems to improve mobility. Recent work includes optimizing human-robot interaction algorithms and publishing in high-impact journals like Nature . Collaborations with labs like the Biodesign Lab and BIONICs Lab highlight cross-disciplinary efforts. Publications: Over 15 articles since 2017 span topics like exoskeleton design, energy expenditure modeling, and Bayesian reinforcement learning. Notable contributions include a 2022 Nature paper on personalized exoskeleton assistance and a 2021 study on navigation aids for impaired vision. Awards & Grants: Students in his group have received prestigious NSF GRFP fellowships and conference awards, reflecting the lab's emphasis on innovation. The lab actively engages in grant-funded projects to advance assistive technology. Lab & Team: The Slade Lab opened at Harvard in 2023 and includes PhD students and postdocs working on devices like robotic exoskeletons and health-tracking systems. Future work focuses on scalable solutions for mobility challenges through interdisciplinary approaches.
Jeannette Bohg is an Assistant Professor of Computer Science at Stanford University, directing the Interactive Perception and Robot Learning Lab. Previously, she was a group leader at the Autonomous Motion Department (AMD) of the MPI for Intelligent Systems (2012-2017). She holds a PhD from KTH Royal Institute of Technology (Stockholm) and degrees from Chalmers University and TU Dresden. Her research focuses on perception, learning, and real-time multi-modal methods for autonomous robotic manipulation and grasping, aiming to bridge principles of human sensorimotor coordination with robotic implementation. Education: PhD in Robotics (KTH), MSc in Art & Technology (Chalmers), Diploma in Computer Science (TU Dresden) Research interests include developing goal-directed, real-time robotic systems capable of meaningful feedback for execution and learning. Key areas are dexterous manipulation, imitation learning, and cross-embodiment policy transfer. Notable contributions include the TidyBot platform and work on force-aware surgical robotics. Awards include the 2019 IEEE ICRA Best Paper Award, 2019 IEEE RA Early Career Award, and 2020 RSS Early Career Award. Her lab explores intersections of robotics, ML, and computer vision. Advising: Actively mentoring students/postdocs in manipulation, perception, and learning. Grants and collaborations span NSF, Stanford AI Lab, and industry partnerships. Future work emphasizes robust real-world deployment and human-robot collaboration. Labs/Teams: Leads the Interactive Perception and Robot Learning Lab, contributing to Stanford’s AI ecosystem. Previously managed the MPI AMD group, fostering interdisciplinary research in autonomous systems.