Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Prof. Ingmar Posner is a leading figure in applied artificial intelligence at the University of Oxford, where he serves as Principal Investigator for the Applied Artificial Intelligence Lab (A2I) and founding Director of the Oxford Robotics Institute. His work focuses on enabling robots to operate effectively in complex real-world environments through experience-driven learning. Key research areas: robot learning, scene interpretation, data-efficient learning, and transfer learning Applications in manipulation, autonomous driving, logistics, and space exploration His team has produced groundbreaking work in world models, sim-to-real transfer, and constraint-based manipulation systems (e.g., COMBO-Grasp). Notable contributions include the TWIST distillation framework and foundational research in tactile data generation (TactGen). He has received multiple best paper awards at top robotics venues. Publications reveal evolving research themes: 2025 work emphasizes language-conditioned learning (Lumos) and multi-agent decision-making, while 2024 focused on diffusion models for locomotion and differentiable simulators. Earlier work spans from urban scene analysis to physically plausible scene synthesis (RELATE).
Justin Johnson is an Assistant Professor at the University of Michigan's College of Engineering, Department of Electrical Engineering and Computer Science, and a Research Scientist at Facebook AI Research (FAIR). His work bridges computer vision, machine learning, and deep learning, focusing on visual reasoning, vision-language tasks, image generation, and 3D reasoning using neural networks. PhD, Stanford University (advised by Fei-Fei Li) His research interests span visual reasoning , vision and language , 3D vision , and image generation , with a focus on innovative applications of deep neural networks. Recent publications highlight work on 3D consistency, self-supervised learning, and multimodal integration of vision and text. Notable contributions include PyTorch3D for 3D data processing, and foundational work in visual question answering , neural style transfer , and scene graph-based image generation . Publications span top conferences like ICCV, CVPR, and NeurIPS. He teaches courses including EECS 498/598: Deep Learning for Computer Vision and EECS 442: Computer Vision at University of Michigan, with prior involvement in Stanford's CS 231N in co-teaching roles. Software projects include open-source frameworks like fast-neural-style for real-time artistic style transfer, and PyTorch3D for efficient 3D deep learning. These tools demonstrate his commitment to practical implementations and community-driven research.
Prof. Christian Holz is an Associate Professor at the Department of Computer Science and Deputy Head of the Institute of Intelligent Interactive Systems at ETH Zürich. His work focuses on advancing human-computer interaction through innovations in wearable technologies, mixed reality systems, and sensor-driven applications. Key research areas include motion capture, physiological signal processing, and adaptive user interfaces. Holz leads the SIPLab (siplab.ethz.ch), producing influential work at the intersection of computer science and biomedical engineering. His research explores cutting-edge topics such as egocentric vision systems, wearable health monitoring devices, and VR/AR applications. Recent studies investigate cybersickness detection via EEG, heart rate estimation from eye-tracking cameras, and scalable motion capture using inertial/UWB sensors. Holz's work emphasizes practical applications in healthcare, education, and human-centered computing. Publications reflect a strong focus on interdisciplinary solutions, combining machine learning with sensor data analysis. Notable contributions include the EgoSim multi-view simulator, WildPPG biomedical dataset, and MiBOT cardiovascular modulation device. His research bridges theoretical advancements with real-world usability in domains like emergency response training, chronic disease monitoring, and immersive education.
Shrikanth (Shri) Narayanan is a University Professor and holder of the Niki and Max Nikias Chair in Engineering at the University of Southern California (USC), serving as the inaugural Vice President for Presidential Initiatives. He leads the Signal Analysis and Interpretation Lab (SAIL) and holds joint appointments in Computer Science, Linguistics, Psychology, Neuroscience, Pediatrics, and Otolaryngology-Head and Neck Surgery. His research focuses on speech and audio processing, behavioral signal processing, and real-time MRI of speech production, with applications in healthcare, education, and technology. Education: B.E. in Electrical Engineering from College of Engineering, Guindy (Chennai, India, 1988); M.S., Engineer, and Ph.D. in Electrical Engineering from UCLA (1990, 1992, 1995). Research interests span computational linguistics, machine learning, and multimodal human behavior analysis. He pioneered technologies for speech biomarkers in mental health, real-time MRI of speech production, and wearable sensor systems for longitudinal health studies. His work in speech emotion recognition, forensic interviews, and clinical applications has been recognized through over 40 awards, including the IEEE Flanagan Award and ISCA Medal. He has published extensively in journals like Proceedings of the IEEE , Journal of the Acoustical Society of America , and PLOS One . Key Grants: NSF CAREER, Okawa Research, IBM Faculty, Google/Amazon awards. Labs/Teams: Signal Analysis & Interpretation Lab (SAIL), USC Information Sciences Institute (ISI), Google Visiting Faculty Researcher.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Kristen Grauman is a Full Professor in the Department of Computer Science at the University of Texas at Austin, where she leads the UT Computer Vision Group. Her research focuses on computer vision and machine learning, with applications in visual recognition, video analysis, and multi-modal perception. She received her B.A. from Boston College and her Ph.D. from MIT. Her research interests span visual recognition, image and video search, video analysis, first-person vision, embodied and multi-modal perception, and interactive machine learning. She has made significant contributions to the field, particularly in developing algorithms for understanding visual content and human activities from video, including foundational work on the Pyramid Match Kernel and relative attributes. Her recent publications reveal a strong emphasis on egocentric (first-person) vision, audio-visual learning, and view-invariant representations. There is a clear trajectory toward multi-modal integration (vision, audio, language) and real-world applications in instructional videos, human activity understanding, and embodied AI systems. She has received numerous awards including: AAAI Fellow (2019) J. K. Aggarwal Prize, International Association for Pattern Recognition (2018) Helmholtz Prize (2017) UT Austin Academy of Distinguished Teachers (2017) Best Paper Award, Asian Conference on Computer Vision (2016) Presidential Early Career Award for Scientists and Engineers (2014) Computers and Thought Award, International Joint Conferences on Artificial Intelligence (2013) Pattern Analysis and Machine Intelligence Young Researcher Award (2013) Alfred P. Sloan Research Fellow (2012) Marr Prize (2011) Prof. Grauman serves as Associate Editor-in-Chief for the IEEE Transactions on Pattern Analysis and Machine Intelligence. She has secured substantial research funding including the Presidential Early Career Award, NSF grants, and industry partnerships. Her advising has produced numerous influential publications and students who are now leaders in computer vision. She leads the UT Computer Vision Group, which collaborates closely with the Electrical and Computer Engineering Department. The group is pioneering large-scale egocentric video research through projects like Ego4D and Ego-Exo4D, focusing on real-world applications in human activity understanding, audio-visual perception, and interactive systems.
Sara Ng is a Visiting Assistant Professor in the Department of Linguistics at Western Washington University. She holds a PhD in Linguistics from the University of Washington (2024) and an MS in Computational Linguistics (2023), alongside a B.A. in Linguistics and B.S. in Applied Mathematics from the University of Utah (2017). Her research focuses on computational models of prosody, speech perception, and their integration with speech technologies like automatic speech recognition (ASR). She investigates how prosody conveys pragmatic meaning, influences conversational dynamics, and impacts listeners with hearing impairments. Her work bridges computational methods and linguistic theory, addressing challenges in clinical and technological applications. Education PhD in Linguistics, University of Washington, 2024 MS in Computational Linguistics, University of Washington, 2023 B.A. in Linguistics (Honors), B.S. in Applied Mathematics, University of Utah, 2017 Research Interests Ng’s research explores computational linguistics, prosody modeling, speech technology, and hearing impairment studies. She develops methods to leverage prosodic cues for tasks like punctuation prediction and investigates the impact of hearing loss on speech perception. Grants & Awards 2024: Nominated for UW Excellence in Teaching Award 2023: Excellence in Linguistic Research Graduate Fellowship 2017: University of Utah Top Scholar Award Teaching Ng teaches courses in phonetics, computational linguistics, and linguistics for honors students. She emphasizes pedagogical innovation and inclusivity, with experience as an instructor of record and teaching assistant at both Western Washington University and the University of Washington. Labs & Collaborations She collaborates with the TIAL Lab (University of Washington), Phonetics Lab, Hearing Aid Laboratory (Northwestern University), and CLILLAC-ARP (Université Paris Cité). Her work intersects with clinical and engineering domains, addressing real-world applications of speech technology.
Hao Su is an Associate Professor in the Department of Computer Science and Engineering at University of California, San Diego . He serves as Chairman & CTO of Hillbot Inc , and leads the SU Lab which focuses on building autonomous systems that learn actively in physical environments. His affiliations include the Institute for Learning-enabled Optimization at Scale , Artificial Intelligence Group , Contextual Robotics Institute , Halicioğlu Data Science Institute , and Center for Visual Computing . As a researcher in Computer Vision, Robotics, and Neural Geometry , he has made significant contributions to 3D foundation models, reward-free world models, diffusion policy frameworks, and GPU-accelerated simulation environments. His 2024-2025 publications include advancements in hand-eye calibration, dynamic mesh reconstruction, and multi-stage robotic manipulation. His scientific awards include: Frontiers of Science Award (2025) TPAMI Young Research Award (2025) NSF CAREER Award (2023) ACM SIGGRAPH Best Doctorate Thesis Honorable Mention (2019) He has served as Program Chair for CVPR 2025 and Area Chair for ICLR 2022 and NeurIPS 2023 , while previously serving as Publication Chair for 3DV 2016 and Program Committee for SIGGRAPH Asia Workshops .
Nadia Figueroa is the Shalini and Rajeev Misra Presidential Assistant Professor in the Mechanical Engineering and Applied Mechanics (MEAM) Department at the University of Pennsylvania. She holds secondary appointments in Computer and Information Science (CIS) and Electrical and Systems Engineering (ESE), and is a core faculty member at the GRASP Lab. Prior to Penn, she was a Postdoctoral Associate at MIT’s CSAIL and earned her PhD in Robotics from EPFL under Prof. Aude Billard. Her research focuses on physical and perceptual adaptive intelligence for robots, enabling fluid collaboration with humans in dynamic environments. Key applications include robot learning from demonstration , human-robot co-manipulation , safe navigation in human-centric spaces , and rehabilitation robotics . Her work integrates machine learning control theory artificial intelligence biomechanics psychology with guarantees of stability, safety, and robustness . Recent publications highlight advancements in reactive collision avoidance dynamical system learning intent estimation EEG-driven assistive control origami-based reconfigurable robots across platforms like autonomous vehicles and humanoid robots. She has authored a 2022 textbook on dynamical systems for robot control and received the Presidential Assistant Professorship at Penn.
Professor Andrew Davison holds the position of Professor of Robot Vision at Imperial College London's Department of Computing. He leads the Dyson Robotics Laboratory and the Robot Vision Research Group, focusing on advancing SLAM (Simultaneous Localization and Mapping) and Spatial AI. His groundbreaking work includes the MonoSLAM algorithm (2003), enabling real-time 3D vision for robotics and AR/VR. Current research emphasizes scalable, semantic-rich Spatial AI systems, as outlined in his FutureMapping papers (2018–2019). Education: BA in Physics (Oxford, 1994), D.Phil. (Oxford, 1998). Postdoctoral work at AIST, Japan (1998–2000), followed by a lectureship at Imperial (2002–present). Industrial collaborations include SLAMcore, a Spatial AI startup, and Dyson Robotics Lab. Over 18 PhD students supervised, many now leading roles at Meta, NVIDIA, SLAMcore, and academia. Notable contributions include DTAM, KinectFusion, and Event Camera SLAM. Recognized for software tools like SceneLib and contributions to robotics benchmarks (SLAMBench). Active on Twitter (@AjdDavison) for research updates.
Prof. Luke Zettlemoyer is an Adjunct Professor of Computer Science and Engineering at the University of Washington, with affiliations to the Department of Linguistics. He focuses on machine learning, natural language processing, and multimodal systems, contributing to advancements in large language models, ethical AI, and scalable architectures. His research addresses challenges in model alignment, generalization, and cross-domain integration. Key research interests include multimodal reward models, efficient tokenization strategies, and model optimization techniques. He has explored topics such as neural trajectories for robot learning, content-adaptive image processing, and ethical mitigation of verbatim data reproduction. His publications span 2023–2025, emphasizing practical applications of AI in robotics, vision-language systems, and scalable retrieval-based models. While no formal awards are listed, his work reflects significant contributions to foundational AI research.
Lerrel Pinto is an Assistant Professor of Computer Science at the Courant Institute of Mathematical Sciences at New York University (NYU), where he leads the General-purpose Robotics and AI Lab (GRAIL) as part of the CILVR research group. His work bridges the gap between theoretical machine learning and practical robotics applications, with a focus on enabling robots to generalize and adapt in real-world environments. Dr. Pinto received his undergraduate degree from IIT Guwahati, followed by a PhD from the Robotics Institute at Carnegie Mellon University (CMU). He then completed a postdoctoral fellowship at the University of California, Berkeley before joining NYU as faculty. His research program centers on robot learning and decision making, with several key thrusts that demonstrate his innovative approach to robotics. Pinto's work emphasizes large-scale learning techniques that leverage both extensive data and sophisticated model architectures. A significant portion of his research focuses on representation learning for sensory data, particularly developing methods that enable robots to make sense of visual, tactile, and auditory inputs. His lab has made notable contributions to reinforcement learning algorithms that allow robots to adapt to new scenarios with minimal retraining. Pinto also champions open-source robotics , developing affordable robot platforms that democratize access to robotics research. Analysis of Pinto's recent publications reveals a strong trend toward multimodal perception in robotics, integrating visual, tactile, and auditory information to create more robust robot systems. His work increasingly focuses on zero-shot and few-shot learning capabilities, enabling robots to handle novel situations without extensive retraining. There's also a clear progression toward general-purpose robotics , moving away from task-specific solutions toward more flexible systems that can handle diverse real-world challenges. Dr. Pinto's scientific contributions have been recognized with several prestigious awards: Sloan Research Fellowship (2025) NSF CAREER Award (2024) RAL Early Career Award (2024) Best Student Paper Award at ICRA (2016) Outstanding Paper Award at MFM-EAI workshop at ICML (2024) Best Paper Award at NGSM workshop at ICML (2024) Best Student Paper Award at RSS (2023) As an advisor, Pinto has mentored numerous students who have gone on to impactful careers in both academia and industry. His former PhD student Denis Yarats co-founded Perplexity.AI, while Mahi Shafiullah became a postdoc at UC Berkeley and Meta AI. Many of his Masters students have pursued PhDs at top institutions like CMU, MIT, and Stanford, or joined leading robotics companies including 1X, Fauna Robotics, and NVIDIA. Pinto's lab has secured significant research funding, including the NSF CAREER award and likely other grants supporting his robotics research program. The General-purpose Robotics and AI Lab (GRAIL) that Pinto leads brings together a diverse team of researchers working on cutting-edge robotics challenges. The lab maintains strong collaborations with industry partners and other academic institutions, facilitating technology transfer and real-world impact. GRAIL's research spans multiple robotics platforms and focuses on developing algorithms that enable robots to learn from diverse experiences and generalize across environments.
Sanjiv Singh is a Research Professor at the Robotics Institute within Carnegie Mellon University's School of Computer Science. His academic journey at CMU spans from Systems Scientist (1995-2001), to Senior Research Scientist (2001-2003), Associate Research Professor (2003-2007), and finally Research Professor since 2007. He also holds an adjunct faculty position in Mechanical Engineering since 2009. Singh serves as Editor-in-Chief of the Journal of Field Robotics, demonstrating his leadership in the robotics community. His educational background includes a Ph.D. and M.S. in Robotics from Carnegie Mellon University (1995, 1992), an M.S. in Electrical Engineering from Lehigh University (1985), and a B.S. in Computer Science from the University of Denver (1983). Dr. Singh's research focuses on three primary themes: Autonomous Navigation (developing motion planning and control for ground and air vehicles with applications in agriculture, exploration, and low-flying aircraft), Coordinated Multi-Robots (examining team-based tasks like structure assembly and search/rescue operations), and Forceful Interaction with the world (using physical models to enable robots to handle complex, high-force interactions). His work spans aerial robotics, agricultural and forestry robotics, mining robotics, 3D vision, sensing and perception, visual servoing, motion planning, and field service robotics. Analysis of his recent publications (2016-2020) reveals a strong focus on collision avoidance algorithms, sensor fusion techniques, and real-time navigation systems. His research demonstrates consistent advancement in SLAM (Simultaneous Localization and Mapping) technologies, particularly in GPS-denied environments, with increasing sophistication in handling complex aerial maneuvers and multi-robot coordination. Editor-in-Chief of Journal of Field Robotics Dr. Singh has advised numerous graduate students throughout his career, with current advisees including Matt Aasted (Ph.D), Andrew Chambers (M.S), Hugh Cover (M.S), Michael Dille (Ph.D), and Justin Haines (M.S). His past students include prominent researchers like Sebastian Scherer, Joe Djugash, Fred Heger, Geoff Hollinger, and Ji Zhang who have gone on to make significant contributions in robotics. His research has been supported through various projects including CASC (agricultural applications), Riverine, Transformer, Trestle, and Ember. His laboratory work focuses on developing practical robotic systems capable of operating in challenging real-world environments, with particular emphasis on agricultural applications, search and rescue operations, and coordinated multi-robot teams that can work effectively alongside humans.
Olga Sorkine-Hornung is a Professor of Computer Science at ETH Zurich and head of the Institute of Visual Computing. She leads the Interactive Geometry Lab, focusing on theoretical and practical advancements in digital content creation, geometry processing, and shape modeling. Current position: ETH Zurich, Department of Computer Science Previous roles: Courant Institute (NYU), Technical University of Berlin Education: BSc and PhD from Tel Aviv University, postdoc at TU Berlin Her research spans shape representation, digital fabrication, computer animation, and fundamental geometry processing. Key contributions include Laplacian surface editing, as-rigid-as-possible deformation, and generalized winding numbers. She works on applications in VR, AR, and autonomous systems. Awards include Test of Time Awards (2024), ACM Fellow (2020), ERC Consolidator Grant (2020), and EUROGRAPHICS Young Researcher Award (2008). She has supervised numerous students and co-developed software libraries like libigl and Instant Meshes . Co-chair roles for SIGGRAPH, Eurographics, and Pacific Graphics Editorial board member for ACM Transactions on Graphics and other journals Keynote speaker at VMV, CVPR, and SIAM conferences Her work bridges mathematical rigor with practical implementation, advancing computer graphics and geometry processing through intuitive algorithms that maintain surface detail while enabling efficient computation.