Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Christian Theobalt is a Professor of Computer Science at Saarland University and Scientific Director of the Visual Computing and Artificial Intelligence Department at the Max Planck Institute for Informatics . He leads the Saarbruecken Center for Visual Computing as a strategic partnership between Google and MPI. PhD in Computer Science (2005) from MPI-INF/Saarland University Postdoctoral Researcher at MPI (2005-2007) Visiting Assistant Professor at Stanford (2007-2009) His research focuses on the intersection of Computer Graphics, Computer Vision, and Artificial Intelligence , with specializations in: 3D/4D Human Reconstruction Neural Rendering Performance Capture Geometric Deep Learning Volumetric Video Quantum Visual Computing Recent article trends show emphasis on: Quantum computing applications in visual reconstruction Neural rendering with radiance fields Human motion capture from egocentric views Gesture-language interaction modeling Multi-modal scene understanding Real-time free-viewpoint rendering Scientific Awards : Fellow of EUROGRAPHICS (2022) CVPR Best Student Paper Honorable Mention (2020) ERC Consolidator Grant (2017) Karl Heinz Beckurts Award (2017) Busy Beaver Teaching Award (2016) ERC Starting Grant (2013) Advising over 50 students and researchers including: Current researchers: Viktor Rudnev, Linjie Lyu, Mohit Mendiratta Postdocs: Kwang In Kim, Kiran Varanasi Alumni: Franziska Mueller (Google), Dushyant Mehta (Qualcomm), Ayush Tewari (MIT) Grants include ERC grants, Google Glass Research Award, and multiple industry partnerships. His lab maintains cutting-edge facilities with: Multi-camera capture systems Quantum annealing infrastructure HDR radiance field technology GPU clusters for AI research Time-of-flight imaging systems Event camera arrays
Luca Carlone is the Boeing Career Development Associate Professor in the Department of Aeronautics and Astronautics at MIT and a Principal Investigator at the Laboratory for Information & Decision Systems (LIDS) . He leads the SPARK Lab , focusing on developing certifiable perception algorithms for autonomous systems. PhD in Mechatronics (Polytechnic University of Turin, 2012) Research spans robotics, computer vision, and optimization Research Interests : Certifiable Perception algorithms for high-integrity systems High-level Perception (geometric, semantic, physical understanding) Efficient Perception methods for resource-constrained robots Scientific Contributions include: 2024 Outstanding Systems Paper Award (RSS) 2023 IEEE Transactions on Robotics King-Sun Fu Award 2021 NSF CAREER Award 2020 AIAA Advising Award 2019 Amazon Research Award Advising : Teaches graduate courses like Visual Navigation for Autonomous Vehicles and Robotics: Science and Systems . Collaborates with institutions including JPL, Caltech, and KAIST through the DARPA SubT Challenge.
Nima Fazeli is an Assistant Professor of Robotics at the University of Michigan (2020–Present), holding courtesy appointments in Computer Science & Engineering (CSE) and Mechanical Engineering. He directs the Manipulation and Machine Intelligence (MMint) Lab, focusing on enabling dexterous robotic manipulation through multimodal representation learning, tactile sensing, and model-based reasoning. His work integrates mechanics, perception, controls, and planning to achieve autonomous interaction with uncertain environments. Education: PhD, MIT (2019); MSc, University of Maryland (2014); BSc, Amirkabir University of Technology (2011) Research interests emphasize embodied intelligence , including visuo-tactile fusion, contact dynamics modeling, and cross-modal learning. Recent work explores tactile shadows, deformable object manipulation, and language-guided robot control. His research is supported by the NSF CAREER grant and National Robotics Initiative, with applications in manufacturing, assistive robotics, and space systems. Publications span topics like tactile sensing hardware (e.g., GelSlim 4.0), visuo-tactile implicit representations (ViTaSCOPE), and failure recovery policies (Racer). His team’s work has been featured in outlets like The New York Times and BBC. Key Awards: NSF CAREER Grant (2024) Teaching includes Introduction to Robotic Manipulation . Collaborations involve cross-disciplinary projects with mechanical, electrical, and biomedical engineering groups.
Christian Rupprecht is an Associate Professor at the Department of Computer Science, University of Oxford, specializing in computer vision and machine learning. His research focuses on unsupervised learning, 3D reconstruction, and visual understanding. His work includes contributions to conferences such as GCPR'25, ICCV'25, and CVPR'25, with papers spanning topics like correspondence estimation, animal pose modeling, and synthetic data generation. He leads projects within the prestigious Visual Geometry Group (VGG). Notably, his paper VGGT received the Best Paper Award at CVPR'25. His research integrates deep learning and geometric modeling, emphasizing robustness and generalization in visual systems. Best Paper Award at CVPR'25
Alexander Schwing is an Associate Professor in the Department of Electrical and Computer Engineering and Computer Science at the University of Illinois at Urbana-Champaign, affiliated with the Coordinated Science Laboratory. His research focuses on machine learning and computer vision with applications in 3D scene understanding, generative modeling, and multi-agent systems. Education: Diploma in Electrical Engineering and Information Technology, Technical University of Munich (TUM) PhD in Computer Science, ETH Zurich Postdoctoral Fellow, University of Toronto Research Interests: Structured prediction in deep learning Generative adversarial networks and stability Multi-modal vision-language models 3D scene reconstruction from single images Embodied agent collaboration Semantic segmentation with temporal coherence Recent Publications: Highlight trends in neural rendering, video object segmentation, and reinforcement learning with applications to 3D modeling and multi-agent systems. Notable innovations include SAIL-VOS dataset for amodal segmentation and NeRFDeformer for single-view scene transformation. Scientific Awards: NSF CAREER Award, 3M and Amazon research awards, multiple student recognition awards, ETH Zurich PhD medal, and best paper at Intelligent Tutoring Systems 2014. Teaching: Offers graduate courses in Pattern Recognition (ECE 544) and Machine Learning (CS 446/ECE 449). Previously taught at University of Toronto and ETH Zurich. Labs & Collaborations: Leads research at Coordinated Science Laboratory (UIUC) with collaborations across University of Toronto, ETH Zurich, and industry partners like Samsung SAIT and Amazon.
Prof. Stefan Leutenegger is a tenure-track Assistant Professor at Technische Universität München (TUM), leading the Machine Learning for Robotics group within the TUM School of Computation, Information, and Technology. Previously, he held roles as Senior Lecturer (2018–2021) and Lecturer (2014–2018) at Imperial College London's Dyson Robotics Lab, where he founded the Smart Robotics Lab. He earned his PhD (2014) from ETH Zurich, focusing on autonomous solar-powered aircraft navigation, and holds BSc (2006) and MSc (2009) in Mechanical Engineering from ETH Zurich. His research centers on mobile robotics, particularly enabling robots (e.g., drones) to perceive and navigate complex environments using machine learning and sensor data fusion. Key focus areas include SLAM, event-based vision, 3D reconstruction, and autonomous exploration. He has pioneered algorithms like BRISK (2011), OKVIS (2014), and ElasticFusion (2016), advancing real-time robotics perception. Notable Awards: Imperial College President's Award (2018), Best ECCV Paper (2016), ETH Medal for Dissertations (2015). Labs: TUM's Machine Learning for Robotics Group, Imperial's Smart Robotics Lab. Publications: Over 100 papers, including seminal works in CVPR, ECCV, and Robotics: Science and Systems. Current projects include DigiForests (forest inventory via robotics), aerial additive manufacturing, and object-centric semantic mapping. His work bridges theory and practice, with applications in autonomous drones, construction robotics, and human-robot interaction.
Mathieu Salzmann is a Senior Scientist and Lecturer at École Polytechnique Fédérale de Lausanne (EPFL), affiliated with the Computer Vision Laboratory (CVLAB) in the School of Computer and Communication Sciences (IC). He also holds a courtesy appointment with the EPFL College of Humanities and serves as Deputy Chief Data Scientist at the Swiss Data Science Center (SDSC). He has held concurrent roles in teaching units including SIN, SODH, and SSC, reflecting his interdisciplinary engagement. His research focuses on the intersection of machine learning and computer vision, particularly in deep learning for 2D and 3D visual scene understanding, efficient and robust models, domain adaptation, and interpretable AI. These interests are evident across his extensive publication record in top-tier venues. His recent publications (2023–2024) show a consistent trend in advancing deep learning methods for visual recognition, with strong representation at CVPR, ICCV, ECCV, ICML, ICLR, and NeurIPS. Topics include domain generalization, 3D understanding, model robustness, and multimodal learning, often with applications in real-world systems. His editorial roles as Associate Editor for IEEE TPAMI and Action Editor for TMLR further highlight his leadership in the field. Area Chair: ICML 2023, CVPR 2023, ICCV 2023, NeurIPS 2023, AAAI 2024, ECCV 2024 Associate Editor: IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) Action Editor: Transactions on Machine Learning Research (TMLR) Mathieu Salzmann has supervised numerous PhD students at EPFL, both current and past, including Bouquet Yann Yanis, Javed Saqib, Li Shuangqi, and others. He has also been involved in research grants and collaborative projects, such as his work with S. Süsstrunk and R. Baroni on comics reconfiguration. His part-time role as Senior GNC Engineer at ClearSpace (2020–2024) illustrates his applied research engagement in aerospace systems. He is actively involved in EPFL’s data science and AI research ecosystem through SDSC and multiple labs.
Lourdes Agapito is a Professor of 3D Vision at the Department of Computer Science, University College London (UCL), within the Faculty of Engineering Sciences. She leads research in Non-Rigid Structure from Motion (NR-SFM) and 3D reconstruction from monocular video sequences. Her work addresses dynamic scenes, deformable objects, and articulated structures, with applications in robotics and computer vision. She holds an ERC Starting Grant (2008–2014) and led the EU Horizon 2020-funded Second Hands project (2014–2019), collaborating with institutions like EPFL and KIT to develop robots with 3D visual perception for maintenance tasks. Her research group focuses on dense optical flow estimation, video registration, and deformable tracking. Agapito’s research interests include monocular 3D reconstruction, non-rigid motion analysis, and neural approaches to 3D modeling. She has supervised multiple PhD students and postdocs, including notable researchers such as Ravi Garg and Marco Paladini (Sullivan Prize recipient). Her contributions to conferences include roles as Program Chair for CVPR 2016 and CVPR 2017, and she has authored influential papers on topics like Video-Popup (ECCV 2014) and Modal Space (CVPR 2017). Current projects involve advancing neural parametric models and real-time 3D reconstruction techniques. Awards include the ERC Starting Grant and recognition for her team’s work in non-rigid reconstruction. She actively mentors students and collaborates on grants, with recent openings for postdocs and PhD candidates in 3D vision and robotics.
Angjoo Kanazawa is an Assistant Professor in the Department of Electrical Engineering and Computer Sciences (EECS) at the University of California, Berkeley. She leads the Kanazawa AI Research (KAIR) lab under the Berkeley Artificial Intelligence Research (BAIR) umbrella and serves on the advisory board of Wonder Dynamics. Her research focuses on the intersection of computer vision, computer graphics, and machine learning, with a particular emphasis on 4D reconstruction of dynamic scenes, neural radiance fields (NeRF), and systems that model human-environment interactions from 2D visual data. Education : Ph.D., Computer Science (2017), University of Maryland, College Park BA, Mathematics and Computer Science (2012), New York University (NYU) Her work aims to build systems that can capture, perceive, and understand complex 3D/4D worlds from photographs and videos, enabling applications in scene reconstruction, motion analysis, and generative modeling. She has pioneered techniques for scaling NeRFs across GPUs (NeRF-XL), developing open-source tools like nerfstudio and gsplat. Her recent publications focus on topics like self-occluded avatar recovery (SOAR), decentralized diffusion models, and 4D reconstruction of articulated objects for robotics. Kanazawa's research has been recognized with prestigious awards including the IEEE CS TCPAMI Young Researcher Award (2024) , Sloan Research Fellowship (2023) , and Google Faculty Research Award (2021) . Her lab has trained numerous students who now hold positions at leading institutions and companies like Anthropic, Meta Reality Labs, and Luma AI. Key Scientific Awards : IEEE CS TCPAMI Young Researcher Award (2024) Sloan Research Fellow (2023) Hellman Fellow (2022) Bakar Fellows Spark Award (2022) Google Faculty Research Award (2021) Her KAIR lab collaborates extensively with industry partners and academic institutions, including the Max Planck Institute and Google Research. She has served as an advisor for PhD students and postdocs who now lead teams at UC Berkeley, MIT, Stanford, and Luma AI, while her teaching includes graduate courses like CS 280A (Computer Vision) and CS 294-173 (Learning for 3D Vision).
Ajmal Mian is a Professor of Computer Science at the University of Western Australia (UWA), affiliated with the School of Physics, Maths and Computing. He holds an Australian Research Council Future Fellowship (2022) and leads research in Artificial Intelligence, Computer Vision, and Machine Learning. His work focuses on 3D computer vision, adversarial AI defense, and explainable AI. His research interests include 3D point cloud analysis, face recognition, human action recognition, and remote sensing. He has published over 300 papers and secured major grants from ARC, NHMRC, and DARPA, totaling millions in funding. He has supervised 29 PhD students and mentored 12 postdoctoral researchers. Key projects include 3D diffusion models for scene generation, robust 3D vision systems, and defense against AI deception attacks. He serves as a fellow of IAPR, an ACM Distinguished Speaker, and has editorial roles at IEEE Transactions on Neural Networks and Pattern Recognition. Research Awards: HBF Mid-Career Scientist of the Year, West Australian Early Career Scientist of the Year, IAPR Best Scientific Paper Award. Grants: ARC Discovery Projects, National Intelligence & Security Discovery grants, DARPA grants for AI security. His teaching spans computer vision, machine learning, and programming courses. Collaborations include defense, medical, and agricultural applications.
Alexei A. Efros is the Howard Friesen Professor in the EECS Department at the University of California, Berkeley, and a core member of the Berkeley Artificial Intelligence Research (BAIR) Lab. Previously, he spent a decade at Carnegie Mellon University’s Robotics Institute. His research focuses on data-driven computer vision, self-supervised learning, computational photography, and generative models. He has pioneered advancements in visual representation learning, including seminal work on neural radiance fields and generative adversarial networks. Education Background: Efros holds a PhD in Computer Science from MIT, though specific details of his academic journey are not explicitly provided in the text. His career includes postdoctoral research at the University of Oxford with Andrew Zisserman and collaborative work with Team WILLOW at INRIA Paris. Research Interests: Efros explores how vast uncurated visual data can be leveraged for understanding and synthesizing the visual world. Key areas include self-supervised learning, generative models, and applications in robotics and art. His lab has contributed influential techniques such as Style Transfer, GAN-based image synthesis, and neural scene representation learning. Recent work emphasizes real-time adaptation (Test-Time Training), 3D perception models, and ethical AI implications of generative systems. Publications: Over 150+ publications span topics like Generative Adversarial Networks (GANs), unsupervised learning, and visual-linguistic models. Notable works include Unpaired Image-to-Image Translation (CUT/GAU), Style Transfer , and Swapping Autoencoder . His research has significant industry impact, with techniques adopted in Adobe’s software and generative AI applications. Grants & Collaborations: Efros has secured major funding from NSF, DARPA, and industry partnerships (e.g., Adobe, NVIDIA). He co-leads projects on scalable vision models, ethical AI, and real-world perception systems. Current collaborations include work with MIT, NYU, and INRIA Paris. Labs & Teams: Leads the BAIR Vision Group at Berkeley, fostering interdisciplinary research between computer vision, graphics, and robotics. The group emphasizes Slow Science principles, prioritizing deep exploration over rapid publication.
Maria Gorlatova is an Associate Professor of Electrical and Computer Engineering at Duke University's Pratt School of Engineering, where she leads the Intelligent Interactive Internet of Things (I3T) Lab. She also holds a secondary affiliation as Faculty Network Member of the Duke Institute for Brain Sciences and has previously served as Assistant Professor of Computer Science. Dr. Gorlatova earned her Ph.D. in Electrical Engineering from Columbia University (2013), following M.Sc. and B.Sc. (Summa Cum Laude) degrees in Electrical Engineering from University of Ottawa, Canada. Prior to joining Duke, she was an Associate Research Scholar in the Electrical Engineering Department and Associate Director of the Princeton EDGE Lab at Princeton University (2016-2018). She also has industry experience with Telcordia Technologies, IBM, and D. E. Shaw Research. Her research focuses on advancing intelligent behavior in Internet of Things systems and applications, particularly in mobile pervasive systems and the Internet of Things. Her work crosses traditional discipline boundaries, requiring thinking across multiple layers of system and protocol stacks. Current research themes include breaking barriers for technologies that enable fundamentally new deployments and experiences, such as energy harvesting, artificial intelligence adapted to IoT constraints, and augmented reality. Her lab specifically develops edge- and IoT-enabled intelligent augmented reality platforms, with applications in healthcare and human-robot collaboration. Analyzing her recent publications reveals a strong focus on augmented reality systems, particularly for medical applications. Her work spans computer vision for AR, spatial tracking, SLAM systems, vision-language models for AR security, and VR/AR applications in neurosurgery and rehabilitation. A significant portion of her recent work addresses challenges in mixed reality for medical procedures, demonstrating the translational impact of her research. Google Anita Borg USA Fellowship Canadian Graduate Scholar CGS NSERC Fellowships Columbia University Presidential Fellowship Columbia University Jury Award for Outstanding Achievement in Communications ACM SenSys Best Student Demonstration Award IEEE Communications Society Young Author Best Paper Award IEEE Communications Society Award for Advances in Communications Best Research Artifact Award, IEEE IPSN (2020) N2 Women Rising Star, Networking Networking Women (N2Women) (2019) Dr. Gorlatova's research has been supported by various funding sources that enable her work on edge computing for augmented reality, IoT systems, and medical applications. She actively mentors graduate students who frequently appear as first authors on her publications, indicating strong student involvement in her research. Her I3T Lab at Duke focuses on creating human-facing pervasive mobile computing platforms that enable transformative applications, with recent emphasis on creating advanced augmented reality platforms that integrate edge computing and IoT technologies. The I3T Lab is developing next-generation AR systems with capabilities in edge AI, collaborative spatial awareness, AR user cognitive context sensing, and AR QoS/QoE evaluation. Current projects include applications in healthcare (particularly neurosurgery guidance and rehabilitation) and human-robot collaboration scenarios, demonstrating the lab's focus on real-world impact of pervasive computing technologies.
Shubham Tulsiani is an Assistant Professor at Carnegie Mellon University's Robotics Institute, where he leads the Computer Vision group and the Physical Perception Lab. His research focuses on inferring physically and spatially grounded representations from perceptual inputs, with applications in 3D vision, robot manipulation, and neural scene reconstruction. He directs an active research group with multiple PhD and Master's students. Research interests center on 3D scene understanding , robot learning , and generative modeling , with specific emphasis on: self-supervised perception, neural rendering, multi-view geometry, manipulation from visual inputs, and physics-based reasoning. The lab develops methods that leverage physical world constraints as supervisory signals. Recent publications demonstrate strong focus on diffusion models for 3D tasks , sparse-view reconstruction , and robotic manipulation transfer . Key trends include neural inverse rendering, view synthesis from limited observations, and translating human interactions to robot actions. Awards include: Best Student Paper Award at CVPR 2015 Advising includes supervision of 5 PhD students, 4 MS students, and undergraduates. Lab alumni hold positions at Google, Stanford, Meta, and Princeton. The Physical Perception Lab collaborates with FAIR Pittsburgh and the CMU Computer Vision group.
Prof. Matthias Nießner is a Professor at the Technical University of Munich , where he leads the Visual Computing Lab . Prior to this, he held a Visiting Assistant Professor position at Stanford University . His work bridges computer vision , graphics , and machine learning , focusing on 3D reconstruction , semantic scene understanding , and AI-driven video synthesis . Prof. Nießner has published over 150 works in top venues like SIGGRAPH , CVPR , and ECCV , with several receiving best paper awards (SIGCHI’14, HPG’15, SPG’18, SIGGRAPH’16 Emerging Tech). His research has garnered international media attention, including features in the New York Times , Wall Street Journal , and MIT Technological Review , as well as TV demonstrations (e.g., Jimmy Kimmel Live for Face2Face technology). Awards : TUM-IAS Rudolph Moessbauer Fellowship (2017–ongoing) Google Faculty Award (2017) Nvidia Professor Partnership Award (2018) ERC Starting Grant (2018, €1.5M) Eurographics Young Researcher Award (2019) Research Trends : 3D Gaussian Splatting for real-time rendering Neural Radiance Fields (NeRF) with mesh supervision Audio-driven facial animation via diffusion models Latent space diffusion for 3D scenes Self-supervised and zero-shot methods for 3D and image analysis As a co-founder and director of Synthesia Inc. , he drives democratization of synthetic media. His YouTube channel has over 5 million views, reflecting his impact beyond academia.