Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Prof. Ingmar Posner is a leading figure in applied artificial intelligence at the University of Oxford, where he serves as Principal Investigator for the Applied Artificial Intelligence Lab (A2I) and founding Director of the Oxford Robotics Institute. His work focuses on enabling robots to operate effectively in complex real-world environments through experience-driven learning. Key research areas: robot learning, scene interpretation, data-efficient learning, and transfer learning Applications in manipulation, autonomous driving, logistics, and space exploration His team has produced groundbreaking work in world models, sim-to-real transfer, and constraint-based manipulation systems (e.g., COMBO-Grasp). Notable contributions include the TWIST distillation framework and foundational research in tactile data generation (TactGen). He has received multiple best paper awards at top robotics venues. Publications reveal evolving research themes: 2025 work emphasizes language-conditioned learning (Lumos) and multi-agent decision-making, while 2024 focused on diffusion models for locomotion and differentiable simulators. Earlier work spans from urban scene analysis to physically plausible scene synthesis (RELATE).
Justin Johnson is an Assistant Professor at the University of Michigan's College of Engineering, Department of Electrical Engineering and Computer Science, and a Research Scientist at Facebook AI Research (FAIR). His work bridges computer vision, machine learning, and deep learning, focusing on visual reasoning, vision-language tasks, image generation, and 3D reasoning using neural networks. PhD, Stanford University (advised by Fei-Fei Li) His research interests span visual reasoning , vision and language , 3D vision , and image generation , with a focus on innovative applications of deep neural networks. Recent publications highlight work on 3D consistency, self-supervised learning, and multimodal integration of vision and text. Notable contributions include PyTorch3D for 3D data processing, and foundational work in visual question answering , neural style transfer , and scene graph-based image generation . Publications span top conferences like ICCV, CVPR, and NeurIPS. He teaches courses including EECS 498/598: Deep Learning for Computer Vision and EECS 442: Computer Vision at University of Michigan, with prior involvement in Stanford's CS 231N in co-teaching roles. Software projects include open-source frameworks like fast-neural-style for real-time artistic style transfer, and PyTorch3D for efficient 3D deep learning. These tools demonstrate his commitment to practical implementations and community-driven research.
Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Christian Theobalt is a Professor of Computer Science at Saarland University and Scientific Director of the Visual Computing and Artificial Intelligence Department at the Max Planck Institute for Informatics . He leads the Saarbruecken Center for Visual Computing as a strategic partnership between Google and MPI. PhD in Computer Science (2005) from MPI-INF/Saarland University Postdoctoral Researcher at MPI (2005-2007) Visiting Assistant Professor at Stanford (2007-2009) His research focuses on the intersection of Computer Graphics, Computer Vision, and Artificial Intelligence , with specializations in: 3D/4D Human Reconstruction Neural Rendering Performance Capture Geometric Deep Learning Volumetric Video Quantum Visual Computing Recent article trends show emphasis on: Quantum computing applications in visual reconstruction Neural rendering with radiance fields Human motion capture from egocentric views Gesture-language interaction modeling Multi-modal scene understanding Real-time free-viewpoint rendering Scientific Awards : Fellow of EUROGRAPHICS (2022) CVPR Best Student Paper Honorable Mention (2020) ERC Consolidator Grant (2017) Karl Heinz Beckurts Award (2017) Busy Beaver Teaching Award (2016) ERC Starting Grant (2013) Advising over 50 students and researchers including: Current researchers: Viktor Rudnev, Linjie Lyu, Mohit Mendiratta Postdocs: Kwang In Kim, Kiran Varanasi Alumni: Franziska Mueller (Google), Dushyant Mehta (Qualcomm), Ayush Tewari (MIT) Grants include ERC grants, Google Glass Research Award, and multiple industry partnerships. His lab maintains cutting-edge facilities with: Multi-camera capture systems Quantum annealing infrastructure HDR radiance field technology GPU clusters for AI research Time-of-flight imaging systems Event camera arrays
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Hao Su is an Associate Professor in the Department of Computer Science and Engineering at University of California, San Diego . He serves as Chairman & CTO of Hillbot Inc , and leads the SU Lab which focuses on building autonomous systems that learn actively in physical environments. His affiliations include the Institute for Learning-enabled Optimization at Scale , Artificial Intelligence Group , Contextual Robotics Institute , Halicioğlu Data Science Institute , and Center for Visual Computing . As a researcher in Computer Vision, Robotics, and Neural Geometry , he has made significant contributions to 3D foundation models, reward-free world models, diffusion policy frameworks, and GPU-accelerated simulation environments. His 2024-2025 publications include advancements in hand-eye calibration, dynamic mesh reconstruction, and multi-stage robotic manipulation. His scientific awards include: Frontiers of Science Award (2025) TPAMI Young Research Award (2025) NSF CAREER Award (2023) ACM SIGGRAPH Best Doctorate Thesis Honorable Mention (2019) He has served as Program Chair for CVPR 2025 and Area Chair for ICLR 2022 and NeurIPS 2023 , while previously serving as Publication Chair for 3DV 2016 and Program Committee for SIGGRAPH Asia Workshops .
Kyros Kutulakos is a Professor in the Department of Computer Science at the University of Toronto, where he leads research in computational imaging and 3D sensing. His affiliations include the Toronto Computational Imaging Group, Computer Vision Group, Dynamic Graphics Project (DGP), and Vector Institute Group. He teaches graduate and undergraduate courses such as CSC320 (Introduction to Visual Computing) and CSC2530 (Computational Imaging & 3D Sensing). His research interests span computational imaging, non-line-of-sight imaging, single-photon detectors, 3D sensing, and neural rendering. Notable contributions include advancements in structured-light imaging, time-of-flight systems, and super-oscillatory microscopy. He has advised numerous PhD and MSc students, fostering cutting-edge research in imaging technologies. Kutulakos has received prestigious awards, including the Dean’s Research Excellence Award (2023) and multiple best paper prizes (e.g., Marr Prize at ICCV 2023). He has served as program chair for ICCV 2013, ICCP 2010, and CVPR 2003, contributing to academic leadership in computer vision. His work bridges optics, photonics, and computation, with applications in autonomous systems, medical imaging, and astronomy. Current research focuses on extreme imaging scenarios, such as imaging in pitch-black environments and around corners, leveraging novel sensor designs and computational techniques.
Noah Snavely is a Professor of Computer Science at Cornell Tech and the Cornell Ann S. Bowers College of Computing and Information Science. His research spans computer vision and graphics with a focus on 3D scene understanding from image collections. He also works at Google DeepMind in NYC and leads the Cornell Graphics and Vision Group. His research interests include computer vision and computer graphics , particularly in recovering 3D structure from large photo collections, neural radiance fields, image-based rendering, and scene understanding. His work enables applications in mapping technologies, immersive VR experiences, and synthesizing 3D worlds from text prompts with implications for game design and filmmaking. Recent publications show a strong trend toward generative 3D modeling and neural rendering , with significant contributions in neural radiance fields, view synthesis, and dynamic scene reconstruction. His research group explores unstructured photo collections to develop new technology for 3D world modeling and image analysis. PECASE Microsoft New Faculty Fellowship Alfred P. Sloan Fellowship SIGGRAPH Significant New Researcher Award ACM Fellow (2023) IEEE Fellow NSF CAREER Award Helmholtz Prize Snavely has mentored numerous PhD students and postdocs including Ruojin Cai, Qianqian Wang, and Zhengqi Li. His teaching includes CS5670: Computer Vision across multiple semesters. He has received the Cornell College of Engineering Mr. and Mrs. Richard F. Tucker Teaching Excellence Award (2012). At Cornell Tech, his research group develops technology for modeling the world in 3D from online photo collections and analyzing images for computer graphics applications. His work has been featured in major news outlets including the New York Times, New Scientist, and MIT Technology Review.
Nima Fazeli is an Assistant Professor of Robotics at the University of Michigan (2020–Present), holding courtesy appointments in Computer Science & Engineering (CSE) and Mechanical Engineering. He directs the Manipulation and Machine Intelligence (MMint) Lab, focusing on enabling dexterous robotic manipulation through multimodal representation learning, tactile sensing, and model-based reasoning. His work integrates mechanics, perception, controls, and planning to achieve autonomous interaction with uncertain environments. Education: PhD, MIT (2019); MSc, University of Maryland (2014); BSc, Amirkabir University of Technology (2011) Research interests emphasize embodied intelligence , including visuo-tactile fusion, contact dynamics modeling, and cross-modal learning. Recent work explores tactile shadows, deformable object manipulation, and language-guided robot control. His research is supported by the NSF CAREER grant and National Robotics Initiative, with applications in manufacturing, assistive robotics, and space systems. Publications span topics like tactile sensing hardware (e.g., GelSlim 4.0), visuo-tactile implicit representations (ViTaSCOPE), and failure recovery policies (Racer). His team’s work has been featured in outlets like The New York Times and BBC. Key Awards: NSF CAREER Grant (2024) Teaching includes Introduction to Robotic Manipulation . Collaborations involve cross-disciplinary projects with mechanical, electrical, and biomedical engineering groups.
Olga Sorkine-Hornung is a Professor of Computer Science at ETH Zurich and head of the Institute of Visual Computing. She leads the Interactive Geometry Lab, focusing on theoretical and practical advancements in digital content creation, geometry processing, and shape modeling. Current position: ETH Zurich, Department of Computer Science Previous roles: Courant Institute (NYU), Technical University of Berlin Education: BSc and PhD from Tel Aviv University, postdoc at TU Berlin Her research spans shape representation, digital fabrication, computer animation, and fundamental geometry processing. Key contributions include Laplacian surface editing, as-rigid-as-possible deformation, and generalized winding numbers. She works on applications in VR, AR, and autonomous systems. Awards include Test of Time Awards (2024), ACM Fellow (2020), ERC Consolidator Grant (2020), and EUROGRAPHICS Young Researcher Award (2008). She has supervised numerous students and co-developed software libraries like libigl and Instant Meshes . Co-chair roles for SIGGRAPH, Eurographics, and Pacific Graphics Editorial board member for ACM Transactions on Graphics and other journals Keynote speaker at VMV, CVPR, and SIAM conferences Her work bridges mathematical rigor with practical implementation, advancing computer graphics and geometry processing through intuitive algorithms that maintain surface detail while enabling efficient computation.
Prof. Matthias Nießner is a Professor at the Technical University of Munich, leading the Visual Computing Lab. His research intersects computer graphics, vision, and AI, focusing on 3D reconstruction, semantic understanding, and AI-driven video synthesis. He holds a PhD from the University of Erlangen-Nuremberg (2013) and was a Visiting Assistant Professor at Stanford University (2013–2017). Notable awards include the ERC Starting Grant (2018), Nvidia Professorship Award, and Eurographics Young Researcher Award (2019). His work has been featured in mainstream media and led to startups like Synthesia Inc. Research spans Gaussian splatting, neural radiance fields, and generative AI for 3D avatars. Over 150 publications include SIGGRAPH, CVPR, and ECCV, with best paper awards. Projects like Face2Face and ScanNet have driven innovation in facial reenactment and 3D scene datasets. Education: PhD in Computer Science, University of Erlangen-Nuremberg (2013) Diploma in Computer Science, University of Erlangen-Nuremberg (2010) Research Interests: 3D digitization, neural rendering, generative AI, non-rigid reconstruction, and applications in AR/VR. Awards: ERC Starting Grant (2018) Nvidia Professorship Award (2018) Google Faculty Award (2018) SIGGRAPH Best Emerging Tech Award (2016) Grants: Over €1.5M from ERC and industry partnerships. Labs/Teams: Visual Computing Lab at TUM and Synthesia Inc. (co-founder). Key projects include ScanNet (large 3D indoor dataset), Face2Face (real-time facial reenactment), and Gaussian-based 3D avatars. Current work focuses on diffusion models, neural radiance fields, and AI-generated media detection.
Alexei A. Efros is a Professor in the Department of Electrical Engineering and Computer Sciences (EECS) at UC Berkeley, where he holds the Howard Friesen Professorship and is affiliated with the Berkeley Artificial Intelligence Research (BAIR) Lab. He previously served on the faculty at the Robotics Institute of Carnegie Mellon University (CMU) and completed a postdoctoral fellowship at the University of Oxford. His research spans data-driven computer vision, self-supervised learning, computational photography, and applications to computer graphics and robotics. His research interests include: Data-Driven Computer Vision Self-Supervised and Unsupervised Learning Generative Models and Image Synthesis Visual Representation Learning Applications in Robotics and Human-Computer Interaction Intersections with Human Vision and the Humanities The recent publications highlight a strong trend toward self-supervised learning, visual reasoning, and generative modeling, particularly diffusion models and 3D scene understanding. His work increasingly bridges computer vision with language, robotics, and cognitive science, emphasizing interpretability and real-world applicability. There is a clear focus on leveraging unlabeled data and developing methods for robust, generalizable AI systems. His scientific awards and recognitions include: Berkeley Fellowship Google Fellowship Soros Fellowship NSF Fellowship SIGGRAPH Outstanding Doctoral Dissertation Award Facebook Fellowship Adobe Fellowship CMU School of Computer Science Distinguished Dissertation Award ACM Doctoral Dissertation Honorable Mention Alexei Efros has advised numerous PhD students and postdocs, many of whom have gone on to faculty positions at top institutions including CMU, Stanford, MIT, Columbia, NYU, and Georgia Tech. His lab has received research funding from major tech companies and federal agencies, though specific grants are not detailed in the text. He teaches core computer vision and machine learning courses at both undergraduate and graduate levels at UC Berkeley. His research group is highly active, with ongoing projects in 3D perception, generative modeling, and vision-language systems. He leads a vibrant research lab at UC Berkeley, part of the BAIR consortium, collaborating with leading researchers such as Jitendra Malik, Trevor Darrell, Pieter Abbeel, and Angjoo Kanazawa. His lab fosters strong interdisciplinary connections with institutions worldwide, including Oxford, INRIA, and École Normale Supérieure.
Anthony Rowe is the Siewiorek and Walker Family Professor of Electrical and Computer Engineering at Carnegie Mellon University (CMU) and a Chief Scientist at Bosch Research. His primary affiliation is with the CyLab and the Wireless, Sensing and Embedded Systems (WiSE Lab) at CMU. He specializes in networked embedded systems, sensor networks, and extended reality (XR) technologies. His research emphasizes energy-efficient sensing, real-time localization, and XR integration with physical systems. Research Focus: His work spans XR systems (e.g., AR/VR edge networking in ARENA), mmWave radar for sensing (e.g., tire wear monitoring via Osprey), distributed edge computing (Silverline), and low-power wide-area networking (OpenChirp). Recent efforts include AI-integrated XR platforms (XaiR) and radar tomography (DART). Grants & Projects: Leads the CONIX Research Center ($27.5M NSF/DARPA grant), Bosch-funded edge computing projects, and DOE initiatives on microgrids. Notable projects include ARENA (XR edge architecture), GridBallast (smart grid control), and rural microgrid deployments in Haiti. Awards: Best Student Paper (ISMAR 2024), Best Paper (IPSN 2020), and the Steven J. Fenves Research Award (2015). Recognized for innovations in localization (MobiCom 2021), radar (ICRA 2023), and energy systems (BuildSys 2010). Teaching: Teaches courses on embedded systems (18-349/18-449), real-time systems, and mixed reality (18-453). Courses emphasize hands-on design and real-world applications. Labs & Teams: Directs the WiSE Lab, collaborating with Bosch Research and industry partners. The lab develops open-source frameworks like ARENA and OpenChirp, and contributes to standards for edge computing and sensing.
Mathieu Salzmann is a Senior Scientist and Lecturer at École Polytechnique Fédérale de Lausanne (EPFL), affiliated with the Computer Vision Laboratory (CVLAB) in the School of Computer and Communication Sciences (IC). He also holds a courtesy appointment with the EPFL College of Humanities and serves as Deputy Chief Data Scientist at the Swiss Data Science Center (SDSC). He has held concurrent roles in teaching units including SIN, SODH, and SSC, reflecting his interdisciplinary engagement. His research focuses on the intersection of machine learning and computer vision, particularly in deep learning for 2D and 3D visual scene understanding, efficient and robust models, domain adaptation, and interpretable AI. These interests are evident across his extensive publication record in top-tier venues. His recent publications (2023–2024) show a consistent trend in advancing deep learning methods for visual recognition, with strong representation at CVPR, ICCV, ECCV, ICML, ICLR, and NeurIPS. Topics include domain generalization, 3D understanding, model robustness, and multimodal learning, often with applications in real-world systems. His editorial roles as Associate Editor for IEEE TPAMI and Action Editor for TMLR further highlight his leadership in the field. Area Chair: ICML 2023, CVPR 2023, ICCV 2023, NeurIPS 2023, AAAI 2024, ECCV 2024 Associate Editor: IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) Action Editor: Transactions on Machine Learning Research (TMLR) Mathieu Salzmann has supervised numerous PhD students at EPFL, both current and past, including Bouquet Yann Yanis, Javed Saqib, Li Shuangqi, and others. He has also been involved in research grants and collaborative projects, such as his work with S. Süsstrunk and R. Baroni on comics reconfiguration. His part-time role as Senior GNC Engineer at ClearSpace (2020–2024) illustrates his applied research engagement in aerospace systems. He is actively involved in EPFL’s data science and AI research ecosystem through SDSC and multiple labs.