Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Katerina Fragkiadaki is the JPMorgan Chase Associate Professor of Computer Science in the Machine Learning Department at Carnegie Mellon University. She works at the intersection of Artificial Intelligence, Computer Vision, Machine Learning, Language Understanding, and Robotics. PhD from GRASP Lab, University of Pennsylvania Postdoctoral researcher at UC Berkeley (with Jitendra Malik) and Google Research Recipient of NSF CAREER, DARPA Young Investigator, Amazon, Google, Sony, UPMC, and AFOSR awards Organizer of CoRL 2023 Workshop on Generalist Robots ICLR 2024 Program Chair, multiple area chair roles Her research group focuses on developing machines that autonomously improve world models through human-environment interactions, with specific emphasis on: Representation learning and video understanding 2D/3D unified vision-language models Generative simulation and reinforcement learning Real2Sim/Sim2Real robot learning Continual learning and spatial common sense 3D scene reconstruction and dynamics Recent publications highlight advancements in: 3D mesh generation with compositional transformers Unified 2D/3D perception frameworks Physics-aware generative models Diffusion-based robotic manipulation policies Embodied agents with memory prompting Awards include: 2024: DARPA Young Investigator Award 2023: Amazon Faculty Award 2022: Sony Faculty Research Award 2021: UPMC Faculty Research Award 2020: NSF CAREER Award 2019: Google Faculty Award Key collaborations span institutions including UC Berkeley, Google Research, Stanford, MIT, and University of Tsukuba. Her work bridges theoretical innovation with practical applications in: Autonomous robot manipulation 4D world modeling Language-grounded perception Visual dynamics prediction Embodied program synthesis Physics-based simulation engines
Andrea Tagliasacchi is an Associate Professor in the School of Computing Science at Simon Fraser University (SFU), where he holds the Visual Computing Research Chair. He is also a part-time (20%) Staff Research Scientist at Google DeepMind in Toronto and holds an associate professor (status only) appointment in the Department of Computer Science at the University of Toronto. Education: PhD in Computing Science – Simon Fraser University (NSERC Alexander Graham Bell Fellow) Postdoctoral Research – École Polytechnique Fédérale de Lausanne (EPFL) MSc in Computer Science – Politecnico di Milano (Gold Medalist) His research lies at the intersection of computer vision, computer graphics, and machine learning, with a focus on 3D visual perception. Key areas include neural radiance fields (NeRF), 3D Gaussian splatting, inverse rendering, and geometric deep learning, with applications in robotics, augmented reality, and autonomous systems. His work emphasizes robust and efficient scene understanding and reconstruction from visual data. His recent publications, appearing in top venues like CVPR, SIGGRAPH, NeurIPS, and ECCV, demonstrate a strong emphasis on neural fields, 3D reconstruction, and generative modeling. Trends include improving rendering efficiency, enhancing robustness to noise and distractors, and enabling controllable and 3D-aware generation. His group has made significant contributions to Gaussian splatting, NeRF optimization, and diffusion-based 3D/4D synthesis. Scientific Awards: 2024 CVPR Best Paper Award (Honorable Mention) 2020 CVPR Best Student Paper Award 2015 SGP Best Paper Award NSERC Alexander Graham Bell Canada Graduate Scholarship MITACS Best Paper Award (SIGGRAPH Asia 2009) NSF Best Poster Award (SGP 2012) He has advised numerous PhD and MSc students, many of whom are now researchers at leading institutions and companies. His research has been supported through collaborations with Google, Intel, and academic partners. He serves the community as a Senior Area Chair for CVPR 2025, Associate Editor for IEEE TPAMI (2024–2026), Guest Editor for IEEE TPAMI on 3D GenAI, and Program Chair for 3DV 2024. He leads a vibrant research lab at SFU focused on pushing the boundaries of 3D scene understanding with machine learning.
Andrew Zisserman is a Royal Society Research Professor at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). His research focuses on computer vision, artificial intelligence, and neural networks, with significant contributions to multimodal learning, video understanding, and 3D scene analysis. He leads projects exploring visual-language models, audio-visual synchronization, and clinical imaging applications. Key research areas include: Video analysis and temporal modeling Multimodal systems for sign language translation and action recognition 3D shape estimation and physical property inference Foundation models and cross-modal retrieval Recent work highlights: Developed Flamingo and Tapir models for video-language tasks Advancements in spinal MRI analysis and clinical imaging Leadership in EGO4D and VoxCeleb challenges Honors include Fellowship of the Royal Society (FRS) and the ISSLS Prize in Clinical Science 2023 for spinal analysis innovations. His lab collaborates globally, emphasizing real-world applications in healthcare and autonomous systems.
Angjoo Kanazawa is an Assistant Professor in the Department of Electrical Engineering and Computer Sciences (EECS) at the University of California, Berkeley. She leads the Kanazawa AI Research (KAIR) lab under the Berkeley Artificial Intelligence Research (BAIR) umbrella and serves on the advisory board of Wonder Dynamics. Her research focuses on the intersection of computer vision, computer graphics, and machine learning, with a particular emphasis on 4D reconstruction of dynamic scenes, neural radiance fields (NeRF), and systems that model human-environment interactions from 2D visual data. Education : Ph.D., Computer Science (2017), University of Maryland, College Park BA, Mathematics and Computer Science (2012), New York University (NYU) Her work aims to build systems that can capture, perceive, and understand complex 3D/4D worlds from photographs and videos, enabling applications in scene reconstruction, motion analysis, and generative modeling. She has pioneered techniques for scaling NeRFs across GPUs (NeRF-XL), developing open-source tools like nerfstudio and gsplat. Her recent publications focus on topics like self-occluded avatar recovery (SOAR), decentralized diffusion models, and 4D reconstruction of articulated objects for robotics. Kanazawa's research has been recognized with prestigious awards including the IEEE CS TCPAMI Young Researcher Award (2024) , Sloan Research Fellowship (2023) , and Google Faculty Research Award (2021) . Her lab has trained numerous students who now hold positions at leading institutions and companies like Anthropic, Meta Reality Labs, and Luma AI. Key Scientific Awards : IEEE CS TCPAMI Young Researcher Award (2024) Sloan Research Fellow (2023) Hellman Fellow (2022) Bakar Fellows Spark Award (2022) Google Faculty Research Award (2021) Her KAIR lab collaborates extensively with industry partners and academic institutions, including the Max Planck Institute and Google Research. She has served as an advisor for PhD students and postdocs who now lead teams at UC Berkeley, MIT, Stanford, and Luma AI, while her teaching includes graduate courses like CS 280A (Computer Vision) and CS 294-173 (Learning for 3D Vision).
Achuta Kadambi, Ph.D., is an Associate Professor at UCLA in Electrical Engineering and Computer Science, leading an interdisciplinary research group focused on AI, computational imaging, and bias mitigation in medical technologies. He recruits PhD students from EE, CS, and Bioengineering departments and has commercialized research through two California-based companies. His research investigates the intersection of physics and artificial intelligence, with a focus on unbiased low-level vision systems. Current projects explore how light transport interacts with human skin variations to identify and correct imaging biases in facial recognition and medical devices. His work has produced over 70 patents, with 30+ issued, and a textbook Computational Imaging (MIT Press, 2022). NSF CAREER Award (2021) for light transport bias research DARPA Young Faculty Award (2021) for AI and medical imaging innovations ARO Young Investigator Program (2021) for computational sensing IEEE-HKN Under 35 Award (2022) for inclusive EECS inventions Forbes 30 Under 30 recognition His recent publications focus on polarization imaging, 3D Gaussian splatting, synthetic data generation for healthcare, and bias mitigation in machine learning. Collaborations with UCLA medical school faculty, including Dr. Laleh Jalilian, aim to deploy these innovations in clinical settings. Current teaching includes ECE 149: Foundations of Computer Vision (Fall 2024, Spring 2025) and ECE 102: Signals and Systems (Winter 2024).
David B. Lindell is an Assistant Professor in the Department of Computer Science at the University of Toronto, with affiliations to the Vector Institute and AXL. He is a founding member of the Toronto Computational Imaging Group. His research focuses on physically based intelligent sensing, integrating physical models, signal processing, and AI to advance sensing systems. Notable projects include imaging around corners, through scattering media, and developing machine learning algorithms for 3D scene reconstruction. Education: Ph.D. in Computational Imaging from Stanford University (advisor: Gordon Wetzstein). Awards include the 2024 Ontario Early Researcher Award and the Best Student Paper at CVPR 2025. His work combines computational imaging with applications in computer graphics and autonomous systems. Research interests span non-line-of-sight imaging, single-photon sensing, and neural representations. Key contributions include the Light-Cone Transform (Nature 2018), confocal diffuse tomography (Nature Communications 2020), and AutoInt (CVPR 2021). His lab develops systems for 3D reconstruction, transient imaging, and photon-efficient sensors. Selected grants and support: NSF CAREER Award, DARPA REVEAL program, and KAUST Visual Computing Center funding. Active collaborations with industry and academic institutions on autonomous driving and medical imaging applications.
Prof. Dr. Dennis Säring is a faculty member at the University of Applied Sciences Wedel , specifically affiliated with the School of Engineering. His academic and research activities focus on Deep Learning , Medical Image Analysis , and applications of Artificial Intelligence in healthcare and biomedical imaging. He has led seminars on Deep Learning topics and supervised student projects in Autonomous Driving at Audi's AADC 2018 competition. Research Highlights : Cardiovascular imaging, forensic age estimation via MRI, neural network-based bone segmentation, and cerebrovascular aneurysm analysis. Technical Expertise : Cardiac MRI, 3D/4D image processing, parametric mapping, and spatiotemporal data fusion. His recent publications (2018-2023) emphasize 3D MR segmentation for age assessment, CMR strain analysis in athletes, and T1/T2 mapping for myocarditis. Key collaborations include institutions like the University Medical Center Hamburg-Eppendorf and Wedler Hochschulbund, with funding for autonomous vehicle research. While no explicit scientific awards are listed, his work spans clinical cardiology, forensic radiology, and AI-driven medical diagnostics.
Cyrill Stachniss is a Full Professor at the University of Bonn , where he heads the Lab for Photogrammetry and Robotics and is affiliated with the Lamarr Institute for Machine Learning and Artificial Intelligence . He was previously a Visiting Professor in Engineering at the University of Oxford until 2025. His academic journey includes positions at the University of Freiburg, University of Zaragoza, and the Swiss Federal Institute of Technology. University: University of Bonn School: Faculty of Engineering Department: Department of Photogrammetry Academic Rank: Professor His research spans robotics, photogrammetry, SLAM, autonomous navigation, perception systems, agricultural robotics, and unmanned aerial vehicles . He emphasizes probabilistic techniques for mobile robots and has made significant contributions to visual and LiDAR-based localization, scene understanding, and 3D reconstruction. The recent publications highlight a strong trend toward neural implicit representations, agricultural phenotyping, radar-based perception, and active learning . His team develops robust systems for real-world deployment in dynamic and unstructured environments, particularly in precision farming and autonomous vehicles. Scientific Awards: IEEE RAS Early Career Award (2013) Microsoft Research Faculty Fellow (2010) 7th EURON Georges Giralt Award (2008) Multiple Best Paper Awards at ICRA, IROS, RSS, and RAL Faculty Teaching Award, University of Freiburg He has advised numerous students and leads the DFG Cluster of Excellence PhenoRob and Research Unit FOR 1505 Mapping on Demand . His lab has co-founded three startups, reflecting strong industry and societal impact. He also runs the educational video series 5 Minutes with Cyrill , explaining key robotics concepts.
Gordon Wetzstein is an Associate Professor of Electrical Engineering and, by courtesy, Computer Science at Stanford University. He leads the Stanford Computational Imaging Lab and co-directs the Stanford Center for Image Systems Engineering (SCIEN). His research focuses on computational imaging, wearable computing, and neural rendering, blending computer graphics, vision, AI, and optics. Education: Ph.D., Computer Science, University of British Columbia (2011) Diploma, Media Systems Science, Bauhaus University (2006) Research Interests: His work spans computational displays , holography , non-line-of-sight imaging , and AI-driven optical systems . Key projects include Autofocals (gaze-contingent eyeglasses) and neural holography systems. He explores applications in AR/VR, medical imaging, and scientific visualization. Publications: Recent work includes advances in 3D holography, gaze-tracking systems, and AI-optics integration. His papers address challenges in display efficiency, light-field processing, and real-time imaging. Awards: Fellow of Optica NSF CAREER Award (2016) PECASE (2019) ACM SIGGRAPH Significant New Researcher Award (2018) Advising & Grants: He advises over 20 doctoral and postdoctoral students. His lab collaborates with industry (e.g., Raxium, Google) and has secured grants from NSF, DARPA, and private foundations. Labs & Teams: His lab develops cutting-edge systems like neural holography and non-line-of-sight imaging. The SCIEN center fosters interdisciplinary image systems research.
Dylan Campbell is a Lecturer in Computing at the Australian National University (ANU), affiliated with the ANU College of Systems & Society. His research focuses on computer vision, optimization, and robotics, particularly in 3D vision and deep learning applications. He has held prior roles as a Research Fellow at the University of Oxford’s Visual Geometry Group and ANU’s Australian Centre for Robotic Vision. Campbell holds a PhD from ANU (2018) and a BE in Mechatronic Engineering from UNSW (2012). Research interests include geometric sensor alignment, neural radiance fields, and differentiable optimization layers. He actively supervises students (7 PhD/DPhil, 3 MEng, 9 honours) and teaches advanced courses in computer vision and robotics. Notable awards include the Marr Prize Honourable Mention (2017) and the IEEE Australia Council Postgraduate Student Paper Competition (2018). He has organized workshops at ECCV and CVPR, served as a reviewer for top conferences like CVPR/ICCV/ECCV, and contributed to datasets like SEED4D and RefRef. His work emphasizes efficient training of neural networks and leveraging symmetries in data for long-range connections.
Manuel Kaufmann is a Lecturer in the Department of Computer Science at ETH Zürich. His work focuses on advanced 3D human motion capture, sensor-based systems, and computer vision applications. He is affiliated with the Institute of Informatics (inf.ethz.ch) and contributes to research in real-time motion tracking, dataset development, and machine learning integration for human-robot interaction. Research interests include holistic human-scene reconstruction from monocular videos, gaze estimation using EEG signals, and expressive avatar creation. His projects emphasize practical applications in robotics, sports analytics, and biomedical engineering, often leveraging electromagnetic and inertial sensors for high-precision data acquisition. His publications reflect a trend toward multi-modal data fusion, real-world dataset creation (e.g., WorldPose, ARCTIC), and addressing challenges in loose garment modeling (Reloo). These efforts aim to improve markerless motion capture, crowd analysis, and human-robot collaboration. No scientific awards or grants are explicitly listed. He has no documented advisees, though his research may involve collaborations with students or teams. His office is located at OAT X 23, Andreasstrasse 5, Zürich, Switzerland, and contact details include a phone number and professional email.
Associate Professor Mohsen Kalantari is a Geospatial Engineering academic at the University of New South Wales (UNSW) School of Civil and Environmental Engineering , with concurrent roles as co-founder of the startup Faramoon . His career spans roles at the University of Melbourne's Department of Infrastructure Engineering and Victorian government's land administration initiatives through DELWP. Education : PhD in Geomatics Engineering (2008, University of Melbourne), Master of GIS Engineering (2004), Bachelor of Surveying Engineering (2001) His research bridges geospatial engineering with construction automation , focusing on 3D cadastre , BIM-GIS integration , and smart cities . Recent publications show trends in underground land administration , digital twins , and LADM standard implementations . Scientific Awards : National educational recognition (2019), Victorian educational grants (2018), and prestigious fellowships (2012) As a supervisor , he guides PhD candidates in topics ranging from BIM for waste management to underground cadastral systems . His industry engagement includes partnerships with the United Nations , Open Geospatial Consortium , and Singapore Land Authority .
Kevin C. Zhou is an Assistant Professor in the Department of Biomedical Engineering at the University of Michigan. His research focuses on developing high-performance computational optical imaging systems with unprecedented spatiotemporal throughput, integrating advanced optical instrumentation with machine learning-driven algorithms to analyze big data in biology and medicine. His lab specializes in creating imaging systems capable of capturing high-resolution, high-speed, and high-dimensional datasets. Dr. Zhou holds a Ph.D. in Biomedical Engineering from Duke University (NSF GRFP Fellow) and a B.S. in Biomedical Engineering from Yale University (Barry Goldwater Scholar). Prior to joining U-M, he was a Schmidt Science Fellow and postdoctoral researcher at UC Berkeley. Key research areas include: High-throughput microscopy (gigapixel-scale systems) 3D tomographic imaging Light field and Fourier-based imaging modalities Machine learning for image reconstruction and analysis Biomedical applications in cellular/molecular imaging His recent work has advanced technologies like multi-camera array microscopes (MCAM/MCAS) and Fourier light field mesoscopes, achieving video-rate 3D imaging of freely moving organisms. These innovations enable applications in digital cytopathology, behavioral tracking, and high-content biological studies. Notable awards include the NSF Graduate Research Fellowship and Barry Goldwater Scholarship. His research has been featured in top journals and conferences with a focus on advancing optical imaging hardware and computational pipelines.
Chuang Gan is an Assistant Professor at the University of Massachusetts Amherst, affiliated with the College of Information and Computer Sciences and the Department of Computer Science. His work focuses on advancing artificial intelligence, robotics, computer vision, and embodied agents through interdisciplinary research combining neural networks, physical simulations, and multimodal learning. Research interests include generative models, reinforcement learning, vision-language integration, and scalable autonomous systems. He explores topics like world modeling for robots, adaptive policy learning, and physics-driven AI. His projects often involve creating systems that learn from visual, auditory, and tactile inputs to perform complex tasks such as object manipulation, navigation, and decision-making in dynamic environments. Recent research trends emphasize embodied AI systems capable of long-horizon planning, compositional reasoning, and efficient learning from limited data. His work bridges theory and practice, with applications in robotics, simulation platforms, and multimodal generation. Key contributions include frameworks for 3D scene understanding, adaptive world models, and novel training paradigms for large language models. His research has been applied to robotics platforms like RoboDreamer and UBSoft, focusing on unbounded soft environments. Collaborations involve designing benchmarks for physical scene understanding (e.g., Physion++), and creating tools like DiffTactile for tactile simulation. His work often integrates principles from differential geometry, PDE dynamics, and game theory. Chuang Gan’s research group develops open-source tools and benchmarks, such as the SoftZoo robot co-design platform and the SOK-Bench situated reasoning benchmark. His team emphasizes scalable alignment methods beyond human supervision and explores ethical AI through principles like symmetry-enhanced training.