Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Noah Snavely is a Professor of Computer Science at Cornell Tech and the Cornell Ann S. Bowers College of Computing and Information Science. His research spans computer vision and graphics with a focus on 3D scene understanding from image collections. He also works at Google DeepMind in NYC and leads the Cornell Graphics and Vision Group. His research interests include computer vision and computer graphics , particularly in recovering 3D structure from large photo collections, neural radiance fields, image-based rendering, and scene understanding. His work enables applications in mapping technologies, immersive VR experiences, and synthesizing 3D worlds from text prompts with implications for game design and filmmaking. Recent publications show a strong trend toward generative 3D modeling and neural rendering , with significant contributions in neural radiance fields, view synthesis, and dynamic scene reconstruction. His research group explores unstructured photo collections to develop new technology for 3D world modeling and image analysis. PECASE Microsoft New Faculty Fellowship Alfred P. Sloan Fellowship SIGGRAPH Significant New Researcher Award ACM Fellow (2023) IEEE Fellow NSF CAREER Award Helmholtz Prize Snavely has mentored numerous PhD students and postdocs including Ruojin Cai, Qianqian Wang, and Zhengqi Li. His teaching includes CS5670: Computer Vision across multiple semesters. He has received the Cornell College of Engineering Mr. and Mrs. Richard F. Tucker Teaching Excellence Award (2012). At Cornell Tech, his research group develops technology for modeling the world in 3D from online photo collections and analyzing images for computer graphics applications. His work has been featured in major news outlets including the New York Times, New Scientist, and MIT Technology Review.
California Institute of Technology (Caltech)United States
Xingxing Zuo is an Assistant Professor (tenure-track) in the Robotics Department at MBZUAI. He holds a PhD from Zhejiang University (2021) and a Bachelor’s from UESTC (2016). Previously, he was a Postdoctoral Scholar at Caltech (2024–2025), a Postdoc at ETH Zurich (2019–2021), and held visiting roles at TU Munich, University of Delaware, and University of Technology Sydney. His research focuses on robotics, 3D computer vision, and embodied AI, with emphasis on robot-human collaboration, state estimation, and sensor fusion. Educations: PhD in Robotics, Zhejiang University (2021, with honors) Bachelor’s in Computer Science, University of Electronic Science and Technology of China (2016, with honors) Research Highlights: Develops novel methods for LiDAR-camera-inertial fusion, neural radiance fields, and radar-cameras systems Pioneered techniques like Flying Co-Stereo (long-range aerial mapping) and FMGS (vision-language embedded 3D splatting) Focuses on real-time SLAM, robust depth estimation, and photorealistic scene reconstruction Awards & Recognition: Best Paper Finalist at ICRA 2021 (CodeVIO) Oral Presentation at ICCV 2021 (MBA-VO) Recipient of Google Visiting Faculty Researcher (2023) Grants & Labs: Organized Thermal Infrared in Robotics workshop at ICRA 2025 Leads research on embodied AI and multi-sensor SLAM systems Develops open-source tools like LIC-Fusion and Coco-LIC frameworks
University of Illinois Urbana-ChampaignUnited States
Derek W Hoiem is a Professor in the Siebel School for Computing and Data Science at the University of Illinois Urbana-Champaign, where he has been a faculty member since 2009. His research focuses on computer vision and related areas, and he is also the co-founder and Chief Science Officer of Reconstruct, an AI-based construction technology company. His educational background includes: PhD in Robotics, Carnegie Mellon University (2007) Beckman Postdoctoral Fellowship (2008) Prof. Hoiem's research spans computer vision, with a focus on object recognition, scene understanding, and graphics. His work also extends to mobile robotics and 3D scene reconstruction. He has made significant contributions in areas such as visual recognition, 3D modeling, and the application of computer vision in construction monitoring. His recent publications (2023-2025) demonstrate a strong focus on advancing multimodal understanding, particularly in region-based representations, 3D vision, and neural radiance fields. There is a clear trend towards integrating language and vision, improving efficiency in neural networks, and applying computer vision to real-world problems such as construction progress monitoring. His scientific awards and honors are extensive and include: IEEE Fellow (2022) University Scholar (2022) Koendrink Prize (2022) Dean's Award for Excellence in Research, Associate Professor (2021) Campus Distinguished Promotion Award (2015) Best Paper Award: IEEE Winter Conference on Applications in Computer Vision (WACV) (2015) CW Gear Junior Faculty Award (2014) IEEE PAMI Young Researcher Award (2014) Dean's Award for Excellence in Research, Assistant Professor (2014) Sloan Research Fellowship (2013) Intel Early Career Faculty Honor Program Award (2012) NSF CAREER Award (2011) ACM Doctoral Dissertation Award, Honorable Mention (2008) Carnegie Mellon University SCS Distinguished Dissertation Award (2008) Best Paper Award: IEEE Computer Vision and Pattern Recognition (CVPR) (2006) Prof. Hoiem has secured significant research funding, including an NSF CAREER award and an Intel Early Career Faculty award. He is also actively involved in technology transfer, having co-founded Reconstruct where he serves as Chief Science Officer. His teaching excellence is reflected in multiple "List of Teachers Ranked as Excellent" awards spanning from 2010 to 2021. Prof. Hoiem leads a research group at UIUC focused on computer vision and 3D scene understanding. Additionally, he co-founded and serves as Chief Science Officer at Reconstruct, which develops AI-based solutions for construction monitoring.
Haibin Ling is the SUNY Empire Innovation Professor in the Department of Computer Science at Stony Brook University, part of the College of Engineering and Applied Sciences. His research focuses on computer vision, medical image analysis, augmented reality, and AI applications in science. He holds a Ph.D. from the University of Maryland (2006) and prior degrees from Peking University. Previously, he worked at Temple University (2008–2019) and held roles at Siemens Corporate Research, UCLA, and Microsoft Research Asia. Professor Ling's work spans biomedical imaging, AI for science, and human-computer interaction. He leads the CV Lab and collaborates with the AI Institute at Stony Brook. Awards include the NSF CAREER Award (2014), Best Student Paper (ACM UIST 2003), and IEEE Fellow (2020). He serves on editorial boards for IEEE Trans. PAMI, Pattern Recognition, and CVIU, and chairs major conferences like CVPR. His research group includes over 50 students and alumni, with active projects in tracking benchmarks (LaSOT), Leafsnap, and medical imaging tools. Notable publications address OCTA flow estimation, backdoor attacks on vision models, and topology-guided medical learning. Collaborations involve institutions like Temple University and Stony Brook's Department of Applied Mathematics and Statistics.
California Institute of Technology (Caltech)United States
Georgia Gkioxari is an Assistant Professor in the Division of Computing and Mathematical Sciences at Caltech , with a part-time affiliation at Meta AI . Her work focuses on extending visual perception models through advanced 2D and 3D representation learning, spatial reasoning, and generative models. Education: Not explicitly mentioned in the text Research interests span 3D perception , spatial reasoning , and vision-language integration , with projects like Visual Agentic AI for Spatial Reasoning and Token-by-Token Multimodal Alignment . Her publications emphasize 3D object detection , reconstruction , and generative modeling techniques including diffusion models and transformers . Scientific recognition includes the Meta LLM Evaluation Research Grant , Okawa Research Grant , Google Faculty Scholar Award 2024 , and Amazon Research Award . She teaches courses like Large Language & Vision Models (EE/CS 148) and Learning & 3D (CS 101) at Caltech. Labs & Teams: Leads Glab with members including Ilona Demler, Ziqi Ma, and Damiano Marsili
Xiaoxiao Long is a Tenure-Track Associate Professor at the School of Intelligence Science and Technology, Nanjing University. He joined NJU as an associate professor in February 2024. Previously, he earned his Ph.D. from the University of Hong Kong (HKU) under the supervision of Prof. Wenping Wang (IEEE & ACM Fellow) and Prof. Taku Komura. His educational background includes: Ph.D. in Computer Science from University of Hong Kong Bachelor's degree in Control Science & Engineering from Zhejiang University Dr. Long's research focuses on computer graphics and 3D computer vision, with particular emphasis on 3D Vision, Physical AI, and World Models. His long-term goal is to develop General-Purpose AI with spatial capabilities. His work bridges theoretical understanding of 3D spaces with practical implementations of spatial AI systems, with applications spanning robotics, virtual reality, and augmented environments. He employs innovative neural network approaches and geometric constraints to advance 3D scene understanding and reconstruction. His publication record shows strong momentum with multiple papers accepted to top-tier conferences including CVPR (5 papers in 2025 alone), ICML, ICLR, ECCV, and TPAMI. His research demonstrates a clear progression from foundational geometric estimation techniques (ASN++) toward more comprehensive spatial AI systems. His scientific recognition includes: Excellent Young Scholars Fund (Overseas) from NSFC Dr. Long has successfully mentored numerous students who have published at major venues and gone on to pursue advanced degrees at prestigious institutions including USTC, Beihang University, HKU, UCAS, Virginia Tech, and HKUST. He is currently recruiting Ph.D. and master's students for Fall 2026, seeking candidates interested in pushing the boundaries of 3D computer vision and spatial AI. His laboratory focuses on developing advanced techniques for 3D scene understanding, neural rendering, and physical AI. Current projects span Gaussian-based representations, neural radiance fields, and geometric estimation, with applications in robotics, virtual environments, and spatial reasoning systems.
Clark Olson is a Professor in the Division of Computing & Software Systems at the University of Washington Bothell, part of the School of Science, Technology, Engineering & Mathematics. He earned his Ph.D. in Computer Science from UC Berkeley (1994), M.S. in Electrical Engineering (1990), and B.S. in Computer Engineering (1989) from the University of Washington, Seattle. Education: Ph.D. in Computer Science (2017) from University of California, Berkeley M.S. in Electrical Engineering (1990) from University of Washington, Seattle B.S. in Computer Engineering (1989) from University of Washington, Seattle His research focuses on computer vision, robot navigation, and clustering algorithms. He has developed techniques for Mars rover terrain mapping, subspace clustering, and geometric feature matching. His work bridges theory and application in autonomous systems and image analysis. Analysis of his publications reveals expertise in computer vision (8 papers), clustering algorithms (4 papers), and robotics (5 papers). Key subtopics include Mars exploration (3 papers), Hough transforms (3 papers), and probabilistic methods (3 papers). Professor Olson teaches courses ranging from introductory programming (CSS 161-162) to advanced topics in computer vision (CSS 487-587) and algorithm design (CSS 549). He also advises on the CSSE Capstone (CSS 497) projects requiring rigorous prerequisites and structured evaluation criteria.
Prof. Christian Heipke is a distinguished academic serving as Dean of the Faculty of Civil Engineering and Geodetic Science at Leibniz University Hannover, Germany. He also holds the position of Executive Director at the Institute of Photogrammetry and GeoInformation (IPI), one of the leading research institutions in geospatial sciences within the faculty. His leadership extends across multiple committees including the Curriculum and Teaching Committee, Admissions and Examination Boards for Geodetic Science and Geoinformatics, and Navigation and Environmental Robotics. As a Professor at IPI, he maintains active research while overseeing significant academic and administrative responsibilities at the university. Professor Heipke's research spans multiple domains within geospatial sciences, with particular emphasis on: Advanced photogrammetric techniques and algorithms Remote sensing applications for environmental monitoring Computer vision approaches for geospatial data analysis Urban development monitoring using satellite imagery Machine learning applications in geoinformatics Disaster prediction and management systems His recent scholarly output reveals a strong focus on integrating cutting-edge computer vision and deep learning techniques with traditional photogrammetric methods. Analysis of his 15 most recent publications shows a clear trajectory toward more sophisticated AI-driven approaches for processing geospatial data, with particular attention to time-series analysis, uncertainty quantification, and multi-view systems. His work bridges theoretical advancements with practical applications in flood forecasting, deforestation monitoring, urban planning, and construction materials analysis. The geographic scope of his research has expanded significantly, with recent projects focusing on international case studies in the Philippines and tropical regions. Professor Heipke leads the Institute of Photogrammetry and GeoInformation, a major research hub that has celebrated 75 years of contributions to the field. His leadership extends to the Graduiertenkolleg 2159: "Integrity and Collaboration in Dynamic Sensor Networks," where he serves as a professor overseeing doctoral research. The institute maintains state-of-the-art facilities for processing satellite imagery, aerial photography, and developing novel algorithms for geospatial data analysis. Under his direction, the institute has strengthened its international collaborations and interdisciplinary research approaches, particularly in addressing Sustainable Development Goals through geospatial technologies.
Dr. Robert J. Teather is an Associate Professor and the Director of the School of Information Technology at Carleton University in Ottawa, Canada. He previously served as an interim Director of the School of Information Technology during the 2022-23 academic year. His academic journey includes a PhD in Computer Science from York University (2013) and a postdoctoral fellowship at McMaster University (2015). Dr. Teather's educational background includes: PhD in Computer Science from York University (2013) Master's Thesis: "Comparing 2D and 3D Direct Manipulation Interfaces" from York University (2008), which was awarded the Joseph Liu Thesis Award Dr. Teather's research broadly falls under the field of human-computer interaction, with specialization in 3D user interfaces, virtual reality, and user interfaces for computer games. His work establishes methods for direct comparison of 2D and 3D interfaces for conceptually equivalent tasks, such as selection and manipulation interfaces. He investigates factors influencing human performance in VR, including stereo 3D graphics, haptic feedback, and head-tracking. His research also evaluates novel user interfaces like tilt control or touchscreens, and examines human performance with game input devices in complex tasks involving navigation, selection, and manipulation of objects in game environments. His research has been published extensively in top venues including IEEE VR, ACM SUI, and Graphics Interface. Among his notable scientific achievements are: NSERC Postgraduate Scholarship during his PhD studies Ontario Graduate Scholarship during his PhD studies Best Paper Honourable Mention at the ACM Symposium on Applied Perception 2020 Best Demo Award for SUI 2017 Joseph Liu Thesis Award (2008) Dr. Teather actively supervises graduate students at Carleton University, currently overseeing multiple PhD and Master's students in the areas of human-computer interaction and interactive digital media. His research is supported by NSERC and the Canada Foundation for Innovation, providing funding for his students and laboratory equipment. His students have produced research spanning VR as a persuasive tool to improve vaccine confidence, selection performance using smartphones in VR, and text entry methods in virtual reality environments. Dr. Teather leads a well-equipped CFI-supported lab focused on virtual and augmented reality research. His team works collaboratively on projects related to interactive virtual reality systems, computer game user interfaces, and input devices for 3D interaction. The lab environment fosters interdisciplinary research with opportunities for students to work on cutting-edge VR/AR technologies and contribute to the growing field of spatial computing.
Oisin Mac Aodha is a Reader (Associate Professor) in Machine Learning at the School of Informatics, University of Edinburgh. He is also an ELLIS Scholar and founder of the Turing interest group on biodiversity monitoring and forecasting, having previously served as a Turing Fellow from 2021-2025. Mac Aodha completed his undergraduate degree in electronic engineering from the University of Galway in Ireland, followed by his MSc and PhD at University College London (UCL). His academic journey includes postdoctoral positions at UCL (2013-2016) working with Prof. Gabriel Brostow and Prof. Kate Jones, and at Caltech (2016-2019) in Prof. Pietro Perona's Computational Vision Lab as part of the Visipedia team. His research centers on computer vision and machine learning with emphasis on 3D understanding, human-in-the-loop methods, and AI for conservation and biodiversity monitoring. He has made significant contributions to monocular depth estimation (including the influential Monodepth2 paper), fine-grained visual categorization, and biodiversity monitoring systems. His work bridges theoretical machine learning with practical ecological applications, developing tools for species identification, range estimation, and conservation efforts. Recent publications reveal a strong trend toward ecological applications while maintaining fundamental contributions to 3D vision and representation learning. His major scientific achievements include: Turing Fellow (2021-2025) ELLIS Scholar Founder of the Turing interest group on biodiversity monitoring and forecasting Co-organizer of the Fine-Grained Visual Categorization (FGVC) workshop series at major vision conferences Mac Aodha advises multiple PhD students and postdocs working on computer vision for biodiversity monitoring, 3D understanding, and human-in-the-loop learning. His team has developed practical tools like Whombat (an open-source annotation tool for bioacoustics) and contributed to field-deployed biodiversity monitoring systems. He has served as Area Chair for top conferences including NeurIPS, CVPR, ICCV, and ICML, demonstrating his standing in the computer vision community. His research group collaborates extensively with ecologists at University College London, particularly with Prof. Kate Jones' team, bridging machine learning expertise with ecological domain knowledge. The Vision at Edinburgh group he contributes to focuses on developing practical AI tools that address real-world conservation challenges while advancing fundamental computer vision research.
Swiss Federal Institute of Technology in LausanneSwitzerland
Sabine Süsstrunk is a Full Professor and Director of the Images and Visual Representation Laboratory (IVRL) at EPFL's School of Computer and Communication Sciences. She holds a BS in Scientific Photography from ETH Zürich, MS from Rochester Institute of Technology, and PhD from University of East Anglia. Her career includes positions at Hewlett-Packard Labs and Corbis Corporation. Her research explores computational imaging, computational photography, color processing, computer vision, and image quality. Key interests include near-infrared applications, multispectral imaging, and computational aesthetics. Her work bridges hardware and software solutions for imaging challenges. Publications demonstrate consistent focus on advancing generative models (diffusion models, neural cellular automata), 3D reconstruction (NeRF variants), and media integrity (DeepFake detection). Recent trends show increased emphasis on 3D vision, robustness in generative AI, and video analysis. Awards & Honors: IS&T/SPIE Electronic Imaging Scientist of the Year (2013) Raymond C. Bowman Teaching Award (2018) EPFL AGEPoly IC Polysphere Award (2020) 8 Best Paper/Demo Awards Fellowships: IEEE, IS&T, ELLIS, AIIA She leads the IVRL lab and advises PhD candidates while serving as President of the Swiss Science Council. Research is supported through competitive grants and industry collaborations.
Swiss Federal Institute of Technology in LausanneSwitzerland
Gabriele Facciolo is a Professor at the Centre Borelli, ENS Paris-Saclay, France. He is a Senior Member of the Institut Universitaire de France (IUF) and holds an Innovation Chair (2025). His research focuses on image and video processing, remote sensing, and super-resolution techniques. Current affiliations: Centre Borelli (ENS Paris-Saclay), Institut Universitaire de France His research explores advanced algorithms for satellite stereo pipelines, real-time deblurring, denoising, and explainable AI systems for legal evidence enhancement. He coordinates projects like ANR SURECAVI (Super-resolution for visible camera systems) and ANR IMPROVED (video enhancement for judicial use), with recent work on Gaussian Splatting for Earth Observation and multi-date satellite super-resolution. Notable scientific achievements include the IGARSS 2025 Top 10 Student Paper Award and leadership in projects funded by ANR (€890k) and Prime Minister's entities (SGDSN/ANSSI). His work bridges computational imaging, defense applications, and digital forensics. Project leadership: SURECAVI, IMPROVED, BOFOR Key technologies: GPU acceleration, real-time processing, optical flow estimation, RPC refinement Gabriele actively contributes to open-source tools like S2P (Satellite Stereo Pipeline), MGM (MultiGlobal Matching), and OMNIflip. He teaches in the Master MVA program and collaborates across institutions (ENPC, UPF).
Anup Basu is a Professor in the Department of Computing Science at the University of Alberta. His research focuses on computer graphics, computer vision, and multimedia communications. He holds an B.S. in Math & Statistics from the Indian Statistical Institute (1980), an M.E. in Computer Science from the Indian Statistical Institute (1983), and a Ph.D. in Computing Science from the University of Maryland (1990). His work emphasizes Quality of Service (QoS) in multimedia delivery for e-commerce and telelearning, adaptive bandwidth monitoring, and 3D visualization tools. He pioneered foveated image compression and stereo visualization techniques, contributing to MPEG-4 coding standards. He leads major initiatives like the ASRA/TelePhotogenics/IBM 3D Medical Imaging project ($2M+ funding) and developed patented SHR Stereo/3D scanning technologies. Awards include the American Neurological Association Fellowship. He has held leadership roles as General Chair for IEEE International Conferences on SMC (2017), Multimedia & Expo (2013), and SMC (2014). His research integrates interdisciplinary collaborations across universities and industry partners, leveraging advanced equipment like the CAVE system for immersive visualization.
Iro Laina is a Departmental Lecturer in Computer Vision at the University of Oxford's Visual Geometry Group. She holds a PhD (Dr. rer. nat.) from the Technical University of Munich (TUM), where her dissertation earned the ECVA PhD Award. Her research focuses on unsupervised and language-supervised learning for 3D scene understanding, image/video perception systems, and geometric reconstruction. Education: PhD in Computer Science (TUM), MSc in Biomedical Computing (TUM), Diploma in Electrical & Computer Engineering (NTUA). Research Interests: 3D Reconstruction and Generation Unsupervised Learning Multi-View and Video Analysis Generative Diffusion Models Geometry-Aware Networks Her recent work emphasizes scalable 3D scene synthesis, training-free methods, and cross-modal fusion with LLMs. Over 15+ publications since 2021 reflect her leadership in geometric deep learning. Awards: ECVA PhD Award (2020), Recognized in multiple international conferences. Advising: Mentors DPhil students in creative AI applications (e.g., gameplay design). Active in Oxford's Robotics and Biomedical Engineering networks. Labs/Tech: Core member of the Visual Geometry Group, collaborating on projects like IMAD2025 with the ZERO Institute.