Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Noah Snavely is a Professor of Computer Science at Cornell Tech and the Cornell Ann S. Bowers College of Computing and Information Science. His research spans computer vision and graphics with a focus on 3D scene understanding from image collections. He also works at Google DeepMind in NYC and leads the Cornell Graphics and Vision Group. His research interests include computer vision and computer graphics , particularly in recovering 3D structure from large photo collections, neural radiance fields, image-based rendering, and scene understanding. His work enables applications in mapping technologies, immersive VR experiences, and synthesizing 3D worlds from text prompts with implications for game design and filmmaking. Recent publications show a strong trend toward generative 3D modeling and neural rendering , with significant contributions in neural radiance fields, view synthesis, and dynamic scene reconstruction. His research group explores unstructured photo collections to develop new technology for 3D world modeling and image analysis. PECASE Microsoft New Faculty Fellowship Alfred P. Sloan Fellowship SIGGRAPH Significant New Researcher Award ACM Fellow (2023) IEEE Fellow NSF CAREER Award Helmholtz Prize Snavely has mentored numerous PhD students and postdocs including Ruojin Cai, Qianqian Wang, and Zhengqi Li. His teaching includes CS5670: Computer Vision across multiple semesters. He has received the Cornell College of Engineering Mr. and Mrs. Richard F. Tucker Teaching Excellence Award (2012). At Cornell Tech, his research group develops technology for modeling the world in 3D from online photo collections and analyzing images for computer graphics applications. His work has been featured in major news outlets including the New York Times, New Scientist, and MIT Technology Review.
Xingxing Zuo is an Assistant Professor (tenure-track) in the Robotics Department at MBZUAI. He holds a PhD from Zhejiang University (2021) and a Bachelor’s from UESTC (2016). Previously, he was a Postdoctoral Scholar at Caltech (2024–2025), a Postdoc at ETH Zurich (2019–2021), and held visiting roles at TU Munich, University of Delaware, and University of Technology Sydney. His research focuses on robotics, 3D computer vision, and embodied AI, with emphasis on robot-human collaboration, state estimation, and sensor fusion. Educations: PhD in Robotics, Zhejiang University (2021, with honors) Bachelor’s in Computer Science, University of Electronic Science and Technology of China (2016, with honors) Research Highlights: Develops novel methods for LiDAR-camera-inertial fusion, neural radiance fields, and radar-cameras systems Pioneered techniques like Flying Co-Stereo (long-range aerial mapping) and FMGS (vision-language embedded 3D splatting) Focuses on real-time SLAM, robust depth estimation, and photorealistic scene reconstruction Awards & Recognition: Best Paper Finalist at ICRA 2021 (CodeVIO) Oral Presentation at ICCV 2021 (MBA-VO) Recipient of Google Visiting Faculty Researcher (2023) Grants & Labs: Organized Thermal Infrared in Robotics workshop at ICRA 2025 Leads research on embodied AI and multi-sensor SLAM systems Develops open-source tools like LIC-Fusion and Coco-LIC frameworks
Derek W Hoiem is a Professor in the Siebel School for Computing and Data Science at the University of Illinois Urbana-Champaign, where he has been a faculty member since 2009. His research focuses on computer vision and related areas, and he is also the co-founder and Chief Science Officer of Reconstruct, an AI-based construction technology company. His educational background includes: PhD in Robotics, Carnegie Mellon University (2007) Beckman Postdoctoral Fellowship (2008) Prof. Hoiem's research spans computer vision, with a focus on object recognition, scene understanding, and graphics. His work also extends to mobile robotics and 3D scene reconstruction. He has made significant contributions in areas such as visual recognition, 3D modeling, and the application of computer vision in construction monitoring. His recent publications (2023-2025) demonstrate a strong focus on advancing multimodal understanding, particularly in region-based representations, 3D vision, and neural radiance fields. There is a clear trend towards integrating language and vision, improving efficiency in neural networks, and applying computer vision to real-world problems such as construction progress monitoring. His scientific awards and honors are extensive and include: IEEE Fellow (2022) University Scholar (2022) Koendrink Prize (2022) Dean's Award for Excellence in Research, Associate Professor (2021) Campus Distinguished Promotion Award (2015) Best Paper Award: IEEE Winter Conference on Applications in Computer Vision (WACV) (2015) CW Gear Junior Faculty Award (2014) IEEE PAMI Young Researcher Award (2014) Dean's Award for Excellence in Research, Assistant Professor (2014) Sloan Research Fellowship (2013) Intel Early Career Faculty Honor Program Award (2012) NSF CAREER Award (2011) ACM Doctoral Dissertation Award, Honorable Mention (2008) Carnegie Mellon University SCS Distinguished Dissertation Award (2008) Best Paper Award: IEEE Computer Vision and Pattern Recognition (CVPR) (2006) Prof. Hoiem has secured significant research funding, including an NSF CAREER award and an Intel Early Career Faculty award. He is also actively involved in technology transfer, having co-founded Reconstruct where he serves as Chief Science Officer. His teaching excellence is reflected in multiple "List of Teachers Ranked as Excellent" awards spanning from 2010 to 2021. Prof. Hoiem leads a research group at UIUC focused on computer vision and 3D scene understanding. Additionally, he co-founded and serves as Chief Science Officer at Reconstruct, which develops AI-based solutions for construction monitoring.
Haibin Ling is the SUNY Empire Innovation Professor in the Department of Computer Science at Stony Brook University, part of the College of Engineering and Applied Sciences. His research focuses on computer vision, medical image analysis, augmented reality, and AI applications in science. He holds a Ph.D. from the University of Maryland (2006) and prior degrees from Peking University. Previously, he worked at Temple University (2008–2019) and held roles at Siemens Corporate Research, UCLA, and Microsoft Research Asia. Professor Ling's work spans biomedical imaging, AI for science, and human-computer interaction. He leads the CV Lab and collaborates with the AI Institute at Stony Brook. Awards include the NSF CAREER Award (2014), Best Student Paper (ACM UIST 2003), and IEEE Fellow (2020). He serves on editorial boards for IEEE Trans. PAMI, Pattern Recognition, and CVIU, and chairs major conferences like CVPR. His research group includes over 50 students and alumni, with active projects in tracking benchmarks (LaSOT), Leafsnap, and medical imaging tools. Notable publications address OCTA flow estimation, backdoor attacks on vision models, and topology-guided medical learning. Collaborations involve institutions like Temple University and Stony Brook's Department of Applied Mathematics and Statistics.
Georgia Gkioxari is an Assistant Professor in the Division of Computing and Mathematical Sciences at Caltech , with a part-time affiliation at Meta AI . Her work focuses on extending visual perception models through advanced 2D and 3D representation learning, spatial reasoning, and generative models. Education: Not explicitly mentioned in the text Research interests span 3D perception , spatial reasoning , and vision-language integration , with projects like Visual Agentic AI for Spatial Reasoning and Token-by-Token Multimodal Alignment . Her publications emphasize 3D object detection , reconstruction , and generative modeling techniques including diffusion models and transformers . Scientific recognition includes the Meta LLM Evaluation Research Grant , Okawa Research Grant , Google Faculty Scholar Award 2024 , and Amazon Research Award . She teaches courses like Large Language & Vision Models (EE/CS 148) and Learning & 3D (CS 101) at Caltech. Labs & Teams: Leads Glab with members including Ilona Demler, Ziqi Ma, and Damiano Marsili
Xiaoxiao Long is a Tenure-Track Associate Professor at the School of Intelligence Science and Technology, Nanjing University. He joined NJU as an associate professor in February 2024. Previously, he earned his Ph.D. from the University of Hong Kong (HKU) under the supervision of Prof. Wenping Wang (IEEE & ACM Fellow) and Prof. Taku Komura. His educational background includes: Ph.D. in Computer Science from University of Hong Kong Bachelor's degree in Control Science & Engineering from Zhejiang University Dr. Long's research focuses on computer graphics and 3D computer vision, with particular emphasis on 3D Vision, Physical AI, and World Models. His long-term goal is to develop General-Purpose AI with spatial capabilities. His work bridges theoretical understanding of 3D spaces with practical implementations of spatial AI systems, with applications spanning robotics, virtual reality, and augmented environments. He employs innovative neural network approaches and geometric constraints to advance 3D scene understanding and reconstruction. His publication record shows strong momentum with multiple papers accepted to top-tier conferences including CVPR (5 papers in 2025 alone), ICML, ICLR, ECCV, and TPAMI. His research demonstrates a clear progression from foundational geometric estimation techniques (ASN++) toward more comprehensive spatial AI systems. His scientific recognition includes: Excellent Young Scholars Fund (Overseas) from NSFC Dr. Long has successfully mentored numerous students who have published at major venues and gone on to pursue advanced degrees at prestigious institutions including USTC, Beihang University, HKU, UCAS, Virginia Tech, and HKUST. He is currently recruiting Ph.D. and master's students for Fall 2026, seeking candidates interested in pushing the boundaries of 3D computer vision and spatial AI. His laboratory focuses on developing advanced techniques for 3D scene understanding, neural rendering, and physical AI. Current projects span Gaussian-based representations, neural radiance fields, and geometric estimation, with applications in robotics, virtual environments, and spatial reasoning systems.
Clark Olson is a Professor in the Division of Computing & Software Systems at the University of Washington Bothell, part of the School of Science, Technology, Engineering & Mathematics. He earned his Ph.D. in Computer Science from UC Berkeley (1994), M.S. in Electrical Engineering (1990), and B.S. in Computer Engineering (1989) from the University of Washington, Seattle. Education: Ph.D. in Computer Science (2017) from University of California, Berkeley M.S. in Electrical Engineering (1990) from University of Washington, Seattle B.S. in Computer Engineering (1989) from University of Washington, Seattle His research focuses on computer vision, robot navigation, and clustering algorithms. He has developed techniques for Mars rover terrain mapping, subspace clustering, and geometric feature matching. His work bridges theory and application in autonomous systems and image analysis. Analysis of his publications reveals expertise in computer vision (8 papers), clustering algorithms (4 papers), and robotics (5 papers). Key subtopics include Mars exploration (3 papers), Hough transforms (3 papers), and probabilistic methods (3 papers). Professor Olson teaches courses ranging from introductory programming (CSS 161-162) to advanced topics in computer vision (CSS 487-587) and algorithm design (CSS 549). He also advises on the CSSE Capstone (CSS 497) projects requiring rigorous prerequisites and structured evaluation criteria.
Prof. Christian Heipke is a distinguished academic serving as Dean of the Faculty of Civil Engineering and Geodetic Science at Leibniz University Hannover, Germany. He also holds the position of Executive Director at the Institute of Photogrammetry and GeoInformation (IPI), one of the leading research institutions in geospatial sciences within the faculty. His leadership extends across multiple committees including the Curriculum and Teaching Committee, Admissions and Examination Boards for Geodetic Science and Geoinformatics, and Navigation and Environmental Robotics. As a Professor at IPI, he maintains active research while overseeing significant academic and administrative responsibilities at the university. Professor Heipke's research spans multiple domains within geospatial sciences, with particular emphasis on: Advanced photogrammetric techniques and algorithms Remote sensing applications for environmental monitoring Computer vision approaches for geospatial data analysis Urban development monitoring using satellite imagery Machine learning applications in geoinformatics Disaster prediction and management systems His recent scholarly output reveals a strong focus on integrating cutting-edge computer vision and deep learning techniques with traditional photogrammetric methods. Analysis of his 15 most recent publications shows a clear trajectory toward more sophisticated AI-driven approaches for processing geospatial data, with particular attention to time-series analysis, uncertainty quantification, and multi-view systems. His work bridges theoretical advancements with practical applications in flood forecasting, deforestation monitoring, urban planning, and construction materials analysis. The geographic scope of his research has expanded significantly, with recent projects focusing on international case studies in the Philippines and tropical regions. Professor Heipke leads the Institute of Photogrammetry and GeoInformation, a major research hub that has celebrated 75 years of contributions to the field. His leadership extends to the Graduiertenkolleg 2159: "Integrity and Collaboration in Dynamic Sensor Networks," where he serves as a professor overseeing doctoral research. The institute maintains state-of-the-art facilities for processing satellite imagery, aerial photography, and developing novel algorithms for geospatial data analysis. Under his direction, the institute has strengthened its international collaborations and interdisciplinary research approaches, particularly in addressing Sustainable Development Goals through geospatial technologies.
Dr. Robert J. Teather is an Associate Professor and the Director of the School of Information Technology at Carleton University in Ottawa, Canada. He previously served as an interim Director of the School of Information Technology during the 2022-23 academic year. His academic journey includes a PhD in Computer Science from York University (2013) and a postdoctoral fellowship at McMaster University (2015). Dr. Teather's educational background includes: PhD in Computer Science from York University (2013) Master's Thesis: "Comparing 2D and 3D Direct Manipulation Interfaces" from York University (2008), which was awarded the Joseph Liu Thesis Award Dr. Teather's research broadly falls under the field of human-computer interaction, with specialization in 3D user interfaces, virtual reality, and user interfaces for computer games. His work establishes methods for direct comparison of 2D and 3D interfaces for conceptually equivalent tasks, such as selection and manipulation interfaces. He investigates factors influencing human performance in VR, including stereo 3D graphics, haptic feedback, and head-tracking. His research also evaluates novel user interfaces like tilt control or touchscreens, and examines human performance with game input devices in complex tasks involving navigation, selection, and manipulation of objects in game environments. His research has been published extensively in top venues including IEEE VR, ACM SUI, and Graphics Interface. Among his notable scientific achievements are: NSERC Postgraduate Scholarship during his PhD studies Ontario Graduate Scholarship during his PhD studies Best Paper Honourable Mention at the ACM Symposium on Applied Perception 2020 Best Demo Award for SUI 2017 Joseph Liu Thesis Award (2008) Dr. Teather actively supervises graduate students at Carleton University, currently overseeing multiple PhD and Master's students in the areas of human-computer interaction and interactive digital media. His research is supported by NSERC and the Canada Foundation for Innovation, providing funding for his students and laboratory equipment. His students have produced research spanning VR as a persuasive tool to improve vaccine confidence, selection performance using smartphones in VR, and text entry methods in virtual reality environments. Dr. Teather leads a well-equipped CFI-supported lab focused on virtual and augmented reality research. His team works collaboratively on projects related to interactive virtual reality systems, computer game user interfaces, and input devices for 3D interaction. The lab environment fosters interdisciplinary research with opportunities for students to work on cutting-edge VR/AR technologies and contribute to the growing field of spatial computing.
Oisin Mac Aodha is a Reader (Associate Professor) in Machine Learning at the School of Informatics, University of Edinburgh. He is also an ELLIS Scholar and founder of the Turing interest group on biodiversity monitoring and forecasting, having previously served as a Turing Fellow from 2021-2025. Mac Aodha completed his undergraduate degree in electronic engineering from the University of Galway in Ireland, followed by his MSc and PhD at University College London (UCL). His academic journey includes postdoctoral positions at UCL (2013-2016) working with Prof. Gabriel Brostow and Prof. Kate Jones, and at Caltech (2016-2019) in Prof. Pietro Perona's Computational Vision Lab as part of the Visipedia team. His research centers on computer vision and machine learning with emphasis on 3D understanding, human-in-the-loop methods, and AI for conservation and biodiversity monitoring. He has made significant contributions to monocular depth estimation (including the influential Monodepth2 paper), fine-grained visual categorization, and biodiversity monitoring systems. His work bridges theoretical machine learning with practical ecological applications, developing tools for species identification, range estimation, and conservation efforts. Recent publications reveal a strong trend toward ecological applications while maintaining fundamental contributions to 3D vision and representation learning. His major scientific achievements include: Turing Fellow (2021-2025) ELLIS Scholar Founder of the Turing interest group on biodiversity monitoring and forecasting Co-organizer of the Fine-Grained Visual Categorization (FGVC) workshop series at major vision conferences Mac Aodha advises multiple PhD students and postdocs working on computer vision for biodiversity monitoring, 3D understanding, and human-in-the-loop learning. His team has developed practical tools like Whombat (an open-source annotation tool for bioacoustics) and contributed to field-deployed biodiversity monitoring systems. He has served as Area Chair for top conferences including NeurIPS, CVPR, ICCV, and ICML, demonstrating his standing in the computer vision community. His research group collaborates extensively with ecologists at University College London, particularly with Prof. Kate Jones' team, bridging machine learning expertise with ecological domain knowledge. The Vision at Edinburgh group he contributes to focuses on developing practical AI tools that address real-world conservation challenges while advancing fundamental computer vision research.
Sabine Süsstrunk is a Full Professor and Director of the Images and Visual Representation Laboratory (IVRL) at EPFL's School of Computer and Communication Sciences. She holds a BS in Scientific Photography from ETH Zürich, MS from Rochester Institute of Technology, and PhD from University of East Anglia. Her career includes positions at Hewlett-Packard Labs and Corbis Corporation. Her research explores computational imaging, computational photography, color processing, computer vision, and image quality. Key interests include near-infrared applications, multispectral imaging, and computational aesthetics. Her work bridges hardware and software solutions for imaging challenges. Publications demonstrate consistent focus on advancing generative models (diffusion models, neural cellular automata), 3D reconstruction (NeRF variants), and media integrity (DeepFake detection). Recent trends show increased emphasis on 3D vision, robustness in generative AI, and video analysis. Awards & Honors: IS&T/SPIE Electronic Imaging Scientist of the Year (2013) Raymond C. Bowman Teaching Award (2018) EPFL AGEPoly IC Polysphere Award (2020) 8 Best Paper/Demo Awards Fellowships: IEEE, IS&T, ELLIS, AIIA She leads the IVRL lab and advises PhD candidates while serving as President of the Swiss Science Council. Research is supported through competitive grants and industry collaborations.
Gabriele Facciolo is a Professor at the Centre Borelli, ENS Paris-Saclay, France. He is a Senior Member of the Institut Universitaire de France (IUF) and holds an Innovation Chair (2025). His research focuses on image and video processing, remote sensing, and super-resolution techniques. Current affiliations: Centre Borelli (ENS Paris-Saclay), Institut Universitaire de France His research explores advanced algorithms for satellite stereo pipelines, real-time deblurring, denoising, and explainable AI systems for legal evidence enhancement. He coordinates projects like ANR SURECAVI (Super-resolution for visible camera systems) and ANR IMPROVED (video enhancement for judicial use), with recent work on Gaussian Splatting for Earth Observation and multi-date satellite super-resolution. Notable scientific achievements include the IGARSS 2025 Top 10 Student Paper Award and leadership in projects funded by ANR (€890k) and Prime Minister's entities (SGDSN/ANSSI). His work bridges computational imaging, defense applications, and digital forensics. Project leadership: SURECAVI, IMPROVED, BOFOR Key technologies: GPU acceleration, real-time processing, optical flow estimation, RPC refinement Gabriele actively contributes to open-source tools like S2P (Satellite Stereo Pipeline), MGM (MultiGlobal Matching), and OMNIflip. He teaches in the Master MVA program and collaborates across institutions (ENPC, UPF).
Chao Liu is a Research Scientist at CNRS (French National Center for Scientific Research) since 2008, affiliated with the DEXTER team and the Department of Robotics, LIRMM at University of Montpellier, France. He earned his Ph.D. in Electrical & Electronic Engineering from Nanyang Technological University, Singapore (2006). Current research focuses on surgical robotics , haptics , teleoperation , and nonlinear control theory with applications in computer vision. His work addresses challenges in robotic-assisted telesurgery, including: Stable and transparent human-robot interaction through wave variable compensators and passivity filters Physiological motion compensation using spatio-temporal LSTM and dual Kalman filters EMG-based motion recognition for surgical skill assessment 3D soft-tissue reconstruction with stereo-endoscopes and deep learning Dr. Liu leads European and French projects like: TS2RT (CNRS-funded): Safer teleoperation with motion compensation ROBACUS (ANR-funded): Needle positioning with MPC control HaTUMoCo (CNRS-funded): Haptic teleoperation with uncertainty handling ARAKNES (EU-funded): Microrobotic systems for endoluminal surgery Scientific honors include Senior Member of IEEE and Member of Sigma Xi . He supervises Ph.D. and Master's students working on topics such as concentric tube robot optimization, haptic teleoperation, and EMG-based force estimation. Dr. Liu serves on IEEE Technical Committees for Telerobotics and Haptics , and as Technical Editor of IEEE/ASME Transactions on Mechatronics.
Professor David Taubman is a distinguished academic serving as Professor and Deputy Head of School (Research) at the School of Electrical Engineering and Telecommunications (EE&T) at the University of New South Wales (UNSW) in Sydney, Australia. He is also co-director of Kakadu Software Pty. Ltd. and its affiliates Kakadu R&D and Kakadu GPU. With a career spanning over three decades, Professor Taubman has made significant contributions to the field of image and video compression, most notably as the author of the EBCOT coding algorithm adopted in the JPEG2000 international standard. Professor Taubman earned his B.Sc. in Mathematics and Computer Science (1986) and B.E. (Medal) in Electrical Engineering (1988) from the University of Sydney, followed by an M.Sc. (1992) and Ph.D. (1994) in Electrical Engineering from the University of California at Berkeley. His professional journey includes engineering work at the Electricity Commission of N.S.W. (1988-1990), research positions at Hewlett-Packard Laboratories in Palo Alto (1994-1998), and an academic career at UNSW where he progressed from Senior Lecturer (1998-2003) to Associate Professor (2004-2009) and finally to Professor (2009-present). He has held various leadership roles including Head of the EE&T Telecommunications Research Group (2003-2014), Head of the EE&T Signal Processing Research Group (2014-present), Director of Research for the School of EE&T (2011-2016), and Deputy Head of School (Research) since 2017. Professor Taubman's research interests center on image and video compression, with particular expertise in JPEG2000 standards and implementations. His work spans signal processing, wavelet transforms, scalable video coding, motion modeling, and multimedia systems. He has pioneered numerous compression algorithms and frameworks, including the EBCOT coding algorithm that became central to the JPEG2000 standard. His recent research focuses on efficient motion modeling with cuboidal partitioning, learned lifting-based transform structures, and high-throughput implementations of JPEG2000 for video applications. His work bridges theoretical foundations with practical implementations, as evidenced by the commercially successful Kakadu Software tools that have garnered around 500 commercial licensees. Analysis of Professor Taubman's recent publications reveals a consistent focus on advancing compression technologies with particular emphasis on scalability, efficiency, and adaptability. His work spans traditional image compression (JPEG2000 extensions), video coding (cuboid-based partitioning for UHD/360-degree video), and emerging applications (nanopore sequencing data compression). A notable trend is the integration of machine learning techniques with traditional compression frameworks, as seen in his work on learned lifting-based transform structures. His research maintains strong connections to real-world applications across diverse domains including medical imaging, astronomical data processing, and genomic sequencing. IEEE Fellow Engineers Australia Fellow (by invitation) Professor Taubman has served as Associate Editor for the IEEE Transactions on Image Processing for two four-year appointments (2003-2005 and 2010-2013). He has been actively involved in numerous research grants focused on image and video compression technologies, particularly those related to the JPEG2000 standard and its extensions. His work has received significant industry support, reflected in his consultancy with various U.S., Japanese, and Australian corporations. He has also contributed to international standards development as a member of Standards Australia Technical Committee MS-065 (mirroring ISO TC42 on Digital Photography) and as a constitutional member of Standards Australia Technical Committee IT-029 (Coded Representation of Picture, Audio and Multimedia/Hypermedia Information). Professor Taubman co-directs Kakadu Software Pty. Ltd. and its research affiliates Kakadu R&D and Kakadu GPU, which have developed the commercially successful Kakadu Software tools for JPEG2000. His research group at UNSW focuses on advanced image and video compression techniques, with particular expertise in wavelet-based methods, scalable coding, and motion modeling. The group maintains strong industry connections and has contributed significantly to the development and standardization of image compression technologies worldwide.