Xingang Pan is an Assistant Professor in the College of Computing and Data Science at Nanyang Technological University (NTU), leading the MMLab@NTU. His research focuses on generative AI and visual content creation, particularly in generative models, 3D vision, computer graphics, and computer vision. Prior to NTU, he was a postdoc at the Max Planck Institute for Informatics and earned his Ph.D. from the Chinese University of Hong Kong (2021) and B.Sc. from Tsinghua University (2016). His work emphasizes generative intelligence, exploring long-term world simulation, diffusion models, and multi-scale 3D generation. Notable contributions include WORLDMEM (2025), Alias-free Latent Diffusion (2025), and SAR3D (2025). His research has been published in top venues like CVPR, ICCV, and SIGGRAPH. Xingang Pan oversees the MMLab@NTU, which actively recruits students globally without nationality constraints. The lab’s projects include GAN2Shape (unsupervised 3D reconstruction from 2D GANs) and LN3Diff (scalable 3D generation).
Alex Wong is an Assistant Professor of Computer Science at Yale University, specializing in computer vision, robotics, and medical imaging. His research focuses on sensor fusion, unsupervised learning, 3D vision, robust perception under adverse conditions, and medical image analysis. He holds degrees from the University of California, Los Angeles (UCLA), including a B.S., M.S., and Ph.D. in Computer Science. Wong’s work bridges theoretical advances with practical applications, particularly in depth estimation, autonomous systems, and medical diagnostics. He has received prestigious awards such as the NeurIPS Outstanding Student Paper Award (2011) and the ICRA Best Paper Award in Robot Vision (2019). His research often addresses challenges in unstructured environments, emphasizing robustness and adaptability. Recent projects include developing novel frameworks for unsupervised depth completion, adversarial robustness in vision systems, and multimodal fusion techniques. His contributions span conferences like CVPR, ICCV, and ICRA, with a strong focus on advancing AI for real-world applications in healthcare and robotics. Education: B.S., Computer Science, UCLA M.S., Computer Science, UCLA Ph.D., Computer Science, UCLA Awards: NeurIPS Outstanding Student Paper Award (2011) ICRA Best Paper Award in Robot Vision (2019) His lab at Yale Engineering focuses on AI-driven solutions for perception challenges, collaborating across disciplines to advance medical imaging and autonomous systems. Current efforts explore generative models, continual learning, and vision-language integration for robust scene understanding.
Ajmal Mian is a Professor of Computer Science at the University of Western Australia (UWA), affiliated with the School of Physics, Maths and Computing. He holds an Australian Research Council Future Fellowship (2022) and leads research in Artificial Intelligence, Computer Vision, and Machine Learning. His work focuses on 3D computer vision, adversarial AI defense, and explainable AI. His research interests include 3D point cloud analysis, face recognition, human action recognition, and remote sensing. He has published over 300 papers and secured major grants from ARC, NHMRC, and DARPA, totaling millions in funding. He has supervised 29 PhD students and mentored 12 postdoctoral researchers. Key projects include 3D diffusion models for scene generation, robust 3D vision systems, and defense against AI deception attacks. He serves as a fellow of IAPR, an ACM Distinguished Speaker, and has editorial roles at IEEE Transactions on Neural Networks and Pattern Recognition. Research Awards: HBF Mid-Career Scientist of the Year, West Australian Early Career Scientist of the Year, IAPR Best Scientific Paper Award. Grants: ARC Discovery Projects, National Intelligence & Security Discovery grants, DARPA grants for AI security. His teaching spans computer vision, machine learning, and programming courses. Collaborations include defense, medical, and agricultural applications.
Daniel M. Roy is a Full Professor at the University of Toronto, with cross-appointments in the Department of Computer Science, Department of Statistical Sciences, and Department of Electrical and Computer Engineering. He serves as Research Director at the Vector Institute and holds the CIFAR Canada AI Chair. Research Focus: Foundational principles of prediction, inference, and decision-making under uncertainty across machine learning, statistics, mathematical logic, applied probability, and computer science. Scientific Contributions: Key work in learning theory, statistical network analysis, probabilistic programming, and information-theoretic frameworks for generalization. Awards: ICML 2024 Best Paper Award for "Information Complexity of Stochastic Convex Optimization" and promotion to Full Professor in 2024. Student Advising: Actively mentors Ph.D. candidates and postdoctoral researchers with strong quantitative backgrounds, particularly at the intersection of machine learning, statistics, and computer science. Email: daniel.roy@utoronto.ca
Deva Ramanan is a Professor at the Robotics Institute of Carnegie Melllon University, where he leads research in computer vision and machine learning. His work focuses on modeling human visual perception, leveraging large-scale visual data, and developing systems for 3D understanding, neural rendering, and autonomous systems. He advises a large group of PhD students and has mentored numerous postdoctoral researchers now in leading roles across industry and academia. His research interests include computer vision, machine learning, human perception modeling, 3D scene understanding, neural rendering, autonomous driving, video understanding, and multimodal foundation models. These areas reflect his focus on both foundational models and their application to real-world problems in robotics and AI. The recent publications highlight a strong trend toward multimodal and 3D-aware models, with increasing use of diffusion models, neural fields, and large vision-language systems. Key themes include scene flow, 3D reconstruction from monocular video, autonomous driving perception, and robust evaluation of vision-language models. There is a clear emphasis on both methodological innovation and practical deployment in dynamic environments. Marr Prize, Honorable Mention (ICCV 2021) Best Paper, Honorable Mention (ECCV 2020) Best Paper Finalist (WACV 2024) Best Paper Award (WACV 2016) Best Industrial Paper, Honorable Mention (BMVC 2017) Marr Prize winner (ICCV 2009) Deva Ramanan has advised numerous PhD and master’s students, many of whom are now at top institutions and companies including Apple, Meta, Google, Nvidia, OpenAI, and Princeton. He has received substantial funding from IARPA, DARPA, NSF, Intel, Google, and Facebook for projects in video analytics, dispersed computing, visual cloud systems, and multi-task recognition. His group has developed influential datasets and benchmarks used widely in the community. He leads a vibrant research lab focused on advancing computer vision through deep learning and multimodal integration. His team works on core challenges in perception, including 3D reconstruction, motion modeling, object detection, and scene understanding, with applications in robotics and autonomous systems.
Shubham Tulsiani is an Assistant Professor at Carnegie Mellon University's Robotics Institute, where he leads the Computer Vision group and the Physical Perception Lab. His research focuses on inferring physically and spatially grounded representations from perceptual inputs, with applications in 3D vision, robot manipulation, and neural scene reconstruction. He directs an active research group with multiple PhD and Master's students. Research interests center on 3D scene understanding , robot learning , and generative modeling , with specific emphasis on: self-supervised perception, neural rendering, multi-view geometry, manipulation from visual inputs, and physics-based reasoning. The lab develops methods that leverage physical world constraints as supervisory signals. Recent publications demonstrate strong focus on diffusion models for 3D tasks , sparse-view reconstruction , and robotic manipulation transfer . Key trends include neural inverse rendering, view synthesis from limited observations, and translating human interactions to robot actions. Awards include: Best Student Paper Award at CVPR 2015 Advising includes supervision of 5 PhD students, 4 MS students, and undergraduates. Lab alumni hold positions at Google, Stanford, Meta, and Princeton. The Physical Perception Lab collaborates with FAIR Pittsburgh and the CMU Computer Vision group.
Matthew O'Toole is an Associate Professor at Carnegie Mellon University's School of Computer Science, holding joint appointments in the Robotics Institute and Computer Science Department. His research focuses on computational imaging, integrating optics, electronics, and computational processing to innovate visual information capture and display. Education: PhD (Computer Science, University of Toronto, 2016), MSc (2009), BSc (Honors Computer Science and Mathematics, University of British Columbia, 2007). Prior roles include Banting Postdoctoral Fellow at Stanford University and visiting scholar at MIT Media Lab's Camera Culture group. Research interests emphasize programmable imaging systems, transient imaging, non-line-of-sight sensing, and holographic displays. Key innovations include vibration sensing via dual-shutter optics and radar super-resolution for autonomous vehicles. Awards include runner-up best paper recognitions at ICCV 2007, CVPR 2014, and SIGGRAPH 2017 dissertation honors. Advisees include Dorian Chan and Arjun Teh. Grants supported by Canadian Banting Fellowships. Active in workshop organization (CVPR Computational Cameras 2016-2017) and course development on computational imaging at SIGGRAPH 2014. Labs/Teams: Leads research in computational imaging and robotics at CMU, collaborating with industry partners like NVIDIA and MDA. Current projects explore LiDAR-radar fusion, holographic projection systems, and dynamic scene reconstruction.
Prof. Matthias Nießner is a Professor at the Technical University of Munich , where he leads the Visual Computing Lab . Prior to this, he held a Visiting Assistant Professor position at Stanford University . His work bridges computer vision , graphics , and machine learning , focusing on 3D reconstruction , semantic scene understanding , and AI-driven video synthesis . Prof. Nießner has published over 150 works in top venues like SIGGRAPH , CVPR , and ECCV , with several receiving best paper awards (SIGCHI’14, HPG’15, SPG’18, SIGGRAPH’16 Emerging Tech). His research has garnered international media attention, including features in the New York Times , Wall Street Journal , and MIT Technological Review , as well as TV demonstrations (e.g., Jimmy Kimmel Live for Face2Face technology). Awards : TUM-IAS Rudolph Moessbauer Fellowship (2017–ongoing) Google Faculty Award (2017) Nvidia Professor Partnership Award (2018) ERC Starting Grant (2018, €1.5M) Eurographics Young Researcher Award (2019) Research Trends : 3D Gaussian Splatting for real-time rendering Neural Radiance Fields (NeRF) with mesh supervision Audio-driven facial animation via diffusion models Latent space diffusion for 3D scenes Self-supervised and zero-shot methods for 3D and image analysis As a co-founder and director of Synthesia Inc. , he drives democratization of synthetic media. His YouTube channel has over 5 million views, reflecting his impact beyond academia.
Andrew Childs is a Professor at the University of Maryland, affiliated with the Department of Computer Science and the Institute for Advanced Computer Studies (UMIACS). He serves as Director of the NSF Quantum Leap Challenge Institute for Robust Quantum Simulation (RQS) and is a Fellow at the Joint Center for Quantum Information and Computer Science (QuICS). His research focuses on quantum algorithms for simulating physical systems, algebraic problems, and quantum walk protocols, with applications in quantum computing and computational complexity. University of Maryland Institute for Advanced Computer Studies (UMIACS) Joint Center for Quantum Information and Computer Science (QuICS) NSF Quantum Leap Challenge Institute for Robust Quantum Simulation Childs' research spans quantum simulation, quantum Fourier transform, phase estimation, and Hamiltonian dynamics. He has developed techniques to reduce quantum computational resources for simulating quantum systems and explored limitations of quantum computers through hidden subgroup problems and non-unitary dynamics. His publications cover diverse areas including quantum walk optimization, Hamiltonian simulation methods, and applications to cryptography and condensed matter physics. Recent works address spatial search algorithms, product formulas for commutators, and quantum routing protocols. As an educator, Childs has taught courses on quantum algorithms and information processing at both the University of Maryland and University of Waterloo, with lecture notes and materials spanning multiple years. Contact: amchilds@umd.edu | Office: ATL 3359 | Affiliated with University of Maryland's quantum research institutes.
David W. Jacobs is a Professor in the Department of Computer Science at the University of Maryland, with a joint appointment at the University of Maryland Institute for Advanced Computer Studies (UMIACS). He also served as the interim Director of the University of Maryland Center for Machine Learning starting in 2018. University: University of Maryland School: College of Computer, Mathematical, and Natural Sciences Department: Department of Computer Science Academic Rank: Professor Education: He received his B.A. from Yale University, and M.S. and Ph.D. in Computer Science from MIT. Research Interests: His research primarily focuses on computer vision and machine learning, particularly visual object recognition, lighting variation modeling, 3D reconstruction, perceptual organization, motion understanding, and the integration of vision with graphics and human-computer interaction. A major applied contribution is the development of Leafsnap , an electronic field guide app for plant identification, which has been downloaded over 1.5 million times and used in biodiversity and educational contexts. Publication Trends: His recent scholarly output centers on deep learning, convolutional networks, residual architectures, generative models (especially GANs), and interpretability. His work often bridges theoretical insights with practical applications in vision and AI. Scientific Awards: Honorable Mention, Best Paper Award, CVPR 2000 Best Student Paper Award, UIST 2003 Best Paper Award, Eurographics 2016 2011 Edward O. Wilson Biodiversity Technology Pioneer Award for Leafsnap Teaching and Advising: He has taught advanced courses such as CMSC 422 (Introduction to Machine Learning) and CMSC 828L (Deep Learning). He mentors students through course projects and research, though specific advisees are not listed. He has collaborated with institutions like Columbia University and the Smithsonian on impactful interdisciplinary projects. Labs and Teams: He is affiliated with UMIACS and leads research efforts in vision and learning, contributing to the University of Maryland Center for Machine Learning. His team has developed several mobile applications including Leafsnap, Birdsnap, and Dogsnap, demonstrating a strong focus on real-world deployment of vision technology.
David Duvenaud is an Associate Professor at the University of Toronto , holding a Canada Research Chair in Generative Models and a Schwartz Reisman Chair in Technology and Society . He is cross-appointed to the Department of Computer Science and Department of Statistical Sciences . A Sloan Research Fellow and founding member of the Vector Institute , his work bridges deep probabilistic models , AI safety , and scientific computing . PhD in Machine Learning (University of Cambridge, 2014) Postdoc in Hyperparameter Optimization (Harvard University, 2016) Co-founded Invenia (energy forecasting company) His research spans foundational Neural Ordinary Differential Equations (NeurIPS 2018 Best Paper) and Automatic Chemical Design (ACS Central Science 2018) to recent work on AGI governance (2025) and AI safety (2024). Key contributions include stochastic variational inference , implicit differentiation frameworks , and antisymmetrization layers for quantum Monte Carlo. Recent publications (2024-2025) focus on systemic existential risks from AI , many-shot jailbreaking attacks , and epistemic uncertainty quantification . His group trains energy-based models with scalable MCMC samplers and develops invertible neural architectures (e.g., Residual Flows NeurIPS 2019). He also explores human-AI alignment through LLM Processes (NeurIPS 2024) and Sycophancy in Language Models (ICLR 2024). Canada Research Chair (2025) NSERC Grant (2025) Sloan Research Fellow (2021) Schwartz Reisman Chair (2021) Best Paper Award (NeurIPS 2018) Distinguished Paper Award (ICFP 2021) His students include James Requeima , Jesse Bettencourt , and Raymond Douglas . He teaches courses on Statistical Methods for Machine Learning and Differentiable Inference . Current work (2025) investigates systemic human disempowerment through incremental AI capabilities and sabotage risk mitigation via hyperparameter-aware evaluations.
Max Planck Institute for Intelligent SystemsGermany
Andreas Geiger is a Professor and Head of the Department of Computer Science at the University of Tübingen, Germany. He leads the Autonomous Vision Group (AVG) within CyberValley and is a core faculty member of the Tübingen AI Center. His roles also include PI in the ML in Science Excellence Cluster and the CRC Robust Vision, as well as ELLIS Fellow and coordinator of the ELLIS PhD program. He specializes in machine learning models for computer vision, robotics, and autonomous systems, with applications in self-driving cars, VR/AR, and scientific document analysis. Educational background: While not explicitly detailed, his positions imply a Ph.D. in Computer Science or related field. His work spans interdisciplinary collaborations with institutions like ETH Zürich, Microsoft, and the University of Bonn. Research focuses on 3D scene understanding, Gaussian splatting, generative models, and reliable autonomous systems. Notable contributions include the KITTI dataset and foundational work in neural radiance fields. Awards include the Sage 10-Year Impact Award (2024), ERC Starting Grant (2019), and IEEE PAMI Young Researcher Award (2018). Key projects include the Scholar Inbox paper recommender platform, ReSim (reliable world simulation), and advancements in 3D scene generation (e.g., UrbanCAD, PrITTI). His lab maintains a strong focus on open-source tools and datasets, such as the CARLA Route Generator. Grants and funding include support from Vector Stiftung (MINT innovation program) and EU initiatives like the ML in Science Cluster. His team collaborates internationally, with recent work presented at CVPR, SIGGRAPH, and NeurIPS.
Massachusetts Institute of TechnologyUnited States
Edward H. Adelson is the John and Dorothy Wilson Professor of Vision Science at MIT, affiliated with the Department of Brain and Cognitive Sciences and the Computer Science and Artificial Intelligence Laboratory (CSAIL). His research spans computer vision, human vision science, and robotics, with a focus on artificial touch sensing and tactile robotics. He has pioneered technologies like the GelSight tactile sensor, enabling high-resolution touch sensing for robots surpassing human skin sensitivity. Adelson holds a PhD in Experimental Psychology from the University of Michigan (1979) and a BA in Physics and Philosophy from Yale University (1974). His career includes roles at MIT since 1987, progressing from Associate Professor to Professor and later the Wilson Chair. He contributed to early vision theories, including the plenoptic function and motion energy models, and has been recognized with prestigious awards like the Helmholtz Prize (2013) and Rank Prize (1992). His research interests include material perception, optical sensing, and the integration of vision and touch. Key innovations include the plenoptic camera, layered representation techniques for motion analysis, and tactile sensors for robotics. Adelson has authored over 300 publications and holds numerous patents in imaging, vision, and robotics. Awards and honors include membership in the National Academy of Sciences and the American Academy of Arts and Sciences. His work bridges neuroscience and engineering, advancing both fundamental understanding and practical applications in robotics and computer vision.
Yi Fang is an Associate Professor of Computer Engineering and an affiliated Associate Professor of Computer Science at New York University Abu Dhabi (NYUAD), and a Global Network Associate Professor at NYU Tandon. He is a core faculty member in the Division of Engineering, specializing in Electrical and Computer Engineering. His research is centered at the intersection of Embodied AI, Robotics, and AI-driven assistive technologies, with strong support from agencies such as the US NSF, UAE ADEK, and ASPIRE. PhD, Purdue University Yi Fang's research interests span 3D Computer Vision, Multimedia Processing, Machine Learning, Deep Learning, and Embodied AI . He focuses on AI-driven perception, learning, and real-world applications, particularly in engineering, medicine, and accessibility. His lab, the Embodied AI and Robotics (AIR) Lab, develops intelligent robotic systems that integrate perception, learning, and decision-making to solve complex societal challenges. His work emphasizes large-scale visual computing, deep visual learning, and cross-domain/multimodal foundation models , with recent innovations in assistive AI for the Deaf and Hard-of-Hearing community. The 15 most recent publications reflect a consistent focus on 3D vision, sketch-based 3D retrieval, point cloud learning, and assistive computer vision . His work leverages deep learning, adversarial training, metric learning, and generative models to bridge modalities such as sketches, depth images, and 3D models. There is a clear trend toward cross-modal understanding, unsupervised representation learning, and real-world assistive applications , especially for visually impaired individuals. Yi Fang actively contributes to the academic community as an Area Chair for top-tier conferences including CVPR, ECCV, ICCV, IJCAI, and IROS. He also serves in peer review and mentoring roles, shaping the future of AI and robotics research. As a dedicated educator, he teaches foundational and advanced courses such as Computer Vision, Applied Machine Learning, Data Structures, and Capstone Design . He mentors students through research seminars and honors projects, fostering innovation and technical excellence. His research is supported by major grants from US NSF, UAE ADEK, and ASPIRE, enabling high-impact interdisciplinary collaborations. He founded and directs the Embodied AI and Robotics (AIR) Lab at NYU Abu Dhabi, a dedicated research space for developing intelligent systems that seamlessly integrate perception, learning, and decision-making. The lab promotes interdisciplinary collaboration across engineering, medicine, and social sciences, advancing the frontiers of Embodied AI.
California Institute of Technology (Caltech)United States
Georgia Gkioxari is an Assistant Professor in the Division of Computing and Mathematical Sciences at Caltech , with a part-time affiliation at Meta AI . Her work focuses on extending visual perception models through advanced 2D and 3D representation learning, spatial reasoning, and generative models. Education: Not explicitly mentioned in the text Research interests span 3D perception , spatial reasoning , and vision-language integration , with projects like Visual Agentic AI for Spatial Reasoning and Token-by-Token Multimodal Alignment . Her publications emphasize 3D object detection , reconstruction , and generative modeling techniques including diffusion models and transformers . Scientific recognition includes the Meta LLM Evaluation Research Grant , Okawa Research Grant , Google Faculty Scholar Award 2024 , and Amazon Research Award . She teaches courses like Large Language & Vision Models (EE/CS 148) and Learning & 3D (CS 101) at Caltech. Labs & Teams: Leads Glab with members including Ilona Demler, Ziqi Ma, and Damiano Marsili