Dr. Chang Xu is an Associate Professor in Machine Learning and Computer Vision at the University of Sydney's School of Computer Science. He holds a Bachelor of Engineering from Tianjin University and a PhD from Peking University. His research focuses on machine learning, data mining, and their applications in AI and computer vision, including multi-view learning, visual search, and face recognition. He is an ARC Future Fellow and a member of the Sydney Southeast Asia Centre and The Net Zero Institute. Education: B.E. in Engineering (Tianjin University), Ph.D. in Computer Science (Peking University). His research interests emphasize handling heterogeneous data, exploring data variety, and developing algorithms for robust AI systems. His work includes adversarial robustness, neural architecture search, and efficient deep learning models. Research trends in his articles include adversarial robustness in neural architectures, efficient vision transformers, multimodal 3D style transfer, and underwater image restoration. Key contributions span image restoration, video super-resolution, and lightweight network design. He has advised multiple PhD and master's students on topics like diffusion models, radar image synthesis, and graph similarity. Awards: ARC Future Fellow. Collaborations focus on cross-domain data integration and AI applications. His labs and teams explore generative models, robust learning, and scalable robotics policies. Recent work includes diffusion models for action segmentation and robust vision-language systems.
Jean Oh is a Researcher at the Robotics Institute of Carnegie Mellon University (CMU) , leading the interdisciplinary Bot Intelligence Group (BIG) . Her work focuses on developing persistent robots that co-exist and collaborate with humans in shared environments, emphasizing continuous improvement through training, exploration, and human interaction. Education: Ph.D. in Language and Information Technologies, CMU M.S. in Computer Science, Columbia University B.S. in Biotechnology, Yonsei University Oh's research integrates vision, language, and planning systems in robotics, with applications in human-robot teaming , self-driving cars , disaster response , eldercare , and creative robotics . She has pioneered projects like socially-compliant robot navigation in human crowds and AI-driven robotic painting systems. Recent publication trends highlight her work in vision-language planning , social navigation , computational creativity , and human-robot collaboration . Notable contributions include the StyleCLIPDraw algorithm for text-to-art generation and Social-PatteRNN for human-like trajectory prediction. Scientific Awards: Best Paper Award in Cognitive Robotics (ICRA'18, ICRA'15) Best Systems Paper Finalist (HRI'25) Best Oral Paper Finalist (Humanoids'24) Best Paper in Entertainment (IROS'24) Argoverse Challenge Winner (CVPR'24) Best Student Paper (AIAA'24) Best Demo Finalist (RoboSoft'24) Oh mentors a diverse team of PhD, MS, and undergraduate students from CMU departments including Robotics, Computer Science, and Mechanical Engineering. Her research is funded by US Army Research Lab , DiDi Chuxing , and DARPA , with collaborations across industry and academia .
Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Justin Johnson is an Assistant Professor at the University of Michigan's College of Engineering, Department of Electrical Engineering and Computer Science, and a Research Scientist at Facebook AI Research (FAIR). His work bridges computer vision, machine learning, and deep learning, focusing on visual reasoning, vision-language tasks, image generation, and 3D reasoning using neural networks. PhD, Stanford University (advised by Fei-Fei Li) His research interests span visual reasoning , vision and language , 3D vision , and image generation , with a focus on innovative applications of deep neural networks. Recent publications highlight work on 3D consistency, self-supervised learning, and multimodal integration of vision and text. Notable contributions include PyTorch3D for 3D data processing, and foundational work in visual question answering , neural style transfer , and scene graph-based image generation . Publications span top conferences like ICCV, CVPR, and NeurIPS. He teaches courses including EECS 498/598: Deep Learning for Computer Vision and EECS 442: Computer Vision at University of Michigan, with prior involvement in Stanford's CS 231N in co-teaching roles. Software projects include open-source frameworks like fast-neural-style for real-time artistic style transfer, and PyTorch3D for efficient 3D deep learning. These tools demonstrate his commitment to practical implementations and community-driven research.
Rhenish Friedrich Wilhelm University of BonnGermany
Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Prof. Konrad Schindler holds the position of Full Professor at the Department of Civil, Environmental and Geomatic Engineering at ETH Zürich. He is also the Head of the Institute of Geodesy and Photogrammetry (IGP), leading research and educational activities in geomatics and computer vision. His career spans roles as a Photogrammetric Engineer, scientific assistant, postdoc researcher, and academic faculty across institutions including Graz University of Technology, Monash University, and TU Darmstadt before joining ETH Zürich in 2010. Education: Undergraduate studies in Geodesy (1992–1995), Graz University of Technology, Austria MEng in Photogrammetry and Geoinformation (1995–1999), Vienna University of Technology, Austria PhD in Computer Science (2001–2003), Graz University of Technology, Austria Research focuses on Photogrammetry , Remote Sensing , Computer Vision , and Image Understanding with interdisciplinary applications in environmental monitoring, geospatial analysis, and disaster response. He develops computational methods for 3D reconstruction, fusion of multi-modal data, and AI-driven solutions for satellite imagery interpretation. His work bridges geomatic engineering and machine learning to address challenges in urban mapping, climate modeling, and biological systems analysis. Publications reflect expertise in geospatial AI, diffusion models, and benchmarking datasets for disaster resilience. Notable works include Marigold (image analysis adaptation) and BRIGHT (building damage assessment). His research emphasizes practicality and scalability, such as affordable depth estimation and global biomass datasets. He has received the 2013 Marr Prize Honourable Mention (IEEE) and the 2012 U.V. Helava Award (ISPRS), alongside several Best Presentation Awards. His contributions span technical leadership, editorial roles (ISPRS Journal), and service to Swiss remote sensing commissions. Advising and grants: While no specific advisee names or grant details are listed, his career trajectory includes mentoring postdocs and junior faculty. He teaches advanced courses in Photogrammetry , Image Interpretation , and Machine Vision , integrating cutting-edge AI techniques into curricula. His research group collaborates on global-scale projects like canopy height mapping and satellite-based climate variable assessments. Labs/Teams: As Institute Head, he oversees the IGP lab at ETH Zürich, with prior affiliations including the Digital Perception Lab (Monash University) and the Computer Vision Lab (ETH Zurich). His work often involves multi-institutional collaborations focused on geospatial AI and environmental science.
Shrikanth (Shri) Narayanan is a University Professor and holder of the Niki and Max Nikias Chair in Engineering at the University of Southern California (USC), serving as the inaugural Vice President for Presidential Initiatives. He leads the Signal Analysis and Interpretation Lab (SAIL) and holds joint appointments in Computer Science, Linguistics, Psychology, Neuroscience, Pediatrics, and Otolaryngology-Head and Neck Surgery. His research focuses on speech and audio processing, behavioral signal processing, and real-time MRI of speech production, with applications in healthcare, education, and technology. Education: B.E. in Electrical Engineering from College of Engineering, Guindy (Chennai, India, 1988); M.S., Engineer, and Ph.D. in Electrical Engineering from UCLA (1990, 1992, 1995). Research interests span computational linguistics, machine learning, and multimodal human behavior analysis. He pioneered technologies for speech biomarkers in mental health, real-time MRI of speech production, and wearable sensor systems for longitudinal health studies. His work in speech emotion recognition, forensic interviews, and clinical applications has been recognized through over 40 awards, including the IEEE Flanagan Award and ISCA Medal. He has published extensively in journals like Proceedings of the IEEE , Journal of the Acoustical Society of America , and PLOS One . Key Grants: NSF CAREER, Okawa Research, IBM Faculty, Google/Amazon awards. Labs/Teams: Signal Analysis & Interpretation Lab (SAIL), USC Information Sciences Institute (ISI), Google Visiting Faculty Researcher.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Hao Su is an Associate Professor in the Department of Computer Science and Engineering at University of California, San Diego . He serves as Chairman & CTO of Hillbot Inc , and leads the SU Lab which focuses on building autonomous systems that learn actively in physical environments. His affiliations include the Institute for Learning-enabled Optimization at Scale , Artificial Intelligence Group , Contextual Robotics Institute , Halicioğlu Data Science Institute , and Center for Visual Computing . As a researcher in Computer Vision, Robotics, and Neural Geometry , he has made significant contributions to 3D foundation models, reward-free world models, diffusion policy frameworks, and GPU-accelerated simulation environments. His 2024-2025 publications include advancements in hand-eye calibration, dynamic mesh reconstruction, and multi-stage robotic manipulation. His scientific awards include: Frontiers of Science Award (2025) TPAMI Young Research Award (2025) NSF CAREER Award (2023) ACM SIGGRAPH Best Doctorate Thesis Honorable Mention (2019) He has served as Program Chair for CVPR 2025 and Area Chair for ICLR 2022 and NeurIPS 2023 , while previously serving as Publication Chair for 3DV 2016 and Program Committee for SIGGRAPH Asia Workshops .
Michael Kaess is an Associate Professor at the Robotics Institute, Carnegie Mellon University (CMU), within the School of Computer Science. He leads the Robot Perception Lab (RPL) and contributes to the Field Robotics Center (FRC) and Computer Vision Group (CV). His research focuses on efficient perception algorithms for mobile robots, particularly in 3D mapping, SLAM, and sensor fusion using vision, LiDAR, inertial, and sonar data. Kaess holds a PhD in Computer Science from Georgia Tech and was a postdoc at MIT's Marine Robotics Lab. Education: Georgia Institute of Technology, PhD in Computer Science (2008) MIT, Postdoctoral Associate (2008–2010) Research Interests: Kaess develops algorithms for robust and efficient inference in robotics, emphasizing factor graphs and linear algebra. His work spans underwater robotics, aerial systems, tactile SLAM, and multi-sensor integration. Key areas include SLAM with planes/lines, imaging sonar reconstruction, and neural field methods for LiDAR-visual fusion. Publications: Over 145 papers, including work on EDPLVO (visual odometry), HoloOcean (underwater simulation), and neural radiance fields with LiDAR. Recent trends focus on robust incremental smoothing, acoustic-optical fusion, and real-time volumetric mapping. Awards: Recognized with the RSS Test of Time Award (2020), Outstanding Associate Editor (2022), and paper awards at ICRA/ICRA. Active in conference organization (IROS/ICRA program committees). Advising & Grants: Supervises 10+ current PhD/MSc students, with past advisees contributing to CoRL/ICRA work. Manages grants in perception, autonomy, and marine robotics. Teaches courses like Robot Localization and Mapping (16-833). Labs/Teams: Directs RPL, collaborates with FRC on field robotics. Develops open-source tools like GTSAM (GNU Toolkit for Smoothing and Mapping).
Kenji Kawaguchi is the Presidential Young Professor in the Department of Computer Science at the National University of Singapore (NUS), where he leads the Deep Learning Lab and is a faculty affiliate at the NUS Institute of Data Science. His research bridges theoretical and applied machine learning, focusing on deep learning, large language models, and physics-informed neural networks. His educational background includes a Ph.D. and S.M. in Computer Science and Electrical Engineering from the Massachusetts Institute of Technology (MIT), advised by Leslie Pack Kaelbling, and a postdoctoral fellowship at Harvard University’s Center of Mathematical Sciences and Applications. Dr. Kawaguchi’s research interests center on the theoretical foundations of deep learning, optimization, generalization, and applications in areas such as molecular modeling, AI safety, and efficient training of large models. He has made significant contributions to understanding in-context learning, diffusion models, and neural operators for partial differential equations. His recent publications (2023–2025) reflect a strong trend toward improving the efficiency, robustness, and interpretability of large-scale models, particularly in language and scientific domains. Key themes include LLM alignment and safety, diffusion model optimization, and physics-informed learning for high-dimensional problems. Presidential Young Professor He has served as Area Chair and PC Member for top-tier conferences including NeurIPS, ICML, ICLR, AAAI, and UAI, and as reviewer for journals such as JMLR and Annals of Statistics. He has delivered invited talks at Harvard, MIT, Stanford, CMU, Brown, and Google Research, reflecting his international recognition. He actively mentors students and welcomes PhD candidates and postdocs to join his research group.
Tien Tsin Wong is a Professor in the Department of Data Science & AI at Monash University, Australia. Previously, he served as a Professor at the Chinese University of Hong Kong (1999–2024) and held a Visiting Assistant Professor position at the Hong Kong University of Science and Technology (1998–1999). His research focuses on Generative AI, Computer Graphics, Computer Vision, and Computational Manga, with significant contributions to GPU techniques, image-based rendering, and multimedia compression. Education: He earned a B.Sc. (1992), MPhil (1994), and PhD (1998) in Computer Science from the Chinese University of Hong Kong. Research Interests: His work bridges computational techniques with artistic applications, particularly in manga and animation. Notable areas include generative models, diffusion-based video synthesis, and physically plausible scene generation. His research aligns with UN Sustainable Development Goals through innovations in education and digital accessibility. Awards : He has received the 2004 Young Researcher Award, 2005 IEEE Transactions on Multimedia Prize Paper Award, and two international invention medals (Geneva 2018, Asia Hong Kong 2019). Editorial Roles : He serves as an Associate Editor for Computer Graphics Forum , IEEE Transactions on Visualization and Computer Graphics , and Computational Visual Media . His editorial work underscores his influence in advancing visualization and graphics research. Labs/Teams : While not explicitly named, his collaborations span global institutions, focusing on computational manga, generative AI, and GPU-optimized techniques. His work often involves interdisciplinary teams addressing challenges in digital media and AI.
Kyros Kutulakos is a Professor in the Department of Computer Science at the University of Toronto, where he leads research in computational imaging and 3D sensing. His affiliations include the Toronto Computational Imaging Group, Computer Vision Group, Dynamic Graphics Project (DGP), and Vector Institute Group. He teaches graduate and undergraduate courses such as CSC320 (Introduction to Visual Computing) and CSC2530 (Computational Imaging & 3D Sensing). His research interests span computational imaging, non-line-of-sight imaging, single-photon detectors, 3D sensing, and neural rendering. Notable contributions include advancements in structured-light imaging, time-of-flight systems, and super-oscillatory microscopy. He has advised numerous PhD and MSc students, fostering cutting-edge research in imaging technologies. Kutulakos has received prestigious awards, including the Dean’s Research Excellence Award (2023) and multiple best paper prizes (e.g., Marr Prize at ICCV 2023). He has served as program chair for ICCV 2013, ICCP 2010, and CVPR 2003, contributing to academic leadership in computer vision. His work bridges optics, photonics, and computation, with applications in autonomous systems, medical imaging, and astronomy. Current research focuses on extreme imaging scenarios, such as imaging in pitch-black environments and around corners, leveraging novel sensor designs and computational techniques.
California Institute of Technology (Caltech)United States
Xingxing Zuo is an Assistant Professor (tenure-track) in the Robotics Department at MBZUAI. He holds a PhD from Zhejiang University (2021) and a Bachelor’s from UESTC (2016). Previously, he was a Postdoctoral Scholar at Caltech (2024–2025), a Postdoc at ETH Zurich (2019–2021), and held visiting roles at TU Munich, University of Delaware, and University of Technology Sydney. His research focuses on robotics, 3D computer vision, and embodied AI, with emphasis on robot-human collaboration, state estimation, and sensor fusion. Educations: PhD in Robotics, Zhejiang University (2021, with honors) Bachelor’s in Computer Science, University of Electronic Science and Technology of China (2016, with honors) Research Highlights: Develops novel methods for LiDAR-camera-inertial fusion, neural radiance fields, and radar-cameras systems Pioneered techniques like Flying Co-Stereo (long-range aerial mapping) and FMGS (vision-language embedded 3D splatting) Focuses on real-time SLAM, robust depth estimation, and photorealistic scene reconstruction Awards & Recognition: Best Paper Finalist at ICRA 2021 (CodeVIO) Oral Presentation at ICCV 2021 (MBA-VO) Recipient of Google Visiting Faculty Researcher (2023) Grants & Labs: Organized Thermal Infrared in Robotics workshop at ICRA 2025 Leads research on embodied AI and multi-sensor SLAM systems Develops open-source tools like LIC-Fusion and Coco-LIC frameworks
Adriana Kovashka is an Associate Professor at the University of Pittsburgh , affiliated with the School of Computing and Information and serving as Department Chair . Her academic journey began with BA degrees in Computer Science and Media Studies from Pomona College (2008) and a PhD in Computer Science from The University of Texas at Austin (2014). Joined Pitt’s faculty in January 2015 NSF CAREER awardee (2021) Google Faculty Research Award recipient Dr. Kovashka’s research spans Computer Vision , Machine Learning , and Natural Language Processing , focusing on visual rhetoric, weak multimodal supervision, and domain adaptation. She pioneered techniques for analyzing political imagery, developing robust object detection frameworks, and exploring the intersection of visual and textual persuasion through large-scale annotated datasets. Her recent work emphasizes geographic diversity in vision-language systems, audio-visual fusion for domain generalization, and shape-texture bias mitigation in CNNs. Key publications include groundbreaking studies on symbolic reasoning, multimodal dialogue systems, and ethical AI applications in education. Scientific honors include: NSF CRII Award (2016) NSF CAREER Award (2021) Pitt CRDF Award (2016, 2018) Best Paper at ECV Workshop (2021) Google Faculty Research Award (2016, 2018) Dr. Kovashka actively mentors students in multimodal learning projects and collaborates with interdisciplinary teams on NSF-funded initiatives. She co-organizes workshops like the first CVPR workshop on advertisement understanding and leads research groups exploring human-AI co-learning systems.