Andrea Vedaldi is a Professor of Computer Vision and Machine Learning at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). He specializes in unsupervised methods for understanding images and videos, focusing on 3D geometry and semantics. His research bridges foundational AI and practical applications, with contributions to generative models, neural fields, and self-supervised learning. Education: PhD in Computer Science (2008), University of California, Los Angeles MSc in Computer Science (2005), UCLA BSc in Information Engineering (2003), University of Padua Research Interests: Unsupervised learning, 3D perception, generative AI, neural rendering, and scalable vision systems. His work emphasizes ethical, responsible AI aligned with ERC-funded projects like UNION (ERC Consolidator Grant). Key Contributions: Co-developer of VLFeat and MatConvNet libraries Leader in 3D reconstruction and diffusion models (e.g., CatFree3D) Recipient of the PAMI Thomas S. Huang Prize and multiple best paper awards Grants & Service: Principal Investigator on £2.3M ERC Consolidator Grant (UNION) Co-organizer of major conferences (ECCV 2020 Program Chair, CVPR 2023 Area Chair) Reviewer for top journals/conferences (PAMI, CVPR, NeurIPS) Labs & Teams: VGG Group at Oxford, collaborating on projects like Meta 3D Gen and Common Objects in 3D (CO3D).
Jean Oh is a Researcher at the Robotics Institute of Carnegie Mellon University (CMU) , leading the interdisciplinary Bot Intelligence Group (BIG) . Her work focuses on developing persistent robots that co-exist and collaborate with humans in shared environments, emphasizing continuous improvement through training, exploration, and human interaction. Education: Ph.D. in Language and Information Technologies, CMU M.S. in Computer Science, Columbia University B.S. in Biotechnology, Yonsei University Oh's research integrates vision, language, and planning systems in robotics, with applications in human-robot teaming , self-driving cars , disaster response , eldercare , and creative robotics . She has pioneered projects like socially-compliant robot navigation in human crowds and AI-driven robotic painting systems. Recent publication trends highlight her work in vision-language planning , social navigation , computational creativity , and human-robot collaboration . Notable contributions include the StyleCLIPDraw algorithm for text-to-art generation and Social-PatteRNN for human-like trajectory prediction. Scientific Awards: Best Paper Award in Cognitive Robotics (ICRA'18, ICRA'15) Best Systems Paper Finalist (HRI'25) Best Oral Paper Finalist (Humanoids'24) Best Paper in Entertainment (IROS'24) Argoverse Challenge Winner (CVPR'24) Best Student Paper (AIAA'24) Best Demo Finalist (RoboSoft'24) Oh mentors a diverse team of PhD, MS, and undergraduate students from CMU departments including Robotics, Computer Science, and Mechanical Engineering. Her research is funded by US Army Research Lab , DiDi Chuxing , and DARPA , with collaborations across industry and academia .
Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Christian Theobalt is a Professor of Computer Science at Saarland University and Scientific Director of the Visual Computing and Artificial Intelligence Department at the Max Planck Institute for Informatics . He leads the Saarbruecken Center for Visual Computing as a strategic partnership between Google and MPI. PhD in Computer Science (2005) from MPI-INF/Saarland University Postdoctoral Researcher at MPI (2005-2007) Visiting Assistant Professor at Stanford (2007-2009) His research focuses on the intersection of Computer Graphics, Computer Vision, and Artificial Intelligence , with specializations in: 3D/4D Human Reconstruction Neural Rendering Performance Capture Geometric Deep Learning Volumetric Video Quantum Visual Computing Recent article trends show emphasis on: Quantum computing applications in visual reconstruction Neural rendering with radiance fields Human motion capture from egocentric views Gesture-language interaction modeling Multi-modal scene understanding Real-time free-viewpoint rendering Scientific Awards : Fellow of EUROGRAPHICS (2022) CVPR Best Student Paper Honorable Mention (2020) ERC Consolidator Grant (2017) Karl Heinz Beckurts Award (2017) Busy Beaver Teaching Award (2016) ERC Starting Grant (2013) Advising over 50 students and researchers including: Current researchers: Viktor Rudnev, Linjie Lyu, Mohit Mendiratta Postdocs: Kwang In Kim, Kiran Varanasi Alumni: Franziska Mueller (Google), Dushyant Mehta (Qualcomm), Ayush Tewari (MIT) Grants include ERC grants, Google Glass Research Award, and multiple industry partnerships. His lab maintains cutting-edge facilities with: Multi-camera capture systems Quantum annealing infrastructure HDR radiance field technology GPU clusters for AI research Time-of-flight imaging systems Event camera arrays
Justin Johnson is an Assistant Professor at the University of Michigan's College of Engineering, Department of Electrical Engineering and Computer Science, and a Research Scientist at Facebook AI Research (FAIR). His work bridges computer vision, machine learning, and deep learning, focusing on visual reasoning, vision-language tasks, image generation, and 3D reasoning using neural networks. PhD, Stanford University (advised by Fei-Fei Li) His research interests span visual reasoning , vision and language , 3D vision , and image generation , with a focus on innovative applications of deep neural networks. Recent publications highlight work on 3D consistency, self-supervised learning, and multimodal integration of vision and text. Notable contributions include PyTorch3D for 3D data processing, and foundational work in visual question answering , neural style transfer , and scene graph-based image generation . Publications span top conferences like ICCV, CVPR, and NeurIPS. He teaches courses including EECS 498/598: Deep Learning for Computer Vision and EECS 442: Computer Vision at University of Michigan, with prior involvement in Stanford's CS 231N in co-teaching roles. Software projects include open-source frameworks like fast-neural-style for real-time artistic style transfer, and PyTorch3D for efficient 3D deep learning. These tools demonstrate his commitment to practical implementations and community-driven research.
Yuki M. Asano is a full Professor at the University of Technology Nuremberg , leading the Fundamental AI (FunAI) Lab . Previously, he led the QUVA Lab at the University of Amsterdam and earned his PhD at the Visual Geometry Group (VGG) of the University of Oxford under Andrea Vedaldi and Christian Rupprecht. University of Technology Nuremberg (2024–present) University of Amsterdam (prior to 2024) University of Oxford (PhD, 2020) His research spans Artificial Intelligence , Machine Learning , and Computer Vision , with a focus on Causal Representation Learning , Self-Supervised Learning , and Efficient Model Adaptation . He pioneered techniques like BISCUIT (causal variable identification) and VeRA (parameter-efficient fine-tuning). His work extends to Medical Imaging and Environmental Monitoring through applications in fetal ultrasound analysis and marine debris detection. Recent publications (2023–2025) highlight advancements in Self-Supervised Learning , Vision-Language Models , and 3D Understanding . Notable papers include TWIST & SCOUT (multimodal LLM grounding), SIGMA (masked video modeling), and GeneralAD (anomaly detection). His ICCV 2023 work on Self-Ordering Point Clouds and MoSiC (optimal-transport motion trajectories) underscores his interdisciplinary approach. He received the JUPITER compute grant (2025) and an Outstanding Paper Award at ICLR 2024 . His collaborations span institutions like MIT-IBM Watson AI Lab, Qualcomm AI Research, and University of Amsterdam.
Kristen Grauman is a Full Professor in the Department of Computer Science at the University of Texas at Austin, where she leads the UT Computer Vision Group. Her research focuses on computer vision and machine learning, with applications in visual recognition, video analysis, and multi-modal perception. She received her B.A. from Boston College and her Ph.D. from MIT. Her research interests span visual recognition, image and video search, video analysis, first-person vision, embodied and multi-modal perception, and interactive machine learning. She has made significant contributions to the field, particularly in developing algorithms for understanding visual content and human activities from video, including foundational work on the Pyramid Match Kernel and relative attributes. Her recent publications reveal a strong emphasis on egocentric (first-person) vision, audio-visual learning, and view-invariant representations. There is a clear trajectory toward multi-modal integration (vision, audio, language) and real-world applications in instructional videos, human activity understanding, and embodied AI systems. She has received numerous awards including: AAAI Fellow (2019) J. K. Aggarwal Prize, International Association for Pattern Recognition (2018) Helmholtz Prize (2017) UT Austin Academy of Distinguished Teachers (2017) Best Paper Award, Asian Conference on Computer Vision (2016) Presidential Early Career Award for Scientists and Engineers (2014) Computers and Thought Award, International Joint Conferences on Artificial Intelligence (2013) Pattern Analysis and Machine Intelligence Young Researcher Award (2013) Alfred P. Sloan Research Fellow (2012) Marr Prize (2011) Prof. Grauman serves as Associate Editor-in-Chief for the IEEE Transactions on Pattern Analysis and Machine Intelligence. She has secured substantial research funding including the Presidential Early Career Award, NSF grants, and industry partnerships. Her advising has produced numerous influential publications and students who are now leaders in computer vision. She leads the UT Computer Vision Group, which collaborates closely with the Electrical and Computer Engineering Department. The group is pioneering large-scale egocentric video research through projects like Ego4D and Ego-Exo4D, focusing on real-world applications in human activity understanding, audio-visual perception, and interactive systems.
Jennifer Olsen, PhD, is an Assistant Professor of Computer Science at the University of San Diego since 2020. She holds a PhD, MS, and BS in Human-Computer Interaction and Cognitive Science from Carnegie Mellon University, followed by postdoctoral research at the Swiss Federal Institute of Technology (EPFL), Lausanne, Switzerland. Her research focuses on the intersection of human-computer interaction, cognition, and education, emphasizing collaborative learning and educational technology design from both learner and instructor perspectives. Education: PhD in Human-Computer Interaction, Carnegie Mellon University MS in Human-Computer Interaction, Carnegie Mellon University BS in Cognitive Science, Carnegie Mellon University Research Interests: Dr. Olsen explores how collaboration supports learning, designs technologies to enhance educational practices, and investigates gaze-based metrics for understanding collaborative problem-solving. Her work spans gamified robotics, AI-driven orchestration systems, and virtual reality applications in vocational training. She emphasizes learner-centered design and the integration of social robots and virtual agents in pedagogical settings. Grants/Advising: While no specific grants or advisees are listed, her prolific publication record indicates active involvement in educational technology research and development. Her work addresses challenges in classroom orchestration, multimodal data analysis, and accessibility in educational robotics. Labs/Teams: Collaborates with interdisciplinary teams focused on educational technology, human-robot interaction, and adaptive learning systems. Her research leverages tools like FROG orchestration graphs and eye-tracking technologies to develop practical classroom solutions.
Adrian Weller is a prominent researcher and academic at the University of Cambridge, serving as a Director of Research in Machine Learning within the Department of Engineering. He holds multiple significant leadership roles including Programme Director for Trust and Society at the Leverhulme Centre for the Future of Intelligence (CFI), and previously served as Programme Director for AI at The Alan Turing Institute, the UK national institute for data science and AI. His work bridges theoretical machine learning research with practical applications and societal implications of artificial intelligence. Weller's research interests span a broad spectrum of AI and machine learning topics with a particular focus on ensuring beneficial societal outcomes. His work encompasses explainability, fairness, robustness, scalability, privacy, safety, and ethics in AI systems. He has made significant contributions to trustworthy machine learning, including developing frameworks for AI governance, certification, and human-AI collaboration. His research group actively investigates neuro-symbolic approaches, privacy-preserving techniques, and methods for improving the reliability and interpretability of AI systems. His recent publications demonstrate a strong trend toward addressing the practical challenges of deploying AI systems in real-world contexts, particularly focusing on certification frameworks, governance mechanisms, and human-centered approaches. His work spans theoretical advances in machine learning architectures while maintaining a strong connection to societal impact, with publications appearing in top venues across AI, machine learning, and interdisciplinary applications. Scientific Awards: MBE for services to digital innovation (2022 Queen's Birthday Honours) Turing AI Fellowship for Trustworthy Machine Learning Weller actively supervises a large group of PhD students and postdocs, with current students including Juyeon Heo, Yanzhi Chen, Katie Collins, Isaac Reid, Yichao Liang, Herbie Bradley, and Shoaib Siddiqui. His former students have gone on to positions at leading institutions including Google DeepMind, ETH Zurich, NYU, and MPI-IS Tübingen. He has served on numerous advisory boards including the Centre for Data Ethics and Innovation, UNESCO's expert group on AI ethics, and the World Economic Forum's Global Future Council on AI. His research has been supported through his Turing AI Fellowship and various collaborative projects focused on safe and ethical AI development. Weller leads a vibrant research group focused on trustworthy machine learning, which actively organizes workshops and conferences including ICML 2024 (where he served as Program Chair), multiple workshops on responsible AI, and events through the ELLIS network. His group collaborates extensively across disciplines, working with researchers in computer science, social sciences, law, and policy to address the multifaceted challenges of developing beneficial AI systems.
Tien Tsin Wong is a Professor in the Department of Data Science & AI at Monash University, Australia. Previously, he served as a Professor at the Chinese University of Hong Kong (1999–2024) and held a Visiting Assistant Professor position at the Hong Kong University of Science and Technology (1998–1999). His research focuses on Generative AI, Computer Graphics, Computer Vision, and Computational Manga, with significant contributions to GPU techniques, image-based rendering, and multimedia compression. Education: He earned a B.Sc. (1992), MPhil (1994), and PhD (1998) in Computer Science from the Chinese University of Hong Kong. Research Interests: His work bridges computational techniques with artistic applications, particularly in manga and animation. Notable areas include generative models, diffusion-based video synthesis, and physically plausible scene generation. His research aligns with UN Sustainable Development Goals through innovations in education and digital accessibility. Awards : He has received the 2004 Young Researcher Award, 2005 IEEE Transactions on Multimedia Prize Paper Award, and two international invention medals (Geneva 2018, Asia Hong Kong 2019). Editorial Roles : He serves as an Associate Editor for Computer Graphics Forum , IEEE Transactions on Visualization and Computer Graphics , and Computational Visual Media . His editorial work underscores his influence in advancing visualization and graphics research. Labs/Teams : While not explicitly named, his collaborations span global institutions, focusing on computational manga, generative AI, and GPU-optimized techniques. His work often involves interdisciplinary teams addressing challenges in digital media and AI.
Shoudong Huang is a Professor at the School of Mechanical and Mechatronic Engineering , University of Technology Sydney, and Deputy Director of the UTS Robotics Institute. His research focuses on mobile robot navigation , SLAM , nonlinear state estimation , and surgical robotics . He has published over 200 papers and is recognized as one of the 100 Most Influential Scholars in Robotics (Aminer, 2018). PhD in Automatic Control, Northeastern University (China) Postdoctoral Research Fellow, University of Hong Kong (1998-2000) Research Fellow, Australian National University (2001-2003) Full-time academic roles at UTS since 2004 His work addresses challenges in robot localization across extreme environments (underwater, underground mining, surgical settings) and develops globally optimal SLAM algorithms with guaranteed performance. He has secured over $4 million AUD in external funding, including ARC Discovery grants and industry partnerships. Recent publications emphasize cross-modal calibration (camera-LiDAR), interval analysis for bounded noise , and template-based deformable surface reconstruction . These span applications in autonomous driving, surgical navigation, and UAV guidance. Chancellor’s Medal for Research Excellence (2020) Supervisor of the Year (2023) Best Paper Award (2016 ICARCV) Huang serves as Associate Editor for IEEE Transactions on Robotics and International Journal of Robotics Research , and has held leadership roles in top robotics conferences like IROS and RSS. His collaborations span MIT, USC, Zhejiang University, and industry partners including PMSW Research Pty Ltd and Multiplex Constructions Pty Ltd.
Huan Zhang serves as an Assistant Professor in the Department of Electrical and Computer Engineering at the University of Illinois Urbana-Champaign (UIUC), with affiliate appointments in the Department of Computer Science and the Coordinated Science Laboratory. His research focuses on building trustworthy AI systems with formal verification techniques to provide provable guarantees for safety-critical applications, particularly in machine learning and neural networks. Dr. Zhang received his Ph.D. in Computer Science from UCLA in 2020, advised by Professor Cho-Jui Hsieh. His academic journey includes an M.S. in Computer Engineering from UC Davis (2014) and a Bachelor of Engineering from Zhejiang University (2012). Prior to joining UIUC, he completed a postdoctoral fellowship at Carnegie Mellon University (2021-2023) with Professor Zico Kolter. Huan Zhang's research program centers on formal verification of machine learning systems, with particular emphasis on neural network verification, AI safety, robustness, and reliability. He pioneered the linear bound propagation-based verification framework that enables formal verification for networks with millions of neurons. His work spans five major research categories: formal verification of machine learning, training trustworthy ML models, machine learning safety and adversarial attacks, reinforcement learning safety, and optimization for scalable machine learning. His CROWN framework (NeurIPS 2018) established a foundational approach for neural network verification through efficient linear bound propagation. His recent publications demonstrate a strategic expansion from foundational verification techniques toward increasingly complex systems including large language models, vision-language models, and robotic control systems. The research trajectory shows a clear progression from theoretical frameworks to practical implementations with real-world applications, particularly in safety-critical domains. His work increasingly bridges formal methods with practical AI deployment requirements. Winner of International Verification of Neural Networks Competition (VNN-COMP) as team leader (2021-2024) Schmidt Futures AI2050 Early Career Fellowship ($300,000 research grant) Adversarial Machine Learning (AdvML) Rising Star Award (2021) IBM PhD Fellowship (2018) Dr. Zhang leads the development of α,β-CROWN, a neural network verifier that has won VNN-COMP 2021-2023, and auto_LiRPA, a PyTorch-based library for perturbation analysis on general computational graphs. He has mentored numerous graduate students from CMU, UCLA, UIUC, and Columbia University. His research is supported by significant funding including the Schmidt Futures fellowship and industry collaborations. He teaches courses including ECE 120, ECE 484, ECE 584, and ECE 598 HZ on topics ranging from computing fundamentals to safe autonomy and machine learning. Dr. Zhang maintains active research collaborations across multiple institutions and is affiliated with UIUC's Coordinated Science Laboratory. His work has significant implications for safety-critical AI applications in autonomous systems, healthcare, and other mission-critical domains where reliability guarantees are essential. He regularly gives guest lectures at institutions including Yale, Stony Brook, and the University of Nebraska Lincoln on formal verification techniques.
Noah Snavely is a Professor of Computer Science at Cornell Tech and the Cornell Ann S. Bowers College of Computing and Information Science. His research spans computer vision and graphics with a focus on 3D scene understanding from image collections. He also works at Google DeepMind in NYC and leads the Cornell Graphics and Vision Group. His research interests include computer vision and computer graphics , particularly in recovering 3D structure from large photo collections, neural radiance fields, image-based rendering, and scene understanding. His work enables applications in mapping technologies, immersive VR experiences, and synthesizing 3D worlds from text prompts with implications for game design and filmmaking. Recent publications show a strong trend toward generative 3D modeling and neural rendering , with significant contributions in neural radiance fields, view synthesis, and dynamic scene reconstruction. His research group explores unstructured photo collections to develop new technology for 3D world modeling and image analysis. PECASE Microsoft New Faculty Fellowship Alfred P. Sloan Fellowship SIGGRAPH Significant New Researcher Award ACM Fellow (2023) IEEE Fellow NSF CAREER Award Helmholtz Prize Snavely has mentored numerous PhD students and postdocs including Ruojin Cai, Qianqian Wang, and Zhengqi Li. His teaching includes CS5670: Computer Vision across multiple semesters. He has received the Cornell College of Engineering Mr. and Mrs. Richard F. Tucker Teaching Excellence Award (2012). At Cornell Tech, his research group develops technology for modeling the world in 3D from online photo collections and analyzing images for computer graphics applications. His work has been featured in major news outlets including the New York Times, New Scientist, and MIT Technology Review.
Xingxing Zuo is an Assistant Professor (tenure-track) in the Robotics Department at MBZUAI. He holds a PhD from Zhejiang University (2021) and a Bachelor’s from UESTC (2016). Previously, he was a Postdoctoral Scholar at Caltech (2024–2025), a Postdoc at ETH Zurich (2019–2021), and held visiting roles at TU Munich, University of Delaware, and University of Technology Sydney. His research focuses on robotics, 3D computer vision, and embodied AI, with emphasis on robot-human collaboration, state estimation, and sensor fusion. Educations: PhD in Robotics, Zhejiang University (2021, with honors) Bachelor’s in Computer Science, University of Electronic Science and Technology of China (2016, with honors) Research Highlights: Develops novel methods for LiDAR-camera-inertial fusion, neural radiance fields, and radar-cameras systems Pioneered techniques like Flying Co-Stereo (long-range aerial mapping) and FMGS (vision-language embedded 3D splatting) Focuses on real-time SLAM, robust depth estimation, and photorealistic scene reconstruction Awards & Recognition: Best Paper Finalist at ICRA 2021 (CodeVIO) Oral Presentation at ICCV 2021 (MBA-VO) Recipient of Google Visiting Faculty Researcher (2023) Grants & Labs: Organized Thermal Infrared in Robotics workshop at ICRA 2025 Leads research on embodied AI and multi-sensor SLAM systems Develops open-source tools like LIC-Fusion and Coco-LIC frameworks