Andrea Vedaldi is a Professor of Computer Vision and Machine Learning at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). He specializes in unsupervised methods for understanding images and videos, focusing on 3D geometry and semantics. His research bridges foundational AI and practical applications, with contributions to generative models, neural fields, and self-supervised learning. Education: PhD in Computer Science (2008), University of California, Los Angeles MSc in Computer Science (2005), UCLA BSc in Information Engineering (2003), University of Padua Research Interests: Unsupervised learning, 3D perception, generative AI, neural rendering, and scalable vision systems. His work emphasizes ethical, responsible AI aligned with ERC-funded projects like UNION (ERC Consolidator Grant). Key Contributions: Co-developer of VLFeat and MatConvNet libraries Leader in 3D reconstruction and diffusion models (e.g., CatFree3D) Recipient of the PAMI Thomas S. Huang Prize and multiple best paper awards Grants & Service: Principal Investigator on £2.3M ERC Consolidator Grant (UNION) Co-organizer of major conferences (ECCV 2020 Program Chair, CVPR 2023 Area Chair) Reviewer for top journals/conferences (PAMI, CVPR, NeurIPS) Labs & Teams: VGG Group at Oxford, collaborating on projects like Meta 3D Gen and Common Objects in 3D (CO3D).
Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Professor Andrew Davison holds the position of Professor of Robot Vision at Imperial College London's Department of Computing. He leads the Dyson Robotics Laboratory and the Robot Vision Research Group, focusing on advancing SLAM (Simultaneous Localization and Mapping) and Spatial AI. His groundbreaking work includes the MonoSLAM algorithm (2003), enabling real-time 3D vision for robotics and AR/VR. Current research emphasizes scalable, semantic-rich Spatial AI systems, as outlined in his FutureMapping papers (2018–2019). Education: BA in Physics (Oxford, 1994), D.Phil. (Oxford, 1998). Postdoctoral work at AIST, Japan (1998–2000), followed by a lectureship at Imperial (2002–present). Industrial collaborations include SLAMcore, a Spatial AI startup, and Dyson Robotics Lab. Over 18 PhD students supervised, many now leading roles at Meta, NVIDIA, SLAMcore, and academia. Notable contributions include DTAM, KinectFusion, and Event Camera SLAM. Recognized for software tools like SceneLib and contributions to robotics benchmarks (SLAMBench). Active on Twitter (@AjdDavison) for research updates.
Noah Snavely is a Professor of Computer Science at Cornell Tech and the Cornell Ann S. Bowers College of Computing and Information Science. His research spans computer vision and graphics with a focus on 3D scene understanding from image collections. He also works at Google DeepMind in NYC and leads the Cornell Graphics and Vision Group. His research interests include computer vision and computer graphics , particularly in recovering 3D structure from large photo collections, neural radiance fields, image-based rendering, and scene understanding. His work enables applications in mapping technologies, immersive VR experiences, and synthesizing 3D worlds from text prompts with implications for game design and filmmaking. Recent publications show a strong trend toward generative 3D modeling and neural rendering , with significant contributions in neural radiance fields, view synthesis, and dynamic scene reconstruction. His research group explores unstructured photo collections to develop new technology for 3D world modeling and image analysis. PECASE Microsoft New Faculty Fellowship Alfred P. Sloan Fellowship SIGGRAPH Significant New Researcher Award ACM Fellow (2023) IEEE Fellow NSF CAREER Award Helmholtz Prize Snavely has mentored numerous PhD students and postdocs including Ruojin Cai, Qianqian Wang, and Zhengqi Li. His teaching includes CS5670: Computer Vision across multiple semesters. He has received the Cornell College of Engineering Mr. and Mrs. Richard F. Tucker Teaching Excellence Award (2012). At Cornell Tech, his research group develops technology for modeling the world in 3D from online photo collections and analyzing images for computer graphics applications. His work has been featured in major news outlets including the New York Times, New Scientist, and MIT Technology Review.
Xingxing Zuo is an Assistant Professor (tenure-track) in the Robotics Department at MBZUAI. He holds a PhD from Zhejiang University (2021) and a Bachelor’s from UESTC (2016). Previously, he was a Postdoctoral Scholar at Caltech (2024–2025), a Postdoc at ETH Zurich (2019–2021), and held visiting roles at TU Munich, University of Delaware, and University of Technology Sydney. His research focuses on robotics, 3D computer vision, and embodied AI, with emphasis on robot-human collaboration, state estimation, and sensor fusion. Educations: PhD in Robotics, Zhejiang University (2021, with honors) Bachelor’s in Computer Science, University of Electronic Science and Technology of China (2016, with honors) Research Highlights: Develops novel methods for LiDAR-camera-inertial fusion, neural radiance fields, and radar-cameras systems Pioneered techniques like Flying Co-Stereo (long-range aerial mapping) and FMGS (vision-language embedded 3D splatting) Focuses on real-time SLAM, robust depth estimation, and photorealistic scene reconstruction Awards & Recognition: Best Paper Finalist at ICRA 2021 (CodeVIO) Oral Presentation at ICCV 2021 (MBA-VO) Recipient of Google Visiting Faculty Researcher (2023) Grants & Labs: Organized Thermal Infrared in Robotics workshop at ICRA 2025 Leads research on embodied AI and multi-sensor SLAM systems Develops open-source tools like LIC-Fusion and Coco-LIC frameworks
Katerina Fragkiadaki is the JPMorgan Chase Associate Professor of Computer Science in the Machine Learning Department at Carnegie Mellon University. She works at the intersection of Artificial Intelligence, Computer Vision, Machine Learning, Language Understanding, and Robotics. PhD from GRASP Lab, University of Pennsylvania Postdoctoral researcher at UC Berkeley (with Jitendra Malik) and Google Research Recipient of NSF CAREER, DARPA Young Investigator, Amazon, Google, Sony, UPMC, and AFOSR awards Organizer of CoRL 2023 Workshop on Generalist Robots ICLR 2024 Program Chair, multiple area chair roles Her research group focuses on developing machines that autonomously improve world models through human-environment interactions, with specific emphasis on: Representation learning and video understanding 2D/3D unified vision-language models Generative simulation and reinforcement learning Real2Sim/Sim2Real robot learning Continual learning and spatial common sense 3D scene reconstruction and dynamics Recent publications highlight advancements in: 3D mesh generation with compositional transformers Unified 2D/3D perception frameworks Physics-aware generative models Diffusion-based robotic manipulation policies Embodied agents with memory prompting Awards include: 2024: DARPA Young Investigator Award 2023: Amazon Faculty Award 2022: Sony Faculty Research Award 2021: UPMC Faculty Research Award 2020: NSF CAREER Award 2019: Google Faculty Award Key collaborations span institutions including UC Berkeley, Google Research, Stanford, MIT, and University of Tsukuba. Her work bridges theoretical innovation with practical applications in: Autonomous robot manipulation 4D world modeling Language-grounded perception Visual dynamics prediction Embodied program synthesis Physics-based simulation engines
Christian Rupprecht is an Associate Professor at the Department of Computer Science, University of Oxford, specializing in computer vision and machine learning. His research focuses on unsupervised learning, 3D reconstruction, and visual understanding. His work includes contributions to conferences such as GCPR'25, ICCV'25, and CVPR'25, with papers spanning topics like correspondence estimation, animal pose modeling, and synthetic data generation. He leads projects within the prestigious Visual Geometry Group (VGG). Notably, his paper VGGT received the Best Paper Award at CVPR'25. His research integrates deep learning and geometric modeling, emphasizing robustness and generalization in visual systems. Best Paper Award at CVPR'25
Olga Sorkine-Hornung is a Professor of Computer Science at ETH Zurich and head of the Institute of Visual Computing. She leads the Interactive Geometry Lab, focusing on theoretical and practical advancements in digital content creation, geometry processing, and shape modeling. Current position: ETH Zurich, Department of Computer Science Previous roles: Courant Institute (NYU), Technical University of Berlin Education: BSc and PhD from Tel Aviv University, postdoc at TU Berlin Her research spans shape representation, digital fabrication, computer animation, and fundamental geometry processing. Key contributions include Laplacian surface editing, as-rigid-as-possible deformation, and generalized winding numbers. She works on applications in VR, AR, and autonomous systems. Awards include Test of Time Awards (2024), ACM Fellow (2020), ERC Consolidator Grant (2020), and EUROGRAPHICS Young Researcher Award (2008). She has supervised numerous students and co-developed software libraries like libigl and Instant Meshes . Co-chair roles for SIGGRAPH, Eurographics, and Pacific Graphics Editorial board member for ACM Transactions on Graphics and other journals Keynote speaker at VMV, CVPR, and SIAM conferences Her work bridges mathematical rigor with practical implementation, advancing computer graphics and geometry processing through intuitive algorithms that maintain surface detail while enabling efficient computation.
Alexander Schwing is an Associate Professor in the Department of Electrical and Computer Engineering and Computer Science at the University of Illinois at Urbana-Champaign, affiliated with the Coordinated Science Laboratory. His research focuses on machine learning and computer vision with applications in 3D scene understanding, generative modeling, and multi-agent systems. Education: Diploma in Electrical Engineering and Information Technology, Technical University of Munich (TUM) PhD in Computer Science, ETH Zurich Postdoctoral Fellow, University of Toronto Research Interests: Structured prediction in deep learning Generative adversarial networks and stability Multi-modal vision-language models 3D scene reconstruction from single images Embodied agent collaboration Semantic segmentation with temporal coherence Recent Publications: Highlight trends in neural rendering, video object segmentation, and reinforcement learning with applications to 3D modeling and multi-agent systems. Notable innovations include SAIL-VOS dataset for amodal segmentation and NeRFDeformer for single-view scene transformation. Scientific Awards: NSF CAREER Award, 3M and Amazon research awards, multiple student recognition awards, ETH Zurich PhD medal, and best paper at Intelligent Tutoring Systems 2014. Teaching: Offers graduate courses in Pattern Recognition (ECE 544) and Machine Learning (CS 446/ECE 449). Previously taught at University of Toronto and ETH Zurich. Labs & Collaborations: Leads research at Coordinated Science Laboratory (UIUC) with collaborations across University of Toronto, ETH Zurich, and industry partners like Samsung SAIT and Amazon.
Wei-Lun (Harry) Chao is an Associate Professor in the Department of Computer Science and Engineering at the Ohio State University (OSU), College of Engineering. Promoted to this role in May 2025, he is also an Innovation Scholar and Distinguished Assistant Professor of Engineering Inclusive Excellence. His work spans machine learning, computer vision, and their applications in autonomous driving, healthcare, biology, and natural language processing. Research Focus: Machine learning with imperfect data, interpretable and personalized learning, robust perception for autonomous systems, and visual recognition in real-world scenarios. Awards: 2025 OSU Early Career Distinguished Scholar Award, CVPR Best Student Paper Award (2024), CSE Faculty Teaching Award (2024), Lumley Research Award (2023). Grants: Funded by NSF, NIH, ONR, Cisco, AWS, and Google. Notable Research Trends: The 15 most recent articles highlight his work on vision foundation models, federated learning, diffusion models for biological species generation, interpretable vision transformers, and robust perception systems for autonomous driving. Key subfields include sparse autoencoders, 3D object detection, semi-supervised learning, and anomaly detection in scientific domains. Scientific Awards: 2025 Early Career Distinguished Scholar Award (OSU) CVPR Best Student Paper Award (2024) CSE Faculty Teaching Award (2024) Lumley Research Award (2023) Mentoring & Grants: As an advisor for the OSU Buckeye AutoDrive Team and AI Club, he mentors graduate and undergraduate students. His research is supported by major grants from NSF, NIH, ONR, and industry partners like Cisco and Google.
Prof. Stefan Leutenegger is a tenure-track Assistant Professor at Technische Universität München (TUM), leading the Machine Learning for Robotics group within the TUM School of Computation, Information, and Technology. Previously, he held roles as Senior Lecturer (2018–2021) and Lecturer (2014–2018) at Imperial College London's Dyson Robotics Lab, where he founded the Smart Robotics Lab. He earned his PhD (2014) from ETH Zurich, focusing on autonomous solar-powered aircraft navigation, and holds BSc (2006) and MSc (2009) in Mechanical Engineering from ETH Zurich. His research centers on mobile robotics, particularly enabling robots (e.g., drones) to perceive and navigate complex environments using machine learning and sensor data fusion. Key focus areas include SLAM, event-based vision, 3D reconstruction, and autonomous exploration. He has pioneered algorithms like BRISK (2011), OKVIS (2014), and ElasticFusion (2016), advancing real-time robotics perception. Notable Awards: Imperial College President's Award (2018), Best ECCV Paper (2016), ETH Medal for Dissertations (2015). Labs: TUM's Machine Learning for Robotics Group, Imperial's Smart Robotics Lab. Publications: Over 100 papers, including seminal works in CVPR, ECCV, and Robotics: Science and Systems. Current projects include DigiForests (forest inventory via robotics), aerial additive manufacturing, and object-centric semantic mapping. His work bridges theory and practice, with applications in autonomous drones, construction robotics, and human-robot interaction.
Fadel Adib is an Associate Professor with Tenure at MIT, holding a joint appointment in the MIT Media Lab and the Department of Electrical Engineering and Computer Science (EECS). He leads the Signal Kinetics research group at the Media Lab and is the founder and CEO of Cartesian Systems. His work focuses on novel wireless technologies for sensing and connectivity in challenging environments, including underwater communication, health monitoring, and robotics. Education: B.Sc. from American University of Beirut (2011), M.Sc. and Ph.D. in Computer Science from MIT (2013 and 2016), where he won best thesis awards. His research has led to startups such as Cartesian Systems and Emerald Innovations, addressing applications in climate monitoring, healthcare, and logistics. Research Interests: Underwater backscatter networking, batteryless systems, non-line-of-sight perception, RFID localization, and wireless health monitoring. His innovations include the first battery-free underwater camera and systems for contactless vital signs monitoring. Awards: Sloan Research Fellowship (2021), NSF CAREER Award (2019), ACM SIGMOBILE Rockstar Award (2022), and Technology Review 35 Under 35 (2014). Grants: ONR Young Investigator Award, Google Faculty Research Award. Teams: Signal Kinetics research group at MIT Media Lab. Labs and Projects: Focus on cross-disciplinary initiatives such as wireless mapping (Cartesian Systems), underwater-to-air communication, and implantable medical devices.
Lourdes Agapito is a Professor of 3D Vision at the Department of Computer Science, University College London (UCL), within the Faculty of Engineering Sciences. She leads research in Non-Rigid Structure from Motion (NR-SFM) and 3D reconstruction from monocular video sequences. Her work addresses dynamic scenes, deformable objects, and articulated structures, with applications in robotics and computer vision. She holds an ERC Starting Grant (2008–2014) and led the EU Horizon 2020-funded Second Hands project (2014–2019), collaborating with institutions like EPFL and KIT to develop robots with 3D visual perception for maintenance tasks. Her research group focuses on dense optical flow estimation, video registration, and deformable tracking. Agapito’s research interests include monocular 3D reconstruction, non-rigid motion analysis, and neural approaches to 3D modeling. She has supervised multiple PhD students and postdocs, including notable researchers such as Ravi Garg and Marco Paladini (Sullivan Prize recipient). Her contributions to conferences include roles as Program Chair for CVPR 2016 and CVPR 2017, and she has authored influential papers on topics like Video-Popup (ECCV 2014) and Modal Space (CVPR 2017). Current projects involve advancing neural parametric models and real-time 3D reconstruction techniques. Awards include the ERC Starting Grant and recognition for her team’s work in non-rigid reconstruction. She actively mentors students and collaborates on grants, with recent openings for postdocs and PhD candidates in 3D vision and robotics.
April Yi Wang is a tenure-track Assistant Professor in the Department of Computer Science at ETH Zürich, where she directs the Programming, Education, and Computer-Human Interaction Lab (PEACH Lab). She is a core faculty member at the Institute for Intelligent Interactive Systems and associated with the ETH AI Center. Wang is also an active member of ETH HCI and Swiss CHI communities, contributing significantly to human-computer interaction and educational technology research. Dr. Wang's educational background includes: Ph.D. in Information Science from University of Michigan (2023), advised by Steve Oney and Christopher Brooks M.Sc. in Computer Science from Simon Fraser University (2018), advised by Parmit Chilana B.Eng in Computer Science from Zhejiang University (2016) Dr. Wang's research focuses on human-centered approaches to programming and data science. Her work reimagines programming as a form of literature that communicates with both machines and people, exploring creative representations like text, shapes, animations, and everyday objects. She investigates how to make programming more natural and intuitive through literate programming environments, with applications in professional and educational contexts. Her research spans human-computer interaction, educational technology, and AI-assisted programming tools. Analysis of Dr. Wang's recent publications reveals a strong focus on AI-enhanced educational tools, particularly for programming and data literacy. Her work increasingly integrates large language models to scaffold learning while maintaining user agency. There's a clear trajectory toward developing situated learning approaches that connect abstract concepts to real-world contexts through augmented reality and tangible interfaces. Her research bridges HCI, education, and AI to create more accessible and engaging technical learning experiences. Dr. Wang has received numerous prestigious awards including: 2023 Gary M. Olson Award and Honourable Mention Award at ACM CHI 2022 Rising Stars in EECS and Heidelberg Laureate Forum Young Researcher 2020 Best Short Paper Award at IEEE VL/HCC and Honourable Mention at ACM CHI 2019 Best Paper Award at ACM CSCW Dr. Wang actively mentors students through thesis projects at ETH Zürich, supervising numerous bachelor's and master's students on topics ranging from AI-assisted programming to data literacy tools. Her lab, PEACH Lab, has secured funding including the recent innovedum funding for the Coducate project. She serves on program committees for major conferences including CHI and UIST, and regularly reviews for top HCI and education journals. The PEACH Lab, directed by Dr. Wang, focuses on creating expressive, intelligent, and human-centered systems that make technical topics more accessible. The lab explores textual, visual, and embodied representations for programming, with emphasis on enhancing communication, collaboration, and learning. Current research directions include balancing automation with user agency, supporting diverse learning needs, and developing tools for interdisciplinary technical communication.