California Institute of Technology (Caltech)United States
Xingxing Zuo is an Assistant Professor (tenure-track) in the Robotics Department at MBZUAI. He holds a PhD from Zhejiang University (2021) and a Bachelor’s from UESTC (2016). Previously, he was a Postdoctoral Scholar at Caltech (2024–2025), a Postdoc at ETH Zurich (2019–2021), and held visiting roles at TU Munich, University of Delaware, and University of Technology Sydney. His research focuses on robotics, 3D computer vision, and embodied AI, with emphasis on robot-human collaboration, state estimation, and sensor fusion. Educations: PhD in Robotics, Zhejiang University (2021, with honors) Bachelor’s in Computer Science, University of Electronic Science and Technology of China (2016, with honors) Research Highlights: Develops novel methods for LiDAR-camera-inertial fusion, neural radiance fields, and radar-cameras systems Pioneered techniques like Flying Co-Stereo (long-range aerial mapping) and FMGS (vision-language embedded 3D splatting) Focuses on real-time SLAM, robust depth estimation, and photorealistic scene reconstruction Awards & Recognition: Best Paper Finalist at ICRA 2021 (CodeVIO) Oral Presentation at ICCV 2021 (MBA-VO) Recipient of Google Visiting Faculty Researcher (2023) Grants & Labs: Organized Thermal Infrared in Robotics workshop at ICRA 2025 Leads research on embodied AI and multi-sensor SLAM systems Develops open-source tools like LIC-Fusion and Coco-LIC frameworks
Katerina Fragkiadaki is the JPMorgan Chase Associate Professor of Computer Science in the Machine Learning Department at Carnegie Mellon University. She works at the intersection of Artificial Intelligence, Computer Vision, Machine Learning, Language Understanding, and Robotics. PhD from GRASP Lab, University of Pennsylvania Postdoctoral researcher at UC Berkeley (with Jitendra Malik) and Google Research Recipient of NSF CAREER, DARPA Young Investigator, Amazon, Google, Sony, UPMC, and AFOSR awards Organizer of CoRL 2023 Workshop on Generalist Robots ICLR 2024 Program Chair, multiple area chair roles Her research group focuses on developing machines that autonomously improve world models through human-environment interactions, with specific emphasis on: Representation learning and video understanding 2D/3D unified vision-language models Generative simulation and reinforcement learning Real2Sim/Sim2Real robot learning Continual learning and spatial common sense 3D scene reconstruction and dynamics Recent publications highlight advancements in: 3D mesh generation with compositional transformers Unified 2D/3D perception frameworks Physics-aware generative models Diffusion-based robotic manipulation policies Embodied agents with memory prompting Awards include: 2024: DARPA Young Investigator Award 2023: Amazon Faculty Award 2022: Sony Faculty Research Award 2021: UPMC Faculty Research Award 2020: NSF CAREER Award 2019: Google Faculty Award Key collaborations span institutions including UC Berkeley, Google Research, Stanford, MIT, and University of Tsukuba. Her work bridges theoretical innovation with practical applications in: Autonomous robot manipulation 4D world modeling Language-grounded perception Visual dynamics prediction Embodied program synthesis Physics-based simulation engines
Christian Rupprecht is an Associate Professor at the Department of Computer Science, University of Oxford, specializing in computer vision and machine learning. His research focuses on unsupervised learning, 3D reconstruction, and visual understanding. His work includes contributions to conferences such as GCPR'25, ICCV'25, and CVPR'25, with papers spanning topics like correspondence estimation, animal pose modeling, and synthetic data generation. He leads projects within the prestigious Visual Geometry Group (VGG). Notably, his paper VGGT received the Best Paper Award at CVPR'25. His research integrates deep learning and geometric modeling, emphasizing robustness and generalization in visual systems. Best Paper Award at CVPR'25
Sham Kakade is the Rampell Family Professor of Computer Science and Professor of Statistics at Harvard University, co-director of the Kempner Institute. His research focuses on advancing artificial general intelligence through foundational work in reinforcement learning, large-scale learning systems, and autonomous agent architectures. He earned his PhD in 2003 from the Gatsby Computational Neuroscience Unit at University College London. His work emphasizes scalable optimization algorithms, distributed systems for foundation models, and understanding emergent capabilities in neural architectures. Research interests include full-stack training pipelines for foundation models, mathematical principles of large-scale learning systems, and bridging language models with embodied intelligence. He advises prospective students with backgrounds in applied deep learning or theoretical computer science, offering access to the Kempner Institute's computational resources. He serves on committees for the ACM Prize in Computing and Sloan Research Fellowships, co-organizes the Simons Symposium on Theoretical Machine Learning, and chaired COLT 2011. His lab works at the intersection of theory and practice, addressing challenges in AI's societal impact and technical scalability. Labs/Teams: Co-directs the Kempner Institute, fostering collaborations between AI researchers and social scientists. Active in Harvard's SEAS community.
Olga Sorkine-Hornung is a Professor of Computer Science at ETH Zurich and head of the Institute of Visual Computing. She leads the Interactive Geometry Lab, focusing on theoretical and practical advancements in digital content creation, geometry processing, and shape modeling. Current position: ETH Zurich, Department of Computer Science Previous roles: Courant Institute (NYU), Technical University of Berlin Education: BSc and PhD from Tel Aviv University, postdoc at TU Berlin Her research spans shape representation, digital fabrication, computer animation, and fundamental geometry processing. Key contributions include Laplacian surface editing, as-rigid-as-possible deformation, and generalized winding numbers. She works on applications in VR, AR, and autonomous systems. Awards include Test of Time Awards (2024), ACM Fellow (2020), ERC Consolidator Grant (2020), and EUROGRAPHICS Young Researcher Award (2008). She has supervised numerous students and co-developed software libraries like libigl and Instant Meshes . Co-chair roles for SIGGRAPH, Eurographics, and Pacific Graphics Editorial board member for ACM Transactions on Graphics and other journals Keynote speaker at VMV, CVPR, and SIAM conferences Her work bridges mathematical rigor with practical implementation, advancing computer graphics and geometry processing through intuitive algorithms that maintain surface detail while enabling efficient computation.
Alexander Schwing is an Associate Professor in the Department of Electrical and Computer Engineering and Computer Science at the University of Illinois at Urbana-Champaign, affiliated with the Coordinated Science Laboratory. His research focuses on machine learning and computer vision with applications in 3D scene understanding, generative modeling, and multi-agent systems. Education: Diploma in Electrical Engineering and Information Technology, Technical University of Munich (TUM) PhD in Computer Science, ETH Zurich Postdoctoral Fellow, University of Toronto Research Interests: Structured prediction in deep learning Generative adversarial networks and stability Multi-modal vision-language models 3D scene reconstruction from single images Embodied agent collaboration Semantic segmentation with temporal coherence Recent Publications: Highlight trends in neural rendering, video object segmentation, and reinforcement learning with applications to 3D modeling and multi-agent systems. Notable innovations include SAIL-VOS dataset for amodal segmentation and NeRFDeformer for single-view scene transformation. Scientific Awards: NSF CAREER Award, 3M and Amazon research awards, multiple student recognition awards, ETH Zurich PhD medal, and best paper at Intelligent Tutoring Systems 2014. Teaching: Offers graduate courses in Pattern Recognition (ECE 544) and Machine Learning (CS 446/ECE 449). Previously taught at University of Toronto and ETH Zurich. Labs & Collaborations: Leads research at Coordinated Science Laboratory (UIUC) with collaborations across University of Toronto, ETH Zurich, and industry partners like Samsung SAIT and Amazon.
Prof. Matthias Nießner is a Professor at the Technical University of Munich, leading the Visual Computing Lab. His research intersects computer graphics, vision, and AI, focusing on 3D reconstruction, semantic understanding, and AI-driven video synthesis. He holds a PhD from the University of Erlangen-Nuremberg (2013) and was a Visiting Assistant Professor at Stanford University (2013–2017). Notable awards include the ERC Starting Grant (2018), Nvidia Professorship Award, and Eurographics Young Researcher Award (2019). His work has been featured in mainstream media and led to startups like Synthesia Inc. Research spans Gaussian splatting, neural radiance fields, and generative AI for 3D avatars. Over 150 publications include SIGGRAPH, CVPR, and ECCV, with best paper awards. Projects like Face2Face and ScanNet have driven innovation in facial reenactment and 3D scene datasets. Education: PhD in Computer Science, University of Erlangen-Nuremberg (2013) Diploma in Computer Science, University of Erlangen-Nuremberg (2010) Research Interests: 3D digitization, neural rendering, generative AI, non-rigid reconstruction, and applications in AR/VR. Awards: ERC Starting Grant (2018) Nvidia Professorship Award (2018) Google Faculty Award (2018) SIGGRAPH Best Emerging Tech Award (2016) Grants: Over €1.5M from ERC and industry partnerships. Labs/Teams: Visual Computing Lab at TUM and Synthesia Inc. (co-founder). Key projects include ScanNet (large 3D indoor dataset), Face2Face (real-time facial reenactment), and Gaussian-based 3D avatars. Current work focuses on diffusion models, neural radiance fields, and AI-generated media detection.
Alexei A. Efros is a Professor in the Department of Electrical Engineering and Computer Sciences (EECS) at UC Berkeley, where he holds the Howard Friesen Professorship and is affiliated with the Berkeley Artificial Intelligence Research (BAIR) Lab. He previously served on the faculty at the Robotics Institute of Carnegie Mellon University (CMU) and completed a postdoctoral fellowship at the University of Oxford. His research spans data-driven computer vision, self-supervised learning, computational photography, and applications to computer graphics and robotics. His research interests include: Data-Driven Computer Vision Self-Supervised and Unsupervised Learning Generative Models and Image Synthesis Visual Representation Learning Applications in Robotics and Human-Computer Interaction Intersections with Human Vision and the Humanities The recent publications highlight a strong trend toward self-supervised learning, visual reasoning, and generative modeling, particularly diffusion models and 3D scene understanding. His work increasingly bridges computer vision with language, robotics, and cognitive science, emphasizing interpretability and real-world applicability. There is a clear focus on leveraging unlabeled data and developing methods for robust, generalizable AI systems. His scientific awards and recognitions include: Berkeley Fellowship Google Fellowship Soros Fellowship NSF Fellowship SIGGRAPH Outstanding Doctoral Dissertation Award Facebook Fellowship Adobe Fellowship CMU School of Computer Science Distinguished Dissertation Award ACM Doctoral Dissertation Honorable Mention Alexei Efros has advised numerous PhD students and postdocs, many of whom have gone on to faculty positions at top institutions including CMU, Stanford, MIT, Columbia, NYU, and Georgia Tech. His lab has received research funding from major tech companies and federal agencies, though specific grants are not detailed in the text. He teaches core computer vision and machine learning courses at both undergraduate and graduate levels at UC Berkeley. His research group is highly active, with ongoing projects in 3D perception, generative modeling, and vision-language systems. He leads a vibrant research lab at UC Berkeley, part of the BAIR consortium, collaborating with leading researchers such as Jitendra Malik, Trevor Darrell, Pieter Abbeel, and Angjoo Kanazawa. His lab fosters strong interdisciplinary connections with institutions worldwide, including Oxford, INRIA, and École Normale Supérieure.
Anca Dragan is an Associate Professor in the Department of Electrical Engineering and Computer Sciences at the University of California, Berkeley, where she runs the InterACT Lab focused on algorithms for human-AI and human-robot interaction. Currently on leave from Berkeley, she leads AI Safety and Alignment at Google DeepMind, overseeing safety for Gemini models and preparing for future advancements. Dragan has been a co-PI of the Center for Human-Compatible AI and served on the steering committee for the Berkeley AI Research (BAIR) Lab. B.Sc. in Computer Science from Jacobs University Bremen, Germany Ph.D. from Carnegie Mellon University Dr. Dragan's research focuses on enabling AI agents to work effectively with, around, and in support of people. Her work bridges robotics, machine learning, and game theory to create systems that better understand human preferences and coordinate with users. Key areas include AI alignment (ensuring AI does what people actually want), learning reward functions from diverse human feedback forms, and developing algorithms for human-AI collaboration across domains like autonomous vehicles, brain-machine interfaces, and recommender systems. Her research emphasizes maintaining uncertainty about human preferences and accounting for the plurality of human values. Dr. Dragan's recent publications reveal a strong focus on addressing fundamental challenges in AI safety and alignment. Her work spans theoretical foundations of reward learning, practical implementations for human-AI coordination, and critical examinations of limitations in current approaches. There's a clear trajectory toward more robust, safe, and value-aligned AI systems that can handle complex human preferences while avoiding both present-day harms and potential catastrophic risks. IEEE RAS Early Academic Career Award in Robotics and Automation (2021) McEntyre Award for Excellence in Teaching (2020) PECASE (Presidential Early Career Award for Science and Engineering) (2019) Sloan Fellowship (2018) NSF CAREER Award (2017) Okawa Foundation Award (2017) MIT Tech Review 35 Innovators Under 35 (2017) Multiple best paper awards at top robotics and AI conferences Dr. Dragan has mentored numerous successful students who have gone on to faculty positions at MIT, Stanford, CMU, and Princeton, as well as industry roles at DeepMind, Waymo, and Meta. Her advising philosophy emphasizes both technical rigor and consideration of broader societal impacts. She has secured significant research funding including NSF CAREER, ONR Young Investigator, and Okawa Foundation awards, supporting work on human-AI interaction and alignment. Dragan has also consulted for Waymo for six years, helping develop roadmaps for deploying increasingly learning-based safety-critical systems. Dr. Dragan leads the InterACT Lab at UC Berkeley, which has produced influential work on Cooperative Inverse Reinforcement Learning, Inverse Reward Design, and other foundational concepts in human-AI interaction. The lab's research has significantly shaped the field of AI alignment, with applications spanning autonomous vehicles that coordinate with human drivers, brain-machine interfaces that adapt to user needs, and language models that better understand human preferences. Current work focuses on scaling safety approaches as AI capabilities advance, ensuring alignment keeps pace with technological progress.
University of Illinois Urbana-ChampaignUnited States
Derek W Hoiem is a Professor in the Siebel School for Computing and Data Science at the University of Illinois Urbana-Champaign, where he has been a faculty member since 2009. His research focuses on computer vision and related areas, and he is also the co-founder and Chief Science Officer of Reconstruct, an AI-based construction technology company. His educational background includes: PhD in Robotics, Carnegie Mellon University (2007) Beckman Postdoctoral Fellowship (2008) Prof. Hoiem's research spans computer vision, with a focus on object recognition, scene understanding, and graphics. His work also extends to mobile robotics and 3D scene reconstruction. He has made significant contributions in areas such as visual recognition, 3D modeling, and the application of computer vision in construction monitoring. His recent publications (2023-2025) demonstrate a strong focus on advancing multimodal understanding, particularly in region-based representations, 3D vision, and neural radiance fields. There is a clear trend towards integrating language and vision, improving efficiency in neural networks, and applying computer vision to real-world problems such as construction progress monitoring. His scientific awards and honors are extensive and include: IEEE Fellow (2022) University Scholar (2022) Koendrink Prize (2022) Dean's Award for Excellence in Research, Associate Professor (2021) Campus Distinguished Promotion Award (2015) Best Paper Award: IEEE Winter Conference on Applications in Computer Vision (WACV) (2015) CW Gear Junior Faculty Award (2014) IEEE PAMI Young Researcher Award (2014) Dean's Award for Excellence in Research, Assistant Professor (2014) Sloan Research Fellowship (2013) Intel Early Career Faculty Honor Program Award (2012) NSF CAREER Award (2011) ACM Doctoral Dissertation Award, Honorable Mention (2008) Carnegie Mellon University SCS Distinguished Dissertation Award (2008) Best Paper Award: IEEE Computer Vision and Pattern Recognition (CVPR) (2006) Prof. Hoiem has secured significant research funding, including an NSF CAREER award and an Intel Early Career Faculty award. He is also actively involved in technology transfer, having co-founded Reconstruct where he serves as Chief Science Officer. His teaching excellence is reflected in multiple "List of Teachers Ranked as Excellent" awards spanning from 2010 to 2021. Prof. Hoiem leads a research group at UIUC focused on computer vision and 3D scene understanding. Additionally, he co-founded and serves as Chief Science Officer at Reconstruct, which develops AI-based solutions for construction monitoring.
Swiss Federal Institute of Technology in LausanneSwitzerland
Mathieu Salzmann is a Senior Scientist and Lecturer at École Polytechnique Fédérale de Lausanne (EPFL), affiliated with the Computer Vision Laboratory (CVLAB) in the School of Computer and Communication Sciences (IC). He also holds a courtesy appointment with the EPFL College of Humanities and serves as Deputy Chief Data Scientist at the Swiss Data Science Center (SDSC). He has held concurrent roles in teaching units including SIN, SODH, and SSC, reflecting his interdisciplinary engagement. His research focuses on the intersection of machine learning and computer vision, particularly in deep learning for 2D and 3D visual scene understanding, efficient and robust models, domain adaptation, and interpretable AI. These interests are evident across his extensive publication record in top-tier venues. His recent publications (2023–2024) show a consistent trend in advancing deep learning methods for visual recognition, with strong representation at CVPR, ICCV, ECCV, ICML, ICLR, and NeurIPS. Topics include domain generalization, 3D understanding, model robustness, and multimodal learning, often with applications in real-world systems. His editorial roles as Associate Editor for IEEE TPAMI and Action Editor for TMLR further highlight his leadership in the field. Area Chair: ICML 2023, CVPR 2023, ICCV 2023, NeurIPS 2023, AAAI 2024, ECCV 2024 Associate Editor: IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) Action Editor: Transactions on Machine Learning Research (TMLR) Mathieu Salzmann has supervised numerous PhD students at EPFL, both current and past, including Bouquet Yann Yanis, Javed Saqib, Li Shuangqi, and others. He has also been involved in research grants and collaborative projects, such as his work with S. Süsstrunk and R. Baroni on comics reconfiguration. His part-time role as Senior GNC Engineer at ClearSpace (2020–2024) illustrates his applied research engagement in aerospace systems. He is actively involved in EPFL’s data science and AI research ecosystem through SDSC and multiple labs.
Swiss Federal Institute of Technology in LausanneSwitzerland
Lior Wolf is a Professor at the School of Computer Science, Tel Aviv University. Previously, he was a postdoctoral researcher at MIT's Center for Biological and Computational Learning (CBCL) under Prof. Tomaso Poggio and earned his PhD from Hebrew University of Jerusalem with Prof. Amnon Shashua. His educational background includes: PhD in Computer Science, Hebrew University of Jerusalem Postdoctoral Research, MIT CBCL Prof. Wolf's research centers on artificial intelligence with seminal contributions to deep learning, computer vision, and natural language processing. His work bridges theoretical foundations (e.g., attention mechanisms, transformer analysis) with practical applications in medical imaging, speech processing, and sign language technology. He pioneered methods for neural network interpretability, efficient sequence modeling, and multimodal fusion. Analysis of his 2023-2025 publications reveals dominant trends in large language model optimization (neuron pruning, attention analysis), efficient video generation, and cross-modal learning. His work increasingly integrates medical applications (fMRI/EEG analysis) while maintaining theoretical rigor in model architecture design. His scientific achievements include: Best paper award at EMNLP 2024 for 'Backward Lens: Projecting Language Model Gradients into the Vocabulary Space' Best paper award at SCIA 2023 for 'Gradient Adjusting Networks for Domain Inversion' Best paper award at FG 2021 for 'Generating Master Faces for Dictionary Attacks' Prof. Wolf mentors graduate students in the School of Computer Science and leads research at the ICRC building laboratory. His team collaborates with 'the friends of TAU' on projects spanning biometric security, medical imaging, and generative AI. Current work focuses on efficient transformers, neural network interpretability, and multimodal medical diagnostics.
Lourdes Agapito is a Professor of 3D Vision at the Department of Computer Science, University College London (UCL), within the Faculty of Engineering Sciences. She leads research in Non-Rigid Structure from Motion (NR-SFM) and 3D reconstruction from monocular video sequences. Her work addresses dynamic scenes, deformable objects, and articulated structures, with applications in robotics and computer vision. She holds an ERC Starting Grant (2008–2014) and led the EU Horizon 2020-funded Second Hands project (2014–2019), collaborating with institutions like EPFL and KIT to develop robots with 3D visual perception for maintenance tasks. Her research group focuses on dense optical flow estimation, video registration, and deformable tracking. Agapito’s research interests include monocular 3D reconstruction, non-rigid motion analysis, and neural approaches to 3D modeling. She has supervised multiple PhD students and postdocs, including notable researchers such as Ravi Garg and Marco Paladini (Sullivan Prize recipient). Her contributions to conferences include roles as Program Chair for CVPR 2016 and CVPR 2017, and she has authored influential papers on topics like Video-Popup (ECCV 2014) and Modal Space (CVPR 2017). Current projects involve advancing neural parametric models and real-time 3D reconstruction techniques. Awards include the ERC Starting Grant and recognition for her team’s work in non-rigid reconstruction. She actively mentors students and collaborates on grants, with recent openings for postdocs and PhD candidates in 3D vision and robotics.
Andrea Tagliasacchi is an Associate Professor in the School of Computing Science at Simon Fraser University (SFU), where he holds the Visual Computing Research Chair. He is also a part-time (20%) Staff Research Scientist at Google DeepMind in Toronto and holds an associate professor (status only) appointment in the Department of Computer Science at the University of Toronto. Education: PhD in Computing Science – Simon Fraser University (NSERC Alexander Graham Bell Fellow) Postdoctoral Research – École Polytechnique Fédérale de Lausanne (EPFL) MSc in Computer Science – Politecnico di Milano (Gold Medalist) His research lies at the intersection of computer vision, computer graphics, and machine learning, with a focus on 3D visual perception. Key areas include neural radiance fields (NeRF), 3D Gaussian splatting, inverse rendering, and geometric deep learning, with applications in robotics, augmented reality, and autonomous systems. His work emphasizes robust and efficient scene understanding and reconstruction from visual data. His recent publications, appearing in top venues like CVPR, SIGGRAPH, NeurIPS, and ECCV, demonstrate a strong emphasis on neural fields, 3D reconstruction, and generative modeling. Trends include improving rendering efficiency, enhancing robustness to noise and distractors, and enabling controllable and 3D-aware generation. His group has made significant contributions to Gaussian splatting, NeRF optimization, and diffusion-based 3D/4D synthesis. Scientific Awards: 2024 CVPR Best Paper Award (Honorable Mention) 2020 CVPR Best Student Paper Award 2015 SGP Best Paper Award NSERC Alexander Graham Bell Canada Graduate Scholarship MITACS Best Paper Award (SIGGRAPH Asia 2009) NSF Best Poster Award (SGP 2012) He has advised numerous PhD and MSc students, many of whom are now researchers at leading institutions and companies. His research has been supported through collaborations with Google, Intel, and academic partners. He serves the community as a Senior Area Chair for CVPR 2025, Associate Editor for IEEE TPAMI (2024–2026), Guest Editor for IEEE TPAMI on 3D GenAI, and Program Chair for 3DV 2024. He leads a vibrant research lab at SFU focused on pushing the boundaries of 3D scene understanding with machine learning.
Chris Atkeson is a Professor at the Robotics Institute of Carnegie Mellon University. His research focuses on achieving human-level competence in machines through humanoid robotics and human-aware environments. He explores machine learning techniques such as reinforcement learning, nonparametric methods, and memory-based learning to develop robots capable of complex tasks like manipulation, locomotion, and perception. His work emphasizes bridging the gap between simulation and real-world applications (sim2real transfer), with contributions to tactile sensing (e.g., FingerVision), dynamic walking control, and human-robot collaboration. Notable projects include participation in the DARPA Robotics Challenge with Team WPI-CMU, where his team developed reliable humanoid behavior for disaster response scenarios. Atkeson’s research spans robotics, computer vision, and control systems, with a focus on enabling robots to perceive, learn, and act in unstructured environments. His recent work includes advancements in 3D scene capture, soft robotics, and energy-based planning for compositional tasks.