Prof. Ingmar Posner is a leading figure in applied artificial intelligence at the University of Oxford, where he serves as Principal Investigator for the Applied Artificial Intelligence Lab (A2I) and founding Director of the Oxford Robotics Institute. His work focuses on enabling robots to operate effectively in complex real-world environments through experience-driven learning. Key research areas: robot learning, scene interpretation, data-efficient learning, and transfer learning Applications in manipulation, autonomous driving, logistics, and space exploration His team has produced groundbreaking work in world models, sim-to-real transfer, and constraint-based manipulation systems (e.g., COMBO-Grasp). Notable contributions include the TWIST distillation framework and foundational research in tactile data generation (TactGen). He has received multiple best paper awards at top robotics venues. Publications reveal evolving research themes: 2025 work emphasizes language-conditioned learning (Lumos) and multi-agent decision-making, while 2024 focused on diffusion models for locomotion and differentiable simulators. Earlier work spans from urban scene analysis to physically plausible scene synthesis (RELATE).
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Massachusetts Institute of TechnologyUnited States
Luca Carlone is the Boeing Career Development Associate Professor in the Department of Aeronautics and Astronautics at MIT and a Principal Investigator at the Laboratory for Information & Decision Systems (LIDS) . He leads the SPARK Lab , focusing on developing certifiable perception algorithms for autonomous systems. PhD in Mechatronics (Polytechnic University of Turin, 2012) Research spans robotics, computer vision, and optimization Research Interests : Certifiable Perception algorithms for high-integrity systems High-level Perception (geometric, semantic, physical understanding) Efficient Perception methods for resource-constrained robots Scientific Contributions include: 2024 Outstanding Systems Paper Award (RSS) 2023 IEEE Transactions on Robotics King-Sun Fu Award 2021 NSF CAREER Award 2020 AIAA Advising Award 2019 Amazon Research Award Advising : Teaches graduate courses like Visual Navigation for Autonomous Vehicles and Robotics: Science and Systems . Collaborates with institutions including JPL, Caltech, and KAIST through the DARPA SubT Challenge.
Adriana Kovashka is an Associate Professor at the University of Pittsburgh , affiliated with the School of Computing and Information and serving as Department Chair . Her academic journey began with BA degrees in Computer Science and Media Studies from Pomona College (2008) and a PhD in Computer Science from The University of Texas at Austin (2014). Joined Pitt’s faculty in January 2015 NSF CAREER awardee (2021) Google Faculty Research Award recipient Dr. Kovashka’s research spans Computer Vision , Machine Learning , and Natural Language Processing , focusing on visual rhetoric, weak multimodal supervision, and domain adaptation. She pioneered techniques for analyzing political imagery, developing robust object detection frameworks, and exploring the intersection of visual and textual persuasion through large-scale annotated datasets. Her recent work emphasizes geographic diversity in vision-language systems, audio-visual fusion for domain generalization, and shape-texture bias mitigation in CNNs. Key publications include groundbreaking studies on symbolic reasoning, multimodal dialogue systems, and ethical AI applications in education. Scientific honors include: NSF CRII Award (2016) NSF CAREER Award (2021) Pitt CRDF Award (2016, 2018) Best Paper at ECV Workshop (2021) Google Faculty Research Award (2016, 2018) Dr. Kovashka actively mentors students in multimodal learning projects and collaborates with interdisciplinary teams on NSF-funded initiatives. She co-organizes workshops like the first CVPR workshop on advertisement understanding and leads research groups exploring human-AI co-learning systems.
Nima Fazeli is an Assistant Professor of Robotics at the University of Michigan (2020–Present), holding courtesy appointments in Computer Science & Engineering (CSE) and Mechanical Engineering. He directs the Manipulation and Machine Intelligence (MMint) Lab, focusing on enabling dexterous robotic manipulation through multimodal representation learning, tactile sensing, and model-based reasoning. His work integrates mechanics, perception, controls, and planning to achieve autonomous interaction with uncertain environments. Education: PhD, MIT (2019); MSc, University of Maryland (2014); BSc, Amirkabir University of Technology (2011) Research interests emphasize embodied intelligence , including visuo-tactile fusion, contact dynamics modeling, and cross-modal learning. Recent work explores tactile shadows, deformable object manipulation, and language-guided robot control. His research is supported by the NSF CAREER grant and National Robotics Initiative, with applications in manufacturing, assistive robotics, and space systems. Publications span topics like tactile sensing hardware (e.g., GelSlim 4.0), visuo-tactile implicit representations (ViTaSCOPE), and failure recovery policies (Racer). His team’s work has been featured in outlets like The New York Times and BBC. Key Awards: NSF CAREER Grant (2024) Teaching includes Introduction to Robotic Manipulation . Collaborations involve cross-disciplinary projects with mechanical, electrical, and biomedical engineering groups.
Ole Winther is Professor in High dimensional biological data analysis/Machine learning at the Department of Biology, University of Copenhagen and Professor in Data science and complexity at DTU Compute, Technical University of Denmark. He serves as CRO and co-founder of raffle.ai, CTO and co-founder of FindZebra, Head of ELLIS Unit Copenhagen, and co-PI of the Machine Learning for Life Science Center. His research spans Bioinformatics , Machine Learning , and AI for Science , focusing on applying deep learning to biological sequence analysis, latent variable models, and medical NLP. Winther's work develops predictive and generative models for bioinformatics, with significant contributions to protein localization tools (SignalP, DeepLoc, DeepTMHMM), single-cell genomics, and novel deep learning architectures like variational autoencoders and diffusion models. Analysis of Winther's recent publications (2023-2025) reveals a strong trend toward integrating protein language models with traditional bioinformatics approaches and applying diffusion models to scientific problems. His work bridges theoretical machine learning advancements with practical applications in biology and medicine, particularly in protein sequence analysis, medical search engines, and scientific simulation acceleration. Winther currently supervises a diverse research group including Panagiotis Antoniadis, Rachael M. DeVries, Jun Wang, Beatrix M. G. Nielsen, Felix G. Teufel, Irene R. Rodriguez, Anders Christensen, and Christopher Heje Grønbech. His former students have established successful careers at institutions including Google, Apple, and various startups, with notable alumni like Casper Sønderby (Google Brain) and Søren Sønderby (Apple). He leads significant research initiatives including the ELLIS Unit Copenhagen and the Machine Learning for Life Science Center, while maintaining active industry partnerships through his co-founded companies raffle.ai (enterprise search using NLP) and FindZebra (search engine for rare diseases). His teaching includes Deep Learning courses at both DTU (02456) and University of Copenhagen (NDAK24002U).
University of Illinois Urbana-ChampaignUnited States
Derek W Hoiem is a Professor in the Siebel School for Computing and Data Science at the University of Illinois Urbana-Champaign, where he has been a faculty member since 2009. His research focuses on computer vision and related areas, and he is also the co-founder and Chief Science Officer of Reconstruct, an AI-based construction technology company. His educational background includes: PhD in Robotics, Carnegie Mellon University (2007) Beckman Postdoctoral Fellowship (2008) Prof. Hoiem's research spans computer vision, with a focus on object recognition, scene understanding, and graphics. His work also extends to mobile robotics and 3D scene reconstruction. He has made significant contributions in areas such as visual recognition, 3D modeling, and the application of computer vision in construction monitoring. His recent publications (2023-2025) demonstrate a strong focus on advancing multimodal understanding, particularly in region-based representations, 3D vision, and neural radiance fields. There is a clear trend towards integrating language and vision, improving efficiency in neural networks, and applying computer vision to real-world problems such as construction progress monitoring. His scientific awards and honors are extensive and include: IEEE Fellow (2022) University Scholar (2022) Koendrink Prize (2022) Dean's Award for Excellence in Research, Associate Professor (2021) Campus Distinguished Promotion Award (2015) Best Paper Award: IEEE Winter Conference on Applications in Computer Vision (WACV) (2015) CW Gear Junior Faculty Award (2014) IEEE PAMI Young Researcher Award (2014) Dean's Award for Excellence in Research, Assistant Professor (2014) Sloan Research Fellowship (2013) Intel Early Career Faculty Honor Program Award (2012) NSF CAREER Award (2011) ACM Doctoral Dissertation Award, Honorable Mention (2008) Carnegie Mellon University SCS Distinguished Dissertation Award (2008) Best Paper Award: IEEE Computer Vision and Pattern Recognition (CVPR) (2006) Prof. Hoiem has secured significant research funding, including an NSF CAREER award and an Intel Early Career Faculty award. He is also actively involved in technology transfer, having co-founded Reconstruct where he serves as Chief Science Officer. His teaching excellence is reflected in multiple "List of Teachers Ranked as Excellent" awards spanning from 2010 to 2021. Prof. Hoiem leads a research group at UIUC focused on computer vision and 3D scene understanding. Additionally, he co-founded and serves as Chief Science Officer at Reconstruct, which develops AI-based solutions for construction monitoring.
Jiaxuan You is an Assistant Professor at the Siebel School of Computing and Data Science, University of Illinois at Urbana-Champaign, leading the U Lab focused on achieving Artificial General Intelligence (AGI) in digital environments. His research spans graph neural networks (GNNs), relational data, foundation models, and machine learning systems. PhD and MS in Computer Science from Stanford University (2021) Developed GraphGym and PyTorch Geometric (PyG) for graph learning Core member at Kumo AI (2021-2023) His research explores: Graph-enhanced LLMs: Integrating relational structures into foundation models AGI Development: Self-optimizing AI agents and tool utilization ML Systems: Scalable architectures and redundancy-free computation Interdisciplinary Applications: Financial networks, crop yield prediction, and metro systems Recent publications focus on temporal reasoning, multi-agent dynamics, and hybrid architectures for LLMs. He actively develops open-source tools like DBGYM and GraphRouter. Scientific recognition includes: JPMC PhD Fellowship Baidu Scholarship Best Student Paper at AAAI 2017 World Bank Big Data Innovation Challenge winner He mentors PhD and intern students, emphasizing machine learning systems expertise. His lab collaborates on AGI workshops (e.g., ICLR 2024) and industry projects.
Shubham Tulsiani is an Assistant Professor at Carnegie Mellon University's Robotics Institute, where he leads the Computer Vision group and the Physical Perception Lab. His research focuses on inferring physically and spatially grounded representations from perceptual inputs, with applications in 3D vision, robot manipulation, and neural scene reconstruction. He directs an active research group with multiple PhD and Master's students. Research interests center on 3D scene understanding , robot learning , and generative modeling , with specific emphasis on: self-supervised perception, neural rendering, multi-view geometry, manipulation from visual inputs, and physics-based reasoning. The lab develops methods that leverage physical world constraints as supervisory signals. Recent publications demonstrate strong focus on diffusion models for 3D tasks , sparse-view reconstruction , and robotic manipulation transfer . Key trends include neural inverse rendering, view synthesis from limited observations, and translating human interactions to robot actions. Awards include: Best Student Paper Award at CVPR 2015 Advising includes supervision of 5 PhD students, 4 MS students, and undergraduates. Lab alumni hold positions at Google, Stanford, Meta, and Princeton. The Physical Perception Lab collaborates with FAIR Pittsburgh and the CMU Computer Vision group.
Swiss Federal Institute of Technology in LausanneSwitzerland
Pierre Vandergheynst is a Full Professor at the Swiss Federal Institute of Technology Lausanne (EPFL) in the Department of Electrical Engineering, with a courtesy appointment in Computer and Communication Sciences. He serves as EPFL’s Vice-Provost for Education since 2015 and leads the Signal Processing Laboratory 2 (LTS2). His research spans harmonic analysis, sparse approximations, mathematical data processing, and applications in signal/image processing, computer vision, machine learning, and graph-based data analysis. PhD in Mathematical Physics (1998), Université catholique de Louvain Postdoctoral Researcher at EPFL (1998-2001) Assistant Professor at EPFL (2002-2007) His research explores geometry/symmetry in high-dimensional data, redundant dictionaries for dimensionality reduction, and computational harmonic analysis on manifolds. Recent work focuses on protein structure modeling, geometric deep learning, and graph-based signal processing. Key article trends include graph neural networks for protein analysis, geometric deep learning in neuroscience, and structured knowledge priors in neural models. His 2023-2025 publications emphasize interpretable AI, long-range dependencies in graphs, and molecular representation learning. Scientific Awards: IEEE Signal Processing Magazine Best Paper Award (2023) Signal Processing Society Best Paper Award (2022) Apple ARTS Award (2007) De Boelpaepe Prize, Royal Academy of Sciences of Belgium (2009-2010) He has supervised over 30 PhD theses and contributed to foundational work in graph signal processing, compressive sensing, and geometric deep learning. His lab develops tools for data science on non-Euclidean structures, with applications in medicine, astronomy, and wireless systems.
Sebastian Scherer is an Associate Research Professor at the Robotics Institute (RI), Carnegie Mellon University (CMU), where he leads cutting-edge research in autonomous aerial systems and robotics. His work focuses on enabling unmanned rotorcraft to operate safely and efficiently in cluttered, low-altitude, and extreme environments. Education: Ph.D. in Robotics, Carnegie Mellon University (2010) MS in Robotics, Carnegie Mellon University (2007) BS in Computer Science (Minor in Robotics), Carnegie Mellon University (2004) His research interests span robotics, artificial intelligence, autonomous navigation, obstacle avoidance, SLAM, visual-inertial odometry, energy infrastructure, and public policy . He has made seminal contributions to UAV autonomy, including the first obstacle avoidance for micro aerial vehicles in natural environments (2008) and the first automatic landing zone detection and landing on a full-size helicopter (2010). His recent publications (2023–2025) demonstrate a strong focus on resilient autonomy, multi-robot exploration, foundation models for robotics, and large-scale dataset development. His team has released key datasets like TartanGround , BETTY , and SubT-MRS , and simulation tools like Pegasus Simulator , indicating a systems-level approach to advancing real-world autonomy. The research trends emphasize self-supervised learning, robust perception, risk-aware planning, and multi-modal fusion for off-road and urban environments. Scientific Awards: Popular Science Best of What's New 2010 Award AIAA@Infotech Best Paper Runner-up Award (2010) Siebel Scholar Dr. Scherer has advised numerous students and leads a vibrant research group focused on high-impact robotics applications. He has secured significant grants related to UAV autonomy, energy infrastructure, and urban air mobility. His lab develops experimental infrastructure such as AIrTonomy for testing next-generation autonomous aerial vehicles. He is actively involved in advancing SLAM and localization in extreme environments, notably through participation in the DARPA Subterranean Challenge. His team develops large-scale datasets and benchmarking frameworks to push the boundaries of robustness and generalization in mobile robotics.
David Alvarez-Melis is an Assistant Professor of Computer Science at Harvard University's John A. Paulson School of Engineering and Applied Sciences (SEAS). He leads the Data-Centric Machine Learning (DCML) group and holds affiliations with the Kempner Institute, Harvard Data Science Initiative, and the Center for Research on Computation and Society. His research focuses on making machine learning more data-efficient and trustworthy, with applications in natural and medical sciences. He also serves as a researcher at Microsoft Research New England. Affiliations: SEAS, Kempner Institute, Harvard Data Science Initiative, CRCS Education: PhD in Computer Science (MIT), MS in Mathematics (NYU Courant), BSc in Applied Mathematics (ITAM) Research Interests: Optimal Transport, dataset distillation, interpretable AI, medical imaging, robustness, and large language models. His work bridges theory and applications, emphasizing geometric and probabilistic methods. Recent Trends in Publications: Focused on advancing optimal transport for data manipulation, distributional deep equilibrium models, and repurposing LLMs for specialized domains. Key themes include synthetic dataset generation, gradient flows in probability spaces, and robust interpretability frameworks. Awards: Aramont Fellowship, Dean’s Competitive Fund, Top Reviewer awards at major conferences (ICLR, NeurIPS, ICML). Grants: Supported by the Aramont Fund and Harvard’s Dean’s Fund. His lab advises students across Harvard and MIT, with notable contributions to medical imaging, NLP, and foundational ML theory. He actively mentors interns and fosters collaborations with industry and academia.
Haibin Ling is the SUNY Empire Innovation Professor in the Department of Computer Science at Stony Brook University, part of the College of Engineering and Applied Sciences. His research focuses on computer vision, medical image analysis, augmented reality, and AI applications in science. He holds a Ph.D. from the University of Maryland (2006) and prior degrees from Peking University. Previously, he worked at Temple University (2008–2019) and held roles at Siemens Corporate Research, UCLA, and Microsoft Research Asia. Professor Ling's work spans biomedical imaging, AI for science, and human-computer interaction. He leads the CV Lab and collaborates with the AI Institute at Stony Brook. Awards include the NSF CAREER Award (2014), Best Student Paper (ACM UIST 2003), and IEEE Fellow (2020). He serves on editorial boards for IEEE Trans. PAMI, Pattern Recognition, and CVIU, and chairs major conferences like CVPR. His research group includes over 50 students and alumni, with active projects in tracking benchmarks (LaSOT), Leafsnap, and medical imaging tools. Notable publications address OCTA flow estimation, backdoor attacks on vision models, and topology-guided medical learning. Collaborations involve institutions like Temple University and Stony Brook's Department of Applied Mathematics and Statistics.
Animesh Garg is an Assistant Professor at the School of Interactive Computing at Georgia Tech, where he leads the People, AI, and Robotics (PAIR) research group . He holds a Senior Researcher position at Nvidia Research and has courtesy appointments at the University of Toronto and Vector Institute. Previously, he served as Chief Scientific Officer at Apptronik (2024-2025) and Senior Staff Research Scientist at Nvidia Research (2018-2024). Education : Ph.D. in Operations Research from UC Berkeley (2011-2016), MS in Computer Science and Industrial Engineering from Georgia Tech and University of Delhi. Research Focus : Building Generalizable Autonomy through Reinforcement Learning , Control Theory , and 3D Vision , with applications in Surgical Robotics , Self-Driving Labs , and Manufacturing . Key Article Themes : His recent work emphasizes Foundation Models for robotics, Differentiable Simulation , Language-Guided Autonomy , and Structured Inductive Biases in sequential decision-making. Scientific Awards : Stephen Fleming Early Career Professorship at Georgia Tech. Teaching : Courses on AI, Deep Reinforcement Learning, and Algorithmic Intelligence in Robotics at Georgia Tech. Labs & Collaborations : Affiliated with Institute for Robotics and Intelligent Machines (IRIM) and ML@GT at Georgia Tech; collaborates intensively with Nvidia Robotics.
Adriana Schulz is an Assistant Professor in the Department of Computer Science & Engineering at the University of Washington's College of Engineering. She leads a research group focused on computational design, computer-aided design (CAD), and digital fabrication. Her work bridges computer science with practical applications in manufacturing, robotics, and sustainable design. Dr. Schulz received her Ph.D. in Computer Science from MIT in 2018 under the supervision of Professor Wojciech Matusik. Prior to her doctoral studies, she earned a Master's degree in Mathematics from IMPA (Instituto Nacional de Matemática Pura e Aplicada) in Rio de Janeiro, where she worked with Professor Luiz Velho, and a Bachelor's degree in Electronics Engineering from UFRJ (Federal University of Rio de Janeiro). Her research interests center around computational tools that enhance design and manufacturing processes. She develops novel algorithms for CAD systems, computational fabrication techniques, and sustainable design approaches. Her work spans multiple domains including robotics, textiles, electronics, and architecture, with a strong emphasis on creating practical tools that designers and engineers can use in real-world applications. She explores how machine learning, particularly neurosymbolic approaches, can improve design workflows and enable new capabilities in computational design systems. Analysis of her recent publications reveals a strong trend toward more intelligent and user-centered design tools. Her research increasingly integrates machine learning with traditional CAD systems to create more intuitive interfaces, supports sustainable design practices with computational tools, and develops novel fabrication techniques that push the boundaries of what's possible with digital manufacturing. She has made significant contributions to zero-waste fashion design, immersion cooling for high-performance computing, and CAD program understanding through novel representation learning techniques. Innovators Under 35 - MIT Technology Review Bolsa Aluno Nota 10 from FAPERJ Engineer 20000 award Dr. Schulz actively mentors several PhD students and postdoctoral researchers, including Haisen Zhao, Ben Jones, Yuxuan Mei, and others, often in collaboration with colleagues across different departments. Her research has attracted significant media attention, with coverage in major outlets including MIT News, BBC, IEEE Spectrum, Wired, and TechCrunch. Her work on Interactive Robogami was noted as the most read article in the International Journal of Robotics Research in its publication year. She leads a vibrant research group at the University of Washington that focuses on computational design systems, with particular emphasis on creating tools that bridge the gap between digital design and physical fabrication. Her team develops novel algorithms for CAD systems, computational fabrication techniques, and sustainable design approaches that have practical applications across multiple industries.