Haibin Ling is the SUNY Empire Innovation Professor in the Department of Computer Science at Stony Brook University, part of the College of Engineering and Applied Sciences. His research focuses on computer vision, medical image analysis, augmented reality, and AI applications in science. He holds a Ph.D. from the University of Maryland (2006) and prior degrees from Peking University. Previously, he worked at Temple University (2008–2019) and held roles at Siemens Corporate Research, UCLA, and Microsoft Research Asia. Professor Ling's work spans biomedical imaging, AI for science, and human-computer interaction. He leads the CV Lab and collaborates with the AI Institute at Stony Brook. Awards include the NSF CAREER Award (2014), Best Student Paper (ACM UIST 2003), and IEEE Fellow (2020). He serves on editorial boards for IEEE Trans. PAMI, Pattern Recognition, and CVIU, and chairs major conferences like CVPR. His research group includes over 50 students and alumni, with active projects in tracking benchmarks (LaSOT), Leafsnap, and medical imaging tools. Notable publications address OCTA flow estimation, backdoor attacks on vision models, and topology-guided medical learning. Collaborations involve institutions like Temple University and Stony Brook's Department of Applied Mathematics and Statistics.
Cyrill Stachniss is a Full Professor at the University of Bonn , where he heads the Lab for Photogrammetry and Robotics and is affiliated with the Lamarr Institute for Machine Learning and Artificial Intelligence . He was previously a Visiting Professor in Engineering at the University of Oxford until 2025. His academic journey includes positions at the University of Freiburg, University of Zaragoza, and the Swiss Federal Institute of Technology. University: University of Bonn School: Faculty of Engineering Department: Department of Photogrammetry Academic Rank: Professor His research spans robotics, photogrammetry, SLAM, autonomous navigation, perception systems, agricultural robotics, and unmanned aerial vehicles . He emphasizes probabilistic techniques for mobile robots and has made significant contributions to visual and LiDAR-based localization, scene understanding, and 3D reconstruction. The recent publications highlight a strong trend toward neural implicit representations, agricultural phenotyping, radar-based perception, and active learning . His team develops robust systems for real-world deployment in dynamic and unstructured environments, particularly in precision farming and autonomous vehicles. Scientific Awards: IEEE RAS Early Career Award (2013) Microsoft Research Faculty Fellow (2010) 7th EURON Georges Giralt Award (2008) Multiple Best Paper Awards at ICRA, IROS, RSS, and RAL Faculty Teaching Award, University of Freiburg He has advised numerous students and leads the DFG Cluster of Excellence PhenoRob and Research Unit FOR 1505 Mapping on Demand . His lab has co-founded three startups, reflecting strong industry and societal impact. He also runs the educational video series 5 Minutes with Cyrill , explaining key robotics concepts.
Andreas Geiger is a Professor and Head of the Department of Computer Science at the University of Tübingen, Germany. He leads the Autonomous Vision Group (AVG) within CyberValley and is a core faculty member of the Tübingen AI Center. His roles also include PI in the ML in Science Excellence Cluster and the CRC Robust Vision, as well as ELLIS Fellow and coordinator of the ELLIS PhD program. He specializes in machine learning models for computer vision, robotics, and autonomous systems, with applications in self-driving cars, VR/AR, and scientific document analysis. Educational background: While not explicitly detailed, his positions imply a Ph.D. in Computer Science or related field. His work spans interdisciplinary collaborations with institutions like ETH Zürich, Microsoft, and the University of Bonn. Research focuses on 3D scene understanding, Gaussian splatting, generative models, and reliable autonomous systems. Notable contributions include the KITTI dataset and foundational work in neural radiance fields. Awards include the Sage 10-Year Impact Award (2024), ERC Starting Grant (2019), and IEEE PAMI Young Researcher Award (2018). Key projects include the Scholar Inbox paper recommender platform, ReSim (reliable world simulation), and advancements in 3D scene generation (e.g., UrbanCAD, PrITTI). His lab maintains a strong focus on open-source tools and datasets, such as the CARLA Route Generator. Grants and funding include support from Vector Stiftung (MINT innovation program) and EU initiatives like the ML in Science Cluster. His team collaborates internationally, with recent work presented at CVPR, SIGGRAPH, and NeurIPS.
Angel Xuan Chang is an Associate Professor at Simon Fraser University's School of Computing Science, affiliated with labs including 3DLG, GrUVi, SFU NatLang, SFU AI/ML, and VINCI. He holds a Canada CIFAR AI Chair and was a TUM-IAS Hans Fischer Fellow (2018-2022). His research bridges natural language processing (NLP), 3D scene understanding, and embodied AI, focusing on language-grounded 3D generation and biodiversity monitoring via DNA barcodes. Recent work includes NuiScene (unbounded outdoor scene generation), ViGiL3D (3D visual grounding dataset), and CLIBD (vision-genomics biodiversity analysis). He advises students in projects like BIOSCAN-5M insect dataset and embodied AI navigation. His 2025 highlights include multiple ICCV and ICLR papers, workshops at ICML and CVPR, and a CRV invited talk. Education: Ph.D. in Computer Science from Stanford University (2014), advised by Chris Manning. Previous roles include visiting research scientist at Facebook AI Research and researcher at Eloquent Labs.
Freda Shi is an Assistant Professor at the David R. Cheriton School of Computer Science at the University of Waterloo and a Faculty Member at the Vector Institute, where she holds a Canada CIFAR AI Chair. She joined the University of Waterloo in July 2024 after completing her Ph.D. at the Toyota Technological Institute at Chicago. Educational Background: Ph.D. in Computer Science, Toyota Technological Institute at Chicago (2024), advised by Professors Karen Livescu and Kevin Gimpel Bachelor's degree in Intelligence Science and Technology (Computer Science Track) with a minor in Sociology, Peking University (2018) Dr. Shi's research focuses on computational linguistics and natural language processing, particularly on deeper understandings of natural language and the human language processing mechanism. She is especially interested in learning language through grounding, computational multilingualism, and related machine learning aspects. Her work aims to inform the design of more efficient, effective, safe, and trustworthy NLP systems. She leads the CompLING Lab at the University of Waterloo, which investigates how language models process spatial relationships and acquire linguistic structures through grounded experiences. Her publication record shows a consistent trajectory of high-impact research, with recent work focusing on spatial reasoning in vision-language models, multilingual chain-of-thought capabilities, and grounded language acquisition. She has published in top-tier conferences including ACL, EMNLP, ICLR, and NAACL, with several papers receiving notable recognition including Best Paper Nominee status at multiple venues. Her research bridges theoretical linguistics with practical NLP applications, demonstrating how linguistic insights can improve AI systems. Scientific Recognition: Canada CIFAR AI Chair (2024) Google Ph.D. Fellowship Thesis of Distinction for her doctoral work Multiple Best Paper Nominee awards at major NLP conferences Dr. Shi teaches CS 784: Computational Linguistics and CS 486/686: Introduction to Artificial Intelligence at the University of Waterloo. She actively contributes to the NLP research community through conference participation, program committee service, and collaborative projects. Her research has significant implications for creating more robust, human-like language understanding systems and advancing the field of grounded language learning in artificial intelligence.
Yi Fang is an Associate Professor of Computer Engineering and an affiliated Associate Professor of Computer Science at New York University Abu Dhabi (NYUAD), and a Global Network Associate Professor at NYU Tandon. He is a core faculty member in the Division of Engineering, specializing in Electrical and Computer Engineering. His research is centered at the intersection of Embodied AI, Robotics, and AI-driven assistive technologies, with strong support from agencies such as the US NSF, UAE ADEK, and ASPIRE. PhD, Purdue University Yi Fang's research interests span 3D Computer Vision, Multimedia Processing, Machine Learning, Deep Learning, and Embodied AI . He focuses on AI-driven perception, learning, and real-world applications, particularly in engineering, medicine, and accessibility. His lab, the Embodied AI and Robotics (AIR) Lab, develops intelligent robotic systems that integrate perception, learning, and decision-making to solve complex societal challenges. His work emphasizes large-scale visual computing, deep visual learning, and cross-domain/multimodal foundation models , with recent innovations in assistive AI for the Deaf and Hard-of-Hearing community. The 15 most recent publications reflect a consistent focus on 3D vision, sketch-based 3D retrieval, point cloud learning, and assistive computer vision . His work leverages deep learning, adversarial training, metric learning, and generative models to bridge modalities such as sketches, depth images, and 3D models. There is a clear trend toward cross-modal understanding, unsupervised representation learning, and real-world assistive applications , especially for visually impaired individuals. Yi Fang actively contributes to the academic community as an Area Chair for top-tier conferences including CVPR, ECCV, ICCV, IJCAI, and IROS. He also serves in peer review and mentoring roles, shaping the future of AI and robotics research. As a dedicated educator, he teaches foundational and advanced courses such as Computer Vision, Applied Machine Learning, Data Structures, and Capstone Design . He mentors students through research seminars and honors projects, fostering innovation and technical excellence. His research is supported by major grants from US NSF, UAE ADEK, and ASPIRE, enabling high-impact interdisciplinary collaborations. He founded and directs the Embodied AI and Robotics (AIR) Lab at NYU Abu Dhabi, a dedicated research space for developing intelligent systems that seamlessly integrate perception, learning, and decision-making. The lab promotes interdisciplinary collaboration across engineering, medicine, and social sciences, advancing the frontiers of Embodied AI.
Mark Yatskar is an Assistant Professor in the Department of Computer and Information Science at the University of Pennsylvania. His research focuses on the intersection of natural language processing, computer vision, and fairness in machine learning. He earned his PhD from the University of Washington under advisors Luke Zettlemoyer and Ali Farhadi, and previously worked as a Young Investigator at the Allen Institute for Artificial Intelligence. Education: PhD in Computer Science, University of Washington (Advisor: Luke Zettlemoyer & Ali Farhadi) Research Interests: Yatskar's work explores how language can structure visual perception and mitigate human biases in machine learning systems. Key themes include: Natural language as a scaffold for visual intelligence Bias characterization and control in machine learning systems His lab currently investigates projects like language-guided bottlenecks, annotator cognitive heuristics, and gender bias amplification. Teaching: CIS 5300: Computational Linguistics (2021-2024) CIS 7000: Language and Vision (2020) CIS 6300: Efficient NLP (2023, 2025) Awards: Best Paper Award at EMNLP (Gender Bias Amplification Research) Advising & Grants: Yatskar advises a team of PhD/Master's students and actively seeks motivated researchers. His group has explored funding in areas like interpretable AI, multimodal reasoning, and dataset bias mitigation. Labs/Teams: Leads the Penn NLP & Vision Lab, focusing on projects like MolMo/PixMo open models, ViUniT visual unit tests, and bias mitigation frameworks.
Fei-Fei Li is the Sequoia Capital Professor in Computer Science at Stanford University and Founding Co-Director of the Stanford Institute for Human-Centered AI (HAI). She holds courtesy appointments in the Graduate School of Business and is a Senior Fellow at HAI. Her work bridges AI research with interdisciplinary applications in healthcare, robotics, and policy. Dr. Li pioneered the ImageNet dataset, instrumental in the AI revolution, and co-founded World Labs to advance spatial intelligence and generative AI. Education: B.A. in Physics, Princeton University (1999) Ph.D. in Electrical Engineering, Caltech (2005) Doctorate (Honorary), Harvey Mudd College (2022) Research Interests: AI ethics, computer vision, robotic learning, healthcare applications, and human-AI collaboration. Her teams developed frameworks like MOMA for activity recognition and BEHAVIOR for embodied AI benchmarks. She advocates for diversity in tech and co-founded AI4All to mentor underrepresented students. Key Contributions: ImageNet and ImageNet Challenge Stanford Vision and Learning Lab (SVL) Policy advisory roles for U.S. Senate, UN Secretary-General, and California Governor Labs & Initiatives: Leads the People, AI & Robots Group (PAIR), Partnership in AI-Assisted Care (PAC), and the Human-Centered AI Institute. Her work emphasizes ethical AI deployment and societal impact. Awards: VinFuture Prize (2024), IEEE Fellow, National Academy memberships (Engineering, Medicine, Arts & Sciences), and recognition as one of Time’s AI100 Influencers.
Marc Pollefeys is a Full Professor at the Department of Computer Science, ETH Zurich, and Director of the Microsoft Mixed Reality and AI Lab. His work focuses on advanced perception systems for HoloLens, 3D computer vision, robotics, and machine learning. Key contributions include automated 3D modeling from video, real-time reconstruction pipelines, and vision-based autonomous systems. Education: PhD from KU Leuven (1999) Previous Affiliation: Professor at UNC Chapel Hill Research interests span 3D reconstruction , computer vision , robotics , SLAM , augmented reality , and privacy-preserving mapping . His work often integrates geometric modeling , feature matching , and deep learning . Recent projects emphasize implicit 3D representations , open-vocabulary scene understanding , and robust estimation using neural-guided algorithms. Recent publications highlight advancements in neural implicit fields , line-based correspondence , and vision-language integration . Trends include hybrid point-line methods, differentiable RANSAC, and privacy-aware localization frameworks. Scientific recognition includes: IEEE Fellow (2012) David Marr Prize (ICCV 1998) DAGM Best Paper Award (1999) Advisees include current and alumni PhD students such as Yagız Aksoy, Federico Camposeco, and Sudipta Sinha. Collaborations span institutions like UNC Chapel Hill, ETH Zurich, and Microsoft Zurich. Research sponsors include Microsoft, Google, and European research initiatives.
Nima Mesgarani is an Associate Professor of Electrical Engineering at Columbia Engineering, Columbia University, affiliated with the Sense, Collect and Move Data Committee. His research bridges engineering and neuroscience through reverse-engineering neural signal processing mechanisms, leading to advancements in brain-machine interfaces, neural prosthetics, and speech processing algorithms. He received his PhD in Electrical Engineering from the University of Maryland and completed postdoctoral training at Johns Hopkins University's Center for Language and Speech Processing and UC San Francisco's Neurosurgery Department. Research Focus Professor Mesgarani's lab integrates computational neuroscience and engineering to study acoustic signal processing. Key areas include: Neural decoding of speech and auditory attention in multi-talker environments Development of brain-controlled hearing technologies Novel speech separation and synthesis algorithms inspired by cortical processing Cross-modal learning between auditory and visual systems Applications of large language models in neural signal interpretation Publication Trends Analysis of his 15 most recent articles (2025) reveals dominant themes: neural decoding techniques using intracranial EEG, brain-inspired speech separation models (e.g., Mamba architectures), applications of large language models in auditory neuroscience, cross-modal distillation methods, and clinical translation of audio processing algorithms. A strong emphasis emerges on real-time brain-computer interfaces and noise-robust speech processing. Laboratory and Collaborations Mesgarani directs an interdisciplinary lab developing neurotechnology for hearing restoration. His team collaborates with neurosurgery departments and speech processing centers, focusing on translating theoretical models into clinical brain-machine interfaces. The lab's work has yielded patents for brain-informed speech separation systems and attention-decoding frameworks.
Joydeep Biswas is an Associate Professor in the Computer Science Department at the University of Texas at Austin, where he serves as the Director of the Autonomous Mobile Robotics Laboratory (AMRL). He is also affiliated with Texas Robotics, the UT Machine Learning Laboratory, and UT Good Systems. Previously, he was an Assistant Professor in the College of Information and Computer Sciences at the University of Massachusetts Amherst. Dr. Biswas earned his PhD in Robotics from Carnegie Mellon University in 2014 and his B.Tech in Engineering Physics from the Indian Institute of Technology Bombay in 2008. His educational background has provided him with a strong foundation in both theoretical and applied aspects of robotics and artificial intelligence. Dr. Biswas's research focuses on enabling long-term autonomy for mobile robots operating in human environments. His work spans robot perception, motion planning, control systems, and AI, with the ultimate goal of creating self-sufficient autonomous mobile robots that can perform tasks accurately and robustly in real-world settings. He is particularly interested in perception, planning, and failure recovery for autonomous mobile robots, which supports his vision of having autonomous service mobile robots deployed at campus-to-city scale, both indoors and outdoors, performing assistive tasks over deployments spanning years. His IJCAI 2019 Early Career Spotlight talk summarizes much of his research to date and ongoing interests. His recent research has shown a strong trend toward social navigation, human-robot interaction, and the application of machine learning techniques to robotics problems. There's a clear progression from fundamental robotics research toward more complex, real-world applications that require robots to understand and navigate human social spaces effectively. His work increasingly integrates large language models and other advanced AI techniques with traditional robotics approaches, as evidenced by his recent publications on topics like preference-conditioned navigation, social navigation benchmarks, and instruction-following navigation systems. Dr. Biswas has received numerous prestigious awards including the NSF CAREER Award (2021), J.P. Morgan Faculty Research Award (2019), Amazon Research Award (2019), and a grant from Northrop Grumman Mission Systems (2018). These awards recognize his innovative contributions to the field of robotics and autonomous systems. As a dedicated educator and mentor, Dr. Biswas actively supervises PhD and master's students, with his PhD student Sadegh Rabiee winning the student poster award at the Northrop Grumman University Symposium 2019. He has secured significant grant funding from the National Science Foundation for projects including 'Introspective Perception and Planning for Long-Term Autonomy' and 'Interactive Synthesis and Repair For Robot Programs,' demonstrating his ability to secure competitive research funding and his commitment to advancing the field. Dr. Biswas leads the Autonomous Mobile Robotics Laboratory (AMRL), which serves as a hub for interdisciplinary research in mobile robotics. The lab has developed notable resources such as the UT Campus Object Dataset (CODA) for 3D perception research and SOCIALGYM, a framework for benchmarking social robot navigation. His team regularly deploys robots on the UT Austin campus and in urban environments to test and refine their approaches in realistic settings, bridging the gap between simulation and real-world application.
Professor Gabriel Brostow is a faculty member in the Department of Computer Science at University College London (UCL), where he leads research in Computer Vision and Human-Computer Interaction. He also serves as Chief Research Scientist and Senior Director of the R&D Team at Niantic, the company behind Pokémon GO. His work bridges academic research and industry applications, focusing on developing AI systems that enhance human capabilities through what he terms 'Human in the Loop AI'—now commonly referred to as Human-Centered AI. Brostow completed his BS in Electrical Engineering at UT Austin, followed by a PhD with Irfan Essa at Georgia Tech. He then pursued postdoctoral research with Roberto Cipolla's Computer Vision & Robotics Group at Cambridge University as a Marshall Sherfield Fellow, and with Marc Pollefeys in ETH Zurich's CVG Group. His research explores how AI, particularly Computer Vision, can serve as 'super-tools' for professionals across various domains including filmmaking, architecture, robotics, and scientific research. Specific interests include assistive technology for everyday life, authoring systems that maximize user effort, 3D reconstruction, depth estimation, and vision-language models. His work often involves creating systems that are validated through real-world human interaction to ensure practical utility. Analysis of his recent publications reveals a strong focus on practical applications of Computer Vision that directly interact with humans. His research spans 3D scene understanding, depth estimation, sketch-based interfaces, and multimodal AI systems. There's a clear emphasis on creating benchmarks and tools that facilitate human-AI collaboration, with applications in assistive technology, urban planning, filmmaking, and biodiversity monitoring. His work frequently appears at top conferences including CVPR, NeurIPS, ECCV, and CHI. Marshall Sherfield Fellowship Brostow actively mentors PhD students, with current advisees including Ross Murphy, Skanda Koppula, Gizem Unlu, Omiros Pantazis, and Jamie Watson. His alumni include numerous PhD graduates and MSc students who have gone on to successful careers in academia and industry. He emphasizes selecting students based on passion and potential rather than just academic credentials, valuing traits like helpfulness, drive, and hunger to learn. His research is supported through collaborations with major institutions and companies including DeepMind, MIT, and the University of Edinburgh. He leads a research group at UCL that collaborates closely with Niantic's R&D team, creating a unique bridge between academic research and industry application. His team's work frequently involves developing novel Computer Vision techniques that are validated through real-world human interaction, ensuring practical utility alongside technical innovation. The group explores blue-sky research problems with applications ranging from assistive technology to professional tools for filmmakers, architects, and scientists studying diverse environments.
Dr. George Stamou is a Professor at the School of Electrical and Computer Engineering of the National Technical University of Athens (NTUA), serving as Director of the Artificial Intelligence and Learning Systems Laboratory (AILS). His expertise spans knowledge representation, machine learning, neural networks, and semantic technologies. He leads interdisciplinary initiatives such as the postgraduate program 'Data Science and Machine Learning' (2018–2022). Research Interests: Focuses on knowledge graphs, interpretable AI, semantic web applications, and multimodal learning. His work integrates formal logic systems (e.g., description logics) with modern deep learning techniques, addressing challenges in explainability, bias detection, and ethical AI applications. Publications: Over 150 articles in AI journals/conferences with an h-index of 34 (Google Scholar). Notable contributions include datasets like CHORDONOMICON (music analysis), GOSt-MT (gender bias in MT), and methodologies for counterfactual explanations in machine learning. Awards & Committees: Active in W3C and RuleML standardization bodies. Co-organized major AI conferences. Recognized for contributions to semantic interoperability and knowledge-based systems. Labs & Teams: Directs AILS-NTUA lab and collaborates with CISRI (Computer & Information Systems Research Institute). Engages in EU projects like CultureLabs (cultural heritage digitalization) andsmarty4covid (health data analysis).
Ellen Riloff serves as Department Head and Professor in the Department of Computer Science at the University of Arizona, where she leads research at the intersection of natural language processing (NLP) and artificial intelligence. Her work bridges theoretical advancements with real-world applications in social computing, planetary science, and crisis response systems. Education: Ph.D. in Computer Science, University of Massachusetts at Amherst (1994) Research Focus: Dr. Riloff specializes in affective computing and information extraction , developing techniques to recognize emotion, social cues, and embodied expressions in text. Her methodologies frequently employ bootstrapping, stacked learning, and semantic lexicon induction. Recent projects address crisis informatics (e.g., social cue recognition in emergencies) and interdisciplinary applications like the Mars Target Encyclopedia for planetary science data extraction. Publication Trends: Analysis of her 15 most recent publications (2021–2025) reveals three dominant trajectories: (1) affective event modeling in social contexts with applications to crisis response; (2) domain-specific NLP for planetary science and food systems; and (3) advanced language model techniques including retrieval-augmented generation and multi-view prompting. Her work increasingly integrates deep learning with traditional linguistic features. Grants and Leadership: Dr. Riloff has directed multiple NSF-funded projects, including RI: Small: Recognizing Implicit Personal States in Natural Language (2016) and RI: Small: Acquiring Domain Knowledge from Text through Cooperative Bootstrapping (2010). These initiatives pioneered bootstrapping frameworks for affective event recognition and information extraction. She also co-organized the Workshop on Pattern-based Approaches to NLP (2023), highlighting her leadership in advancing hybrid NLP methodologies. Collaborative Infrastructure: She co-developed the Mars Target Encyclopedia—a large-scale information extraction system that processes planetary science literature to create structured databases of Mars surface targets. This project demonstrates her commitment to building reusable scientific infrastructure through NLP.
Dylan Campbell is a Lecturer in Computing at the Australian National University (ANU), affiliated with the ANU College of Systems & Society. His research focuses on computer vision, optimization, and robotics, particularly in 3D vision and deep learning applications. He has held prior roles as a Research Fellow at the University of Oxford’s Visual Geometry Group and ANU’s Australian Centre for Robotic Vision. Campbell holds a PhD from ANU (2018) and a BE in Mechatronic Engineering from UNSW (2012). Research interests include geometric sensor alignment, neural radiance fields, and differentiable optimization layers. He actively supervises students (7 PhD/DPhil, 3 MEng, 9 honours) and teaches advanced courses in computer vision and robotics. Notable awards include the Marr Prize Honourable Mention (2017) and the IEEE Australia Council Postgraduate Student Paper Competition (2018). He has organized workshops at ECCV and CVPR, served as a reviewer for top conferences like CVPR/ICCV/ECCV, and contributed to datasets like SEED4D and RefRef. His work emphasizes efficient training of neural networks and leveraging symmetries in data for long-range connections.