Rachel Rudinger is an Assistant Professor at the University of Maryland, affiliated with the Department of Computer Science and the University of Maryland Institute for Advanced Computer Studies (UMIACS). Her research focuses on Natural Language Processing (NLP), Machine Learning, and AI ethics, particularly addressing sociocultural biases and fairness in large language models (LLMs). She holds a PhD from Johns Hopkins University (2019) and a B.S. from Yale University (2013). Rudinger's work explores equitable cultural alignment in AI systems, common ground misalignment in dialog systems, and the mutual influence of gender and occupation in LLMs. She received the NSF CAREER Award in 2024 for her project on robust, fair, and culturally aware commonsense reasoning. Her recent publications investigate empathy gaps in LLMs, synthetic data effectiveness in disaster response, and bias measurement techniques across domains. As an advisor, she guides seven PhD students including Christabel Acquaye and Haozhe An. Her research spans diverse topics from legal language analysis to maternal health question answering, reflecting her commitment to interdisciplinary AI ethics. She actively contributes to workshops on commonsense representation and serves as a reviewer for top conferences in NLP and AI.
Lorraine (Xiang) Li is an Assistant Professor in the Department of Computer Science at the University of Pittsburgh’s School of Computing and Information (SCI). Her research focuses on the intersection of natural language processing, commonsense reasoning, knowledge representation, and machine learning, particularly in designing probabilistic models and evaluation methods for implicit commonsense knowledge in language. Li holds a PhD from the University of Massachusetts, Amherst, and previously worked as a young investigator with the Mosaic team at AI2. She has an M.S. in Computer Science from the University of Chicago, where she conducted research at TTIC. Her work emphasizes advancing AI’s ability to reason contextually and generate robust, human-like understanding through probabilistic frameworks. Key research themes include bias detection in reasoning models, iterative model editing, domain adaptation with LLMs, and evaluating commonsense through probabilistic measures. Her recent publications explore challenges like confirmation bias in chain-of-thought reasoning and geographical robustness in object recognition. Li actively contributes to the NLP community, serving on program committees for ACL, EMNLP, NAACL, and ARR. Though no formal awards are listed, her prolific publication record reflects her impact in AI research. She currently leads research in procedural knowledge models (e.g., Plasma) and long-tail knowledge generation, advancing foundational AI methodologies.
Raquel Fernández is Full Professor of Computational Linguistics and Dialogue Systems at the University of Amsterdam, where she leads the Dialogue Modelling Group at the Institute for Logic, Language & Computation (ILLC). As Vice-Director for Research at ILLC and a Fellow of the ELLIS Society, she bridges computational linguistics, cognitive science, and artificial intelligence through her research on language use in multimodal and conversational contexts. PhD in Computational Linguistics from King's College London Prior research positions at University of Potsdam and Stanford University's CSLI Her work explores how cognitive constraints, social interaction, and perception shape language use, with a focus on: Visually-grounded language processing Multimodal dialogue modeling Model uncertainty and calibration Language grounding in multimodal data Language learning and semantic change Dialogue reference resolution Recent publications analyze multimodal reasoning limitations, cross-lingual knowledge consistency, and uncertainty modeling in dialogue systems. She has received multiple accolades including an ERC Consolidator Grant , NWO VENI/VIDI/Aspasia fellowships , and EMNLP/GenBench awards . Outstanding Paper Award (EMNLP 2023) Best Data Award (GenBench Workshop 2023) ELLIS Society Fellow ERC Consolidator Grant #819455 recipient NWO VENI/VIDI/Aspasia awardee As a leader in academic service, she serves on the SIGDAT Executive Committee and chairs multiple conference committees. Her lab develops models for multimodal dialogue, visual storytelling, and grounded language understanding.
Teruko Mitamura is a prominent researcher at Carnegie Mellon University with over three decades of contributions to natural language processing, computational linguistics, and artificial intelligence. Her work spans from foundational research in event representation to advanced applications in multimodal systems and question answering. Her research interests focus on event detection and understanding, question answering systems, information retrieval, and multimodal processing. She has made significant contributions to event coreference resolution, timeline construction, and cross-document event analysis, developing methodologies that have become standard in the field. Her work often bridges theoretical advances with practical applications, particularly in complex information environments requiring deep semantic understanding. Natural Language Processing : Specializing in event extraction, coreference resolution, and narrative understanding with over 179 publications Question Answering Systems : Developing advanced techniques for complex question answering, particularly through NTCIR QA Lab and PoliInfo tasks Multimodal Processing : Integrating textual, visual, and temporal information for richer understanding in systems like ProMQA Evaluation Methodologies : Creating robust frameworks for assessing NLP systems through TAC KBP Event Tracks Her recent publication trends show a strong focus on leveraging large language models for event understanding, multimodal question answering, and timeline construction. She has expanded her research into specialized domains including patent analysis and novelty examination, demonstrating the breadth of her research impact across academic and practical applications. Active participant in major NLP conferences including ACL, EMNLP, NAACL, and AAAI with consistent publications Long-standing collaborator with researchers at CMU's Language Technologies Institute including Eduard H. Hovy and Eric Nyberg Contributor to shared tasks that have shaped research directions in event processing and question answering Organizer of multiple NTCIR QA Lab tasks focused on political information question answering Dr. Mitamura has mentored numerous researchers who have gone on to make their own contributions to the field, as evidenced by her extensive co-authorship network and the progression of her former students and collaborators into faculty and research positions. Her work continues to evolve with the field while maintaining her focus on deep semantic understanding of events and narratives.
Dylan Campbell is a Lecturer in Computing at the Australian National University (ANU), affiliated with the ANU College of Systems & Society. His research focuses on computer vision, optimization, and robotics, particularly in 3D vision and deep learning applications. He has held prior roles as a Research Fellow at the University of Oxford’s Visual Geometry Group and ANU’s Australian Centre for Robotic Vision. Campbell holds a PhD from ANU (2018) and a BE in Mechatronic Engineering from UNSW (2012). Research interests include geometric sensor alignment, neural radiance fields, and differentiable optimization layers. He actively supervises students (7 PhD/DPhil, 3 MEng, 9 honours) and teaches advanced courses in computer vision and robotics. Notable awards include the Marr Prize Honourable Mention (2017) and the IEEE Australia Council Postgraduate Student Paper Competition (2018). He has organized workshops at ECCV and CVPR, served as a reviewer for top conferences like CVPR/ICCV/ECCV, and contributed to datasets like SEED4D and RefRef. His work emphasizes efficient training of neural networks and leveraging symmetries in data for long-range connections.
Tiancheng Zhao is a principal researcher at the Binjiang Institute of Zhejiang University and founder of the Om Artificial Intelligence Laboratory (Om AI Lab), dedicated to frontier open multimodal AGI research for building next-generation agents that transform work and life through advanced human-machine interaction. His academic credentials include: Ph.D. in Computer Science from Carnegie Mellon University (2016-2019) under Prof. Maxine Eskenazi, Prof. Louis-Philippe Morency, Prof. William W. Cohen, and Dr. Dilek Hakkani-Tur, with pioneering dissertation “Learning to Converse With Latent Actions” in end-to-end generative conversational models M.S. in Computer Science from Carnegie Mellon University (2014-2016) B.S. in Electrical Engineering from UCLA (2010-2014) with Summa Cum Laude, focusing on speech signal processing under Prof. Abeer Alwan Dr. Zhao’s research centers on multimodal foundation models and agents, tackling three core challenges: Multimodal Models for cross-modal representation learning in high-dimensional data, Learning to Learn for effective skill acquisition from diverse signals (supervised labels, rewards, meta-learning), and AI Agents for open-world understanding and complex decision-making. His work bridges computer vision, natural language processing, and real-world applications including healthcare analytics and remote sensing. Analysis of his 50+ publications reveals accelerating innovation in multimodal large language models (2024-2025), with emphasis on stable vision-language architectures (VLM-R1), agent orchestration frameworks, and domain-specific applications in geospatial analysis and healthcare. Key trends include solving long-tail distribution challenges in satellite imagery, developing human-like zooming capabilities for multimodal LLMs, and creating unified benchmarks for autonomous GUI testing. His scientific recognition includes: National Breakthrough Technology Award by Ministry of Science and Technology (2021) Microsoft Research Best & Brightest PhD (2018) BEST PAPER AWARD at SIGDIAL 2018 Best Paper Nomination at SIGDIAL 2016 Top 1 Outstanding Bachelor of Science Award at UCLA (2014) As Om AI Lab founder, Dr. Zhao leads research teams developing computational building blocks for human-AI collaboration. While specific student mentorship details aren’t public, his extensive publication record with junior co-authors indicates active research supervision. Current projects focus on practical system implementations for real-world multimodal agent deployment across diverse domains.
Giorgio Grisetti is a Full Professor at Sapienza University of Rome within the Department of Systems and Computer Science, maintaining active research roles in the RoCoCo lab at Sapienza since November 2010 and the Autonomous Intelligent Systems Lab at Freiburg University where he previously served as a Post Doc under Wolfram Burgard starting in 2006. His educational background includes a M.Sc. in Computer Engineering from the University of Rome (2001) and a Ph.D. from Sapienza University of Rome's Intelligent Systems Lab (2006), supervised by Daniele Nardi. His doctoral thesis focused on SLAM using Rao-Blackwellized particle filters. Dr. Grisetti's research centers on mobile robotics with emphasis on robust solutions for autonomous navigation systems. His work spans theoretical and practical advancements in Simultaneous Localization and Mapping (SLAM), robot localization, path planning, and sensor fusion, particularly leveraging LiDAR and multi-sensor configurations. Recent publications demonstrate strong focus on optimization techniques, sensor calibration, and real-time performance for autonomous systems operating in complex environments. His publication trends reveal deep specialization in LiDAR-based SLAM (7 of 15 recent articles), bundle adjustment methods (4 articles), and sensor calibration/perception (3 articles), with consistent contributions to top robotics venues like IEEE Robotics and Automation Letters and ICRA. Key recognitions include: Nomination for the best IROS paper award (2010) Open Source achievement award from Willow Garage (2010) Best paper award at the International Conference and Exhibition on Unmanned Areal Vehicles (2010) Best Paper award at ICRA 2009 (2009) His research is conducted through the RoCoCo lab at Sapienza University of Rome and the Autonomous Intelligent Systems Lab at Freiburg University, focusing on developing foundational algorithms for mobile robot autonomy. Current projects emphasize robust perception systems, optimization frameworks for sensor fusion, and practical implementations for real-world navigation challenges.
Andrew Markham is a Professor of Computer Science at the University of Oxford , affiliated with Kellogg College . He leads a research group focusing on Cyber Physical Systems (CPS) , specializing in sensors, signal processing, and machine learning to enable machines to better perceive the physical world. His work emphasizes cross-disciplinary collaboration, notably in wildlife tracking and indoor positioning systems. He has held roles as a Postdoctoral Fellow (2008-2012), Associate Professor (2013), and Full Professor (2021). Education : PhD in Electrical Engineering (University of Cape Town, 2008), BSc (Hons) in Electrical Engineering (2004). Research Interests : Tracking and localization in GPS-denied environments (e.g., underground, indoors), magneto-inductive systems, physics-informed machine learning, and data-driven approaches for noisy sensor data. His projects include wildlife monitoring via wireless sensor networks and mmWave radar for human motion capture. Key Projects : CARACAL acoustic monitoring system, mmPoint dense human tracking, and RandLA-Net for large-scale point cloud segmentation. His work spans robotics, environmental sensing, and biomedical applications. Advising & Grants : Supervises over 30 students and collaborates with industrial partners. Research teams include Cyber Physical Systems, Autonomous Ubiquitous Sensing, and Wildlife Monitoring initiatives. Labs/Teams : Leads the CPS research group, focusing on sensor networks, inertial navigation, and multimodal fusion systems. Collaborates with zoology and earth science disciplines on applied projects.
Oswald Lanz is a tenured full professor at the Faculty of Engineering of the Free University of Bozen-Bolzano , leading the Visual Computing Lab . He holds a Ph.D. in Computer Science and a Mathematics degree from the University of Trento. Prior to his current role, he was a researcher and head of research at FBK Trento. He is an endowed professor collaborating with Covision Lab , an AI hub in Bressanone, and coordinates the board of professors for the PhD in Computer Science program since 2025. His research focuses on Computer Vision, Deep Learning, and Video Analytics , with applications in sports technology, medical imaging, and industrial automation. Key achievements include the Amazon AWS Machine Learning Research Award (2020) , ACM Multimedia Best Paper (2015) , and Best Student Paper at ICIAP (2007) . He co-organized the ELLIS-VISMAC Winter School (2025) and chaired ICIAP 2019 . His work spans novel view synthesis, action recognition, and anomaly detection, supported by patents in video tracking and detection. He teaches courses like Deep Learning and Artificial Intelligence in undergraduate and graduate programs. Recent projects such as 5VREAL integrate 5G, edge computing, and AI for sports analysis. His collaborations bridge academia and industry, exemplified by his role in Covision Lab and multidisciplinary initiatives like DSS4LCO for food supply chains. Lanz’s publications emphasize spatiotemporal modeling, neural architecture search, and hybrid machine vision systems.
Paola Cascante-Bonilla is an Assistant Professor in the Department of Computer Science at Stony Brook University, with expertise in computer vision, natural language processing, and embodied AI. Her research focuses on developing systems for compositional reasoning, common-sense inference, and trustworthy AI using vision-language models, while addressing cultural bias and explainability challenges.
Jiajun Wu is an Assistant Professor of Computer Science and, by courtesy, of Psychology at Stanford University. He holds multiple affiliations including membership in Bio-X, Faculty Affiliate status at the Institute for Human-Centered Artificial Intelligence (HAI), and membership in both the Wu Tsai Human Performance Alliance and Wu Tsai Neurosciences Institute. Dr. Wu earned his Ph.D. and S.M. in Electrical Engineering and Computer Science from the Massachusetts Institute of Technology before joining Stanford. Dr. Wu's research program focuses on creating AI systems that understand and interact with the physical world through the integration of computer vision, machine learning, robotics, and cognitive science. His work emphasizes physics-based modeling combined with deep learning to develop systems capable of perceiving, reasoning about, and predicting physical interactions. Key research areas include 3D scene understanding, neurosymbolic AI approaches, multimodal perception (combining vision, sound, and language), and embodied intelligence for robotics applications. His lab develops novel frameworks that bridge the gap between neural networks and symbolic reasoning to create more interpretable and robust AI systems. Analysis of Dr. Wu's recent publications reveals a strong trajectory toward integrated multimodal understanding for embodied AI. His work increasingly combines vision, sound, and language processing with physical reasoning to create systems that can interact meaningfully with the physical world. There's a clear progression from foundational computer vision research toward practical robotics applications, with significant emphasis on foundation models for robotics, sim2real transfer techniques, and creating comprehensive datasets for embodied AI research. Dr. Wu's exceptional contributions have been recognized with numerous prestigious awards including the NSF CAREER award (2024), Young Investigator Programs from ONR (2024) and AFOSR (2023), the Okawa research grant (2024), and being named to IEEE Intelligent Systems' 'AI's 10 to Watch' (2024). He has received multiple best paper awards at leading conferences including ICRA (2024), SIGGRAPH Asia (2023), and CoRL (2023). Dr. Wu actively mentors a large cohort of students across multiple levels, serving as primary advisor for doctoral candidates, master's students, and numerous independent researchers. His research is supported by substantial funding from major technology companies including Google, Meta, Amazon, Samsung, and J.P. Morgan, as well as government agencies like NSF, ONR, and AFOSR, reflecting the significance and impact of his work in physical AI and multimodal perception systems. Dr. Wu leads a dynamic research group at Stanford that collaborates extensively with the Wu Tsai Neurosciences Institute and Institute for Human-Centered AI. Current projects include developing neurosymbolic models for computer graphics, creating multisensory datasets like OBJECTFOLDER 2.0 for sim2real transfer in robotics, and building foundation models for embodied intelligence that can understand and manipulate objects with human-like physical intuition.
Dima Damen is a Professor of Computer Vision at the School of Computer Science, University of Bristol, where she leads the Machine Learning and Computer Vision Group. She also serves as a Senior Research Scientist at Google DeepMind. As an EPSRC Early Career Fellow (2020-2025) and a Fellow of ELLIS for Europe, her research focuses on advancing computer vision, particularly in egocentric (first-person) vision, video understanding, and action recognition. Her educational background and professional journey have positioned her as a leader in the field of computer vision, with a particular emphasis on understanding human activities from wearable cameras. She has received numerous awards including Best Paper at ACCV 2024 and Outstanding Paper at ICASSP 2021. Professor Damen's research interests span multiple areas of computer vision and machine learning. She specializes in egocentric vision, where she has made significant contributions to understanding human activities from first-person perspectives. Her work explores video understanding, action recognition, hand-object interactions, and the development of vision-language models that can interpret and generate instructions from visual data. She has pioneered approaches to unique video captioning, long video understanding through active memory representations, and spatial reasoning from egocentric videos. Her research often bridges the gap between theoretical computer vision and practical applications in human-centered AI. Her recent publications demonstrate a strong focus on egocentric vision, with papers like "AMEGO: Active Memory from long EGOcentric videos" (ECCV 2024) and "HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision" (2024) advancing the state of the art in understanding long-form first-person videos. She has also contributed to vision-language models with works like "ShowHowTo: Generating Scene-Conditioned Step-by-Step Visual Instructions" (CVPR 2025) and "It's Just Another Day: Unique Video Captioning by Discriminitave Prompting" (ACCV 2024, Best Paper). Her research shows a consistent trajectory toward building systems that can understand human activities in natural environments with human-like capabilities. Professor Damen has received significant recognition for her work, including: EPSRC Early Career Fellow (2020-2025) ELLIS Fellow for Europe (Nov 2024) Best Paper at ACCV 2024 Outstanding Paper at ICASSP 2021 (awarded to only 3 out of 1700 papers) Outstanding Reviewer for CVPR 2020 and 2021 Program Chair for ICCV 2021 She has successfully advised numerous PhD students and postdoctoral researchers, many of whom have gone on to prestigious positions in academia and industry. Her group has secured significant research funding including the EPSRC Programme Grant Visual AI and the EPSRC UMPIRE grant. She actively collaborates with industry partners including Google DeepMind, Adobe, and Meta, ensuring her research has practical impact. Professor Damen leads the Machine Learning and Computer Vision Group at the University of Bristol, which focuses on egocentric vision, video understanding, and the development of vision-language models. The group has created influential datasets like EPIC-KITCHENS, which has become a standard benchmark in egocentric vision research. Her team regularly participates in and organizes workshops at major computer vision conferences including CVPR, ICCV, and ECCV.
David J. Crandall is the Luddy Professor of Computer Science at Indiana University's Luddy School of Informatics, Computing, and Engineering. He serves as Director of the Luddy Artificial Intelligence Center and leads the IU Computer Vision Lab. With joint appointments in Informatics, Cognitive Science, Data Science, and Statistics, his work spans computer vision, machine learning, and AI. He holds a Ph.D. from Cornell University and previously worked at Eastman Kodak Research Labs. His research focuses on developing statistical and machine learning methods to analyze visual information, including object recognition, human activity analysis in video, 3D reconstruction, social media mining, and computational studies of visual attention. Key applications include egocentric vision systems, social robotics for healthcare, and cross-disciplinary collaborations with developmental psychology. Recent publications demonstrate strong emphasis on egocentric video analysis (Ego4D), human-robot interaction (CHI/HRI), and explainable AI (IJCAI). Medical imaging, nanoscale security systems, and computational social science represent emerging interdisciplinary directions. His work consistently integrates deep learning with real-world applications in health, environmental monitoring, and cultural analytics. Tracy M. Sonneborn Award (2024) Distinguished Member of the ACM (2023) Luddy Professorship (2021) NSF CAREER Grant (2013) Trustees Teaching Award (2017) He has advised over 20 Ph.D. graduates, with current students working on computer vision, robotics, and AI ethics. Major grants include $20M for the NSF AI Institute on Engaged Learning, $4.4M for trusted AI research, and funding from NIH, Google, ONR, and NASA. He directs the Computer Vision Lab and collaborates with Selma Sabanović's robotics group on social agents for older adults.
Dr. Cesar Dario Cadena Lerma is a Lecturer at the Department of Mechanical and Process Engineering and a tenured Senior Scientist at the Institute of Robotics and Intelligent Systems (IRIS) at ETH Zurich. He leads the Perception, Mapping and Navigation team within the Robotics Systems Lab (RSL), co-founded and directs the ETH RobotX initiative focusing on educational robotics, and previously held roles at ETH Zurich's Autonomous Systems Lab, University of Adelaide, and George Mason University. His research focuses on robotics perception, particularly in SLAM (Simultaneous Localization and Mapping), semantic scene understanding, and robust perception systems for dynamic environments. Education: PhD in Computer Science and System Engineering from the University of Zaragoza, followed by postdoctoral research at George Mason University and The University of Adelaide. Professional roles include managing director of ETH RobotX and leadership in multi-modal mapping frameworks like maplab 2.0. Research interests emphasize integrating perception and learning in robotics, with a focus on semantic mapping, data association, place recognition, and navigation in unstructured environments. His work bridges traditional SLAM techniques with modern deep learning approaches to create robust, modular systems. Key contributions include the PHASER registration algorithm, SCIM obstacle avoidance framework, and C-Blox dense mapping system. Awards include the Best Paper Award at the 2017 IEEE International Symposium on Safety, Security, and Rescue Robotics. His articles span topics like semantic pointcloud filtering, volumetric mapping, and embodied domain adaptation. He collaborates widely, with over 50 peer-reviewed publications in top venues such as IEEE Robotics and Automation Letters and International Journal of Robotics Research.
Matthieu Cord is a Professor at Sorbonne University and Scientific Director of valeo.ai, leading research in computer vision, deep learning, and computational cooking. He heads the MLIA team at ISIR Lab, focusing on multimodal models, transformers, and efficient architectures. Research areas include computer vision, large language models with vision, and AI-driven food analytics. Key projects: VISA-DEEP AI chair, Foundation VaViM models, and SmolVLA collaboration with Hugging Face. His recent work examines scalable multimodal models , trajectory prediction , and diffusion-based segmentation , with studies on in-context learning and biased shortcut learning in visual question answering. Articles highlight DeiT variants , fishr for OoD generalization , and STEEX for counterfactual explanations . Scientific awards include IUF Honorary Membership (2009), BMVC 2017 Best Paper, and ICIP 2018 Best Paper. As an advisor, he supervised PhD theses on topics like GAN editing , semantic segmentation , and multimodal retrieval . Current roles involve mentoring the 'Research Band' at MLIA and leading EU-funded initiatives like SCAPE. His work bridges theoretical AI exploration with practical applications in autonomous driving and food technology.