Ari Holtzman is an Assistant Professor of Computer Science at the University of Chicago. His research spans dialogue systems, text generation, and foundational AI methodologies, including the development of Nucleus Sampling and contributions to the Amazon Alexa Prize. He holds an interdisciplinary degree from NYU in Computer Science and Philosophy of Language, and is nearing completion of his PhD at the University of Washington. Research interests include generative models, alignment challenges in LLMs, evaluation metrics like CLIPScore, and model efficiency techniques such as Qlora finetuning. His work bridges theoretical insights with practical applications, emphasizing both technical innovation and ethical considerations in AI. Key awards include the 2017 Amazon Alexa Prize and Phi Beta Kappa honors at NYU. His recent publications focus on benchmarking frameworks, cache optimization for large models, and understanding model limitations through AbsenceBench. Research contributions extend to multimodal systems, computational creativity, and machine unlearning protocols.
Julian McAuley is a Professor in the Department of Computer Science and Engineering at the University of California, San Diego's Jacobs School of Engineering. His research spans recommender systems, machine learning, natural language processing, music information retrieval, and multimodal learning. He maintains an active research group with numerous PhD students and postdocs working on cutting-edge AI problems. His research interests focus on developing advanced algorithms for personalized recommendation systems, with particular emphasis on sequential recommendation, multimodal learning, and integrating large language models with traditional recommendation approaches. His work bridges the gap between theoretical machine learning and practical applications across multiple domains including e-commerce, music, and healthcare. McAuley has published extensively in top-tier conferences including NeurIPS, ICML, KDD, SIGIR, and ACL, with his most recent work exploring the intersection of large language models and recommendation systems. His publications reveal a strong trend toward multimodal approaches that combine text, vision, and audio for more comprehensive understanding and recommendation. He has received significant research funding from major technology companies including Google, Amazon, Facebook, Adobe, and Samsung, as well as government agencies like the National Science Foundation and Department of Defense. His work has practical applications across multiple industries, with a focus on improving user experience through better personalization. McAuley advises numerous PhD students who have gone on to successful careers at leading technology companies and academic institutions. His former students include Wang-Cheng Kang and Jianmo Ni at Google DeepMind, Chris Donahue and Zachary Lipton as assistant professors at CMU, and Ruining He at Google Deepmind.
William Yang Wang serves as the Mellichamp Professor of Artificial Intelligence at the University of California, Santa Barbara (2019-present). He directs the UCSB Center for Responsible Machine Learning, the Mind and Machine Intelligence Initiative, and the UCSB NLP Group. His research focuses on theoretical foundations and practical algorithms for AI, particularly in NLP, LLMs, and neuro-symbolic reasoning. PhD in Computer Science from Carnegie Mellon University Active in AI theory and applications (2016-present) Research interests span multiple AI domains, with special emphasis on NLP and responsible machine learning. He has pioneered datasets like HybridQA, TabFact, and VaTeX, enabling advancements in multi-hop QA, fact verification, and video-language tasks. His work combines statistical relational learning with modern deep learning paradigms. Recent publications center around multimodal reasoning, knowledge graph integration, and responsible AI development. He has received numerous accolades including the IEEE SPS Pierre-Simon Laplace Award (2024) and NSF CAREER Award (2021). Karen Sparck Jones Award (2022) DARPA Young Faculty Award (2018) IBM Faculty Award Mentoring 15+ PhD and postdoc researchers who now hold positions at Microsoft Research, Amazon, Meta GenAI, and academic institutions like Arizona and Rutgers. His lab maintains active collaborations with industry partners through initiatives like ChipAgents.ai, which he founded as CEO.
Raquel Fernández is Full Professor of Computational Linguistics and Dialogue Systems at the University of Amsterdam, where she leads the Dialogue Modelling Group at the Institute for Logic, Language & Computation (ILLC). As Vice-Director for Research at ILLC and a Fellow of the ELLIS Society, she bridges computational linguistics, cognitive science, and artificial intelligence through her research on language use in multimodal and conversational contexts. PhD in Computational Linguistics from King's College London Prior research positions at University of Potsdam and Stanford University's CSLI Her work explores how cognitive constraints, social interaction, and perception shape language use, with a focus on: Visually-grounded language processing Multimodal dialogue modeling Model uncertainty and calibration Language grounding in multimodal data Language learning and semantic change Dialogue reference resolution Recent publications analyze multimodal reasoning limitations, cross-lingual knowledge consistency, and uncertainty modeling in dialogue systems. She has received multiple accolades including an ERC Consolidator Grant , NWO VENI/VIDI/Aspasia fellowships , and EMNLP/GenBench awards . Outstanding Paper Award (EMNLP 2023) Best Data Award (GenBench Workshop 2023) ELLIS Society Fellow ERC Consolidator Grant #819455 recipient NWO VENI/VIDI/Aspasia awardee As a leader in academic service, she serves on the SIGDAT Executive Committee and chairs multiple conference committees. Her lab develops models for multimodal dialogue, visual storytelling, and grounded language understanding.
Teruko Mitamura is a prominent researcher at Carnegie Mellon University with over three decades of contributions to natural language processing, computational linguistics, and artificial intelligence. Her work spans from foundational research in event representation to advanced applications in multimodal systems and question answering. Her research interests focus on event detection and understanding, question answering systems, information retrieval, and multimodal processing. She has made significant contributions to event coreference resolution, timeline construction, and cross-document event analysis, developing methodologies that have become standard in the field. Her work often bridges theoretical advances with practical applications, particularly in complex information environments requiring deep semantic understanding. Natural Language Processing : Specializing in event extraction, coreference resolution, and narrative understanding with over 179 publications Question Answering Systems : Developing advanced techniques for complex question answering, particularly through NTCIR QA Lab and PoliInfo tasks Multimodal Processing : Integrating textual, visual, and temporal information for richer understanding in systems like ProMQA Evaluation Methodologies : Creating robust frameworks for assessing NLP systems through TAC KBP Event Tracks Her recent publication trends show a strong focus on leveraging large language models for event understanding, multimodal question answering, and timeline construction. She has expanded her research into specialized domains including patent analysis and novelty examination, demonstrating the breadth of her research impact across academic and practical applications. Active participant in major NLP conferences including ACL, EMNLP, NAACL, and AAAI with consistent publications Long-standing collaborator with researchers at CMU's Language Technologies Institute including Eduard H. Hovy and Eric Nyberg Contributor to shared tasks that have shaped research directions in event processing and question answering Organizer of multiple NTCIR QA Lab tasks focused on political information question answering Dr. Mitamura has mentored numerous researchers who have gone on to make their own contributions to the field, as evidenced by her extensive co-authorship network and the progression of her former students and collaborators into faculty and research positions. Her work continues to evolve with the field while maintaining her focus on deep semantic understanding of events and narratives.
Bryan A. Plummer is an Assistant Professor in the Department of Computer Science at Boston University, affiliated with the IVC Group and the Artificial Intelligence Research (AIR) initiative at the Rafik B. Hariri Institute. He holds a PhD from the University of Illinois at Urbana-Champaign, specializing in computer vision. His research focuses on multimodal machine learning, efficient neural architectures, explainable AI, and robust ML systems. Plummer's work bridges vision and language, addressing challenges in domain generalization, synthetic data utilization, and model efficiency. Notable contributions include the Flickr30K Entities dataset and advancements in vision-language model robustness against web artifacts. He has advised over 20 students, with several securing roles at top institutions like NVIDIA and Google. His recent awards include the 3M Foundation Fellowship and NSF GRFP honorable mention. Plummer actively serves on conference committees (NeurIPS, CVPR, ICCV) and leads initiatives like the 1st Findings Workshop at ICCV'25.
Prof. Anya Belz is Full Professor of Computer Science at Dublin City University's School of Computing and Science Lead at ADAPT Research Centre. A leading NLP researcher with PhD-level expertise, she specializes in natural language generation, evaluation methodologies, and multimodal systems. Recipient of multiple best paper awards and NAACL Test of Time Award nomination. Research innovations include foundational work on statistical language generation (deployed in weather forecasting systems), comparative evaluation frameworks, vision-language integration, and reproducibility quantification. Current EPSRC-funded ReproHum project coordinates 20 global labs studying evaluation consistency. Achievements : Developed industry-deployed generation systems for accessibility applications Pioneered cross-modal alignment techniques for image description Authored 100+ publications spanning generation, evaluation, and reproducibility
Rita Cucchiara is a Full Professor at the Department of Engineering 'Enzo Ferrari' of the University of Modena and Reggio Emilia. She leads the AImageLab research laboratory, part of the Artificial Intelligence Research and Innovation Center (AIRI) in Modena. Her research focuses on Computer Vision, Pattern Recognition, Machine Learning, and Multimedia, with applications in video surveillance, medical imaging, human-centered AI, and generative models. She is actively involved in interdisciplinary projects like ELIAS (European Lighthouse for AI Sustainability) and ELSA (European Lighthouse on Secure AI). Her recent roles include being elected Rector of the University of Modena and Reggio Emilia in 2025. She has organized and participated in major AI events, including workshops at NeurIPS, CVPR, and ECCV, and has contributed to advancements in multimodal models, deepfake detection, and trustworthy AI. Education details are not explicitly provided, but her extensive academic and research experience at the University of Modena underscores her expertise. She collaborates with institutions like NVIDIA, CINECA, and industry partners such as Digital Design and NVIDIA's AI Technology Center. AImageLab's projects include developing systems for medical imaging, ethical AI, and generative adversarial networks (GANs) for design surfaces. She co-organizes initiatives like the ELLIS Summer School on Large-Scale AI and contributes to policy discussions on AI ethics and societal impact. Her work spans from foundational research (e.g., vision transformers, continual learning) to applied projects (e.g., DDGan system for surface printing). Key grants and collaborations include the FAIR project and PNRR-M4C2 initiatives. She advises students and researchers in AI, with 7 PhD positions funded under national programs. Her leadership roles in AIRI and AImageLab highlight her commitment to bridging academia and industry, fostering innovation in AI-driven solutions for sustainability and healthcare.
Adam Yala is an Assistant Professor of Computational Precision Health, Statistics, and Electrical Engineering and Computer Science at UC Berkeley and UCSF. He is also the Founder & CEO of Voio Inc., a company focused on clinical translation of AI tools. PhD in Computer Science from MIT (2022) His research lies at the intersection of Machine Learning and Precision Medicine, with a focus on robust AI tools for clinical deployment, personalized screening policies, and private data sharing. Current work includes multi-modal imaging analysis, decision guarantees in clinical workflows, and prospective trials in oncology and radiology. Recent publications highlight advancements in AI for cancer risk prediction, vision-language models in healthcare, and data privacy techniques. Tools like Mirai are implemented in 66 hospitals across 30 countries. Bakar Fellows Spark Award (2024) Eppy Award: Investigative Reporting (2022) Falling Walls Finalist: Life Science (2022) NSF Fellowship (2016) He advises PhD students in AI-driven healthcare and collaborates with hospital systems globally. His lab emphasizes clinical translation of machine learning methods in radiology and oncology.
Mani Golparvar Fard is a Professor at the University of Illinois at Urbana-Champaign, holding joint appointments in the Siebel School of Computing and Data Science and the Department of Civil and Environmental Engineering. He also contributes to the Technology Entrepreneur Center. His research focuses on integrating artificial intelligence, computer vision, and data analytics to advance construction management, infrastructure monitoring, and automation. Key areas include BIM integration, reality capture systems, and deep learning-based progress tracking. His work emphasizes automated construction progress monitoring through semantic segmentation, vision-language models, and UAV-based data collection. He has pioneered methods like Scan2BIM-NET for converting point clouds into BIM models and developed frameworks for worker safety analysis using machine learning. Awards: Walter L. Huber Civil Engineering Research Prize (2018) Daniel W. Halpin Award for Scholarship (2016) Advising & Grants: While no specific grant details are provided, his research is supported by collaborations with industry and government initiatives, such as the Japanese national bridge inspection project. He advises a team focused on AI-driven construction solutions and maintains active partnerships with engineering firms. Labs & Teams: Leads research groups in vision-based construction analytics, automated scheduling systems, and BIM integration. His work is disseminated through platforms like the VisualSiteDiary system and the InstaDam open-source platform for structural damage analysis.
Shih-Fu Chang is the Dean of Columbia Engineering and holds the Morris A. and Alma Schapiro Professorship at Columbia University. His research focuses on computer vision, machine learning, and multimedia information retrieval. He is recognized as a foundational figure in the field of content-based visual search and has pioneered innovations in image/video search engines, crime prevention systems, and brain-machine interfaces. His leadership roles include Chair of Columbia's Electrical Engineering Department (2007-2010), Editor-in-Chief of the IEEE Signal Processing Magazine (2006-2008), and Senior Executive Vice Dean at Columbia Engineering, where he drives strategic planning and international collaboration. Dr. Chang has received prestigious awards including the ACM Multimedia Technical Achievement Award, IEEE Signal Processing Technical Achievement Award, and IEEE Kiyo Tomiyasu Award. He is a Fellow of AAAS, ACM, and IEEE, and an Academician of Academia Sinica. His recent work emphasizes multimodal reasoning, few-shot learning, and vision-language systems, with applications in healthcare diagnostics and multimedia benchmarking. His research spans cross-modal understanding, event extraction, and adaptive AI systems. Key contributions include systems like Ferret-v2 for multimodal grounding and RESIN for schema-guided event tracking. He has advised multiple startups and actively contributes to curriculum development in AI and engineering education.
Paola Cascante-Bonilla is an Assistant Professor in the Department of Computer Science at Stony Brook University, with expertise in computer vision, natural language processing, and embodied AI. Her research focuses on developing systems for compositional reasoning, common-sense inference, and trustworthy AI using vision-language models, while addressing cultural bias and explainability challenges.
Dima Damen is a Professor of Computer Vision at the School of Computer Science, University of Bristol, where she leads the Machine Learning and Computer Vision Group. She also serves as a Senior Research Scientist at Google DeepMind. As an EPSRC Early Career Fellow (2020-2025) and a Fellow of ELLIS for Europe, her research focuses on advancing computer vision, particularly in egocentric (first-person) vision, video understanding, and action recognition. Her educational background and professional journey have positioned her as a leader in the field of computer vision, with a particular emphasis on understanding human activities from wearable cameras. She has received numerous awards including Best Paper at ACCV 2024 and Outstanding Paper at ICASSP 2021. Professor Damen's research interests span multiple areas of computer vision and machine learning. She specializes in egocentric vision, where she has made significant contributions to understanding human activities from first-person perspectives. Her work explores video understanding, action recognition, hand-object interactions, and the development of vision-language models that can interpret and generate instructions from visual data. She has pioneered approaches to unique video captioning, long video understanding through active memory representations, and spatial reasoning from egocentric videos. Her research often bridges the gap between theoretical computer vision and practical applications in human-centered AI. Her recent publications demonstrate a strong focus on egocentric vision, with papers like "AMEGO: Active Memory from long EGOcentric videos" (ECCV 2024) and "HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision" (2024) advancing the state of the art in understanding long-form first-person videos. She has also contributed to vision-language models with works like "ShowHowTo: Generating Scene-Conditioned Step-by-Step Visual Instructions" (CVPR 2025) and "It's Just Another Day: Unique Video Captioning by Discriminitave Prompting" (ACCV 2024, Best Paper). Her research shows a consistent trajectory toward building systems that can understand human activities in natural environments with human-like capabilities. Professor Damen has received significant recognition for her work, including: EPSRC Early Career Fellow (2020-2025) ELLIS Fellow for Europe (Nov 2024) Best Paper at ACCV 2024 Outstanding Paper at ICASSP 2021 (awarded to only 3 out of 1700 papers) Outstanding Reviewer for CVPR 2020 and 2021 Program Chair for ICCV 2021 She has successfully advised numerous PhD students and postdoctoral researchers, many of whom have gone on to prestigious positions in academia and industry. Her group has secured significant research funding including the EPSRC Programme Grant Visual AI and the EPSRC UMPIRE grant. She actively collaborates with industry partners including Google DeepMind, Adobe, and Meta, ensuring her research has practical impact. Professor Damen leads the Machine Learning and Computer Vision Group at the University of Bristol, which focuses on egocentric vision, video understanding, and the development of vision-language models. The group has created influential datasets like EPIC-KITCHENS, which has become a standard benchmark in egocentric vision research. Her team regularly participates in and organizes workshops at major computer vision conferences including CVPR, ICCV, and ECCV.
Cecilia O. Alm is a Professor in the Department of Psychology within the College of Liberal Arts at Rochester Institute of Technology (RIT), where she serves as the Artificial Intelligence Program Director. She holds multiple leadership roles including Director of the Center for Human-aware AI and Director of the Computational Linguistics and Speech Processing Lab (CLaSP). Her institutional affiliations span the School of Information, Ph.D. Programs in Cognitive Science and Computing and Information Sciences, Department of Computer Science, and MS in Data Science program. Dr. Alm earned her Ph.D. from the University of Illinois at Urbana-Champaign. Her research focuses on human-centered artificial intelligence with particular emphasis on linguistic and multimodal sensing, affective computing, and natural language processing. She investigates how AI systems can better understand and respond to human communication through multimodal dialogue processing, with applications in accessibility, education, and healthcare. Her recent publications demonstrate a strong trend toward developing inclusive AI systems, particularly through projects addressing Deaf community needs (MULTICOLLAB-ASL), subtle emotion recognition (FUSE corpus), and bias mitigation in NLP. The work consistently integrates multimodal data streams (speech, gaze, gesture) to create more responsive human-AI interaction frameworks. Current research directions emphasize diversity in AI education, visual prosody in sign languages, and human-in-the-loop AI development. Dr. Alm leads several significant NSF-funded initiatives including the AWARE-AI program, IRES AI-PROWIL international research experience, and collaborative projects with Gallaudet University focused on Deaf scientist-centered AI research. She has secured over $2.5 million in external funding for her work on human-aware AI systems. She directs the CLaSP lab which provides research opportunities for PhD, MS, and undergraduate students, with graduates employed at major technology companies including Amazon, Apple, Microsoft, and Facebook. The lab focuses on real-world AI applications in accessibility, human-robot interaction, and multimodal communication systems.
Rui Li is an Associate Professor in the Ph.D. program at Rochester Institute of Technology's Golisano College of Computing and Information Sciences. She directs the Lab for Use-inspired Computational Intelligence (LUCI), focusing on AI applications in computational biology and medical imaging. Education includes: B.Sc. in Computer Science, Harbin Institute of Technology M.Sc. in Computer Science, Tianjin University of Technology Ph.D. in Computing and Information Sciences, RIT Research integrates statistical machine learning with computational biology, medical image analysis, and human visual attention modeling. Current projects include deep learning for histopathology, multimodal medical image registration, and gene network inference. Publications demonstrate consistent focus on medical AI applications, with recent advances in unsupervised image registration, interactive segmentation, and multimodal fusion techniques. Key trends include self-supervised learning, uncertainty-aware models, and human-AI collaboration frameworks. Awards include the NSF CAREER Award for developing adaptive machine intelligence systems. Advises multiple PhD students on projects spanning deep learning architectures, biomedical image analysis, and biological network modeling. Leads several NSF-funded projects including human-centered image understanding systems and gene-protein network inference tools. Directs LUCI lab investigating machine learning for healthcare applications and teaches graduate courses in Statistical Machine Learning and Deep Learning.