Professor Haijiang Li is a Chair in BIM for Smart Engineering at Cardiff University's School of Engineering. His roles include leading the Computational Mechanics and Engineering AI Research Group, directing the BIM for Smart Engineering Centre, and overseeing the BIM MSc programme. He holds editorial roles for journals like Construction Innovation and Automation in Construction , and chairs the European Group of Intelligent Computing in Engineering (EG-ICE). Research focuses on smart computational engineering platforms integrating BIM, AI, and big data for sustainable infrastructure. Key areas include digital twins, disaster management, and resilient urban systems. He has secured £40M in research funding, including £9M as PI, and led over 70 research staff and students. Prof. Li is a Standards Committee Technical Executive at buildingSMART, driving international BIM standards. His work includes co-authoring a book on BIM standards across China, the US, and the UK. Awards include Fellowships from the British Computer Society (FBCS) and the Higher Education Academy (FHEA). His research outputs span over 250 publications, covering topics like AI-driven bridge maintenance, ontology-based decision-making, and energy-efficient urban systems. Collaborations with industry and global partners emphasize practical applications of BIM and smart technologies.
Professor Winston Hsu is a distinguished faculty member in the Department of Computer Science and Information Engineering at National Taiwan University, where he has served as a full professor since 2015. He is the co-director of the Communications and Multimedia Laboratory (CMLab) and founder of the MiRA (Multimedia indexing, Retrieval, and Analysis) research group. Additionally, he serves as the Founding Director for NVIDIA AI Lab at NTU, the first such lab in Asia. Professor Hsu received his Ph.D. in Electrical Engineering from Columbia University in 2007 under the supervision of Professor Shih-Fu Chang. Prior to his academic career, he was a founding engineer and research manager at CyberLink Corp., now a public image/video software company. National Taiwan University (2007-Present): Professor (2015-Present), Assistant/Associate Professor (2007-2015) MobileDrive (2021-2024): CTO and Vice President (Joint Venture between Foxconn and Stellantis) IBM TJ Watson Research Center (2016-2017): Visiting Scientist Microsoft Research Redmond (2014): Visiting Researcher Columbia University (2007): Ph.D. in Electrical Engineering Professor Hsu's research focuses on machine learning, computer vision, large-scale image and video search and recognition, and embedded AI. His work spans from fundamental research in visual recognition to practical applications in automotive systems, medical imaging, and e-commerce. He has pioneered work in disguised face recognition, low-resolution face hallucination, 3D model search, and virtual try-on systems. His current research emphasizes Embodied AI, integrating perception, action, and learning technologies for applications in automotive and robotics domains. His research group has produced numerous influential publications, particularly in top computer vision and multimedia conferences like CVPR, where they won first place in the Disguised Face Recognition competition in 2018. Their work spans diverse application areas including security, medical diagnostics, automotive systems, and e-commerce solutions, demonstrating strong translation from academic research to real-world impact. IBM Research Pat Goldberg Memorial Best Paper Award (2018) First Place, IARPA Disguised Faces in the Wild Competition (CVPR 2018) Best Brave New Idea Paper Award, ACM Multimedia 2017 NVIDIA AI LAB Award (First in Asia, 2016) First Place, MSR-Bing Image Retrieval Challenge (2013) World's Top 2% Scientists (2023) Professor Hsu actively mentors students and researchers, with his group consistently recruiting PhD students, postdocs, and research assistants. He has successfully bridged academia and industry through multiple collaborations, including his role as CTO at MobileDrive (a joint venture between Foxconn and Stellantis) from 2021-2024. His research has been supported by significant industry partnerships with Microsoft, IBM, and NVIDIA, as well as government grants from Taiwan's Ministry of Science and Technology. His laboratory, the Communications and Multimedia Laboratory (CMLab), maintains strong industry connections and focuses on cutting-edge research in visual AI. The lab has produced numerous award-winning projects and maintains active collaborations with global technology companies, particularly in the automotive and consumer electronics sectors.
Ari Holtzman is an Assistant Professor of Computer Science at the University of Chicago. His research spans dialogue systems, text generation, and foundational AI methodologies, including the development of Nucleus Sampling and contributions to the Amazon Alexa Prize. He holds an interdisciplinary degree from NYU in Computer Science and Philosophy of Language, and is nearing completion of his PhD at the University of Washington. Research interests include generative models, alignment challenges in LLMs, evaluation metrics like CLIPScore, and model efficiency techniques such as Qlora finetuning. His work bridges theoretical insights with practical applications, emphasizing both technical innovation and ethical considerations in AI. Key awards include the 2017 Amazon Alexa Prize and Phi Beta Kappa honors at NYU. His recent publications focus on benchmarking frameworks, cache optimization for large models, and understanding model limitations through AbsenceBench. Research contributions extend to multimodal systems, computational creativity, and machine unlearning protocols.
Julian McAuley is a Professor in the Department of Computer Science and Engineering at the University of California, San Diego's Jacobs School of Engineering. His research spans recommender systems, machine learning, natural language processing, music information retrieval, and multimodal learning. He maintains an active research group with numerous PhD students and postdocs working on cutting-edge AI problems. His research interests focus on developing advanced algorithms for personalized recommendation systems, with particular emphasis on sequential recommendation, multimodal learning, and integrating large language models with traditional recommendation approaches. His work bridges the gap between theoretical machine learning and practical applications across multiple domains including e-commerce, music, and healthcare. McAuley has published extensively in top-tier conferences including NeurIPS, ICML, KDD, SIGIR, and ACL, with his most recent work exploring the intersection of large language models and recommendation systems. His publications reveal a strong trend toward multimodal approaches that combine text, vision, and audio for more comprehensive understanding and recommendation. He has received significant research funding from major technology companies including Google, Amazon, Facebook, Adobe, and Samsung, as well as government agencies like the National Science Foundation and Department of Defense. His work has practical applications across multiple industries, with a focus on improving user experience through better personalization. McAuley advises numerous PhD students who have gone on to successful careers at leading technology companies and academic institutions. His former students include Wang-Cheng Kang and Jianmo Ni at Google DeepMind, Chris Donahue and Zachary Lipton as assistant professors at CMU, and Ruining He at Google Deepmind.
Raquel Fernández is Full Professor of Computational Linguistics and Dialogue Systems at the University of Amsterdam, where she leads the Dialogue Modelling Group at the Institute for Logic, Language & Computation (ILLC). As Vice-Director for Research at ILLC and a Fellow of the ELLIS Society, she bridges computational linguistics, cognitive science, and artificial intelligence through her research on language use in multimodal and conversational contexts. PhD in Computational Linguistics from King's College London Prior research positions at University of Potsdam and Stanford University's CSLI Her work explores how cognitive constraints, social interaction, and perception shape language use, with a focus on: Visually-grounded language processing Multimodal dialogue modeling Model uncertainty and calibration Language grounding in multimodal data Language learning and semantic change Dialogue reference resolution Recent publications analyze multimodal reasoning limitations, cross-lingual knowledge consistency, and uncertainty modeling in dialogue systems. She has received multiple accolades including an ERC Consolidator Grant , NWO VENI/VIDI/Aspasia fellowships , and EMNLP/GenBench awards . Outstanding Paper Award (EMNLP 2023) Best Data Award (GenBench Workshop 2023) ELLIS Society Fellow ERC Consolidator Grant #819455 recipient NWO VENI/VIDI/Aspasia awardee As a leader in academic service, she serves on the SIGDAT Executive Committee and chairs multiple conference committees. Her lab develops models for multimodal dialogue, visual storytelling, and grounded language understanding.
Bryan A. Plummer is an Assistant Professor in the Department of Computer Science at Boston University, affiliated with the IVC Group and the Artificial Intelligence Research (AIR) initiative at the Rafik B. Hariri Institute. He holds a PhD from the University of Illinois at Urbana-Champaign, specializing in computer vision. His research focuses on multimodal machine learning, efficient neural architectures, explainable AI, and robust ML systems. Plummer's work bridges vision and language, addressing challenges in domain generalization, synthetic data utilization, and model efficiency. Notable contributions include the Flickr30K Entities dataset and advancements in vision-language model robustness against web artifacts. He has advised over 20 students, with several securing roles at top institutions like NVIDIA and Google. His recent awards include the 3M Foundation Fellowship and NSF GRFP honorable mention. Plummer actively serves on conference committees (NeurIPS, CVPR, ICCV) and leads initiatives like the 1st Findings Workshop at ICCV'25.
Prof. Anya Belz is Full Professor of Computer Science at Dublin City University's School of Computing and Science Lead at ADAPT Research Centre. A leading NLP researcher with PhD-level expertise, she specializes in natural language generation, evaluation methodologies, and multimodal systems. Recipient of multiple best paper awards and NAACL Test of Time Award nomination. Research innovations include foundational work on statistical language generation (deployed in weather forecasting systems), comparative evaluation frameworks, vision-language integration, and reproducibility quantification. Current EPSRC-funded ReproHum project coordinates 20 global labs studying evaluation consistency. Achievements : Developed industry-deployed generation systems for accessibility applications Pioneered cross-modal alignment techniques for image description Authored 100+ publications spanning generation, evaluation, and reproducibility
Rita Cucchiara is a Full Professor at the Department of Engineering 'Enzo Ferrari' of the University of Modena and Reggio Emilia. She leads the AImageLab research laboratory, part of the Artificial Intelligence Research and Innovation Center (AIRI) in Modena. Her research focuses on Computer Vision, Pattern Recognition, Machine Learning, and Multimedia, with applications in video surveillance, medical imaging, human-centered AI, and generative models. She is actively involved in interdisciplinary projects like ELIAS (European Lighthouse for AI Sustainability) and ELSA (European Lighthouse on Secure AI). Her recent roles include being elected Rector of the University of Modena and Reggio Emilia in 2025. She has organized and participated in major AI events, including workshops at NeurIPS, CVPR, and ECCV, and has contributed to advancements in multimodal models, deepfake detection, and trustworthy AI. Education details are not explicitly provided, but her extensive academic and research experience at the University of Modena underscores her expertise. She collaborates with institutions like NVIDIA, CINECA, and industry partners such as Digital Design and NVIDIA's AI Technology Center. AImageLab's projects include developing systems for medical imaging, ethical AI, and generative adversarial networks (GANs) for design surfaces. She co-organizes initiatives like the ELLIS Summer School on Large-Scale AI and contributes to policy discussions on AI ethics and societal impact. Her work spans from foundational research (e.g., vision transformers, continual learning) to applied projects (e.g., DDGan system for surface printing). Key grants and collaborations include the FAIR project and PNRR-M4C2 initiatives. She advises students and researchers in AI, with 7 PhD positions funded under national programs. Her leadership roles in AIRI and AImageLab highlight her commitment to bridging academia and industry, fostering innovation in AI-driven solutions for sustainability and healthcare.
Adam Yala is an Assistant Professor of Computational Precision Health, Statistics, and Electrical Engineering and Computer Science at UC Berkeley and UCSF. He is also the Founder & CEO of Voio Inc., a company focused on clinical translation of AI tools. PhD in Computer Science from MIT (2022) His research lies at the intersection of Machine Learning and Precision Medicine, with a focus on robust AI tools for clinical deployment, personalized screening policies, and private data sharing. Current work includes multi-modal imaging analysis, decision guarantees in clinical workflows, and prospective trials in oncology and radiology. Recent publications highlight advancements in AI for cancer risk prediction, vision-language models in healthcare, and data privacy techniques. Tools like Mirai are implemented in 66 hospitals across 30 countries. Bakar Fellows Spark Award (2024) Eppy Award: Investigative Reporting (2022) Falling Walls Finalist: Life Science (2022) NSF Fellowship (2016) He advises PhD students in AI-driven healthcare and collaborates with hospital systems globally. His lab emphasizes clinical translation of machine learning methods in radiology and oncology.
Mani Golparvar Fard is a Professor at the University of Illinois at Urbana-Champaign, holding joint appointments in the Siebel School of Computing and Data Science and the Department of Civil and Environmental Engineering. He also contributes to the Technology Entrepreneur Center. His research focuses on integrating artificial intelligence, computer vision, and data analytics to advance construction management, infrastructure monitoring, and automation. Key areas include BIM integration, reality capture systems, and deep learning-based progress tracking. His work emphasizes automated construction progress monitoring through semantic segmentation, vision-language models, and UAV-based data collection. He has pioneered methods like Scan2BIM-NET for converting point clouds into BIM models and developed frameworks for worker safety analysis using machine learning. Awards: Walter L. Huber Civil Engineering Research Prize (2018) Daniel W. Halpin Award for Scholarship (2016) Advising & Grants: While no specific grant details are provided, his research is supported by collaborations with industry and government initiatives, such as the Japanese national bridge inspection project. He advises a team focused on AI-driven construction solutions and maintains active partnerships with engineering firms. Labs & Teams: Leads research groups in vision-based construction analytics, automated scheduling systems, and BIM integration. His work is disseminated through platforms like the VisualSiteDiary system and the InstaDam open-source platform for structural damage analysis.
Shih-Fu Chang is the Dean of Columbia Engineering and holds the Morris A. and Alma Schapiro Professorship at Columbia University. His research focuses on computer vision, machine learning, and multimedia information retrieval. He is recognized as a foundational figure in the field of content-based visual search and has pioneered innovations in image/video search engines, crime prevention systems, and brain-machine interfaces. His leadership roles include Chair of Columbia's Electrical Engineering Department (2007-2010), Editor-in-Chief of the IEEE Signal Processing Magazine (2006-2008), and Senior Executive Vice Dean at Columbia Engineering, where he drives strategic planning and international collaboration. Dr. Chang has received prestigious awards including the ACM Multimedia Technical Achievement Award, IEEE Signal Processing Technical Achievement Award, and IEEE Kiyo Tomiyasu Award. He is a Fellow of AAAS, ACM, and IEEE, and an Academician of Academia Sinica. His recent work emphasizes multimodal reasoning, few-shot learning, and vision-language systems, with applications in healthcare diagnostics and multimedia benchmarking. His research spans cross-modal understanding, event extraction, and adaptive AI systems. Key contributions include systems like Ferret-v2 for multimodal grounding and RESIN for schema-guided event tracking. He has advised multiple startups and actively contributes to curriculum development in AI and engineering education.
Paola Cascante-Bonilla is an Assistant Professor in the Department of Computer Science at Stony Brook University, with expertise in computer vision, natural language processing, and embodied AI. Her research focuses on developing systems for compositional reasoning, common-sense inference, and trustworthy AI using vision-language models, while addressing cultural bias and explainability challenges.
Rui Li is an Associate Professor in the Ph.D. program at Rochester Institute of Technology's Golisano College of Computing and Information Sciences. She directs the Lab for Use-inspired Computational Intelligence (LUCI), focusing on AI applications in computational biology and medical imaging. Education includes: B.Sc. in Computer Science, Harbin Institute of Technology M.Sc. in Computer Science, Tianjin University of Technology Ph.D. in Computing and Information Sciences, RIT Research integrates statistical machine learning with computational biology, medical image analysis, and human visual attention modeling. Current projects include deep learning for histopathology, multimodal medical image registration, and gene network inference. Publications demonstrate consistent focus on medical AI applications, with recent advances in unsupervised image registration, interactive segmentation, and multimodal fusion techniques. Key trends include self-supervised learning, uncertainty-aware models, and human-AI collaboration frameworks. Awards include the NSF CAREER Award for developing adaptive machine intelligence systems. Advises multiple PhD students on projects spanning deep learning architectures, biomedical image analysis, and biological network modeling. Leads several NSF-funded projects including human-centered image understanding systems and gene-protein network inference tools. Directs LUCI lab investigating machine learning for healthcare applications and teaches graduate courses in Statistical Machine Learning and Deep Learning.
Ruth Fong is a Teaching Professor at the Department of Computer Science, Princeton University , where she teaches foundational and advanced AI/ML courses (COS324, COS126) while leading the Looking Glass Lab in explainable AI research. She collaborates closely with the Visual AI Lab and Professor Olga Russakovsky . Education: PhD in Visual Geometry Group, University of Oxford (advised by Andrea Vedaldi , funded by Rhodes Trust and Open Philanthropy ) MSc in Neuroscience, University of Oxford (with Rafal Bogacz , Ben Willmore , and Nicol Harper ) AB in Computer Science, Harvard University (with David Cox and Walter Scheirer ) Research Focus: Pioneering Explainable AI and ML Fairness , with emphasis on post-hoc model understanding, interpretable-by-design architectures, and human-AI interaction frameworks. Her work spans computer vision, self-supervised learning, and neuroscience-inspired methodologies. Publication Trends: Recent papers (2023-2025) analyze interactive explanations , concept salience , and gender artifacts in vision datasets . Earlier work (2017-2020) established foundational techniques in extremal perturbations , backpropagation saliency , and neural network interpretability . Scientific Awards: Princeton Engineering Council Teaching Award (2025) Keller Center Summer Course Development Grant (2025) CHI Honorable Mention Paper Award (2023) Open Philanthropy AI Fellowship (2018) Rhodes Scholarship (2015) Advising: Directly mentored 10 Princeton undergraduates on IW/senior theses projects spanning generative AI , medical imaging fairness , and interactive visualization tools . Grants include Princeton SEAS and Open Philanthropy funding for the Looking Glass Lab. Lab & Team: Leads the Looking Glass Lab with 6 graduate/postgraduate members including Rawand Aziz , Matthew Barrett , and Ben Wachspress . Collaborates with faculty across Princeton and Oxford.
Elisa Kreiss is an Assistant Professor in the Department of Communication at UCLA, affiliated with the College of Letters & Science. She leads the Coalas Lab (Computation and Language for Society Lab), focusing on advancing understanding of how communicative context shapes language use through natural language processing, psycholinguistics, and human-computer interaction. Her work addresses challenges in image accessibility for visually impaired users, supported by grants from Google, NSF, and Stanford initiatives. Education: Ph.D. in Linguistics, Stanford University (advised by Christopher Potts) Research Interests: Her research bridges AI ethics, image accessibility, and human-centered evaluation of machine learning systems. Key themes include: Generating context-aware image descriptions for accessibility Ethical implications of vision-language models Interpretable AI through causal reasoning Awards: Google Research Awards National Science Foundation Grant Stanford Human-Centered AI Initiative Support Stanford Community Impact Award (2022) Lab & Advocacy: Directs the Coalas Lab, emphasizing inclusive research environments. Advocates for diversity in STEM and accessible technology design.