Yonatan Bisk is an Assistant Professor at Carnegie Mellon University (CMU) in the School of Computer Science , with dual appointments in the Language Technologies Institute and Robotics Institute . His research bridges Natural Language Processing (NLP) with robotics, focusing on grounded and embodied language understanding. Assistant Professor, Language Technologies Institute, CMU (2021–Present) Courtesy Appointment, Robotics Institute, CMU Research Themes : Language as a social codification of embodied experience Interpretable multimodal model training Human-robot collaboration frameworks Embodied question-answering systems Selected Trends : His recent publications show increasing focus on cross-modal attention mechanisms (Vid2Robot), error detection in toolchains (Tools Fail), and theory-of-mind reasoning in language agents (SOTOPIA). Multimodal integration spans vision, audio, and robotic control contexts (ANAVI). Labs & Collaborations : Founder of CLAW Lab (Connecting Language to Action and the World) Collaborations with Microsoft Research, Meta Inc, and CMU's REAL (Robotics, Embodied AI, Learning) community
James Glass is a Senior Research Scientist at the Massachusetts Institute of Technology (MIT) and heads the Spoken Language Systems Group within MIT's Computer Science and Artificial Intelligence Laboratory (CSAIL). He is also affiliated with the Harvard-MIT Division of Health Sciences and Technology. His research spans automatic speech recognition, multimodal learning, and spoken language understanding, with applications in healthcare and video analysis. Education: SM and PhD in Electrical Engineering and Computer Science from MIT His work focuses on paralinguistic speech analysis, health markers in speech, and the intersection of speech and natural language processing. Recent trends emphasize audio-visual alignment, recursive reasoning, and AI applications in cognitive disorder diagnosis. Scientific awards include IEEE Fellow, ISCA Fellow, and Associate Editor for IEEE Transactions on Pattern Analysis and Machine Intelligence. His group explores unsupervised learning, speaker verification, and social text analysis. James leads the Spoken Language Systems Group at CSAIL, collaborating with institutions like IBM and Harvard-MIT Division of Health Sciences and Technology. His research integrates vision-language models, neural audio codecs, and self-supervised frameworks.
Mark Lee is an Adjunct Professor in the People Analytics department at NYU’s Tandon School of Engineering, specializing in Technology Management and Innovation. He holds a Ph.D. in Engineering Psychology from Georgia Institute of Technology (1996). Currently, he serves as Head of Research, Analytics, and Business Development at UL ComplianceWire, focusing on pharmaceutical and medical device manufacturing training. His research leverages large datasets to improve healthcare safety through regulatory compliance and best practices. Courses taught include Human Factors Engineering, Workplace Design, and Predictive Analytics. Education: Ph.D. in Engineering Psychology, Georgia Tech (1996) Key Roles: Adjunct Professor, Head of Research at UL ComplianceWire Research Focus: Human Factors, Training Systems Design, Healthcare Compliance His work spans auditory display systems for aviation (e.g., 3D audio cockpit interfaces) and ergonomic design for industrial products. Recent projects emphasize data-driven solutions for regulatory challenges in life sciences. Publications highlight studies on visual search strategies, age-related cognitive performance, and application of signal detection theory in decision-making. He actively collaborates with industry and government entities, exemplified by the FDA-UL Cooperative Research Agreement.
Toby Jia-Jun Li is an Assistant Professor in the Department of Computer Science and Engineering at the University of Notre Dame, where he leads the SaNDwich Lab. He also serves as the Director of the Human-Centered Responsible AI Lab in the Lucy Family Institute for Data & Society and is a Faculty Fellow at the Institute for Educational Initiatives (IEI). Previously, he was affiliated with Carnegie Mellon University's Human-Computer Interaction Institute (HCII) and GroupLens Research. Dr. Li's research spans the intersection of Human-Computer Interaction (HCI), End-User Software Engineering, Machine Learning (ML), and Natural Language Processing (NLP), with recent work focusing on addressing societal challenges in the future of work through human-AI collaborative approaches. His work has resulted in over 40 publications at premier venues including CHI, UIST, CSCW, ACL, and ICSE, with 8 papers winning Best Paper or Honorable Mention awards. His recent publications demonstrate a strong focus on human-AI collaboration across various domains, including code understanding, privacy, accessibility, and creative tools. The work shows a trajectory toward increasingly sophisticated integration of human-centered design with AI capabilities, particularly using large language models to enhance human productivity and address societal challenges. Google Research Scholar Award recipient Recipient of Yahoo! Fellowship ($100,000/year) Best Paper Award at UIST 2020 Best Paper Honorable Mention Award at CHI 2021 Best Paper Award at CSCW 2024 Best Paper Award at CHI 2025 Dr. Li actively mentors Ph.D. students and has established collaborations with Google, Microsoft Research, IBM Research, Adobe, Verizon, and J.P. Morgan. His research has been supported by NSF, Google Research Scholar Program, AnalytiXIN Initiative, Yahoo! InMind project, and J.P. Morgan. He is currently recruiting Ph.D. students and undergraduate researchers for his SaNDwich Lab, which focuses on developing interactive systems to empower individuals to create, configure, and extend AI-powered computing systems.
Rita Cucchiara is a Full Professor at the Department of Engineering 'Enzo Ferrari' of the University of Modena and Reggio Emilia. She leads the AImageLab research laboratory, part of the Artificial Intelligence Research and Innovation Center (AIRI) in Modena. Her research focuses on Computer Vision, Pattern Recognition, Machine Learning, and Multimedia, with applications in video surveillance, medical imaging, human-centered AI, and generative models. She is actively involved in interdisciplinary projects like ELIAS (European Lighthouse for AI Sustainability) and ELSA (European Lighthouse on Secure AI). Her recent roles include being elected Rector of the University of Modena and Reggio Emilia in 2025. She has organized and participated in major AI events, including workshops at NeurIPS, CVPR, and ECCV, and has contributed to advancements in multimodal models, deepfake detection, and trustworthy AI. Education details are not explicitly provided, but her extensive academic and research experience at the University of Modena underscores her expertise. She collaborates with institutions like NVIDIA, CINECA, and industry partners such as Digital Design and NVIDIA's AI Technology Center. AImageLab's projects include developing systems for medical imaging, ethical AI, and generative adversarial networks (GANs) for design surfaces. She co-organizes initiatives like the ELLIS Summer School on Large-Scale AI and contributes to policy discussions on AI ethics and societal impact. Her work spans from foundational research (e.g., vision transformers, continual learning) to applied projects (e.g., DDGan system for surface printing). Key grants and collaborations include the FAIR project and PNRR-M4C2 initiatives. She advises students and researchers in AI, with 7 PhD positions funded under national programs. Her leadership roles in AIRI and AImageLab highlight her commitment to bridging academia and industry, fostering innovation in AI-driven solutions for sustainability and healthcare.
Dr A. I. Shihab is a Senior Lecturer at Kingston University's Faculty of Engineering, Computing and the Environment, Department of Networks and Digital Media. He teaches programming languages (C++/Java), data structures, web development, and AI/machine learning. His research focuses on affective computing and machine learning applications including: Acoustic event detection in sports environments Audio signal analysis for tennis match modeling Multi-camera visual surveillance systems Medical imaging analysis using fuzzy clustering techniques Publications demonstrate expertise in combining audio/video modalities for sports analytics (tennis rallies, court-shots) and developing Markov models for sound event sequence analysis. Contact: a.shihab@kingston.ac.uk
Ira Kemelmacher-Shlizerman is a Full Professor of Computer Science at the Paul G. Allen School of Computer Science & Engineering at the University of Washington and Director of the UW Reality Lab. She also serves as a Principal Scientist at Google, where she leads the Shopping Gen AI visuals teams focusing on Virtual Try-On, 3D, and product videos. Her research spans computer vision, computer graphics, and Generative AI, with particular contributions to virtual try-on technology, 3D modeling, and augmented reality applications. Professor Kemelmacher-Shlizerman's research interests focus on Generative AI applications in visual computing. Her work bridges the gap between theoretical computer vision and practical applications, particularly in e-commerce and virtual reality. She has made significant contributions to virtual try-on technology, 3D editing with generative models, and AI applications for shopping experiences. Her research combines deep learning with traditional computer vision techniques to solve challenging problems in image and video synthesis. Her recent publications demonstrate a strong trend toward Generative AI applications for visual shopping experiences, virtual try-on technology, and 3D content creation. The work spans multiple top conferences including CVPR, SIGGRAPH, and ICCV, with a focus on practical applications of computer vision and graphics. Her research has evolved from foundational work in face reconstruction and aging to current applications in virtual shopping and 3D content generation. Google faculty award Madrona prize GeekWire Innovation of the Year Award Covers of CACM and SIGGRAPH Best student paper honorable mention at CVPR'21 Best demo runner up MobiSys'22 Senior member of IEEE Distinguished Member of ACM Professor Kemelmacher-Shlizerman has successfully tech-transferred multiple research projects to industry. She founded Dreambit, a startup acquired by Meta, and previously built and launched the Face Movies feature at Google. She currently leads Google's Shopping Gen AI visuals teams, focusing on 10x improvements to shopping journeys. Her UW Reality Lab serves as a hub for AR/VR research with industry partnerships. She has mentored numerous PhD students who have become researchers in both academia and industry, with several publications featuring student co-authors receiving recognition at top conferences. Professor Kemelmacher-Shlizerman leads the Graphics and Imaging Laboratory (GRAIL) and the UW Reality Lab, which focuses on augmented and virtual reality research with industry partnerships including Google. The labs work on cutting-edge projects in virtual try-on, 3D modeling, and immersive experiences, bridging academic research with real-world applications.
Michael J. Black is a Professor and Director at the Max Planck Institute for Intelligent Systems in Tübingen, Germany, where he leads the Perceiving Systems department and serves as Managing Director . He is also an Honorarprofessor at the University of Tübingen 's Faculty of Science . His career spans roles at Brown University (2000-2010), Xerox PARC, and academic-industry collaborations with Amazon and Meshcapade.
Nils Holzenberger is an Assistant Professor at Télécom Paris, France, since February 2023, affiliated with the Data Intelligence Graphs (DIG) research team within the Information Processing and Communication Laboratory (LTcI). His work bridges artificial intelligence, natural language processing, and legal domains through neuro-symbolic approaches to statutory reasoning, particularly in tax law. Education: PhD in Computer Science, Johns Hopkins University (2017-2022) Master's in Engineering, Mines ParisTech (2013-2017) Preparatory Classes, Lycée Louis-le-Grand (2011-2013) Holzenberger's research centers on legal artificial intelligence with emphasis on statutory reasoning limitations in large language models. He pioneered the SARA dataset for tax law reasoning and LegalBench benchmark, developing hybrid symbolic-neural frameworks that expose LLMs' shortcomings in precise legal interpretation. His work integrates Prolog solvers with NLP techniques to create executable tax code mappings and contract analysis tools, establishing foundational methods for verifiable legal AI systems. Analysis of his 15 most recent publications (2019-2024) reveals three dominant research thrusts: (1) Tax law reasoning benchmarks exposing LLM hallucinations, (2) Neuro-symbolic integration for statutory interpretation, and (3) Low-resource template extraction for legal documents. His work consistently demonstrates that pure neural approaches fail at precise legal reasoning, necessitating symbolic grounding for reliable legal AI applications. The DIG research team at Télécom Paris, where Holzenberger leads legal AI initiatives, is actively hiring faculty for neuro-symbolic projects. While specific grant details aren't public, his collaborations with HEC Paris, Copilex startup, and featured podcast appearances indicate substantial industry-academia engagement in legal tech development. Holzenberger directs the legal AI vertical within LTcI's DIG team, focusing on data intelligence for statutory reasoning. His group develops tools for tax minimization strategy discovery, contract analysis, and legal information extraction, maintaining close ties with legal practitioners through projects like the Prolog-based tax code interpreter. The team's infrastructure supports both academic research and startup partnerships in computational law.
Andrea Cavallaro is a Full Professor at the École Polytechnique Fédérale de Lausanne (EPFL) and Director of the Idiap Research Institute. He holds dual appointments in the School of Engineering (STI) within the Institute of Electrical Engineering and Measurements (IEM) and the School of Engineering's Education Unit (SEL-ENS). His research focuses on machine learning for multimodal perception, privacy-preserving AI, and autonomous systems. Cavallaro earned his PhD in Electrical Engineering from EPFL in 2002 and has held leadership roles including Director of Research at Queen Mary University of London and Turing Fellow at The Alan Turing Institute. Education: PhD in Electrical Engineering (EPFL, 2002) Leadership: Idiap Director, Affiliate at ELLIS Society Editorial Roles: Editor-in-Chief of Signal Processing: Image Communication (2020–2023), Senior Area Editor for IEEE Transactions on Image Processing Research Interests: Machine learning for audio-visual sensing, privacy in AI, autonomous systems perception, and ethical AI frameworks. Key projects include AlignAI (trustworthy AI alignment) and CORSMAL (multimodal object manipulation). Recent articles explore privacy-aware AI models, adversarial attacks, and multimodal perception systems. His work bridges theoretical advancements with practical applications in robotics, healthcare, and education. Awards include the Royal Academy of Engineering Teaching Prize and IAPR Fellowship. Teaching: Leads courses on deep learning ethics and multimodal AI at EPFL. Advising: Supervises 11 PhD students in areas like privacy-preserving algorithms and autonomous systems. Labs/Teams: Coordinates Idiap’s Audiovisual Intelligence and Learning Lab (LIDIAP) and collaborates on projects like GraphNEx (explainable AI via graph neural networks).
Matthew B. Blaschko is a Professor in the Department of Electrical Engineering at KU Leuven, Belgium. He serves as director of the KU Leuven ELLIS unit and is a fellow in the ELLIS Health program. He is a Core PI in the Flanders AI Research Program, working as a workpackage lead for Decision Support Systems and Medical Imaging. Blaschko is also a member of the KU Leuven Institute for Artificial Intelligence and one of the leaders of the working group on Machine Learning and Data Science. Professor Blaschko received his B.S. from Columbia University, M.S. from the University of Massachusetts Amherst, and Dr. rer. nat. from Technische Universität Berlin (awarded for work at Max Planck Institutes Tübingen). He was a Newton International Fellow at the University of Oxford and received his Habilitation (HDR) from École Normale Supérieure de Cachan. Prior to joining KU Leuven, he was a Permanent Research Scientist in the INRIA Saclay Research Center and a Faculty Member at Ecole Centrale Paris. His research focuses on machine learning techniques applied to visual data, with particular emphasis on calibration in deep learning, medical image analysis, and federated learning. Blaschko's work bridges theoretical foundations with practical applications, as evidenced by technology developed in his research being incorporated into MONA, software for ophthalmic image analysis. His research group has made significant contributions to the fields of model calibration, uncertainty estimation, and medical imaging analysis, with recent publications showing strong trends toward improving reliability of AI systems in medical contexts and advancing theoretical understanding of calibration metrics. Professor Blaschko has been recognized with several awards including the Université Paris-Saclay STIC Doctoral School Best Scientific Contribution Award, Best Paper Award at CVPR 2008, Main Award at DAGM 2008, and Best Student Paper Award at ECCV 2008. Professor Blaschko has supervised numerous PhD and Master's students, with current and former students including Deniz Soysal, Claire Marchal, Dongli Xu, Sebastian Gruber, Jiameng Li, Marco Mezzina, and many others working on diverse topics from Alzheimer's disease analysis to surgical phase recognition. His research has been supported by various funding sources including the Flanders AI Research Program. He has co-organized several influential workshops including the "Another Brick in the AI Wall: Building Practical Solutions from Theoretical Foundations" at CVPR 2025, Commands 4 Autonomous Vehicles workshop at ECCV 2020, and the Learning from Limited Labeled Data workshop series at NIPS 2017 and ICLR 2019. His laboratory focuses on machine learning for medical image analysis, with applications in ophthalmology, neurology, and surgical robotics. The group maintains active collaborations with medical institutions and participates in international challenges such as the KNee OsteoArthritis Prediction (KNOAP2020) challenge.
Shang-Tse Chen is an Associate Professor at the Department of Computer Science and Information Engineering and Graduate Institute of Networking and Multimedia , National Taiwan University . He leads the NTU AI Security Lab , focusing on applied and theoretical machine learning with emphasis on cybersecurity, adversarial ML, and ML privacy/fairness. Education: PhD in Computer Science (Georgia Tech, 2019), BSc in CSIE (NTU, 2010) Awards: K. T. Li Young Researcher Award (2025), IBM PhD Fellowship (2018), KDD Best Student Paper Runner-Up (2016), NSF SaTC Grant (2017-2021) His research spans adversarial ML, certified defenses, model inversion attacks, and intersection with differential privacy/fairness. Recent work includes physical adversarial attacks on object detectors and practical defenses using JPEG compression. He teaches courses like Security and Privacy of Machine Learning and Introduction to Medical Informatics . Key publication trends show focus on adversarial robustness (ICML/NeurIPS/ICLR), cybersecurity applications (ACSAC), and ML fairness (ACL/EMNLP). Collaborations include industry partnerships with Intel Labs and Symantec. Scientific Awards: K. T. Li Young Researcher Award (2025) ACM TiiS Best Paper Honorable Mention (2020) IBM PhD Fellowship (2018) KDD Audience Appreciation Award Runner-Up (2018) Symantec Fellowship Runner-Up (2016) KDD Best Student Paper Runner-Up (2016) NSF Grant (2017) He advises 13 current students (PhD/MS/Undergrad) and has mentored alumni now at CMU/UC Berkeley. The lab actively recruits postdocs and students across levels.
Gerd Bruder is an Associate Professor at the University of Central Florida , affiliated with the Department of Computer Science . He previously held positions as a Research Associate Professor at UCF's Institute for Simulation and Training (2016-2022), Post-Doc at University of Hamburg (2014-2016), and Post-Doc at University of Würzburg (2011-2014). Habilitation : Computer Science, University of Hamburg (2017) Ph.D. : Computer Science, University of Münster (2011) M.Sc. : Computer Science (minor in Mathematics), University of Münster (2009) Research Interests : Gerd Bruder's work focuses on Extended Reality (XR) systems, spanning Virtual Reality (VR) , Augmented Reality (AR) , and Mixed Reality (MR) . Key areas include human factors , digital twins , 3D user interfaces , and perceptual modeling , with applications in healthcare, military operations, architecture, and smart environments. Recent Publications highlight innovations in XR locomotion , trust calibration with autonomous agents, spatial cognition , and haptic-visual integration . His 2025 articles address task-switching efficiency in XR and collaborative mixed reality design. Scientific Awards : Best Paper at ACM VRST 2023 2021 TechConnect Innovation Award Best Demo at ACM SUI 2021 Multiple Best Paper Awards (2016-2023) Best Poster & Demo Awards (2015-2020) Patents include systems for medical simulation , smart environment interruption management , and adaptive AR interfaces . His work bridges academic research and practical applications across VR, AR, and MR domains.
Habib Ullah is an Associate Professor in Data Science at the Norwegian University of Life Sciences (NMBU), Norway, where he conducts research at the intersection of computer vision and machine learning. He is affiliated with the Institute of Data Science under the Faculty of Science and Technology. He has previously held academic positions at COMSATS University Islamabad, Pakistan, and the University of Ha'il, Saudi Arabia, and served as a postdoctoral researcher at The Arctic University of Norway. Educational Background: PhD in Information and Communication Technology (Computer Vision), University of Trento, Italy (2011–2015) MSc in Electronics and Computer Engineering, Hanyang University, South Korea (2007–2009) BSc in Computer Systems Engineering, NWFP University of Engineering and Technology, Pakistan (2002–2006) Habib Ullah's research is primarily focused on computer vision and machine learning, with applications in aquaculture, agriculture, and human behavior analysis. He investigates underwater fish feeding sounds using audio classification, develops zero-shot learning models for recognizing unseen classes, and applies deep learning to detect stress in salmon via skin dot patterns. He also explores AI-driven controlled environment agriculture, leveraging sensors and automation for optimal crop growth. His work emphasizes practical AI solutions for real-world challenges in environmental and biological domains. The recent publications highlight a strong trend in leveraging deep learning for zero-shot and semi-supervised learning, particularly in computer vision tasks such as sea ice classification, crowd anomaly detection, and agricultural monitoring. His research spans remote sensing, biomedical signal processing, and human activity recognition, demonstrating interdisciplinary versatility. The keywords reflect a focus on robust feature representation, knowledge transfer, and model generalization. Scientific Awards and Funding: Industrial PhD grant 'Advancing Controlled Environment Agriculture AI' from The Research Council of Norway (Project number 354125, 2 million NOK, 2024) Team member (Coordinator-Participant) in the Battery Cell Assembly Twin (BatCAT) project funded by Horizon Europe (7 mEuro, 2023–2027) Development of an AI-Based Image Analysis System for Monitoring Plant Status (Funding: 1.8 mNOK, starting 2025) Habib Ullah actively supervises PhD projects and contributes to academic service through editorial and organizational roles. He has served as an Associate Editor for IEEE Access, Guest Editor for MDPI Remote Sensing, and Editor of the Springer book Machine Learning Techniques and Sensor Applications for Human Emotion, Activity Recognition, and Support (ML-SHEARS) . He has also been a Track Chair and Program Committee Member for several international conferences, reflecting his leadership in the academic community. His research is supported by significant grants and collaborative projects, indicating strong institutional and international engagement. He is involved in multiple research teams and projects, including the BatCAT project on battery manufacturing and AI applications in controlled environment agriculture with RIFT LABS AS. His lab work integrates deep learning, sensor fusion, and data analytics for environmental and biological monitoring systems.
Jiajun Wu is an Assistant Professor of Computer Science and, by courtesy, of Psychology at Stanford University. He holds multiple affiliations including membership in Bio-X, Faculty Affiliate status at the Institute for Human-Centered Artificial Intelligence (HAI), and membership in both the Wu Tsai Human Performance Alliance and Wu Tsai Neurosciences Institute. Dr. Wu earned his Ph.D. and S.M. in Electrical Engineering and Computer Science from the Massachusetts Institute of Technology before joining Stanford. Dr. Wu's research program focuses on creating AI systems that understand and interact with the physical world through the integration of computer vision, machine learning, robotics, and cognitive science. His work emphasizes physics-based modeling combined with deep learning to develop systems capable of perceiving, reasoning about, and predicting physical interactions. Key research areas include 3D scene understanding, neurosymbolic AI approaches, multimodal perception (combining vision, sound, and language), and embodied intelligence for robotics applications. His lab develops novel frameworks that bridge the gap between neural networks and symbolic reasoning to create more interpretable and robust AI systems. Analysis of Dr. Wu's recent publications reveals a strong trajectory toward integrated multimodal understanding for embodied AI. His work increasingly combines vision, sound, and language processing with physical reasoning to create systems that can interact meaningfully with the physical world. There's a clear progression from foundational computer vision research toward practical robotics applications, with significant emphasis on foundation models for robotics, sim2real transfer techniques, and creating comprehensive datasets for embodied AI research. Dr. Wu's exceptional contributions have been recognized with numerous prestigious awards including the NSF CAREER award (2024), Young Investigator Programs from ONR (2024) and AFOSR (2023), the Okawa research grant (2024), and being named to IEEE Intelligent Systems' 'AI's 10 to Watch' (2024). He has received multiple best paper awards at leading conferences including ICRA (2024), SIGGRAPH Asia (2023), and CoRL (2023). Dr. Wu actively mentors a large cohort of students across multiple levels, serving as primary advisor for doctoral candidates, master's students, and numerous independent researchers. His research is supported by substantial funding from major technology companies including Google, Meta, Amazon, Samsung, and J.P. Morgan, as well as government agencies like NSF, ONR, and AFOSR, reflecting the significance and impact of his work in physical AI and multimodal perception systems. Dr. Wu leads a dynamic research group at Stanford that collaborates extensively with the Wu Tsai Neurosciences Institute and Institute for Human-Centered AI. Current projects include developing neurosymbolic models for computer graphics, creating multisensory datasets like OBJECTFOLDER 2.0 for sim2real transfer in robotics, and building foundation models for embodied intelligence that can understand and manipulate objects with human-like physical intuition.