Almut Sophia Koepke is a junior research group leader at the Technical University of Munich and University of Tübingen, focusing on multimodal learning problems integrating sound, vision, and text. Her work bridges foundational research in audio-visual understanding with practical applications in few-shot learning, zero-shot translation, and cross-modal attention mechanisms.
Oliver Kroemer is an Associate Professor at Carnegie Mellon University's Robotics Institute (RI), affiliated with the Intelligent Autonomous Manipulation (IAM) Lab. His research focuses on enabling robots to learn versatile manipulation skills through lifelong frameworks, with applications in elder care, environmental maintenance, and hazardous operations. Developed methods for robot learning via physical interaction and reinforcement learning Created representations for contact states and motor primitives to improve skill generalization Research Interests: Spanning robot learning, tactile sensing, force-velocity control, and lifelong skill acquisition. Projects include Agile and Dynamic Interactions for Mobile Manipulation and Integrated Planning and Learning (Pillar project). Scientific Awards: Finalist, Georges Giralt Ph.D. Award (2015) Education: Masters & Bachelors in Engineering, University of Cambridge (2008) Ph.D., Technische Universitaet Darmstadt (2014) Students & Affiliates: Current PhD: Mark Lee, Sarvesh Patil, Saumya Saxena, Yunus Seker, Zilin Si Past PhD: Alex LaGrassa, Tabitha Lee, Qiao Liang, Shivam Vats, Kevin Zhang
WANG Ye is an Associate Professor in the Department of Computer Science at the School of Computing, National University of Singapore (NUS). He holds a PhD in Information Technology from Tampere University of Technology, Finland, and has been a tenured faculty member at NUS since 2002, following his industry research role at Nokia Research Center. He is the director of the Sound and Music Computing Lab at NUS, leading cutting-edge research in AI-driven music and health technologies. PhD, Information Technology, Tampere University of Technology, Finland (2002) MSc, Telecommunications, Braunschweig University of Technology, Germany (1993) BSc, Telecommunications, South China University of Technology, China (1983) His research is centered on Sound and Music Computing for Human Health and Potential (SMC4HHP) , with a focus on eHealth, eLearning, mobile/wearable computing, and music information retrieval. His work spans AI for stroke rehabilitation, language learning through singing, singing voice synthesis, and automatic music transcription. He has pioneered systems like SLIONS (language learning via karaoke), CocoLyricist (AI co-creation for stroke recovery), and SinTechSVS (expressive singing voice synthesis). The latest articles highlight a strong trend in AI-driven music and health technologies , particularly in controllable lyric generation, singing voice synthesis, automatic pronunciation assessment, and multimodal music transcription. The research increasingly integrates large language models, explainable AI, fairness, and real-world deployment, reflecting a shift from theoretical exploration to practical, human-centered applications in healthcare and education. Dr. Wang has received numerous scientific honors, including: Best Paper Awards at ACM MM, ISMIR, IEEE ISM, and CHI First Prize, Asia Pacific Assistive, Rehabilitative, and Therapeutic Technologies Challenge (2015) Faculty Teaching Excellence Award, NUS School of Computing (2024) Top Paper Award, ACM Multimedia 2022 AI in Medicine Collaborative Grant for CocoLyricist project He has supervised over 11 PhD and 20 MComp students and is currently guiding six PhD candidates. His grants come from MOE, NRF, A*STAR, Nokia, and Smule. He has served as General Chair of ISMIR2017 and TPC Co-Chair of ICOT2017, and is on the editorial boards of IEEE Transactions on Multimedia and Journal of New Music Research. He has also developed and taught the first course on Sound and Music Computing in Singapore. Dr. Wang leads the Sound and Music Computing Lab (SMC Lab) , a multidisciplinary team exploring the synergy of music computing, AI, mobile technology, and cloud systems for health and education. The lab actively collaborates with medical institutions such as NUS Yong Loo Lin School of Medicine, Singapore General Hospital, and Harvard Medical School, and is currently working on projects in AI-supported language learning, stroke rehabilitation, and intelligent music interfaces.
Peter Henderson is an Assistant Professor at Princeton University with joint appointments in the Department of Computer Science and the School of Public and International Affairs. He is affiliated with the Center for Information Technology Policy (CITP), Princeton Language and Intelligence Initiative (PLI), Center for Statistics and Machine Learning (CSML), and Program in Law & Public Policy (PLAW). J.D./Ph.D., Stanford University, 2023 M.Sc., McGill University and Montréal Institute for Learning Algorithms Henderson's research focuses on the critical intersection of artificial intelligence and law, with particular emphasis on AI safety, methods to improve reasoning in foundation models, interdisciplinary approaches to law and AI, and AI governance. His work spans language-grounded reinforcement learning, alignment techniques, strategic decision-making in legal contexts, and public interest artificial intelligence. He investigates how legal frameworks can guide the development of AI systems that benefit society while addressing potential harms. Henderson's recent publications reveal a strong trajectory toward addressing the safety, governance, and legal implications of foundation models. His work combines rigorous technical AI research with deep legal analysis, particularly examining how intellectual property law interacts with AI development, regulatory pathways that balance innovation with harm prevention, and the role of law in shaping responsible AI deployment. A significant thread throughout his research is the development of better evaluation methodologies for AI systems, especially in critical domains like law. SEAS Excellence in Teaching Award (2025) Henderson leads the Princeton Law+Language, AI, & Society (POLARIS) Lab, where he advises students and researchers working at the intersection of AI and legal studies. His research has been supported through collaborations with government agencies, including work with the IRS on audit selection algorithms. His findings have informed practical applications in government efficiency and equity, as well as legal system improvements. Henderson runs the POLARIS Lab at Princeton, which focuses on developing AI systems that work for the public interest, particularly in legal contexts. His team develops sequential decision-making systems for government efficiency, foundation models capable of reasoning about law, and improved safety evaluation frameworks for AI systems. The lab maintains strong connections with legal practitioners and policymakers to ensure research has real-world impact.
Pasquale Bottalico serves as Associate Professor in the Department of Speech and Hearing Science at the University of Illinois, with dual appointments as Associate Professor at the Center for Latin American and Caribbean Studies and Affiliate Faculty in the School of Music. His unique interdisciplinary profile bridges engineering, music performance, and speech science, reflecting his dual academic training and professional artistry. His educational foundation includes: Bachelor's in Telecommunications Engineering from Univeristà Mediterranea di Reggio Calabria, Italy Concurrent Opera Singing degree from F. Cilea Music Academy, Reggio Calabria Master's in Telecommunications Engineering from Politecnico di Torino, Italy Ph.D. in Metrology specializing in acoustics measurement uncertainty and classroom acoustics Dr. Bottalico's research centers on vocal load quantification and professional voice techniques , with significant contributions to understanding vocal fatigue in teachers and singers. His work spans Speech Intelligibility in educational environments, Room Acoustics for performance and learning spaces, and Musical Acoustics of historical vocal styles. A distinctive thread throughout his research examines how acoustic conditions modulate voice production and perception, increasingly incorporating virtual reality and bone conduction technologies for innovative assessment and intervention approaches. His Colombian vocal health study demonstrates cross-cultural applications of his work. Analysis of his 2023-2025 publications reveals three dominant research trajectories: (1) The impact of noise and dysphonia on children's speech processing in educational settings, using multimodal assessment including EEG; (2) Virtual reality applications for voice production research and therapeutic intervention; (3) Cross-cultural validation of vocal fatigue metrics and development of biofeedback systems. His work consistently bridges engineering precision with clinical applicability, particularly for professional voice users in challenging acoustic environments. No scientific awards were documented in the available information. While specific advising relationships aren't detailed, his research collaborations span international institutions including Colombian and Italian universities, suggesting graduate mentorship in interdisciplinary projects. No grant information was provided, though his systematic reviews and cross-cultural studies imply externally funded research activities. Though no dedicated laboratory is specified, his virtual reality voice studies and acoustic parameter assessments suggest affiliations with audio engineering facilities and voice clinics, likely through the Speech and Hearing Science department's research infrastructure.
Timothy M. Hospedales is a Professor of Artificial Intelligence at the Institute of Perception, Action and Behaviour within the School of Informatics at the University of Edinburgh . He also serves as VP AI and Head of Samsung AI Research Centre Europe . His research focuses on efficient and robust AI , emphasizing meta-learning , lifelong transfer-learning , and domain adaptation in both probabilistic and deep learning frameworks. Applications span computer vision , vision and language , reinforcement learning for robotics , and finance . Professor at University of Edinburgh (2020–present) ELLIS Fellow (2021) Head of Samsung AI Research Europe (2020–present) Founding Director of Applied Machine Learning Lab at QMUL (2012–2016) His work includes pioneering contributions to meta-learning , few-shot learning , and self-supervised methods , with notable awards such as the Best Paper Prize at ICML AutoML 2018 and Best Student Paper at ICPR 2018 . He has co-authored 15+ recent papers on topics like Vision-Language Models , Medical AI Fairness , and Diffusion Model Optimization . He served as Program Co-Chair for BMVC 2018 and AAAI 2022 , and authored a book on Visual Adaptation in the Deep Learning Era (2022). Co-Chair, BMVC 2018 Guest Editor, IET CV Special Issue (2016) Keynote Speaker at TASK-CV Workshop (ECCV 2016) Special Issue on Fewer Labels (IEEE PAMI 2020) His leadership extends to organizing workshops like the Learning-to-Learn Workshop at ICLR 2021 , Meta-Learning Workshop at NeurIPS 2020 , and Domain Generalisation Workshop at ICLR 2023 . Current projects include Meta-Omnium (CVPR 2023) for general-purpose meta-learning and MetaAudio (ICANN 2022) for few-shot audio classification benchmarks.
Professor Carlo Harvey is a creative technologist at the School of Digital Arts (SODA), Manchester Metropolitan University. His interdisciplinary research merges games , machine learning , virtual production , and cultural heritage reinterpretation . He leads industry collaborations with entities like Jaguar Land Rover and Epic Games, focusing on AI-driven interactive audio, real-time visualization, and accessibility solutions. Award-winning projects : TIGA, Innovate UK, and Epic Games MegaGrant for Accession Industry partnerships : Automotive sector, cultural institutions His research spans human-computer interaction , multisensory virtual environments , and acoustic-visual cross-modal perception . Recent publications address robotic simulations, motion alignment, and haptic feedback systems. Scientific recognition : TIGA Award, Innovate UK Funding, Epic Games MegaGrant Advocacy : Digital inclusion, creative collaboration, social impact of technology
Prof. Roger Wattenhofer is a Full Professor at the Department of Information Technology and Electrical Engineering, ETH Zurich, Switzerland, and Deputy head of the Computer Engineering and Networks Lab. He holds a doctorate in Computer Science from ETH Zurich (1998) and has held positions at Brown University and Microsoft Research before returning to ETH. His research focuses on distributed computing, wireless networks, and algorithmic systems design, with contributions to Byzantine agreement protocols, blockchain technologies, and neural network architectures. He teaches courses such as Distributed Systems and Computational Thinking. Education: Ph.D. in Computer Science (ETH Zurich, 1998). Research interests include distributed systems, network algorithms, and the intersection of machine learning with distributed computing. His work spans both theoretical foundations and practical implementations, addressing challenges in fault tolerance, consensus mechanisms, and algorithmic efficiency. Recent publications explore topics like adversarial robustness in voting systems, privacy in reinforcement learning, and generative music models. He actively contributes to open-source frameworks and benchmarks for neural algorithmic reasoning.
Najim Dehak is an Associate Professor in the Department of Electrical and Computer Engineering at Johns Hopkins University, part of the Whiting School of Engineering. His research focuses on machine learning applied to speech processing, audio classification, and health applications. He is renowned for developing the I-vector representation for speaker recognition, introduced in 2008 during a workshop at Johns Hopkins’ Center for Language and Speech Processing. Prior to this role, he was a research scientist at MIT’s Computer Science and Artificial Intelligence Laboratory. Dehak holds a PhD from the School of Advanced Technology in Montreal (2009). He is a Senior Member of IEEE and contributes to the IEEE Speech and Language Technical Committee. His work bridges AI, healthcare, and signal processing, with notable contributions to neurodegenerative disease detection via speech and handwriting analysis. Research interests include adversarial attacks on speech systems, multimodal biomarker discovery, and robust speech processing across demographics. His lab’s tools, like the Hermespeech Recorder, enable scalable data collection for clinical and research applications. Education: PhD in Advanced Technology (2009), Montreal Affiliations: Johns Hopkins University, IEEE Labs/Teams: Center for Language and Speech Processing (CLSP) His recent work explores AI’s role in aging research, including Alzheimer’s and Parkinson’s disease detection through speech, eye tracking, and handwriting analysis. Ongoing projects address fairness in speaker verification and robustness against adversarial attacks in ASR systems.
Prof. Matthias Nießner is a Professor at the Technical University of Munich , where he leads the Visual Computing Lab . Prior to this, he held a Visiting Assistant Professor position at Stanford University . His work bridges computer vision , graphics , and machine learning , focusing on 3D reconstruction , semantic scene understanding , and AI-driven video synthesis . Prof. Nießner has published over 150 works in top venues like SIGGRAPH , CVPR , and ECCV , with several receiving best paper awards (SIGCHI’14, HPG’15, SPG’18, SIGGRAPH’16 Emerging Tech). His research has garnered international media attention, including features in the New York Times , Wall Street Journal , and MIT Technological Review , as well as TV demonstrations (e.g., Jimmy Kimmel Live for Face2Face technology). Awards : TUM-IAS Rudolph Moessbauer Fellowship (2017–ongoing) Google Faculty Award (2017) Nvidia Professor Partnership Award (2018) ERC Starting Grant (2018, €1.5M) Eurographics Young Researcher Award (2019) Research Trends : 3D Gaussian Splatting for real-time rendering Neural Radiance Fields (NeRF) with mesh supervision Audio-driven facial animation via diffusion models Latent space diffusion for 3D scenes Self-supervised and zero-shot methods for 3D and image analysis As a co-founder and director of Synthesia Inc. , he drives democratization of synthetic media. His YouTube channel has over 5 million views, reflecting his impact beyond academia.
Xavier Serra is a Full Professor at the Department of Engineering at Universitat Pompeu Fabra (UPF), Barcelona. He is the founder and director of the Music Technology Group (MTG), and leads the UPF-BMAT Chair on AI and Music. He also coordinates the Master in Sound and Music Computing and serves as President of the Phonos Foundation. His research focuses on audio signal processing, sound and music computing, and computational musicology, emphasizing open science and open innovation. Education: BSc in Biology, University of Barcelona (1981) Master in Music, Florida State University (1983) PhD in Computer Music, Stanford University (1989) Research Interests: Audio Signal Processing Data-Driven and Knowledge-Driven Methodologies Music Information Retrieval Cultural Music Analysis (e.g., Carnatic/Turkish/Andalusian Music) Music Education Technology Notable Projects: CompMusic (ERC Advanced Grant, 2010-2017): Multicultural computational music analysis Open datasets: Freesound, Saraga, FSD50K Technologies: Reactable, Vocaloid, Essentia API Recent Trends in Articles: Focus on AI-driven audio processing (neural fingerprints, generative models), cross-cultural music analysis, and explainable music difficulty estimation. Awards: ERC Advanced Grant (2010) for CompMusic Project. Labs/Teams: Director of MTG, Phonos Foundation, and UPF-BMAT Chair. Active in open-source projects and international collaborations.
Elena Maria Baralis is a Full Professor at the Department of Control and Computer Science (DAUIN) at the Polytechnic University of Turin. She serves as Pro-Rector, member of the Board of Directors (without voting rights), member of the Academic Senate (without voting rights), and coordinator of the University's Permanent Observatory for Monitoring the Academic Sector. She chairs the Control and Computer Engineering Department and previously chaired the Computer Engineering School from October 2012 to October 2018. Her research interests focus on database systems and data mining, specifically explainable AI, bias detection in data analytics, and machine learning algorithms for big data. Her work spans various application domains including predictive maintenance, Industry 4.0, and healthcare. Recent publications demonstrate her expertise in speech processing, bias mitigation, and innovative neural network architectures like Kolmogorov-Arnold Networks. Her research output shows a clear trend toward addressing fairness and explainability in AI systems while exploring novel approaches to speech and language understanding. Professor Baralis has received significant recognition including becoming a Fellow of the Academy of Sciences of Turin in 2017. She has served as Editor-in-Chief for IEEE Internet of Things Journal (2016-2019) and Knowledge and Information Systems (2014-present). She actively mentors doctoral students including Claudio Savelli (researching Machine Unlearning), Eleonora Poeta, Giuseppe Gallipoli, Alkis Koudounas, and others. Her research is supported by numerous projects including AI4CTI (Artificial Intelligence for Cyber Threat Intelligence, 2025-2028), Smart manufacturing driven by Machine Learning in Industry 4.0 (2019-2020), and I-REACT (2016-2019).
David W. Jacobs is a Professor in the Department of Computer Science at the University of Maryland, with a joint appointment at the University of Maryland Institute for Advanced Computer Studies (UMIACS). He also served as the interim Director of the University of Maryland Center for Machine Learning starting in 2018. University: University of Maryland School: College of Computer, Mathematical, and Natural Sciences Department: Department of Computer Science Academic Rank: Professor Education: He received his B.A. from Yale University, and M.S. and Ph.D. in Computer Science from MIT. Research Interests: His research primarily focuses on computer vision and machine learning, particularly visual object recognition, lighting variation modeling, 3D reconstruction, perceptual organization, motion understanding, and the integration of vision with graphics and human-computer interaction. A major applied contribution is the development of Leafsnap , an electronic field guide app for plant identification, which has been downloaded over 1.5 million times and used in biodiversity and educational contexts. Publication Trends: His recent scholarly output centers on deep learning, convolutional networks, residual architectures, generative models (especially GANs), and interpretability. His work often bridges theoretical insights with practical applications in vision and AI. Scientific Awards: Honorable Mention, Best Paper Award, CVPR 2000 Best Student Paper Award, UIST 2003 Best Paper Award, Eurographics 2016 2011 Edward O. Wilson Biodiversity Technology Pioneer Award for Leafsnap Teaching and Advising: He has taught advanced courses such as CMSC 422 (Introduction to Machine Learning) and CMSC 828L (Deep Learning). He mentors students through course projects and research, though specific advisees are not listed. He has collaborated with institutions like Columbia University and the Smithsonian on impactful interdisciplinary projects. Labs and Teams: He is affiliated with UMIACS and leads research efforts in vision and learning, contributing to the University of Maryland Center for Machine Learning. His team has developed several mobile applications including Leafsnap, Birdsnap, and Dogsnap, demonstrating a strong focus on real-world deployment of vision technology.
Mark Gales is Professor of Information Engineering at the University of Cambridge and an Official Fellow at Emmanuel College. He is currently on sabbatical leave for the 2024/25 academic year. Prior to his academic career, he worked as a consultant at Roke Manor Research Ltd, developing radar systems, before transitioning to speech and language processing. PhD in 'Model-Based Techniques for Robust Speech Recognition' (University of Cambridge, 1995) BA in Electrical and Information Sciences (University of Cambridge, 1988) His research focuses on speech and language processing , particularly in automated language assessment and low-resource speech technology . He leads the Automated Language Teaching and Assessment (ALTA) Institute , which collaborates with Cambridge University Press & Assessment (CUP&A) to develop commercial tools like Linguaskill and Speak & Improve . These platforms provide automated spoken/written assessment for millions of users globally. Recent publications highlight his work in LLM-driven speech processing , including adversarial attacks on foundation models, end-to-end spoken error correction, and uncertainty estimation frameworks. His team's research spans multilingual capabilities, with deployments in languages ranging from Dholuo to Tok Pisin . Awards : IEEE Fellow, ISCA Fellow Leadership : Fellows' Steward at Emmanuel College Mark has contributed extensively to Hidden Markov Model (HMM) applications in speech recognition, which underpinned early automatic speech systems. His work now bridges LLM-based language assessment with cross-lingual transfer learning and robustness testing for real-world deployments.
Yao Qin is an Assistant Professor in the Department of Electrical and Computer Engineering at the University of California, Santa Barbara (UCSB), with dual affiliation in the Department of Computer Science. She concurrently serves as Co-Director of the REAL AI Initiative at UCSB and holds a Senior Research Scientist position at Google DeepMind, where she contributes to the Gemini Multimodal project. Her academic credentials include a PhD in Computer Science and Engineering from the University of California, San Diego (advised by Prof. Garrison W. Cottrell) and a BS in Electrical Engineering from Dalian University of Technology. During her doctoral studies, she completed internships with pioneering researchers Geoffrey Hinton and Ian Goodfellow. Dr. Qin's research program centers on machine learning robustness, with emphasis on adversarial robustness, out-of-distribution generalization, and fairness. She develops reliable AI systems specifically for healthcare applications, with diabetes management as a primary focus. Her lab explores critical themes including AI safety in multimodal models and diabetes-specific AI solutions, particularly exercise metabolism modeling and glycemic effect prediction. Recent publications reveal a strong trajectory in robust machine learning with cross-domain applications. Her work consistently bridges theoretical robustness concepts with practical healthcare implementations, particularly in diabetes care. Key publication venues include CVPR, ICML, NeurIPS, and ICLR, with notable contributions to out-of-distribution detection, adversarial transfer learning, and multimodal AI safety. Her distinguished recognition includes: EECS Rising Star at MIT (2021) UCSB Regents' Junior Faculty Fellowship Award Helmsley Charitable Trust award for Type 1 diabetes research UCSB Faculty Research Grant American Diabetes Association Abstract Award (ADA-2025) Dr. Qin actively mentors four PhD students—Mehak Dhaliwal, Andong Hua, Kenan Tang, and Youngseok Yoon—on projects spanning LLMs for diabetes, multimodal robustness, and generative time-series modeling. Her research is funded by the Helmsley Charitable Trust and UCSB, with recent grants supporting exercise-specific AID algorithms for diabetes management. As Co-Director of the REAL AI Initiative, she leads a research ecosystem focused on developing reliable artificial intelligence. Current lab activities include organizing workshops at NeurIPS-2024 (AdvML-Frontiers and AIM-FM) and developing next-generation diabetes management tools through collaborations with medical institutions.