Oliver Kroemer is an Associate Professor at Carnegie Mellon University's Robotics Institute (RI), affiliated with the Intelligent Autonomous Manipulation (IAM) Lab. His research focuses on enabling robots to learn versatile manipulation skills through lifelong frameworks, with applications in elder care, environmental maintenance, and hazardous operations. Developed methods for robot learning via physical interaction and reinforcement learning Created representations for contact states and motor primitives to improve skill generalization Research Interests: Spanning robot learning, tactile sensing, force-velocity control, and lifelong skill acquisition. Projects include Agile and Dynamic Interactions for Mobile Manipulation and Integrated Planning and Learning (Pillar project). Scientific Awards: Finalist, Georges Giralt Ph.D. Award (2015) Education: Masters & Bachelors in Engineering, University of Cambridge (2008) Ph.D., Technische Universitaet Darmstadt (2014) Students & Affiliates: Current PhD: Mark Lee, Sarvesh Patil, Saumya Saxena, Yunus Seker, Zilin Si Past PhD: Alex LaGrassa, Tabitha Lee, Qiao Liang, Shivam Vats, Kevin Zhang
WANG Ye is an Associate Professor in the Department of Computer Science at the School of Computing, National University of Singapore (NUS). He holds a PhD in Information Technology from Tampere University of Technology, Finland, and has been a tenured faculty member at NUS since 2002, following his industry research role at Nokia Research Center. He is the director of the Sound and Music Computing Lab at NUS, leading cutting-edge research in AI-driven music and health technologies. PhD, Information Technology, Tampere University of Technology, Finland (2002) MSc, Telecommunications, Braunschweig University of Technology, Germany (1993) BSc, Telecommunications, South China University of Technology, China (1983) His research is centered on Sound and Music Computing for Human Health and Potential (SMC4HHP) , with a focus on eHealth, eLearning, mobile/wearable computing, and music information retrieval. His work spans AI for stroke rehabilitation, language learning through singing, singing voice synthesis, and automatic music transcription. He has pioneered systems like SLIONS (language learning via karaoke), CocoLyricist (AI co-creation for stroke recovery), and SinTechSVS (expressive singing voice synthesis). The latest articles highlight a strong trend in AI-driven music and health technologies , particularly in controllable lyric generation, singing voice synthesis, automatic pronunciation assessment, and multimodal music transcription. The research increasingly integrates large language models, explainable AI, fairness, and real-world deployment, reflecting a shift from theoretical exploration to practical, human-centered applications in healthcare and education. Dr. Wang has received numerous scientific honors, including: Best Paper Awards at ACM MM, ISMIR, IEEE ISM, and CHI First Prize, Asia Pacific Assistive, Rehabilitative, and Therapeutic Technologies Challenge (2015) Faculty Teaching Excellence Award, NUS School of Computing (2024) Top Paper Award, ACM Multimedia 2022 AI in Medicine Collaborative Grant for CocoLyricist project He has supervised over 11 PhD and 20 MComp students and is currently guiding six PhD candidates. His grants come from MOE, NRF, A*STAR, Nokia, and Smule. He has served as General Chair of ISMIR2017 and TPC Co-Chair of ICOT2017, and is on the editorial boards of IEEE Transactions on Multimedia and Journal of New Music Research. He has also developed and taught the first course on Sound and Music Computing in Singapore. Dr. Wang leads the Sound and Music Computing Lab (SMC Lab) , a multidisciplinary team exploring the synergy of music computing, AI, mobile technology, and cloud systems for health and education. The lab actively collaborates with medical institutions such as NUS Yong Loo Lin School of Medicine, Singapore General Hospital, and Harvard Medical School, and is currently working on projects in AI-supported language learning, stroke rehabilitation, and intelligent music interfaces.
Catherine Lai is a Reader (~Associate Professor) in the Department of Linguistics and English Language at the University of Edinburgh, with strong affiliations to the Centre for Speech Technology Research (CSTR) and the Institute for Language, Cognition and Computation (ILCC) in the School of Informatics. She is based in the School of Philosophy, Psychology and Language Sciences and is actively involved in research, teaching, and academic service. Department: Department of Linguistics and English Language School: School of Philosophy, Psychology and Language Sciences Research Institutes: Centre for Speech Technology Research, Institute for Language, Cognition and Computation Email: C.Lai@ed.ac.uk Her research centers on the role of prosody—non-lexical aspects of speech—in spoken communication. She investigates how prosody contributes to discourse structure, information structure, and affect in dialogue, using interdisciplinary methods from linguistics and machine learning. Her work bridges theoretical linguistics and practical speech technology, aiming to improve spoken language understanding and synthesis systems. She is particularly interested in how prosody shapes listener expectations and how affect and topic are expressed and perceived in conversation. Her recent publications reflect a strong focus on self-supervised learning in speech models, emotion recognition, ASR error correction using large language models, cognitive state classification, and ethical considerations in language technology. She explores topics such as the uncanny valley in synthetic speech, gender expression through voice, and community-centered development of language technologies. Prize from Scopus Profile Catherine Lai has supervised several PhD students, including Leimin Tian and Yuanchao Li, and has been involved in significant research projects, such as a Toyota-funded initiative on spoken dialogue for robot companions. She has secured multiple grants and leads a research agenda that integrates theoretical inquiry with real-world applications in assistive technologies and social science. Her academic service includes organizing major conferences like Interspeech and UK and Ireland Speech. She is a key member of research teams at CSTR and ILCC, collaborating across disciplines to advance the understanding of spoken communication and the development of robust, ethical speech technologies.
Juan Pablo Bello is Professor of Music Technology and Computer Science & Engineering at NYU. He directs the Music and Audio Research Lab (MARL) and the Center for Urban Science and Progress (CUSP). He holds a Ph.D. in Electronic Engineering from Queen Mary University of London (2003) and a B.Eng. from Universidad Simón Bolívar (1998). Research integrates signal processing, machine learning and acoustics to analyze music/audio. Current projects include urban sound mapping, music structure discovery, and audio event detection. Funded by NSF and industry partners, with 100+ publications in IEEE/ACM venues. Honors include an NSF CAREER Award and Fulbright scholarship. He teaches graduate courses in music informatics and fosters industry-academia partnerships in audio AI.
Pierre Vandergheynst is a Full Professor at the Swiss Federal Institute of Technology Lausanne (EPFL) in the Department of Electrical Engineering, with a courtesy appointment in Computer and Communication Sciences. He serves as EPFL’s Vice-Provost for Education since 2015 and leads the Signal Processing Laboratory 2 (LTS2). His research spans harmonic analysis, sparse approximations, mathematical data processing, and applications in signal/image processing, computer vision, machine learning, and graph-based data analysis. PhD in Mathematical Physics (1998), Université catholique de Louvain Postdoctoral Researcher at EPFL (1998-2001) Assistant Professor at EPFL (2002-2007) His research explores geometry/symmetry in high-dimensional data, redundant dictionaries for dimensionality reduction, and computational harmonic analysis on manifolds. Recent work focuses on protein structure modeling, geometric deep learning, and graph-based signal processing. Key article trends include graph neural networks for protein analysis, geometric deep learning in neuroscience, and structured knowledge priors in neural models. His 2023-2025 publications emphasize interpretable AI, long-range dependencies in graphs, and molecular representation learning. Scientific Awards: IEEE Signal Processing Magazine Best Paper Award (2023) Signal Processing Society Best Paper Award (2022) Apple ARTS Award (2007) De Boelpaepe Prize, Royal Academy of Sciences of Belgium (2009-2010) He has supervised over 30 PhD theses and contributed to foundational work in graph signal processing, compressive sensing, and geometric deep learning. His lab develops tools for data science on non-Euclidean structures, with applications in medicine, astronomy, and wireless systems.
Professor Carlo Harvey is a creative technologist at the School of Digital Arts (SODA), Manchester Metropolitan University. His interdisciplinary research merges games , machine learning , virtual production , and cultural heritage reinterpretation . He leads industry collaborations with entities like Jaguar Land Rover and Epic Games, focusing on AI-driven interactive audio, real-time visualization, and accessibility solutions. Award-winning projects : TIGA, Innovate UK, and Epic Games MegaGrant for Accession Industry partnerships : Automotive sector, cultural institutions His research spans human-computer interaction , multisensory virtual environments , and acoustic-visual cross-modal perception . Recent publications address robotic simulations, motion alignment, and haptic feedback systems. Scientific recognition : TIGA Award, Innovate UK Funding, Epic Games MegaGrant Advocacy : Digital inclusion, creative collaboration, social impact of technology
Professor Simon Godsill MA PhD FIET FIEEE is a University Professor of Statistical Signal Processing in the Department of Engineering at the University of Cambridge. He heads a research team specializing in statistical signal processing, digital audio restoration, and Bayesian inference. His work addresses the processing and analysis of digital speech, audio, tracking systems, and financial datasets, with a focus on probabilistic modeling and computational methods. Research interests include statistical signal processing , degraded signal restoration , and Bayesian computational methods . Recent publications emphasize Gaussian processes, variational inference, and multi-object tracking for applications in audio enhancement and financial data analysis. He co-founded the audio remastering company CEDAR Audio Ltd in 1988. Scientific awards: Fellow of the Institution of Engineering and Technology (FIET) Fellow of the Institute of Electrical and Electronics Engineers (FIEEE) Outside academia, he enjoys singing, cricket, piano/organ playing, and running. His team at Cambridge's Engineering department focuses on robust tracking algorithms and signal enhancement techniques.
Ali H. Sayed is the Dean of the School of Engineering (Faculté des sciences et techniques de l'ingénieur - STI) at École Polytechnique Fédérale de Lausanne (EPFL), Switzerland, where he also directs the Adaptive Systems Laboratory (Laboratoire de systèmes adaptatifs). Previously, he served as an emeritus professor and chair of the Electrical Engineering Department at UCLA. He is a highly cited researcher and a member of the US National Academy of Engineering and the World Academy of Sciences. Sayed served as president of the IEEE Signal Processing Society in 2018 and 2019. Professor Sayed's research focuses on adaptation and learning theories, data and network sciences, statistical inference, multi-agent systems, adaptive networks, and optimization. His work bridges theoretical foundations with practical applications in signal processing, machine learning, and network science. He has made significant contributions to distributed learning algorithms, social learning over networks, and adaptive signal processing techniques that have influenced both academic research and practical implementations. His recent publications demonstrate a strong focus on multi-agent systems, distributed learning, privacy-preserving techniques, and social learning over networks. The research trends show increasing emphasis on federated learning with privacy guarantees, graph-based learning approaches, and the intersection of social dynamics with information processing. His work consistently addresses fundamental theoretical questions while maintaining relevance to practical applications in communication networks, social media analysis, and distributed artificial intelligence systems. Professor Sayed has received numerous prestigious awards throughout his career, including: IEEE Fourier Award (2022) Norbert Wiener Society Award (2020) IEEE Signal Processing Society Education Award (2015) Papoulis Award from the European Association for Signal Processing (2014) Technical Achievement Award from IEEE Signal Processing Society (2012) Terman Award from the American Society for Engineering Education (2005) IEEE Donald G. Fink Prize (1996) Multiple Best Paper Awards from IEEE and EURASIP Sayed has authored or co-authored over 570 publications and six monographs. He has mentored numerous PhD students and researchers in the fields of signal processing and adaptive systems. His editorial leadership includes serving as Editor-in-Chief of IEEE Transactions on Signal Processing (2003-2005) and EURASIP Journal on Advances in Signal Processing (2006-2007), as well as Founding Editor-in-Chief of the Open Access Book Series on Information and Learning Sciences. At EPFL, Professor Sayed leads the Adaptive Systems Laboratory, which focuses on developing theoretical frameworks and practical algorithms for adaptive systems, networked learning, and distributed signal processing. The lab's research encompasses both fundamental theoretical investigations and applications to real-world problems in communications, social networks, and computational biology.
Professor Simo Särkkä holds a position in Sensor Informatics and Medical Technology at the Department of Electrical Engineering and Automation (EEA), Aalto University. His research focuses on multi-sensor data processing, Bayesian filtering, machine learning, and their applications in medical technology, brain imaging, and inverse problems. He leads research groups including the Helsinki Institute for Information Technology (HIIT) and Sensor Informatics and Medical Technology. His work bridges theoretical advancements in probabilistic methods with practical implementations in healthcare and engineering. Key research interests include Gaussian processes, stochastic differential equations, quantum machine learning, and signal processing. He has contributed to advancements in algorithms for nonlinear state-space models, parallel computing techniques, and medical imaging technologies such as scatter correction in CT scans. His methodologies are applied across domains like autonomous systems, robotics, and bioengineering. Notable publications span topics like quantum-assisted Gaussian regression, physics-informed machine learning for industrial processes, and parallel-in-time numerical methods. His work emphasizes computational efficiency and robustness in high-dimensional and real-time systems.
Xavier Serra is a Full Professor at the Department of Engineering at Universitat Pompeu Fabra (UPF), Barcelona. He is the founder and director of the Music Technology Group (MTG), and leads the UPF-BMAT Chair on AI and Music. He also coordinates the Master in Sound and Music Computing and serves as President of the Phonos Foundation. His research focuses on audio signal processing, sound and music computing, and computational musicology, emphasizing open science and open innovation. Education: BSc in Biology, University of Barcelona (1981) Master in Music, Florida State University (1983) PhD in Computer Music, Stanford University (1989) Research Interests: Audio Signal Processing Data-Driven and Knowledge-Driven Methodologies Music Information Retrieval Cultural Music Analysis (e.g., Carnatic/Turkish/Andalusian Music) Music Education Technology Notable Projects: CompMusic (ERC Advanced Grant, 2010-2017): Multicultural computational music analysis Open datasets: Freesound, Saraga, FSD50K Technologies: Reactable, Vocaloid, Essentia API Recent Trends in Articles: Focus on AI-driven audio processing (neural fingerprints, generative models), cross-cultural music analysis, and explainable music difficulty estimation. Awards: ERC Advanced Grant (2010) for CompMusic Project. Labs/Teams: Director of MTG, Phonos Foundation, and UPF-BMAT Chair. Active in open-source projects and international collaborations.
Changxi Zheng is an Associate Professor in the Department of Computer Science at Columbia University's School of Engineering and Applied Science (SEAS). He directs Columbia's Computer Graphics Group (C2G2) within the Columbia Vision and Graphics Center (CVGC). After receiving his PhD from Cornell University, he joined the faculty of Computer Science Department at Columbia, where he has established himself as a leading researcher in computer graphics and scientific computing. Dr. Zheng's research spans multiple areas of applied computer science with a particular focus on computer graphics and scientific computing. His work centers around developing numerical models for simulating physical phenomena involving complex motions such as fluids, bubbles, and thin rods, along with their resulting acoustic waves. Leveraging computational insights from these models, he devises methods for improving tangible object creation, enabling novel human-computer interactions, and developing software tools for acoustic and photonic devices. His research has attracted significant public interest and media coverage, including projects like FontCode, AirCode, and Computational Metallophone Design. His recent publications reveal a strong interdisciplinary approach, bridging computer graphics, physics simulation, machine learning, and hardware design. His work demonstrates consistent innovation in computational methods for simulating physical phenomena and applying these techniques to practical problems in 3D printing, acoustic modeling, and interactive systems. The breadth of his research spans from fundamental physics-based simulations to practical applications in industry. Columbia SEAS Dean's Fellow (for advised students) NSF Graduate Research Fellow (for Ruilin Xu) Snap Research Fellow (for Rundi Wu) CKGSB Fellow (for Yun Fei) Adobe Research Fellow (for Gabriel Cirio) Marie Sklodowska-Curie Individual Fellow (for Rundi Wu) Best Paper Award at ACM International Conference on Multimedia (ACMMM), 2019 Dr. Zheng actively mentors a diverse group of students, including current PhD candidates and postdoctoral researchers. His research group has received support from various sources that enable their innovative work in computational graphics and physics-based simulation. He has supervised numerous successful students who have gone on to positions at leading technology companies including Adobe, Tencent, Facebook, and academic institutions. As director of Columbia's Computer Graphics Group (C2G2) within the Columbia Vision and Graphics Center (CVGC), Dr. Zheng leads a vibrant research team focused on advancing the state of the art in computer graphics, physics-based simulation, and their applications. The group maintains strong collaborations with industry partners and academic institutions worldwide, fostering an environment of innovation and practical application of theoretical concepts.
Nicolas Mathevon is a Professor at the University of Saint-Etienne, with a focus on Animal Behavior and Bioacoustics . He currently serves as Director of Studies at Ecole Pratique des Hautes Etudes and leads the ENES Team . His research explores acoustic communication across diverse species, including crocodiles, seals, penguins, and humans, with particular emphasis on vocal recognition, signal evolution, and neurobiological underpinnings. Major collaborations include Dr. T. Aubin (CNRS), Dr. I. Charrier (marine mammals), Dr. N. Grimault (crocodilian bioacoustics), and Prof. D. Reby (human communication). He also co-edits a new book, The Voices of Nature , published by Princeton University Press. His recent work spans ecoacoustic monitoring for conservation, multi-modal communication in crocodiles, and cross-species analysis of distress signals in bonobos, chimps, and humans. Neurobiological studies include brain activation patterns in response to vocalizations and sound localization mechanisms in reptiles. Scientific Awards include senior membership in the Institut universitaire de France .
Alanson Sample is an Associate Professor in the Department of Electrical Engineering and Computer Science at the University of Michigan (since 2018), leading the Interactive Sensing and Computing Lab. His research focuses on Human-Computer Interaction, Cyber-Physical Systems, and wireless technology innovations. He holds a PhD in Electrical Engineering from the University of Washington (2011), with prior postdoctoral work there developing implantable medical devices. Before academia, Sample held senior roles at Disney Research (2013-2018) as Executive Lab Director and Principal Research Scientist, leading teams in Robotics, AI, Computer Vision, and HCI. Earlier, he worked at Intel Labs (2008-2013) on energy harvesting for wearables and IoT. Key research themes include wireless power systems, embedded sensing, and privacy-aware technologies. Over 150+ publications span topics like magnetic sensing (MagDesk), acoustic gesture recognition (HandSAW), and privacy-preserving activity tracking (PrivacyMic). His work bridges academic and industry innovation, particularly in scalable interactive systems and medical technology. Notable projects include the Quasistatic Cavity Resonance wireless charging system, RFID-based sensing platforms (WISP), and wearable health monitoring devices. His lab develops hardware-software co-design frameworks for energy-efficient embedded systems. Current research emphasizes medical applications, smart infrastructure, and ubiquitous computing interfaces.
Michele Zorzi is a Professor of Telecommunications at the School of Engineering, University of Padova, Italy, where he has held a faculty position since 2003. He leads the SIGNET (Signal processing and Networking) Research Group, focusing on cutting-edge wireless networking challenges including mmWave communications and underwater networks. His extensive publication record exceeds 600 papers in top-tier journals and conferences, reflecting significant contributions to the field through both theoretical and experimental work. He received his Laurea Degree (1990) and Ph.D. (1994) in Electrical Engineering from the University of Padova. Prior academic appointments include Politecnico di Milano (1993-1996), University of California San Diego (1995-1998), and University of Ferrara (1998-2003), where he progressed from Associate Professor to full Professor. His educational trajectory demonstrates deep roots in Italian academia with international exposure. Professor Zorzi's research spans wireless communications and networking, with current emphases on mmWave networking for vehicular systems, underwater acoustic/optical communications, non-terrestrial networks, and AI-driven networking solutions. His group conducts experimental validations including at-sea trials for underwater systems and testbeds for vehicular networks. Key projects include PRATA for predictive QoS in autonomous driving and IoT-based environmental monitoring of the Venice Lagoon, demonstrating practical applications of theoretical work. Analysis of his 2022-2025 publications reveals strong trends in applying artificial intelligence to networking challenges across diverse environments. There is significant emphasis on vehicular networks (predictive QoS, teleoperated driving), underwater systems (acoustic/optical communications, AUV swarms), and satellite networks (Starlink integration, NTN security). Experimental validation in real-world scenarios like the Venice Lagoon monitoring project and underwater sea trials characterizes his applied research approach. His scientific accolades include: IEEE Fellow (2007) IEEE Communications Society Best Tutorial Paper Award (2008, 2019) Stephen O. Rice Best Paper Award (2018) Multiple best paper awards at IEEE conferences (2005-2020) As principal investigator for numerous European and US research projects plus 20+ industry-funded initiatives, Professor Zorzi has mentored over 35 PhD students and post-docs. Graduates now hold prominent positions at institutions including Stanford, UCSD, CTTC, and Huawei. His SIGNET group maintains active international collaborations and contributes to open-source networking tools via GitHub, demonstrating commitment to community engagement. The SIGNET Research Group, housed within the Department of Information Engineering, operates specialized experimental facilities for mmWave and underwater communications. Current initiatives include AI-based predictive QoS frameworks for vehicular networks, underwater optical communication systems using ultraviolet light, and large-scale IoT deployments for environmental monitoring. The group's GitHub presence indicates strong open-science practices, while recent sea trials confirm hands-on experimental capabilities beyond theoretical work.
David Lindlbauer is an Assistant Professor at the Human-Computer Interaction Institute (HCII) of Carnegie Mellon University, where he leads the Augmented Perception Lab and co-directs the CMU Extended Reality Technology Center. His research bridges human perception, extended reality (AR/VR), and computational interaction techniques, focusing on developing systems that dynamically adapt interface elements based on environmental context, user cognition, and task requirements. He completed his PhD at TU Berlin under Prof. Marc Alexa and held a postdoctoral position at ETH Zurich's Advanced Interactive Technologies lab. His work has been published extensively at top venues including ACM CHI, UIST, and IEEE VR, with research themes spanning gaze tracking, spatial audio optimization, haptic feedback, and multimodal notification systems. Media outlets like MIT Technology Review and Fast Company Design have featured his innovations. Dr. Lindlbauer has received prestigious grants from Meta, NSF, and ETH Zurich, and serves on program committees for CHI, UIST, and ISMAR. He has been recognized with Best Paper awards at ISS 2023 and CHI 2016, and his lab develops tools like MineXR for personalized XR interfaces and RealityReplay for temporal change visualization in mixed reality environments.