Mark Y. Liberman is the Christopher H. Browne Distinguished Professor of Linguistics and Trustee Professor at the University of Pennsylvania. He holds a joint appointment in the Department of Linguistics and the Department of Computer and Information Science. His roles include Director of the Linguistic Data Consortium (LDC), Faculty Director of Ware College House, and former Director of the Institute for Research in Cognitive Science. Education: A.B. in Linguistics and Applied Mathematics from Harvard University (1965–1969), M.S. (1972) and Ph.D. (1975) in Linguistics from MIT. Research focuses on corpus-based phonetics, clinical linguistics applications, tonal phonology, formal models for linguistic annotation, and computational linguistics. He explores speech production, prosody, and interdisciplinary topics like language evolution and neurobiology of speech. Recent articles highlight advancements in speech biomarkers for neurodegenerative diseases, autism analysis, and computational linguistics. Awards include Fellowships from the AAAS and Linguistic Society of America. He advises graduate students and leads large-scale language resource initiatives like LDC, contributing to open-access linguistic datasets. Labs/Teams: Linguistic Data Consortium (LDC), Institute for Research in Cognitive Science (IRCS), and collaborations in computational linguistics and neuroscience.
Jan Østergaard is a Full Professor in Information Theory and Signal Processing at Aalborg University's Department of Electronic Systems. He leads the AI and Sound research section and directs the CASPR center. His expertise spans AI-driven acoustic signal processing, information theory, and EEG signal analysis. Østergaard holds a M.Sc. from Aalborg University and a PhD (cum laude) from Delft University of Technology. Major awards include the Danish Young Researcher’s Award and a EURASIP Best Thesis honor. His work focuses on speech enhancement, sound zone technologies, and neural tracking of auditory attention. Recent research emphasizes low-latency speech transmission, deep learning for sound field control, and robust voice activity detection. He serves on editorial boards and national committees, advancing Denmark’s sound technology initiatives. Education: M.Sc. (Aalborg, 1999), PhD (Delft, 2007) Research interests emphasize practical AI applications in sound systems, including hearing aid improvements, data-efficient acoustic modeling, and feedback control in networked systems. Over 210 publications and 17 active projects reflect his interdisciplinary impact across academia and industry.
Mark Steedman is a Professor in the School of Informatics at the University of Edinburgh, where he conducts research in Artificial Intelligence, Computational Cognitive and Social Science, and Natural Language and Speech Processing. He is affiliated with the Institute for Language, Cognition and Computation (ILCC), the Centre for Speech Technology Research (CSTR), and the Human Communications Research Center (HCRC). He also holds an adjunct professorship in Computer and Information Science at the University of Pennsylvania. His research focuses on Combinatory Categorial Grammar (CCG) , computational linguistics , prosody and intonation , temporal semantics , gesture in communication , and computational music analysis . He has authored foundational books including Surface Structure and Interpretation , The Syntactic Process , and Taking Scope . The recent publications reflect a strong trend toward integrating formal grammatical frameworks like CCG with modern neural and distributional models, particularly in semantic parsing, entailment reasoning, and cognitive modeling. His work bridges symbolic and statistical approaches in NLP, often focusing on robust, wide-coverage parsing and semantic interpretation. Best Paper Award at AACL/IJCNLP 2023 for 'Smoothing Entailment Graphs with Language Models' Best Paper Award at ACL 2023 for 'Extrinsic Evaluation of Machine Translation Metrics' Influential Paper Award 2017 from IFAAMAS for 'Animated Conversation' Mark Steedman has supervised numerous PhD students and collaborated widely across institutions. He leads research in formal grammar applications to cognitive modeling, dialogue, and multimodal communication. His lab contributes to CCG software and semantic parsing tools, and he continues to be actively involved in advancing the integration of symbolic and neural AI.
Lior Wolf is a Professor at the School of Computer Science, Tel Aviv University. Previously, he was a postdoctoral researcher at MIT's Center for Biological and Computational Learning (CBCL) under Prof. Tomaso Poggio and earned his PhD from Hebrew University of Jerusalem with Prof. Amnon Shashua. His educational background includes: PhD in Computer Science, Hebrew University of Jerusalem Postdoctoral Research, MIT CBCL Prof. Wolf's research centers on artificial intelligence with seminal contributions to deep learning, computer vision, and natural language processing. His work bridges theoretical foundations (e.g., attention mechanisms, transformer analysis) with practical applications in medical imaging, speech processing, and sign language technology. He pioneered methods for neural network interpretability, efficient sequence modeling, and multimodal fusion. Analysis of his 2023-2025 publications reveals dominant trends in large language model optimization (neuron pruning, attention analysis), efficient video generation, and cross-modal learning. His work increasingly integrates medical applications (fMRI/EEG analysis) while maintaining theoretical rigor in model architecture design. His scientific achievements include: Best paper award at EMNLP 2024 for 'Backward Lens: Projecting Language Model Gradients into the Vocabulary Space' Best paper award at SCIA 2023 for 'Gradient Adjusting Networks for Domain Inversion' Best paper award at FG 2021 for 'Generating Master Faces for Dictionary Attacks' Prof. Wolf mentors graduate students in the School of Computer Science and leads research at the ICRC building laboratory. His team collaborates with 'the friends of TAU' on projects spanning biometric security, medical imaging, and generative AI. Current work focuses on efficient transformers, neural network interpretability, and multimodal medical diagnostics.
Simon King is a Professor of Speech Processing at the University of Edinburgh , affiliated with the School of Philosophy, Psychology and Language Sciences . He serves as Director of the Centre for Speech Technology Research (CSTR) and teaches courses like Speech Processing and Speech Synthesis , while directing the MSc in Speech and Language Processing . Research Interests His research focuses on: Developing new acoustic models (e.g., Linear Dynamical Models, factorial-HMMs) for speech recognition Advancing unit selection and HMM-based speech synthesis Integrating articulatory measurement data for enhanced modeling Exploring perceptual measures in synthesis criteria Building multilingual speech systems to identify universal speech building blocks Publication Trends Simon's recent work emphasizes deep learning (DNNs, LSTMs) in speech synthesis, multilingual frameworks , and articulatory-acoustic feature integration . His studies often bridge grapheme-based modeling , perceptual error reduction , and noise-robust synthesis . Scientific Awards EPSRC Advanced Research Fellowship (2005-2009) Students & Collaborations He has supervised numerous PhD students including Rasmus Dall, Tom Merritt, and Srikanth Ronanki. Current research fellows like Mirjam Wester and Zhizheng Wu contribute to projects such as Natural Speech Technology (NST) and Simple4All .
WANG Ye is an Associate Professor in the Department of Computer Science at the School of Computing, National University of Singapore (NUS). He holds a PhD in Information Technology from Tampere University of Technology, Finland, and has been a tenured faculty member at NUS since 2002, following his industry research role at Nokia Research Center. He is the director of the Sound and Music Computing Lab at NUS, leading cutting-edge research in AI-driven music and health technologies. PhD, Information Technology, Tampere University of Technology, Finland (2002) MSc, Telecommunications, Braunschweig University of Technology, Germany (1993) BSc, Telecommunications, South China University of Technology, China (1983) His research is centered on Sound and Music Computing for Human Health and Potential (SMC4HHP) , with a focus on eHealth, eLearning, mobile/wearable computing, and music information retrieval. His work spans AI for stroke rehabilitation, language learning through singing, singing voice synthesis, and automatic music transcription. He has pioneered systems like SLIONS (language learning via karaoke), CocoLyricist (AI co-creation for stroke recovery), and SinTechSVS (expressive singing voice synthesis). The latest articles highlight a strong trend in AI-driven music and health technologies , particularly in controllable lyric generation, singing voice synthesis, automatic pronunciation assessment, and multimodal music transcription. The research increasingly integrates large language models, explainable AI, fairness, and real-world deployment, reflecting a shift from theoretical exploration to practical, human-centered applications in healthcare and education. Dr. Wang has received numerous scientific honors, including: Best Paper Awards at ACM MM, ISMIR, IEEE ISM, and CHI First Prize, Asia Pacific Assistive, Rehabilitative, and Therapeutic Technologies Challenge (2015) Faculty Teaching Excellence Award, NUS School of Computing (2024) Top Paper Award, ACM Multimedia 2022 AI in Medicine Collaborative Grant for CocoLyricist project He has supervised over 11 PhD and 20 MComp students and is currently guiding six PhD candidates. His grants come from MOE, NRF, A*STAR, Nokia, and Smule. He has served as General Chair of ISMIR2017 and TPC Co-Chair of ICOT2017, and is on the editorial boards of IEEE Transactions on Multimedia and Journal of New Music Research. He has also developed and taught the first course on Sound and Music Computing in Singapore. Dr. Wang leads the Sound and Music Computing Lab (SMC Lab) , a multidisciplinary team exploring the synergy of music computing, AI, mobile technology, and cloud systems for health and education. The lab actively collaborates with medical institutions such as NUS Yong Loo Lin School of Medicine, Singapore General Hospital, and Harvard Medical School, and is currently working on projects in AI-supported language learning, stroke rehabilitation, and intelligent music interfaces.
Catherine Lai is a Reader (~Associate Professor) in the Department of Linguistics and English Language at the University of Edinburgh, with strong affiliations to the Centre for Speech Technology Research (CSTR) and the Institute for Language, Cognition and Computation (ILCC) in the School of Informatics. She is based in the School of Philosophy, Psychology and Language Sciences and is actively involved in research, teaching, and academic service. Department: Department of Linguistics and English Language School: School of Philosophy, Psychology and Language Sciences Research Institutes: Centre for Speech Technology Research, Institute for Language, Cognition and Computation Email: C.Lai@ed.ac.uk Her research centers on the role of prosody—non-lexical aspects of speech—in spoken communication. She investigates how prosody contributes to discourse structure, information structure, and affect in dialogue, using interdisciplinary methods from linguistics and machine learning. Her work bridges theoretical linguistics and practical speech technology, aiming to improve spoken language understanding and synthesis systems. She is particularly interested in how prosody shapes listener expectations and how affect and topic are expressed and perceived in conversation. Her recent publications reflect a strong focus on self-supervised learning in speech models, emotion recognition, ASR error correction using large language models, cognitive state classification, and ethical considerations in language technology. She explores topics such as the uncanny valley in synthetic speech, gender expression through voice, and community-centered development of language technologies. Prize from Scopus Profile Catherine Lai has supervised several PhD students, including Leimin Tian and Yuanchao Li, and has been involved in significant research projects, such as a Toyota-funded initiative on spoken dialogue for robot companions. She has secured multiple grants and leads a research agenda that integrates theoretical inquiry with real-world applications in assistive technologies and social science. Her academic service includes organizing major conferences like Interspeech and UK and Ireland Speech. She is a key member of research teams at CSTR and ILCC, collaborating across disciplines to advance the understanding of spoken communication and the development of robust, ethical speech technologies.
Pierre Vandergheynst is a Full Professor at the Swiss Federal Institute of Technology Lausanne (EPFL) in the Department of Electrical Engineering, with a courtesy appointment in Computer and Communication Sciences. He serves as EPFL’s Vice-Provost for Education since 2015 and leads the Signal Processing Laboratory 2 (LTS2). His research spans harmonic analysis, sparse approximations, mathematical data processing, and applications in signal/image processing, computer vision, machine learning, and graph-based data analysis. PhD in Mathematical Physics (1998), Université catholique de Louvain Postdoctoral Researcher at EPFL (1998-2001) Assistant Professor at EPFL (2002-2007) His research explores geometry/symmetry in high-dimensional data, redundant dictionaries for dimensionality reduction, and computational harmonic analysis on manifolds. Recent work focuses on protein structure modeling, geometric deep learning, and graph-based signal processing. Key article trends include graph neural networks for protein analysis, geometric deep learning in neuroscience, and structured knowledge priors in neural models. His 2023-2025 publications emphasize interpretable AI, long-range dependencies in graphs, and molecular representation learning. Scientific Awards: IEEE Signal Processing Magazine Best Paper Award (2023) Signal Processing Society Best Paper Award (2022) Apple ARTS Award (2007) De Boelpaepe Prize, Royal Academy of Sciences of Belgium (2009-2010) He has supervised over 30 PhD theses and contributed to foundational work in graph signal processing, compressive sensing, and geometric deep learning. His lab develops tools for data science on non-Euclidean structures, with applications in medicine, astronomy, and wireless systems.
Professor Simo Särkkä holds a position in Sensor Informatics and Medical Technology at the Department of Electrical Engineering and Automation (EEA), Aalto University. His research focuses on multi-sensor data processing, Bayesian filtering, machine learning, and their applications in medical technology, brain imaging, and inverse problems. He leads research groups including the Helsinki Institute for Information Technology (HIIT) and Sensor Informatics and Medical Technology. His work bridges theoretical advancements in probabilistic methods with practical implementations in healthcare and engineering. Key research interests include Gaussian processes, stochastic differential equations, quantum machine learning, and signal processing. He has contributed to advancements in algorithms for nonlinear state-space models, parallel computing techniques, and medical imaging technologies such as scatter correction in CT scans. His methodologies are applied across domains like autonomous systems, robotics, and bioengineering. Notable publications span topics like quantum-assisted Gaussian regression, physics-informed machine learning for industrial processes, and parallel-in-time numerical methods. His work emphasizes computational efficiency and robustness in high-dimensional and real-time systems.
Xavier Serra is a Full Professor at the Department of Engineering at Universitat Pompeu Fabra (UPF), Barcelona. He is the founder and director of the Music Technology Group (MTG), and leads the UPF-BMAT Chair on AI and Music. He also coordinates the Master in Sound and Music Computing and serves as President of the Phonos Foundation. His research focuses on audio signal processing, sound and music computing, and computational musicology, emphasizing open science and open innovation. Education: BSc in Biology, University of Barcelona (1981) Master in Music, Florida State University (1983) PhD in Computer Music, Stanford University (1989) Research Interests: Audio Signal Processing Data-Driven and Knowledge-Driven Methodologies Music Information Retrieval Cultural Music Analysis (e.g., Carnatic/Turkish/Andalusian Music) Music Education Technology Notable Projects: CompMusic (ERC Advanced Grant, 2010-2017): Multicultural computational music analysis Open datasets: Freesound, Saraga, FSD50K Technologies: Reactable, Vocaloid, Essentia API Recent Trends in Articles: Focus on AI-driven audio processing (neural fingerprints, generative models), cross-cultural music analysis, and explainable music difficulty estimation. Awards: ERC Advanced Grant (2010) for CompMusic Project. Labs/Teams: Director of MTG, Phonos Foundation, and UPF-BMAT Chair. Active in open-source projects and international collaborations.
Changxi Zheng is an Associate Professor in the Department of Computer Science at Columbia University's School of Engineering and Applied Science (SEAS). He directs Columbia's Computer Graphics Group (C2G2) within the Columbia Vision and Graphics Center (CVGC). After receiving his PhD from Cornell University, he joined the faculty of Computer Science Department at Columbia, where he has established himself as a leading researcher in computer graphics and scientific computing. Dr. Zheng's research spans multiple areas of applied computer science with a particular focus on computer graphics and scientific computing. His work centers around developing numerical models for simulating physical phenomena involving complex motions such as fluids, bubbles, and thin rods, along with their resulting acoustic waves. Leveraging computational insights from these models, he devises methods for improving tangible object creation, enabling novel human-computer interactions, and developing software tools for acoustic and photonic devices. His research has attracted significant public interest and media coverage, including projects like FontCode, AirCode, and Computational Metallophone Design. His recent publications reveal a strong interdisciplinary approach, bridging computer graphics, physics simulation, machine learning, and hardware design. His work demonstrates consistent innovation in computational methods for simulating physical phenomena and applying these techniques to practical problems in 3D printing, acoustic modeling, and interactive systems. The breadth of his research spans from fundamental physics-based simulations to practical applications in industry. Columbia SEAS Dean's Fellow (for advised students) NSF Graduate Research Fellow (for Ruilin Xu) Snap Research Fellow (for Rundi Wu) CKGSB Fellow (for Yun Fei) Adobe Research Fellow (for Gabriel Cirio) Marie Sklodowska-Curie Individual Fellow (for Rundi Wu) Best Paper Award at ACM International Conference on Multimedia (ACMMM), 2019 Dr. Zheng actively mentors a diverse group of students, including current PhD candidates and postdoctoral researchers. His research group has received support from various sources that enable their innovative work in computational graphics and physics-based simulation. He has supervised numerous successful students who have gone on to positions at leading technology companies including Adobe, Tencent, Facebook, and academic institutions. As director of Columbia's Computer Graphics Group (C2G2) within the Columbia Vision and Graphics Center (CVGC), Dr. Zheng leads a vibrant research team focused on advancing the state of the art in computer graphics, physics-based simulation, and their applications. The group maintains strong collaborations with industry partners and academic institutions worldwide, fostering an environment of innovation and practical application of theoretical concepts.
Michele Zorzi is a Professor of Telecommunications at the School of Engineering, University of Padova, Italy, where he has held a faculty position since 2003. He leads the SIGNET (Signal processing and Networking) Research Group, focusing on cutting-edge wireless networking challenges including mmWave communications and underwater networks. His extensive publication record exceeds 600 papers in top-tier journals and conferences, reflecting significant contributions to the field through both theoretical and experimental work. He received his Laurea Degree (1990) and Ph.D. (1994) in Electrical Engineering from the University of Padova. Prior academic appointments include Politecnico di Milano (1993-1996), University of California San Diego (1995-1998), and University of Ferrara (1998-2003), where he progressed from Associate Professor to full Professor. His educational trajectory demonstrates deep roots in Italian academia with international exposure. Professor Zorzi's research spans wireless communications and networking, with current emphases on mmWave networking for vehicular systems, underwater acoustic/optical communications, non-terrestrial networks, and AI-driven networking solutions. His group conducts experimental validations including at-sea trials for underwater systems and testbeds for vehicular networks. Key projects include PRATA for predictive QoS in autonomous driving and IoT-based environmental monitoring of the Venice Lagoon, demonstrating practical applications of theoretical work. Analysis of his 2022-2025 publications reveals strong trends in applying artificial intelligence to networking challenges across diverse environments. There is significant emphasis on vehicular networks (predictive QoS, teleoperated driving), underwater systems (acoustic/optical communications, AUV swarms), and satellite networks (Starlink integration, NTN security). Experimental validation in real-world scenarios like the Venice Lagoon monitoring project and underwater sea trials characterizes his applied research approach. His scientific accolades include: IEEE Fellow (2007) IEEE Communications Society Best Tutorial Paper Award (2008, 2019) Stephen O. Rice Best Paper Award (2018) Multiple best paper awards at IEEE conferences (2005-2020) As principal investigator for numerous European and US research projects plus 20+ industry-funded initiatives, Professor Zorzi has mentored over 35 PhD students and post-docs. Graduates now hold prominent positions at institutions including Stanford, UCSD, CTTC, and Huawei. His SIGNET group maintains active international collaborations and contributes to open-source networking tools via GitHub, demonstrating commitment to community engagement. The SIGNET Research Group, housed within the Department of Information Engineering, operates specialized experimental facilities for mmWave and underwater communications. Current initiatives include AI-based predictive QoS frameworks for vehicular networks, underwater optical communication systems using ultraviolet light, and large-scale IoT deployments for environmental monitoring. The group's GitHub presence indicates strong open-science practices, while recent sea trials confirm hands-on experimental capabilities beyond theoretical work.
Mark Gales is Professor of Information Engineering at the University of Cambridge and an Official Fellow at Emmanuel College. He is currently on sabbatical leave for the 2024/25 academic year. Prior to his academic career, he worked as a consultant at Roke Manor Research Ltd, developing radar systems, before transitioning to speech and language processing. PhD in 'Model-Based Techniques for Robust Speech Recognition' (University of Cambridge, 1995) BA in Electrical and Information Sciences (University of Cambridge, 1988) His research focuses on speech and language processing , particularly in automated language assessment and low-resource speech technology . He leads the Automated Language Teaching and Assessment (ALTA) Institute , which collaborates with Cambridge University Press & Assessment (CUP&A) to develop commercial tools like Linguaskill and Speak & Improve . These platforms provide automated spoken/written assessment for millions of users globally. Recent publications highlight his work in LLM-driven speech processing , including adversarial attacks on foundation models, end-to-end spoken error correction, and uncertainty estimation frameworks. His team's research spans multilingual capabilities, with deployments in languages ranging from Dholuo to Tok Pisin . Awards : IEEE Fellow, ISCA Fellow Leadership : Fellows' Steward at Emmanuel College Mark has contributed extensively to Hidden Markov Model (HMM) applications in speech recognition, which underpinned early automatic speech systems. His work now bridges LLM-based language assessment with cross-lingual transfer learning and robustness testing for real-world deployments.
Anthony Rollett is a Professor in the Department of Materials Science and Engineering at Carnegie Mellon University , where he has been a faculty member since 1995. He serves as the Principal Investigator and Co-Director of the NASA-supported Institute for Model-Based Qualification & Certification of Additive Manufacturing (IMQCAM) and co-director of the Next Manufacturing Center . Prior to CMU, he held leadership roles at Los Alamos National Laboratory (1991-1995). Education: Ph.D., Materials Engineering, Drexel University (1987) MA, Metallurgy and Materials Science, Cambridge University (1977) Research Interests: Rollett’s work focuses on microstructural evolution and microstructure-property relationships in 3D using experiments and simulations. His expertise spans additive manufacturing , metal 3D printing , materials for energy systems , grain growth , recrystallization , and stereology , with techniques like high-energy diffraction microscopy (HEDM) and dynamic x-ray radiography (DXR) . Scientific Contributions: He has over 320 peer-reviewed publications and an h-index >80 . His recent articles highlight machine learning for laser processing , fatigue analysis of additively manufactured alloys, and design optimization for heat exchangers in supercritical CO2 and solar thermal applications . Scientific Awards: Fellow of ASM International (1996) Fellow of the Institute of Physics (UK) (2004) Fellow of The Minerals, Metals & Materials Society (TMS) (2011) Cyril Stanley Smith Award (TMS, 2014) Member of Honor, French Metallurgical Society (2015) US Steel Professor (2017) Francqui International Professor (2020-2021) International FAME Award (2023) Leadership & Impact: Rollett co-led the development of a NASA Space Technology Research Institute for additive manufacturing and established a new master’s program in additive manufacturing (2018). His research group is funded by industry , federal agencies , and Pennsylvania state grants . He also serves on the Basic Energy Science Advisory Committee and Defense Programs Advisory Committee for the Department of Energy.
Nima Mesgarani is an Associate Professor of Electrical Engineering at Columbia Engineering, Columbia University, affiliated with the Sense, Collect and Move Data Committee. His research bridges engineering and neuroscience through reverse-engineering neural signal processing mechanisms, leading to advancements in brain-machine interfaces, neural prosthetics, and speech processing algorithms. He received his PhD in Electrical Engineering from the University of Maryland and completed postdoctoral training at Johns Hopkins University's Center for Language and Speech Processing and UC San Francisco's Neurosurgery Department. Research Focus Professor Mesgarani's lab integrates computational neuroscience and engineering to study acoustic signal processing. Key areas include: Neural decoding of speech and auditory attention in multi-talker environments Development of brain-controlled hearing technologies Novel speech separation and synthesis algorithms inspired by cortical processing Cross-modal learning between auditory and visual systems Applications of large language models in neural signal interpretation Publication Trends Analysis of his 15 most recent articles (2025) reveals dominant themes: neural decoding techniques using intracranial EEG, brain-inspired speech separation models (e.g., Mamba architectures), applications of large language models in auditory neuroscience, cross-modal distillation methods, and clinical translation of audio processing algorithms. A strong emphasis emerges on real-time brain-computer interfaces and noise-robust speech processing. Laboratory and Collaborations Mesgarani directs an interdisciplinary lab developing neurotechnology for hearing restoration. His team collaborates with neurosurgery departments and speech processing centers, focusing on translating theoretical models into clinical brain-machine interfaces. The lab's work has yielded patents for brain-informed speech separation systems and attention-decoding frameworks.