Shrikanth (Shri) Narayanan is a University Professor and holder of the Niki and Max Nikias Chair in Engineering at the University of Southern California (USC), serving as the inaugural Vice President for Presidential Initiatives. He leads the Signal Analysis and Interpretation Lab (SAIL) and holds joint appointments in Computer Science, Linguistics, Psychology, Neuroscience, Pediatrics, and Otolaryngology-Head and Neck Surgery. His research focuses on speech and audio processing, behavioral signal processing, and real-time MRI of speech production, with applications in healthcare, education, and technology. Education: B.E. in Electrical Engineering from College of Engineering, Guindy (Chennai, India, 1988); M.S., Engineer, and Ph.D. in Electrical Engineering from UCLA (1990, 1992, 1995). Research interests span computational linguistics, machine learning, and multimodal human behavior analysis. He pioneered technologies for speech biomarkers in mental health, real-time MRI of speech production, and wearable sensor systems for longitudinal health studies. His work in speech emotion recognition, forensic interviews, and clinical applications has been recognized through over 40 awards, including the IEEE Flanagan Award and ISCA Medal. He has published extensively in journals like Proceedings of the IEEE , Journal of the Acoustical Society of America , and PLOS One . Key Grants: NSF CAREER, Okawa Research, IBM Faculty, Google/Amazon awards. Labs/Teams: Signal Analysis & Interpretation Lab (SAIL), USC Information Sciences Institute (ISI), Google Visiting Faculty Researcher.
Professor Bernd Möbius is a leading academic in Phonetics and Phonology at the Department of Language Science and Technology, Saarland University. His research bridges phonetic theory with speech technology applications, focusing on text-to-speech systems, prosody modeling, and computational simulations of speech processes. Current research projects: DFG SFB 1102, C1: Information density and phonetic structure predictability DFG SFB 1102, C4: Slavic intercomprehension and surprisal theory (INCOMSLAV) Research Themes: Key areas include text-to-speech synthesis, speech prosody analysis, experimental methods in speech production/perception, information density in phonetics, and cross-linguistic studies of Slavic-Germanic languages. Scientific Contributions: Recent work explores Parkinson-induced dysarthria detection, breath noise acoustics, surprisal-driven speech behaviors, multilingual BERT models for idiomaticity, and perceptual consequences of acoustic adjustments.
Meredith Tamminga is an Associate Professor of Linguistics at the University of Pennsylvania, where she directs the Language Variation and Cognition Lab . Her work bridges sociolinguistics , psycholinguistics , and theoretical linguistics , focusing on how linguistic variability is represented in mental grammar and how social context influences speech perception and production. She co-leads the Philadelphia Signs Project , exploring sign language sociolinguistics with collaborators at Penn and Gallaudet University. PhD in Linguistics (2014), University of Pennsylvania BA in Linguistics (2009), McGill University Her research integrates experimental methods with naturalistic speech analysis to study individual and group-level language change. Keywords include phonetics-phonology mapping , intraspeaker variation , and quantitative modeling . Notable grants include NSF awards and SAS Dean’s Mentorship Award . Key affiliations include MindCORE , Penn’s hub for integrative mind sciences, and collaborative ties to the Phonetics Lab , Child Language Lab , and Cultural Evolution of Language Lab . Her lab emphasizes cross-departmental research and welcomes undergraduate RAs and PhD students.
Simon King is a Professor of Speech Processing at the University of Edinburgh , affiliated with the School of Philosophy, Psychology and Language Sciences . He serves as Director of the Centre for Speech Technology Research (CSTR) and teaches courses like Speech Processing and Speech Synthesis , while directing the MSc in Speech and Language Processing . Research Interests His research focuses on: Developing new acoustic models (e.g., Linear Dynamical Models, factorial-HMMs) for speech recognition Advancing unit selection and HMM-based speech synthesis Integrating articulatory measurement data for enhanced modeling Exploring perceptual measures in synthesis criteria Building multilingual speech systems to identify universal speech building blocks Publication Trends Simon's recent work emphasizes deep learning (DNNs, LSTMs) in speech synthesis, multilingual frameworks , and articulatory-acoustic feature integration . His studies often bridge grapheme-based modeling , perceptual error reduction , and noise-robust synthesis . Scientific Awards EPSRC Advanced Research Fellowship (2005-2009) Students & Collaborations He has supervised numerous PhD students including Rasmus Dall, Tom Merritt, and Srikanth Ronanki. Current research fellows like Mirjam Wester and Zhizheng Wu contribute to projects such as Natural Speech Technology (NST) and Simple4All .
Jens Edlund is a Professor at KTH Royal Institute of Technology's Division of Speech, Music and Hearing. His research focuses on speech technology, dialogue systems, prosody, and evolutionary phonetics. He has contributed to foundational work on speech synthesis, conversational interaction, and multimodal corpora like the D64 corpus. Key projects include the MonAMI Reminder system and analysis of primate vocalizations to understand speech evolution. Edlund has collaborated extensively with global researchers, producing over 150 peer-reviewed works. His work integrates computational methods with linguistic and biological insights, emphasizing human-like dialogue systems and cross-species vocal analysis. Education: Ph.D. in Speech Technology (2011, KTH) Grants: Multiple EU and Swedish Research Council grants for speech technology and interdisciplinary studies Research labs include the KTH Speech, Music and Hearing Lab and collaborations with institutions like Max Planck Institute for Evolutionary Anthropology. Current work explores evolutionary origins of speech biomechanics and AI-driven speech synthesis evaluation.
Mark Gales is Professor of Information Engineering at the University of Cambridge and an Official Fellow at Emmanuel College. He is currently on sabbatical leave for the 2024/25 academic year. Prior to his academic career, he worked as a consultant at Roke Manor Research Ltd, developing radar systems, before transitioning to speech and language processing. PhD in 'Model-Based Techniques for Robust Speech Recognition' (University of Cambridge, 1995) BA in Electrical and Information Sciences (University of Cambridge, 1988) His research focuses on speech and language processing , particularly in automated language assessment and low-resource speech technology . He leads the Automated Language Teaching and Assessment (ALTA) Institute , which collaborates with Cambridge University Press & Assessment (CUP&A) to develop commercial tools like Linguaskill and Speak & Improve . These platforms provide automated spoken/written assessment for millions of users globally. Recent publications highlight his work in LLM-driven speech processing , including adversarial attacks on foundation models, end-to-end spoken error correction, and uncertainty estimation frameworks. His team's research spans multilingual capabilities, with deployments in languages ranging from Dholuo to Tok Pisin . Awards : IEEE Fellow, ISCA Fellow Leadership : Fellows' Steward at Emmanuel College Mark has contributed extensively to Hidden Markov Model (HMM) applications in speech recognition, which underpinned early automatic speech systems. His work now bridges LLM-based language assessment with cross-lingual transfer learning and robustness testing for real-world deployments.
Jianjing Kuang is an Associate Professor of Linguistics at the University of Pennsylvania, affiliated with the School of Arts and Sciences. As Director of the Penn Phonetics Laboratory, they lead research in phonetics, laboratory phonology, and tonal language studies. Kuang holds a Ph.D. from UCLA (2013) and specializes in the interplay between production and perception in speech, particularly focusing on tonal systems, prosody, and cross-linguistic fieldwork. Their work integrates behavioral experiments, corpus studies, and computational modeling to explore phonological contrasts, voice quality, and sound change. Key affiliations include MindCORE and the Center for East Asian Studies. Research interests include multidimensional cues in tone processing, glottal articulations, mapping production-perception relationships, and prosodic sentence processing. Kuang has conducted fieldwork on languages such as Yi, Q’anjob’al, and Mandarin. Their studies address topics like tonal splitting in Yi, cue-changing in Korean stops, and prosodic patterns in Mayan languages. Recent work explores voice quality’s role in pitch perception and prosodic boundary detection. Awards and grants are not explicitly listed, but Kuang’s extensive publications (over 50 peer-reviewed papers) reflect sustained academic impact. They advise graduate students in phonetics and phonology, with active collaborations in computational linguistics and speech technology. The Penn Phonetics Laboratory, under Kuang’s direction, emphasizes interdisciplinary research bridging experimental and theoretical phonetics.
Tara McAllister is an Associate Professor and Director of the Doctoral Program in Communicative Sciences and Disorders at New York University’s Steinhardt School. She leads the Biofeedback Intervention Technology for Speech (BITS) Lab , focusing on speech learning mechanisms and biofeedback treatments for speech disorders. Her work emphasizes acoustic and ultrasound biofeedback efficacy in resolving residual speech sound disorders, particularly in children. McAllister directs development of the staRt iOS app, expanding access to biofeedback training. She holds degrees from Harvard, MIT, and Boston University, with clinical expertise in speech-language pathology. Education: A.B./A.M., Linguistics, Harvard University (2003) M.S., Communication Disorders, Boston University (2007) Ph.D., Linguistics, MIT (2009) Research Interests: Speech motor control, perception-production links, bilingual phonological development, and technology-driven interventions. Her NIH-funded studies investigate biofeedback applications for speech disorders and crowdsourcing methodologies for perceptual analysis. Grants & Labs: NIH/NIDCD-funded BITS Lab research staRt app development since 2014 Teaching: Courses include Critical Evaluation of Research and Speech Science Instrumentation , emphasizing evidence-based practices in communication sciences.
John Kingston is a Professor of Linguistics and Director of the Phonetics Lab at the University of Massachusetts Amherst, where he has been since 1990. He holds a BA and MA from the University of Chicago (1976–1977) and a PhD from UC Berkeley (1985). His research focuses on the interplay between phonetics and phonology, particularly speech perception and its influence on phonological representations. He co-founded the Laboratory Phonology Conference series in 1987 and has conducted fieldwork on Otomanguean languages. His work emphasizes experimental methods to study phonological questions, including studies on vowel perception, tone systems, and cross-linguistic phonetic patterns. Kingston’s academic journey includes roles at the University of Texas, Austin (1984–1986) and Cornell University (1986–1990). His research explores how auditory processing and linguistic knowledge shape speech perception, with notable contributions to understanding tonogenesis, perceptual contrast effects, and vowel category learning in second languages. He collaborates on grants examining Ganong effects and phonological inventories, advocating for theories that bridge perceptual and structural aspects of language. His lab, the Phonetics Lab, supports experimental work on speech perception and production. Kingston is also the Honors Program Coordinator, mentoring students in linguistics and related fields. Despite no explicit awards listed, his extensive publications and conference leadership reflect his scholarly impact.
Jonathan David Bobaljik is a leading linguist whose research spans formal syntax, phonology, and the documentation of endangered languages, particularly Itelmen. He is actively involved in multiple international research collaborations and regularly presents at major linguistic conferences across Europe and North America. His work is published in top journals such as Natural Language and Linguistic Theory and Language . His research interests lie at the intersection of theoretical and descriptive linguistics. He investigates universal patterns in morphology and syntax, including suppletion, vowel harmony, ergativity, and control structures, often using data from understudied languages like Itelmen. His work combines rigorous formal analysis with deep empirical grounding in language documentation. The recent publications reflect a strong trend toward integrating typological diversity with formal theory. Key areas include morphosyntactic universals (e.g., suppletion in pronouns), information structure in verb-final languages, and the phonology of ejective consonants. There is also a growing emphasis on sociolinguistic and community-based aspects, as seen in collaborative work on language revitalization and the social life of the Itelmen language. He has received recognition through numerous scholarly publications and invited talks, though specific awards are not listed in the provided text. His collaborative projects suggest active grant involvement, particularly in language documentation and cross-linguistic typology. Bobaljik plays a central role in the Itelmen documentation and revitalization project, contributing to the creation of a dictionary app, the publication of historical manuscripts, and public exhibits on Siberian Indigenous knowledge. He collaborates with a team including David Koester, Chikako Ono, Tatiana Degai, and Maria Pupynina.
Zainab Hermes is an Associate Instructional Professor in the Department of Near Eastern Studies at the University of Chicago, where she has taught since 2018. She holds a PhD in Linguistics from the University of Illinois at Urbana-Champaign (2018), earned through experimental research combining phonetics and neuroscience methodologies. Education PhD in Linguistics, University of Illinois at Urbana-Champaign (2018) Research Interests Her work focuses on experimental linguistics, particularly the articulation of Arabic dialects (Cairene, Lebanese, Jordanian, Saudi). Key projects involve: MRI studies of vocal tract configurations for back consonants MEG investigations of neurological representations during Arabic diglossic code-switching Self-paced reading studies on L2 acquisition of Arabic grammatical gender Publications & Methodologies Her research employs advanced imaging techniques (rt-MRI, EMA) and neurolinguistic methods (MEG) to analyze speech production and processing. Her teaching includes courses on Arabic linguistics, dialectology, and academic reading. Academic Contributions She collaborates on interdisciplinary studies bridging phonetics, neuroscience, and language acquisition frameworks.
Ricardo Gutierrez-Osuna is a Professor in the Department of Computer Science and Engineering at Texas A&M University, part of the College of Engineering. He leads the PSI Lab and focuses on machine learning, speech processing, and digital health applications. His research spans topics like wearable sensors, foreign accent conversion, and physiological monitoring. Education: Ph.D. (Computer Engineering, NC State, 1998), M.S. (Computer Engineering, NC State, 1995), B.S. (Electrical Engineering, Universidad Politécnica de Madrid, 1992). Research interests include intelligent sensors, speech processing, machine learning, neuromorphic computation, and mobile robotics. His work bridges computer science and biomedical engineering, with applications in health monitoring and human-computer interaction. Awards: NSF CAREER Award (2002) Ramón y Cajal Award (2005-2010) Texas A&M Barbara and Ralph Cox Fellow (2009) Multiple teaching awards (2009-2010) His lab develops innovative technologies like stress-detecting wearables, biofeedback games, and systems for non-native speech improvement. He collaborates on projects involving voice conversion, glucose prediction algorithms, and multi-modal sensing devices.
Prof. Marianne Pouplier is a Professor at the Institute of Phonetics and Speech Processing (IPS) at Ludwig Maximilian University of Munich. Her research focuses on speech production mechanisms, particularly coarticulation, phonetic universals, and language-specific variations. She investigates articulatory timing, speech errors, and cross-linguistic differences in consonant clusters using advanced methodologies like real-time MRI and electromagnetic articulography. Her work bridges theoretical models (e.g., Articulatory Phonology) with empirical data, emphasizing the interplay between phonological representation and motor execution. Key contributions include studies on nasal coarticulation, larynx dynamics, and the role of articulatory effort in speech production. Selected publications highlight her expertise in analyzing speech motor control through interdisciplinary approaches. She collaborates internationally, contributing to projects like the Bavarian Archive for Speech Signals (BAS). No awards or grants are explicitly listed, but her extensive publication record underscores her scholarly impact.
Georgia Zellou is an Associate Professor in the Department of Linguistics at the University of California, Davis, where she co-directs the Phonetics Lab and conducts award-winning research at the intersection of phonetics, speech perception, and human-AI interaction. Her work investigates how phonetic detail is cognitively represented through variations in speech production, with significant contributions to understanding speech alignment with voice assistants, face-masked speech intelligibility, and cross-linguistic perception of synthetic voices. Her academic credentials include a Ph.D. in Linguistics from the University of Colorado at Boulder (2012), an M.A. in Linguistics from Stony Brook University (2007), and a B.A. in Linguistics & Anthropology from the University of Florida (2005, Cum Laude, Phi Beta Kappa). Ph.D., Linguistics, University of Colorado at Boulder (2012) M.A., Linguistics, Stony Brook University (2007) B.A., Linguistics & Anthropology, University of Florida (2005) Professor Zellou's research program centers on laboratory phonology approaches to real-world communication challenges, examining how acoustic-phonetic details influence speech perception across contexts. Her studies span speech alignment with voice-AI systems (e.g., Amazon Alexa), sociophonetic variation in bilingual speech, and the cognitive mechanisms underlying perceptual compensation for coarticulation. She employs experimental methods including eye-tracking, acoustic analysis, and perceptual testing to uncover how phonetic variation functions pragmatically in human communication and human-machine interaction. Analysis of her 15 most recent publications (2023-2025) reveals three dominant research trajectories: (1) human-AI voice interaction dynamics, including prosodic alignment and social evaluation of TTS voices; (2) intelligibility optimization in challenging contexts (face masks, clear speech for diverse listeners); and (3) cross-linguistic phonetic variation in vowelless words and consonant clusters. These works consistently bridge theoretical phonology with applied speech technology, demonstrating how fine-grained phonetic detail influences communication effectiveness in both human-human and human-machine contexts. Her scientific recognition includes: Fulbright Scholar (2022) for research in France Chancellor’s Award for Excellence in Undergraduate Mentoring (2019) Fellow of the Linguistic Society of America (2020) Amazon Faculty Research Award (2019) for Alexa-related speech studies Dean’s Fellow designation at UC Davis (2020-2023) Professor Zellou maintains an active mentoring practice recognized with the Chancellor’s Award, supervising undergraduate researchers in the Phonetics Lab while teaching core linguistics courses from introductory to advanced graduate levels. Her research program is supported by competitive grants including NSF funding, Amazon Research Awards, and UC Davis internal grants (Hellman Foundation, ISS Junior Faculty Grant), reflecting the translational value of her work for speech technology development. She has co-directed major initiatives including the 2019 LSA Linguistic Institute. The Phonetics Lab she co-leads serves as a hub for experimental phonetics research, focusing on speech production-perception relationships through projects investigating vocal accommodation to voice assistants, nasal coarticulation dynamics, and cross-linguistic prosody. Current collaborations with industry partners aim to implement human speech adaptation principles into voice assistant design to enhance naturalness and engagement.
Max Planck Institute of Colloids and InterfacesGermany
Shrikanth (Shri) Narayanan is University Professor and Niki & C. L. Max Nikias Chair in Engineering at the University of Southern California (USC), with appointments spanning Electrical & Computer Engineering, Computer Science, Linguistics, Psychology, Neuroscience, Pediatrics, and Otolaryngology-Head & Neck Surgery. He serves as Research Director of the Information Sciences Institute and Director of the Ming Hsieh Institute. PhD in Electrical Engineering (UCLA, 1995) Engineer and MS in Electrical Engineering (UCLA, 1992 and 1990) BE in Electrical Engineering (Anna University, India, 1988) His interdisciplinary research focuses on human-centered signal processing and machine intelligence , addressing societal challenges in health, education, defense, and media arts. Key areas include: Behavioral signal processing Affective computing Multimodal signal processing Computational speech science Biomedical applications Scientific Awards : IEEE James L. Flanagan Speech and Audio Processing Award (2025) Edward J. McCluskey Technical Achievement Award (2024) ISCA Medal for Scientific Achievement (2023) Claude Shannon-Harry Nyquist Technical Achievement Award (2023) ACM ICMI Sustained Accomplishment Award (2020) USC Distinguished Faculty Service Award With over 1,000 publications and 19 patents , his work has been commercialized through startups like Behavioral Signals Technologies and Lyssn . He leads transformative university initiatives and has served in editorial roles for top journals including Computer Speech and Language and IEEE Transactions on Affective Computing .