Nicholas Evans is Distinguished Professor of Linguistics and Director of the ARC Centre of Excellence for the Dynamics of Language (CoEDL) at the Australian National University’s School of Culture, History & Language. His work bridges fieldwork-based language documentation with theoretical questions in typology, cultural evolution, and social cognition. Focus on endangered Australian and Papuan languages Director of ARC Laureate Project on 'The Wellsprings of Linguistic Diversity' Co-leader of SCOPIC (Social Cognition Parallax Corpus) study Collaborator in global linguistic diversity initiatives His research explores how micro-level community multilingualism shapes macro-level linguistic diversity, with fieldwork spanning seven years in remote Indigenous communities. Recent projects include PARABANK (paradigm syncretism analysis) and Southern New Guinea language studies, particularly Nen and Yam family languages. Scientific recognition includes the Ken Hale Award (Linguistic Society of America), Anneliese Maier Forschungspreis, and fellowships in the Australian Academy of Humanities, Australian Social Sciences Academy, and the British Academy.
Justin Johnson is an Assistant Professor at the University of Michigan's College of Engineering, Department of Electrical Engineering and Computer Science, and a Research Scientist at Facebook AI Research (FAIR). His work bridges computer vision, machine learning, and deep learning, focusing on visual reasoning, vision-language tasks, image generation, and 3D reasoning using neural networks. PhD, Stanford University (advised by Fei-Fei Li) His research interests span visual reasoning , vision and language , 3D vision , and image generation , with a focus on innovative applications of deep neural networks. Recent publications highlight work on 3D consistency, self-supervised learning, and multimodal integration of vision and text. Notable contributions include PyTorch3D for 3D data processing, and foundational work in visual question answering , neural style transfer , and scene graph-based image generation . Publications span top conferences like ICCV, CVPR, and NeurIPS. He teaches courses including EECS 498/598: Deep Learning for Computer Vision and EECS 442: Computer Vision at University of Michigan, with prior involvement in Stanford's CS 231N in co-teaching roles. Software projects include open-source frameworks like fast-neural-style for real-time artistic style transfer, and PyTorch3D for efficient 3D deep learning. These tools demonstrate his commitment to practical implementations and community-driven research.
Professor Bernd Möbius is a leading academic in Phonetics and Phonology at the Department of Language Science and Technology, Saarland University. His research bridges phonetic theory with speech technology applications, focusing on text-to-speech systems, prosody modeling, and computational simulations of speech processes. Current research projects: DFG SFB 1102, C1: Information density and phonetic structure predictability DFG SFB 1102, C4: Slavic intercomprehension and surprisal theory (INCOMSLAV) Research Themes: Key areas include text-to-speech synthesis, speech prosody analysis, experimental methods in speech production/perception, information density in phonetics, and cross-linguistic studies of Slavic-Germanic languages. Scientific Contributions: Recent work explores Parkinson-induced dysarthria detection, breath noise acoustics, surprisal-driven speech behaviors, multilingual BERT models for idiomaticity, and perceptual consequences of acoustic adjustments.
Prof. Matthias Nießner is a Professor at the Technical University of Munich, leading the Visual Computing Lab. His research intersects computer graphics, vision, and AI, focusing on 3D reconstruction, semantic understanding, and AI-driven video synthesis. He holds a PhD from the University of Erlangen-Nuremberg (2013) and was a Visiting Assistant Professor at Stanford University (2013–2017). Notable awards include the ERC Starting Grant (2018), Nvidia Professorship Award, and Eurographics Young Researcher Award (2019). His work has been featured in mainstream media and led to startups like Synthesia Inc. Research spans Gaussian splatting, neural radiance fields, and generative AI for 3D avatars. Over 150 publications include SIGGRAPH, CVPR, and ECCV, with best paper awards. Projects like Face2Face and ScanNet have driven innovation in facial reenactment and 3D scene datasets. Education: PhD in Computer Science, University of Erlangen-Nuremberg (2013) Diploma in Computer Science, University of Erlangen-Nuremberg (2010) Research Interests: 3D digitization, neural rendering, generative AI, non-rigid reconstruction, and applications in AR/VR. Awards: ERC Starting Grant (2018) Nvidia Professorship Award (2018) Google Faculty Award (2018) SIGGRAPH Best Emerging Tech Award (2016) Grants: Over €1.5M from ERC and industry partnerships. Labs/Teams: Visual Computing Lab at TUM and Synthesia Inc. (co-founder). Key projects include ScanNet (large 3D indoor dataset), Face2Face (real-time facial reenactment), and Gaussian-based 3D avatars. Current work focuses on diffusion models, neural radiance fields, and AI-generated media detection.
Jackie Chit Kit Cheung is an Associate Professor in the School of Computer Science at McGill University, where he co-directs the Reasoning and Learning Lab. He holds the Canada CIFAR AI Chair and serves as an Associate Scientific Co-Director at the Mila Quebec AI Institute. He is also a consulting researcher at Microsoft Research Montreal. Academic Background: Ph.D. in Computer Science, University of Toronto (2010–2014) M.Sc. in Computer Science, University of Toronto (2008–2010) B.Sc. (Honours) in Computer Science, minors in Linguistics and German, University of British Columbia (2004–2008) Research Interests: Jackie Cheung's research lies at the intersection of natural language processing, machine learning, and cognitive science. He focuses on natural language generation , automatic summarization , commonsense and pragmatic reasoning , and the evaluation of NLP systems . His work aims to build language models that reflect real-world structure and reasoning, with applications in health, education, and language revitalization. He is particularly interested in how implicit meaning is processed in context, a core concern in linguistic pragmatics. Publication Trends: His recent publications (2024–2025) show a strong emphasis on evaluation methodologies (e.g., COSMIC, ECBD), hallucination and factuality in language models, coreference and reasoning , and long-context modeling . He frequently collaborates across institutions and integrates insights from linguistics and psychology into NLP system design and analysis. Scientific Awards: Best Paper Award, ACL 2018 Outstanding Paper Award, NAACL 2018 Student Research Workshop SAC Award, ACL 2024 (for COSMIC) Best Poster Award, CMDO AI-Health Symposium 2024 Advising and Grants: He advises a large and diverse group of graduate students, including PhD and Master’s candidates, often in co-supervision with other faculty. His group has received support from major AI and health research initiatives, including CIFAR and Mila. He has trained alumni who have gone on to faculty and industry research positions. He has served in leadership roles in top NLP conferences, including as Senior Area Chair, Workshop Chair, and Program Chair of Canadian AI 2018. Labs and Teams: He co-directs the Reasoning and Learning Lab at McGill and is deeply involved with Mila – Quebec AI Institute . He founded the NLP Reading Group at McGill, which brings together researchers from computer science, linguistics, and information studies to discuss theoretical and applied NLP topics.
Swiss Federal Institute of Technology in LausanneSwitzerland
Lior Wolf is a Professor at the School of Computer Science, Tel Aviv University. Previously, he was a postdoctoral researcher at MIT's Center for Biological and Computational Learning (CBCL) under Prof. Tomaso Poggio and earned his PhD from Hebrew University of Jerusalem with Prof. Amnon Shashua. His educational background includes: PhD in Computer Science, Hebrew University of Jerusalem Postdoctoral Research, MIT CBCL Prof. Wolf's research centers on artificial intelligence with seminal contributions to deep learning, computer vision, and natural language processing. His work bridges theoretical foundations (e.g., attention mechanisms, transformer analysis) with practical applications in medical imaging, speech processing, and sign language technology. He pioneered methods for neural network interpretability, efficient sequence modeling, and multimodal fusion. Analysis of his 2023-2025 publications reveals dominant trends in large language model optimization (neuron pruning, attention analysis), efficient video generation, and cross-modal learning. His work increasingly integrates medical applications (fMRI/EEG analysis) while maintaining theoretical rigor in model architecture design. His scientific achievements include: Best paper award at EMNLP 2024 for 'Backward Lens: Projecting Language Model Gradients into the Vocabulary Space' Best paper award at SCIA 2023 for 'Gradient Adjusting Networks for Domain Inversion' Best paper award at FG 2021 for 'Generating Master Faces for Dictionary Attacks' Prof. Wolf mentors graduate students in the School of Computer Science and leads research at the ICRC building laboratory. His team collaborates with 'the friends of TAU' on projects spanning biometric security, medical imaging, and generative AI. Current work focuses on efficient transformers, neural network interpretability, and multimodal medical diagnostics.
Andrew Zisserman is a Royal Society Research Professor at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). His research focuses on computer vision, artificial intelligence, and neural networks, with significant contributions to multimodal learning, video understanding, and 3D scene analysis. He leads projects exploring visual-language models, audio-visual synchronization, and clinical imaging applications. Key research areas include: Video analysis and temporal modeling Multimodal systems for sign language translation and action recognition 3D shape estimation and physical property inference Foundation models and cross-modal retrieval Recent work highlights: Developed Flamingo and Tapir models for video-language tasks Advancements in spinal MRI analysis and clinical imaging Leadership in EGO4D and VoxCeleb challenges Honors include Fellowship of the Royal Society (FRS) and the ISSLS Prize in Clinical Science 2023 for spinal analysis innovations. His lab collaborates globally, emphasizing real-world applications in healthcare and autonomous systems.
Jiatao Gu is an Assistant Professor in the Department of Computer and Information Science (CIS) at the University of Pennsylvania, with a part-time role as Staff Research Scientist at Apple (MLR). He holds a Ph.D. in Electrical and Electronic Engineering from the University of Hong Kong (2018) and a B.Eng. in Electronic Engineering from Tsinghua University (2014). His research focuses on generative machine learning and AI agent interaction with the physical world, emphasizing multi-modal systems spanning language, images, videos, and 3D. Key themes include efficient modeling , flexible architecture design , and scalable decision-making frameworks . 2025: ICLR paper on DART framework 2024: TMLR work on GFlowNet alignment 2023: NeurIPS research on diffusion stability 2022: ACL papers on speech translation Recent publications explore diffusion models for text-to-image synthesis, 3D reconstruction, and efficient sampling techniques. His work addresses fundamental challenges in attention mechanisms, entropy collapse, and multi-stage distillation while advancing non-autoregressive translation and vision-language reasoning . Prospective students can apply through his recruitment process at UPenn. Prior affiliations include Meta AI (FAIR Labs) and academic collaborations with institutions like New York University's CILVR Lab.
Simon King is a Professor of Speech Processing at the University of Edinburgh , affiliated with the School of Philosophy, Psychology and Language Sciences . He serves as Director of the Centre for Speech Technology Research (CSTR) and teaches courses like Speech Processing and Speech Synthesis , while directing the MSc in Speech and Language Processing . Research Interests His research focuses on: Developing new acoustic models (e.g., Linear Dynamical Models, factorial-HMMs) for speech recognition Advancing unit selection and HMM-based speech synthesis Integrating articulatory measurement data for enhanced modeling Exploring perceptual measures in synthesis criteria Building multilingual speech systems to identify universal speech building blocks Publication Trends Simon's recent work emphasizes deep learning (DNNs, LSTMs) in speech synthesis, multilingual frameworks , and articulatory-acoustic feature integration . His studies often bridge grapheme-based modeling , perceptual error reduction , and noise-robust synthesis . Scientific Awards EPSRC Advanced Research Fellowship (2005-2009) Students & Collaborations He has supervised numerous PhD students including Rasmus Dall, Tom Merritt, and Srikanth Ronanki. Current research fellows like Mirjam Wester and Zhizheng Wu contribute to projects such as Natural Speech Technology (NST) and Simple4All .
Catherine Lai is a Reader (~Associate Professor) in the Department of Linguistics and English Language at the University of Edinburgh, with strong affiliations to the Centre for Speech Technology Research (CSTR) and the Institute for Language, Cognition and Computation (ILCC) in the School of Informatics. She is based in the School of Philosophy, Psychology and Language Sciences and is actively involved in research, teaching, and academic service. Department: Department of Linguistics and English Language School: School of Philosophy, Psychology and Language Sciences Research Institutes: Centre for Speech Technology Research, Institute for Language, Cognition and Computation Email: C.Lai@ed.ac.uk Her research centers on the role of prosody—non-lexical aspects of speech—in spoken communication. She investigates how prosody contributes to discourse structure, information structure, and affect in dialogue, using interdisciplinary methods from linguistics and machine learning. Her work bridges theoretical linguistics and practical speech technology, aiming to improve spoken language understanding and synthesis systems. She is particularly interested in how prosody shapes listener expectations and how affect and topic are expressed and perceived in conversation. Her recent publications reflect a strong focus on self-supervised learning in speech models, emotion recognition, ASR error correction using large language models, cognitive state classification, and ethical considerations in language technology. She explores topics such as the uncanny valley in synthetic speech, gender expression through voice, and community-centered development of language technologies. Prize from Scopus Profile Catherine Lai has supervised several PhD students, including Leimin Tian and Yuanchao Li, and has been involved in significant research projects, such as a Toyota-funded initiative on spoken dialogue for robot companions. She has secured multiple grants and leads a research agenda that integrates theoretical inquiry with real-world applications in assistive technologies and social science. Her academic service includes organizing major conferences like Interspeech and UK and Ireland Speech. She is a key member of research teams at CSTR and ILCC, collaborating across disciplines to advance the understanding of spoken communication and the development of robust, ethical speech technologies.
Timothy M. Hospedales is a Professor of Artificial Intelligence at the Institute of Perception, Action and Behaviour within the School of Informatics at the University of Edinburgh . He also serves as VP AI and Head of Samsung AI Research Centre Europe . His research focuses on efficient and robust AI , emphasizing meta-learning , lifelong transfer-learning , and domain adaptation in both probabilistic and deep learning frameworks. Applications span computer vision , vision and language , reinforcement learning for robotics , and finance . Professor at University of Edinburgh (2020–present) ELLIS Fellow (2021) Head of Samsung AI Research Europe (2020–present) Founding Director of Applied Machine Learning Lab at QMUL (2012–2016) His work includes pioneering contributions to meta-learning , few-shot learning , and self-supervised methods , with notable awards such as the Best Paper Prize at ICML AutoML 2018 and Best Student Paper at ICPR 2018 . He has co-authored 15+ recent papers on topics like Vision-Language Models , Medical AI Fairness , and Diffusion Model Optimization . He served as Program Co-Chair for BMVC 2018 and AAAI 2022 , and authored a book on Visual Adaptation in the Deep Learning Era (2022). Co-Chair, BMVC 2018 Guest Editor, IET CV Special Issue (2016) Keynote Speaker at TASK-CV Workshop (ECCV 2016) Special Issue on Fewer Labels (IEEE PAMI 2020) His leadership extends to organizing workshops like the Learning-to-Learn Workshop at ICLR 2021 , Meta-Learning Workshop at NeurIPS 2020 , and Domain Generalisation Workshop at ICLR 2023 . Current projects include Meta-Omnium (CVPR 2023) for general-purpose meta-learning and MetaAudio (ICANN 2022) for few-shot audio classification benchmarks.
WANG Ye is an Associate Professor in the Department of Computer Science at the School of Computing, National University of Singapore (NUS). He holds a PhD in Information Technology from Tampere University of Technology, Finland, and has been a tenured faculty member at NUS since 2002, following his industry research role at Nokia Research Center. He is the director of the Sound and Music Computing Lab at NUS, leading cutting-edge research in AI-driven music and health technologies. PhD, Information Technology, Tampere University of Technology, Finland (2002) MSc, Telecommunications, Braunschweig University of Technology, Germany (1993) BSc, Telecommunications, South China University of Technology, China (1983) His research is centered on Sound and Music Computing for Human Health and Potential (SMC4HHP) , with a focus on eHealth, eLearning, mobile/wearable computing, and music information retrieval. His work spans AI for stroke rehabilitation, language learning through singing, singing voice synthesis, and automatic music transcription. He has pioneered systems like SLIONS (language learning via karaoke), CocoLyricist (AI co-creation for stroke recovery), and SinTechSVS (expressive singing voice synthesis). The latest articles highlight a strong trend in AI-driven music and health technologies , particularly in controllable lyric generation, singing voice synthesis, automatic pronunciation assessment, and multimodal music transcription. The research increasingly integrates large language models, explainable AI, fairness, and real-world deployment, reflecting a shift from theoretical exploration to practical, human-centered applications in healthcare and education. Dr. Wang has received numerous scientific honors, including: Best Paper Awards at ACM MM, ISMIR, IEEE ISM, and CHI First Prize, Asia Pacific Assistive, Rehabilitative, and Therapeutic Technologies Challenge (2015) Faculty Teaching Excellence Award, NUS School of Computing (2024) Top Paper Award, ACM Multimedia 2022 AI in Medicine Collaborative Grant for CocoLyricist project He has supervised over 11 PhD and 20 MComp students and is currently guiding six PhD candidates. His grants come from MOE, NRF, A*STAR, Nokia, and Smule. He has served as General Chair of ISMIR2017 and TPC Co-Chair of ICOT2017, and is on the editorial boards of IEEE Transactions on Multimedia and Journal of New Music Research. He has also developed and taught the first course on Sound and Music Computing in Singapore. Dr. Wang leads the Sound and Music Computing Lab (SMC Lab) , a multidisciplinary team exploring the synergy of music computing, AI, mobile technology, and cloud systems for health and education. The lab actively collaborates with medical institutions such as NUS Yong Loo Lin School of Medicine, Singapore General Hospital, and Harvard Medical School, and is currently working on projects in AI-supported language learning, stroke rehabilitation, and intelligent music interfaces.
University of Illinois Urbana-ChampaignUnited States
Pasquale Bottalico serves as Associate Professor in the Department of Speech and Hearing Science at the University of Illinois, with dual appointments as Associate Professor at the Center for Latin American and Caribbean Studies and Affiliate Faculty in the School of Music. His unique interdisciplinary profile bridges engineering, music performance, and speech science, reflecting his dual academic training and professional artistry. His educational foundation includes: Bachelor's in Telecommunications Engineering from Univeristà Mediterranea di Reggio Calabria, Italy Concurrent Opera Singing degree from F. Cilea Music Academy, Reggio Calabria Master's in Telecommunications Engineering from Politecnico di Torino, Italy Ph.D. in Metrology specializing in acoustics measurement uncertainty and classroom acoustics Dr. Bottalico's research centers on vocal load quantification and professional voice techniques , with significant contributions to understanding vocal fatigue in teachers and singers. His work spans Speech Intelligibility in educational environments, Room Acoustics for performance and learning spaces, and Musical Acoustics of historical vocal styles. A distinctive thread throughout his research examines how acoustic conditions modulate voice production and perception, increasingly incorporating virtual reality and bone conduction technologies for innovative assessment and intervention approaches. His Colombian vocal health study demonstrates cross-cultural applications of his work. Analysis of his 2023-2025 publications reveals three dominant research trajectories: (1) The impact of noise and dysphonia on children's speech processing in educational settings, using multimodal assessment including EEG; (2) Virtual reality applications for voice production research and therapeutic intervention; (3) Cross-cultural validation of vocal fatigue metrics and development of biofeedback systems. His work consistently bridges engineering precision with clinical applicability, particularly for professional voice users in challenging acoustic environments. No scientific awards were documented in the available information. While specific advising relationships aren't detailed, his research collaborations span international institutions including Colombian and Italian universities, suggesting graduate mentorship in interdisciplinary projects. No grant information was provided, though his systematic reviews and cross-cultural studies imply externally funded research activities. Though no dedicated laboratory is specified, his virtual reality voice studies and acoustic parameter assessments suggest affiliations with audio engineering facilities and voice clinics, likely through the Speech and Hearing Science department's research infrastructure.
Adriana Tapus is a Full Professor at ENSTA Paris, affiliated with Institut Polytechnique de Paris, leading the Autonomous Systems and Robotics Laboratory (SAR) within the Computer Science and Systems Engineering Unit (U2IS). She holds an HDR (Habilitation) and a PhD from EPFL, Switzerland, with postdoctoral experience at USC. Her research focuses on socially assistive robotics, human-robot interaction (HRI), and personalized therapy for individuals with physical/cognitive impairments. She directs the IP Paris Doctoral School and coordinates national/international projects like the EU-funded ENRICHME and SWEET. Education: PhD in Mobile Robotics, EPFL (2005) Habilitation (HDR), ENSTA Paris (2011) M.S. Computer Science, University Joseph Fourier Engineer, Politehnica University of Bucharest Research Interests: Tapus pioneers socially assistive robotics, integrating machine learning, human modeling, and multimodal communication (verbal/non-verbal/para-verbal). Her work addresses adaptive therapies for vulnerable populations using robotics, physiological data interpretation, and context-aware interaction. Key themes include: Human-robot cooperation and trust Emotion recognition and expression Personalized rehabilitation systems AI ethics and human-centered design Publications: Over 150 articles, with recent work exploring humor in HRI, teleoperation trust models, and cross-cultural intelligent vehicles. Notable 2025 contributions include studies on robot laughter efficacy and multimodal facial expression frameworks. Awards: 2025: 4 IROS papers accepted 2016: 25 Women in Robotics recognition 2010: Romanian Academy Award Multiple conference best paper awards (RO-MAN, ICRA, etc.) Advising & Grants: Supervised over 20 PhD students and led projects like EU Horizon 2020 ENRICHME. Current students focus on teleoperation dynamics, robot humor, and haptic interfaces. Active in editorial roles (IJSR, THRI) and conference organization (HRI General Chair 2019). Labs/Teams: Founder of RoboticsByDesign lab and co-initiator of the Hi! Paris interdisciplinary AI center. The SAR lab develops systems for healthcare, education, and human-robot collaboration.
Najim Dehak is an Associate Professor in the Department of Electrical and Computer Engineering at Johns Hopkins University, part of the Whiting School of Engineering. His research focuses on machine learning applied to speech processing, audio classification, and health applications. He is renowned for developing the I-vector representation for speaker recognition, introduced in 2008 during a workshop at Johns Hopkins’ Center for Language and Speech Processing. Prior to this role, he was a research scientist at MIT’s Computer Science and Artificial Intelligence Laboratory. Dehak holds a PhD from the School of Advanced Technology in Montreal (2009). He is a Senior Member of IEEE and contributes to the IEEE Speech and Language Technical Committee. His work bridges AI, healthcare, and signal processing, with notable contributions to neurodegenerative disease detection via speech and handwriting analysis. Research interests include adversarial attacks on speech systems, multimodal biomarker discovery, and robust speech processing across demographics. His lab’s tools, like the Hermespeech Recorder, enable scalable data collection for clinical and research applications. Education: PhD in Advanced Technology (2009), Montreal Affiliations: Johns Hopkins University, IEEE Labs/Teams: Center for Language and Speech Processing (CLSP) His recent work explores AI’s role in aging research, including Alzheimer’s and Parkinson’s disease detection through speech, eye tracking, and handwriting analysis. Ongoing projects address fairness in speaker verification and robustness against adversarial attacks in ASR systems.