Mark Y. Liberman is the Christopher H. Browne Distinguished Professor of Linguistics and Trustee Professor at the University of Pennsylvania. He holds a joint appointment in the Department of Linguistics and the Department of Computer and Information Science. His roles include Director of the Linguistic Data Consortium (LDC), Faculty Director of Ware College House, and former Director of the Institute for Research in Cognitive Science. Education: A.B. in Linguistics and Applied Mathematics from Harvard University (1965–1969), M.S. (1972) and Ph.D. (1975) in Linguistics from MIT. Research focuses on corpus-based phonetics, clinical linguistics applications, tonal phonology, formal models for linguistic annotation, and computational linguistics. He explores speech production, prosody, and interdisciplinary topics like language evolution and neurobiology of speech. Recent articles highlight advancements in speech biomarkers for neurodegenerative diseases, autism analysis, and computational linguistics. Awards include Fellowships from the AAAS and Linguistic Society of America. He advises graduate students and leads large-scale language resource initiatives like LDC, contributing to open-access linguistic datasets. Labs/Teams: Linguistic Data Consortium (LDC), Institute for Research in Cognitive Science (IRCS), and collaborations in computational linguistics and neuroscience.
Justin Johnson is an Assistant Professor at the University of Michigan's College of Engineering, Department of Electrical Engineering and Computer Science, and a Research Scientist at Facebook AI Research (FAIR). His work bridges computer vision, machine learning, and deep learning, focusing on visual reasoning, vision-language tasks, image generation, and 3D reasoning using neural networks. PhD, Stanford University (advised by Fei-Fei Li) His research interests span visual reasoning , vision and language , 3D vision , and image generation , with a focus on innovative applications of deep neural networks. Recent publications highlight work on 3D consistency, self-supervised learning, and multimodal integration of vision and text. Notable contributions include PyTorch3D for 3D data processing, and foundational work in visual question answering , neural style transfer , and scene graph-based image generation . Publications span top conferences like ICCV, CVPR, and NeurIPS. He teaches courses including EECS 498/598: Deep Learning for Computer Vision and EECS 442: Computer Vision at University of Michigan, with prior involvement in Stanford's CS 231N in co-teaching roles. Software projects include open-source frameworks like fast-neural-style for real-time artistic style transfer, and PyTorch3D for efficient 3D deep learning. These tools demonstrate his commitment to practical implementations and community-driven research.
Jan Østergaard is a Full Professor in Information Theory and Signal Processing at Aalborg University's Department of Electronic Systems. He leads the AI and Sound research section and directs the CASPR center. His expertise spans AI-driven acoustic signal processing, information theory, and EEG signal analysis. Østergaard holds a M.Sc. from Aalborg University and a PhD (cum laude) from Delft University of Technology. Major awards include the Danish Young Researcher’s Award and a EURASIP Best Thesis honor. His work focuses on speech enhancement, sound zone technologies, and neural tracking of auditory attention. Recent research emphasizes low-latency speech transmission, deep learning for sound field control, and robust voice activity detection. He serves on editorial boards and national committees, advancing Denmark’s sound technology initiatives. Education: M.Sc. (Aalborg, 1999), PhD (Delft, 2007) Research interests emphasize practical AI applications in sound systems, including hearing aid improvements, data-efficient acoustic modeling, and feedback control in networked systems. Over 210 publications and 17 active projects reflect his interdisciplinary impact across academia and industry.
Mark Steedman is a Professor in the School of Informatics at the University of Edinburgh, where he conducts research in Artificial Intelligence, Computational Cognitive and Social Science, and Natural Language and Speech Processing. He is affiliated with the Institute for Language, Cognition and Computation (ILCC), the Centre for Speech Technology Research (CSTR), and the Human Communications Research Center (HCRC). He also holds an adjunct professorship in Computer and Information Science at the University of Pennsylvania. His research focuses on Combinatory Categorial Grammar (CCG) , computational linguistics , prosody and intonation , temporal semantics , gesture in communication , and computational music analysis . He has authored foundational books including Surface Structure and Interpretation , The Syntactic Process , and Taking Scope . The recent publications reflect a strong trend toward integrating formal grammatical frameworks like CCG with modern neural and distributional models, particularly in semantic parsing, entailment reasoning, and cognitive modeling. His work bridges symbolic and statistical approaches in NLP, often focusing on robust, wide-coverage parsing and semantic interpretation. Best Paper Award at AACL/IJCNLP 2023 for 'Smoothing Entailment Graphs with Language Models' Best Paper Award at ACL 2023 for 'Extrinsic Evaluation of Machine Translation Metrics' Influential Paper Award 2017 from IFAAMAS for 'Animated Conversation' Mark Steedman has supervised numerous PhD students and collaborated widely across institutions. He leads research in formal grammar applications to cognitive modeling, dialogue, and multimodal communication. His lab contributes to CCG software and semantic parsing tools, and he continues to be actively involved in advancing the integration of symbolic and neural AI.
Prof. Matthias Nießner is a Professor at the Technical University of Munich, leading the Visual Computing Lab. His research intersects computer graphics, vision, and AI, focusing on 3D reconstruction, semantic understanding, and AI-driven video synthesis. He holds a PhD from the University of Erlangen-Nuremberg (2013) and was a Visiting Assistant Professor at Stanford University (2013–2017). Notable awards include the ERC Starting Grant (2018), Nvidia Professorship Award, and Eurographics Young Researcher Award (2019). His work has been featured in mainstream media and led to startups like Synthesia Inc. Research spans Gaussian splatting, neural radiance fields, and generative AI for 3D avatars. Over 150 publications include SIGGRAPH, CVPR, and ECCV, with best paper awards. Projects like Face2Face and ScanNet have driven innovation in facial reenactment and 3D scene datasets. Education: PhD in Computer Science, University of Erlangen-Nuremberg (2013) Diploma in Computer Science, University of Erlangen-Nuremberg (2010) Research Interests: 3D digitization, neural rendering, generative AI, non-rigid reconstruction, and applications in AR/VR. Awards: ERC Starting Grant (2018) Nvidia Professorship Award (2018) Google Faculty Award (2018) SIGGRAPH Best Emerging Tech Award (2016) Grants: Over €1.5M from ERC and industry partnerships. Labs/Teams: Visual Computing Lab at TUM and Synthesia Inc. (co-founder). Key projects include ScanNet (large 3D indoor dataset), Face2Face (real-time facial reenactment), and Gaussian-based 3D avatars. Current work focuses on diffusion models, neural radiance fields, and AI-generated media detection.
Swiss Federal Institute of Technology in LausanneSwitzerland
Lior Wolf is a Professor at the School of Computer Science, Tel Aviv University. Previously, he was a postdoctoral researcher at MIT's Center for Biological and Computational Learning (CBCL) under Prof. Tomaso Poggio and earned his PhD from Hebrew University of Jerusalem with Prof. Amnon Shashua. His educational background includes: PhD in Computer Science, Hebrew University of Jerusalem Postdoctoral Research, MIT CBCL Prof. Wolf's research centers on artificial intelligence with seminal contributions to deep learning, computer vision, and natural language processing. His work bridges theoretical foundations (e.g., attention mechanisms, transformer analysis) with practical applications in medical imaging, speech processing, and sign language technology. He pioneered methods for neural network interpretability, efficient sequence modeling, and multimodal fusion. Analysis of his 2023-2025 publications reveals dominant trends in large language model optimization (neuron pruning, attention analysis), efficient video generation, and cross-modal learning. His work increasingly integrates medical applications (fMRI/EEG analysis) while maintaining theoretical rigor in model architecture design. His scientific achievements include: Best paper award at EMNLP 2024 for 'Backward Lens: Projecting Language Model Gradients into the Vocabulary Space' Best paper award at SCIA 2023 for 'Gradient Adjusting Networks for Domain Inversion' Best paper award at FG 2021 for 'Generating Master Faces for Dictionary Attacks' Prof. Wolf mentors graduate students in the School of Computer Science and leads research at the ICRC building laboratory. His team collaborates with 'the friends of TAU' on projects spanning biometric security, medical imaging, and generative AI. Current work focuses on efficient transformers, neural network interpretability, and multimodal medical diagnostics.
Andrew Zisserman is a Royal Society Research Professor at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). His research focuses on computer vision, artificial intelligence, and neural networks, with significant contributions to multimodal learning, video understanding, and 3D scene analysis. He leads projects exploring visual-language models, audio-visual synchronization, and clinical imaging applications. Key research areas include: Video analysis and temporal modeling Multimodal systems for sign language translation and action recognition 3D shape estimation and physical property inference Foundation models and cross-modal retrieval Recent work highlights: Developed Flamingo and Tapir models for video-language tasks Advancements in spinal MRI analysis and clinical imaging Leadership in EGO4D and VoxCeleb challenges Honors include Fellowship of the Royal Society (FRS) and the ISSLS Prize in Clinical Science 2023 for spinal analysis innovations. His lab collaborates globally, emphasizing real-world applications in healthcare and autonomous systems.
Simon King is a Professor of Speech Processing at the University of Edinburgh , affiliated with the School of Philosophy, Psychology and Language Sciences . He serves as Director of the Centre for Speech Technology Research (CSTR) and teaches courses like Speech Processing and Speech Synthesis , while directing the MSc in Speech and Language Processing . Research Interests His research focuses on: Developing new acoustic models (e.g., Linear Dynamical Models, factorial-HMMs) for speech recognition Advancing unit selection and HMM-based speech synthesis Integrating articulatory measurement data for enhanced modeling Exploring perceptual measures in synthesis criteria Building multilingual speech systems to identify universal speech building blocks Publication Trends Simon's recent work emphasizes deep learning (DNNs, LSTMs) in speech synthesis, multilingual frameworks , and articulatory-acoustic feature integration . His studies often bridge grapheme-based modeling , perceptual error reduction , and noise-robust synthesis . Scientific Awards EPSRC Advanced Research Fellowship (2005-2009) Students & Collaborations He has supervised numerous PhD students including Rasmus Dall, Tom Merritt, and Srikanth Ronanki. Current research fellows like Mirjam Wester and Zhizheng Wu contribute to projects such as Natural Speech Technology (NST) and Simple4All .
Raman Arora is an Associate Professor in the Department of Computer Science at Johns Hopkins University, with affiliations to the Mathematical Institute for Data Science (MINDS), the Center for Language and Speech Processing (CLSP), and the Institute for Data-Intensive Engineering and Science (IDIES). His research spans theoretical and practical aspects of machine learning, focusing on robustness, privacy, representation learning, and optimization. Research Interests: Machine Learning Theory Representation Learning (e.g., Deep CCA, Multi-view Learning) Privacy-Preserving Machine Learning (Differential Privacy) Robustness in Deep Learning Online and Reinforcement Learning Stochastic Optimization Algorithms His recent publications, primarily in top-tier venues like NeurIPS, ICML, and ICLR, demonstrate a strong focus on the theoretical foundations of adversarial robustness, multi-task learning, offline reinforcement learning, and differentially private optimization. His work often bridges theory and practice, with applications in speech, language, and data-intensive systems. Scientific Awards and Honors: NSF CAREER Award (2020) ICML Test-of-Time Award Finalist (2023) for Deep CCA Member, Institute for Advanced Study (2019–2020) Visiting Scientist, Simons Institute (2019, 2020, 2022) Advising and Grants: Raman Arora has advised numerous PhD and master’s students, many of whom are now researchers at leading tech companies like Google, Meta, and Microsoft. His research is supported by significant grants from the NSF (including CAREER, BIGDATA, TRIPODS, and CRCNS awards), DARPA, and other agencies, focusing on foundational aspects of machine learning such as inductive biases, privacy, robustness, and computational neuroscience. Laboratory and Research Group: He leads a dynamic research group at Johns Hopkins, comprising current PhD students and postdoctoral researchers working on the intersection of theory and applications in machine learning. The group is actively involved in projects related to adversarial robustness, meta-learning, offline reinforcement learning, and private optimization.
Catherine Lai is a Reader (~Associate Professor) in the Department of Linguistics and English Language at the University of Edinburgh, with strong affiliations to the Centre for Speech Technology Research (CSTR) and the Institute for Language, Cognition and Computation (ILCC) in the School of Informatics. She is based in the School of Philosophy, Psychology and Language Sciences and is actively involved in research, teaching, and academic service. Department: Department of Linguistics and English Language School: School of Philosophy, Psychology and Language Sciences Research Institutes: Centre for Speech Technology Research, Institute for Language, Cognition and Computation Email: C.Lai@ed.ac.uk Her research centers on the role of prosody—non-lexical aspects of speech—in spoken communication. She investigates how prosody contributes to discourse structure, information structure, and affect in dialogue, using interdisciplinary methods from linguistics and machine learning. Her work bridges theoretical linguistics and practical speech technology, aiming to improve spoken language understanding and synthesis systems. She is particularly interested in how prosody shapes listener expectations and how affect and topic are expressed and perceived in conversation. Her recent publications reflect a strong focus on self-supervised learning in speech models, emotion recognition, ASR error correction using large language models, cognitive state classification, and ethical considerations in language technology. She explores topics such as the uncanny valley in synthetic speech, gender expression through voice, and community-centered development of language technologies. Prize from Scopus Profile Catherine Lai has supervised several PhD students, including Leimin Tian and Yuanchao Li, and has been involved in significant research projects, such as a Toyota-funded initiative on spoken dialogue for robot companions. She has secured multiple grants and leads a research agenda that integrates theoretical inquiry with real-world applications in assistive technologies and social science. Her academic service includes organizing major conferences like Interspeech and UK and Ireland Speech. She is a key member of research teams at CSTR and ILCC, collaborating across disciplines to advance the understanding of spoken communication and the development of robust, ethical speech technologies.
Timothy M. Hospedales is a Professor of Artificial Intelligence at the Institute of Perception, Action and Behaviour within the School of Informatics at the University of Edinburgh . He also serves as VP AI and Head of Samsung AI Research Centre Europe . His research focuses on efficient and robust AI , emphasizing meta-learning , lifelong transfer-learning , and domain adaptation in both probabilistic and deep learning frameworks. Applications span computer vision , vision and language , reinforcement learning for robotics , and finance . Professor at University of Edinburgh (2020–present) ELLIS Fellow (2021) Head of Samsung AI Research Europe (2020–present) Founding Director of Applied Machine Learning Lab at QMUL (2012–2016) His work includes pioneering contributions to meta-learning , few-shot learning , and self-supervised methods , with notable awards such as the Best Paper Prize at ICML AutoML 2018 and Best Student Paper at ICPR 2018 . He has co-authored 15+ recent papers on topics like Vision-Language Models , Medical AI Fairness , and Diffusion Model Optimization . He served as Program Co-Chair for BMVC 2018 and AAAI 2022 , and authored a book on Visual Adaptation in the Deep Learning Era (2022). Co-Chair, BMVC 2018 Guest Editor, IET CV Special Issue (2016) Keynote Speaker at TASK-CV Workshop (ECCV 2016) Special Issue on Fewer Labels (IEEE PAMI 2020) His leadership extends to organizing workshops like the Learning-to-Learn Workshop at ICLR 2021 , Meta-Learning Workshop at NeurIPS 2020 , and Domain Generalisation Workshop at ICLR 2023 . Current projects include Meta-Omnium (CVPR 2023) for general-purpose meta-learning and MetaAudio (ICANN 2022) for few-shot audio classification benchmarks.
WANG Ye is an Associate Professor in the Department of Computer Science at the School of Computing, National University of Singapore (NUS). He holds a PhD in Information Technology from Tampere University of Technology, Finland, and has been a tenured faculty member at NUS since 2002, following his industry research role at Nokia Research Center. He is the director of the Sound and Music Computing Lab at NUS, leading cutting-edge research in AI-driven music and health technologies. PhD, Information Technology, Tampere University of Technology, Finland (2002) MSc, Telecommunications, Braunschweig University of Technology, Germany (1993) BSc, Telecommunications, South China University of Technology, China (1983) His research is centered on Sound and Music Computing for Human Health and Potential (SMC4HHP) , with a focus on eHealth, eLearning, mobile/wearable computing, and music information retrieval. His work spans AI for stroke rehabilitation, language learning through singing, singing voice synthesis, and automatic music transcription. He has pioneered systems like SLIONS (language learning via karaoke), CocoLyricist (AI co-creation for stroke recovery), and SinTechSVS (expressive singing voice synthesis). The latest articles highlight a strong trend in AI-driven music and health technologies , particularly in controllable lyric generation, singing voice synthesis, automatic pronunciation assessment, and multimodal music transcription. The research increasingly integrates large language models, explainable AI, fairness, and real-world deployment, reflecting a shift from theoretical exploration to practical, human-centered applications in healthcare and education. Dr. Wang has received numerous scientific honors, including: Best Paper Awards at ACM MM, ISMIR, IEEE ISM, and CHI First Prize, Asia Pacific Assistive, Rehabilitative, and Therapeutic Technologies Challenge (2015) Faculty Teaching Excellence Award, NUS School of Computing (2024) Top Paper Award, ACM Multimedia 2022 AI in Medicine Collaborative Grant for CocoLyricist project He has supervised over 11 PhD and 20 MComp students and is currently guiding six PhD candidates. His grants come from MOE, NRF, A*STAR, Nokia, and Smule. He has served as General Chair of ISMIR2017 and TPC Co-Chair of ICOT2017, and is on the editorial boards of IEEE Transactions on Multimedia and Journal of New Music Research. He has also developed and taught the first course on Sound and Music Computing in Singapore. Dr. Wang leads the Sound and Music Computing Lab (SMC Lab) , a multidisciplinary team exploring the synergy of music computing, AI, mobile technology, and cloud systems for health and education. The lab actively collaborates with medical institutions such as NUS Yong Loo Lin School of Medicine, Singapore General Hospital, and Harvard Medical School, and is currently working on projects in AI-supported language learning, stroke rehabilitation, and intelligent music interfaces.
Sebastian Scherer is an Associate Research Professor at the Robotics Institute (RI), Carnegie Mellon University (CMU), where he leads cutting-edge research in autonomous aerial systems and robotics. His work focuses on enabling unmanned rotorcraft to operate safely and efficiently in cluttered, low-altitude, and extreme environments. Education: Ph.D. in Robotics, Carnegie Mellon University (2010) MS in Robotics, Carnegie Mellon University (2007) BS in Computer Science (Minor in Robotics), Carnegie Mellon University (2004) His research interests span robotics, artificial intelligence, autonomous navigation, obstacle avoidance, SLAM, visual-inertial odometry, energy infrastructure, and public policy . He has made seminal contributions to UAV autonomy, including the first obstacle avoidance for micro aerial vehicles in natural environments (2008) and the first automatic landing zone detection and landing on a full-size helicopter (2010). His recent publications (2023–2025) demonstrate a strong focus on resilient autonomy, multi-robot exploration, foundation models for robotics, and large-scale dataset development. His team has released key datasets like TartanGround , BETTY , and SubT-MRS , and simulation tools like Pegasus Simulator , indicating a systems-level approach to advancing real-world autonomy. The research trends emphasize self-supervised learning, robust perception, risk-aware planning, and multi-modal fusion for off-road and urban environments. Scientific Awards: Popular Science Best of What's New 2010 Award AIAA@Infotech Best Paper Runner-up Award (2010) Siebel Scholar Dr. Scherer has advised numerous students and leads a vibrant research group focused on high-impact robotics applications. He has secured significant grants related to UAV autonomy, energy infrastructure, and urban air mobility. His lab develops experimental infrastructure such as AIrTonomy for testing next-generation autonomous aerial vehicles. He is actively involved in advancing SLAM and localization in extreme environments, notably through participation in the DARPA Subterranean Challenge. His team develops large-scale datasets and benchmarking frameworks to push the boundaries of robustness and generalization in mobile robotics.
Srijan Kumar is an Assistant Professor in the School of Computational Science and Engineering at Georgia Institute of Technology's College of Computing. His research focuses on data science, AI for security, and online safety, addressing challenges in detecting malicious users, misinformation, and enhancing AI robustness. His work has been deployed in platforms like Flipkart and Wikipedia, and recognized through awards such as the NSF CAREER and Forbes 30 Under 30. Education: B.Tech from Indian Institute of Technology, Kharagpur Ph.D. in Computer Science from University of Maryland, College Park Postdoctoral training at Stanford University Research Interests: Multi-modal/multi-lingual detection of harmful content and users Adversarial robustness of AI models Graph and network analysis for early detection Responsible recommender systems Awards & Grants: NSF CAREER Award (2023) Kavli Fellow (2022) Facebook/Adobe Faculty Awards NSF Convergence Accelerator Phase II grant ($5M) Advising & Labs: Leads the CLAWS Lab, advising over 20 students. Active in mentoring through conferences and NSF-funded projects.
Najim Dehak is an Associate Professor in the Department of Electrical and Computer Engineering at Johns Hopkins University, part of the Whiting School of Engineering. His research focuses on machine learning applied to speech processing, audio classification, and health applications. He is renowned for developing the I-vector representation for speaker recognition, introduced in 2008 during a workshop at Johns Hopkins’ Center for Language and Speech Processing. Prior to this role, he was a research scientist at MIT’s Computer Science and Artificial Intelligence Laboratory. Dehak holds a PhD from the School of Advanced Technology in Montreal (2009). He is a Senior Member of IEEE and contributes to the IEEE Speech and Language Technical Committee. His work bridges AI, healthcare, and signal processing, with notable contributions to neurodegenerative disease detection via speech and handwriting analysis. Research interests include adversarial attacks on speech systems, multimodal biomarker discovery, and robust speech processing across demographics. His lab’s tools, like the Hermespeech Recorder, enable scalable data collection for clinical and research applications. Education: PhD in Advanced Technology (2009), Montreal Affiliations: Johns Hopkins University, IEEE Labs/Teams: Center for Language and Speech Processing (CLSP) His recent work explores AI’s role in aging research, including Alzheimer’s and Parkinson’s disease detection through speech, eye tracking, and handwriting analysis. Ongoing projects address fairness in speaker verification and robustness against adversarial attacks in ASR systems.