Mark Y. Liberman is the Christopher H. Browne Distinguished Professor of Linguistics and Trustee Professor at the University of Pennsylvania. He holds a joint appointment in the Department of Linguistics and the Department of Computer and Information Science. His roles include Director of the Linguistic Data Consortium (LDC), Faculty Director of Ware College House, and former Director of the Institute for Research in Cognitive Science. Education: A.B. in Linguistics and Applied Mathematics from Harvard University (1965–1969), M.S. (1972) and Ph.D. (1975) in Linguistics from MIT. Research focuses on corpus-based phonetics, clinical linguistics applications, tonal phonology, formal models for linguistic annotation, and computational linguistics. He explores speech production, prosody, and interdisciplinary topics like language evolution and neurobiology of speech. Recent articles highlight advancements in speech biomarkers for neurodegenerative diseases, autism analysis, and computational linguistics. Awards include Fellowships from the AAAS and Linguistic Society of America. He advises graduate students and leads large-scale language resource initiatives like LDC, contributing to open-access linguistic datasets. Labs/Teams: Linguistic Data Consortium (LDC), Institute for Research in Cognitive Science (IRCS), and collaborations in computational linguistics and neuroscience.
Antoine Bosselut is a Tenure Track Assistant Professor at École Polytechnique Fédérale de Lausanne (EPFL) in the School of Computer and Communication Sciences, where he leads the EPFL NLP group. His research focuses on developing AI reasoning agents that can model, represent, and reason about human and world knowledge, with applications in health, education, and global fairness. His research interests span multiple critical areas in modern AI: LLM Representations of Knowledge: Understanding what language models know and how they represent that knowledge internally Reasoning Algorithms: Developing methods to improve LLMs' reasoning capabilities through symbolic systems, neuroscience, and cognitive science Large-scale AI Development: Creating open-source foundation models with multilingual capabilities AI Democratization: Ensuring equitable access to AI technologies across different cultural and regulatory contexts His recent publications reveal a strong focus on the intersection of language models and cognitive science, with multiple studies examining how LLMs align with human cognition. He's also deeply engaged in practical applications of NLP technology, particularly in education and multilingual settings, as evidenced by his work on evaluating AI's impact on higher education and developing culturally-aware language models. His scientific achievements have been recognized with several prestigious awards: Outstanding Paper Award at NAACL 2025 ELLIS Scholar designation in 2024 Outstanding Paper Award at ACL 2023 Inclusion in Forbes 30 Under 30 list for Science & Healthcare in 2021 Bosselut actively mentors numerous PhD students across diverse NLP research areas and has secured significant recognition for his work in both academic and media circles. His lab has been featured in major publications including Communications of the ACM, The Atlantic, and Quanta Magazine, highlighting the societal impact of his research on commonsense reasoning and AI capabilities.
Prof. Luke Zettlemoyer is an Adjunct Professor of Computer Science and Engineering at the University of Washington, with affiliations to the Department of Linguistics. He focuses on machine learning, natural language processing, and multimodal systems, contributing to advancements in large language models, ethical AI, and scalable architectures. His research addresses challenges in model alignment, generalization, and cross-domain integration. Key research interests include multimodal reward models, efficient tokenization strategies, and model optimization techniques. He has explored topics such as neural trajectories for robot learning, content-adaptive image processing, and ethical mitigation of verbatim data reproduction. His publications span 2023–2025, emphasizing practical applications of AI in robotics, vision-language systems, and scalable retrieval-based models. While no formal awards are listed, his work reflects significant contributions to foundational AI research.
Walid Magdy is a Professor at the School of Informatics, The University of Edinburgh, where he is a faculty member at the Institute for Language, Cognition and Computation (ILCC). He holds a PhD from Dublin City University's School of Computing and has been a faculty fellow at The Alan Turing Institute. With substantial industry experience, he has worked for Qatar Computing Research Institute (QCRI), Microsoft, and IBM, where he filed nine patents. His research spans computational social science and natural language processing. In computational social science, he focuses on social content analysis, political bias detection, polarization, and users' behavior analysis/prediction, with particular attention to understudied communities. His natural language processing work centers on social applications including sentiment analysis, sarcasm detection, and hate-speech identification, with specialized expertise in Arabic and its dialects as well as Persian. Professor Magdy's recent publications demonstrate a consistent focus on Arabic language processing and social media analysis, with growing attention to cultural factors in NLP systems. His work bridges computational methods with social science questions, particularly regarding misinformation, cultural differences in online behavior, and the impact of platform policies on diverse linguistic communities. He has expanded his research to include Persian language processing and continues to explore social dynamics across different cultural contexts. His significant contributions have been recognized with multiple awards: Best Demo Award for TwiXplorer at CSCW 2024 Outstanding Paper Award at ACL 2024 Best Paper Award Runner-up at SocInfo 2022 Best Paper Award at WANLP-EMNLP 2022 Best Paper Award at AIRS 2015 Best Dataset Award at ICWSM 2013 Professor Magdy serves as an editor for Springer's Social Network Analysis and Mining journal and has held leadership roles in major conferences including ACL 2023, ArabicNLP 2023, ICWSM 2022, and ASONAM 2022. He has co-organized numerous shared tasks focused on Arabic language processing and social media analysis, contributing significantly to community resources and benchmarks. He founded and directs the Social Media Analysis and Support for Humanity (SMASH) research group at the University of Edinburgh, which has organized the Summer School in Computational Social Science (SICSS-Edinburgh) in 2022, 2023, and 2025, training researchers in computational social science methodologies.
Professor Anna Korhonen is a leading academic at the University of Cambridge, holding positions as Professor of Natural Language Processing, Co-Director of the Language Technology Laboratory (LTL), Director of the Centre for Human-Inspired Artificial Intelligence (CHIA), and Fellow of Churchill College. Her work bridges computational linguistics, artificial intelligence, and interdisciplinary applications. Her research focuses on human-centric NLP with core interests in multilingual/low-resource systems, conversational AI, and responsible technology development. She emphasizes applications for social and global good, including healthcare, climate science, and equitable language technologies. Her methodology integrates cognitive modeling with machine learning to create interpretable, fair NLP systems. Key projects include ERC-funded initiatives like MultiConvAI (multilingual conversational AI) and Towards Globally Equitable Language Technologies , alongside Innovate UK's ESG RoboFactory and EPSRC's Modeling Idiomaticity project. Her work spans biomedical text mining ( PheneBank , LION ), educational technology ( EF Education First Research Lab ), and cross-lingual transfer learning. Fellow of the Association for Computational Linguistics (ACL) Fellow of ELLIS (European Laboratory for Learning and Intelligent Systems) Royal Society University Research Fellow (2005-2014) Google Faculty Award recipient EACL Chair Elect Korhonen actively supervises PhD/MPhil students in Computation, Cognition and Language programs and leads interdisciplinary collaborations across computer science, linguistics, and domain sciences. Her lab (LTL) develops foundational NLP techniques while addressing real-world challenges through partnerships with healthcare, environmental science, and education sectors. Current strategic initiatives include the Institute for Technology and Humanity and CHIA's human-inspired AI framework.
Alan Ritter is an Associate Professor at the School of Interactive Computing , Georgia Institute of Technology, with additional affiliation to the Machine Learning Center . His research focuses on Natural Language Processing , particularly robust models across domains/languages with fewer labels and efficient resource use, plus data-driven dialogue agents for open-topic conversations. Research Interests : Robust NLP models, cross-lingual transfer, resource-efficient learning, dialogue systems, cultural bias measurement, and privacy-aware language models Students : Mentors Ph.D. students in Georgia Tech's ML and CS programs, including Junmo Kang, Yang Chen, and Duong Minh Le. Alumni include Fan Bai (Ph.D. 2023), Yang Chen (Ph.D. 2024), and Andrew Li (M.S. 2024). Awards : NSF CAREER Award, Amazon Research Award, ACL 2024 Best Social Impact Paper, IUI 2009 Best Student Paper. Recent Work : Studies training budget allocation between supervised and preference-based finetuning, cross-lingual information extraction, cultural bias in LLMs, and privacy risk mitigation in social media disclosures. Service : Served as Program Chair for NAACL 2025, Area Chair for multiple top-tier conferences (COLM, EMNLP, ACL, EACL, AAAI). Email : alan.ritter@cc.gatech.edu
Sean Cao serves as Associate Professor (with tenure) at the Robert H. Smith School of Business, University of Maryland, where he is Director and Co-founder of the AI Initiative for Capital Market Research. He also holds an affiliation as professor at Harvard Business School's D 3 Institute. His academic journey began with a Ph.D. from the University of Illinois at Urbana-Champaign. Dr. Cao's research focuses on the intersection of artificial intelligence and capital markets, with particular expertise in how machine learning transforms financial analysis, corporate disclosure practices, and investment decision-making. His work examines the evolving relationship between human analysts and AI systems, blockchain applications in financial reporting, and the strategic adaptation of corporate communications for machine readership. He has pioneered research on the "AI divide" among investor groups and developed frameworks for human-AI collaborative stock analysis. His publication portfolio spans top journals including Journal of Financial Economics, Review of Financial Studies, Journal of Accounting Research, and Management Science. The research demonstrates consistent thematic progression toward increasingly sophisticated AI applications in finance, with recent work exploring distributed ledger technologies for auditing, machine learning for extracting private information from disclosures, and the economics of greenwashing in ESG funds. His studies frequently combine textual analysis with traditional financial metrics to uncover novel market insights. Fama-DFA Prize from Journal of Financial Economics for best paper in capital markets and asset pricing Michael J. Brennan Award from Review of Financial Studies Deloitte Initiative for AI and Learning award for developing trustworthy AI for social equity PanAgora Asset Management's Dr. Richard A. Crowell Memorial Prize Multiple best paper awards from Midwest Finance Association, Global AI Finance Conference, and Asian Finance Association Dr. Cao has delivered over 200 invited research talks at major institutions including the Central Bank of Japan, Central Bank of Thailand, and U.S. Securities and Exchange Commission. He serves as Guest Associate Editor for Management Science and has co-chaired Review of Financial Studies conferences on FinTech and Machine Learning. His educational initiatives include a widely adopted free AI textbook for finance and accounting that has been implemented at universities worldwide including Indiana University, UT Dallas, and University of Minnesota. As Director of the AI Initiative for Capital Market Research, Dr. Cao leads a multidisciplinary team exploring practical AI applications in finance. The initiative has secured significant funding including a $150,000 grant from GRF CPAs & Advisors. His research group maintains strong industry connections through partnerships with regulatory bodies, financial institutions, and technology companies, facilitating the translation of academic research into practical financial applications.
Sham Kakade is the Rampell Family Professor of Computer Science and Professor of Statistics at Harvard University, co-director of the Kempner Institute. His research focuses on advancing artificial general intelligence through foundational work in reinforcement learning, large-scale learning systems, and autonomous agent architectures. He earned his PhD in 2003 from the Gatsby Computational Neuroscience Unit at University College London. His work emphasizes scalable optimization algorithms, distributed systems for foundation models, and understanding emergent capabilities in neural architectures. Research interests include full-stack training pipelines for foundation models, mathematical principles of large-scale learning systems, and bridging language models with embodied intelligence. He advises prospective students with backgrounds in applied deep learning or theoretical computer science, offering access to the Kempner Institute's computational resources. He serves on committees for the ACM Prize in Computing and Sloan Research Fellowships, co-organizes the Simons Symposium on Theoretical Machine Learning, and chaired COLT 2011. His lab works at the intersection of theory and practice, addressing challenges in AI's societal impact and technical scalability. Labs/Teams: Co-directs the Kempner Institute, fostering collaborations between AI researchers and social scientists. Active in Harvard's SEAS community.
Wei-Lun (Harry) Chao is an Associate Professor in the Department of Computer Science and Engineering at the Ohio State University (OSU), College of Engineering. Promoted to this role in May 2025, he is also an Innovation Scholar and Distinguished Assistant Professor of Engineering Inclusive Excellence. His work spans machine learning, computer vision, and their applications in autonomous driving, healthcare, biology, and natural language processing. Research Focus: Machine learning with imperfect data, interpretable and personalized learning, robust perception for autonomous systems, and visual recognition in real-world scenarios. Awards: 2025 OSU Early Career Distinguished Scholar Award, CVPR Best Student Paper Award (2024), CSE Faculty Teaching Award (2024), Lumley Research Award (2023). Grants: Funded by NSF, NIH, ONR, Cisco, AWS, and Google. Notable Research Trends: The 15 most recent articles highlight his work on vision foundation models, federated learning, diffusion models for biological species generation, interpretable vision transformers, and robust perception systems for autonomous driving. Key subfields include sparse autoencoders, 3D object detection, semi-supervised learning, and anomaly detection in scientific domains. Scientific Awards: 2025 Early Career Distinguished Scholar Award (OSU) CVPR Best Student Paper Award (2024) CSE Faculty Teaching Award (2024) Lumley Research Award (2023) Mentoring & Grants: As an advisor for the OSU Buckeye AutoDrive Team and AI Club, he mentors graduate and undergraduate students. His research is supported by major grants from NSF, NIH, ONR, and industry partners like Cisco and Google.
Dr. Jimeng Sun is a Health Innovation Professor at the Siebel School of Computing and Data Science and Carle Illinois College of Medicine at the University of Illinois Urbana-Champaign. Co-founder of Keiji AI , he leads groundbreaking research at the intersection of artificial intelligence and healthcare, actively deploying clinical AI systems and developing frameworks like PyHealth and Therapeutics Data Commons . His research spans four major areas: Clinical AI Systems : Developing interpretable models (e.g., RETAIN) for patient similarity, temporal event prediction, medication recommendation, and clinical outcome forecasting Drug Discovery : Creating molecular optimization frameworks, drug-target interaction models, and AI-driven platforms Clinical Trials : Pioneering patient-trial matching, outcome prediction, and optimization frameworks using deep learning and graph neural networks Biosignal Analysis : Advancing sleep staging, seizure classification, and automated EEG/Cardiac monitoring systems With over 500 top-tier publications (including in Nature , NEJM AI , and leading AI conferences) and an h-index of 99, his work has been recognized with the Top 100 AI Leaders in Drug Discovery and Advanced Healthcare award. He maintains active collaborations with institutions like Massachusetts General Hospital , Medidata Solutions , and OSF Healthcare . His recent publications reveal a strong focus on: Reinforcement learning applications in medical data analysis Large language model adaptation for clinical tasks Knowledge graph integration with AI systems Synthetic data generation for healthcare Multi-modal learning in clinical contexts Explainable AI for medical applications Dr. Sun's lab ( Sunlab ) emphasizes practical impact over theoretical work, actively collaborating with hospitals and healthtech companies. He welcomes contributions from clinicians, researchers, and industry partners through initiatives like his AI for Health webinar series .
Simon King is a Professor of Speech Processing at the University of Edinburgh , affiliated with the School of Philosophy, Psychology and Language Sciences . He serves as Director of the Centre for Speech Technology Research (CSTR) and teaches courses like Speech Processing and Speech Synthesis , while directing the MSc in Speech and Language Processing . Research Interests His research focuses on: Developing new acoustic models (e.g., Linear Dynamical Models, factorial-HMMs) for speech recognition Advancing unit selection and HMM-based speech synthesis Integrating articulatory measurement data for enhanced modeling Exploring perceptual measures in synthesis criteria Building multilingual speech systems to identify universal speech building blocks Publication Trends Simon's recent work emphasizes deep learning (DNNs, LSTMs) in speech synthesis, multilingual frameworks , and articulatory-acoustic feature integration . His studies often bridge grapheme-based modeling , perceptual error reduction , and noise-robust synthesis . Scientific Awards EPSRC Advanced Research Fellowship (2005-2009) Students & Collaborations He has supervised numerous PhD students including Rasmus Dall, Tom Merritt, and Srikanth Ronanki. Current research fellows like Mirjam Wester and Zhizheng Wu contribute to projects such as Natural Speech Technology (NST) and Simple4All .
Min Yen Kan is an Associate Professor and Vice Dean of Undergraduate Studies at the National University of Singapore's School of Computing, Department of Computer Science. With a PhD from Columbia University (2002), he leads the Web Information Retrieval / Natural Language Processing Group (WING.NUS) and serves as ACL Ethics Committee co-chair. His research spans Natural Language Processing , Large Language Models , Digital Libraries , and Information Retrieval , with specific focus on scientific discourse analysis, fact verification, and multimodal systems. Current projects include Scholarly Document Information Extraction (TRL 6), Task-Oriented Dialogue Systems (TRL 4), and Recommendation Systems (TRL 5). Recent publications reveal strong trends in LLM limitations (bias, hallucination, evaluation), conversational recommendation systems , and misinformation detection . His work consistently bridges theoretical NLP with real-world applications in digital libraries and scientific communication. Award highlights include: CIKM 2019 Best Paper Award ACL Distinguished Service Awards Vannevar Bush Best Paper Award (JCDL 2012) ACM Distinguished Speaker designation Kan mentors PhD students with placements at Google and USTC, and serves as associate editor for Information Retrieval and survey editor for Journal of AI Research . His lab WING.NUS develops practical tools like SciWING for scientific document processing and FANG for fake news detection. Media engagements include commentary on AI regulations in Southeast Asia and workforce implications in the AI era.
Sara Stymne is a Senior Lecturer in Computational Linguistics at the Department of Linguistics and Philology, Uppsala University, where she has been working since 2012. She initially joined as a post-doc (2012-2015), then worked as a researcher (2015-2017), and served as an assistant professor (2017-2023) before her current position as Senior Lecturer. Prior to Uppsala, she was a researcher at Linköping University's Department of Computer and Information Science. Dr. Stymne earned her PhD in Computational Linguistics from Linköping University in 2012 with the thesis 'Text Harmonization Strategies for Phrase-Based Statistical Machine Translation,' following a Licentiate degree in Computational Linguistics (2009) and a Master's degree in Cognitive Science (2006), both also from Linköping University. During her doctoral studies, she spent the autumn of 2010 and spring of 2009 at Xerox Research Centre Europe in Grenoble, France. Her primary research interests focus on cross-lingual natural language processing and digital humanities, with particular emphasis on multilingual dependency parsing. Dr. Stymne is passionate about applying computational linguistics to solve research questions in other fields, including language history, literary analysis, and political science. Her earlier work concentrated on machine translation, with specific interests in discourse-aware translation, compound processing, and error analysis. She has made significant contributions to the development of language technology tools for analyzing dialogue, narrative, and stylistic features in literature. Analysis of Dr. Stymne's recent publications reveals a strong focus on cross-lingual and cross-domain natural language processing. Her work spans multiple subfields including dependency parsing across genres and topics, discourse relation analysis in low-resource languages like Egyptian Arabic, direct speech identification in Swedish literature, and causality detection in governmental documents. A notable trend is her application of NLP techniques to digital humanities problems, particularly in analyzing literary texts and historical language change. Her research often involves creating and utilizing specialized datasets for specific linguistic phenomena across multiple languages. Dr. Stymne actively supervises graduate students, having guided numerous master's and bachelor's theses on topics ranging from speech recognition to multilingual parsing and causality detection. She leads or participates in several research projects including 'Fictional prose and language change' (funded by VR, 2021-2023) and 'Enabling climate-resilient development' (funded by Marianne and Marcus Wallenberg Foundation, 2023-2027), demonstrating her commitment to interdisciplinary research with practical applications. Her work has resulted in several notable software resources including uuPronPred for cross-lingual pronoun prediction, uuparser for dependency parsing, and Docent for document-level machine translation. Within the Computational Linguistics and Language Technology group at Uppsala University, Dr. Stymne contributes to multiple research initiatives focused on developing language technology tools for digital humanities applications. Her team works closely with literary scholars and historians to create computational methods for analyzing large corpora of literary texts, particularly focusing on Swedish literature across different historical periods. Her research bridges the gap between theoretical computational linguistics and practical applications in the humanities, creating new methodologies for quantitative analysis of literary and historical texts.
Prof. Dr. Gülşen Eryiğit is a Professor at Istanbul Technical University within the Faculty of Computer and Informatics , Department of Artificial Intelligence and Data Engineering . She founded and directs the ITU Natural Language Processing Group , Turkey's leading team in Turkish-language NLP, and serves as Senior Action Editor for ACL RR , Director of ITU TÖMER (Turkish Language Teaching Center), and Co-Chair of the EU UniDive Cost Action WG3. Education: PhD in Computer Engineering from ITU (2007), MSc and BSc from ITU and Marmara University Research: Focuses on Natural Language Processing for Turkish, including dependency parsing , coreference resolution , multiword expressions , and language education technology Her recent work involves multilingual transfer learning , LLM applications for Turkish text simplification, and gamification for morphology education. She has received prestigious awards like the Siemens Excellence Award and TÜBİTAK's Above Threshold Award . Her research has produced 65+ publications and 29+ projects funded by EU, TÜBİTAK, and industry partners. Scientific Awards: Siemens Excellence Award (2007) Above Threshold Award (TÜBİTAK, 2015) Certificate of Appreciation (EU 7th Framework, 2012) Thank You Plaque (ITU, 2017) The ITU NLP Group under her leadership has developed Turkey's first licensed NLP software exported internationally. She collaborates with European institutions through COST actions and participates in ACL, CoNLL, and LREC conferences. Her lab focuses on language technology for Turkish , including sign language processing and social media normalization.
Mark Gales is Professor of Information Engineering at the University of Cambridge and an Official Fellow at Emmanuel College. He is currently on sabbatical leave for the 2024/25 academic year. Prior to his academic career, he worked as a consultant at Roke Manor Research Ltd, developing radar systems, before transitioning to speech and language processing. PhD in 'Model-Based Techniques for Robust Speech Recognition' (University of Cambridge, 1995) BA in Electrical and Information Sciences (University of Cambridge, 1988) His research focuses on speech and language processing , particularly in automated language assessment and low-resource speech technology . He leads the Automated Language Teaching and Assessment (ALTA) Institute , which collaborates with Cambridge University Press & Assessment (CUP&A) to develop commercial tools like Linguaskill and Speak & Improve . These platforms provide automated spoken/written assessment for millions of users globally. Recent publications highlight his work in LLM-driven speech processing , including adversarial attacks on foundation models, end-to-end spoken error correction, and uncertainty estimation frameworks. His team's research spans multilingual capabilities, with deployments in languages ranging from Dholuo to Tok Pisin . Awards : IEEE Fellow, ISCA Fellow Leadership : Fellows' Steward at Emmanuel College Mark has contributed extensively to Hidden Markov Model (HMM) applications in speech recognition, which underpinned early automatic speech systems. His work now bridges LLM-based language assessment with cross-lingual transfer learning and robustness testing for real-world deployments.