Mark Liberman is a Trustee Professor at the University of Pennsylvania , holding appointments in the Department of Linguistics and Department of Computer and Information Science . He serves as Director of the Linguistic Data Consortium and Faculty Director of Ware College House . His career spans linguistics, speech technology, and computational methods. Education: Harvard University (1965-1969), MIT (M.S. 1972, Ph.D. 1975) Professional Experience: AT&T Bell Laboratories (1975-1990), University of Pennsylvania (1990-present) His research interests include: Corpus-based Phonetics : Analyzing speech patterns via large-scale datasets. Clinical Applications : Developing speech biomarkers for neurodegenerative diseases. Tonal Phonology : Studying lexical tone and intonation in languages like Yoruba and Mandarin. Formal Annotation Models : Creating frameworks for linguistic data standardization. Recent publications highlight automated speech analysis, cross-linguistic prosody, and digital biomarkers for conditions like ALS-FTD and Alzheimer’s. His collaborations span computational linguistics , neurology , and cognitive science . Scientific awards include the IEEE James L. Flanagan Award (2017), Antonio Zampolli Prize (2010), and fellowships from the AAAS and Linguistic Society of America . He advises PhD students May Chan and Jonathan Him Nok Lee and contributes to editorial boards for journals like Cognition and Annual Review of Linguistics . His work bridges speech science , language technology , and neurocognitive research .
Nicholas Evans is a Distinguished Professor in the Department of Linguistics at the School of Culture, History & Language, Australian National University. He is an ARC Laureate Fellow and serves as Director of the ARC Centre of Excellence for the Dynamics of Language (CoEDL), reflecting his leadership in integrating typology, descriptive linguistics, and cognitive science. His work spans linguistic diversity, endangered languages, and the interplay between language and culture. His research focuses on Australian and Papuan languages, with extensive fieldwork in remote communities. Key interests include linguistic typology, historical and contact linguistics, semantics, and the cognitive underpinnings of grammar. He has published foundational grammars and dictionaries of Kayardild, Bininj Gun-wok, and Dalabon, contributing significantly to the documentation of endangered languages. His recent publications reveal a strong trend in cross-linguistic studies of social cognition (SCOPIC), syntactic typology, and phylogenetic modeling of the Yam languages. These works integrate linguistic, cognitive, and genetic data, highlighting interdisciplinary approaches to understanding language evolution and human diversity. ARC Laureate Fellow Fellow of the Australian Academy of the Humanities (FAHA) Fellow of the Academy of the Social Sciences in Australia (FASSA) Fellow of the British Academy (FBA) Evans supervises research students and leads major projects such as 'The Wellsprings of Linguistic Diversity' and CoEDL, which involve extensive collaboration across institutions and disciplines. His grants include modeling Pacific creoles and phonemic analysis of Tok Pisin. He is also involved in digital archiving initiatives like PARADISEC and the Language Data Commons of Australia, ensuring long-term preservation of linguistic data. He leads and contributes to research teams focused on Southern New Guinea languages, paradigm syncretisms (PARABANK), and cultural heritage preservation. His work emphasizes reciprocal engagement with speech communities and training the next generation of language documenters.
Nicholas Evans is Distinguished Professor of Linguistics and Director of the ARC Centre of Excellence for the Dynamics of Language (CoEDL) at the Australian National University’s School of Culture, History & Language. His work bridges fieldwork-based language documentation with theoretical questions in typology, cultural evolution, and social cognition. Focus on endangered Australian and Papuan languages Director of ARC Laureate Project on 'The Wellsprings of Linguistic Diversity' Co-leader of SCOPIC (Social Cognition Parallax Corpus) study Collaborator in global linguistic diversity initiatives His research explores how micro-level community multilingualism shapes macro-level linguistic diversity, with fieldwork spanning seven years in remote Indigenous communities. Recent projects include PARABANK (paradigm syncretism analysis) and Southern New Guinea language studies, particularly Nen and Yam family languages. Scientific recognition includes the Ken Hale Award (Linguistic Society of America), Anneliese Maier Forschungspreis, and fellowships in the Australian Academy of Humanities, Australian Social Sciences Academy, and the British Academy.
Mark Y. Liberman is the Christopher H. Browne Distinguished Professor of Linguistics and Trustee Professor at the University of Pennsylvania. He holds a joint appointment in the Department of Linguistics and the Department of Computer and Information Science. His roles include Director of the Linguistic Data Consortium (LDC), Faculty Director of Ware College House, and former Director of the Institute for Research in Cognitive Science. Education: A.B. in Linguistics and Applied Mathematics from Harvard University (1965–1969), M.S. (1972) and Ph.D. (1975) in Linguistics from MIT. Research focuses on corpus-based phonetics, clinical linguistics applications, tonal phonology, formal models for linguistic annotation, and computational linguistics. He explores speech production, prosody, and interdisciplinary topics like language evolution and neurobiology of speech. Recent articles highlight advancements in speech biomarkers for neurodegenerative diseases, autism analysis, and computational linguistics. Awards include Fellowships from the AAAS and Linguistic Society of America. He advises graduate students and leads large-scale language resource initiatives like LDC, contributing to open-access linguistic datasets. Labs/Teams: Linguistic Data Consortium (LDC), Institute for Research in Cognitive Science (IRCS), and collaborations in computational linguistics and neuroscience.
Marlyse Baptista is the President's Distinguished Professor of Linguistics at the University of Pennsylvania, Department of Linguistics. She is affiliated with the School of Arts & Sciences and MindCore initiative. Her research focuses on language contact, creolization processes, bilingualism, and experimental methods in creole studies. Baptista leads the Language Contact and Cognition Lab, previously known as the Cognition, Convergence and Language Emergence (CCLE) group at the University of Michigan. Education includes a PhD in Linguistics (Harvard, 1997), MA degrees from Harvard and the Université de Bordeaux III, and extensive training in multilingual education. Her work bridges generative syntax with experimental approaches, including artificial language learning to study language convergence. Key themes include the cognitive underpinnings of creole formation, bidirectional influences in Cape Verdean Creole, and the role of congruence in language acquisition. Research highlights include studies on Cape Verdean Creole's grammatical properties, genetic-linguistic admixture correlations, and pedagogical resources for creole language education. Recent projects explore experimental validation of creole genesis theories through multilingual acquisition studies. Scientific Awards: President's Distinguished Professorship (University of Pennsylvania) Labs/Teams: Language Contact and Cognition Lab, MindCore Grants/Projects: MULTI Project (creole language educational resources), NSF-funded studies on creole genesis mechanisms Baptista's work emphasizes interdisciplinary approaches, integrating syntax theory, psycholinguistics, and sociolinguistics to address foundational questions in contact linguistics and creolistics.
Alan Ritter is an Associate Professor at the School of Interactive Computing , Georgia Institute of Technology, with additional affiliation to the Machine Learning Center . His research focuses on Natural Language Processing , particularly robust models across domains/languages with fewer labels and efficient resource use, plus data-driven dialogue agents for open-topic conversations. Research Interests : Robust NLP models, cross-lingual transfer, resource-efficient learning, dialogue systems, cultural bias measurement, and privacy-aware language models Students : Mentors Ph.D. students in Georgia Tech's ML and CS programs, including Junmo Kang, Yang Chen, and Duong Minh Le. Alumni include Fan Bai (Ph.D. 2023), Yang Chen (Ph.D. 2024), and Andrew Li (M.S. 2024). Awards : NSF CAREER Award, Amazon Research Award, ACL 2024 Best Social Impact Paper, IUI 2009 Best Student Paper. Recent Work : Studies training budget allocation between supervised and preference-based finetuning, cross-lingual information extraction, cultural bias in LLMs, and privacy risk mitigation in social media disclosures. Service : Served as Program Chair for NAACL 2025, Area Chair for multiple top-tier conferences (COLM, EMNLP, ACL, EACL, AAAI). Email : alan.ritter@cc.gatech.edu
Meredith Tamminga is an Associate Professor of Linguistics at the University of Pennsylvania, where she directs the Language Variation and Cognition Lab . Her work bridges sociolinguistics , psycholinguistics , and theoretical linguistics , focusing on how linguistic variability is represented in mental grammar and how social context influences speech perception and production. She co-leads the Philadelphia Signs Project , exploring sign language sociolinguistics with collaborators at Penn and Gallaudet University. PhD in Linguistics (2014), University of Pennsylvania BA in Linguistics (2009), McGill University Her research integrates experimental methods with naturalistic speech analysis to study individual and group-level language change. Keywords include phonetics-phonology mapping , intraspeaker variation , and quantitative modeling . Notable grants include NSF awards and SAS Dean’s Mentorship Award . Key affiliations include MindCORE , Penn’s hub for integrative mind sciences, and collaborative ties to the Phonetics Lab , Child Language Lab , and Cultural Evolution of Language Lab . Her lab emphasizes cross-departmental research and welcomes undergraduate RAs and PhD students.
David A. Smith is an Associate Professor at the Khoury College of Computer Sciences, Northeastern University. His research focuses on Natural Language Processing (NLP) and computational linguistics, with applications in machine translation, information retrieval, digital humanities, and social sciences. He is a founding member of the NULab for Texts, Maps, and Networks, a research center focused on digital humanities and computational social sciences. Smith's work has been funded by grants from the Mellon Foundation, NEH, and IMLS, supporting projects such as the Viral Texts initiative analyzing 19th-century newspaper networks and the Oceanic Exchanges project tracking transnational information flows. He has contributed to advancements in OCR for historical texts, text reuse detection, and computational analysis of classical languages. He has advised numerous PhD students, including Shijia Liu, Si Wu, and Ryan Muther, and teaches courses like Natural Language Processing and Information Retrieval. His research has been featured in outlets like Wired and the Economist .
Kevin D. Ashley is a Professor of Law at the University of Pittsburgh School of Law. He is also a senior scientist at the Learning Research and Development Center and an adjunct professor of computer science at the University of Pittsburgh. His interdisciplinary work bridges artificial intelligence, legal analytics, and ethical reasoning. JD, Harvard Law School M.A. and PhD, University of Massachusetts BA, Princeton University His research focuses on computational modeling of legal reasoning , AI and ethics , legal text analytics , and case-based reasoning . He has pioneered AI applications for legal decision support, automated argumentation, and bias detection in legal corpora. Recent publications include studies on large language models in legal annotation, argument mining, and empirical legal analysis. His work has been supported by multiple National Science Foundation grants, emphasizing fairness in AI and access to justice. 2022 Codex Prize for Computational Law 2015 University of Massachusetts Outstanding Achievement Award 2002 AAAI Fellow for AI in Law contributions 2000 Chancellor’s Distinguished Research Award A former President of the International Association for Artificial Intelligence and Law, Ashley has held visiting roles at the University of Bologna, Stanford CodeX, and IBM Watson Research Center. He co-edits the journal Artificial Intelligence and Law and teaches courses on applied legal analytics.
Sara Stymne is a Senior Lecturer in Computational Linguistics at the Department of Linguistics and Philology, Uppsala University, where she has been working since 2012. She initially joined as a post-doc (2012-2015), then worked as a researcher (2015-2017), and served as an assistant professor (2017-2023) before her current position as Senior Lecturer. Prior to Uppsala, she was a researcher at Linköping University's Department of Computer and Information Science. Dr. Stymne earned her PhD in Computational Linguistics from Linköping University in 2012 with the thesis 'Text Harmonization Strategies for Phrase-Based Statistical Machine Translation,' following a Licentiate degree in Computational Linguistics (2009) and a Master's degree in Cognitive Science (2006), both also from Linköping University. During her doctoral studies, she spent the autumn of 2010 and spring of 2009 at Xerox Research Centre Europe in Grenoble, France. Her primary research interests focus on cross-lingual natural language processing and digital humanities, with particular emphasis on multilingual dependency parsing. Dr. Stymne is passionate about applying computational linguistics to solve research questions in other fields, including language history, literary analysis, and political science. Her earlier work concentrated on machine translation, with specific interests in discourse-aware translation, compound processing, and error analysis. She has made significant contributions to the development of language technology tools for analyzing dialogue, narrative, and stylistic features in literature. Analysis of Dr. Stymne's recent publications reveals a strong focus on cross-lingual and cross-domain natural language processing. Her work spans multiple subfields including dependency parsing across genres and topics, discourse relation analysis in low-resource languages like Egyptian Arabic, direct speech identification in Swedish literature, and causality detection in governmental documents. A notable trend is her application of NLP techniques to digital humanities problems, particularly in analyzing literary texts and historical language change. Her research often involves creating and utilizing specialized datasets for specific linguistic phenomena across multiple languages. Dr. Stymne actively supervises graduate students, having guided numerous master's and bachelor's theses on topics ranging from speech recognition to multilingual parsing and causality detection. She leads or participates in several research projects including 'Fictional prose and language change' (funded by VR, 2021-2023) and 'Enabling climate-resilient development' (funded by Marianne and Marcus Wallenberg Foundation, 2023-2027), demonstrating her commitment to interdisciplinary research with practical applications. Her work has resulted in several notable software resources including uuPronPred for cross-lingual pronoun prediction, uuparser for dependency parsing, and Docent for document-level machine translation. Within the Computational Linguistics and Language Technology group at Uppsala University, Dr. Stymne contributes to multiple research initiatives focused on developing language technology tools for digital humanities applications. Her team works closely with literary scholars and historians to create computational methods for analyzing large corpora of literary texts, particularly focusing on Swedish literature across different historical periods. Her research bridges the gap between theoretical computational linguistics and practical applications in the humanities, creating new methodologies for quantitative analysis of literary and historical texts.
Prof. Dr. Gülşen Eryiğit is a Professor at Istanbul Technical University within the Faculty of Computer and Informatics , Department of Artificial Intelligence and Data Engineering . She founded and directs the ITU Natural Language Processing Group , Turkey's leading team in Turkish-language NLP, and serves as Senior Action Editor for ACL RR , Director of ITU TÖMER (Turkish Language Teaching Center), and Co-Chair of the EU UniDive Cost Action WG3. Education: PhD in Computer Engineering from ITU (2007), MSc and BSc from ITU and Marmara University Research: Focuses on Natural Language Processing for Turkish, including dependency parsing , coreference resolution , multiword expressions , and language education technology Her recent work involves multilingual transfer learning , LLM applications for Turkish text simplification, and gamification for morphology education. She has received prestigious awards like the Siemens Excellence Award and TÜBİTAK's Above Threshold Award . Her research has produced 65+ publications and 29+ projects funded by EU, TÜBİTAK, and industry partners. Scientific Awards: Siemens Excellence Award (2007) Above Threshold Award (TÜBİTAK, 2015) Certificate of Appreciation (EU 7th Framework, 2012) Thank You Plaque (ITU, 2017) The ITU NLP Group under her leadership has developed Turkey's first licensed NLP software exported internationally. She collaborates with European institutions through COST actions and participates in ACL, CoNLL, and LREC conferences. Her lab focuses on language technology for Turkish , including sign language processing and social media normalization.
Mark Gales is Professor of Information Engineering at the University of Cambridge and an Official Fellow at Emmanuel College. He is currently on sabbatical leave for the 2024/25 academic year. Prior to his academic career, he worked as a consultant at Roke Manor Research Ltd, developing radar systems, before transitioning to speech and language processing. PhD in 'Model-Based Techniques for Robust Speech Recognition' (University of Cambridge, 1995) BA in Electrical and Information Sciences (University of Cambridge, 1988) His research focuses on speech and language processing , particularly in automated language assessment and low-resource speech technology . He leads the Automated Language Teaching and Assessment (ALTA) Institute , which collaborates with Cambridge University Press & Assessment (CUP&A) to develop commercial tools like Linguaskill and Speak & Improve . These platforms provide automated spoken/written assessment for millions of users globally. Recent publications highlight his work in LLM-driven speech processing , including adversarial attacks on foundation models, end-to-end spoken error correction, and uncertainty estimation frameworks. His team's research spans multilingual capabilities, with deployments in languages ranging from Dholuo to Tok Pisin . Awards : IEEE Fellow, ISCA Fellow Leadership : Fellows' Steward at Emmanuel College Mark has contributed extensively to Hidden Markov Model (HMM) applications in speech recognition, which underpinned early automatic speech systems. His work now bridges LLM-based language assessment with cross-lingual transfer learning and robustness testing for real-world deployments.
Dr. Ly Fie Sugianto is an Associate Professor in the Department of Accounting at Monash Business School, Monash University. Her research focuses on the integration of data analytics, artificial intelligence, and machine learning in accounting and business systems, with applications in the energy sector and organizational behavior. Monash Business School, Monash University Department of Accounting Specialization: Accounting Information Systems, Data Analytics, AI Her research interests span Accounting Information Systems , Agent-Based Simulation , Decision Support Systems , and Technology Adoption . She applies computational methods to study competitive dynamics in deregulated electricity markets and the impact of digital tools on employee well-being and organizational resilience. The recent publications reflect a strong trend in using AI and simulation to analyze complex socio-technical systems, particularly in energy markets and leadership dynamics. Keywords across her work include agent-based modeling , data analytics , servant leadership , and enterprise social media , indicating interdisciplinary research at the intersection of information systems, management, and public policy. Her scientific awards include competitive grants from the ARC (SPIRT/Linkage) , the Australia Indonesia Governance Research Partnership (AIGRP) , and the Sumitomo Foundation . ARC Grant: Dispatch Optimisation in the Australian National Electricity Market Sumitomo Foundation: Technology Use and Employee Well-Being AIGRP: Governance and MSME Resilience during Pandemic Dr. Sugianto has advised research projects and collaborated with industry partners such as Western Power , Ecogen Energy , and AEMO . She is currently accepting PhD students and leads externally funded research initiatives. Her work contributes to UN Sustainable Development Goals related to industry innovation and responsible consumption. She is affiliated with research teams focusing on intelligent decision support systems and digital transformation in business , with active collaborations in Australia and Indonesia.
Gerold Schneider is an Associate Professor at the University of Zurich , affiliated with the Department of Computational Linguistics under the Faculty of Arts and Social Sciences and Faculty of Business, Economics and Informatics . He leads the Text Crunching Center (TCC) , focusing on interdisciplinary research at the intersection of NLP, Digital Humanities, and Health Data Science. Research Interests His work spans Text Analytics , Digital Humanities , Corpus Linguistics , and Health Data Science , with applications in: Biomedical NLP (e.g., Alzheimer’s detection, clinical trials) Digital Humanities projects (e.g., analyzing Charles Dickens, UN archives) Migration discourse framing across languages Adversarial data collection for hate speech detection Interdisciplinary methodologies for digital unstructured data Recent Publications 2025–2024 research highlights include annotated corpora for preclinical and neurological studies, AI-driven analysis of historical linguistic variation, and innovative tools for language learners. His NLP applications address health diagnostics, ethical AI, and cross-lingual political discourse. Labs & Teams As TCC leader, he spearheads collaborative projects within the Digital Society Initiative (DSI) communities (AI & Law, Health, Ethics, etc.), integrating computational methods with humanities and health research.
Mohammad Aliannejadi is an Assistant Professor at the IRLab (formerly ILPS) within the Informatics Institute at the University of Amsterdam. His research focuses on Information Retrieval (IR), machine learning, natural language processing (NLP), and conversational systems, particularly in modeling user information needs on mobile devices and conversational search systems. He holds a Ph.D. in Informatics from Università della Svizzera italiana (USI), Lugano, Switzerland, and a M.Sc. in Computer Engineering from Tehran Polytechnic. During his Ph.D., he visited the CIIR Lab at the University of Massachusetts Amherst, USA. His research interests include conversational search systems, recommender systems, unified search frameworks, and user-centric evaluation methodologies. Notable contributions include work on clarifying questions in open-domain dialogues, contextual suggestion systems, and cross-market recommendation. Aliannejadi has organized major shared tasks and workshops, including the IGLU Contest (NeurIPS 2021) and XMRec Workshop (RecSys 2021). He serves on program committees of top IR conferences like SIGIR, CIKM, and ECIR, and has authored over 50 peer-reviewed publications in these areas. His work has received recognition, including top performance in TREC Contextual Suggestion tracks (2015, 2016). He actively contributes to the IR community through teaching, including courses on Information Retrieval and Human-in-the-Loop Machine Learning at the University of Amsterdam.