Paul RaysonView profile
Professor
Professor Paul Rayson is a Professor of Natural Language Processing in the School of Computing & Communications at Lancaster University, UK. He serves as Director of the UCREL (University Centre for Computer Corpus Research on Language) interdisciplinary research centre and is affiliated with multiple research institutes including Security Lancaster, the Lancaster Centre for Digital Humanities, and the Data Science Institute. Education: PhD in Computer Science, Lancaster University (2003) BSc (Hons) Computer Science and Mathematics, Lancaster University (1990) Professor Rayson's research focuses on semantic multilingual Natural Language Processing (NLP) in challenging linguistic environments with noisy language data, including historical texts, learner language, speech, email, and other computer-mediated communication. His work spans applications in dementia detection, mental health analysis, online child protection, cyber security, learner dictionaries, and text mining of biomedical literature, historical corpora, and financial narratives. He has developed semantic tagging tools like USAS (UCREL Semantic Analysis System) and Wmatrix for corpus analysis. Major Awards and Honors: FHEA (Fellow of the Higher Education Academy) MBCS (Member of the British Computer Society) Professor Rayson has supervised numerous PhD students in NLP and corpus linguistics, with eight current students and seven completed doctorates. He has led or co-investigated multiple major research projects including the £3.5m ESRC-funded Centre for Corpus Approaches to Social Science (CASS), the National Corpus of Contemporary Welsh, and projects related to mental health forums, financial narrative analysis, and cyber security. His research has been supported by ESRC, EPSRC, and other funding bodies. As Director of UCREL, he oversees research in corpus linguistics and NLP. He is also active in the Cyber Security Research Centre, Digital Health Group, and multiple Data Science Institute initiatives. His lab has developed several widely-used NLP tools including CLAWS for English POS tagging, USAS semantic analysis system, Wmatrix corpus analysis tool, and the Variant Detector (VARD) for historical texts.











