Colin Raffel , currently an Associate Professor at the University of Toronto and Associate Research Director at the Vector Institute , is a leading researcher in machine learning and natural language processing . His career spans roles at Hugging Face (Faculty Researcher), Google Brain (Senior Research Scientist), and UNC Chapel Hill (Assistant Professor). Education: PhD in Electrical Engineering (Columbia), MA in Music/Science (Stanford), BA in Mathematics (Oberlin) Key affiliations: Google Brain (2016-2020), Hugging Face (2021-present), Vector Institute (2023-present) His research focuses on language model development , attention mechanisms , efficient machine learning , and music information retrieval . Recent work explores model merging , parameter-efficient fine-tuning , and data-constrained language models . Teaching : Has instructed courses at University of Toronto and UNC Chapel Hill on Neural Networks , Deep Learning , and Information Theory . Academic service includes organizing ICLR workshops and serving as Senior Area Chair for NeurIPS and EMNLP . Notable awards : NSF CAREER (2022), Caspar Bowden Award (2023), NeurIPS Outstanding Paper (2023) Key contributions : Core developer of WT5 , Git-Theta , and mir_eval software
Adriana Kovashka is an Associate Professor at the University of Pittsburgh , affiliated with the School of Computing and Information and serving as Department Chair . Her academic journey began with BA degrees in Computer Science and Media Studies from Pomona College (2008) and a PhD in Computer Science from The University of Texas at Austin (2014). Joined Pitt’s faculty in January 2015 NSF CAREER awardee (2021) Google Faculty Research Award recipient Dr. Kovashka’s research spans Computer Vision , Machine Learning , and Natural Language Processing , focusing on visual rhetoric, weak multimodal supervision, and domain adaptation. She pioneered techniques for analyzing political imagery, developing robust object detection frameworks, and exploring the intersection of visual and textual persuasion through large-scale annotated datasets. Her recent work emphasizes geographic diversity in vision-language systems, audio-visual fusion for domain generalization, and shape-texture bias mitigation in CNNs. Key publications include groundbreaking studies on symbolic reasoning, multimodal dialogue systems, and ethical AI applications in education. Scientific honors include: NSF CRII Award (2016) NSF CAREER Award (2021) Pitt CRDF Award (2016, 2018) Best Paper at ECV Workshop (2021) Google Faculty Research Award (2016, 2018) Dr. Kovashka actively mentors students in multimodal learning projects and collaborates with interdisciplinary teams on NSF-funded initiatives. She co-organizes workshops like the first CVPR workshop on advertisement understanding and leads research groups exploring human-AI co-learning systems.
Professor Bernd Möbius is a leading academic in Phonetics and Phonology at the Department of Language Science and Technology, Saarland University. His research bridges phonetic theory with speech technology applications, focusing on text-to-speech systems, prosody modeling, and computational simulations of speech processes. Current research projects: DFG SFB 1102, C1: Information density and phonetic structure predictability DFG SFB 1102, C4: Slavic intercomprehension and surprisal theory (INCOMSLAV) Research Themes: Key areas include text-to-speech synthesis, speech prosody analysis, experimental methods in speech production/perception, information density in phonetics, and cross-linguistic studies of Slavic-Germanic languages. Scientific Contributions: Recent work explores Parkinson-induced dysarthria detection, breath noise acoustics, surprisal-driven speech behaviors, multilingual BERT models for idiomaticity, and perceptual consequences of acoustic adjustments.
Simon King is a Professor of Speech Processing at the University of Edinburgh , affiliated with the School of Philosophy, Psychology and Language Sciences . He serves as Director of the Centre for Speech Technology Research (CSTR) and teaches courses like Speech Processing and Speech Synthesis , while directing the MSc in Speech and Language Processing . Research Interests His research focuses on: Developing new acoustic models (e.g., Linear Dynamical Models, factorial-HMMs) for speech recognition Advancing unit selection and HMM-based speech synthesis Integrating articulatory measurement data for enhanced modeling Exploring perceptual measures in synthesis criteria Building multilingual speech systems to identify universal speech building blocks Publication Trends Simon's recent work emphasizes deep learning (DNNs, LSTMs) in speech synthesis, multilingual frameworks , and articulatory-acoustic feature integration . His studies often bridge grapheme-based modeling , perceptual error reduction , and noise-robust synthesis . Scientific Awards EPSRC Advanced Research Fellowship (2005-2009) Students & Collaborations He has supervised numerous PhD students including Rasmus Dall, Tom Merritt, and Srikanth Ronanki. Current research fellows like Mirjam Wester and Zhizheng Wu contribute to projects such as Natural Speech Technology (NST) and Simple4All .
Bo Li is an Associate Professor at the University of Illinois at Urbana-Champaign, affiliated with the Siebel School of Computing and Data Science. Her research focuses on trustworthy machine learning, emphasizing robustness, privacy, and security in AI systems. She leads the Secure Learning Lab (SL²), exploring adversarial attacks and defenses across digital and physical domains. Key contributions include foundational work on adversarial examples, robust learning frameworks, and privacy-preserving techniques. Her academic roles include advisory board positions at the Center for Artificial Intelligence Innovation (CAII) and membership in the Information Trust Institute (ITI). She collaborates with institutions like the Advanced Digital Science Center (ADSC) and the Quantum Information Science and Technology Center (IQUIST). Notable recognitions include the IJCAI Computers and Thought Award (2022), MIT Technology Review's 35 Innovators Under 35 (2020), and multiple best paper awards. Recent work addresses AI safety through frameworks like ShieldAgent and AutoRedTeamer, aiming to enhance system resilience against adversarial threats. Her research spans theoretical guarantees, practical defenses, and ethical AI deployment. Students advised include Chulin Xie, Linyi Li, and Boxin Wang, who have received prestigious fellowships such as the IBM PhD Fellowship and Rising Stars in ML Awards. Advising: Guides PhD students in adversarial ML, privacy, and security. Labs/Teams: Secure Learning Lab (SL²), collaboration with ALERT program. Grants/Funding: NSF CAREER Award, Amazon/Google Faculty Awards, and industry partnerships.
Srijan Kumar is an Assistant Professor in the School of Computational Science and Engineering at Georgia Institute of Technology's College of Computing. His research focuses on data science, AI for security, and online safety, addressing challenges in detecting malicious users, misinformation, and enhancing AI robustness. His work has been deployed in platforms like Flipkart and Wikipedia, and recognized through awards such as the NSF CAREER and Forbes 30 Under 30. Education: B.Tech from Indian Institute of Technology, Kharagpur Ph.D. in Computer Science from University of Maryland, College Park Postdoctoral training at Stanford University Research Interests: Multi-modal/multi-lingual detection of harmful content and users Adversarial robustness of AI models Graph and network analysis for early detection Responsible recommender systems Awards & Grants: NSF CAREER Award (2023) Kavli Fellow (2022) Facebook/Adobe Faculty Awards NSF Convergence Accelerator Phase II grant ($5M) Advising & Labs: Leads the CLAWS Lab, advising over 20 students. Active in mentoring through conferences and NSF-funded projects.
David Alvarez-Melis is an Assistant Professor of Computer Science at Harvard University's John A. Paulson School of Engineering and Applied Sciences (SEAS). He leads the Data-Centric Machine Learning (DCML) group and holds affiliations with the Kempner Institute, Harvard Data Science Initiative, and the Center for Research on Computation and Society. His research focuses on making machine learning more data-efficient and trustworthy, with applications in natural and medical sciences. He also serves as a researcher at Microsoft Research New England. Affiliations: SEAS, Kempner Institute, Harvard Data Science Initiative, CRCS Education: PhD in Computer Science (MIT), MS in Mathematics (NYU Courant), BSc in Applied Mathematics (ITAM) Research Interests: Optimal Transport, dataset distillation, interpretable AI, medical imaging, robustness, and large language models. His work bridges theory and applications, emphasizing geometric and probabilistic methods. Recent Trends in Publications: Focused on advancing optimal transport for data manipulation, distributional deep equilibrium models, and repurposing LLMs for specialized domains. Key themes include synthetic dataset generation, gradient flows in probability spaces, and robust interpretability frameworks. Awards: Aramont Fellowship, Dean’s Competitive Fund, Top Reviewer awards at major conferences (ICLR, NeurIPS, ICML). Grants: Supported by the Aramont Fund and Harvard’s Dean’s Fund. His lab advises students across Harvard and MIT, with notable contributions to medical imaging, NLP, and foundational ML theory. He actively mentors interns and fosters collaborations with industry and academia.
Sara Stymne is a Senior Lecturer in Computational Linguistics at the Department of Linguistics and Philology, Uppsala University, where she has been working since 2012. She initially joined as a post-doc (2012-2015), then worked as a researcher (2015-2017), and served as an assistant professor (2017-2023) before her current position as Senior Lecturer. Prior to Uppsala, she was a researcher at Linköping University's Department of Computer and Information Science. Dr. Stymne earned her PhD in Computational Linguistics from Linköping University in 2012 with the thesis 'Text Harmonization Strategies for Phrase-Based Statistical Machine Translation,' following a Licentiate degree in Computational Linguistics (2009) and a Master's degree in Cognitive Science (2006), both also from Linköping University. During her doctoral studies, she spent the autumn of 2010 and spring of 2009 at Xerox Research Centre Europe in Grenoble, France. Her primary research interests focus on cross-lingual natural language processing and digital humanities, with particular emphasis on multilingual dependency parsing. Dr. Stymne is passionate about applying computational linguistics to solve research questions in other fields, including language history, literary analysis, and political science. Her earlier work concentrated on machine translation, with specific interests in discourse-aware translation, compound processing, and error analysis. She has made significant contributions to the development of language technology tools for analyzing dialogue, narrative, and stylistic features in literature. Analysis of Dr. Stymne's recent publications reveals a strong focus on cross-lingual and cross-domain natural language processing. Her work spans multiple subfields including dependency parsing across genres and topics, discourse relation analysis in low-resource languages like Egyptian Arabic, direct speech identification in Swedish literature, and causality detection in governmental documents. A notable trend is her application of NLP techniques to digital humanities problems, particularly in analyzing literary texts and historical language change. Her research often involves creating and utilizing specialized datasets for specific linguistic phenomena across multiple languages. Dr. Stymne actively supervises graduate students, having guided numerous master's and bachelor's theses on topics ranging from speech recognition to multilingual parsing and causality detection. She leads or participates in several research projects including 'Fictional prose and language change' (funded by VR, 2021-2023) and 'Enabling climate-resilient development' (funded by Marianne and Marcus Wallenberg Foundation, 2023-2027), demonstrating her commitment to interdisciplinary research with practical applications. Her work has resulted in several notable software resources including uuPronPred for cross-lingual pronoun prediction, uuparser for dependency parsing, and Docent for document-level machine translation. Within the Computational Linguistics and Language Technology group at Uppsala University, Dr. Stymne contributes to multiple research initiatives focused on developing language technology tools for digital humanities applications. Her team works closely with literary scholars and historians to create computational methods for analyzing large corpora of literary texts, particularly focusing on Swedish literature across different historical periods. Her research bridges the gap between theoretical computational linguistics and practical applications in the humanities, creating new methodologies for quantitative analysis of literary and historical texts.
Zhiyong Huang is an Associate Professor at the National University of Singapore (NUS) School of Computing. He holds multiple leadership roles, including Deputy Director of the NUS Business Analytics Centre, Director of the Computing Translational Research & Development (C-TReND) Centre, and Assistant Dean (Industry Relations). He is also a Senior Principal Investigator at the NUS Chongqing Research Institute. Education: PhD in Computer Science from École Polytechnique Fédérale de Lausanne (EPFL), MEng and BEng in Computer Engineering from Tsinghua University. Leadership: Senior Member of ACM and IEEE, Pioneer Member of ACM SIGGRAPH, and Chair of the Singapore ACM SIGGRAPH Chapter. Research Interests: His work spans Data Analytics, Machine Learning, Computer Vision, Human-Robot Interaction, and Computer Graphics. Key projects include the NUS Digital Twin initiative and secure data analytics pipelines. Article Trends: Recent publications focus on time series generation, cryptocurrency benchmarks, medical image registration, and phishing detection. These works integrate machine learning, computer vision, and multimodal systems. Scientific Awards: Finalist, World Technology Summit & Awards (Entertainment, 2010) Bronze, National Science and Technology Progress Award (1992) Tsinghua 12.9 Distinguished Young Teacher Award (1989) Grants & Service: Extensive involvement in Singapore's IT Standards Committee, review panels for EDB SIIRD projects, and editorial roles. He has served as PC co-chair, local chair, and reviewer for numerous conferences and journals.
Dr. Muhammad Abdul-Mageed is an Associate Professor in the School of Information at The University of British Columbia, with joint appointments in Linguistics and an associate membership in Computer Science. He holds the Canada Research Chair in Natural Language Processing and Machine Learning. His research focuses on deep learning, socio-pragmatics, and speech/language technologies, particularly for Arabic and African languages. He leads the UBC Deep Learning & NLP Group and co-directs SSHRC-funded grants like I Trust AI and Ensuring Full Literacy. He is a founding member of the Center for Artificial Intelligence Decision making and Action and a member of the Institute for Computing, Information, and Cognitive Systems. His work spans automatic speech recognition, machine translation, computational socio-pragmatics, and low-resource language technologies. Notable projects include developing Arabic speech recognition systems, multidialectal Arabic benchmarks, and tools for African language processing. He has authored over 100 peer-reviewed papers and leads initiatives like the NADI Arabic Dialect Identification shared task and the NileChat project for culturally-aware LLMs. His research aims to create equitable, socially-aware AI systems for health, social media, and information management.
William Yang Wang serves as the Mellichamp Professor of Artificial Intelligence at the University of California, Santa Barbara (2019-present). He directs the UCSB Center for Responsible Machine Learning, the Mind and Machine Intelligence Initiative, and the UCSB NLP Group. His research focuses on theoretical foundations and practical algorithms for AI, particularly in NLP, LLMs, and neuro-symbolic reasoning. PhD in Computer Science from Carnegie Mellon University Active in AI theory and applications (2016-present) Research interests span multiple AI domains, with special emphasis on NLP and responsible machine learning. He has pioneered datasets like HybridQA, TabFact, and VaTeX, enabling advancements in multi-hop QA, fact verification, and video-language tasks. His work combines statistical relational learning with modern deep learning paradigms. Recent publications center around multimodal reasoning, knowledge graph integration, and responsible AI development. He has received numerous accolades including the IEEE SPS Pierre-Simon Laplace Award (2024) and NSF CAREER Award (2021). Karen Sparck Jones Award (2022) DARPA Young Faculty Award (2018) IBM Faculty Award Mentoring 15+ PhD and postdoc researchers who now hold positions at Microsoft Research, Amazon, Meta GenAI, and academic institutions like Arizona and Rutgers. His lab maintains active collaborations with industry partners through initiatives like ChipAgents.ai, which he founded as CEO.
Raquel Fernández is Full Professor of Computational Linguistics and Dialogue Systems at the University of Amsterdam, where she leads the Dialogue Modelling Group at the Institute for Logic, Language & Computation (ILLC). As Vice-Director for Research at ILLC and a Fellow of the ELLIS Society, she bridges computational linguistics, cognitive science, and artificial intelligence through her research on language use in multimodal and conversational contexts. PhD in Computational Linguistics from King's College London Prior research positions at University of Potsdam and Stanford University's CSLI Her work explores how cognitive constraints, social interaction, and perception shape language use, with a focus on: Visually-grounded language processing Multimodal dialogue modeling Model uncertainty and calibration Language grounding in multimodal data Language learning and semantic change Dialogue reference resolution Recent publications analyze multimodal reasoning limitations, cross-lingual knowledge consistency, and uncertainty modeling in dialogue systems. She has received multiple accolades including an ERC Consolidator Grant , NWO VENI/VIDI/Aspasia fellowships , and EMNLP/GenBench awards . Outstanding Paper Award (EMNLP 2023) Best Data Award (GenBench Workshop 2023) ELLIS Society Fellow ERC Consolidator Grant #819455 recipient NWO VENI/VIDI/Aspasia awardee As a leader in academic service, she serves on the SIGDAT Executive Committee and chairs multiple conference committees. Her lab develops models for multimodal dialogue, visual storytelling, and grounded language understanding.
Margaret E. Roberts is a Professor in the Department of Political Science at the University of California, San Diego. She co-directs the China Data Lab at the 21st Century China Center and serves as an affiliate at the UC Institute on Global Conflict and Cooperation. Her academic appointments reflect her interdisciplinary approach combining political science, statistics, and computational methods. University of California, San Diego, Department of Political Science (Current) Co-director, China Data Lab at the 21st Century China Center Affiliate, UC Institute on Global Conflict and Cooperation Roberts earned her PhD in Government from Harvard University (2014), MS in Statistics from Stanford University (2009), and BA in International Relations and Economics from Stanford University (2009). Her educational background bridges political science, statistics, and computational methods, forming the foundation for her interdisciplinary research approach. Professor Roberts' research focuses on the intersection of political methodology and the politics of information, with specific expertise in automated content analysis and the politics of censorship and propaganda in China. Her work employs innovative methods including social media analysis, online experiments, and large-scale text analysis to understand how censorship and propaganda influence information access and political beliefs. She has made significant contributions to text-as-data methodologies, developing tools like the Structural Topic Model (stm) R package that have become widely used in social science research. Roberts' research portfolio demonstrates consistent focus on authoritarian information control, particularly in China, while expanding into broader applications of text analysis in political science. Her publications span top journals in political science, computer science, and interdisciplinary fields, reflecting the cross-disciplinary nature of her work. Goldsmith Book Award Best Book Award in the Human Rights Section Best Book Award in Information Technology and Politics Section Best Book Award of the last decade in the Political Communication Section of the American Political Science Association Chancellor's Associates Endowed Chair at UCSD Foreign Affairs Best Books of 2018 Professor Roberts has secured significant research funding supporting her work on Chinese censorship, propaganda, and text analysis methodologies. Her research has practical applications for understanding digital authoritarianism, content moderation, and the development of computational tools for social science research. She has mentored numerous students and collaborators, contributing to the next generation of scholars working at the intersection of political science and computational methods. Roberts also leads the China Data Lab, which serves as a hub for research on Chinese politics and society using digital methods.
Arman Cohan is an Assistant Professor of Computer Science at Yale University, affiliated with the School of Engineering & Applied Science. His research focuses on the intersection of Machine Learning and Natural Language Processing (NLP), particularly in language modeling, representation learning, retrieval systems, and applications in specialized domains such as scientific text processing. He earned his Ph.D. in Computer Science from Georgetown University and has received notable awards, including the Dr. Harold N. Glassman Distinguished Doctoral Dissertation Award (2019) and the EMNLP 2017 Best Long Paper Award. His work emphasizes ethical AI, robustness of LLMs, and interdisciplinary applications in healthcare, science, and education. Cohan's research group, the Yale NLP Lab, develops advanced techniques for multi-document summarization, adversarial fact-checking, and LLM-driven tools for scientific discovery. Recent projects include frameworks like SciBERT, Longformer, and ChemAgent, which enhance domain-specific reasoning and safety in AI systems. His publications address challenges in table reasoning, uncertainty expression, and multimodal reasoning, with applications in medical decision-making and educational problem-solving. He collaborates on initiatives like the Roberts Innovation Fund to advance AI in healthcare and environmental technology.
Dr. Aditya Joshi is a Senior Lecturer in the School of Computer Science & Engineering at the University of New South Wales (UNSW). He specializes in Natural Language Processing (NLP), with a focus on sarcasm detection, dialectal NLP, and ethical AI applications in public health and cybersecurity. He joined UNSW in 2023 following industry roles at SEEK, Notiv, and Fractal Analytics, where he developed NLP systems for recommendation engines and meeting analytics. His research has garnered over 3,000 citations (h-index 26) and secured $3.1M in grants, including Defence Trailblazer and Google exploreCSR awards. Education: Joint PhD (2018) from IIT Bombay (India) and Monash University (Australia); MTech in CSE (2011) from IIT Bombay. Research Interests: Making NLP models robust for non-native English speakers and the LGBTI+ community, algorithmic enhancements to transformers, and applications in public health, cybersecurity, and societal issues. His work spans epidemic intelligence (collaborations with EPIWATCH and IFCYBER), cybersecurity tools like AuditNet, and inclusive AI initiatives such as queer-inclusive workshops funded by Google. He designed UNSW's new NLP course (COMP6713) and co-authored a Wiley textbook on NLP. Notable grants include the A$1.4M 'Comprehensive Defence Data Platform' (Lead CI) and A$92K Google exploreCSR grant for benchmarking dialectal sentiment. His awards include the Best PhD Thesis from IITB-Monash and Best Paper accolades at FAccT 2023 and MoMM 2020. He supervises projects on kernel-based attention reformulation, prompt-based sarcasm detection, and multilingual small-scale LLMs. His service roles include Executive Committee Member at ALTA and arXiv moderator for computational linguistics.