Sunita Sarawagi is a Professor at Computer Science and Engineering , IIT Bombay, and a member of AI Labs@CSE . She is also associated with the Center for Machine Intelligence and Data Science (CMInDS), which she founded in 2020. Education: PhD in Computer Science from UC Berkeley (Thesis: Query Processing in Tertiary Memory Databases), BTech in Computer Science from IIT Kharagpur Research Interests span machine learning , data analytics , graphical models , and structured learning , with applications in text segmentation, sequence modeling, domain adaptation, and human-in-the-loop systems. Her publications reveal a strong focus on integrating data mining with database systems , temporal data analysis , and information extraction using probabilistic methods. Professional Activities include serving on the IEEE John Von Neumann Medal committee (2017-), VLDB 2011 Research Track Co-chair , and multiple program committee roles at top conferences like ICML, KDD, and SIGMOD. Labs & Teams : Leads the SS Lab , a research group focused on probabilistic graphical models, sequence modeling, and data integration techniques.
Antoine Bosselut is a Tenure Track Assistant Professor at École Polytechnique Fédérale de Lausanne (EPFL) in the School of Computer and Communication Sciences, where he leads the EPFL NLP group. His research focuses on developing AI reasoning agents that can model, represent, and reason about human and world knowledge, with applications in health, education, and global fairness. His research interests span multiple critical areas in modern AI: LLM Representations of Knowledge: Understanding what language models know and how they represent that knowledge internally Reasoning Algorithms: Developing methods to improve LLMs' reasoning capabilities through symbolic systems, neuroscience, and cognitive science Large-scale AI Development: Creating open-source foundation models with multilingual capabilities AI Democratization: Ensuring equitable access to AI technologies across different cultural and regulatory contexts His recent publications reveal a strong focus on the intersection of language models and cognitive science, with multiple studies examining how LLMs align with human cognition. He's also deeply engaged in practical applications of NLP technology, particularly in education and multilingual settings, as evidenced by his work on evaluating AI's impact on higher education and developing culturally-aware language models. His scientific achievements have been recognized with several prestigious awards: Outstanding Paper Award at NAACL 2025 ELLIS Scholar designation in 2024 Outstanding Paper Award at ACL 2023 Inclusion in Forbes 30 Under 30 list for Science & Healthcare in 2021 Bosselut actively mentors numerous PhD students across diverse NLP research areas and has secured significant recognition for his work in both academic and media circles. His lab has been featured in major publications including Communications of the ACM, The Atlantic, and Quanta Magazine, highlighting the societal impact of his research on commonsense reasoning and AI capabilities.
David Bamman is an Associate Professor in the School of Information at UC Berkeley, specializing in applying Natural Language Processing (NLP) and machine learning to cultural and social science questions. He leads research in born-literary NLP, computational humanities, and cultural analytics, with affiliated roles in EECS, Linguistics, and Computational Precision Health. Bamman holds degrees from Carnegie Mellon (Ph.D., 2015), Boston University (M.A., 2006), and University of Wisconsin-Madison (B.A., 1998). His work is supported by NEH, NSF, and industry grants. Educations: Ph.D. in Computer Science (2015), Carnegie Mellon University M.A. in Applied Linguistics (2006), Boston University B.A. in Classics (1998), University of Wisconsin-Madison Research Interests: NLP for underserved domains (e.g., literature, social media), coreference resolution, cultural analytics, and computational methods for studying literature and culture. Projects include LitBank and BookNLP datasets. Grants & Awards: Hellman Fellow (2019), Amazon Research Award (2017), NSF CAREER Award, and NEH funding. Teaching: Courses include Natural Language Processing (Info 159/259), Computational Humanities (INFO 190), and Applied NLP (INFO 256). His research group explores topics like racial representation in high school literature, Hollywood diversity metrics, and the sociocultural implications of LLMs. Bamman advises multiple PhD students and collaborates on datasets like CMU Book Summaries and 11K Latin Books.
Alan Ritter is an Associate Professor at the School of Interactive Computing , Georgia Institute of Technology, with additional affiliation to the Machine Learning Center . His research focuses on Natural Language Processing , particularly robust models across domains/languages with fewer labels and efficient resource use, plus data-driven dialogue agents for open-topic conversations. Research Interests : Robust NLP models, cross-lingual transfer, resource-efficient learning, dialogue systems, cultural bias measurement, and privacy-aware language models Students : Mentors Ph.D. students in Georgia Tech's ML and CS programs, including Junmo Kang, Yang Chen, and Duong Minh Le. Alumni include Fan Bai (Ph.D. 2023), Yang Chen (Ph.D. 2024), and Andrew Li (M.S. 2024). Awards : NSF CAREER Award, Amazon Research Award, ACL 2024 Best Social Impact Paper, IUI 2009 Best Student Paper. Recent Work : Studies training budget allocation between supervised and preference-based finetuning, cross-lingual information extraction, cultural bias in LLMs, and privacy risk mitigation in social media disclosures. Service : Served as Program Chair for NAACL 2025, Area Chair for multiple top-tier conferences (COLM, EMNLP, ACL, EACL, AAAI). Email : alan.ritter@cc.gatech.edu
Professor Ihab Ilyas is a leading figure in data management and machine learning at the Cheriton School of Computer Science , University of Waterloo , where he holds the Thomson Reuters Research Chair in Data Quality . He is currently on leave from the university, serving as a Distinguished Engineer, Proactive Intelligence at Apple Inc . His research focuses on data cleaning , large-scale data integration , knowledge graphs , and machine learning applications in data quality . He has pioneered systems like HoloClean and Saga , with commercial impacts through co-founded startups Inductiv (acquired by Apple) and Tamr . PhD in Computer Science from Purdue University Co-founder of Inductiv (acquired by Apple) Co-founder of Tamr Research Interests include: Probabilistic and uncertain data management Machine learning for data quality and enrichment Big data systems and information extraction Knowledge graph construction and optimization Scientific Awards and Recognitions include: Fellow of the Royal Society of Canada (2024) C.C. Gotlieb Computer Award (2024) IEEE Fellow (2021) ACM Fellow (2020) Cheriton Faculty Fellowship (2013-2016) Ontario Early Researcher Award
Maya Ramanath is an Associate Professor in the Department of Computer Science and Engineering at Indian Institute of Technology (IIT) Delhi. She joined IIT Delhi in 2011 after a postdoctoral research stint at the Max-Planck Institute for Informatics in Germany. Her research interests focus on database systems, information retrieval, semantic web technologies, and knowledge graph construction and applications. Education: PhD in Computer Science, Indian Institute of Science, Bangalore M.Sc.(Engg.) in Computer Science, Indian Institute of Science, Bangalore B.E. in Computer Science and Engineering, Bangalore University, Bangalore Her recent work emphasizes efficient query processing over large-scale graphs, knowledge graph applications, and natural language interfaces for semantic data. Notable contributions include algorithms for reachability approximation in web-scale graphs, speculative query planning for knowledge graphs, and exploratory querying techniques. She has collaborated extensively on projects like NAGA, ESTHETE, and KlusTree, advancing the state of the art in graph-based data management and semantic search. Publications span conferences such as ICDE, ECIR, EDBT, and VLDB, reflecting a strong focus on database systems, graph algorithms, and semantic web applications. Her work bridges theoretical foundations with practical implementations, addressing scalability and efficiency challenges in modern data management systems. Research and advising activities include supervision of projects on distributed graph processing, query optimization, and knowledge representation. She has contributed to open-source tools like LegoDB and StatiX, and her lab focuses on interdisciplinary approaches to data-centric AI.
Christos Faloutsos is the Fredkin Professor of Computer Science at Carnegie Mellon University, with a courtesy appointment in Electrical and Computer Engineering. He holds a B.Sc. from the National Technical University of Athens and M.Sc./Ph.D. from the University of Toronto. His research focuses on data mining, graph analysis, fractals, and database systems. Notable contributions include foundational work on R-trees, graph mining laws (e.g., Kronecker graphs), and applications in medical imaging, network security, and fraud detection. Key projects include PEGASUS (petascale graph mining), fraud detection in online auctions (NetProbe), and tools for human trafficking analysis (TrafficVis). He has led NSF-funded projects on tensor mining, network anomaly detection, and bioinformatics. Over 300 refereed publications highlight his contributions across databases, data mining, and networks. Awards include the KDD Best Paper (2005, 2016), SIGMOD Test-of-Time Award, and recognition as a top nurturer in IT. His lab collaborations span the Parallel Data Lab (PDL), Machine Learning Department, and Computational Biology.
David A. Smith is an Associate Professor at the Khoury College of Computer Sciences, Northeastern University. His research focuses on Natural Language Processing (NLP) and computational linguistics, with applications in machine translation, information retrieval, digital humanities, and social sciences. He is a founding member of the NULab for Texts, Maps, and Networks, a research center focused on digital humanities and computational social sciences. Smith's work has been funded by grants from the Mellon Foundation, NEH, and IMLS, supporting projects such as the Viral Texts initiative analyzing 19th-century newspaper networks and the Oceanic Exchanges project tracking transnational information flows. He has contributed to advancements in OCR for historical texts, text reuse detection, and computational analysis of classical languages. He has advised numerous PhD students, including Shijia Liu, Si Wu, and Ryan Muther, and teaches courses like Natural Language Processing and Information Retrieval. His research has been featured in outlets like Wired and the Economist .
Byron Wallace is the Sy and Laurie Sternberg Interdisciplinary Associate Professor at Northeastern University's Khoury College of Computer Sciences, where he also serves as Associate Dean for Research and Director of the BS in Data Science Program. His research focuses on Natural Language Processing and Machine Learning applications in healthcare. Education Details of formal education are not explicitly provided in the available text, though he holds a PhD from Tufts University (mentioned in thesis award context). Research Interests His work centers on developing NLP and ML models for health applications, with particular emphasis on: Biomedical evidence synthesis automation Electronic Health Record processing Model interpretability and trustworthiness Human-in-the-loop systems Learning with limited supervision Research Trends Recent publications demonstrate strong focus on large language model applications in healthcare, including factuality evaluation for medical summarization, evidence extraction from clinical trials, and interpretable risk prediction models. Notable contributions include work on GPT-3 applications in medical evidence synthesis and neural methods for EHR analysis. Scientific Awards ACL Outstanding Paper Award (2022) ICLR Spotlight (top 5% acceptance) (2024) Best Student-led Paper Award at AMIA 2021 NSF CAREER Award (2018-2023) Advising and Grants Currently advises 4 PhD students and has mentored numerous others. Major funding includes: NSF CAREER Award ($500K+) NIH R01 grant for EHR summarization NSF Medium grant for healthcare summarization Support from Army Research Office, Amazon, and Seton Hospital Labs and Teams Leads the Evidence Inference project team working on automated biomedical evidence synthesis. Collaborates with Brigham and Women's Hospital, Mass General Hospital, and Reboot Rx for clinical translation of research.
Philip S. Yu is a Distinguished Professor in the Department of Computer Science at the University of Illinois at Chicago and holds the Wexler Chair in Information Technology. Previously, he led the Software Tools and Techniques department at IBM Thomas J. Watson Research Center. Education: B.S. in Electrical Engineering, National Taiwan University M.S. and Ph.D. in Electrical Engineering, Stanford University M.B.A., New York University His research spans data mining , big data , social networks , privacy-preserving data publishing , graph/network mining , recommender systems , and deep learning . He has authored over 970 papers with 74,500+ citations and an H-index of 127. Recent work focuses on heterogeneous graph representation, quantum walks in network analysis, and federated unlearning. Scientific Honors: ACM SIGKDD 2016 Innovation Award IEEE Computer Society 2013 Technical Achievement Award IEEE ICDM 2003 Research Contributions Award IEEE Region 1 Award (1999) UIC Research of the Year (2013) IBM Master Inventor with 300+ patents AI 2000 Most Influential Scholar Honorable Mentions (2024-2025) He served as Editor-in-Chief for ACM Transactions on Knowledge Discovery from Data and IEEE Transactions on Knowledge and Data Engineering , and on steering committees for ACM KDD and IEEE Data Mining. His work bridges theoretical advances in graph neural networks , deep learning , and privacy-preserving systems with applications in healthcare, social media, and enterprise analytics.
Hadi Meidani is a Clinical Associate Professor at the Carle Illinois College of Medicine , specifically within the Department of Biomedical and Translational Sciences at the University of Illinois at Urbana-Champaign . He teaches courses in Civil and Environmental Engineering, including topics like Systems Engineering & Economics , Machine Learning in CEE , and Uncertainty Quantification . Ph.D., Civil Engineering, University of Southern California (2012) M.S., Electrical Engineering, University of Southern California (2012) M.S., Structural Engineering, Sharif University of Technology (2005) B.S., Civil Engineering, K.N. Toosi University of Technology (2002) Dr. Meidani's research focuses on uncertainty quantification , scientific machine learning , and optimization under uncertainty for engineering systems. His work spans stochastic multiscale analysis , physics-informed machine learning , and model reduction techniques. His recent publications emphasize machine learning for infrastructure systems , graph neural networks , physics-informed models , and traffic assignment . Key trends include deep learning , multi-fidelity modeling , and neural operator transformers applied to metamaterial design , seismic reliability , and autonomous freight delivery .
Karsten Lambers is Professor of Digital and Computational Archaeology at the Faculty of Archaeology, Leiden University, where he leads research and teaching in the application of computational methods to archaeological data. His work integrates machine learning, remote sensing, text mining, and citizen science to advance archaeological prospection and heritage management. He is affiliated with the Department of Archaeological Sciences and plays key roles in research groups and university-wide initiatives such as SAILS and ARCHON. His research interests span Digital Archaeology , Machine Learning in Archaeology , Remote Sensing , Geoarchaeology , and Human-Environment Interaction . He investigates how computational tools can extract meaningful archaeological information from large datasets, including LiDAR imagery and excavation reports. His fieldwork spans Central Europe and Latin America, with a focus on prehistoric landscapes and cultural heritage. The analysis of his recent publications reveals a strong trend toward automated detection using deep learning (e.g., R-CNN, WODAN), named entity recognition in archaeological texts (e.g., ArcheoBERTje), and citizen science integration for data validation. His work bridges archaeology with computer science, geomatics, and environmental science, emphasizing interdisciplinary collaboration and methodological rigor. His scientific awards include: Best Thesis Award (University of Zurich, 2005) EUROPA NOSTRA Award (2020, 2022) Membership in the German Archaeological Institute (since 2022) Lambers actively supervises students and leads major research projects such as ABMA, EXALT, and Heritage Quest. He has secured substantial research funding and collaborates widely with computer scientists, geophysicists, and palaeoecologists. His teaching includes digital methods, modeling, and simulation, often linked to ongoing research. He has also contributed to open educational resources and digital textbooks in archaeology. He leads or participates in several research labs and teams, including the Digital Archaeology Research Group (which he chairs), the Heritage Quest citizen science project, and interdisciplinary teams focusing on alpine terraces and Iraqi prospection. His work emphasizes the integration of digital tools into practical archaeological workflows, advocating for complementary human-computer strategies.
Professor Christian Bizer is a leading figure in web-based systems and data integration at the University of Mannheim , where he chairs Information Systems V: Web-based Systems . His research focuses on integrating data from multiple sources using large language models and LLM-based agents, with applications in product data extraction and DBpedia knowledge graph construction. He co-founded the DBpedia project and initiated the WebDataCommons initiative. Current research areas: Entity matching, schema matching, table annotation, information extraction, data discovery Key projects: WebMall benchmark, WInte.r integration framework, Schema.org analysis His work applies to e-commerce data integration and knowledge graph construction, with empirical studies on schema.org adoption. He supervises PhD students including Alexander Brinkmann and Ralph Peeters. Scientific Awards: Best Paper at iiWAS 2024 SWSA Ten-Year Award at ISWC 2019 Yahoo FREP Award 2015 Semantic Web Challenge winners Teaching includes courses on web data integration, web mining, large language models, and data mining for master's programs. He leads the DWS PhD colloquium and team projects on LLM agents for data integration.
Fabian Suchanek is a full professor at Institut Polytechnique de Paris, specifically affiliated with Télécom Paris. He leads research in the Data, Intelligence, and Graphs (DIG) team within the Computer Science department. His academic career focuses on bridging artificial intelligence with structured knowledge representations. Suchanek's research interests span artificial intelligence, knowledge bases, and natural language processing, with particular emphasis on knowledge graph construction , rule mining , knowledge-based language models , and explainable AI . His work demonstrates how structured knowledge can enhance machine learning systems, particularly large language models, by providing factual grounding and interpretability. The research group he leads develops practical systems that address real-world knowledge management challenges. His recent publications showcase a strong trajectory in knowledge-intensive AI, with notable contributions to knowledge graph completion, rule mining techniques, and neural approaches to knowledge base validation. The research demonstrates increasing integration between symbolic and neural approaches to AI. Best Student Paper Award at KR 2024 for work on contextual reasoning Best Demo Award of IJCAI 2024 for rule mining in knowledge graphs French Open Research Award for the YAGO project Best Paper Award of ESWC 2021 for Neural Knowledge Base Repairs Suchanek has secured significant research funding, evidenced by his active recruitment of PhD students for knowledge-based language model research. He has held visiting positions, including at Nanyang Technological University (June-September 2023), and is recognized internationally through keynote invitations such as the Singapore ACM SIGKDD Symposium 2023. He has deliberately stepped back from administrative duties at Institut Polytechnique de Paris to focus on research. His laboratory maintains strong industry connections through open-source software projects including the YAGO knowledge base, AMIE for rule mining, STACI for explainable AI, and several other tools that have become standard in knowledge representation research.
Brendan O'Connor is an Associate Professor in the Manning College of Information and Computer Sciences (CICS) at the University of Massachusetts Amherst. His research focuses on computational social science and natural language processing (NLP), particularly exploring how social factors influence language technologies and using text analysis to understand societal trends. His work includes studies on racial bias in NLP, political event analysis, and social media linguistics. He holds a PhD in Machine Learning from Carnegie Mellon University (2014) and dual MS/BS in Symbolic Systems from Stanford University (2006). Education: PhD in Machine Learning, Carnegie Mellon University (2014) MS in Symbolic Systems, Stanford University (2006) BS in Symbolic Systems, Stanford University (2006) Research interests span AI ethics, social media analysis, and computational methods for studying language and society. Notably, he investigates racial disparities in NLP systems, linguistic variation in African American English, and event detection in news and social media. His work has been recognized with NSF CAREER and Google Faculty awards, and his research has been cited thousands of times. His lab, the Statistical Social Language Analysis Lab, develops tools for analyzing large-scale text data. He is affiliated with the Center for Data Science, Center for Intelligent Information Retrieval, and Computational Social Science Institute. Recent projects include analyzing global news coverage of critical events and developing frameworks for zero-shot argument explication. Awards and Honors: NSF CAREER Award Google Faculty Research Award Best Paper Award Advising and Grants: O'Connor has advised projects on social media polling representativeness and demographic analysis. His grants include collaborative research on sociopolitical event extraction and bias mitigation in AI systems. He has also contributed to platforms like Rookie for news archive exploration and ezCoref for coreference resolution. Labs/Teams: Leads the Statistical Social Language Analysis Lab and collaborates with the UMass NLP Group and Harvard Institute for Quantitative Social Science. His work bridges NLP with social science methodologies, emphasizing transparency in algorithms and causal inference using text data.