Arman Cohan is an Assistant Professor in the Department of Computer Science at Yale University, where he leads the Yale NLP Lab since its founding in January 2023. His research spans natural language processing and machine learning with emphasis on language modeling, representation learning, retrieval systems, and specialized domain applications including scientific discovery and AI for science. His primary research interests include: Natural Language Processing Machine Learning Large Language Models Information Retrieval AI for Science Scientific Problem-Solving Recent publications (2025) demonstrate intense focus on evaluating and advancing LLM capabilities across multimodal reasoning, scientific claim verification, financial domain applications, and biological modeling. The lab consistently produces high-impact work accepted at top-tier conferences including ACL, EMNLP, and ICLR, with 11 papers at ACL 2025 alone. Scientific awards include: Best Paper Award at AI4Research Workshop (IJCAI 2024) Outstanding Paper Award at EACL 2023 Best Paper Award at ACL 2024 for Olmo language model research Professor Cohan actively advises PhD students including Kaili Liu, Jacob Dunefsky, Alan Li, Yilun Zhao, and has graduated researchers such as Linyong Nan (now at Zoom) and Ansong Ni (now at Meta). The lab maintains strong industry partnerships and receives substantial research funding as evidenced by its prolific output and conference presence. The Yale NLP Lab hosts the annual New England NLP Workshop and regularly features speakers from leading institutions including Meta AI, Allen Institute for AI, and DeepMind, fostering a collaborative environment for advancing NLP research.
Greg Durrett is an Associate Professor in the Department of Computer Science at University of Texas at Austin, leading the TAUR Lab (Text Analysis, Understanding, and Reasoning ). His research focuses on advancing Large Language Models (LLMs) for knowledge-intensive tasks in medical information processing scientific discovery legal reasoning . He received his B.S. in Computer Science and Mathematics from MIT (2010) and Ph.D. in Computer Science from UC Berkeley (2016). His work develops techniques to train LLMs with new capabilities augment models for reliability assess model outputs improve reasoning frameworks . His 15 most recent publications (2021-2025) span knowledge propagation in LLMs chain-of-thought reasoning code generation benchmarks multi-modal reasoning fact verification discourse analysis . Scientific honors include NSF CAREER Award (2024) NSF grants (2018, 2024) Bloomberg Data Science Grant (2017) Facebook Fellowship (2014) Best Paper Finalist (EMNLP 2013) . Teaching: CS388: Natural Language Processing (graduate) CS371N: NLP (undergraduate) High school NLP module .
Fabian Suchanek is a full professor at Institut Polytechnique de Paris, specifically affiliated with Télécom Paris. He leads research in the Data, Intelligence, and Graphs (DIG) team within the Computer Science department. His academic career focuses on bridging artificial intelligence with structured knowledge representations. Suchanek's research interests span artificial intelligence, knowledge bases, and natural language processing, with particular emphasis on knowledge graph construction , rule mining , knowledge-based language models , and explainable AI . His work demonstrates how structured knowledge can enhance machine learning systems, particularly large language models, by providing factual grounding and interpretability. The research group he leads develops practical systems that address real-world knowledge management challenges. His recent publications showcase a strong trajectory in knowledge-intensive AI, with notable contributions to knowledge graph completion, rule mining techniques, and neural approaches to knowledge base validation. The research demonstrates increasing integration between symbolic and neural approaches to AI. Best Student Paper Award at KR 2024 for work on contextual reasoning Best Demo Award of IJCAI 2024 for rule mining in knowledge graphs French Open Research Award for the YAGO project Best Paper Award of ESWC 2021 for Neural Knowledge Base Repairs Suchanek has secured significant research funding, evidenced by his active recruitment of PhD students for knowledge-based language model research. He has held visiting positions, including at Nanyang Technological University (June-September 2023), and is recognized internationally through keynote invitations such as the Singapore ACM SIGKDD Symposium 2023. He has deliberately stepped back from administrative duties at Institut Polytechnique de Paris to focus on research. His laboratory maintains strong industry connections through open-source software projects including the YAGO knowledge base, AMIE for rule mining, STACI for explainable AI, and several other tools that have become standard in knowledge representation research.
Zhonghai Lu is a Professor of Electronic Systems Design (specializing in Dependable and Autonomous Systems) at KTH Royal Institute of Technology, part of the Department of Electrical Engineering in the School of Electrical Engineering and Computer Science (EECS). He serves as Program Director for KTH's Embedded Systems master's program and Director of Studies at the Division of Electronics and Embedded Systems. His research focuses on Network-on-Chip (NoC), computer architecture, embedded systems, and Prognostics and Health Management (PHM) of power electronics. He leads a research group exploring in-network processing and embedded intelligence, transforming passive networks into active computational frameworks. Lu holds a BSc from Beijing Normal University (1989), MSc and PhD from KTH (2002, 2007), and an MBA in Innovation and Growth from the University of Turku (2012). He has authored over 240 scientific papers, including journal articles and peer-reviewed conferences, with notable recognitions such as Best Paper Awards at NOCS’2015 and EU HiPEAC, and a Featured Paper in IEEE Transactions on Computers (2020). He serves as Associate Editor for ACM Transactions on Architecture and Code Optimization (TACO) and has chaired major conferences like HiPEAC’2017 and NOCS’2018. His research group’s recent work includes integrating AI into hardware acceleration, fault-tolerant neural networks, and RUL estimation for power electronics using recurrent neural networks. Lu has secured grants from the Swedish Research Council and Intel Corporation and developed courses like IL2230 (Hardware Architectures for Deep Learning) and IL2233 (Embedded Intelligence), pioneering embedded AI education at KTH. Education: BSc (Beijing Normal University), MSc/PhD (KTH), MBA (University of Turku) Awards: Best Paper Awards (NOCS, EU HiPEAC), Swedish Research Council Grants, Intel Research Gifts Labs/Teams: Research Group on In-Network Processing and Embedded Intelligence
Thomas Hacker is a Professor in the Department of Computer and Information Technology at Purdue Polytechnic Institute, Purdue University. His research focuses on cloud computing, high-performance computing, operating systems, computer networking, and cyber infrastructure . He holds a Ph.D. and M.S. in Computer Science & Engineering from the University of Michigan, along with dual B.S. degrees in Computer Science and Physics from Oakland University. Education: PhD (Computer Science & Engineering), University of Michigan (2004) MS (Computer Science & Engineering), University of Michigan (1993) BS (Computer Science, Mathematics Minor), Oakland University (1989) BS (Physics), Oakland University (1989) Dr. Hacker's research spans cloud and grid computing, operating systems, and distributed systems , with applications in earthquake engineering data systems and AI-driven infrastructure analysis. His recent work explores extended layer 2 networking for bare-metal provisioning ( 2023 IEEE Cloud Summit ) and machine-supported bridge inspection using artificial intelligence ( Transportation Research Record, 2023 ). Notable scientific contributions include 15+ publications on topics like cyberinfrastructure for earthquake engineering, container-based virtualization, and data-intensive systems. His work has been recognized with awards such as the NSF CAREER Award (2010) and multiple Purdue Seed for Success Awards . Key Scientific Awards: NSF CAREER Award (2010) Purdue Seed for Success Awards (2008-2013) ASEE Information Systems Division Best Paper Award (2012) College of Technology Outstanding Faculty in Discovery Award (2010) He has held leadership roles at Purdue, including Department Head (2018-2021) and Interim Department Head (2011-2016) . His career spans academic positions at Indiana University, University of Michigan, and industry roles at Storage Technology Corporation.
Ann B. Lee is a Professor and Co-Director of the PhD Program in Statistics at Carnegie Mellon University , with a joint appointment in the Department of Statistics & Data Science and the Machine Learning Department. Prior to joining CMU, she held positions as a J.W. Gibbs Assistant Professor at Yale University and a visiting research associate at Brown University. PhD in Physics, Brown University MSc/BSc in Engineering Physics, Chalmers University of Technology, Sweden Her research focuses on statistical methodology for complex data in the physical sciences , emphasizing trustworthy inference, uncertainty quantification, and integration of classical statistics with machine learning. Recent work includes likelihood-free inference, calibrated forecasting, and diagnostics for generative models. The STAMPS research group , which she co-founded in 2018, hosts weekly meetings and public webinars. In Fall 2024, STAMPS will transition into a CMU Research Center. Recent publications span likelihood-free inference , climate modeling , and astronomy . Notable collaborations include applications to hurricane intensity guidance , galaxy redshift estimation , and cosmological parameter biases . She mentors PhD students and has advised multiple award-winning researchers, including ASA Best Student Paper Award winners. Her teaching includes advanced courses on probability, regression, and AI for climate sciences.
Prof. Dr. Andrea Stocco is a Professor at the Technische Universität München (TUM), affiliated with the TUM School of Computation, Information and Technology. His research focuses on the intersection of software engineering and deep learning, particularly addressing the robustness and reliability of data-intensive systems. Key areas include autonomous vehicles, web application testing, and automated functional oracles for deep learning systems. He leads initiatives such as the Lehrstuhl für Software und Systems Engineering , collaborating on projects like CrESt and SUPPRA – Algorand Center of Excellence . Research interests encompass monitoring techniques for AI-driven systems, test suite maintainability, and scenario-based testing of cyber-physical systems (CPS). His work emphasizes practical applications, such as improving testing frameworks for evolving web applications and enhancing interoperability in autonomous driving systems (ADS). Recent efforts include leveraging large language models (LLMs) for secure code assessment and benchmarking generative AI for test input generation. No scientific awards are explicitly mentioned in the provided texts. His publications reflect a strong focus on testing methodologies, with over 40 articles since 2013, covering domains like web test automation, dependency-aware testing, and safety-critical failure prediction in autonomous systems. Advising and grants details are not detailed in the current data, but his lab contributes to TUM's broader efforts in software engineering and systems reliability.
Andrew Rice is a Professor of Computer Science at the University of Cambridge's Department of Computer Science and Technology, and holds the Hassabis Fellowship in Computer Science. He is also the Director of Studies in Computer Science at Queens' College. His research focuses on programming languages, software engineering, and machine learning applications in software development. He leads projects like Isaac Computer Science and ALTA (Automated Language Teaching and Assessment), advancing adaptive learning technologies. His work includes static analysis tools such as Error Prone at Google, energy efficiency studies in computing infrastructure, and contributions to the Computing for the Future of the Planet initiative. His teaching emphasizes practical skill development through flipped classrooms and video lectures, earning him the 2014 Pilkington Prize for teaching excellence. He has held visiting roles at Google and collaborated on energy consumption research for mobile devices and data centers. His research spans systems, networking, and natural language processing, with a strong focus on applying computational methods to real-world challenges. Key Projects: Isaac Physics/Computer Science, ALTA, Error Prone Static Analysis Research Themes: Programming Languages, Machine Learning, Energy Efficiency Awards: Pilkington Prize (2014)
Elena Simperl is a Professor of Computer Science and Deputy Head of Department for Enterprise and Engagement at King's College London's Department of Informatics. She co-directs the King's Institute for Artificial Intelligence and serves as Director of Research for the Open Data Institute. As a Hans Fischer Senior Fellow at the Technical University of Munich's Institute for Advanced Study, she leads the Trustworthy Knowledge Graphs focus group and contributes to advancing human-centric AI research across European institutions. Professor Simperl obtained her doctoral degree in Computer Science from the Free University of Berlin and her diploma from the Technical University of Munich. Prior to joining King's, she held academic positions in Germany, Austria, and at the University of Southampton, and was a Turing Fellow. Her career trajectory demonstrates consistent leadership in bridging academic research with practical applications in data ecosystems. Her research sits at the critical intersection of AI and social computing, focusing on human-centric approaches to building sociotechnical systems that integrate data, algorithms, and human capabilities. She investigates how to make knowledge engineering more accessible, how to leverage collective intelligence for data quality improvement, and how to design participatory AI systems that address societal challenges like misinformation. Her work spans knowledge graphs, semantic technologies, crowdsourcing, and open data, with particular emphasis on the social dimensions of data-intensive systems and the governance frameworks needed for trustworthy AI deployment. Analysis of her recent publications reveals a strong evolution toward integrating large language models with traditional knowledge engineering practices while maintaining human oversight. There's a clear trajectory from foundational work on knowledge representation toward increasingly applied research addressing real-world challenges in media ecosystems, citizen science, and data governance, with growing attention to policy implications of AI technologies. Fellow of the British Computer Society Fellow of the Royal Society of Arts Hans Fischer Senior Fellow at TUM-IAS (2023) Ranked among top 100 most influential scholars in knowledge engineering of the last decade Included in Women in AI 2000 ranking Professor Simperl has led 14 major European and national research projects totaling millions in funding, including MediaFutures (a Horizon 2020 program tackling online misinformation), QROWD, ODINE, Data Pitch, and ACTION. She currently co-chairs the Croissant working group in ML Commons developing data standards for AI, and serves as president of the Semantic Web Science Association. Her research has directly influenced the development of data ecosystems supporting startups and citizen science initiatives across Europe, demonstrating exceptional ability to translate theoretical advances into practical impact. As Director of Research at the Open Data Institute, she oversees initiatives connecting data entrepreneurs with artists and civic organizations. Her leadership in the MediaFutures project established a data-driven innovation hub that supported 51 startups/SMEs and 43 artists through three open calls, creating a sustainable model for arts-technology collaborations addressing media challenges. Her work with the ODINE project helped create a European ecosystem for data-driven startups, demonstrating her commitment to building practical applications of open data principles.
Prof. Alexander Pretschner is a Professor of Software & Systems Engineering at the Technical University of Munich (TUM) and Founding Director of the Bavarian Research Institute for Digital Transformation (bidt). He also serves as Scientific Director of fortiss, a Bavarian research institute for software-intensive systems. His research focuses on software engineering, testing, information security, and ethical software development. Pretschner holds a PhD from TUM and has held academic positions at Karlsruhe Institute of Technology (KIT) and TU Kaiserslautern. He is a co-editor of several prestigious journals, including IEEE Transactions on Reliability and the Journal of Software Testing, Verification and Reliability. Education: PhD in Computer Science, Technical University of Munich MSc in Computer Science, University of Kansas (on Fulbright Scholarship) Diplom in Computer Science, RWTH Aachen University Research Interests: His work spans testing methodologies, secure software design, and ethical considerations in agile development. Notable contributions include frameworks for metamorphic testing, distributed data usage control, and accountability mechanisms for cyber-physical systems. Awards: IBM Faculty Award (2012, 2013) Google Focused Research Award (2011, 2012) EARTO Innovation Prize (2014) 2nd Platz Supervisory Award (2020) Advising & Grants: Pretschner has supervised numerous PhD and Master’s students, contributing to over 200 publications. He leads projects like EDAP (Ethical Deliberation in Agile Processes) and collaborates with industry partners on cybersecurity and AI ethics initiatives. Labs & Teams: His work is anchored in bidt, fortiss, and TUM’s Chair of Software & Systems Engineering, focusing on societal impacts of digitalization and trustworthy AI systems.
Qizhen Zhang is a Professor in the Department of Computer Science at the University of Toronto's Faculty of Arts and Science. Specializing in hyperscale data processing systems, Zhang leads research at the intersection of cloud computing, distributed systems, and data center networking. Recent work focuses on network-centric designs for efficient large-scale data processing, including pioneering contributions to disaggregated data center architectures. Research interests center on hyperscale data processing , network-aware system design , and disaggregated infrastructure . Current projects investigate DPU-accelerated systems (dpBento, DPDPU), memory disaggregation (Cowbird, Redy), and blockchain scalability (FlexChain). Zhang's approach systematically integrates network characteristics into distributed system optimization, addressing challenges in trillion-item workloads through novel shuffle layers (TeShu) and compute pushdown mechanisms (TELEPORT). Zhang advises multiple graduate students working on satellite networking (SaTE), federated learning, and DPU-optimized storage (DDS). Professional service includes program committees for SIGMOD, VLDB, NSDI, and EuroSys (2023-2026), plus journal reviews for ACM TODS and IEEE/ACM Transactions on Networking. Industrial collaborations with Microsoft Research have yielded production-oriented systems like Redy and CompuCache. Key contributions include MimicNet for scalable network simulation (SIGCOMM 2021), GraphRex for network-aware graph processing (SIGMOD 2019), and foundational work on disaggregated data centers (CIDR 2020, VLDB 2020). Teaching responsibilities encompass graduate courses CSC2235 (Cloud-native Data Management) and undergraduate CSCC43 (Databases) at the University of Toronto.
Yongjoo Park is an Assistant Professor in the Department of Computer Science at the University of Illinois at Urbana-Champaign (UIUC), affiliated with the Grainger College of Engineering. He leads research in data-intensive AI systems as a member of the Data and Information Systems (DAIS) lab, focusing on novel data systems that bridge database theory and practical AI applications. His work emphasizes open-source contributions through GitHub and direct societal impact. Research interests center on systems for data-intensive AI , particularly efficient Retrieval-Augmented Generation (RAG) systems for exploratory AI, data science versioning, and in-storage computing. Key projects include Kishu (the world's first undoable Jupyter notebook with time-travel capabilities), CARE (a causal-relational system for structured/unstructured data), and AirDB/AirIndex (serverless transactions and automatic index optimization). His group develops tools enabling scalable, optimized AI workflows from storage layers to LLM inference. Recent publications reveal a strong focus on interactive data systems (85% of recent work), with significant contributions to notebook environments (Kishu), vector databases (ISCA'25), and RAG optimization. Awards highlight technical innovation, including SIGMOD 2025 Best Demo Award and NSF CAREER funding. His open-source philosophy drives GitHub releases of all major systems. SIGMOD 2025 Best Demo Award (Kishu) NSF CAREER Award (Novel data science systems) SIGMOD'23 Best Artifact Award Honorable Mention (DeepOLA) IBM-Illinois Project Selection (VectorDB/RAG) Mentorship spans 12 current PhD/MS students and 6 graduated advisees, including Supawit Chockchowwat (now Postdoc at Google, future Assistant Professor at CMKL University). He teaches advanced courses like CS511 (Advanced Data Management) and recruits 1-2 new PhD students annually, prioritizing data systems research. His lab emphasizes diversity, individual respect, and concrete outcomes in a collaborative workspace.
Ping Yang is a Professor and Associate Director for Research and Graduate Programs in the School of Computing at Binghamton University (SUNY). She holds a Ph.D. in Computer Science from Stony Brook University, an ME from the Chinese Academy of Sciences, and a BS from Zhongshan University. Her research focuses on cybersecurity, AI-based security, virtual machine security, privacy policy analysis, and formal methods. She directs the Center for Information Assurance and Cybersecurity and coordinates cybersecurity programs at both undergraduate and graduate levels. Education: BS in Computer Science, Zhongshan University ME in Computer Science, Chinese Academy of Sciences MS and PhD in Computer Science, State University of New York at Stony Brook Research Interests: Dr. Yang's work spans information and systems security, security in virtualized computing, access control mechanisms, privacy policies, and formal methods for security verification. Her projects include blockchain-based provenance storage, real-time anomaly detection in workflows, and privacy-preserving virtual machine migration. She has led NSF-funded initiatives on security in cloud environments and scientific workflows. Awards: Not explicitly listed in the provided materials. Advising & Grants: Advised over 30 PhD/Master’s students and contributed to grants including NSF Scholarship for Service and GenCyber programs. Her team develops tools like RBAC-PAT for access control analysis. Labs/Teams: Leads the Center for Information Assurance and Cybersecurity and collaborates on projects involving secure data workflows and blockchain applications in scientific research.
Kenneth Ross is a Professor in the Computer Science Department at Columbia University in New York City. His primary appointment is within the Department of Computer Science, with affiliations including the Foundations of Data Science Committee. His work bridges theoretical database research and practical system implementation. His research focuses on database systems with particular expertise in query processing, query language design, data warehousing, and architecture-sensitive database system design. Additional research spans computational biology, especially analysis of large genomic data sets. Current projects include Linear Algebra Operators in Databases for machine learning workloads and Repeats and Somatic Mutation analysis in genomics. His work consistently addresses the intersection of hardware capabilities and database system design. Ross leads the Database Research Lab at Columbia, which has produced significant work on query optimization, GPU database processing, and hardware-conscious database systems. His recent publications demonstrate strong focus on adapting database systems to modern hardware including GPUs, SIMD processors, and persistent memory. His scientific recognition includes: Packard Foundation Fellowship Sloan Foundation Fellowship NSF Young Investigator Award Distinguished Faculty Teaching Award (2008) Ross actively advises undergraduate engineering students (juniors with last names P-Z) and has taught foundational courses including Introduction to Databases and Programming and Problem Solving for over two decades. His teaching portfolio shows consistent engagement with both theoretical concepts and practical implementation challenges in computer science education.
Yi Ding is an Assistant Professor in the Elmore Family School of Electrical and Computer Engineering at Purdue University, where they lead the STYLE (Sustainable computing Systems and LEarning) Lab. Dr. Ding joined Purdue in August 2023 after completing a postdoctoral fellowship at MIT CSAIL as an NSF Computing Innovation Fellow, mentored by Michael Carbin. During their postdoc, they also held a visiting position at Meta Infra Data Center to improve server maintenance efficiency in hyperscale datacenters. They received their Ph.D. in Computer Science from the University of Chicago, advised by Henry Hoffmann. Dr. Ding's research focuses on computer systems, computer architecture, and AI/ML, with strong emphasis on applications in sustainability and healthcare. Their work spans sustainable computing, including energy efficiency in datacenters and LLM serving, as well as healthcare applications such as mental health prediction and EEG analysis. Their recent publications demonstrate a strong trend toward addressing environmental impacts of computing, particularly in datacenters and AI systems, while also exploring innovative healthcare applications. Their research bridges systems, sustainability, and health domains, creating novel solutions for pressing societal challenges. Dr. Ding has received notable recognition including the Seed Funding for High-Impact Review Papers (2024), the Meta Research Award (2021), and was selected as a Computing Innovation Fellow by CRA/CCC (2020). Seed Funding for High-Impact Review Papers (2024) with Inez Hua 1st Place in Research Talk in CoE at Fall 2024 Undergrad Research Expo (awarded to Gavin Fortwendel) 2020 Computing Innovation Fellow by CRA/CCC Meta Research Award on Statistics for Improving Insights, Models, and Decisions (2021) Dr. Ding is actively recruiting self-motivated Ph.D. students interested in AI/ML systems research. They have secured funding for undergraduate research projects through DUIRI, focusing on sustainable AI computing and energy use in training autonomous vehicles. Their lab collaborates with various institutions including MIT, Meta, and interdisciplinary partners at Purdue. The STYLE Lab under Dr. Ding's leadership is actively engaged in multiple research initiatives addressing sustainable computing and healthcare applications, with strong industry connections and funding support from both internal university sources and external partners.