Professor Ihab Ilyas is a leading figure in data management and machine learning at the Cheriton School of Computer Science , University of Waterloo , where he holds the Thomson Reuters Research Chair in Data Quality . He is currently on leave from the university, serving as a Distinguished Engineer, Proactive Intelligence at Apple Inc . His research focuses on data cleaning , large-scale data integration , knowledge graphs , and machine learning applications in data quality . He has pioneered systems like HoloClean and Saga , with commercial impacts through co-founded startups Inductiv (acquired by Apple) and Tamr . PhD in Computer Science from Purdue University Co-founder of Inductiv (acquired by Apple) Co-founder of Tamr Research Interests include: Probabilistic and uncertain data management Machine learning for data quality and enrichment Big data systems and information extraction Knowledge graph construction and optimization Scientific Awards and Recognitions include: Fellow of the Royal Society of Canada (2024) C.C. Gotlieb Computer Award (2024) IEEE Fellow (2021) ACM Fellow (2020) Cheriton Faculty Fellowship (2013-2016) Ontario Early Researcher Award
Maya Ramanath is an Associate Professor in the Department of Computer Science and Engineering at Indian Institute of Technology (IIT) Delhi. She joined IIT Delhi in 2011 after a postdoctoral research stint at the Max-Planck Institute for Informatics in Germany. Her research interests focus on database systems, information retrieval, semantic web technologies, and knowledge graph construction and applications. Education: PhD in Computer Science, Indian Institute of Science, Bangalore M.Sc.(Engg.) in Computer Science, Indian Institute of Science, Bangalore B.E. in Computer Science and Engineering, Bangalore University, Bangalore Her recent work emphasizes efficient query processing over large-scale graphs, knowledge graph applications, and natural language interfaces for semantic data. Notable contributions include algorithms for reachability approximation in web-scale graphs, speculative query planning for knowledge graphs, and exploratory querying techniques. She has collaborated extensively on projects like NAGA, ESTHETE, and KlusTree, advancing the state of the art in graph-based data management and semantic search. Publications span conferences such as ICDE, ECIR, EDBT, and VLDB, reflecting a strong focus on database systems, graph algorithms, and semantic web applications. Her work bridges theoretical foundations with practical implementations, addressing scalability and efficiency challenges in modern data management systems. Research and advising activities include supervision of projects on distributed graph processing, query optimization, and knowledge representation. She has contributed to open-source tools like LegoDB and StatiX, and her lab focuses on interdisciplinary approaches to data-centric AI.
Asaf Cidon is an Associate Professor at Columbia University, jointly affiliated with the Department of Electrical Engineering and Computer Science, and a member of the Data Science Institute. His research focuses on software systems , storage , large-scale machine learning , and cybersecurity . Stanford University - PhD in Electrical Engineering Stanford University - MS in Electrical Engineering Technion - BS in Computer and Software Engineering His work in distributed storage systems has been commercialized by companies such as Facebook, Tibco, and Rubrik. He has led projects like Sentinel and Forensics during his industry career. Recent publications highlight advancements in software-based radiation protection (ASPLOS'26), PCIe pooling with CXL (HotOS'25), and AI phishing detection (IMC'25). These reflect his expertise in system architecture , security , and networking . Scientific recognitions include best paper awards at OSDI, Usenix Security, CIDR, and ATC, along with NSF CAREER and ARO Young Investigator Awards . His papers Cookie Monster (SOSP'24) and Chablis (CIDR'24) received notable accolades. He has mentored numerous PhD and Master's students , including Edward Guo, Harry Wang, and Teng Jiang, many of whom now hold roles at Google, Meta, Amazon, and academic institutions. His lab at Columbia is actively recruiting CS and EE PhD students. Prior to Columbia, he founded and led the startup Sookasa to acquisition and served as Senior Vice President of Email Protection at Barracuda Networks , managing a $200M business with 100 engineers.
Panos Ipeirotis is a Professor at the Leonard N. Stern School of Business at New York University, affiliated with the Department of Technology, Operations, and Statistics. He also serves as the George A. Kellner Faculty Fellow and is associated with the Center for Data Science and Computer Science departments at NYU. PhD in Computer Science (Columbia University, 2004) MSc in Computer Science (Columbia University, 2001) BSc in Computer Engineering & Informatics (University of Patras, 1999) His research spans crowdsourcing, machine learning, human-AI collaboration, online labor markets, and social media analytics. He pioneered human-machine loop systems that combine human and machine intelligence to achieve superior outcomes. His work has applications in data quality assurance, visual media search (e.g., Google Project Glass), and economic valuation of user-generated content. Recent publications focus on algorithmic fairness in hiring systems, occupational segregation analysis, and theoretical advancements in crowdsourcing consensus mechanisms. Earlier work includes foundational studies on data quality in crowdsourcing platforms, economic impacts of product reviews, and query optimization for text-centric tasks. 2015 Lagrange Prize in Complex Systems NSF CAREER Award SIGKDD Test of Time Award (2020) Multiple Best Paper awards (WWW 2011, KDD 2008, SIGMOD 2006) He has received significant grants, including a $1.5 million Google Research Grant (2013) for integrating crowdsourcing with machine learning algorithms. His work bridges computer science, economics, and social psychology, with implications for policy-making and business strategy.
Beng Chin Ooi is a Lee Kong Chian Centennial Professor at the National University of Singapore (NUS), School of Computing (SoC). He holds a B.Sc. (1st Class Honors) and Ph.D. in Computer Science from Monash University. His research focuses on database systems, large-scale analytics, and distributed computing. He has held leadership roles including Dean of School of Computing (2007–2013) and Director of Smart Systems Institute (2011–2021). Key achievements include the Singapore President’s Science Award (2011), ACM Fellow (2011), IEEE Fellow (2009), and multiple best paper awards. His work emphasizes scalable data management, blockchain systems, and healthcare data analytics. Education: Monash University (B.Sc., Ph.D.) Leadership: Dean (SoC), Director (Smart Systems Institute) Awards: Over 15 major honors including ACM SIGMOD E.F. Codd Innovations Award (2020) His research spans distributed databases, big data systems, and innovative applications of blockchain technology. Recent work includes NASI (neural architecture search) and Rafiki (ML-as-a-service).
Dr. Khurram Aziz is a Senior Instructor in the Faculty of Computer Science at Dalhousie University , Halifax, Canada. He is actively engaged in teaching and research, with a focus on optical networks, data center interconnects, and network performance modeling. Education: PhD in Electrical Engineering, Vienna University of Technology, Austria (2008) MSc in Electrical Engineering, National University of Singapore (2003) BSc (Hons) in Electrical Engineering, University of Engineering and Technology, Lahore, Pakistan (1998) His research interests include optical packet and burst switched networks , optical interconnects for data centers , analytical modeling and simulation , and network routing and switching . He has contributed extensively to the design and performance evaluation of scalable optical switches and hybrid switching systems. The recent publications reflect a strong trend in data center optical networks , focusing on performance, blocking probability, signal degradation, and architectural classification. His work bridges theoretical modeling with practical simulation frameworks, such as CloudNetSim++ in OMNeT++, contributing to cloud and high-capacity network research. Dr. Aziz has no listed scientific awards in the provided text. He teaches several core computer science courses including CSCI 2141: Intro to Database Systems , CSCI 3171: Network Computing , CSCI 3132: Object Orientation and Generic Programming , and CSCI 3120: Operating Systems . There is no mention of graduate student supervision or external research grants. He has co-authored book chapters in major handbooks on data centers and switched systems. Dr. Aziz has not listed any formal lab or research team affiliations in the provided content.
Khuzaima Daudjee is a Professor and David R. Cheriton Faculty Fellow in the Cheriton School of Computer Science at the University of Waterloo. His research focuses on systems-oriented problems at the intersection of systems and data management, particularly building large-scale systems, storage infrastructure in the cloud, and modern hardware applications. He leads projects in distributed database systems, elastic scaling, and resource optimization. His recent work includes Caerus (geo-replicated transactions), Tiresias (predictive storage), and MorphoSys (automatic physical design metamorphosis). Daudjee has chaired major conferences including ICDE 2026 and serves on editorial boards for VLDB, SIGMOD, and IEEE TKDE journals. His awards include ACM Distinguished Scientist and multiple best paper awards. Educational initiatives include developing distributed systems teaching materials and supervising graduate students across database and distributed systems domains. His industry collaborations involve cloud infrastructure optimization and scalable data processing frameworks.
Andy Pavlo is an Associate Professor with Indefinite Tenure in the Computer Science Department at Carnegie Mellon University's School of Computer Science. He is an active member of the CMU Database Group and the Parallel Data Laboratory, where he leads research in database management systems with a focus on self-driving architectures, transaction processing, and large-scale analytics. His work bridges academic research and industry applications through projects like NoisePage, OtterTune (which he co-founded and served as CEO before it ceased operations), and Peloton. Dr. Pavlo's research interests span database management systems with particular emphasis on autonomous database architectures that can self-tune and optimize without human intervention. His work explores transaction processing systems that can handle high-throughput workloads while maintaining consistency, and large-scale data analytics techniques that efficiently process massive datasets. He has made significant contributions to query optimization, database extensibility, and automatic database tuning using machine learning techniques. His recent work on database extensibility revealed critical issues in PostgreSQL's extension ecosystem, showing that approximately 16% of extensions are incompatible with at least one other extension due to API violations and memory errors. His research output demonstrates a consistent focus on practical database systems challenges, with recent publications examining database extensibility, user-defined function optimization, and the cyclical nature of database research. The articles show a strong trend toward making database systems more autonomous, with increasing integration of machine learning techniques for automatic tuning and optimization. His work often combines deep theoretical analysis with practical implementation in open-source systems. Dijkstra Award 2024 for contributions to database systems research Dr. Pavlo actively mentors graduate students, with current advisees including Wan Shen Lim, William Zhang, and Sam Arch (co-advised with Todd Mowry). His former students have gone on to successful careers in both industry and academia. He has secured significant research funding through CMU's affiliate program with major database companies including ClickHouse, DataStax, dbt, Firebolt, MotherDuck, RelationalAI, SingleStore, Spiral, PingCAP/TiDB, Yellowbrick, and Yugabyte. His research is supported by these industry partnerships and likely includes NSF funding given his active participation in the database research community. At CMU, Dr. Pavlo leads the Database Group and organizes several seminar series including "SQL or Death," "Database Building Blocks," and "ML⇄DB Technical Talks." These seminars bring together researchers and practitioners to discuss cutting-edge developments in database systems. He also runs a summer research internship program that has attracted students for multiple consecutive years, indicating a strong research group with ongoing projects and funding.
Daniel Abadi is the Darnell-Kanal Professor of Computer Science at the University of Maryland, College Park, with an appointment in the University of Maryland Institute for Advanced Computer Studies. He leads the Data Systems Lab at Maryland (DSLAM) and has made significant contributions to database system architecture and implementation, particularly in scalable and distributed systems. Prof. Abadi's research focuses on database system architecture, especially at the intersection with scalable and distributed systems. He is best-known for the development of the storage and query execution engines of the C-Store (column-oriented database) prototype, which was commercialized by Vertica and eventually acquired by Hewlett-Packard, and for his HadoopDB research on fault tolerant scalable analytical database systems which was commercialized by Hadapt and acquired by Teradata. His current work includes deterministic distributed systems like Calvin and SLOG, which provide strictly serializable, low-latency, geographically replicated database transactions. Analysis of his recent publications reveals a strong focus on modern database challenges including transaction processing, distributed systems architecture, data mesh concepts, schema evolution, and IoT data management. His work consistently addresses the tension between consistency, availability, and performance in distributed database systems, with recent emphasis on moving beyond traditional two-phase commit protocols and exploring novel approaches to data architecture like data mesh. ACM Fellow Sloan Research Fellowship Churchill Scholarship NSF CAREER Award VLDB Best Paper Award Two VLDB Test of Time Awards (for C-Store and HadoopDB) 2008 SIGMOD Jim Gray Doctoral Dissertation Award 2013-2014 Yale Provost's Teaching Prize 2013 VLDB Early Career Researcher Award PhD dissertation advisor for Alexander Thomson and Jose Falerio, whose dissertations won SIGMOD Jim Gray Doctoral Dissertation Awards in 2015 and 2020 respectively Prof. Abadi has advised multiple PhD students, including Alexander Thomson and Jose Falerio, both of whom received the prestigious SIGMOD Jim Gray Doctoral Dissertation Award. His research has been supported by numerous grants including NSF funding for projects like SLOG. He maintains an active presence in the database community through his widely-read blog DBMS Musings and through service on program committees for major conferences including SIGMOD, VLDB, and CIDR. He leads the DSLAM research group at the University of Maryland, which focuses on cutting-edge database system research with strong industry connections and practical impact.
State University of New York at BuffaloUnited States
Zhuoyue Zhao is an Assistant Professor in the Department of Computer Science and Engineering at the University at Buffalo, School of Engineering and Applied Sciences. His office is located at 338I Davis Hall, Buffalo, NY 14260, and he can be reached at zzhao35@buffalo.edu or by phone at (716) 645-4735. Dr. Zhao received his PhD in Computer Science from the University of Utah in 2021, where he was advised by Prof. Feifei Li and Prof. Jeff Phillips. Prior to that, he earned his BS in Computer Science from Shanghai Jiao Tong University in 2016, where he was part of the prestigious ACM Class. During his undergraduate studies, he conducted research under Prof. Kenny Zhu and spent Fall 2015 as a research assistant at Hong Kong Polytechnic University supervised by Prof. Eric Lo. Dr. Zhao's research focuses on database management systems, with specific emphasis on traditional and approximate query processing, query optimization, database systems on modern hardware, transaction processing, indexing, and storage. His work bridges theoretical foundations with practical implementations, often resulting in systems that address real-world database challenges. He has made significant contributions to probabilistic query processing, transaction scheduling, and learned indexing techniques. His recent publications demonstrate a clear trajectory toward optimizing database performance in hybrid transactional/analytical processing environments. His research increasingly integrates systems techniques with machine learning approaches, particularly in the area of learned indexes. There's also a strong focus on making database operations more efficient through innovative scheduling mechanisms and query processing techniques that can handle concurrent updates. Google PhD Fellowship (2019-2021) Best Paper Award at SIGMOD 2016 for "Wander Join: Online Aggregation via Random Walks" Best Paper Award at SIGMOD 2025 for "Low-Latency Transaction Scheduling via Userspace Interrupts" Dr. Zhao currently advises several PhD students including Yunnan Yu, Congying Wang, Gaoxiang Liu (co-advised with Prof. Ziming Zhao), and Zhuoran Li. He has successfully guided MS student Nithin Sastry Tellapuri to graduation (Fall 2023), who is now employed at AirPay. His research is supported by significant funding including an NSF CAREER award (#2339596) totaling $599,977 for research on "Speedy and Reliable Approximate Queries in Hybrid Transactional/Analytical Systems" (2024-2029) and an unrestricted Google gift of $30,000 (2021). Dr. Zhao leads the ADBLab research group at UB, where students work on cutting-edge database systems research. His lab focuses on building practical database systems that address real-world challenges in query processing, transaction management, and indexing. The lab maintains strong connections with industry partners and regularly contributes to open-source database projects.
Aditya Parameswaran is an Associate Professor in the Electrical Engineering and Computer Sciences (EECS) department at the University of California, Berkeley. He co-directs the EPIC Data Lab and the Police Records Access project, focusing on simplifying data science at scale through human-in-the-loop systems, LLM-powered tools, and scalable data systems. His research spans database systems, human-computer interaction, and machine learning, with notable contributions in tools like Lux, Modin, and DataSpread. Education : PhD in Computer Science from Stanford University (2013) BTech in Computer Science and Engineering from IIT Bombay (2007) Research Interests : Parameswaran's work centers on empowering end-users with intuitive data tools. Recent projects include LLM-powered systems for document processing (DocETL, TWIX), proactive data systems, and benchmarking frameworks. He emphasizes democratizing data science through low/no-code solutions and improving production ML workflows. Articles Trends : His recent work (2023–2025) prioritizes LLM integration into data systems, focusing on robust pipelines, assertion generation (SPADE), and debugging tools (RAGGY). Earlier contributions include visualization recommendation (Lux), scalable dataframes (Modin), and spreadsheet optimization (DataSpread). Awards : Recipient of the VLDB Early Career Award (2019), Sloan Research Fellowship (2020), NSF CAREER Award (2017), and multiple best paper/demonstration awards at top venues like SIGMOD and VLDB. Advising & Grants : Guides over 20 PhD/postdoc alumni, many now in academia (e.g., Madelon Hulsebos at CWI) and industry leadership roles. Active in securing grants (e.g., NSF, Army Research Office) and industry partnerships (e.g., Snowflake, LangChain). Labs/Teams : Leads the EPIC Data Lab, focusing on agentic data systems, and co-founded Ponder (acquired by Snowflake). Collaborates on the Police Records Access initiative, building transparency tools for public records.
Daniel Abadi is the Darnell-Kanal Professor of Computer Science at the University of Maryland, College Park. He leads the Data Systems Lab at Maryland (DSLAM) and is widely recognized for his groundbreaking contributions to database system architecture and implementation. Previously, he was a faculty member at Yale University where he received the Provost's Teaching Prize. Abadi's research primarily focuses on database system architecture, particularly at the intersection with scalable and distributed systems. He is best known for developing the storage and query execution engines of the C-Store prototype (a column-oriented database system commercialized by Vertica and later acquired by Hewlett-Packard), HadoopDB research (commercialized by Hadapt and acquired by Teradata), and deterministic distributed transactional systems like Calvin (currently being commercialized by Fauna). His work bridges theoretical innovation with practical industrial impact. Analysis of his recent publications reveals a consistent trajectory toward solving fundamental challenges in distributed database systems. His research has evolved from foundational work on column-stores and hybrid database architectures to cutting-edge innovations in geo-replicated transactions, concurrency control mechanisms, and the integration of machine learning with database systems. The trend shows increasing focus on practical implementations that address real-world scalability and performance challenges in large-scale data processing environments. ACM Fellow Churchill Scholarship recipient NSF CAREER Award winner Sloan Research Fellowship recipient VLDB Best Paper Award winner Two VLDB Test of Time Awards (for C-Store and HadoopDB) 2008 SIGMOD Jim Gray Doctoral Dissertation Award 2013-2014 Yale Provost's Teaching Prize 2013 VLDB Early Career Researcher Award Professor Abadi has successfully mentored several PhD students, most notably Alexander Thomson and Jose Falerio, both of whom won the prestigious SIGMOD Jim Gray Doctoral Dissertation Award for their work under his supervision. His research has been generously supported by multiple NSF grants including BIGDATA awards and other funding mechanisms that have enabled significant advances in database technology. He actively collaborates with industry partners, with several of his research projects leading directly to commercial products. At the University of Maryland, Abadi directs the Data Systems Lab at Maryland (DSLAM), which focuses on developing innovative database technologies that address contemporary challenges in data management. The lab's research spans distributed transaction processing, database architecture, and the integration of database systems with emerging computing paradigms. Notable projects include SLOG (Serializable, Low-latency, Geo-replicated Transactions), which eliminates traditional tradeoffs in distributed database design, and ongoing work in deterministic database systems that provide strong consistency guarantees without sacrificing performance.
Swiss Federal Institute of Technology in LausanneSwitzerland
Babak Falsafi is a Full Professor at the School of Computer and Communication Sciences (IC) at EPFL, leading the Parallel Systems Architecture Laboratory (PARSA). He is a renowned expert in computer architecture, datacenter systems, and cloud-native server design. His research focuses on post-Moore era computing, emphasizing heterogeneous architectures, energy efficiency, and scalable IT infrastructure. Falsafi is the founder of EcoCloud, an EC-sponsored industrial-academic consortium investigating sustainable information technology. He holds ACM and IEEE fellowships, a Sloan Research Fellowship, and has contributed to major projects like Optimus Prime (data transformation acceleration), AstriFlash (flash-based online service systems), and Midgard (virtual memory re-design). His work spans hardware-software co-design, memory systems, and security. Falsafi advises numerous PhD students and collaborates with industry partners such as Google and Cavium. Key achievements include pioneering scalable multiprocessor architectures, snoop filters in IBM BlueGene, and spatial memory streaming in ARM cores. His lab develops open-source tools like QFlex for server simulation. He frequently presents at top conferences (HPCA, ISCA, MICRO) and chairs workshops on post-Moore infrastructure. Teaching roles include leading courses in computer architecture and parallel systems across multiple EPFL departments (SIN, EDIC, SSC, SMA). His work addresses datacenter challenges like the 'data tax' and mitigating latency through specialized accelerators.
Hyunghoon Cho is an Assistant Professor at Yale School of Medicine in the Department of Biomedical Informatics & Data Science, with a secondary appointment in the Department of Computer Science. He received his PhD in Electrical Engineering and Computer Science from MIT (2019) and MS/BS in Computer Science from Stanford University (2013). His research focuses on computational challenges in biomedical data privacy, single-cell genomics, and network biology. Assistant Professor (Primary): Biomedical Informatics & Data Science Assistant Professor (Secondary): Computer Science Appointments: Yale School of Medicine | Broad Institute (Schmidt Fellow) Research Themes: Privacy-Enhancing Technologies for genomic and health data Scalable AI/ML tools for omics data analysis Structured biological modeling for system-level discovery His work includes secure GWAS, transcriptomic privacy assessment, and sfkit - a federated genomic analysis toolkit. He received the NIH Director's Early Independence Award and leads NSF-funded projects on confidential genome analytics. Awards: NIH Director's Early Independence Award Lab Members: Haris Smajlović (Postdoc), Vincent Angelo (CBB MS), Denis Loginov (Senior Software Engineer), Lucy Zheng (CBB PhD)
Yang Zhou is an Associate Professor in the Department of Computer Science and Software Engineering at Auburn University, part of the Samuel Ginn College of Engineering. His research focuses on big data algorithms, machine learning, data mining, and distributed computing. He has contributed to advancements in federated learning frameworks, graph mining tools, and spatial machine learning for environmental applications like flood mapping. Education includes a Ph.D. in Computer Science from Georgia Tech (2021), M.E. in Computer Application Technology from Chongqing University (2016), and B.E. in Engineering from Jiangnan University (2014). His work emphasizes scalable algorithms for large-scale systems, with tools like DirDense for dense subgraph mining and FedASMU for federated learning optimization. Recent publications explore adversarial robustness, blockchain strategies in IoT, and curriculum-based learning for large language models. He advises on interdisciplinary projects at the intersection of AI and environmental science.