Sunita Sarawagi is a Professor at Computer Science and Engineering , IIT Bombay, and a member of AI Labs@CSE . She is also associated with the Center for Machine Intelligence and Data Science (CMInDS), which she founded in 2020. Education: PhD in Computer Science from UC Berkeley (Thesis: Query Processing in Tertiary Memory Databases), BTech in Computer Science from IIT Kharagpur Research Interests span machine learning , data analytics , graphical models , and structured learning , with applications in text segmentation, sequence modeling, domain adaptation, and human-in-the-loop systems. Her publications reveal a strong focus on integrating data mining with database systems , temporal data analysis , and information extraction using probabilistic methods. Professional Activities include serving on the IEEE John Von Neumann Medal committee (2017-), VLDB 2011 Research Track Co-chair , and multiple program committee roles at top conferences like ICML, KDD, and SIGMOD. Labs & Teams : Leads the SS Lab , a research group focused on probabilistic graphical models, sequence modeling, and data integration techniques.
Stefano Ermon is an Associate Professor in the Department of Computer Science at Stanford University, affiliated with the Artificial Intelligence Laboratory and a Senior Fellow at the Woods Institute for the Environment. His research focuses on advancing machine learning and generative AI techniques to address societal and environmental challenges, including computational sustainability, geospatial analysis, and climate science. He holds a Ph.D. from Cornell University (2015). Education: Ph.D. in Computer Science, Cornell University (2015). Research Interests: Ermon’s work bridges foundational machine learning (e.g., diffusion models, generative AI, and optimization) with applications in sustainability, geospatial analysis (via satellite imagery), and earth observation systems. Notable contributions include predicting poverty using satellite data and developing scalable methods for molecule generation. Articles Trends: His recent work emphasizes diffusion models for generative tasks (e.g., text-to-image, molecule design), geospatial AI (e.g., environmental monitoring), and ethical AI (e.g., bias mitigation in LLMs). He also explores applications in robotics and scientific computing. Awards: He has received prestigious awards, including the ICML 2024 Best Paper Award, Sloan Research Fellowship, Microsoft Research Faculty Fellowship, and the IJCAI Computers and Thought Award. Advising and Grants: Ermon teaches courses like Probabilistic Graphical Models (CS228) and has secured grants from NSF, ONR, AFOSR, and private foundations. His lab develops tools for climate science and sustainable development. Labs/Teams: Leads the Stanford AI Lab group focused on computational sustainability and generative AI, collaborating with institutions like the Woods Institute for environmental applications.
Mark Schmidt is a Professor in the Department of Computer Science at the University of British Columbia, Faculty of Science. He holds the Canada Research Chair in Large-Scale Machine Learning and is a Canada CIFAR AI Chair at the Alberta Machine Intelligence Institute. His research spans multiple centers including CAIDA (Centre for Artificial Intelligence Decision-making and Action), the Data Science Institute, and the Machine Intelligence Learning Discovery (MILD) group. Dr. Schmidt's educational background includes a Ph.D. from UBC (2005-2010), an M.Sc. from the University of Alberta (2003-2005), and a B.Sc. from the University of Alberta (2000-2003). His academic career progressed from Postdoc positions at Simon Fraser University, Ecole Normale Superieure, and UBC to Assistant Professor (2014-2019), Associate Professor (2019-2024), and current Professor (2024-present) at UBC. His research focuses on machine learning optimization, with particular emphasis on improving numerical algorithms for large-scale machine learning applications. His work bridges theoretical optimization and practical applications across computer vision, natural language processing, and reinforcement learning. He has developed numerous optimization techniques including variants of stochastic gradient methods, coordinate descent algorithms, and natural gradient approaches. Analysis of his recent publications reveals a strong focus on optimization for over-parameterized models, particularly transformers and large language models. His work addresses critical challenges in step-size selection, convergence guarantees, and efficient implementation of optimization algorithms. His research has significant implications for training deep neural networks more effectively and understanding why certain optimization methods outperform others in practice. Among his notable awards are the Dorothy Killam Fellowship (2025), Arthur B. McDonald Fellowship (2024), Sloan Research Fellowship (2017), and multiple UBC teaching awards. He has also received the Lagrange Prize in Continuous Optimization and Best Paper Award at AISTATS 2021. Dr. Schmidt actively supervises numerous graduate students, with over 40 PhD and Master's students listed as current or alumni members of his research group. His laboratory maintains strong connections with industry partners, with many alumni securing positions at leading AI companies including Google, Amazon, Meta, and Microsoft.
Jimmy Ba is an Assistant Professor in the Department of Computer Science at the University of Toronto and a CIFAR AI Chair. His research develops efficient learning algorithms for deep neural networks, with applications in reinforcement learning and AI. He completed his PhD under Geoffrey Hinton and holds multiple fellowships including the Facebook Graduate Fellowship. Research Focus: Neural network efficiency, reinforcement learning architectures, and optimization methods for deep learning systems. Teaching: Courses on Neural Networks, Deep Learning, and Inference Algorithms at University of Toronto. Awards: Facebook Graduate Fellowship (2016-2018) Massey College Junior Fellowship
Boris Hanin is an Associate Professor at Princeton University's Department of Operations Research and Financial Engineering (ORFE) and Associated Faculty at the Program in Applied and Computational Mathematics (PACM). Prior to Princeton, he held academic positions at Texas A&M University (Assistant Professor of Mathematics), MIT (NSF Postdoctoral Fellow), and Northwestern University (PhD in Mathematics under Steve Zelditch). He works part-time at Foundry, an AI/computing startup, leading the Foundry Institute. PhD in Mathematics, Northwestern University NSF Postdoctoral Fellowship, MIT Mathematics Associate Professor, Princeton ORFE Part-Time Leader, Foundry Institute His research spans machine learning, probability theory, and mathematical physics, focusing on neural network theory (approximation power, optimization guarantees), random matrix theory, and spectral asymptotics. He has made foundational contributions to understanding gradient behavior, initialization, and infinite-width limits in deep learning. Boris has supervised numerous PhD students and postdocs, with former members securing prestigious positions at Harvard, MIT, and Huawei. He serves as Associate Editor for journals like Pure and Applied Analysis and Mathematics of Operations Research , and has taught short courses at Oxford, Luxembourg, and Tor Vergata on statistical physics of neural networks and deep learning theory. 2024 Sloan Fellowship in Mathematics NSF CAREER grant DMS-2143754 NSF grant DMS-2133806 His research group collaborates on topics including hyperparameter transfer, architecture-aware scaling, and spectral analysis of random waves and neural networks. He actively contributes to theoretical machine learning through foundational publications in venues like Probability Theory and Related Fields , Journal of Machine Learning Research , and Communications in Mathematical Physics .
Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Jason D. Lee is an associate professor of Electrical Engineering and Computer Sciences (EECS) and Statistics at the University of California, Berkeley. Previously, he held academic positions at Princeton University as an associate professor, and was a research scientist at Google DeepMind. He completed his Ph.D. in Computational and Mathematical Engineering at Stanford University under the advisement of Trevor Hastie and Jonathan Taylor. For students and collaborators, his primary contact email is jasonlee@princeton.edu, though prospective students and postdocs are asked to include "filter_student" in the subject line. Ph.D., Computational and Mathematical Engineering, Stanford University (2015) B.Sc., Mathematics, Duke University (2010) Lee's research lies at the intersection of machine learning, statistics, and optimization, focusing on the theoretical foundations of artificial intelligence. His work addresses fundamental questions in deep learning, including optimization landscapes, representation learning, and reinforcement learning theory. He has made significant contributions to understanding how gradient descent operates in neural network training and has developed provably efficient algorithms for various learning scenarios. His ten most recent publications represent a diverse yet coherent body of work across machine learning theory, focusing on topics such as Gaussian multi-index models, transformer learning capabilities, shallow neural networks, and optimization techniques. These publications appear in top venues including COLT, ICML, NeurIPS, and JMLR. Among his notable accolades are: Samsung AI Researcher of the Year Award (2023) NSF Career Award (2022) ONR Young Investigator Award (2021) Sloan Research Fellow in Computer Science (2019) NIPS Best Student Paper Award (2016) Princeton Commendation for Outstanding Teaching (ECE538B) Lee has advised numerous students including Alex Damian, Wenhao Zhan, Eshaan Nichani, Tianle Cai, Zixuan Wang, and Yunwei Ren. Former advisees have gone on to positions at institutions like NYU Courant, Facebook AI Research, UW, MIT, Duke, and Microsoft Research. His research group and collaborators span multiple institutions, working on theoretical and applied aspects of machine learning and artificial intelligence, with a particular focus on the optimization and learning dynamics of neural networks and transformer models.
Yee Whye Teh is a Professor at the Department of Statistics, University of Oxford, and a research scientist at DeepMind. His work focuses on statistical machine learning, including probabilistic learning, Bayesian nonparametrics, deep learning, and Monte Carlo methods. He co-directs the ELLIS programme on Robust Machine Learning and has held roles such as Programme Co-chair for ICML 2017. Teh has delivered keynotes at UAI 2019, an IMS Medallion Lecture at JSM 2019, and the Breiman Lecture in 2017. His research emphasizes scalable inference algorithms, hierarchical models, and applications in genetics and natural language processing. Teh's educational background includes a PhD from the University of Toronto (2003) and a Master's from the same institution (2000). He has contributed to widely used software tools like the Sequence Memoizer and has been recognized for his work through prestigious lectureships. Research interests span Bayesian nonparametric models, MCMC methods, and their applications in genetics and data compression. His lab collaborates on projects like fragmentation-coagulation processes for genetic variation modeling and Mondrian forests for online learning. Teh advises students through Oxford's graduate programs, though he notes high demand for mentorship. His work often bridges theory and practice, addressing challenges in big data learning and small data problems.
Fanny Yang is an Assistant Professor in the Computer Science Department at ETH Zurich . She previously held postdoctoral positions at Stanford University and a Junior Fellowship at the Institute for Theoretical Studies at ETH Zurich, advised by Nicolai Meinshausen . Her PhD was completed at the EECS Department of UC Berkeley , supervised by Martin Wainwright . Research Interests : Theoretical foundations of machine learning and statistics , particularly focusing on overparameterized models and high-dimensional data . Developing trustworthy ML models with emphasis on distributional robustness , domain generalization , and interpretability . Applications in medical diagnostics and treatment effect analysis , aiming to address reliability issues in real-world domains. Recent Work Trends : 2025 Articles : Theoretical analysis of semi-supervised multi-objective learning , test-time scaling with verifiers, and foundation models for efficient randomized experiments . 2024 Articles : Studies on robust mixture learning , privacy-preserving data synthesis via optimal transport , confounding quantification in causal inference , and semi-private learning frameworks. 2023 Articles : Investigations into active vs. passive learning in high dimensions, inductive bias in noisy interpolation , and semi-supervised novelty detection using model ensembles .
Benjamin Van Roy is a Professor at Stanford University since 1998, affiliated with the Departments of Electrical Engineering and Management Science and Engineering, and the Institute for Computational and Mathematical Engineering. He leads the Efficient Agent Team at Google DeepMind and previously held leadership roles at Morgan Stanley, Unica, and Enuvis. He holds SB, SM, and PhD degrees in Computer Science and Electrical Engineering from MIT, advised by John Tsitsiklis. His research focuses on reinforcement learning, alignment, and information theory, with contributions to machine learning foundations, optimization, and finance. He has authored over 150 publications, including influential works on Thompson Sampling, approximate dynamic programming, and exploration strategies. His honors include INFORMS and IEEE Fellowships and the INFORMS Lanchester Prize. Van Roy advises doctoral students across academia and industry, with graduates at top institutions and companies like Meta, Tesla, and Citadel. He teaches courses on reinforcement learning, stochastic control, and optimization. His open-source projects include Epistemic Neural Networks and the Neural Testbed for evaluating machine learning models. Key contributions include foundational work in reinforcement learning theory, scalable methods for recommendation systems, and applications in finance and resource allocation. His research bridges theoretical insights with practical applications, emphasizing alignment and safety of AI systems.
Sewoong Oh is a Professor at the Paul G. Allen School of Computer Science & Engineering at the University of Washington, where he has been faculty since 2019. His research focuses on the foundations of machine learning with particular emphasis on differential privacy, secure and robust machine learning, and federated learning. Prior to joining UW, he was at the Department of Industrial and Enterprise Systems Engineering at the University of Illinois at Urbana-Champaign from 2012-2019. He is affiliated with multiple NSF AI Institutes including ACTION (Agent-based Cyber Threat Intelligence and Operation), AI-EDGE (Future Edge Networks and Distributed Intelligence), and IFML (Foundations of Machine Learning). Oh received his PhD in Electrical Engineering from Stanford University in 2011 under Andrea Montanari, followed by postdoctoral work at MIT's Laboratory for Information and Decision Systems under Devavrat Shah. His educational background spans top institutions in theoretical computer science and electrical engineering. His research spans the critical intersection of machine learning security, privacy, and robustness. Oh's work addresses fundamental challenges in making AI systems reliable and trustworthy, with particular focus on defending against backdoor attacks, developing privacy-preserving algorithms, and creating efficient tokenization methods for language models. His recent SuperBPE work demonstrates how moving beyond traditional subword tokenization can significantly improve language model efficiency and performance. His research consistently bridges theoretical foundations with practical implementations, as evidenced by numerous open-source repositories containing production-ready code. Oh's publications reveal a consistent trajectory toward addressing security and privacy concerns in increasingly complex machine learning systems, with recent work focusing on language model tokenization, data-centric AI, and federated learning frameworks. His research shows strong interdisciplinary connections between theoretical computer science, statistics, and practical machine learning systems. ACM SIGMETRICS best paper award (2015) NSF CAREER award (2016) ACM SIGMETRICS rising star award (2017) GOOGLE Faculty Research Awards (2017, 2020) 2024 ICML Best Paper Award (for work by advisee Jon Hayase) Professor Oh maintains an active research group with numerous PhD students, postdocs, and undergraduate researchers. His students have secured prestigious positions at leading tech companies including Amazon, Google, and Snap, as well as academic appointments at institutions like University of Wisconsin-Madison and Shanghai Jiao Tong University. He has received substantial research funding through his participation in multiple NSF AI Institutes and industry research awards from Google. His lab maintains several active GitHub repositories implementing cutting-edge research in backdoor defense, robust statistics, and privacy-preserving machine learning.
Adrian Weller is a prominent researcher and academic at the University of Cambridge, serving as a Director of Research in Machine Learning within the Department of Engineering. He holds multiple significant leadership roles including Programme Director for Trust and Society at the Leverhulme Centre for the Future of Intelligence (CFI), and previously served as Programme Director for AI at The Alan Turing Institute, the UK national institute for data science and AI. His work bridges theoretical machine learning research with practical applications and societal implications of artificial intelligence. Weller's research interests span a broad spectrum of AI and machine learning topics with a particular focus on ensuring beneficial societal outcomes. His work encompasses explainability, fairness, robustness, scalability, privacy, safety, and ethics in AI systems. He has made significant contributions to trustworthy machine learning, including developing frameworks for AI governance, certification, and human-AI collaboration. His research group actively investigates neuro-symbolic approaches, privacy-preserving techniques, and methods for improving the reliability and interpretability of AI systems. His recent publications demonstrate a strong trend toward addressing the practical challenges of deploying AI systems in real-world contexts, particularly focusing on certification frameworks, governance mechanisms, and human-centered approaches. His work spans theoretical advances in machine learning architectures while maintaining a strong connection to societal impact, with publications appearing in top venues across AI, machine learning, and interdisciplinary applications. Scientific Awards: MBE for services to digital innovation (2022 Queen's Birthday Honours) Turing AI Fellowship for Trustworthy Machine Learning Weller actively supervises a large group of PhD students and postdocs, with current students including Juyeon Heo, Yanzhi Chen, Katie Collins, Isaac Reid, Yichao Liang, Herbie Bradley, and Shoaib Siddiqui. His former students have gone on to positions at leading institutions including Google DeepMind, ETH Zurich, NYU, and MPI-IS Tübingen. He has served on numerous advisory boards including the Centre for Data Ethics and Innovation, UNESCO's expert group on AI ethics, and the World Economic Forum's Global Future Council on AI. His research has been supported through his Turing AI Fellowship and various collaborative projects focused on safe and ethical AI development. Weller leads a vibrant research group focused on trustworthy machine learning, which actively organizes workshops and conferences including ICML 2024 (where he served as Program Chair), multiple workshops on responsible AI, and events through the ELLIS network. His group collaborates extensively across disciplines, working with researchers in computer science, social sciences, law, and policy to address the multifaceted challenges of developing beneficial AI systems.
Kenji Kawaguchi is the Presidential Young Professor in the Department of Computer Science at the National University of Singapore (NUS), where he leads the Deep Learning Lab and is a faculty affiliate at the NUS Institute of Data Science. His research bridges theoretical and applied machine learning, focusing on deep learning, large language models, and physics-informed neural networks. His educational background includes a Ph.D. and S.M. in Computer Science and Electrical Engineering from the Massachusetts Institute of Technology (MIT), advised by Leslie Pack Kaelbling, and a postdoctoral fellowship at Harvard University’s Center of Mathematical Sciences and Applications. Dr. Kawaguchi’s research interests center on the theoretical foundations of deep learning, optimization, generalization, and applications in areas such as molecular modeling, AI safety, and efficient training of large models. He has made significant contributions to understanding in-context learning, diffusion models, and neural operators for partial differential equations. His recent publications (2023–2025) reflect a strong trend toward improving the efficiency, robustness, and interpretability of large-scale models, particularly in language and scientific domains. Key themes include LLM alignment and safety, diffusion model optimization, and physics-informed learning for high-dimensional problems. Presidential Young Professor He has served as Area Chair and PC Member for top-tier conferences including NeurIPS, ICML, ICLR, AAAI, and UAI, and as reviewer for journals such as JMLR and Annals of Statistics. He has delivered invited talks at Harvard, MIT, Stanford, CMU, Brown, and Google Research, reflecting his international recognition. He actively mentors students and welcomes PhD candidates and postdocs to join his research group.
Manik Varma is a Distinguished Scientist and Vice President at Microsoft Research India, and an Adjunct Professor at the Indian Institute of Technology Delhi. He is a Fellow of the Indian Academies of Science (IASc, INSA, NASI), the Indian National Academy of Engineering (INAE), and the Association for Computing Machinery (ACM). He has received prestigious awards such as the Shanti Swarup Bhatnagar Prize and Microsoft Gold Star Award. Education : BSc in Physics from St. Stephen's College (David Raja Ram Prize) BA in Theoretical Physics from the University of Oxford (Rhodes Scholar) DPhil in Computer Vision and Machine Learning from the University of Oxford (University Scholar) Post-doctoral Fellow at the Mathematical Sciences Research Institute (MSRI), Berkeley Visiting Miller Professor at UC Berkeley His research focuses on Machine Learning (Extreme Classification, Resource-efficient ML, Supervised Learning), Information Retrieval (Computational Advertising, Dense Retrieval, Recommender Systems), and Computer Vision (Image Search, Object Recognition). Recent work includes graph-regularized encoders, label variance reduction, and multimodal classification frameworks. His publications span extreme classification algorithms like NGAME , SiameseXML , and DECAF , with applications in IoT, web search, and recommendation systems. He leads a research group at Microsoft Research India and advises PhD students at IIT Delhi. Scientific Awards : Shanti Swarup Bhatnagar Prize (Government of India) Microsoft Gold Star and Achievement Awards WSDM 2019 Best Paper Prize BuildSys 2019 Best Paper Runner-up Fellow of ACM, IASc, INSA, NASI, INAE He has supervised numerous PhD students, including Sonu Mehta and Suchith Prabhu, and collaborates with institutions like Microsoft Research India, IIT Delhi, and UC Berkeley. His research has led to scalable solutions for billion-label classification and resource-constrained IoT applications.
Aarti Singh is a Professor in the Machine Learning Department at Carnegie Mellon University and Director of the NSF AI Institute for Societal Decision Making. She leads research at the intersection of machine learning, statistics, and decision making, with applications to scientific and societal domains. Her work focuses on designing principled interactive algorithms for learning and decision making under uncertainty. Education: Ph.D. in Electrical Engineering, University of Wisconsin-Madison (2008) M.S. in Electrical Engineering, University of Wisconsin-Madison (2003) B.E. in Electronics and Communication Engineering, University of Delhi (2001) Research Interests: Professor Singh's research centers on developing interactive machine learning algorithms that go beyond finding input-output associations to make higher-level decisions about the most informative data and actions. Her work spans autonomous decision making, including active sampling, stochastic optimization, bandits, and reinforcement learning that are statistically optimal, computationally tractable, and robust. She also investigates human factors in decision making, designing algorithms that model and leverage human feedback while accounting for bias, memory effects, and calibration. Her research has applications in material science, cosmology, and peer review systems. Research Trends: Professor Singh's recent publications demonstrate a strong focus on reinforcement learning, particularly in developing more efficient and robust algorithms for decision making under uncertainty. Her work bridges theoretical foundations with practical applications, spanning from fundamental algorithm development to real-world implementation in scientific domains. There's a clear trajectory toward integrating human factors into decision-making algorithms, with significant contributions to peer review systems and preference learning. Scientific Awards: NSF Career Award United States Air Force Young Investigator Award A. Nico Habermann Faculty Chair Award Harold A. Peterson Best Dissertation Award Multiple paper awards Advising and Grants: Professor Singh has advised numerous PhD and master's students, many of whom have gone on to faculty positions or research roles at leading institutions. Her research is supported by prestigious grants from ONR, Simons Foundation, AFRL, ARL, and NSF. She serves as General Chair (2025) and Program Chair (2020) for the International Conference on Machine Learning (ICML) and has held leadership roles in multiple professional organizations. Research Team: Professor Singh leads a vibrant research group within the Machine Learning Department at CMU, with current PhD students working on topics including reinforcement learning, human-AI collaboration, and decision making under uncertainty. She also directs the NSF AI Institute for Societal Decision Making, which brings together researchers from multiple disciplines to develop AI systems that support human decision making in societal contexts.