Sunita Sarawagi is a Professor at Computer Science and Engineering , IIT Bombay, and a member of AI Labs@CSE . She is also associated with the Center for Machine Intelligence and Data Science (CMInDS), which she founded in 2020. Education: PhD in Computer Science from UC Berkeley (Thesis: Query Processing in Tertiary Memory Databases), BTech in Computer Science from IIT Kharagpur Research Interests span machine learning , data analytics , graphical models , and structured learning , with applications in text segmentation, sequence modeling, domain adaptation, and human-in-the-loop systems. Her publications reveal a strong focus on integrating data mining with database systems , temporal data analysis , and information extraction using probabilistic methods. Professional Activities include serving on the IEEE John Von Neumann Medal committee (2017-), VLDB 2011 Research Track Co-chair , and multiple program committee roles at top conferences like ICML, KDD, and SIGMOD. Labs & Teams : Leads the SS Lab , a research group focused on probabilistic graphical models, sequence modeling, and data integration techniques.
Kaiming He is an Associate Professor with tenure in the Department of Electrical Engineering and Computer Science (EECS) at the Massachusetts Institute of Technology (MIT), holding the Douglas Ross (1954) Career Development Professor of Software Technology chair. He also works part-time as a Distinguished Scientist at Google DeepMind. Prior to joining MIT in 2024, he was a research scientist at Facebook AI Research (FAIR) from 2016 to 2024, and a researcher at Microsoft Research Asia (MSRA) from 2011 to 2016. Dr. He received his PhD from the Chinese University of Hong Kong in 2011 and his Bachelor of Science from Tsinghua University in 2007. His academic journey reflects a strong foundation in computer science and engineering that has led to transformative contributions in artificial intelligence. His research primarily focuses on computer vision and deep learning, with pioneering work on deep residual networks (ResNets), visual object detection and segmentation, and self-supervised learning. He is best known for his work on Deep Residual Networks (ResNets), recognized as the most-cited paper of the twenty-first century. The residual connections he pioneered are now fundamental components in modern deep learning architectures including Transformers, AlphaGo Zero, AlphaFold, and various generative AI models. His recent publications demonstrate continued innovation across generative models, transformer architectures, and cross-disciplinary AI applications. His work bridges theoretical advances in neural network design with practical implementations that address real-world challenges in physics, biology, and other scientific domains. PAMI Young Researcher Award (2018) Best Paper Award, CVPR (2009, 2016) Best Paper Award, ICCV (2017) Best Student Paper Award, ICCV (2017) Everingham Prize, ICCV (2021) Most-cited paper of the twenty-first century Dr. He advises graduate students including Jake Austin, Xingjian Bai, and Mingyang Deng, and teaches advanced courses such as "6.S978: Deep Generative Models" (Fall 2024) and "6.8300/6.8301: Advances in Computer Vision" (Spring 2024). His research group actively explores how AI can serve as a unifying framework across scientific disciplines, breaking down traditional barriers between fields through shared methodologies and tools.
Wei Xu is an Associate Professor at Georgia Institute of Technology's College of Computing and School of Interactive Computing, with affiliations to the Machine Learning Center. Their research bridges machine learning, natural language processing, and social media with focus areas in large language models, cultural bias mitigation, multilingual capabilities, and human-AI collaboration in text evaluation. NSF CAREER and Google Academic Research Award recipient Director of NLP X Lab PhD from New York University, BSMS from Tsinghua University Research interests span: Multilingual Multicultural LLMs addressing representational gaps and cultural adaptation in language models (NAACL 2025, ACL 2024); Robustness and Reasoning through dynamic AGI evaluations (ACL 2024, EMNLP 2024); Interdisciplinary NLP applications in security, healthcare, and law (EMNLP 2024, ACL 2024). Recent publications focus on multilingual alignment (NAACL 2025), privacy risk estimation (arXiv 2025), cultural bias analysis (ACL 2024), and medical text simplification (EMNLP 2024). Key themes include bias mitigation, multimodal processing, and practical LLM evaluation. Scientific Awards : NSF CAREER, Google/Sony/Criteo research awards, ACL'24 Best Social Impact Award, COLING'18 Best Paper Advising 15 PhD/MS/BSMS students including Yao Dou (human-centered LLM evaluation), Tarek Naous (multilingual LLMs), and alumni like Chao Jiang (Apple AI/ML) and Yang Chen (NVIDIA research scientist). Teaches graduate courses on NLP and LLMs.
Andrea Vedaldi is a Professor of Computer Vision and Machine Learning at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). He specializes in unsupervised methods for understanding images and videos, focusing on 3D geometry and semantics. His research bridges foundational AI and practical applications, with contributions to generative models, neural fields, and self-supervised learning. Education: PhD in Computer Science (2008), University of California, Los Angeles MSc in Computer Science (2005), UCLA BSc in Information Engineering (2003), University of Padua Research Interests: Unsupervised learning, 3D perception, generative AI, neural rendering, and scalable vision systems. His work emphasizes ethical, responsible AI aligned with ERC-funded projects like UNION (ERC Consolidator Grant). Key Contributions: Co-developer of VLFeat and MatConvNet libraries Leader in 3D reconstruction and diffusion models (e.g., CatFree3D) Recipient of the PAMI Thomas S. Huang Prize and multiple best paper awards Grants & Service: Principal Investigator on £2.3M ERC Consolidator Grant (UNION) Co-organizer of major conferences (ECCV 2020 Program Chair, CVPR 2023 Area Chair) Reviewer for top journals/conferences (PAMI, CVPR, NeurIPS) Labs & Teams: VGG Group at Oxford, collaborating on projects like Meta 3D Gen and Common Objects in 3D (CO3D).
Jean Oh is a Researcher at the Robotics Institute of Carnegie Mellon University (CMU) , leading the interdisciplinary Bot Intelligence Group (BIG) . Her work focuses on developing persistent robots that co-exist and collaborate with humans in shared environments, emphasizing continuous improvement through training, exploration, and human interaction. Education: Ph.D. in Language and Information Technologies, CMU M.S. in Computer Science, Columbia University B.S. in Biotechnology, Yonsei University Oh's research integrates vision, language, and planning systems in robotics, with applications in human-robot teaming , self-driving cars , disaster response , eldercare , and creative robotics . She has pioneered projects like socially-compliant robot navigation in human crowds and AI-driven robotic painting systems. Recent publication trends highlight her work in vision-language planning , social navigation , computational creativity , and human-robot collaboration . Notable contributions include the StyleCLIPDraw algorithm for text-to-art generation and Social-PatteRNN for human-like trajectory prediction. Scientific Awards: Best Paper Award in Cognitive Robotics (ICRA'18, ICRA'15) Best Systems Paper Finalist (HRI'25) Best Oral Paper Finalist (Humanoids'24) Best Paper in Entertainment (IROS'24) Argoverse Challenge Winner (CVPR'24) Best Student Paper (AIAA'24) Best Demo Finalist (RoboSoft'24) Oh mentors a diverse team of PhD, MS, and undergraduate students from CMU departments including Robotics, Computer Science, and Mechanical Engineering. Her research is funded by US Army Research Lab , DiDi Chuxing , and DARPA , with collaborations across industry and academia .
Antoine Bosselut is a Tenure Track Assistant Professor at École Polytechnique Fédérale de Lausanne (EPFL) in the School of Computer and Communication Sciences, where he leads the EPFL NLP group. His research focuses on developing AI reasoning agents that can model, represent, and reason about human and world knowledge, with applications in health, education, and global fairness. His research interests span multiple critical areas in modern AI: LLM Representations of Knowledge: Understanding what language models know and how they represent that knowledge internally Reasoning Algorithms: Developing methods to improve LLMs' reasoning capabilities through symbolic systems, neuroscience, and cognitive science Large-scale AI Development: Creating open-source foundation models with multilingual capabilities AI Democratization: Ensuring equitable access to AI technologies across different cultural and regulatory contexts His recent publications reveal a strong focus on the intersection of language models and cognitive science, with multiple studies examining how LLMs align with human cognition. He's also deeply engaged in practical applications of NLP technology, particularly in education and multilingual settings, as evidenced by his work on evaluating AI's impact on higher education and developing culturally-aware language models. His scientific achievements have been recognized with several prestigious awards: Outstanding Paper Award at NAACL 2025 ELLIS Scholar designation in 2024 Outstanding Paper Award at ACL 2023 Inclusion in Forbes 30 Under 30 list for Science & Healthcare in 2021 Bosselut actively mentors numerous PhD students across diverse NLP research areas and has secured significant recognition for his work in both academic and media circles. His lab has been featured in major publications including Communications of the ACM, The Atlantic, and Quanta Magazine, highlighting the societal impact of his research on commonsense reasoning and AI capabilities.
Stefano Ermon is an Associate Professor in the Department of Computer Science at Stanford University, affiliated with the Artificial Intelligence Laboratory and a Senior Fellow at the Woods Institute for the Environment. His research focuses on advancing machine learning and generative AI techniques to address societal and environmental challenges, including computational sustainability, geospatial analysis, and climate science. He holds a Ph.D. from Cornell University (2015). Education: Ph.D. in Computer Science, Cornell University (2015). Research Interests: Ermon’s work bridges foundational machine learning (e.g., diffusion models, generative AI, and optimization) with applications in sustainability, geospatial analysis (via satellite imagery), and earth observation systems. Notable contributions include predicting poverty using satellite data and developing scalable methods for molecule generation. Articles Trends: His recent work emphasizes diffusion models for generative tasks (e.g., text-to-image, molecule design), geospatial AI (e.g., environmental monitoring), and ethical AI (e.g., bias mitigation in LLMs). He also explores applications in robotics and scientific computing. Awards: He has received prestigious awards, including the ICML 2024 Best Paper Award, Sloan Research Fellowship, Microsoft Research Faculty Fellowship, and the IJCAI Computers and Thought Award. Advising and Grants: Ermon teaches courses like Probabilistic Graphical Models (CS228) and has secured grants from NSF, ONR, AFOSR, and private foundations. His lab develops tools for climate science and sustainable development. Labs/Teams: Leads the Stanford AI Lab group focused on computational sustainability and generative AI, collaborating with institutions like the Woods Institute for environmental applications.
Shivaram Kalyanakrishnan is an Associate Professor at the Department of Computer Science and Engineering , Indian Institute of Technology Bombay , specialising in Artificial Intelligence and Machine Learning . His research spans sequential decision making , multiagent learning , multi-armed bandits , and humanoid robotics , with applications in robot soccer , computer games , and online advertising . He teaches advanced courses like CS 747: Foundations of Intelligent and Learning Agents and CS 748: Advances in Intelligent and Learning Agents , focusing on end-to-end system design and theoretical analysis. His scientific awards include the Best Student Paper Award at RoboCup International Symposium 2006 and nomination for Best Student Paper Award at AAMAS 2007 . His work on reinforcement learning and policy iteration has been published in leading venues such as IJCAI , ICML , and COLT , with recent contributions to railway scheduling and bandit algorithms. While no explicit list of advisees is provided, his research projects and publications suggest mentorship of students in collaborative efforts. Contact : shivaram@cse.iitb.ac.in .
Boris Hanin is an Associate Professor at Princeton University's Department of Operations Research and Financial Engineering (ORFE) and Associated Faculty at the Program in Applied and Computational Mathematics (PACM). Prior to Princeton, he held academic positions at Texas A&M University (Assistant Professor of Mathematics), MIT (NSF Postdoctoral Fellow), and Northwestern University (PhD in Mathematics under Steve Zelditch). He works part-time at Foundry, an AI/computing startup, leading the Foundry Institute. PhD in Mathematics, Northwestern University NSF Postdoctoral Fellowship, MIT Mathematics Associate Professor, Princeton ORFE Part-Time Leader, Foundry Institute His research spans machine learning, probability theory, and mathematical physics, focusing on neural network theory (approximation power, optimization guarantees), random matrix theory, and spectral asymptotics. He has made foundational contributions to understanding gradient behavior, initialization, and infinite-width limits in deep learning. Boris has supervised numerous PhD students and postdocs, with former members securing prestigious positions at Harvard, MIT, and Huawei. He serves as Associate Editor for journals like Pure and Applied Analysis and Mathematics of Operations Research , and has taught short courses at Oxford, Luxembourg, and Tor Vergata on statistical physics of neural networks and deep learning theory. 2024 Sloan Fellowship in Mathematics NSF CAREER grant DMS-2143754 NSF grant DMS-2133806 His research group collaborates on topics including hyperparameter transfer, architecture-aware scaling, and spectral analysis of random waves and neural networks. He actively contributes to theoretical machine learning through foundational publications in venues like Probability Theory and Related Fields , Journal of Machine Learning Research , and Communications in Mathematical Physics .
Prof. Dr. Thomas Hofmann is a Full Professor and Head of the Department of Computer Science at ETH Zurich since 2014. He also leads the Institute for Machine Learning. His research focuses on machine learning, deep learning, natural language understanding, and text understanding. Hofmann holds a Ph.D. from the University of Bonn (1997) and has held academic positions at Brown University (1999–2004) and TU Darmstadt. He transitioned to industry as Director of Engineering at Google (2006–2014), leading the Zurich R&D center, before returning to academia. He co-founded Recommind (2000) and 1plusX (Swiss marketing tech company), currently serving as Chief Scientist and board member at 1plusX. Education: Ph.D. in Computer Science, University of Bonn (1997) Postdoctoral Work: MIT (CBCL/AI Lab), UC Berkeley (EECS/ICSI) His research explores advanced machine learning techniques, including diffusion models, generative adversarial networks, and optimization dynamics. Hofmann’s entrepreneurial ventures reflect his focus on applying AI to real-world challenges, such as e-discovery and marketing technology. His work spans theoretical contributions (e.g., neural network training dynamics, continual learning) and applied innovations (e.g., image editing, portrait generation). Hofmann actively bridges academia and industry, influencing both research and commercial AI applications.
Kevin Jamieson is an Associate Professor at the Paul G. Allen School of Computer Science & Engineering and an Adjunct Professor in the Department of Statistics at the University of Washington . His academic journey includes a B.S. (2009) , M.S. (2010) , and Ph.D. (2015) in electrical engineering from the University of Washington, Columbia University, and University of Wisconsin–Madison respectively. He completed a postdoc at UC Berkeley's AMP Lab before joining UW in 2017. Ph.D., Electrical Engineering, University of Wisconsin–Madison (2015) M.S., Electrical Engineering, Columbia University (2010) B.S., Electrical Engineering, University of Washington (2009) Jamieson's research lies at the intersection of interactive machine learning , active learning , and sequential decision making . His work focuses on: Adaptive sampling strategies in multi-armed bandits and reinforcement learning (RL) Developing instance-dependent optimal algorithms that adapt to problem difficulty Applications in robotics , human perception studies , and hyperparameter optimization Representation learning for large models and experimental design frameworks His 15 most recent publications (2025-2022) demonstrate expertise in bandit theory , contextual RL , and game-theoretic learning . Notable trends include sample-efficient optimization , adaptive A/B testing , and sim-to-real transfer in robotics. Jamieson has received: NSF CAREER award for foundational contributions Amazon Faculty Research award for innovation in learning systems He actively recruits graduate students and postdocs , emphasizing collaboration in areas like: Multi-agent RL and strategic actor learning Empirical process suprema and adaptive sampling theory Applications in robotics , large language model finetuning , and biomedical data analysis Jamieson leads the Washington AI Lab (WAIL) and develops open-source learning systems like the NEXT framework for real-world adaptive data collection. He serves as co-PI for the Institute for the Foundations of Data Science (IFDS) and co-organizes the Distinguished Seminar in Optimization & Data .
Robert West is an Associate Professor at EPFL (École polytechnique fédérale de Lausanne) in the School of Computer and Communication Sciences , leading the Data Science Lab (dlab) . His research focuses on Natural Language Processing , Machine Learning , and Computational Social Science , analyzing human-generated data from the web, social media, and online platforms. Education : PhD in Computer Science (2016) - Stanford University MSc in Computer Science (2010) - McGill University BSc in Computer Science (2007) - Technische Universität München Research Interests : West develops algorithms for analyzing large-scale web data, with emphasis on multilingual NLP , social network analysis , and AI ethics . His work bridges machine learning with social science to understand digital human behavior. Scientific Awards : ICWSM’22 Adamic–Glance Distinguished Young Researcher Award Google Faculty Research Award Facebook Research Award Multiple Outstanding Paper Awards at ICWSM and WWW Advising & Grants : He advises 12 PhD students and has secured funding from the Swiss National Science Foundation , Swiss Data Science Center , and industry partners. His lab maintains collaborations with Microsoft Research and CROSS . Labs & Collaborations : West leads the Data Science Lab at EPFL, which focuses on web-scale data analysis , privacy-preserving machine learning , and AI for social good . The lab develops tools like Wikispeedia and Quotebank for public data exploration.
Shimon Whiteson is Professor of Computer Science at the University of Oxford, leading the Whiteson Research Lab focused on reinforcement learning, multi-agent systems, and deep learning. His research develops algorithms for efficient learning in complex environments. Current work explores meta-reinforcement learning frameworks that enable agents to rapidly adapt to new tasks, with applications in autonomous driving simulation and robotics. Recent innovations include novel methods for offline reinforcement learning, multi-agent coordination, and morphology-aware control. Publications demonstrate advances in: Meta-RL algorithm design for few-shot adaptation Multi-agent reinforcement learning environments and benchmarks Imitation learning in autonomous driving Bayesian methods for sample-efficient learning Research outputs include widely used benchmarks and tools including JaxMARL for accelerated multi-agent RL research. Current doctoral supervision focuses on temporal abstraction in RL, multi-agent coordination, and reinforcement learning theory.
Yee Whye Teh is a Professor at the Department of Statistics, University of Oxford, and a research scientist at DeepMind. His work focuses on statistical machine learning, including probabilistic learning, Bayesian nonparametrics, deep learning, and Monte Carlo methods. He co-directs the ELLIS programme on Robust Machine Learning and has held roles such as Programme Co-chair for ICML 2017. Teh has delivered keynotes at UAI 2019, an IMS Medallion Lecture at JSM 2019, and the Breiman Lecture in 2017. His research emphasizes scalable inference algorithms, hierarchical models, and applications in genetics and natural language processing. Teh's educational background includes a PhD from the University of Toronto (2003) and a Master's from the same institution (2000). He has contributed to widely used software tools like the Sequence Memoizer and has been recognized for his work through prestigious lectureships. Research interests span Bayesian nonparametric models, MCMC methods, and their applications in genetics and data compression. His lab collaborates on projects like fragmentation-coagulation processes for genetic variation modeling and Mondrian forests for online learning. Teh advises students through Oxford's graduate programs, though he notes high demand for mentorship. His work often bridges theory and practice, addressing challenges in big data learning and small data problems.
Jason D. Lee is an associate professor of Electrical Engineering and Computer Sciences (EECS) and Statistics at the University of California, Berkeley. Previously, he held academic positions at Princeton University as an associate professor, and was a research scientist at Google DeepMind. He completed his Ph.D. in Computational and Mathematical Engineering at Stanford University under the advisement of Trevor Hastie and Jonathan Taylor. For students and collaborators, his primary contact email is jasonlee@princeton.edu, though prospective students and postdocs are asked to include "filter_student" in the subject line. Ph.D., Computational and Mathematical Engineering, Stanford University (2015) B.Sc., Mathematics, Duke University (2010) Lee's research lies at the intersection of machine learning, statistics, and optimization, focusing on the theoretical foundations of artificial intelligence. His work addresses fundamental questions in deep learning, including optimization landscapes, representation learning, and reinforcement learning theory. He has made significant contributions to understanding how gradient descent operates in neural network training and has developed provably efficient algorithms for various learning scenarios. His ten most recent publications represent a diverse yet coherent body of work across machine learning theory, focusing on topics such as Gaussian multi-index models, transformer learning capabilities, shallow neural networks, and optimization techniques. These publications appear in top venues including COLT, ICML, NeurIPS, and JMLR. Among his notable accolades are: Samsung AI Researcher of the Year Award (2023) NSF Career Award (2022) ONR Young Investigator Award (2021) Sloan Research Fellow in Computer Science (2019) NIPS Best Student Paper Award (2016) Princeton Commendation for Outstanding Teaching (ECE538B) Lee has advised numerous students including Alex Damian, Wenhao Zhan, Eshaan Nichani, Tianle Cai, Zixuan Wang, and Yunwei Ren. Former advisees have gone on to positions at institutions like NYU Courant, Facebook AI Research, UW, MIT, Duke, and Microsoft Research. His research group and collaborators span multiple institutions, working on theoretical and applied aspects of machine learning and artificial intelligence, with a particular focus on the optimization and learning dynamics of neural networks and transformer models.