Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Christian Theobalt is a Professor of Computer Science at Saarland University and Scientific Director of the Visual Computing and Artificial Intelligence Department at the Max Planck Institute for Informatics . He leads the Saarbruecken Center for Visual Computing as a strategic partnership between Google and MPI. PhD in Computer Science (2005) from MPI-INF/Saarland University Postdoctoral Researcher at MPI (2005-2007) Visiting Assistant Professor at Stanford (2007-2009) His research focuses on the intersection of Computer Graphics, Computer Vision, and Artificial Intelligence , with specializations in: 3D/4D Human Reconstruction Neural Rendering Performance Capture Geometric Deep Learning Volumetric Video Quantum Visual Computing Recent article trends show emphasis on: Quantum computing applications in visual reconstruction Neural rendering with radiance fields Human motion capture from egocentric views Gesture-language interaction modeling Multi-modal scene understanding Real-time free-viewpoint rendering Scientific Awards : Fellow of EUROGRAPHICS (2022) CVPR Best Student Paper Honorable Mention (2020) ERC Consolidator Grant (2017) Karl Heinz Beckurts Award (2017) Busy Beaver Teaching Award (2016) ERC Starting Grant (2013) Advising over 50 students and researchers including: Current researchers: Viktor Rudnev, Linjie Lyu, Mohit Mendiratta Postdocs: Kwang In Kim, Kiran Varanasi Alumni: Franziska Mueller (Google), Dushyant Mehta (Qualcomm), Ayush Tewari (MIT) Grants include ERC grants, Google Glass Research Award, and multiple industry partnerships. His lab maintains cutting-edge facilities with: Multi-camera capture systems Quantum annealing infrastructure HDR radiance field technology GPU clusters for AI research Time-of-flight imaging systems Event camera arrays
Prof. Dr. Thomas Hofmann is a Full Professor and Head of the Department of Computer Science at ETH Zurich since 2014. He also leads the Institute for Machine Learning. His research focuses on machine learning, deep learning, natural language understanding, and text understanding. Hofmann holds a Ph.D. from the University of Bonn (1997) and has held academic positions at Brown University (1999–2004) and TU Darmstadt. He transitioned to industry as Director of Engineering at Google (2006–2014), leading the Zurich R&D center, before returning to academia. He co-founded Recommind (2000) and 1plusX (Swiss marketing tech company), currently serving as Chief Scientist and board member at 1plusX. Education: Ph.D. in Computer Science, University of Bonn (1997) Postdoctoral Work: MIT (CBCL/AI Lab), UC Berkeley (EECS/ICSI) His research explores advanced machine learning techniques, including diffusion models, generative adversarial networks, and optimization dynamics. Hofmann’s entrepreneurial ventures reflect his focus on applying AI to real-world challenges, such as e-discovery and marketing technology. His work spans theoretical contributions (e.g., neural network training dynamics, continual learning) and applied innovations (e.g., image editing, portrait generation). Hofmann actively bridges academia and industry, influencing both research and commercial AI applications.
Kevin Jamieson is an Associate Professor at the Paul G. Allen School of Computer Science & Engineering and an Adjunct Professor in the Department of Statistics at the University of Washington . His academic journey includes a B.S. (2009) , M.S. (2010) , and Ph.D. (2015) in electrical engineering from the University of Washington, Columbia University, and University of Wisconsin–Madison respectively. He completed a postdoc at UC Berkeley's AMP Lab before joining UW in 2017. Ph.D., Electrical Engineering, University of Wisconsin–Madison (2015) M.S., Electrical Engineering, Columbia University (2010) B.S., Electrical Engineering, University of Washington (2009) Jamieson's research lies at the intersection of interactive machine learning , active learning , and sequential decision making . His work focuses on: Adaptive sampling strategies in multi-armed bandits and reinforcement learning (RL) Developing instance-dependent optimal algorithms that adapt to problem difficulty Applications in robotics , human perception studies , and hyperparameter optimization Representation learning for large models and experimental design frameworks His 15 most recent publications (2025-2022) demonstrate expertise in bandit theory , contextual RL , and game-theoretic learning . Notable trends include sample-efficient optimization , adaptive A/B testing , and sim-to-real transfer in robotics. Jamieson has received: NSF CAREER award for foundational contributions Amazon Faculty Research award for innovation in learning systems He actively recruits graduate students and postdocs , emphasizing collaboration in areas like: Multi-agent RL and strategic actor learning Empirical process suprema and adaptive sampling theory Applications in robotics , large language model finetuning , and biomedical data analysis Jamieson leads the Washington AI Lab (WAIL) and develops open-source learning systems like the NEXT framework for real-world adaptive data collection. He serves as co-PI for the Institute for the Foundations of Data Science (IFDS) and co-organizes the Distinguished Seminar in Optimization & Data .
Robert West is an Associate Professor at EPFL (École polytechnique fédérale de Lausanne) in the School of Computer and Communication Sciences , leading the Data Science Lab (dlab) . His research focuses on Natural Language Processing , Machine Learning , and Computational Social Science , analyzing human-generated data from the web, social media, and online platforms. Education : PhD in Computer Science (2016) - Stanford University MSc in Computer Science (2010) - McGill University BSc in Computer Science (2007) - Technische Universität München Research Interests : West develops algorithms for analyzing large-scale web data, with emphasis on multilingual NLP , social network analysis , and AI ethics . His work bridges machine learning with social science to understand digital human behavior. Scientific Awards : ICWSM’22 Adamic–Glance Distinguished Young Researcher Award Google Faculty Research Award Facebook Research Award Multiple Outstanding Paper Awards at ICWSM and WWW Advising & Grants : He advises 12 PhD students and has secured funding from the Swiss National Science Foundation , Swiss Data Science Center , and industry partners. His lab maintains collaborations with Microsoft Research and CROSS . Labs & Collaborations : West leads the Data Science Lab at EPFL, which focuses on web-scale data analysis , privacy-preserving machine learning , and AI for social good . The lab develops tools like Wikispeedia and Quotebank for public data exploration.
Jason D. Lee is an associate professor of Electrical Engineering and Computer Sciences (EECS) and Statistics at the University of California, Berkeley. Previously, he held academic positions at Princeton University as an associate professor, and was a research scientist at Google DeepMind. He completed his Ph.D. in Computational and Mathematical Engineering at Stanford University under the advisement of Trevor Hastie and Jonathan Taylor. For students and collaborators, his primary contact email is jasonlee@princeton.edu, though prospective students and postdocs are asked to include "filter_student" in the subject line. Ph.D., Computational and Mathematical Engineering, Stanford University (2015) B.Sc., Mathematics, Duke University (2010) Lee's research lies at the intersection of machine learning, statistics, and optimization, focusing on the theoretical foundations of artificial intelligence. His work addresses fundamental questions in deep learning, including optimization landscapes, representation learning, and reinforcement learning theory. He has made significant contributions to understanding how gradient descent operates in neural network training and has developed provably efficient algorithms for various learning scenarios. His ten most recent publications represent a diverse yet coherent body of work across machine learning theory, focusing on topics such as Gaussian multi-index models, transformer learning capabilities, shallow neural networks, and optimization techniques. These publications appear in top venues including COLT, ICML, NeurIPS, and JMLR. Among his notable accolades are: Samsung AI Researcher of the Year Award (2023) NSF Career Award (2022) ONR Young Investigator Award (2021) Sloan Research Fellow in Computer Science (2019) NIPS Best Student Paper Award (2016) Princeton Commendation for Outstanding Teaching (ECE538B) Lee has advised numerous students including Alex Damian, Wenhao Zhan, Eshaan Nichani, Tianle Cai, Zixuan Wang, and Yunwei Ren. Former advisees have gone on to positions at institutions like NYU Courant, Facebook AI Research, UW, MIT, Duke, and Microsoft Research. His research group and collaborators span multiple institutions, working on theoretical and applied aspects of machine learning and artificial intelligence, with a particular focus on the optimization and learning dynamics of neural networks and transformer models.
Yee Whye Teh is a Professor at the Department of Statistics, University of Oxford, and a research scientist at DeepMind. His work focuses on statistical machine learning, including probabilistic learning, Bayesian nonparametrics, deep learning, and Monte Carlo methods. He co-directs the ELLIS programme on Robust Machine Learning and has held roles such as Programme Co-chair for ICML 2017. Teh has delivered keynotes at UAI 2019, an IMS Medallion Lecture at JSM 2019, and the Breiman Lecture in 2017. His research emphasizes scalable inference algorithms, hierarchical models, and applications in genetics and natural language processing. Teh's educational background includes a PhD from the University of Toronto (2003) and a Master's from the same institution (2000). He has contributed to widely used software tools like the Sequence Memoizer and has been recognized for his work through prestigious lectureships. Research interests span Bayesian nonparametric models, MCMC methods, and their applications in genetics and data compression. His lab collaborates on projects like fragmentation-coagulation processes for genetic variation modeling and Mondrian forests for online learning. Teh advises students through Oxford's graduate programs, though he notes high demand for mentorship. His work often bridges theory and practice, addressing challenges in big data learning and small data problems.
Shrikanth (Shri) Narayanan is a University Professor and holder of the Niki and Max Nikias Chair in Engineering at the University of Southern California (USC), serving as the inaugural Vice President for Presidential Initiatives. He leads the Signal Analysis and Interpretation Lab (SAIL) and holds joint appointments in Computer Science, Linguistics, Psychology, Neuroscience, Pediatrics, and Otolaryngology-Head and Neck Surgery. His research focuses on speech and audio processing, behavioral signal processing, and real-time MRI of speech production, with applications in healthcare, education, and technology. Education: B.E. in Electrical Engineering from College of Engineering, Guindy (Chennai, India, 1988); M.S., Engineer, and Ph.D. in Electrical Engineering from UCLA (1990, 1992, 1995). Research interests span computational linguistics, machine learning, and multimodal human behavior analysis. He pioneered technologies for speech biomarkers in mental health, real-time MRI of speech production, and wearable sensor systems for longitudinal health studies. His work in speech emotion recognition, forensic interviews, and clinical applications has been recognized through over 40 awards, including the IEEE Flanagan Award and ISCA Medal. He has published extensively in journals like Proceedings of the IEEE , Journal of the Acoustical Society of America , and PLOS One . Key Grants: NSF CAREER, Okawa Research, IBM Faculty, Google/Amazon awards. Labs/Teams: Signal Analysis & Interpretation Lab (SAIL), USC Information Sciences Institute (ISI), Google Visiting Faculty Researcher.
Dr. Diyi Yang is an Assistant Professor in the Computer Science Department at Stanford University. She leads the Social and Language Technologies (SALT) Lab, affiliated with the Stanford NLP Group, Stanford HCI Group, Stanford AI Lab (SAIL), and Stanford Human-Centered Artificial Intelligence (HAI). Her research focuses on socially aware natural language processing, large language models (LLMs), and human-AI interaction, aiming to improve human-human and human-computer communication through socially grounded AI systems. Education: Ph.D. in Language Technologies Institute, Carnegie Mellon University (2013–2019) B.S. in ACM Honored Class, Shanghai Jiao Tong University (2009–2013) Research Interests: Dr. Yang’s work bridges computational social science and NLP. She explores how AI can understand social contexts in language use and develop systems that respect cultural norms, ethical standards, and human values. Her lab’s projects include AI companions for skill training (e.g., Rehearsal Dialects), norm-aware LLMs (NormBank), and frameworks for human-AI collaboration (Co-Gym). Recent efforts address bias in AI, societal impacts of LLMs, and ethical evaluation of human-AI systems. Awards & Honors: 2024: Sloan Research Fellowship, ONR Young Investigator Award 2023: Adamic-Glance Young Distinguished Award (ICWSM), Kavli Fellow (NAS) 2022: NSF CAREER Award, Microsoft Research Faculty Fellow 2020: IEEE AI’s 10 to Watch Advising & Grants: Dr. Yang advises over 10 PhD students and postdocs, co-leading projects on LLM evaluation (SWE-bench/SWE-smith), human-AI ethics, and culturally aware NLP. Her research is supported by NSF, Amazon, DARPA, Google, and Stanford’s HAI initiative. Labs & Teams: The SALT Lab collaborates across disciplines, with projects spanning computer science, linguistics, and social sciences. Current initiatives include developing AI tools for mental health support (AI Partner & Mentor) and auditing societal impacts of LLMs.
Benjamin Van Roy is a Professor at Stanford University since 1998, affiliated with the Departments of Electrical Engineering and Management Science and Engineering, and the Institute for Computational and Mathematical Engineering. He leads the Efficient Agent Team at Google DeepMind and previously held leadership roles at Morgan Stanley, Unica, and Enuvis. He holds SB, SM, and PhD degrees in Computer Science and Electrical Engineering from MIT, advised by John Tsitsiklis. His research focuses on reinforcement learning, alignment, and information theory, with contributions to machine learning foundations, optimization, and finance. He has authored over 150 publications, including influential works on Thompson Sampling, approximate dynamic programming, and exploration strategies. His honors include INFORMS and IEEE Fellowships and the INFORMS Lanchester Prize. Van Roy advises doctoral students across academia and industry, with graduates at top institutions and companies like Meta, Tesla, and Citadel. He teaches courses on reinforcement learning, stochastic control, and optimization. His open-source projects include Epistemic Neural Networks and the Neural Testbed for evaluating machine learning models. Key contributions include foundational work in reinforcement learning theory, scalable methods for recommendation systems, and applications in finance and resource allocation. His research bridges theoretical insights with practical applications, emphasizing alignment and safety of AI systems.
Minh Q. Phan is an Associate Professor of Engineering at Dartmouth College's Thayer School of Engineering. His expertise spans system identification, iterative learning control, model predictive control, robotic swarm control, and intelligent control systems. He holds a BS from the University of California, Berkeley, and MS/M.Phil/PhD degrees from Columbia University in Mechanical Engineering. Dr. Phan has contributed to over 50 peer-reviewed publications and serves as an Associate Editor for the Journal of Guidance, Control, and Dynamics. His research focuses on advancing control theory applications in robotics, structural health monitoring, and sustainable construction materials. Key contributions include the development of OKID (Observer/Kalman Filter Identification) methods and bilinear system identification frameworks. Education History: Bachelor of Science in Mechanical Engineering, UC Berkeley, 1985 Master of Science in Mechanical Engineering, Columbia University, 1986 Master of Philosophy in Mechanical Engineering, Columbia University, 1988 Doctor of Philosophy in Mechanical Engineering, Columbia University, 1989 Research Interests: Advanced control methodologies for dynamic systems Model-based predictive control strategies Applications in robotics and aerospace engineering Structural health monitoring via system identification Machine learning for materials science Teaching Responsibilities include courses like ENGG 149 (Systems Identification), ENGS 145 (Modern Control Theory), and ENGG 148 (Structural Mechanics). His work bridges theoretical control systems with practical industrial applications, including automation in food processing and sustainable construction practices. Dr. Phan has collaborated on projects addressing viral epidemiology in Vietnam and coastal erosion mitigation strategies.
Chua Tat Seng is a Professor at the School of Computing, National University of Singapore (NUS), holding the KITHCT Chair Professorship since 2009. He serves as co-Director of the NExT++ Center, a joint research center between NUS and Tsinghua University focused on Extreme Search. His academic career spans over three decades at NUS, where he has held various leadership positions including Acting Dean of the School of Computing (1998-2000) and Acting Head of the Department of Information Systems & Computer Science (1996-1998). Professor Chua's research spans unstructured data analytics , multimedia information retrieval , recommendation and conversation systems , and emerging applications in e-commerce and fintech . He established the Lab for Media Search (LMS) at the School of Computing and has been instrumental in advancing multimodal learning and search technologies. His work bridges theoretical foundations with practical applications, particularly in developing trustable AI systems for real-world deployment. His recent publications demonstrate a strong focus on large language models for recommendation systems , multimodal learning , and generative AI applications . The research trends show increasing emphasis on LLM-based recommendation, multimodal understanding, and addressing fundamental challenges in AI reliability, fairness, and efficiency. His work spans theoretical advancements in representation learning to practical applications in e-commerce, finance, and healthcare domains. ACM SIGMM Technical Achievement Award 2015 Multiple Best Paper Awards across ACM Multimedia, IEEE Transactions, and MMM conferences (2007-2020) Professor Chua has supervised 37 PhD students since 2004, establishing himself as a dedicated mentor in the academic community. His research has been supported by substantial grants including NExT++ ($12 million), Base Metals Price Forecasting ($200,000), and Multilingual Multimodal Knowledge Graph ($500,000). He maintains active collaborations with Tsinghua University, University of Southampton, and industry partners like Four Elements Capital and Singapore Press Holdings. As co-Director of the NExT++ Center, he leads a major research initiative focused on Web Intelligence and User Empowerment. His visiting professorships at Tsinghua University (2017-present) and Zhejiang University (2021-present) reflect his international impact in the field of multimedia and AI research.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Yuki M. Asano is a full Professor at the University of Technology Nuremberg , leading the Fundamental AI (FunAI) Lab . Previously, he led the QUVA Lab at the University of Amsterdam and earned his PhD at the Visual Geometry Group (VGG) of the University of Oxford under Andrea Vedaldi and Christian Rupprecht. University of Technology Nuremberg (2024–present) University of Amsterdam (prior to 2024) University of Oxford (PhD, 2020) His research spans Artificial Intelligence , Machine Learning , and Computer Vision , with a focus on Causal Representation Learning , Self-Supervised Learning , and Efficient Model Adaptation . He pioneered techniques like BISCUIT (causal variable identification) and VeRA (parameter-efficient fine-tuning). His work extends to Medical Imaging and Environmental Monitoring through applications in fetal ultrasound analysis and marine debris detection. Recent publications (2023–2025) highlight advancements in Self-Supervised Learning , Vision-Language Models , and 3D Understanding . Notable papers include TWIST & SCOUT (multimodal LLM grounding), SIGMA (masked video modeling), and GeneralAD (anomaly detection). His ICCV 2023 work on Self-Ordering Point Clouds and MoSiC (optimal-transport motion trajectories) underscores his interdisciplinary approach. He received the JUPITER compute grant (2025) and an Outstanding Paper Award at ICLR 2024 . His collaborations span institutions like MIT-IBM Watson AI Lab, Qualcomm AI Research, and University of Amsterdam.
Kristen Grauman is a Full Professor in the Department of Computer Science at the University of Texas at Austin, where she leads the UT Computer Vision Group. Her research focuses on computer vision and machine learning, with applications in visual recognition, video analysis, and multi-modal perception. She received her B.A. from Boston College and her Ph.D. from MIT. Her research interests span visual recognition, image and video search, video analysis, first-person vision, embodied and multi-modal perception, and interactive machine learning. She has made significant contributions to the field, particularly in developing algorithms for understanding visual content and human activities from video, including foundational work on the Pyramid Match Kernel and relative attributes. Her recent publications reveal a strong emphasis on egocentric (first-person) vision, audio-visual learning, and view-invariant representations. There is a clear trajectory toward multi-modal integration (vision, audio, language) and real-world applications in instructional videos, human activity understanding, and embodied AI systems. She has received numerous awards including: AAAI Fellow (2019) J. K. Aggarwal Prize, International Association for Pattern Recognition (2018) Helmholtz Prize (2017) UT Austin Academy of Distinguished Teachers (2017) Best Paper Award, Asian Conference on Computer Vision (2016) Presidential Early Career Award for Scientists and Engineers (2014) Computers and Thought Award, International Joint Conferences on Artificial Intelligence (2013) Pattern Analysis and Machine Intelligence Young Researcher Award (2013) Alfred P. Sloan Research Fellow (2012) Marr Prize (2011) Prof. Grauman serves as Associate Editor-in-Chief for the IEEE Transactions on Pattern Analysis and Machine Intelligence. She has secured substantial research funding including the Presidential Early Career Award, NSF grants, and industry partnerships. Her advising has produced numerous influential publications and students who are now leaders in computer vision. She leads the UT Computer Vision Group, which collaborates closely with the Electrical and Computer Engineering Department. The group is pioneering large-scale egocentric video research through projects like Ego4D and Ego-Exo4D, focusing on real-world applications in human activity understanding, audio-visual perception, and interactive systems.