Andrea Vedaldi is a Professor of Computer Vision and Machine Learning at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). He specializes in unsupervised methods for understanding images and videos, focusing on 3D geometry and semantics. His research bridges foundational AI and practical applications, with contributions to generative models, neural fields, and self-supervised learning. Education: PhD in Computer Science (2008), University of California, Los Angeles MSc in Computer Science (2005), UCLA BSc in Information Engineering (2003), University of Padua Research Interests: Unsupervised learning, 3D perception, generative AI, neural rendering, and scalable vision systems. His work emphasizes ethical, responsible AI aligned with ERC-funded projects like UNION (ERC Consolidator Grant). Key Contributions: Co-developer of VLFeat and MatConvNet libraries Leader in 3D reconstruction and diffusion models (e.g., CatFree3D) Recipient of the PAMI Thomas S. Huang Prize and multiple best paper awards Grants & Service: Principal Investigator on £2.3M ERC Consolidator Grant (UNION) Co-organizer of major conferences (ECCV 2020 Program Chair, CVPR 2023 Area Chair) Reviewer for top journals/conferences (PAMI, CVPR, NeurIPS) Labs & Teams: VGG Group at Oxford, collaborating on projects like Meta 3D Gen and Common Objects in 3D (CO3D).
Dr. Chang Xu is an Associate Professor in Machine Learning and Computer Vision at the University of Sydney's School of Computer Science. He holds a Bachelor of Engineering from Tianjin University and a PhD from Peking University. His research focuses on machine learning, data mining, and their applications in AI and computer vision, including multi-view learning, visual search, and face recognition. He is an ARC Future Fellow and a member of the Sydney Southeast Asia Centre and The Net Zero Institute. Education: B.E. in Engineering (Tianjin University), Ph.D. in Computer Science (Peking University). His research interests emphasize handling heterogeneous data, exploring data variety, and developing algorithms for robust AI systems. His work includes adversarial robustness, neural architecture search, and efficient deep learning models. Research trends in his articles include adversarial robustness in neural architectures, efficient vision transformers, multimodal 3D style transfer, and underwater image restoration. Key contributions span image restoration, video super-resolution, and lightweight network design. He has advised multiple PhD and master's students on topics like diffusion models, radar image synthesis, and graph similarity. Awards: ARC Future Fellow. Collaborations focus on cross-domain data integration and AI applications. His labs and teams explore generative models, robust learning, and scalable robotics policies. Recent work includes diffusion models for action segmentation and robust vision-language systems.
Jean Oh is a Researcher at the Robotics Institute of Carnegie Mellon University (CMU) , leading the interdisciplinary Bot Intelligence Group (BIG) . Her work focuses on developing persistent robots that co-exist and collaborate with humans in shared environments, emphasizing continuous improvement through training, exploration, and human interaction. Education: Ph.D. in Language and Information Technologies, CMU M.S. in Computer Science, Columbia University B.S. in Biotechnology, Yonsei University Oh's research integrates vision, language, and planning systems in robotics, with applications in human-robot teaming , self-driving cars , disaster response , eldercare , and creative robotics . She has pioneered projects like socially-compliant robot navigation in human crowds and AI-driven robotic painting systems. Recent publication trends highlight her work in vision-language planning , social navigation , computational creativity , and human-robot collaboration . Notable contributions include the StyleCLIPDraw algorithm for text-to-art generation and Social-PatteRNN for human-like trajectory prediction. Scientific Awards: Best Paper Award in Cognitive Robotics (ICRA'18, ICRA'15) Best Systems Paper Finalist (HRI'25) Best Oral Paper Finalist (Humanoids'24) Best Paper in Entertainment (IROS'24) Argoverse Challenge Winner (CVPR'24) Best Student Paper (AIAA'24) Best Demo Finalist (RoboSoft'24) Oh mentors a diverse team of PhD, MS, and undergraduate students from CMU departments including Robotics, Computer Science, and Mechanical Engineering. Her research is funded by US Army Research Lab , DiDi Chuxing , and DARPA , with collaborations across industry and academia .
Stefano Ermon is an Associate Professor in the Department of Computer Science at Stanford University, affiliated with the Artificial Intelligence Laboratory and a Senior Fellow at the Woods Institute for the Environment. His research focuses on advancing machine learning and generative AI techniques to address societal and environmental challenges, including computational sustainability, geospatial analysis, and climate science. He holds a Ph.D. from Cornell University (2015). Education: Ph.D. in Computer Science, Cornell University (2015). Research Interests: Ermon’s work bridges foundational machine learning (e.g., diffusion models, generative AI, and optimization) with applications in sustainability, geospatial analysis (via satellite imagery), and earth observation systems. Notable contributions include predicting poverty using satellite data and developing scalable methods for molecule generation. Articles Trends: His recent work emphasizes diffusion models for generative tasks (e.g., text-to-image, molecule design), geospatial AI (e.g., environmental monitoring), and ethical AI (e.g., bias mitigation in LLMs). He also explores applications in robotics and scientific computing. Awards: He has received prestigious awards, including the ICML 2024 Best Paper Award, Sloan Research Fellowship, Microsoft Research Faculty Fellowship, and the IJCAI Computers and Thought Award. Advising and Grants: Ermon teaches courses like Probabilistic Graphical Models (CS228) and has secured grants from NSF, ONR, AFOSR, and private foundations. His lab develops tools for climate science and sustainable development. Labs/Teams: Leads the Stanford AI Lab group focused on computational sustainability and generative AI, collaborating with institutions like the Woods Institute for environmental applications.
Rhenish Friedrich Wilhelm University of BonnGermany
Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Prof. Dr. Thomas Hofmann is a Full Professor and Head of the Department of Computer Science at ETH Zurich since 2014. He also leads the Institute for Machine Learning. His research focuses on machine learning, deep learning, natural language understanding, and text understanding. Hofmann holds a Ph.D. from the University of Bonn (1997) and has held academic positions at Brown University (1999–2004) and TU Darmstadt. He transitioned to industry as Director of Engineering at Google (2006–2014), leading the Zurich R&D center, before returning to academia. He co-founded Recommind (2000) and 1plusX (Swiss marketing tech company), currently serving as Chief Scientist and board member at 1plusX. Education: Ph.D. in Computer Science, University of Bonn (1997) Postdoctoral Work: MIT (CBCL/AI Lab), UC Berkeley (EECS/ICSI) His research explores advanced machine learning techniques, including diffusion models, generative adversarial networks, and optimization dynamics. Hofmann’s entrepreneurial ventures reflect his focus on applying AI to real-world challenges, such as e-discovery and marketing technology. His work spans theoretical contributions (e.g., neural network training dynamics, continual learning) and applied innovations (e.g., image editing, portrait generation). Hofmann actively bridges academia and industry, influencing both research and commercial AI applications.
Prof. Konrad Schindler holds the position of Full Professor at the Department of Civil, Environmental and Geomatic Engineering at ETH Zürich. He is also the Head of the Institute of Geodesy and Photogrammetry (IGP), leading research and educational activities in geomatics and computer vision. His career spans roles as a Photogrammetric Engineer, scientific assistant, postdoc researcher, and academic faculty across institutions including Graz University of Technology, Monash University, and TU Darmstadt before joining ETH Zürich in 2010. Education: Undergraduate studies in Geodesy (1992–1995), Graz University of Technology, Austria MEng in Photogrammetry and Geoinformation (1995–1999), Vienna University of Technology, Austria PhD in Computer Science (2001–2003), Graz University of Technology, Austria Research focuses on Photogrammetry , Remote Sensing , Computer Vision , and Image Understanding with interdisciplinary applications in environmental monitoring, geospatial analysis, and disaster response. He develops computational methods for 3D reconstruction, fusion of multi-modal data, and AI-driven solutions for satellite imagery interpretation. His work bridges geomatic engineering and machine learning to address challenges in urban mapping, climate modeling, and biological systems analysis. Publications reflect expertise in geospatial AI, diffusion models, and benchmarking datasets for disaster resilience. Notable works include Marigold (image analysis adaptation) and BRIGHT (building damage assessment). His research emphasizes practicality and scalability, such as affordable depth estimation and global biomass datasets. He has received the 2013 Marr Prize Honourable Mention (IEEE) and the 2012 U.V. Helava Award (ISPRS), alongside several Best Presentation Awards. His contributions span technical leadership, editorial roles (ISPRS Journal), and service to Swiss remote sensing commissions. Advising and grants: While no specific advisee names or grant details are listed, his career trajectory includes mentoring postdocs and junior faculty. He teaches advanced courses in Photogrammetry , Image Interpretation , and Machine Vision , integrating cutting-edge AI techniques into curricula. His research group collaborates on global-scale projects like canopy height mapping and satellite-based climate variable assessments. Labs/Teams: As Institute Head, he oversees the IGP lab at ETH Zürich, with prior affiliations including the Digital Perception Lab (Monash University) and the Computer Vision Lab (ETH Zurich). His work often involves multi-institutional collaborations focused on geospatial AI and environmental science.
Chua Tat Seng is a Professor at the School of Computing, National University of Singapore (NUS), holding the KITHCT Chair Professorship since 2009. He serves as co-Director of the NExT++ Center, a joint research center between NUS and Tsinghua University focused on Extreme Search. His academic career spans over three decades at NUS, where he has held various leadership positions including Acting Dean of the School of Computing (1998-2000) and Acting Head of the Department of Information Systems & Computer Science (1996-1998). Professor Chua's research spans unstructured data analytics , multimedia information retrieval , recommendation and conversation systems , and emerging applications in e-commerce and fintech . He established the Lab for Media Search (LMS) at the School of Computing and has been instrumental in advancing multimodal learning and search technologies. His work bridges theoretical foundations with practical applications, particularly in developing trustable AI systems for real-world deployment. His recent publications demonstrate a strong focus on large language models for recommendation systems , multimodal learning , and generative AI applications . The research trends show increasing emphasis on LLM-based recommendation, multimodal understanding, and addressing fundamental challenges in AI reliability, fairness, and efficiency. His work spans theoretical advancements in representation learning to practical applications in e-commerce, finance, and healthcare domains. ACM SIGMM Technical Achievement Award 2015 Multiple Best Paper Awards across ACM Multimedia, IEEE Transactions, and MMM conferences (2007-2020) Professor Chua has supervised 37 PhD students since 2004, establishing himself as a dedicated mentor in the academic community. His research has been supported by substantial grants including NExT++ ($12 million), Base Metals Price Forecasting ($200,000), and Multilingual Multimodal Knowledge Graph ($500,000). He maintains active collaborations with Tsinghua University, University of Southampton, and industry partners like Four Elements Capital and Singapore Press Holdings. As co-Director of the NExT++ Center, he leads a major research initiative focused on Web Intelligence and User Empowerment. His visiting professorships at Tsinghua University (2017-present) and Zhejiang University (2021-present) reflect his international impact in the field of multimedia and AI research.
Yuki M. Asano is a full Professor at the University of Technology Nuremberg , leading the Fundamental AI (FunAI) Lab . Previously, he led the QUVA Lab at the University of Amsterdam and earned his PhD at the Visual Geometry Group (VGG) of the University of Oxford under Andrea Vedaldi and Christian Rupprecht. University of Technology Nuremberg (2024–present) University of Amsterdam (prior to 2024) University of Oxford (PhD, 2020) His research spans Artificial Intelligence , Machine Learning , and Computer Vision , with a focus on Causal Representation Learning , Self-Supervised Learning , and Efficient Model Adaptation . He pioneered techniques like BISCUIT (causal variable identification) and VeRA (parameter-efficient fine-tuning). His work extends to Medical Imaging and Environmental Monitoring through applications in fetal ultrasound analysis and marine debris detection. Recent publications (2023–2025) highlight advancements in Self-Supervised Learning , Vision-Language Models , and 3D Understanding . Notable papers include TWIST & SCOUT (multimodal LLM grounding), SIGMA (masked video modeling), and GeneralAD (anomaly detection). His ICCV 2023 work on Self-Ordering Point Clouds and MoSiC (optimal-transport motion trajectories) underscores his interdisciplinary approach. He received the JUPITER compute grant (2025) and an Outstanding Paper Award at ICLR 2024 . His collaborations span institutions like MIT-IBM Watson AI Lab, Qualcomm AI Research, and University of Amsterdam.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Kristen Grauman is a Full Professor in the Department of Computer Science at the University of Texas at Austin, where she leads the UT Computer Vision Group. Her research focuses on computer vision and machine learning, with applications in visual recognition, video analysis, and multi-modal perception. She received her B.A. from Boston College and her Ph.D. from MIT. Her research interests span visual recognition, image and video search, video analysis, first-person vision, embodied and multi-modal perception, and interactive machine learning. She has made significant contributions to the field, particularly in developing algorithms for understanding visual content and human activities from video, including foundational work on the Pyramid Match Kernel and relative attributes. Her recent publications reveal a strong emphasis on egocentric (first-person) vision, audio-visual learning, and view-invariant representations. There is a clear trajectory toward multi-modal integration (vision, audio, language) and real-world applications in instructional videos, human activity understanding, and embodied AI systems. She has received numerous awards including: AAAI Fellow (2019) J. K. Aggarwal Prize, International Association for Pattern Recognition (2018) Helmholtz Prize (2017) UT Austin Academy of Distinguished Teachers (2017) Best Paper Award, Asian Conference on Computer Vision (2016) Presidential Early Career Award for Scientists and Engineers (2014) Computers and Thought Award, International Joint Conferences on Artificial Intelligence (2013) Pattern Analysis and Machine Intelligence Young Researcher Award (2013) Alfred P. Sloan Research Fellow (2012) Marr Prize (2011) Prof. Grauman serves as Associate Editor-in-Chief for the IEEE Transactions on Pattern Analysis and Machine Intelligence. She has secured substantial research funding including the Presidential Early Career Award, NSF grants, and industry partnerships. Her advising has produced numerous influential publications and students who are now leaders in computer vision. She leads the UT Computer Vision Group, which collaborates closely with the Electrical and Computer Engineering Department. The group is pioneering large-scale egocentric video research through projects like Ego4D and Ego-Exo4D, focusing on real-world applications in human activity understanding, audio-visual perception, and interactive systems.
Prof. Luke Zettlemoyer is an Adjunct Professor of Computer Science and Engineering at the University of Washington, with affiliations to the Department of Linguistics. He focuses on machine learning, natural language processing, and multimodal systems, contributing to advancements in large language models, ethical AI, and scalable architectures. His research addresses challenges in model alignment, generalization, and cross-domain integration. Key research interests include multimodal reward models, efficient tokenization strategies, and model optimization techniques. He has explored topics such as neural trajectories for robot learning, content-adaptive image processing, and ethical mitigation of verbatim data reproduction. His publications span 2023–2025, emphasizing practical applications of AI in robotics, vision-language systems, and scalable retrieval-based models. While no formal awards are listed, his work reflects significant contributions to foundational AI research.
Indian Institute of Technology Hyderabad (IITH)India
Vineeth N Balasubramanian is a Professor in the Department of Computer Science & Engineering at the Indian Institute of Technology Hyderabad, with affiliate faculty status in the Department of Artificial Intelligence. His research focuses on the intersection of deep learning, machine learning, and computer vision, emphasizing explainability, robustness, and real-world applications. He leads Lab 1055, which investigates problems such as Explainable and robust AI/ML systems Lifelong learning in evolving environments Multimodal vision-language models Applications in agriculture, autonomous navigation, and human behavior analysis His recent work includes causal reasoning in transformers, vision-language model capabilities, and drone-based object detection. Funded by organizations like Google, Microsoft, Intel, and DST, he has received multiple awards including the World's Top 2% Scientists (2022-23), INSA/INAE Fellowships, and Best Paper recognitions. Lab 1055 collaborates with institutions like CMU, UBC, and Monash University, contributing to cutting-edge advancements in AI.
Geoffrey E. Hinton is a distinguished Professor in the Department of Computer Science at the University of Toronto. He is renowned for his foundational contributions to machine learning, particularly in the development of deep learning and neural networks. His research focuses on understanding learning processes in both artificial and biological systems, with key contributions including Boltzmann machines, backpropagation, and deep belief networks. He teaches advanced machine learning courses such as CSC2535, emphasizing topics like graphical models, variational inference, and deep learning architectures. His work has been published extensively in top journals and conferences, with recent papers exploring forward-forward algorithms, analog diffusion models, and scalable neural network training methods. Hinton has advised numerous PhD and master's students and collaborates with institutions like Vector Institute. He is a central figure in the global AI community, regularly presenting at conferences (e.g., 2023 talks on CBS 60 Minutes, BBC, and PBS). His lab focuses on advancing machine learning theory and applications, addressing challenges in vision, language, and generative models.
Dr. Daniel J. Hsu is a Professor of Computer Science at Columbia University, affiliated with the Foundations of Data Science Center and TRIPODS Institute. His research focuses on algorithmic statistics, machine learning theory, and their applications in public health informatics. He has advised numerous students and postdocs, and his work bridges foundational theory with practical systems like foodborne illness detection via social media analysis. Key roles: Associate Editor (ACM Transactions on Algorithms), Program Chair (ICML 2025, COLT 2019) Research areas: Foundations of Data Science, Machine Learning Theory, Fairness, and High-dimensional Statistics His work on detecting foodborne illnesses using Yelp reviews has been deployed by NYC Health departments. Recent contributions include advancements in transformer architectures, group fairness algorithms, and multi-group learning frameworks. Scientific awards include the Sloan Fellowship and multiple NSF grants. He has pioneered interactive machine teaching methods and developed algorithms for robust parameter estimation in high-dimensional settings.