Dr. Chang Xu is an Associate Professor in Machine Learning and Computer Vision at the University of Sydney's School of Computer Science. He holds a Bachelor of Engineering from Tianjin University and a PhD from Peking University. His research focuses on machine learning, data mining, and their applications in AI and computer vision, including multi-view learning, visual search, and face recognition. He is an ARC Future Fellow and a member of the Sydney Southeast Asia Centre and The Net Zero Institute. Education: B.E. in Engineering (Tianjin University), Ph.D. in Computer Science (Peking University). His research interests emphasize handling heterogeneous data, exploring data variety, and developing algorithms for robust AI systems. His work includes adversarial robustness, neural architecture search, and efficient deep learning models. Research trends in his articles include adversarial robustness in neural architectures, efficient vision transformers, multimodal 3D style transfer, and underwater image restoration. Key contributions span image restoration, video super-resolution, and lightweight network design. He has advised multiple PhD and master's students on topics like diffusion models, radar image synthesis, and graph similarity. Awards: ARC Future Fellow. Collaborations focus on cross-domain data integration and AI applications. His labs and teams explore generative models, robust learning, and scalable robotics policies. Recent work includes diffusion models for action segmentation and robust vision-language systems.
Rhenish Friedrich Wilhelm University of BonnGermany
Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Russell Epstein is a Professor and Director of Graduate Studies in the Department of Psychology at the University of Pennsylvania. He is affiliated with the Center for Cognitive Neuroscience and Goddard Labs. His research focuses on neural mechanisms underlying visual scene perception, spatial navigation, and memory. Epstein holds a BA in Physics from the University of Chicago and a PhD in Applied Mathematics from Harvard University. Epstein’s research interests include high-level vision, spatial cognition, and the neural basis of environmental representations. His lab uses functional MRI and cognitive neuroscience techniques to study how scenes, objects, landmarks, and spaces are encoded in brain systems such as the parahippocampal place area and retrosplenial cortex. Recent work explores cognitive maps, grid-like neural representations, and the role of multisensory cues in navigation. His articles emphasize spatial navigation strategies, hierarchical cognitive maps, and the interplay between perception and memory. Notable contributions include investigations into hippocampal spatial metrics, olfactory navigation, and the neural underpinnings of environmental learning. Epstein teaches courses on cognitive neuroscience, including PSYC 149 and PSYC 600. He advises two graduate students in Psychology and has no listed scientific awards. His work is supported by grants (unspecified) and conducted within collaborative teams at the Center for Cognitive Neuroscience. Epstein’s research extends to labs focused on spatial cognition and neuroimaging, advancing understanding of how humans mentally map environments through visual and sensory integration.
Kristen Grauman is a Full Professor in the Department of Computer Science at the University of Texas at Austin, where she leads the UT Computer Vision Group. Her research focuses on computer vision and machine learning, with applications in visual recognition, video analysis, and multi-modal perception. She received her B.A. from Boston College and her Ph.D. from MIT. Her research interests span visual recognition, image and video search, video analysis, first-person vision, embodied and multi-modal perception, and interactive machine learning. She has made significant contributions to the field, particularly in developing algorithms for understanding visual content and human activities from video, including foundational work on the Pyramid Match Kernel and relative attributes. Her recent publications reveal a strong emphasis on egocentric (first-person) vision, audio-visual learning, and view-invariant representations. There is a clear trajectory toward multi-modal integration (vision, audio, language) and real-world applications in instructional videos, human activity understanding, and embodied AI systems. She has received numerous awards including: AAAI Fellow (2019) J. K. Aggarwal Prize, International Association for Pattern Recognition (2018) Helmholtz Prize (2017) UT Austin Academy of Distinguished Teachers (2017) Best Paper Award, Asian Conference on Computer Vision (2016) Presidential Early Career Award for Scientists and Engineers (2014) Computers and Thought Award, International Joint Conferences on Artificial Intelligence (2013) Pattern Analysis and Machine Intelligence Young Researcher Award (2013) Alfred P. Sloan Research Fellow (2012) Marr Prize (2011) Prof. Grauman serves as Associate Editor-in-Chief for the IEEE Transactions on Pattern Analysis and Machine Intelligence. She has secured substantial research funding including the Presidential Early Career Award, NSF grants, and industry partnerships. Her advising has produced numerous influential publications and students who are now leaders in computer vision. She leads the UT Computer Vision Group, which collaborates closely with the Electrical and Computer Engineering Department. The group is pioneering large-scale egocentric video research through projects like Ego4D and Ego-Exo4D, focusing on real-world applications in human activity understanding, audio-visual perception, and interactive systems.
Aarti Singh is a Professor in the Machine Learning Department at Carnegie Mellon University and Director of the NSF AI Institute for Societal Decision Making. She leads research at the intersection of machine learning, statistics, and decision making, with applications to scientific and societal domains. Her work focuses on designing principled interactive algorithms for learning and decision making under uncertainty. Education: Ph.D. in Electrical Engineering, University of Wisconsin-Madison (2008) M.S. in Electrical Engineering, University of Wisconsin-Madison (2003) B.E. in Electronics and Communication Engineering, University of Delhi (2001) Research Interests: Professor Singh's research centers on developing interactive machine learning algorithms that go beyond finding input-output associations to make higher-level decisions about the most informative data and actions. Her work spans autonomous decision making, including active sampling, stochastic optimization, bandits, and reinforcement learning that are statistically optimal, computationally tractable, and robust. She also investigates human factors in decision making, designing algorithms that model and leverage human feedback while accounting for bias, memory effects, and calibration. Her research has applications in material science, cosmology, and peer review systems. Research Trends: Professor Singh's recent publications demonstrate a strong focus on reinforcement learning, particularly in developing more efficient and robust algorithms for decision making under uncertainty. Her work bridges theoretical foundations with practical applications, spanning from fundamental algorithm development to real-world implementation in scientific domains. There's a clear trajectory toward integrating human factors into decision-making algorithms, with significant contributions to peer review systems and preference learning. Scientific Awards: NSF Career Award United States Air Force Young Investigator Award A. Nico Habermann Faculty Chair Award Harold A. Peterson Best Dissertation Award Multiple paper awards Advising and Grants: Professor Singh has advised numerous PhD and master's students, many of whom have gone on to faculty positions or research roles at leading institutions. Her research is supported by prestigious grants from ONR, Simons Foundation, AFRL, ARL, and NSF. She serves as General Chair (2025) and Program Chair (2020) for the International Conference on Machine Learning (ICML) and has held leadership roles in multiple professional organizations. Research Team: Professor Singh leads a vibrant research group within the Machine Learning Department at CMU, with current PhD students working on topics including reinforcement learning, human-AI collaboration, and decision making under uncertainty. She also directs the NSF AI Institute for Societal Decision Making, which brings together researchers from multiple disciplines to develop AI systems that support human decision making in societal contexts.
Indian Institute of Technology Hyderabad (IITH)India
Vineeth N Balasubramanian is a Professor in the Department of Computer Science & Engineering at the Indian Institute of Technology Hyderabad, with affiliate faculty status in the Department of Artificial Intelligence. His research focuses on the intersection of deep learning, machine learning, and computer vision, emphasizing explainability, robustness, and real-world applications. He leads Lab 1055, which investigates problems such as Explainable and robust AI/ML systems Lifelong learning in evolving environments Multimodal vision-language models Applications in agriculture, autonomous navigation, and human behavior analysis His recent work includes causal reasoning in transformers, vision-language model capabilities, and drone-based object detection. Funded by organizations like Google, Microsoft, Intel, and DST, he has received multiple awards including the World's Top 2% Scientists (2022-23), INSA/INAE Fellowships, and Best Paper recognitions. Lab 1055 collaborates with institutions like CMU, UBC, and Monash University, contributing to cutting-edge advancements in AI.
Adriana Kovashka is an Associate Professor at the University of Pittsburgh , affiliated with the School of Computing and Information and serving as Department Chair . Her academic journey began with BA degrees in Computer Science and Media Studies from Pomona College (2008) and a PhD in Computer Science from The University of Texas at Austin (2014). Joined Pitt’s faculty in January 2015 NSF CAREER awardee (2021) Google Faculty Research Award recipient Dr. Kovashka’s research spans Computer Vision , Machine Learning , and Natural Language Processing , focusing on visual rhetoric, weak multimodal supervision, and domain adaptation. She pioneered techniques for analyzing political imagery, developing robust object detection frameworks, and exploring the intersection of visual and textual persuasion through large-scale annotated datasets. Her recent work emphasizes geographic diversity in vision-language systems, audio-visual fusion for domain generalization, and shape-texture bias mitigation in CNNs. Key publications include groundbreaking studies on symbolic reasoning, multimodal dialogue systems, and ethical AI applications in education. Scientific honors include: NSF CRII Award (2016) NSF CAREER Award (2021) Pitt CRDF Award (2016, 2018) Best Paper at ECV Workshop (2021) Google Faculty Research Award (2016, 2018) Dr. Kovashka actively mentors students in multimodal learning projects and collaborates with interdisciplinary teams on NSF-funded initiatives. She co-organizes workshops like the first CVPR workshop on advertisement understanding and leads research groups exploring human-AI co-learning systems.
Prof. Stefan Leutenegger is a tenure-track Assistant Professor at Technische Universität München (TUM), leading the Machine Learning for Robotics group within the TUM School of Computation, Information, and Technology. Previously, he held roles as Senior Lecturer (2018–2021) and Lecturer (2014–2018) at Imperial College London's Dyson Robotics Lab, where he founded the Smart Robotics Lab. He earned his PhD (2014) from ETH Zurich, focusing on autonomous solar-powered aircraft navigation, and holds BSc (2006) and MSc (2009) in Mechanical Engineering from ETH Zurich. His research centers on mobile robotics, particularly enabling robots (e.g., drones) to perceive and navigate complex environments using machine learning and sensor data fusion. Key focus areas include SLAM, event-based vision, 3D reconstruction, and autonomous exploration. He has pioneered algorithms like BRISK (2011), OKVIS (2014), and ElasticFusion (2016), advancing real-time robotics perception. Notable Awards: Imperial College President's Award (2018), Best ECCV Paper (2016), ETH Medal for Dissertations (2015). Labs: TUM's Machine Learning for Robotics Group, Imperial's Smart Robotics Lab. Publications: Over 100 papers, including seminal works in CVPR, ECCV, and Robotics: Science and Systems. Current projects include DigiForests (forest inventory via robotics), aerial additive manufacturing, and object-centric semantic mapping. His work bridges theory and practice, with applications in autonomous drones, construction robotics, and human-robot interaction.
Swiss Federal Institute of Technology in LausanneSwitzerland
Mathieu Salzmann is a Senior Scientist and Lecturer at École Polytechnique Fédérale de Lausanne (EPFL), affiliated with the Computer Vision Laboratory (CVLAB) in the School of Computer and Communication Sciences (IC). He also holds a courtesy appointment with the EPFL College of Humanities and serves as Deputy Chief Data Scientist at the Swiss Data Science Center (SDSC). He has held concurrent roles in teaching units including SIN, SODH, and SSC, reflecting his interdisciplinary engagement. His research focuses on the intersection of machine learning and computer vision, particularly in deep learning for 2D and 3D visual scene understanding, efficient and robust models, domain adaptation, and interpretable AI. These interests are evident across his extensive publication record in top-tier venues. His recent publications (2023–2024) show a consistent trend in advancing deep learning methods for visual recognition, with strong representation at CVPR, ICCV, ECCV, ICML, ICLR, and NeurIPS. Topics include domain generalization, 3D understanding, model robustness, and multimodal learning, often with applications in real-world systems. His editorial roles as Associate Editor for IEEE TPAMI and Action Editor for TMLR further highlight his leadership in the field. Area Chair: ICML 2023, CVPR 2023, ICCV 2023, NeurIPS 2023, AAAI 2024, ECCV 2024 Associate Editor: IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) Action Editor: Transactions on Machine Learning Research (TMLR) Mathieu Salzmann has supervised numerous PhD students at EPFL, both current and past, including Bouquet Yann Yanis, Javed Saqib, Li Shuangqi, and others. He has also been involved in research grants and collaborative projects, such as his work with S. Süsstrunk and R. Baroni on comics reconfiguration. His part-time role as Senior GNC Engineer at ClearSpace (2020–2024) illustrates his applied research engagement in aerospace systems. He is actively involved in EPFL’s data science and AI research ecosystem through SDSC and multiple labs.
Andrew Zisserman is a Royal Society Research Professor at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). His research focuses on computer vision, artificial intelligence, and neural networks, with significant contributions to multimodal learning, video understanding, and 3D scene analysis. He leads projects exploring visual-language models, audio-visual synchronization, and clinical imaging applications. Key research areas include: Video analysis and temporal modeling Multimodal systems for sign language translation and action recognition 3D shape estimation and physical property inference Foundation models and cross-modal retrieval Recent work highlights: Developed Flamingo and Tapir models for video-language tasks Advancements in spinal MRI analysis and clinical imaging Leadership in EGO4D and VoxCeleb challenges Honors include Fellowship of the Royal Society (FRS) and the ISSLS Prize in Clinical Science 2023 for spinal analysis innovations. His lab collaborates globally, emphasizing real-world applications in healthcare and autonomous systems.
Max Planck Institute for Intelligent SystemsGermany
Celestine Mendler-Dünner is a Principal Investigator at the ELLIS Institute in Tübingen, co-affiliated with the Max Planck Institute for Intelligent Systems and the Tübingen AI Center. She leads the Algorithms and Society research group, focusing on machine learning in social contexts and the role of prediction in digital economies. Her work bridges theoretical machine learning with practical societal impact, developing tools for safe, reliable, and equitable AI ecosystems. Her educational background includes a PhD from ETH Zurich in collaboration with IBM Research, followed by an SNSF postdoctoral fellowship at UC Berkeley hosted by Moritz Hardt. She was previously a group leader at the Max Planck Institute for Intelligent Systems before joining the ELLIS Institute. Mendler-Dünner's research spans several interconnected themes including performative prediction (where predictions change the behavior they aim to predict), algorithmic collective action (how participants can steer AI systems toward common goals), and the role of LLMs in social science research. Her work combines theoretical foundations with practical implementations, addressing challenges in interactive machine learning, optimization in dynamic environments, and context-specific evaluation of AI systems. She particularly examines how algorithmic predictions mediate services and platforms at societal scale, exploring concepts of economic power in digital markets. Her publication record shows a clear evolution from system-aware machine learning algorithms (including foundational work on IBM Snap ML) toward increasingly sociotechnical questions at the intersection of machine learning, economics, and policy. Recent work focuses on measuring performative power in digital economies, evaluating LLMs as risk scores, and developing frameworks for algorithmic collective action in recommender systems and labor markets. Among her notable recognitions are the ETH Medal for her dissertation, the IBM Research Division Award, the Fritz Kutter Award, and the IBM Eminence and Excellence Award. She is an ELLIS Scholar, a fellow of the Elisabeth-Schiemann-Kolleg, and affiliated with several prestigious research programs including the International Max Planck Research School for Intelligent Systems and the Max Planck ETH Center for Learning Systems. ETH Medal (dissertation award) IBM Research Division Award Fritz Kutter Award IBM Eminence and Excellence Award SNSF Early Postdoc Mobility Fellowship Mendler-Dünner actively mentors the next generation of researchers, advising PhD student Patrik Wolf and supervising research interns including Joachim Baumann, Haiqing Zhu, and Anna Badalyan, as well as Master's student Dorothee Sigg. She serves as core faculty for the International Max Planck Research School and associated faculty for the Max Planck ETH Center for Learning Systems. Her group has secured significant research funding through fellowships and institutional support, enabling work on projects like Powermeter (measuring search engine influence) and Snap ML (resource-efficient machine learning library with over 1 million PyPI downloads). She leads the Algorithms and Society research group, which examines machine learning as part of broader sociotechnical ecosystems. The group explores human-population interactions with algorithmic systems and incorporates these insights into learning system fundamentals. Current projects include investigating economic incentives in digital platforms, developing tools for systematic LLM evaluation in social science contexts, and creating frameworks for collective action in algorithmic systems. Mendler-Dünner also co-organizes the Algorithmic Collective Action workshop at NeurIPS 2025, demonstrating her leadership in emerging research directions at the AI-society interface.
Ajmal Mian is a Professor of Computer Science at the University of Western Australia (UWA), affiliated with the School of Physics, Maths and Computing. He holds an Australian Research Council Future Fellowship (2022) and leads research in Artificial Intelligence, Computer Vision, and Machine Learning. His work focuses on 3D computer vision, adversarial AI defense, and explainable AI. His research interests include 3D point cloud analysis, face recognition, human action recognition, and remote sensing. He has published over 300 papers and secured major grants from ARC, NHMRC, and DARPA, totaling millions in funding. He has supervised 29 PhD students and mentored 12 postdoctoral researchers. Key projects include 3D diffusion models for scene generation, robust 3D vision systems, and defense against AI deception attacks. He serves as a fellow of IAPR, an ACM Distinguished Speaker, and has editorial roles at IEEE Transactions on Neural Networks and Pattern Recognition. Research Awards: HBF Mid-Career Scientist of the Year, West Australian Early Career Scientist of the Year, IAPR Best Scientific Paper Award. Grants: ARC Discovery Projects, National Intelligence & Security Discovery grants, DARPA grants for AI security. His teaching spans computer vision, machine learning, and programming courses. Collaborations include defense, medical, and agricultural applications.
Almut Sophia Koepke is a junior research group leader at the Technical University of Munich and University of Tübingen, focusing on multimodal learning problems integrating sound, vision, and text. Her work bridges foundational research in audio-visual understanding with practical applications in few-shot learning, zero-shot translation, and cross-modal attention mechanisms.
Timothy M. Hospedales is a Professor of Artificial Intelligence at the Institute of Perception, Action and Behaviour within the School of Informatics at the University of Edinburgh . He also serves as VP AI and Head of Samsung AI Research Centre Europe . His research focuses on efficient and robust AI , emphasizing meta-learning , lifelong transfer-learning , and domain adaptation in both probabilistic and deep learning frameworks. Applications span computer vision , vision and language , reinforcement learning for robotics , and finance . Professor at University of Edinburgh (2020–present) ELLIS Fellow (2021) Head of Samsung AI Research Europe (2020–present) Founding Director of Applied Machine Learning Lab at QMUL (2012–2016) His work includes pioneering contributions to meta-learning , few-shot learning , and self-supervised methods , with notable awards such as the Best Paper Prize at ICML AutoML 2018 and Best Student Paper at ICPR 2018 . He has co-authored 15+ recent papers on topics like Vision-Language Models , Medical AI Fairness , and Diffusion Model Optimization . He served as Program Co-Chair for BMVC 2018 and AAAI 2022 , and authored a book on Visual Adaptation in the Deep Learning Era (2022). Co-Chair, BMVC 2018 Guest Editor, IET CV Special Issue (2016) Keynote Speaker at TASK-CV Workshop (ECCV 2016) Special Issue on Fewer Labels (IEEE PAMI 2020) His leadership extends to organizing workshops like the Learning-to-Learn Workshop at ICLR 2021 , Meta-Learning Workshop at NeurIPS 2020 , and Domain Generalisation Workshop at ICLR 2023 . Current projects include Meta-Omnium (CVPR 2023) for general-purpose meta-learning and MetaAudio (ICANN 2022) for few-shot audio classification benchmarks.
Barbara Caputo is a Full Professor at Politecnico di Torino, leading the VANDAL Laboratory and directing the AI@PoliTo Interdepartmental Lab. She holds a double affiliation with the Italian Institute of Technology (IIT) and has held roles at Idiap-EPFL and Sapienza University. Her research focuses on AI, computer vision, domain adaptation, and federated learning. She contributes to national AI policy, including the Italian Strategy on AI and the National PhD on AI for Industry 4.0. She is an ERC Laureate and ELLIS Fellow, co-founding ELLIS society. Her work spans visual place recognition, action recognition, and cross-domain learning. Education: PhD in Computer Science from KTH Royal Institute of Technology (2005). Major roles include Rector’s Advisor on AI at PoliTo, Board Member of ELLIS, and coordinator of the AI & Industry 4.0 vertical in the National PhD program. Awards include ERC Laureate (2017), ELLIS Fellow (2019), and Inspiring Fifty Italy (2018). Her research emphasizes federated learning, domain adaptation, and AI ethics. Recent articles explore domain generalization, resource-efficient federated models, and AI-environment interactions. She collaborates with institutions like MUR, CNR, and the European Commission on AI policy and tech initiatives.