Professor Hongdong Li is a Tenured Professor at the School of Computing, Australian National University (ANU), within the College of Engineering and Computer Science. His research focuses on 3D Computer Vision, Machine Learning, and their applications in dynamic environments. He has held visiting roles at Carnegie Mellon University and has contributed to significant projects like the Australia Bionic Eyes initiative. Education: PhD (Electrical Engineering). Research Interests : 3D Computer Vision fundamentals and applied AI systems Learning-based 3D perception for plant sciences Robot navigation in unfamiliar environments Awards : Marr Prize Honourable Mention CVPR Best Paper Award Advising & Grants : Supervised 40+ PhD students, with funding from ARC, CSIRO, Microsoft, and firms like OPPO/Tencent. Active in projects such as bushfire detection via video analytics and sign language translation systems. Labs/Teams : Co-founder of the Australian Centre for Robotic Vision (ACRV). Collaborates globally on cross-view localization and autonomous systems.
Andrea Tagliasacchi is an Associate Professor in the School of Computing Science at Simon Fraser University (SFU), where he holds the Visual Computing Research Chair. He is also a part-time (20%) Staff Research Scientist at Google DeepMind in Toronto and holds an associate professor (status only) appointment in the Department of Computer Science at the University of Toronto. Education: PhD in Computing Science – Simon Fraser University (NSERC Alexander Graham Bell Fellow) Postdoctoral Research – École Polytechnique Fédérale de Lausanne (EPFL) MSc in Computer Science – Politecnico di Milano (Gold Medalist) His research lies at the intersection of computer vision, computer graphics, and machine learning, with a focus on 3D visual perception. Key areas include neural radiance fields (NeRF), 3D Gaussian splatting, inverse rendering, and geometric deep learning, with applications in robotics, augmented reality, and autonomous systems. His work emphasizes robust and efficient scene understanding and reconstruction from visual data. His recent publications, appearing in top venues like CVPR, SIGGRAPH, NeurIPS, and ECCV, demonstrate a strong emphasis on neural fields, 3D reconstruction, and generative modeling. Trends include improving rendering efficiency, enhancing robustness to noise and distractors, and enabling controllable and 3D-aware generation. His group has made significant contributions to Gaussian splatting, NeRF optimization, and diffusion-based 3D/4D synthesis. Scientific Awards: 2024 CVPR Best Paper Award (Honorable Mention) 2020 CVPR Best Student Paper Award 2015 SGP Best Paper Award NSERC Alexander Graham Bell Canada Graduate Scholarship MITACS Best Paper Award (SIGGRAPH Asia 2009) NSF Best Poster Award (SGP 2012) He has advised numerous PhD and MSc students, many of whom are now researchers at leading institutions and companies. His research has been supported through collaborations with Google, Intel, and academic partners. He serves the community as a Senior Area Chair for CVPR 2025, Associate Editor for IEEE TPAMI (2024–2026), Guest Editor for IEEE TPAMI on 3D GenAI, and Program Chair for 3DV 2024. He leads a vibrant research lab at SFU focused on pushing the boundaries of 3D scene understanding with machine learning.
Michael A Osborne is Professor of Machine Learning at the University of Oxford and leads the Bayesian Exploration Lab . He serves as Director of the EPSRC Centre for Doctoral Training in Autonomous Intelligent Machines and Systems and co-directs the Oxford Martin AI Governance Initiative. His research focuses on Bayesian optimization, Gaussian processes, and probabilistic numerics with applications in quantum devices, battery modeling, and AI governance. Key Positions: Professor of Machine Learning, University of Oxford Official Fellow, Exeter College Co-founder of Mind Foundry Lead Researcher, Oxford Martin Programme on Technology and Employment Research Themes: Probabilistic modeling for quantum systems Uncertainty quantification in energy storage AI safety and societal impact analysis Automated experimental design Quantum device calibration Probabilistic numerical methods Technical Contributions: Bridging reality gap in quantum devices Efficient Bayesian quadrature techniques Personalized neurostimulation algorithms Automated measurement protocols Quantum-classical hybrid ML
Simon King is a Professor of Speech Processing at the University of Edinburgh , affiliated with the School of Philosophy, Psychology and Language Sciences . He serves as Director of the Centre for Speech Technology Research (CSTR) and teaches courses like Speech Processing and Speech Synthesis , while directing the MSc in Speech and Language Processing . Research Interests His research focuses on: Developing new acoustic models (e.g., Linear Dynamical Models, factorial-HMMs) for speech recognition Advancing unit selection and HMM-based speech synthesis Integrating articulatory measurement data for enhanced modeling Exploring perceptual measures in synthesis criteria Building multilingual speech systems to identify universal speech building blocks Publication Trends Simon's recent work emphasizes deep learning (DNNs, LSTMs) in speech synthesis, multilingual frameworks , and articulatory-acoustic feature integration . His studies often bridge grapheme-based modeling , perceptual error reduction , and noise-robust synthesis . Scientific Awards EPSRC Advanced Research Fellowship (2005-2009) Students & Collaborations He has supervised numerous PhD students including Rasmus Dall, Tom Merritt, and Srikanth Ronanki. Current research fellows like Mirjam Wester and Zhizheng Wu contribute to projects such as Natural Speech Technology (NST) and Simple4All .
Andrew Zisserman is a Royal Society Research Professor at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). His research focuses on computer vision, artificial intelligence, and neural networks, with significant contributions to multimodal learning, video understanding, and 3D scene analysis. He leads projects exploring visual-language models, audio-visual synchronization, and clinical imaging applications. Key research areas include: Video analysis and temporal modeling Multimodal systems for sign language translation and action recognition 3D shape estimation and physical property inference Foundation models and cross-modal retrieval Recent work highlights: Developed Flamingo and Tapir models for video-language tasks Advancements in spinal MRI analysis and clinical imaging Leadership in EGO4D and VoxCeleb challenges Honors include Fellowship of the Royal Society (FRS) and the ISSLS Prize in Clinical Science 2023 for spinal analysis innovations. His lab collaborates globally, emphasizing real-world applications in healthcare and autonomous systems.
Alex Wong is an Assistant Professor of Computer Science at Yale University, specializing in computer vision, robotics, and medical imaging. His research focuses on sensor fusion, unsupervised learning, 3D vision, robust perception under adverse conditions, and medical image analysis. He holds degrees from the University of California, Los Angeles (UCLA), including a B.S., M.S., and Ph.D. in Computer Science. Wong’s work bridges theoretical advances with practical applications, particularly in depth estimation, autonomous systems, and medical diagnostics. He has received prestigious awards such as the NeurIPS Outstanding Student Paper Award (2011) and the ICRA Best Paper Award in Robot Vision (2019). His research often addresses challenges in unstructured environments, emphasizing robustness and adaptability. Recent projects include developing novel frameworks for unsupervised depth completion, adversarial robustness in vision systems, and multimodal fusion techniques. His contributions span conferences like CVPR, ICCV, and ICRA, with a strong focus on advancing AI for real-world applications in healthcare and robotics. Education: B.S., Computer Science, UCLA M.S., Computer Science, UCLA Ph.D., Computer Science, UCLA Awards: NeurIPS Outstanding Student Paper Award (2011) ICRA Best Paper Award in Robot Vision (2019) His lab at Yale Engineering focuses on AI-driven solutions for perception challenges, collaborating across disciplines to advance medical imaging and autonomous systems. Current efforts explore generative models, continual learning, and vision-language integration for robust scene understanding.
Professor Anders C. Hansen is a mathematician at the University of Cambridge and University of Oslo, leading the Applied Functional and Harmonic Analysis group. His work bridges functional analysis, artificial intelligence, and computational mathematics, focusing on the Solvability Complexity Index (SCI) hierarchy and stability issues in deep learning. He has held prestigious fellowships, including a Royal Society University Research Fellowship and Peterhouse Bye-Fellowship. Educated at the University of Cambridge, UC Berkeley, and the Norwegian University of Science and Technology Developed groundbreaking theories in compressed sensing and deep learning, revealing algorithmic instability paradoxes Organized workshops on computational mathematics and AI interpretability His research explores the SCI hierarchy , exposing computational barriers in AI, quantum mechanics, and inverse problems. Key projects include Smale’s 18th problem and analyzing neural network stability. His work has transformed understanding of compressed sensing, particularly in medical imaging. Recent scientific awards include the Whitehead Prize (2019), IMA Prize (2018), and Leverhulme Prize (2017). Collaborations span institutions like Caltech, MIT, and the University of Vienna. As an educator, he teaches NST Part IA Mathematical Methods , Part II Numerical Analysis , and a Part III course on Compressed Sensing . His group has mentored 17 PhD and postdoctoral researchers since 2012.
Ram Vasudevan is an Associate Professor and Associate Chair of Graduate Studies in the Department of Robotics at the University of Michigan. His research focuses on developing tools for safe and robust deployment of robotic systems, emphasizing optimization, nonlinear control, and real-world applications. Key areas include legged robot locomotion, shared control systems, and safety-critical autonomous systems. Research Interests: Optimization and control of nonlinear systems, locomotion of legged robots, shared control active safety systems, and automation of diagnostic/rehabilitative tasks. His ROAHM Lab prioritizes mathematical guarantees for robotic performance, with applications in medical robotics, autonomous vehicles, and soft robotics. Recent work emphasizes trajectory optimization, sensor fusion, and safety-aware control strategies. He has contributed to benchmarks for autonomous vehicle perception and novel methods in thermal image restoration using neural radiance fields. Awards: None explicitly listed in provided text. Labs/Teams: Directs the ROAHM Lab, collaborating on projects like robotic tail mechanics, real-time motion planning, and sensor data analysis. Active in academic conferences including RSS and ICRA.
Prof. Laura Leal-Taixé is a Professor at the Technical University of Munich (TUM) in the Department of Informatics, leading the Dynamic Vision and Learning group. She holds the Rudolf Mößbauer Tenure Track Assistant Professorship, promoted to W3 Associate Professorship in 2022. Her research focuses on computer vision and machine learning, particularly video analysis, multi-object tracking, and autonomous driving applications. Education: B.Sc./M.Sc. in Telecommunications Engineering (Technical University of Catalonia, UPC) Ph.D. in Information Processing (Leibniz University of Hannover, Germany) Postdoctoral Research at ETH Zürich (Switzerland) and TUM Research Interests: Laura’s work addresses challenges in video analysis, including motion analysis, semantic segmentation, and integrating social dynamics into urban traffic modeling. Her project socialMaps , funded by the Sofja Kovalevskaja Award, explores decoupling vehicle and pedestrian traffic using dynamic maps. Her research combines optimization techniques, deep learning, and sensor data for real-world applications like autonomous systems. Recent Trends: Her publications highlight advancements in multi-object tracking, 3D LiDAR segmentation, and trajectory forecasting, emphasizing neural networks and geometric approaches. Collaborations with industry and academia underscore her work’s practical impact. Awards: 2017 Sofja Kovalevskaja Award (€1.65M, Humboldt Foundation) 2017 DAAD Australia-German Joint Research Grant Travel grants from Women in CV, CVPR Doctoral Consortium, and Vodafone Foundation Labs & Projects: Leads the Dynamic Vision and Learning group at TUM, focusing on vision-based AI for autonomous systems and environmental monitoring. Active in developing datasets like DynamicEarthNet for semantic change analysis.
Angjoo Kanazawa is an Assistant Professor in the Department of Electrical Engineering and Computer Sciences (EECS) at the University of California, Berkeley. She leads the Kanazawa AI Research (KAIR) lab under the Berkeley Artificial Intelligence Research (BAIR) umbrella and serves on the advisory board of Wonder Dynamics. Her research focuses on the intersection of computer vision, computer graphics, and machine learning, with a particular emphasis on 4D reconstruction of dynamic scenes, neural radiance fields (NeRF), and systems that model human-environment interactions from 2D visual data. Education : Ph.D., Computer Science (2017), University of Maryland, College Park BA, Mathematics and Computer Science (2012), New York University (NYU) Her work aims to build systems that can capture, perceive, and understand complex 3D/4D worlds from photographs and videos, enabling applications in scene reconstruction, motion analysis, and generative modeling. She has pioneered techniques for scaling NeRFs across GPUs (NeRF-XL), developing open-source tools like nerfstudio and gsplat. Her recent publications focus on topics like self-occluded avatar recovery (SOAR), decentralized diffusion models, and 4D reconstruction of articulated objects for robotics. Kanazawa's research has been recognized with prestigious awards including the IEEE CS TCPAMI Young Researcher Award (2024) , Sloan Research Fellowship (2023) , and Google Faculty Research Award (2021) . Her lab has trained numerous students who now hold positions at leading institutions and companies like Anthropic, Meta Reality Labs, and Luma AI. Key Scientific Awards : IEEE CS TCPAMI Young Researcher Award (2024) Sloan Research Fellow (2023) Hellman Fellow (2022) Bakar Fellows Spark Award (2022) Google Faculty Research Award (2021) Her KAIR lab collaborates extensively with industry partners and academic institutions, including the Max Planck Institute and Google Research. She has served as an advisor for PhD students and postdocs who now lead teams at UC Berkeley, MIT, Stanford, and Luma AI, while her teaching includes graduate courses like CS 280A (Computer Vision) and CS 294-173 (Learning for 3D Vision).
Ajmal Mian is a Professor of Computer Science at the University of Western Australia (UWA), affiliated with the School of Physics, Maths and Computing. He holds an Australian Research Council Future Fellowship (2022) and leads research in Artificial Intelligence, Computer Vision, and Machine Learning. His work focuses on 3D computer vision, adversarial AI defense, and explainable AI. His research interests include 3D point cloud analysis, face recognition, human action recognition, and remote sensing. He has published over 300 papers and secured major grants from ARC, NHMRC, and DARPA, totaling millions in funding. He has supervised 29 PhD students and mentored 12 postdoctoral researchers. Key projects include 3D diffusion models for scene generation, robust 3D vision systems, and defense against AI deception attacks. He serves as a fellow of IAPR, an ACM Distinguished Speaker, and has editorial roles at IEEE Transactions on Neural Networks and Pattern Recognition. Research Awards: HBF Mid-Career Scientist of the Year, West Australian Early Career Scientist of the Year, IAPR Best Scientific Paper Award. Grants: ARC Discovery Projects, National Intelligence & Security Discovery grants, DARPA grants for AI security. His teaching spans computer vision, machine learning, and programming courses. Collaborations include defense, medical, and agricultural applications.
Wojciech Matusik is a Professor of Electrical Engineering and Computer Science at MIT's Computer Science and Artificial Intelligence Laboratory (CSAIL). He leads the Computational Design and Fabrication Group and is a member of the Computer Graphics Group. His research spans computer graphics, robotics, and AI-driven manufacturing, with a focus on computational design, tactile sensing, and material science. Matusik holds a PhD in Computer Science from MIT (2003), an MS from MIT (2001), and a BS from UC Berkeley (1997). His work includes groundbreaking projects like differentiable cloth simulation (DiffCloth), AI-enhanced molecular design, and tactile sensing gloves. He has received prestigious awards such as the MIT TR35 (2004), DARPA Young Faculty Award (2012), and Ruth and Joel Spira Teaching Award (2014). Matusik teaches courses on computer graphics, machine learning, and computational fabrication at MIT. Key research themes include: Robotics: Robotic assembly, tactile interaction, and soft robotics Graphics: 3D holography, procedural material generation Manufacturing: Additive fabrication, topology optimization His recent articles explore AI-driven molecular synthesis, holographic displays, and tactile-enabled VR systems. Matusik collaborates on open-source tools like the WiReSens tactile platform and Simit language for sparse systems.
Deva Ramanan is a Professor at the Robotics Institute of Carnegie Melllon University, where he leads research in computer vision and machine learning. His work focuses on modeling human visual perception, leveraging large-scale visual data, and developing systems for 3D understanding, neural rendering, and autonomous systems. He advises a large group of PhD students and has mentored numerous postdoctoral researchers now in leading roles across industry and academia. His research interests include computer vision, machine learning, human perception modeling, 3D scene understanding, neural rendering, autonomous driving, video understanding, and multimodal foundation models. These areas reflect his focus on both foundational models and their application to real-world problems in robotics and AI. The recent publications highlight a strong trend toward multimodal and 3D-aware models, with increasing use of diffusion models, neural fields, and large vision-language systems. Key themes include scene flow, 3D reconstruction from monocular video, autonomous driving perception, and robust evaluation of vision-language models. There is a clear emphasis on both methodological innovation and practical deployment in dynamic environments. Marr Prize, Honorable Mention (ICCV 2021) Best Paper, Honorable Mention (ECCV 2020) Best Paper Finalist (WACV 2024) Best Paper Award (WACV 2016) Best Industrial Paper, Honorable Mention (BMVC 2017) Marr Prize winner (ICCV 2009) Deva Ramanan has advised numerous PhD and master’s students, many of whom are now at top institutions and companies including Apple, Meta, Google, Nvidia, OpenAI, and Princeton. He has received substantial funding from IARPA, DARPA, NSF, Intel, Google, and Facebook for projects in video analytics, dispersed computing, visual cloud systems, and multi-task recognition. His group has developed influential datasets and benchmarks used widely in the community. He leads a vibrant research lab focused on advancing computer vision through deep learning and multimodal integration. His team works on core challenges in perception, including 3D reconstruction, motion modeling, object detection, and scene understanding, with applications in robotics and autonomous systems.
Subhransu Maji is an Associate Professor in the Manning College of Information and Computer Sciences at the University of Massachusetts Amherst, and the co-director of the Computer Vision Lab. He is also affiliated with the Center for Data Science and holds a part-time role as an Amazon Scholar. His research focuses on high-level visual recognition algorithms and interdisciplinary applications in ecology and astronomy. He has received prestigious awards including the NSF CAREER Award (2018), Best Paper at WACV 2015, and the Google Graduate Fellowship (2008). Education: PhD in Computer Science from UC Berkeley (2011), BTech from IIT Kanpur (2006). Prior roles include Research Assistant Professor at Toyota Technological Institute at Chicago (2012-2014). Research Interests: Computer Vision Machine Learning AI Applications in Ecology and Astronomy 3D Shape Understanding Climate Science Grants and Funding: Supported by NSF, NASA, Climate Change AI, and industry grants from Facebook, NVIDIA, Adobe, and Dolby. Current projects include satellite imagery analysis for ecology and material science applications using deep learning. Labs and Teams: Leads the Computer Vision Lab, collaborates with interdisciplinary teams on ecological monitoring (e.g., bird migration tracking via radar data) and material property prediction (e.g., zeolite adsorption modeling).
Alexei A. Efros is the Howard Friesen Professor in the EECS Department at the University of California, Berkeley, and a core member of the Berkeley Artificial Intelligence Research (BAIR) Lab. Previously, he spent a decade at Carnegie Mellon University’s Robotics Institute. His research focuses on data-driven computer vision, self-supervised learning, computational photography, and generative models. He has pioneered advancements in visual representation learning, including seminal work on neural radiance fields and generative adversarial networks. Education Background: Efros holds a PhD in Computer Science from MIT, though specific details of his academic journey are not explicitly provided in the text. His career includes postdoctoral research at the University of Oxford with Andrew Zisserman and collaborative work with Team WILLOW at INRIA Paris. Research Interests: Efros explores how vast uncurated visual data can be leveraged for understanding and synthesizing the visual world. Key areas include self-supervised learning, generative models, and applications in robotics and art. His lab has contributed influential techniques such as Style Transfer, GAN-based image synthesis, and neural scene representation learning. Recent work emphasizes real-time adaptation (Test-Time Training), 3D perception models, and ethical AI implications of generative systems. Publications: Over 150+ publications span topics like Generative Adversarial Networks (GANs), unsupervised learning, and visual-linguistic models. Notable works include Unpaired Image-to-Image Translation (CUT/GAU), Style Transfer , and Swapping Autoencoder . His research has significant industry impact, with techniques adopted in Adobe’s software and generative AI applications. Grants & Collaborations: Efros has secured major funding from NSF, DARPA, and industry partnerships (e.g., Adobe, NVIDIA). He co-leads projects on scalable vision models, ethical AI, and real-world perception systems. Current collaborations include work with MIT, NYU, and INRIA Paris. Labs & Teams: Leads the BAIR Vision Group at Berkeley, fostering interdisciplinary research between computer vision, graphics, and robotics. The group emphasizes Slow Science principles, prioritizing deep exploration over rapid publication.