Jonathan T. Barron is a Researcher at Google DeepMind in San Francisco, specializing in Computer Vision , Neural Rendering , and 3D Scene Reconstruction . He earned his PhD at UC Berkeley under Jitendra Malik and has pioneered advancements in NeRF (Neural Radiance Fields) and diffusion-based 3D generation. Research Interests : Computer Vision, Deep Learning, Generative AI, Image Processing, and 3D Reconstruction via Radiance Fields. His work includes Bolt3D for rapid 3D scene generation, CAT3D/CAT4D for text-to-3D/4D, and Zip-NeRF for anti-aliased radiance fields. He has also developed real-time rendering frameworks like SMERF and NeRF-Casting for reflections. Scientific awards: PAMI Young Researcher Award He has served as Area Chair for CVPR, ICCV, and NeurIPS, and his research is widely adopted in applications like Google's Lens Blur , Portrait Mode , and Jump VR .
Andreas Geiger is a Professor and Head of the Department of Computer Science at the University of Tübingen, Germany. He leads the Autonomous Vision Group (AVG) within CyberValley and is a core faculty member of the Tübingen AI Center. His roles also include PI in the ML in Science Excellence Cluster and the CRC Robust Vision, as well as ELLIS Fellow and coordinator of the ELLIS PhD program. He specializes in machine learning models for computer vision, robotics, and autonomous systems, with applications in self-driving cars, VR/AR, and scientific document analysis. Educational background: While not explicitly detailed, his positions imply a Ph.D. in Computer Science or related field. His work spans interdisciplinary collaborations with institutions like ETH Zürich, Microsoft, and the University of Bonn. Research focuses on 3D scene understanding, Gaussian splatting, generative models, and reliable autonomous systems. Notable contributions include the KITTI dataset and foundational work in neural radiance fields. Awards include the Sage 10-Year Impact Award (2024), ERC Starting Grant (2019), and IEEE PAMI Young Researcher Award (2018). Key projects include the Scholar Inbox paper recommender platform, ReSim (reliable world simulation), and advancements in 3D scene generation (e.g., UrbanCAD, PrITTI). His lab maintains a strong focus on open-source tools and datasets, such as the CARLA Route Generator. Grants and funding include support from Vector Stiftung (MINT innovation program) and EU initiatives like the ML in Science Cluster. His team collaborates internationally, with recent work presented at CVPR, SIGGRAPH, and NeurIPS.
Hao Liu is an incoming Assistant Professor of Machine Learning at Carnegie Mellon University and currently works as a research scientist at Google DeepMind. Previously, he completed his Ph.D. in Computer Science at UC Berkeley under the supervision of Pieter Abbeel. He also spent two years part-time at Google as part of the Google Brain team. His educational background includes: Ph.D. in Computer Science from UC Berkeley Hao Liu's research focuses on solving intelligence through deep learning, neural networks, and innovative learning objectives. His work spans multiple areas including large language models, reinforcement learning, world models, and attention mechanisms for long context processing. He has made significant contributions to making transformer models more efficient and capable of handling extremely long sequences through techniques like Ring Attention and Blockwise Transformers. His recent publications demonstrate a strong focus on extending the capabilities of language and vision models, particularly in handling long sequences and multimodal data. Key themes include attention optimization, tokenization efficiency, and alignment techniques. His work bridges theoretical advances with practical implementations for real-world AI systems, with multiple papers at top conferences including NeurIPS, ICML, and ICLR, often receiving spotlight or oral presentations. Hao is actively involved in open-source AI research, having contributed to projects like Koala and OpenLLaMa, which aim to make advanced language models more accessible to the research community. His work on RingAttention has been implemented as a Python package available on GitHub, demonstrating his commitment to practical implementations and community sharing.
Professor Winston Hsu is a distinguished faculty member in the Department of Computer Science and Information Engineering at National Taiwan University, where he has served as a full professor since 2015. He is the co-director of the Communications and Multimedia Laboratory (CMLab) and founder of the MiRA (Multimedia indexing, Retrieval, and Analysis) research group. Additionally, he serves as the Founding Director for NVIDIA AI Lab at NTU, the first such lab in Asia. Professor Hsu received his Ph.D. in Electrical Engineering from Columbia University in 2007 under the supervision of Professor Shih-Fu Chang. Prior to his academic career, he was a founding engineer and research manager at CyberLink Corp., now a public image/video software company. National Taiwan University (2007-Present): Professor (2015-Present), Assistant/Associate Professor (2007-2015) MobileDrive (2021-2024): CTO and Vice President (Joint Venture between Foxconn and Stellantis) IBM TJ Watson Research Center (2016-2017): Visiting Scientist Microsoft Research Redmond (2014): Visiting Researcher Columbia University (2007): Ph.D. in Electrical Engineering Professor Hsu's research focuses on machine learning, computer vision, large-scale image and video search and recognition, and embedded AI. His work spans from fundamental research in visual recognition to practical applications in automotive systems, medical imaging, and e-commerce. He has pioneered work in disguised face recognition, low-resolution face hallucination, 3D model search, and virtual try-on systems. His current research emphasizes Embodied AI, integrating perception, action, and learning technologies for applications in automotive and robotics domains. His research group has produced numerous influential publications, particularly in top computer vision and multimedia conferences like CVPR, where they won first place in the Disguised Face Recognition competition in 2018. Their work spans diverse application areas including security, medical diagnostics, automotive systems, and e-commerce solutions, demonstrating strong translation from academic research to real-world impact. IBM Research Pat Goldberg Memorial Best Paper Award (2018) First Place, IARPA Disguised Faces in the Wild Competition (CVPR 2018) Best Brave New Idea Paper Award, ACM Multimedia 2017 NVIDIA AI LAB Award (First in Asia, 2016) First Place, MSR-Bing Image Retrieval Challenge (2013) World's Top 2% Scientists (2023) Professor Hsu actively mentors students and researchers, with his group consistently recruiting PhD students, postdocs, and research assistants. He has successfully bridged academia and industry through multiple collaborations, including his role as CTO at MobileDrive (a joint venture between Foxconn and Stellantis) from 2021-2024. His research has been supported by significant industry partnerships with Microsoft, IBM, and NVIDIA, as well as government grants from Taiwan's Ministry of Science and Technology. His laboratory, the Communications and Multimedia Laboratory (CMLab), maintains strong industry connections and focuses on cutting-edge research in visual AI. The lab has produced numerous award-winning projects and maintains active collaborations with global technology companies, particularly in the automotive and consumer electronics sectors.
Yaser Sheikh is an Associate Professor at the Robotics Institute of Carnegie Mellon University (on leave) and Director of the Facebook Reality Lab, Pittsburgh . He holds appointments in the Mechanical Engineering Department and focuses on ' metric telepresence ' for AR/VR interactions. His research spans machine perception , computer vision , computer graphics , and machine learning , with applications in social behavior modeling and dynamic 3D reconstruction. University: Carnegie Mellon University Roles: Associate Professor (Robotics Institute), Director (Facebook Reality Lab) Contact: yaser@cs.cmu.edu, yasers@fb.com Research Interests include: Computer Vision: Pose estimation, 3D reconstruction, camera calibration Computer Graphics: Face/Hand animation, photorealistic rendering Machine Learning: Neural rendering, unsupervised learning for landmark detection AR/VR: Telepresence, immersive social interactions Notable Trends in Publications reveal a focus on real-time pose estimation (e.g., OpenPose), dynamic 3D reconstruction , and codec avatars for VR/AR. Recent works emphasize universal priors and neural rendering for photorealistic avatars. Scientific Awards include: Popular Science’s Best of What’s New Award Honda Initiation Award (2010) Best Paper Awards: WACV (2012), SCA (2010), ICCV THEMIS (2009) Hillman Fellowship for Excellence in Computer Science Research (2004) Advising and Grants: He has advised numerous PhD students (e.g., Hanbyul Joo, Tomas Simon) and received funding from the National Science Foundation , DARPA, and industry partners like Intel , Disney , and Honda . Labs & Teams: Leads the Facebook Reality Lab in Pittsburgh, collaborating with institutions like Carnegie Mellon University and Disney Research.
Joy Arulraj is an Associate Professor in the School of Computer Science within the College of Computing at Georgia Institute of Technology. His research focuses on data systems, machine learning, and database systems, with a particular emphasis on video analytics and adaptive query processing. He leads the Data Systems and Analytics Group and is developing the EVA AI-Relational Data System. Dr. Arulraj's research interests span data systems, machine learning, database systems, video analytics, and adaptive query processing. His work centers on developing systems that efficiently process complex queries, particularly for video analytics and machine learning workloads. He has made significant contributions to GPU database systems, non-volatile memory database management, and adaptive query processing techniques. His research often bridges the gap between theoretical database principles and practical implementations for modern hardware architectures. His recent publications show a strong trend toward video analytics systems, adaptive query processing for machine learning workloads, and GPU-accelerated database systems. The EVA system represents a major focus of his recent work, providing end-to-end exploratory video analytics capabilities. His research also addresses fundamental database concepts like buffer management, query optimization, and storage management, adapting these principles for modern hardware and application requirements. Dr. Arulraj has advised numerous graduate students including Pramod Chunduri, Gaurav Tarkok Kakkar, Jiashen Cao, and Sayan Sinha. His graduated students have gone on to work at companies like ServiceNow, Meta Research, and the Korean Army. He actively teaches database system courses at Georgia Tech, including Database System Implementation (CS 4420/6422) and Advanced Database System Implementation (CS 4423/6423), where students build database systems from scratch using C++ and the BuzzDB framework. He maintains an active research program with consistent publication output across top database and systems conferences. His work spans from theoretical database principles to practical system implementations, with a recent emphasis on video analytics, machine learning integration with database systems, and leveraging modern hardware like GPUs and non-volatile memory for database applications.
Geng Yuan is an Assistant Professor at the University of Georgia's School of Computing, specializing in AI systems, energy-efficient deep learning, and hardware-software co-design. His work bridges machine learning algorithms with emerging hardware technologies like superconducting circuits and ReRAM. He holds a Ph.D. in Computer Engineering from Northeastern University (2023) and a Master's in Electrical & Computer Engineering from Syracuse University (2016). Doctor of Philosophy (Ph.D.) in Computer Engineering, Northeastern University (2023) Master of Science (M.S.) in Electrical & Computer Engineering, Syracuse University (2016) Bachelor of Science (B.S.) in Electrical Engineering, Beijing University of Technology (2014) His research focuses on optimizing deep learning systems for edge computing and mobile platforms through techniques like model compression, sparse training, and hardware-aware neural architecture search. Recent projects include adapting large language models via hybrid-grained pruning and developing ultra-low-power AQFP circuits for binary networks. Geng Yuan's publications span top venues like NeurIPS, CVPR, ICML, ICLR, ISCA, and DAC, with notable awards including a Best Paper Award at ICLR Workshop 2021, Spotlight Papers at ICLR 2023 and NeurIPS 2021, and a Design Contest 1st Place at ISLPED 2020. Best Paper Award (ICLR Workshop'21) Spotlight Paper Award (ICLR'23, NeurIPS'21) Design Contest 1st Place (ISLPED'20) Best Paper Nomination (DATE'21, ISQED'18) He actively recruits Ph.D., Master's students, and interns to his research group, focusing on advancing AI systems through interdisciplinary approaches combining machine learning, computer architecture, and electronic design automation.
Marc Pollefeys is a Full Professor at the Department of Computer Science, ETH Zurich, and Director of the Microsoft Mixed Reality and AI Lab. His work focuses on advanced perception systems for HoloLens, 3D computer vision, robotics, and machine learning. Key contributions include automated 3D modeling from video, real-time reconstruction pipelines, and vision-based autonomous systems. Education: PhD from KU Leuven (1999) Previous Affiliation: Professor at UNC Chapel Hill Research interests span 3D reconstruction , computer vision , robotics , SLAM , augmented reality , and privacy-preserving mapping . His work often integrates geometric modeling , feature matching , and deep learning . Recent projects emphasize implicit 3D representations , open-vocabulary scene understanding , and robust estimation using neural-guided algorithms. Recent publications highlight advancements in neural implicit fields , line-based correspondence , and vision-language integration . Trends include hybrid point-line methods, differentiable RANSAC, and privacy-aware localization frameworks. Scientific recognition includes: IEEE Fellow (2012) David Marr Prize (ICCV 1998) DAGM Best Paper Award (1999) Advisees include current and alumni PhD students such as Yagız Aksoy, Federico Camposeco, and Sudipta Sinha. Collaborations span institutions like UNC Chapel Hill, ETH Zurich, and Microsoft Zurich. Research sponsors include Microsoft, Google, and European research initiatives.
Jeannette Bohg is an Assistant Professor of Computer Science at Stanford University, directing the Interactive Perception and Robot Learning Lab. Previously, she was a group leader at the Autonomous Motion Department (AMD) of the MPI for Intelligent Systems (2012-2017). She holds a PhD from KTH Royal Institute of Technology (Stockholm) and degrees from Chalmers University and TU Dresden. Her research focuses on perception, learning, and real-time multi-modal methods for autonomous robotic manipulation and grasping, aiming to bridge principles of human sensorimotor coordination with robotic implementation. Education: PhD in Robotics (KTH), MSc in Art & Technology (Chalmers), Diploma in Computer Science (TU Dresden) Research interests include developing goal-directed, real-time robotic systems capable of meaningful feedback for execution and learning. Key areas are dexterous manipulation, imitation learning, and cross-embodiment policy transfer. Notable contributions include the TidyBot platform and work on force-aware surgical robotics. Awards include the 2019 IEEE ICRA Best Paper Award, 2019 IEEE RA Early Career Award, and 2020 RSS Early Career Award. Her lab explores intersections of robotics, ML, and computer vision. Advising: Actively mentoring students/postdocs in manipulation, perception, and learning. Grants and collaborations span NSF, Stanford AI Lab, and industry partnerships. Future work emphasizes robust real-world deployment and human-robot collaboration. Labs/Teams: Leads the Interactive Perception and Robot Learning Lab, contributing to Stanford’s AI ecosystem. Previously managed the MPI AMD group, fostering interdisciplinary research in autonomous systems.
Prof. Bernt Schiele is a Max Planck Director at the Max Planck Institute for Informatics and holds a Professorship at Saarland University. His research focuses on understanding multimodal sensor data, with key areas in computer vision, 3D object recognition, and machine learning. He leads the Computer Vision and Machine Learning group, addressing challenges in sensor fusion, scene understanding, and human activity recognition. Schiele has held academic roles at TU Darmstadt, ETH Zurich, and MIT, and contributes to top journals like IEEE Transactions on PAMI and conferences like ECCV. His work emphasizes robust models, interpretability, and domain adaptation for real-world applications. Education: PhD (1997, Grenoble), MSc (1994 Karlsruhe/1993 Grenoble) Key Positions: MIT (1997-2000), ETH Zurich (1999-2004), TU Darmstadt (2004-2010) Research interests span 3D scene understanding, multimodal sensor processing, and machine learning techniques for large-scale data. His recent work advances robust object detection, explainable AI, and domain-invariant training methods. He also chairs major conferences like ECCV 2018 and co-chairs ICCV 2011. Publications highlight innovations in interpretable vision transformers, certified explanations, and test-time adaptation. Despite no listed awards, his contributions shape foundational areas of computer vision and multimodal AI.
Andrea Tagliasacchi is an Associate Professor at Simon Fraser University's School of Computing Science, holding the Visual Computing Research Chair. He is also a part-time (20%) staff research scientist at Google DeepMind (Toronto) and an associate professor (status-only) at the University of Toronto's computer science department. His research focuses on 3D visual perception at the intersection of computer vision, graphics, and machine learning. Education: EPFL – Postdoc Simon Fraser University – PhD (NSERC Alexander Graham Bell Fellow) Politecnico di Milano – MSc (Gold Medalist) Research Interests: His work emphasizes 3D reconstruction, neural fields, and applications in robotics, autonomous systems, and augmented reality. Recent advancements include scalable 3D Gaussian splatting, robust neural rendering techniques, and diffusion models for 4D generation. Notable Articles: Recent work spans real-time differentiable ray tracing, stochastic rasterization for 3D Gaussian splats, and generative image composition using neural fields. His publications often blend theoretical contributions with practical applications in CVPR, SIGGRAPH, and NeurIPS. Awards: 2015 SGP Best Paper Award 2020 CVPR Best Student Paper Award 2024 CVPR Best Paper Honorable Mention Advising & Grants: Advised 14+ PhD/MSc students (e.g., Baptiste Angles, Sara Sabour) and co-advised with notable figures like Geoffrey Hinton. Active in grants involving neural field compression, robotic perception, and generative AI. Labs & Teams: Leads a lab at SFU focused on 3D vision and neural fields, collaborating with industry partners like Google Brain and Samsung Research.
Professor Gabriel Brostow is a faculty member in the Department of Computer Science at University College London (UCL), where he leads research in Computer Vision and Human-Computer Interaction. He also serves as Chief Research Scientist and Senior Director of the R&D Team at Niantic, the company behind Pokémon GO. His work bridges academic research and industry applications, focusing on developing AI systems that enhance human capabilities through what he terms 'Human in the Loop AI'—now commonly referred to as Human-Centered AI. Brostow completed his BS in Electrical Engineering at UT Austin, followed by a PhD with Irfan Essa at Georgia Tech. He then pursued postdoctoral research with Roberto Cipolla's Computer Vision & Robotics Group at Cambridge University as a Marshall Sherfield Fellow, and with Marc Pollefeys in ETH Zurich's CVG Group. His research explores how AI, particularly Computer Vision, can serve as 'super-tools' for professionals across various domains including filmmaking, architecture, robotics, and scientific research. Specific interests include assistive technology for everyday life, authoring systems that maximize user effort, 3D reconstruction, depth estimation, and vision-language models. His work often involves creating systems that are validated through real-world human interaction to ensure practical utility. Analysis of his recent publications reveals a strong focus on practical applications of Computer Vision that directly interact with humans. His research spans 3D scene understanding, depth estimation, sketch-based interfaces, and multimodal AI systems. There's a clear emphasis on creating benchmarks and tools that facilitate human-AI collaboration, with applications in assistive technology, urban planning, filmmaking, and biodiversity monitoring. His work frequently appears at top conferences including CVPR, NeurIPS, ECCV, and CHI. Marshall Sherfield Fellowship Brostow actively mentors PhD students, with current advisees including Ross Murphy, Skanda Koppula, Gizem Unlu, Omiros Pantazis, and Jamie Watson. His alumni include numerous PhD graduates and MSc students who have gone on to successful careers in academia and industry. He emphasizes selecting students based on passion and potential rather than just academic credentials, valuing traits like helpfulness, drive, and hunger to learn. His research is supported through collaborations with major institutions and companies including DeepMind, MIT, and the University of Edinburgh. He leads a research group at UCL that collaborates closely with Niantic's R&D team, creating a unique bridge between academic research and industry application. His team's work frequently involves developing novel Computer Vision techniques that are validated through real-world human interaction, ensuring practical utility alongside technical innovation. The group explores blue-sky research problems with applications ranging from assistive technology to professional tools for filmmakers, architects, and scientists studying diverse environments.
Thomas Hacker is a Professor in the Department of Computer and Information Technology at Purdue Polytechnic Institute, Purdue University. His research focuses on cloud computing, high-performance computing, operating systems, computer networking, and cyber infrastructure . He holds a Ph.D. and M.S. in Computer Science & Engineering from the University of Michigan, along with dual B.S. degrees in Computer Science and Physics from Oakland University. Education: PhD (Computer Science & Engineering), University of Michigan (2004) MS (Computer Science & Engineering), University of Michigan (1993) BS (Computer Science, Mathematics Minor), Oakland University (1989) BS (Physics), Oakland University (1989) Dr. Hacker's research spans cloud and grid computing, operating systems, and distributed systems , with applications in earthquake engineering data systems and AI-driven infrastructure analysis. His recent work explores extended layer 2 networking for bare-metal provisioning ( 2023 IEEE Cloud Summit ) and machine-supported bridge inspection using artificial intelligence ( Transportation Research Record, 2023 ). Notable scientific contributions include 15+ publications on topics like cyberinfrastructure for earthquake engineering, container-based virtualization, and data-intensive systems. His work has been recognized with awards such as the NSF CAREER Award (2010) and multiple Purdue Seed for Success Awards . Key Scientific Awards: NSF CAREER Award (2010) Purdue Seed for Success Awards (2008-2013) ASEE Information Systems Division Best Paper Award (2012) College of Technology Outstanding Faculty in Discovery Award (2010) He has held leadership roles at Purdue, including Department Head (2018-2021) and Interim Department Head (2011-2016) . His career spans academic positions at Indiana University, University of Michigan, and industry roles at Storage Technology Corporation.
Xiaoming Liu is the Anil K. and Nandita Jain Endowed Professor of Engineering and MSU Foundation Professor in the Department of Computer Science and Engineering at Michigan State University . Holding a Ph.D. from Carnegie Mellon University (2004), he leads cutting-edge research in computer vision and machine learning. Research Interests : Computer Vision Pattern Recognition Image and Video Processing Machine Learning Medical Image Analysis Multimedia Retrieval Recent Research Trends : Focus on 3D object detection and depth estimation Development of robust biometric recognition systems Integration of radar-camera fusion for autonomous systems Advancements in self-supervised and multimodal learning Exploration of adversarial AI security Creation of interpretable forgery detection frameworks Teaching : Spring 2013: CSE891-006 Computer Vision Seminar Fall 2012-2015: CSE803 Computer Vision Spring 2014-2017: CSE 471 Media Processing and Multimedia Contact Information : Email: liuxm@cse.msu.edu Office: EB 3137, Michigan State University Phone: +1 (517) 355-2359
Hossam Hassanein is a Professor and Director of the School of Computing at Queen's University. He received his B.Sc. in Electrical Engineering from Kuwait University in 1984, M.Sc. in Computer Engineering from the University of Toronto in 1986, and Ph.D. in Computing Science from the University of Alberta in 1990. He joined Queen's University School of Computing in 1999 and has established himself as a leading researcher in telecommunications and networking. Dr. Hassanein's research interests span wireless sensor networks, mobile ad hoc networks, edge computing, Internet of Things (IoT), radio resource management, and data-centric networks. His seminal contributions include pioneering work on WSN planning, load-balanced routing protocols, and energy-efficient network designs. He has championed research in IoT, developing frameworks for smart spaces that use contextual information to enhance IoT applications in healthcare, transportation, and infrastructure. His recent publications (2023-2025) demonstrate a strong focus on cutting-edge areas including extreme edge computing, vehicular networks, and AI/ML integration in networking. Research trends show increasing emphasis on practical applications in telesurgery, digital twins, and industrial IoT, addressing challenges in resource allocation, task offloading, and real-time processing in constrained environments. Dr. Hassanein has received numerous recognitions for his work: Fellow of the IEEE Queen's University School of Graduate Studies Award for Excellence in Graduate Student Supervision (2015) Multiple best paper awards from top international conferences As founder and director of the Telecommunications Research Lab (TRL), Dr. Hassanein has supervised over 75 students who have made substantial contributions in academia and industry. The TRL is one of Queen's largest research groups with extensive international collaborations. Dr. Hassanein has successfully attracted significant research funding from government and industry sources in the competitive telecommunications field. The Telecommunications Research Lab has developed innovative platforms including SPROUTS, a rugged sensor platform used in mining, steel manufacturing, and smart-grid monitoring. TRL's work has had significant impact in WSN planning, data dissemination, and resource reuse in wireless networks, with contributions featured in IEEE Wireless Communications Magazine.