Prof. Matthias Nießner is a Professor at the Technical University of Munich , where he leads the Visual Computing Lab . Prior to this, he held a Visiting Assistant Professor position at Stanford University . His work bridges computer vision , graphics , and machine learning , focusing on 3D reconstruction , semantic scene understanding , and AI-driven video synthesis . Prof. Nießner has published over 150 works in top venues like SIGGRAPH , CVPR , and ECCV , with several receiving best paper awards (SIGCHI’14, HPG’15, SPG’18, SIGGRAPH’16 Emerging Tech). His research has garnered international media attention, including features in the New York Times , Wall Street Journal , and MIT Technological Review , as well as TV demonstrations (e.g., Jimmy Kimmel Live for Face2Face technology). Awards : TUM-IAS Rudolph Moessbauer Fellowship (2017–ongoing) Google Faculty Award (2017) Nvidia Professor Partnership Award (2018) ERC Starting Grant (2018, €1.5M) Eurographics Young Researcher Award (2019) Research Trends : 3D Gaussian Splatting for real-time rendering Neural Radiance Fields (NeRF) with mesh supervision Audio-driven facial animation via diffusion models Latent space diffusion for 3D scenes Self-supervised and zero-shot methods for 3D and image analysis As a co-founder and director of Synthesia Inc. , he drives democratization of synthetic media. His YouTube channel has over 5 million views, reflecting his impact beyond academia.
David B. Lindell is an Assistant Professor in the Department of Computer Science at the University of Toronto, with affiliations to the Vector Institute and AXL. He is a founding member of the Toronto Computational Imaging Group. His research focuses on physically based intelligent sensing, integrating physical models, signal processing, and AI to advance sensing systems. Notable projects include imaging around corners, through scattering media, and developing machine learning algorithms for 3D scene reconstruction. Education: Ph.D. in Computational Imaging from Stanford University (advisor: Gordon Wetzstein). Awards include the 2024 Ontario Early Researcher Award and the Best Student Paper at CVPR 2025. His work combines computational imaging with applications in computer graphics and autonomous systems. Research interests span non-line-of-sight imaging, single-photon sensing, and neural representations. Key contributions include the Light-Cone Transform (Nature 2018), confocal diffuse tomography (Nature Communications 2020), and AutoInt (CVPR 2021). His lab develops systems for 3D reconstruction, transient imaging, and photon-efficient sensors. Selected grants and support: NSF CAREER Award, DARPA REVEAL program, and KAUST Visual Computing Center funding. Active collaborations with industry and academic institutions on autonomous driving and medical imaging applications.
Jonathan T. Barron is a Researcher at Google DeepMind in San Francisco, specializing in Computer Vision , Neural Rendering , and 3D Scene Reconstruction . He earned his PhD at UC Berkeley under Jitendra Malik and has pioneered advancements in NeRF (Neural Radiance Fields) and diffusion-based 3D generation. Research Interests : Computer Vision, Deep Learning, Generative AI, Image Processing, and 3D Reconstruction via Radiance Fields. His work includes Bolt3D for rapid 3D scene generation, CAT3D/CAT4D for text-to-3D/4D, and Zip-NeRF for anti-aliased radiance fields. He has also developed real-time rendering frameworks like SMERF and NeRF-Casting for reflections. Scientific awards: PAMI Young Researcher Award He has served as Area Chair for CVPR, ICCV, and NeurIPS, and his research is widely adopted in applications like Google's Lens Blur , Portrait Mode , and Jump VR .
Andreas Geiger is a Professor and Head of the Department of Computer Science at the University of Tübingen, Germany. He leads the Autonomous Vision Group (AVG) within CyberValley and is a core faculty member of the Tübingen AI Center. His roles also include PI in the ML in Science Excellence Cluster and the CRC Robust Vision, as well as ELLIS Fellow and coordinator of the ELLIS PhD program. He specializes in machine learning models for computer vision, robotics, and autonomous systems, with applications in self-driving cars, VR/AR, and scientific document analysis. Educational background: While not explicitly detailed, his positions imply a Ph.D. in Computer Science or related field. His work spans interdisciplinary collaborations with institutions like ETH Zürich, Microsoft, and the University of Bonn. Research focuses on 3D scene understanding, Gaussian splatting, generative models, and reliable autonomous systems. Notable contributions include the KITTI dataset and foundational work in neural radiance fields. Awards include the Sage 10-Year Impact Award (2024), ERC Starting Grant (2019), and IEEE PAMI Young Researcher Award (2018). Key projects include the Scholar Inbox paper recommender platform, ReSim (reliable world simulation), and advancements in 3D scene generation (e.g., UrbanCAD, PrITTI). His lab maintains a strong focus on open-source tools and datasets, such as the CARLA Route Generator. Grants and funding include support from Vector Stiftung (MINT innovation program) and EU initiatives like the ML in Science Cluster. His team collaborates internationally, with recent work presented at CVPR, SIGGRAPH, and NeurIPS.
Angel Xuan Chang is an Associate Professor at Simon Fraser University's School of Computing Science, affiliated with labs including 3DLG, GrUVi, SFU NatLang, SFU AI/ML, and VINCI. He holds a Canada CIFAR AI Chair and was a TUM-IAS Hans Fischer Fellow (2018-2022). His research bridges natural language processing (NLP), 3D scene understanding, and embodied AI, focusing on language-grounded 3D generation and biodiversity monitoring via DNA barcodes. Recent work includes NuiScene (unbounded outdoor scene generation), ViGiL3D (3D visual grounding dataset), and CLIBD (vision-genomics biodiversity analysis). He advises students in projects like BIOSCAN-5M insect dataset and embodied AI navigation. His 2025 highlights include multiple ICCV and ICLR papers, workshops at ICML and CVPR, and a CRV invited talk. Education: Ph.D. in Computer Science from Stanford University (2014), advised by Chris Manning. Previous roles include visiting research scientist at Facebook AI Research and researcher at Eloquent Labs.
Georgia Gkioxari is an Assistant Professor in the Division of Computing and Mathematical Sciences at Caltech , with a part-time affiliation at Meta AI . Her work focuses on extending visual perception models through advanced 2D and 3D representation learning, spatial reasoning, and generative models. Education: Not explicitly mentioned in the text Research interests span 3D perception , spatial reasoning , and vision-language integration , with projects like Visual Agentic AI for Spatial Reasoning and Token-by-Token Multimodal Alignment . Her publications emphasize 3D object detection , reconstruction , and generative modeling techniques including diffusion models and transformers . Scientific recognition includes the Meta LLM Evaluation Research Grant , Okawa Research Grant , Google Faculty Scholar Award 2024 , and Amazon Research Award . She teaches courses like Large Language & Vision Models (EE/CS 148) and Learning & 3D (CS 101) at Caltech. Labs & Teams: Leads Glab with members including Ilona Demler, Ziqi Ma, and Damiano Marsili
Massimo Piccardi is a Professor of Natural Language Processing (NLP), Computer Vision, and Machine Learning at the University of Technology Sydney (UTS) , where he has been since 2002. He currently serves as the Head of the School of Electrical and Data Engineering and leads the Big Data Analytics program at the Global Big Data Technologies Centre. His research focuses on advancing NLP, machine learning applications in healthcare, and cybersecurity in IoT systems. He has authored over 200 journal papers and conference proceedings, secured significant ARC and CRC grants, and holds the IEEE Computer Society Distinguished Contributor Award (2022). Education & Professional Roles: Joined UTS in 2002, progressing from Associate Professor (2002–2007) to Professor (2008–present). Serves as Associate Editor for IEEE Transactions on Big Data and Editor for Artificial Intelligence in Medicine. Active in professional societies including IEEE, ACL, and ALTA (President, 2023–2024). Research Interests: Core areas include NLP (translation, summarization, adversarial attacks), healthcare informatics (clinical NLP, health service analysis), and cybersecurity (IoT security, privacy-preserving systems). Cross-cutting themes include generative models, cross-lingual systems, and ethical AI. Grants & Projects: Principal Investigator on ARC Discovery/Linkage projects and CRC grants. Recent projects include controllable machine translation (Amazon), privacy-preserving digital agriculture, and STEM innovation (ASTRID project with NBN Co). Labs & Collaborations: Leads the UTS Global Big Data Technologies Centre, collaborating on projects like adversarial NLP attacks, medical machine translation, and secure IoT frameworks.
Dr. Heather Inwood serves as University Associate Professor in Modern Chinese Literature and Culture within the Faculty of Asian and Middle Eastern Studies at the University of Cambridge. She is also a Fellow and Director of Studies at Trinity Hall, maintaining dual institutional affiliations that support her teaching and research activities across undergraduate and postgraduate programs. Her educational background includes a BA in Chinese Studies from the University of Cambridge (Trinity Hall), followed by advanced studies at Tsinghua University and Peking University, culminating in a PhD in modern Chinese literature from SOAS, University of London in 2008. Prior academic appointments include Assistant Professor at The Ohio State University (2008-2013) and Lecturer in Chinese Cultural Studies at the University of Manchester before returning to Cambridge in 2016. Professor Inwood's research program investigates the dynamic interactions between digital media and literary production in contemporary China, with particular focus on poetry communities, genre fiction, and transmedia storytelling. Her work bridges traditional literary analysis with digital humanities approaches, examining how online platforms transform authorship, reception, and cultural value. She has pioneered scholarship on internet poetry scenes, demonstrating their vitality despite public perceptions of poetry's decline in modern China. Her publication trajectory reveals a clear evolution from early work on poetry communities ( Verse Going Viral: China's New Media Scenes , 2014) toward broader investigations of digital narrative forms, game studies, and Sinophone cyberspace. Recent publications increasingly engage with science fiction, transmedia storytelling, and the algorithmic structures shaping literary production, reflecting both the changing digital landscape and her expanding methodological toolkit. Professor Inwood actively supervises five PhD students working on diverse projects spanning Misty poetry, queer narratives, gender studies, and poetic animality, demonstrating her commitment to nurturing next-generation scholarship in Chinese literary studies. Her research has been supported by British Academy Small Grants and similar funding mechanisms that enable sustained fieldwork and publication. Her teaching responsibilities include undergraduate courses on East Asian media and popular culture, Modern Chinese texts, and Modern Chinese literature, connecting classroom instruction directly to her research expertise. She also maintains public engagement through Chinese-language columns for BBC China and other media outlets, bridging academic and public discourse on contemporary Chinese culture.
Nima Mesgarani is an Associate Professor of Electrical Engineering at Columbia Engineering, Columbia University, affiliated with the Sense, Collect and Move Data Committee. His research bridges engineering and neuroscience through reverse-engineering neural signal processing mechanisms, leading to advancements in brain-machine interfaces, neural prosthetics, and speech processing algorithms. He received his PhD in Electrical Engineering from the University of Maryland and completed postdoctoral training at Johns Hopkins University's Center for Language and Speech Processing and UC San Francisco's Neurosurgery Department. Research Focus Professor Mesgarani's lab integrates computational neuroscience and engineering to study acoustic signal processing. Key areas include: Neural decoding of speech and auditory attention in multi-talker environments Development of brain-controlled hearing technologies Novel speech separation and synthesis algorithms inspired by cortical processing Cross-modal learning between auditory and visual systems Applications of large language models in neural signal interpretation Publication Trends Analysis of his 15 most recent articles (2025) reveals dominant themes: neural decoding techniques using intracranial EEG, brain-inspired speech separation models (e.g., Mamba architectures), applications of large language models in auditory neuroscience, cross-modal distillation methods, and clinical translation of audio processing algorithms. A strong emphasis emerges on real-time brain-computer interfaces and noise-robust speech processing. Laboratory and Collaborations Mesgarani directs an interdisciplinary lab developing neurotechnology for hearing restoration. His team collaborates with neurosurgery departments and speech processing centers, focusing on translating theoretical models into clinical brain-machine interfaces. The lab's work has yielded patents for brain-informed speech separation systems and attention-decoding frameworks.
Gordon Wetzstein is an Associate Professor of Electrical Engineering and, by courtesy, Computer Science at Stanford University. He leads the Stanford Computational Imaging Lab and co-directs the Stanford Center for Image Systems Engineering (SCIEN). His research focuses on computational imaging, wearable computing, and neural rendering, blending computer graphics, vision, AI, and optics. Education: Ph.D., Computer Science, University of British Columbia (2011) Diploma, Media Systems Science, Bauhaus University (2006) Research Interests: His work spans computational displays , holography , non-line-of-sight imaging , and AI-driven optical systems . Key projects include Autofocals (gaze-contingent eyeglasses) and neural holography systems. He explores applications in AR/VR, medical imaging, and scientific visualization. Publications: Recent work includes advances in 3D holography, gaze-tracking systems, and AI-optics integration. His papers address challenges in display efficiency, light-field processing, and real-time imaging. Awards: Fellow of Optica NSF CAREER Award (2016) PECASE (2019) ACM SIGGRAPH Significant New Researcher Award (2018) Advising & Grants: He advises over 20 doctoral and postdoctoral students. His lab collaborates with industry (e.g., Raxium, Google) and has secured grants from NSF, DARPA, and private foundations. Labs & Teams: His lab develops cutting-edge systems like neural holography and non-line-of-sight imaging. The SCIEN center fosters interdisciplinary image systems research.
Prof. Bernt Schiele is a Max Planck Director at the Max Planck Institute for Informatics and holds a Professorship at Saarland University. His research focuses on understanding multimodal sensor data, with key areas in computer vision, 3D object recognition, and machine learning. He leads the Computer Vision and Machine Learning group, addressing challenges in sensor fusion, scene understanding, and human activity recognition. Schiele has held academic roles at TU Darmstadt, ETH Zurich, and MIT, and contributes to top journals like IEEE Transactions on PAMI and conferences like ECCV. His work emphasizes robust models, interpretability, and domain adaptation for real-world applications. Education: PhD (1997, Grenoble), MSc (1994 Karlsruhe/1993 Grenoble) Key Positions: MIT (1997-2000), ETH Zurich (1999-2004), TU Darmstadt (2004-2010) Research interests span 3D scene understanding, multimodal sensor processing, and machine learning techniques for large-scale data. His recent work advances robust object detection, explainable AI, and domain-invariant training methods. He also chairs major conferences like ECCV 2018 and co-chairs ICCV 2011. Publications highlight innovations in interpretable vision transformers, certified explanations, and test-time adaptation. Despite no listed awards, his contributions shape foundational areas of computer vision and multimodal AI.
Professor Gabriel Brostow is a faculty member in the Department of Computer Science at University College London (UCL), where he leads research in Computer Vision and Human-Computer Interaction. He also serves as Chief Research Scientist and Senior Director of the R&D Team at Niantic, the company behind Pokémon GO. His work bridges academic research and industry applications, focusing on developing AI systems that enhance human capabilities through what he terms 'Human in the Loop AI'—now commonly referred to as Human-Centered AI. Brostow completed his BS in Electrical Engineering at UT Austin, followed by a PhD with Irfan Essa at Georgia Tech. He then pursued postdoctoral research with Roberto Cipolla's Computer Vision & Robotics Group at Cambridge University as a Marshall Sherfield Fellow, and with Marc Pollefeys in ETH Zurich's CVG Group. His research explores how AI, particularly Computer Vision, can serve as 'super-tools' for professionals across various domains including filmmaking, architecture, robotics, and scientific research. Specific interests include assistive technology for everyday life, authoring systems that maximize user effort, 3D reconstruction, depth estimation, and vision-language models. His work often involves creating systems that are validated through real-world human interaction to ensure practical utility. Analysis of his recent publications reveals a strong focus on practical applications of Computer Vision that directly interact with humans. His research spans 3D scene understanding, depth estimation, sketch-based interfaces, and multimodal AI systems. There's a clear emphasis on creating benchmarks and tools that facilitate human-AI collaboration, with applications in assistive technology, urban planning, filmmaking, and biodiversity monitoring. His work frequently appears at top conferences including CVPR, NeurIPS, ECCV, and CHI. Marshall Sherfield Fellowship Brostow actively mentors PhD students, with current advisees including Ross Murphy, Skanda Koppula, Gizem Unlu, Omiros Pantazis, and Jamie Watson. His alumni include numerous PhD graduates and MSc students who have gone on to successful careers in academia and industry. He emphasizes selecting students based on passion and potential rather than just academic credentials, valuing traits like helpfulness, drive, and hunger to learn. His research is supported through collaborations with major institutions and companies including DeepMind, MIT, and the University of Edinburgh. He leads a research group at UCL that collaborates closely with Niantic's R&D team, creating a unique bridge between academic research and industry application. His team's work frequently involves developing novel Computer Vision techniques that are validated through real-world human interaction, ensuring practical utility alongside technical innovation. The group explores blue-sky research problems with applications ranging from assistive technology to professional tools for filmmakers, architects, and scientists studying diverse environments.
Ellen Riloff serves as Department Head and Professor in the Department of Computer Science at the University of Arizona, where she leads research at the intersection of natural language processing (NLP) and artificial intelligence. Her work bridges theoretical advancements with real-world applications in social computing, planetary science, and crisis response systems. Education: Ph.D. in Computer Science, University of Massachusetts at Amherst (1994) Research Focus: Dr. Riloff specializes in affective computing and information extraction , developing techniques to recognize emotion, social cues, and embodied expressions in text. Her methodologies frequently employ bootstrapping, stacked learning, and semantic lexicon induction. Recent projects address crisis informatics (e.g., social cue recognition in emergencies) and interdisciplinary applications like the Mars Target Encyclopedia for planetary science data extraction. Publication Trends: Analysis of her 15 most recent publications (2021–2025) reveals three dominant trajectories: (1) affective event modeling in social contexts with applications to crisis response; (2) domain-specific NLP for planetary science and food systems; and (3) advanced language model techniques including retrieval-augmented generation and multi-view prompting. Her work increasingly integrates deep learning with traditional linguistic features. Grants and Leadership: Dr. Riloff has directed multiple NSF-funded projects, including RI: Small: Recognizing Implicit Personal States in Natural Language (2016) and RI: Small: Acquiring Domain Knowledge from Text through Cooperative Bootstrapping (2010). These initiatives pioneered bootstrapping frameworks for affective event recognition and information extraction. She also co-organized the Workshop on Pattern-based Approaches to NLP (2023), highlighting her leadership in advancing hybrid NLP methodologies. Collaborative Infrastructure: She co-developed the Mars Target Encyclopedia—a large-scale information extraction system that processes planetary science literature to create structured databases of Mars surface targets. This project demonstrates her commitment to building reusable scientific infrastructure through NLP.
Dr. George Stamou is a Professor at the School of Electrical and Computer Engineering of the National Technical University of Athens (NTUA), serving as Director of the Artificial Intelligence and Learning Systems Laboratory (AILS). His expertise spans knowledge representation, machine learning, neural networks, and semantic technologies. He leads interdisciplinary initiatives such as the postgraduate program 'Data Science and Machine Learning' (2018–2022). Research Interests: Focuses on knowledge graphs, interpretable AI, semantic web applications, and multimodal learning. His work integrates formal logic systems (e.g., description logics) with modern deep learning techniques, addressing challenges in explainability, bias detection, and ethical AI applications. Publications: Over 150 articles in AI journals/conferences with an h-index of 34 (Google Scholar). Notable contributions include datasets like CHORDONOMICON (music analysis), GOSt-MT (gender bias in MT), and methodologies for counterfactual explanations in machine learning. Awards & Committees: Active in W3C and RuleML standardization bodies. Co-organized major AI conferences. Recognized for contributions to semantic interoperability and knowledge-based systems. Labs & Teams: Directs AILS-NTUA lab and collaborates with CISRI (Computer & Information Systems Research Institute). Engages in EU projects like CultureLabs (cultural heritage digitalization) andsmarty4covid (health data analysis).
Kostas Bekris is a Professor in the Department of Computer Science at Rutgers University, specializing in Robotics and Artificial Intelligence. His research focuses on motion planning, autonomous manipulation, and robot control, with notable contributions to tensegrity robotics, perception-driven systems, and large-scale package handling. He leads a team conducting groundbreaking work in robotics, supported by grants from NSF, NASA, and industry collaborators like ExxonMobil. His group emphasizes interdisciplinary approaches, combining machine learning, topological methods, and differentiable physics modeling to advance robot capabilities in complex environments. Education details are not explicitly stated in the provided texts, but his academic career has included significant mentorship of PhD students and postdoctoral researchers. Key projects involve vision-driven manipulation pipelines, obstacle detection systems (PROBE), and resilient robot designs inspired by biological structures. He has been recognized for his work through prestigious awards including the NASA Early Career Grant and multiple NSF grants, as well as team achievements in robotics competitions like the Amazon Picking Challenge. Research interests span robotics subfields such as: Autonomous manipulation in cluttered environments Learning-based control for dynamic systems Topological data analysis for motion reasoning Tensegrity and soft robotics architectures Sim-to-real transfer in robotic tasks His team's work has produced open-source software tools and datasets, advancing benchmarks in manipulation and perception. Recent articles emphasize scalable solutions for industrial automation and robust navigation strategies in unstructured settings. Scientific achievements include: Development of PROBE for proprioceptive obstacle detection Advances in differentiable physics engines for tensegrity systems NSF-funded projects on robotic rearrangement and modular morphologies Advising contributions span over a decade, with current advisees focusing on topics like non-prehensile manipulation and large-scale storage optimization. Collaborations with industry (e.g., ExxonMobil) and academic partners (Yale University) reflect his commitment to applied robotics research. Labs and teams under his leadership include the Rutgers CS Robotics Group, contributing to projects like the ARIAC challenge platform and packing/industrial automation systems. Future work targets improved robot resilience in disaster scenarios and enhanced human-robot collaboration paradigms.