Jens Edlund is a Professor at KTH Royal Institute of Technology's Division of Speech, Music and Hearing. His research focuses on speech technology, dialogue systems, prosody, and evolutionary phonetics. He has contributed to foundational work on speech synthesis, conversational interaction, and multimodal corpora like the D64 corpus. Key projects include the MonAMI Reminder system and analysis of primate vocalizations to understand speech evolution. Edlund has collaborated extensively with global researchers, producing over 150 peer-reviewed works. His work integrates computational methods with linguistic and biological insights, emphasizing human-like dialogue systems and cross-species vocal analysis. Education: Ph.D. in Speech Technology (2011, KTH) Grants: Multiple EU and Swedish Research Council grants for speech technology and interdisciplinary studies Research labs include the KTH Speech, Music and Hearing Lab and collaborations with institutions like Max Planck Institute for Evolutionary Anthropology. Current work explores evolutionary origins of speech biomechanics and AI-driven speech synthesis evaluation.
Tzu-Mao Li is an Assistant Professor in the Department of Computer Science and Engineering (CSE) at the University of California, San Diego (UCSD), affiliated with the Center for Visual Computing. His research focuses on differentiable graphics algorithms, combining classical visual computing with modern machine learning techniques. He holds a Ph.D. from MIT CSAIL under Frédo Durand and a postdoc at MIT and UC Berkeley with Jonathan Ragan-Kelley. His work spans rendering, programming languages for graphics, Monte Carlo methods, and inverse problems. Education: B.S. and M.S. from National Taiwan University (2011-2013), advised by Yung-Yu Chuang. Ph.D. from MIT CSAIL (Computer Graphics Group), advised by Frédo Durand. Postdoctoral research at MIT and UC Berkeley with Jonathan Ragan-Kelley. Research Interests: Differentiable rendering, Monte Carlo integration, programming language design for visual computing, physical simulation, adversarial machine learning, and applications in computer vision and robotics. Key areas include rendering algorithms (path tracing, bidirectional methods), optimization techniques (MCMC, gradient-based), and neural representations (SDFs, neural fields). Publications focus on advancing rendering algorithms, differentiable systems, and applications in inverse problems. Notable contributions include edge sampling for differentiable rendering, warped-area sampling, and diffusion models for BSDF sampling. Awards: ACM SIGGRAPH 2020 Outstanding Doctoral Dissertation Award, multiple Best Paper Awards at SIGGRAPH, and oral presentations at ICCV. Teaching: Courses include CSE 167 (Computer Graphics), CSE 168 (Rendering), and CSE 272 (Advanced Image Synthesis), emphasizing physically-based methods and programming.
Prof. Matthias Nießner is a Professor at the Technical University of Munich , where he leads the Visual Computing Lab . Prior to this, he held a Visiting Assistant Professor position at Stanford University . His work bridges computer vision , graphics , and machine learning , focusing on 3D reconstruction , semantic scene understanding , and AI-driven video synthesis . Prof. Nießner has published over 150 works in top venues like SIGGRAPH , CVPR , and ECCV , with several receiving best paper awards (SIGCHI’14, HPG’15, SPG’18, SIGGRAPH’16 Emerging Tech). His research has garnered international media attention, including features in the New York Times , Wall Street Journal , and MIT Technological Review , as well as TV demonstrations (e.g., Jimmy Kimmel Live for Face2Face technology). Awards : TUM-IAS Rudolph Moessbauer Fellowship (2017–ongoing) Google Faculty Award (2017) Nvidia Professor Partnership Award (2018) ERC Starting Grant (2018, €1.5M) Eurographics Young Researcher Award (2019) Research Trends : 3D Gaussian Splatting for real-time rendering Neural Radiance Fields (NeRF) with mesh supervision Audio-driven facial animation via diffusion models Latent space diffusion for 3D scenes Self-supervised and zero-shot methods for 3D and image analysis As a co-founder and director of Synthesia Inc. , he drives democratization of synthetic media. His YouTube channel has over 5 million views, reflecting his impact beyond academia.
Haibin Ling is the SUNY Empire Innovation Professor in the Department of Computer Science at Stony Brook University, part of the College of Engineering and Applied Sciences. His research focuses on computer vision, medical image analysis, augmented reality, and AI applications in science. He holds a Ph.D. from the University of Maryland (2006) and prior degrees from Peking University. Previously, he worked at Temple University (2008–2019) and held roles at Siemens Corporate Research, UCLA, and Microsoft Research Asia. Professor Ling's work spans biomedical imaging, AI for science, and human-computer interaction. He leads the CV Lab and collaborates with the AI Institute at Stony Brook. Awards include the NSF CAREER Award (2014), Best Student Paper (ACM UIST 2003), and IEEE Fellow (2020). He serves on editorial boards for IEEE Trans. PAMI, Pattern Recognition, and CVIU, and chairs major conferences like CVPR. His research group includes over 50 students and alumni, with active projects in tracking benchmarks (LaSOT), Leafsnap, and medical imaging tools. Notable publications address OCTA flow estimation, backdoor attacks on vision models, and topology-guided medical learning. Collaborations involve institutions like Temple University and Stony Brook's Department of Applied Mathematics and Statistics.
David B. Lindell is an Assistant Professor in the Department of Computer Science at the University of Toronto, with affiliations to the Vector Institute and AXL. He is a founding member of the Toronto Computational Imaging Group. His research focuses on physically based intelligent sensing, integrating physical models, signal processing, and AI to advance sensing systems. Notable projects include imaging around corners, through scattering media, and developing machine learning algorithms for 3D scene reconstruction. Education: Ph.D. in Computational Imaging from Stanford University (advisor: Gordon Wetzstein). Awards include the 2024 Ontario Early Researcher Award and the Best Student Paper at CVPR 2025. His work combines computational imaging with applications in computer graphics and autonomous systems. Research interests span non-line-of-sight imaging, single-photon sensing, and neural representations. Key contributions include the Light-Cone Transform (Nature 2018), confocal diffuse tomography (Nature Communications 2020), and AutoInt (CVPR 2021). His lab develops systems for 3D reconstruction, transient imaging, and photon-efficient sensors. Selected grants and support: NSF CAREER Award, DARPA REVEAL program, and KAUST Visual Computing Center funding. Active collaborations with industry and academic institutions on autonomous driving and medical imaging applications.
Dr. Arno Solin is a tenured Associate Professor in Machine Learning at Aalto University's Department of Computer Science and an Academy of Finland Research Fellow . He leads the Aalto machine learning research group and serves as Director of the Finnish Doctoral Program Network in AI (AI-DOC) . His work bridges probabilistic modeling with practical applications in sensor fusion and real-time inference. ELLIS Scholar (European Laboratory for Learning and Intelligent Systems) Adjunct Professor at Tampere University Member of Young Academy Finland (2021–2025) Research Interests focus on data-efficient machine learning with probabilistic methods for real-time inference and sensor fusion. Key areas include Gaussian processes, diffusion models, stochastic differential equations, and uncertainty quantification in deep learning. His group develops methods that combine structural constraints with adaptive learning for deployment on resource-limited hardware. Publication Trends show consistent output in top venues (NeurIPS, ICML, ICLR, AISTATS) with emphasis on diffusion models , 3D scene reconstruction , and real-time probabilistic modeling . Recent works explore physics-informed learning, heterophily-aware graph models, and compressed representations for world models in reinforcement learning. Scientific Recognition : Awarded AI Researcher of the Year 2024 by AI Finland Teacher of the Year 2023 at Aalto CS ISIF Jean-Pierre Le Cadre Best Paper Award (2018) MLSP Schizophrenia Classification Challenge Winner (2014) NeurIPS/ICML Reviewer Awards Research Leadership includes coordinating Finland's AI Center of Excellence program and directing the Finnish Doctoral Program Network in AI (AI-DOC). He supervises 15+ doctoral/postdoctoral researchers and has spun off Spectacular AI , a company commercializing sensor fusion technology.
Cyrill Stachniss is a Full Professor at the University of Bonn , where he heads the Lab for Photogrammetry and Robotics and is affiliated with the Lamarr Institute for Machine Learning and Artificial Intelligence . He was previously a Visiting Professor in Engineering at the University of Oxford until 2025. His academic journey includes positions at the University of Freiburg, University of Zaragoza, and the Swiss Federal Institute of Technology. University: University of Bonn School: Faculty of Engineering Department: Department of Photogrammetry Academic Rank: Professor His research spans robotics, photogrammetry, SLAM, autonomous navigation, perception systems, agricultural robotics, and unmanned aerial vehicles . He emphasizes probabilistic techniques for mobile robots and has made significant contributions to visual and LiDAR-based localization, scene understanding, and 3D reconstruction. The recent publications highlight a strong trend toward neural implicit representations, agricultural phenotyping, radar-based perception, and active learning . His team develops robust systems for real-world deployment in dynamic and unstructured environments, particularly in precision farming and autonomous vehicles. Scientific Awards: IEEE RAS Early Career Award (2013) Microsoft Research Faculty Fellow (2010) 7th EURON Georges Giralt Award (2008) Multiple Best Paper Awards at ICRA, IROS, RSS, and RAL Faculty Teaching Award, University of Freiburg He has advised numerous students and leads the DFG Cluster of Excellence PhenoRob and Research Unit FOR 1505 Mapping on Demand . His lab has co-founded three startups, reflecting strong industry and societal impact. He also runs the educational video series 5 Minutes with Cyrill , explaining key robotics concepts.
Ravi Ramamoorthi is the Ronald L. Graham Professor of Computer Science and Director of the UC San Diego Center for Visual Computing. He holds a faculty position in the Department of Computer Science and Engineering (CSE) and is an affiliate of the Department of Electrical and Computer Engineering (ECE). He joined UC San Diego in 2014, previously at UC Berkeley and Columbia University. He also holds a part-time appointment as a Distinguished Research Scientist at NVIDIA. His research focuses on visual computing, including rendering, computer vision, light field cameras, and physics-based modeling. Notable contributions include foundational work on spherical harmonic lighting, neural radiance fields (NeRF), and Monte Carlo rendering techniques. His work bridges graphics, vision, and signal processing with applications in sparse reconstruction, importance sampling, and real-time rendering. He teaches courses like CSE 167 (Computer Graphics), CSE 168 (Rendering), and advanced topics in computer graphics. Awards include ACM and IEEE Fellowships, the Okawa Foundation Grant, and multiple Frontiers of Science Awards. His research is supported by NSF, ONR, and industry collaborators including Adobe, Sony, and Qualcomm. Key projects include the Center for Visual Computing, Light Field research, and educational initiatives like edX MOOCs on computer graphics and rendering. His work has influenced industry tools (e.g., Pixar, RenderMan) and modern real-time rendering pipelines with denoising techniques.
Jonathan T. Barron is a Researcher at Google DeepMind in San Francisco, specializing in Computer Vision , Neural Rendering , and 3D Scene Reconstruction . He earned his PhD at UC Berkeley under Jitendra Malik and has pioneered advancements in NeRF (Neural Radiance Fields) and diffusion-based 3D generation. Research Interests : Computer Vision, Deep Learning, Generative AI, Image Processing, and 3D Reconstruction via Radiance Fields. His work includes Bolt3D for rapid 3D scene generation, CAT3D/CAT4D for text-to-3D/4D, and Zip-NeRF for anti-aliased radiance fields. He has also developed real-time rendering frameworks like SMERF and NeRF-Casting for reflections. Scientific awards: PAMI Young Researcher Award He has served as Area Chair for CVPR, ICCV, and NeurIPS, and his research is widely adopted in applications like Google's Lens Blur , Portrait Mode , and Jump VR .
Andreas Geiger is a Professor and Head of the Department of Computer Science at the University of Tübingen, Germany. He leads the Autonomous Vision Group (AVG) within CyberValley and is a core faculty member of the Tübingen AI Center. His roles also include PI in the ML in Science Excellence Cluster and the CRC Robust Vision, as well as ELLIS Fellow and coordinator of the ELLIS PhD program. He specializes in machine learning models for computer vision, robotics, and autonomous systems, with applications in self-driving cars, VR/AR, and scientific document analysis. Educational background: While not explicitly detailed, his positions imply a Ph.D. in Computer Science or related field. His work spans interdisciplinary collaborations with institutions like ETH Zürich, Microsoft, and the University of Bonn. Research focuses on 3D scene understanding, Gaussian splatting, generative models, and reliable autonomous systems. Notable contributions include the KITTI dataset and foundational work in neural radiance fields. Awards include the Sage 10-Year Impact Award (2024), ERC Starting Grant (2019), and IEEE PAMI Young Researcher Award (2018). Key projects include the Scholar Inbox paper recommender platform, ReSim (reliable world simulation), and advancements in 3D scene generation (e.g., UrbanCAD, PrITTI). His lab maintains a strong focus on open-source tools and datasets, such as the CARLA Route Generator. Grants and funding include support from Vector Stiftung (MINT innovation program) and EU initiatives like the ML in Science Cluster. His team collaborates internationally, with recent work presented at CVPR, SIGGRAPH, and NeurIPS.
David Lindlbauer is an Assistant Professor at the Human-Computer Interaction Institute (HCII) of Carnegie Mellon University, where he leads the Augmented Perception Lab and co-directs the CMU Extended Reality Technology Center. His research bridges human perception, extended reality (AR/VR), and computational interaction techniques, focusing on developing systems that dynamically adapt interface elements based on environmental context, user cognition, and task requirements. He completed his PhD at TU Berlin under Prof. Marc Alexa and held a postdoctoral position at ETH Zurich's Advanced Interactive Technologies lab. His work has been published extensively at top venues including ACM CHI, UIST, and IEEE VR, with research themes spanning gaze tracking, spatial audio optimization, haptic feedback, and multimodal notification systems. Media outlets like MIT Technology Review and Fast Company Design have featured his innovations. Dr. Lindlbauer has received prestigious grants from Meta, NSF, and ETH Zurich, and serves on program committees for CHI, UIST, and ISMAR. He has been recognized with Best Paper awards at ISS 2023 and CHI 2016, and his lab develops tools like MineXR for personalized XR interfaces and RealityReplay for temporal change visualization in mixed reality environments.
Angel Xuan Chang is an Associate Professor at Simon Fraser University's School of Computing Science, affiliated with labs including 3DLG, GrUVi, SFU NatLang, SFU AI/ML, and VINCI. He holds a Canada CIFAR AI Chair and was a TUM-IAS Hans Fischer Fellow (2018-2022). His research bridges natural language processing (NLP), 3D scene understanding, and embodied AI, focusing on language-grounded 3D generation and biodiversity monitoring via DNA barcodes. Recent work includes NuiScene (unbounded outdoor scene generation), ViGiL3D (3D visual grounding dataset), and CLIBD (vision-genomics biodiversity analysis). He advises students in projects like BIOSCAN-5M insect dataset and embodied AI navigation. His 2025 highlights include multiple ICCV and ICLR papers, workshops at ICML and CVPR, and a CRV invited talk. Education: Ph.D. in Computer Science from Stanford University (2014), advised by Chris Manning. Previous roles include visiting research scientist at Facebook AI Research and researcher at Eloquent Labs.
Freda Shi is an Assistant Professor at the David R. Cheriton School of Computer Science at the University of Waterloo and a Faculty Member at the Vector Institute, where she holds a Canada CIFAR AI Chair. She joined the University of Waterloo in July 2024 after completing her Ph.D. at the Toyota Technological Institute at Chicago. Educational Background: Ph.D. in Computer Science, Toyota Technological Institute at Chicago (2024), advised by Professors Karen Livescu and Kevin Gimpel Bachelor's degree in Intelligence Science and Technology (Computer Science Track) with a minor in Sociology, Peking University (2018) Dr. Shi's research focuses on computational linguistics and natural language processing, particularly on deeper understandings of natural language and the human language processing mechanism. She is especially interested in learning language through grounding, computational multilingualism, and related machine learning aspects. Her work aims to inform the design of more efficient, effective, safe, and trustworthy NLP systems. She leads the CompLING Lab at the University of Waterloo, which investigates how language models process spatial relationships and acquire linguistic structures through grounded experiences. Her publication record shows a consistent trajectory of high-impact research, with recent work focusing on spatial reasoning in vision-language models, multilingual chain-of-thought capabilities, and grounded language acquisition. She has published in top-tier conferences including ACL, EMNLP, ICLR, and NAACL, with several papers receiving notable recognition including Best Paper Nominee status at multiple venues. Her research bridges theoretical linguistics with practical NLP applications, demonstrating how linguistic insights can improve AI systems. Scientific Recognition: Canada CIFAR AI Chair (2024) Google Ph.D. Fellowship Thesis of Distinction for her doctoral work Multiple Best Paper Nominee awards at major NLP conferences Dr. Shi teaches CS 784: Computational Linguistics and CS 486/686: Introduction to Artificial Intelligence at the University of Waterloo. She actively contributes to the NLP research community through conference participation, program committee service, and collaborative projects. Her research has significant implications for creating more robust, human-like language understanding systems and advancing the field of grounded language learning in artificial intelligence.
Anand Bhojan is an Associate Professor (Educator Track) at the Department of Computer Science, School of Computing, National University of Singapore (NUS). He is a member of the Communication and Internet Research Lab and serves on the Graduate Studies Committee. Dr. Bhojan is also the founder of Anuflora Systems and Virtual and Augmented Reality Labs (www.varlabs.org), and serves as Associate Editor of Computers and Electrical Engineering Journal, Elsevier, and Vice President of International Researchers Club, Singapore. Dr. Bhojan earned his Ph.D. in Computer Science & Engineering from NUS in 2011, where his thesis was nominated for the Best PhD Thesis Award. He also holds a Professional Master's in Computer Applications from Bharathidasan University (1999), a Bachelor's degree in Computing with Gold Medal (University topper) from Bharathiar University (1994), and a Postgraduate Certificate in Teaching Higher Education from University of Sheffield, UK (2003). His research spans multiple domains including Distributed Computing and Wireless Networks (IoT, Security, Blockchain), Artificial Intelligence (Generative AI, FinTech), and Entertainment Computing with focus on Games, VR/AR/Metaverse technologies. Dr. Bhojan leads the Metaverse Foundry research group which focuses on content generation for games & XR simulations across multiple domains including entertainment, healthcare, and architecture, while also experimenting with innovative teaching methods for entertainment media technologies. Dr. Bhojan's recent research has pioneered hybrid rendering techniques that strategically combine ray tracing and rasterization to create more realistic video game graphics without compromising performance. His work addresses critical challenges in real-time rendering, particularly in depth of field and motion blur effects, with the goal of making Hollywood-quality graphics accessible on a wider range of hardware. Earlier work focused on energy efficiency in mobile gaming, including power management techniques and latency optimization for cloud gaming. Among his notable achievements: 2012 Nominated for Best PhD Thesis Award (Wang Gungwu Medal & Prize), NUS 2011 Dean's Graduate Research Achievement Award (PhD), SoC, NUS 2006 Best R&D Project award from TOTE Board, Singapore 1995 Gold medal for first Rank (out of 4000) in Computing, Bharathiar University Dr. Bhojan has served as Organizing Chair and Program Chair for multiple international conferences and has delivered keynote talks at IEEE/ACM International Conferences. He teaches courses including Computer Networks Practice, Game Development, Interaction Design for Virtual and Augmented Reality, and Game Development Project. He leads the Metaverse Foundry research group and the Virtual and Augmented Reality Labs, where undergraduate and graduate students have won multiple research and innovation awards. His research bridges entertainment, education, and emerging technologies with practical applications in making immersive experiences more accessible across different hardware capabilities while maintaining energy efficiency.
Georgia Gkioxari is an Assistant Professor in the Division of Computing and Mathematical Sciences at Caltech , with a part-time affiliation at Meta AI . Her work focuses on extending visual perception models through advanced 2D and 3D representation learning, spatial reasoning, and generative models. Education: Not explicitly mentioned in the text Research interests span 3D perception , spatial reasoning , and vision-language integration , with projects like Visual Agentic AI for Spatial Reasoning and Token-by-Token Multimodal Alignment . Her publications emphasize 3D object detection , reconstruction , and generative modeling techniques including diffusion models and transformers . Scientific recognition includes the Meta LLM Evaluation Research Grant , Okawa Research Grant , Google Faculty Scholar Award 2024 , and Amazon Research Award . She teaches courses like Large Language & Vision Models (EE/CS 148) and Learning & 3D (CS 101) at Caltech. Labs & Teams: Leads Glab with members including Ilona Demler, Ziqi Ma, and Damiano Marsili