Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Christian Theobalt is a Professor of Computer Science at Saarland University and Scientific Director of the Visual Computing and Artificial Intelligence Department at the Max Planck Institute for Informatics . He leads the Saarbruecken Center for Visual Computing as a strategic partnership between Google and MPI. PhD in Computer Science (2005) from MPI-INF/Saarland University Postdoctoral Researcher at MPI (2005-2007) Visiting Assistant Professor at Stanford (2007-2009) His research focuses on the intersection of Computer Graphics, Computer Vision, and Artificial Intelligence , with specializations in: 3D/4D Human Reconstruction Neural Rendering Performance Capture Geometric Deep Learning Volumetric Video Quantum Visual Computing Recent article trends show emphasis on: Quantum computing applications in visual reconstruction Neural rendering with radiance fields Human motion capture from egocentric views Gesture-language interaction modeling Multi-modal scene understanding Real-time free-viewpoint rendering Scientific Awards : Fellow of EUROGRAPHICS (2022) CVPR Best Student Paper Honorable Mention (2020) ERC Consolidator Grant (2017) Karl Heinz Beckurts Award (2017) Busy Beaver Teaching Award (2016) ERC Starting Grant (2013) Advising over 50 students and researchers including: Current researchers: Viktor Rudnev, Linjie Lyu, Mohit Mendiratta Postdocs: Kwang In Kim, Kiran Varanasi Alumni: Franziska Mueller (Google), Dushyant Mehta (Qualcomm), Ayush Tewari (MIT) Grants include ERC grants, Google Glass Research Award, and multiple industry partnerships. His lab maintains cutting-edge facilities with: Multi-camera capture systems Quantum annealing infrastructure HDR radiance field technology GPU clusters for AI research Time-of-flight imaging systems Event camera arrays
Olga Sorkine-Hornung is a Professor of Computer Science at ETH Zurich and head of the Institute of Visual Computing. She leads the Interactive Geometry Lab, focusing on theoretical and practical advancements in digital content creation, geometry processing, and shape modeling. Current position: ETH Zurich, Department of Computer Science Previous roles: Courant Institute (NYU), Technical University of Berlin Education: BSc and PhD from Tel Aviv University, postdoc at TU Berlin Her research spans shape representation, digital fabrication, computer animation, and fundamental geometry processing. Key contributions include Laplacian surface editing, as-rigid-as-possible deformation, and generalized winding numbers. She works on applications in VR, AR, and autonomous systems. Awards include Test of Time Awards (2024), ACM Fellow (2020), ERC Consolidator Grant (2020), and EUROGRAPHICS Young Researcher Award (2008). She has supervised numerous students and co-developed software libraries like libigl and Instant Meshes . Co-chair roles for SIGGRAPH, Eurographics, and Pacific Graphics Editorial board member for ACM Transactions on Graphics and other journals Keynote speaker at VMV, CVPR, and SIAM conferences Her work bridges mathematical rigor with practical implementation, advancing computer graphics and geometry processing through intuitive algorithms that maintain surface detail while enabling efficient computation.
Prof. Matthias Nießner is a Professor at the Technical University of Munich, leading the Visual Computing Lab. His research intersects computer graphics, vision, and AI, focusing on 3D reconstruction, semantic understanding, and AI-driven video synthesis. He holds a PhD from the University of Erlangen-Nuremberg (2013) and was a Visiting Assistant Professor at Stanford University (2013–2017). Notable awards include the ERC Starting Grant (2018), Nvidia Professorship Award, and Eurographics Young Researcher Award (2019). His work has been featured in mainstream media and led to startups like Synthesia Inc. Research spans Gaussian splatting, neural radiance fields, and generative AI for 3D avatars. Over 150 publications include SIGGRAPH, CVPR, and ECCV, with best paper awards. Projects like Face2Face and ScanNet have driven innovation in facial reenactment and 3D scene datasets. Education: PhD in Computer Science, University of Erlangen-Nuremberg (2013) Diploma in Computer Science, University of Erlangen-Nuremberg (2010) Research Interests: 3D digitization, neural rendering, generative AI, non-rigid reconstruction, and applications in AR/VR. Awards: ERC Starting Grant (2018) Nvidia Professorship Award (2018) Google Faculty Award (2018) SIGGRAPH Best Emerging Tech Award (2016) Grants: Over €1.5M from ERC and industry partnerships. Labs/Teams: Visual Computing Lab at TUM and Synthesia Inc. (co-founder). Key projects include ScanNet (large 3D indoor dataset), Face2Face (real-time facial reenactment), and Gaussian-based 3D avatars. Current work focuses on diffusion models, neural radiance fields, and AI-generated media detection.
Angjoo Kanazawa is an Assistant Professor in the Department of Electrical Engineering and Computer Sciences (EECS) at the University of California, Berkeley. She leads the Kanazawa AI Research (KAIR) lab under the Berkeley Artificial Intelligence Research (BAIR) umbrella and serves on the advisory board of Wonder Dynamics. Her research focuses on the intersection of computer vision, computer graphics, and machine learning, with a particular emphasis on 4D reconstruction of dynamic scenes, neural radiance fields (NeRF), and systems that model human-environment interactions from 2D visual data. Education : Ph.D., Computer Science (2017), University of Maryland, College Park BA, Mathematics and Computer Science (2012), New York University (NYU) Her work aims to build systems that can capture, perceive, and understand complex 3D/4D worlds from photographs and videos, enabling applications in scene reconstruction, motion analysis, and generative modeling. She has pioneered techniques for scaling NeRFs across GPUs (NeRF-XL), developing open-source tools like nerfstudio and gsplat. Her recent publications focus on topics like self-occluded avatar recovery (SOAR), decentralized diffusion models, and 4D reconstruction of articulated objects for robotics. Kanazawa's research has been recognized with prestigious awards including the IEEE CS TCPAMI Young Researcher Award (2024) , Sloan Research Fellowship (2023) , and Google Faculty Research Award (2021) . Her lab has trained numerous students who now hold positions at leading institutions and companies like Anthropic, Meta Reality Labs, and Luma AI. Key Scientific Awards : IEEE CS TCPAMI Young Researcher Award (2024) Sloan Research Fellow (2023) Hellman Fellow (2022) Bakar Fellows Spark Award (2022) Google Faculty Research Award (2021) Her KAIR lab collaborates extensively with industry partners and academic institutions, including the Max Planck Institute and Google Research. She has served as an advisor for PhD students and postdocs who now lead teams at UC Berkeley, MIT, Stanford, and Luma AI, while her teaching includes graduate courses like CS 280A (Computer Vision) and CS 294-173 (Learning for 3D Vision).
Prof. Matthias Nießner is a Professor at the Technical University of Munich , where he leads the Visual Computing Lab . Prior to this, he held a Visiting Assistant Professor position at Stanford University . His work bridges computer vision , graphics , and machine learning , focusing on 3D reconstruction , semantic scene understanding , and AI-driven video synthesis . Prof. Nießner has published over 150 works in top venues like SIGGRAPH , CVPR , and ECCV , with several receiving best paper awards (SIGCHI’14, HPG’15, SPG’18, SIGGRAPH’16 Emerging Tech). His research has garnered international media attention, including features in the New York Times , Wall Street Journal , and MIT Technological Review , as well as TV demonstrations (e.g., Jimmy Kimmel Live for Face2Face technology). Awards : TUM-IAS Rudolph Moessbauer Fellowship (2017–ongoing) Google Faculty Award (2017) Nvidia Professor Partnership Award (2018) ERC Starting Grant (2018, €1.5M) Eurographics Young Researcher Award (2019) Research Trends : 3D Gaussian Splatting for real-time rendering Neural Radiance Fields (NeRF) with mesh supervision Audio-driven facial animation via diffusion models Latent space diffusion for 3D scenes Self-supervised and zero-shot methods for 3D and image analysis As a co-founder and director of Synthesia Inc. , he drives democratization of synthetic media. His YouTube channel has over 5 million views, reflecting his impact beyond academia.
Prof. Dr. Otmar Hilliges is a Full Professor at the Department of Computer Science at ETH Zurich. He leads the AIT lab and serves as the head of the Institute of Intelligent Interactive Systems. His research focuses on spatio-temporal understanding of human movement and interaction, leveraging algorithms and representations from videos, images, and sensor data for applications in Augmented Reality (AR), Virtual Reality (VR), and Human-Robot Interaction. Education: Diplom (MSc) in Computer Science, Technical University of Munich (TUM), Germany PhD in Computer Science, Ludwig Maximilian University of Munich (LMU), Germany (2009) Research Interests: Hilliges' work spans computer vision, robotics, and human-computer interaction. He develops methods for 3D human pose estimation, generative models for realistic avatar creation, and physically plausible simulation of human-object interactions. His research emphasizes practical applications in AR/VR and assistive robotics, aiming to bridge the gap between perception and action. Grants & Contributions: ERC Consolidator Grant (2022-2027): 'AI-Perceive: Robust Human-Centric Computer Vision for Advanced AI-Agents' Google Research Agreement (2020-2025): 'Generative Modelling of Humans' Microsoft Research Grants: Focus on human-centric robotics and interactive technologies Labs & Teams: Leads the AIT Lab at ETH Zurich, which pioneers research in intelligent interactive systems, emphasizing human-centric AI and robotics. The lab collaborates on projects ranging from drone cinematography to haptic feedback systems.
Professor Gabriel Brostow is a faculty member in the Department of Computer Science at University College London (UCL), where he leads research in Computer Vision and Human-Computer Interaction. He also serves as Chief Research Scientist and Senior Director of the R&D Team at Niantic, the company behind Pokémon GO. His work bridges academic research and industry applications, focusing on developing AI systems that enhance human capabilities through what he terms 'Human in the Loop AI'—now commonly referred to as Human-Centered AI. Brostow completed his BS in Electrical Engineering at UT Austin, followed by a PhD with Irfan Essa at Georgia Tech. He then pursued postdoctoral research with Roberto Cipolla's Computer Vision & Robotics Group at Cambridge University as a Marshall Sherfield Fellow, and with Marc Pollefeys in ETH Zurich's CVG Group. His research explores how AI, particularly Computer Vision, can serve as 'super-tools' for professionals across various domains including filmmaking, architecture, robotics, and scientific research. Specific interests include assistive technology for everyday life, authoring systems that maximize user effort, 3D reconstruction, depth estimation, and vision-language models. His work often involves creating systems that are validated through real-world human interaction to ensure practical utility. Analysis of his recent publications reveals a strong focus on practical applications of Computer Vision that directly interact with humans. His research spans 3D scene understanding, depth estimation, sketch-based interfaces, and multimodal AI systems. There's a clear emphasis on creating benchmarks and tools that facilitate human-AI collaboration, with applications in assistive technology, urban planning, filmmaking, and biodiversity monitoring. His work frequently appears at top conferences including CVPR, NeurIPS, ECCV, and CHI. Marshall Sherfield Fellowship Brostow actively mentors PhD students, with current advisees including Ross Murphy, Skanda Koppula, Gizem Unlu, Omiros Pantazis, and Jamie Watson. His alumni include numerous PhD graduates and MSc students who have gone on to successful careers in academia and industry. He emphasizes selecting students based on passion and potential rather than just academic credentials, valuing traits like helpfulness, drive, and hunger to learn. His research is supported through collaborations with major institutions and companies including DeepMind, MIT, and the University of Edinburgh. He leads a research group at UCL that collaborates closely with Niantic's R&D team, creating a unique bridge between academic research and industry application. His team's work frequently involves developing novel Computer Vision techniques that are validated through real-world human interaction, ensuring practical utility alongside technical innovation. The group explores blue-sky research problems with applications ranging from assistive technology to professional tools for filmmakers, architects, and scientists studying diverse environments.
Michael J. Black is a Professor and Honorarprofessor at the University of Tübingen's Faculty of Science, Department of Computer Science, and a founding Director of the Max Planck Institute for Intelligent Systems, leading the Perceiving Systems department. He holds a B.Sc. from the University of British Columbia (1985), M.S. from Stanford (1989), and Ph.D. in Computer Science from Yale (1992). His research focuses on computer vision, 3D human modeling, motion capture, and AI-driven digital humans. Key contributions include the SMPL body model, optical flow algorithms, and datasets like Middlebury Flow and Sintel. He has received major awards such as the PAMI Distinguished Researcher Award, multiple Koenderink and Longuet-Higgins Prizes, and is a member of the German National Academy of Sciences Leopoldina and Royal Swedish Academy of Sciences. His commercial ventures include co-founding Body Labs (acquired by Amazon) and Meshcapade, advancing 3D human generation and interaction technologies. Recent work includes markerless motion capture systems (e.g., MAMMA, PICO), 3D hair and garment synthesis, and AI tools like ChatHuman for 3D human interaction analysis. His research bridges vision, graphics, and robotics, with applications in animation, healthcare, and robotics.
Michael J. Black is a Professor and Director at the Max Planck Institute for Intelligent Systems in Tübingen, Germany, where he leads the Perceiving Systems department and serves as Managing Director . He is also an Honorarprofessor at the University of Tübingen 's Faculty of Science . His career spans roles at Brown University (2000-2010), Xerox PARC, and academic-industry collaborations with Amazon and Meshcapade.
Devi Parikh is an Associate Professor at the School of Interactive Computing, Georgia Institute of Technology, and a Research Director at Meta’s FAIR lab. Her research focuses on generative models, AI for creativity, computer vision, and natural language processing. Education: B.S. in Electrical and Computer Engineering from Rowan University (2005), M.S. and Ph.D. in Electrical and Computer Engineering from Carnegie Mellon University (2007, 2009). Research interests include embodied AI, human-AI collaboration, and creative applications of AI. She has held visiting positions at Cornell, MIT, CMU, and others. Awards include NSF CAREER Award, IJCAI Computers and Thought Award, and multiple fellowships. Led development of Habitat , a platform for embodied AI research, and contributed to the Open Catalyst Project for renewable energy storage.
Michael Riegler is a Researcher at the AI Department, Simula Research Laboratory , focusing on interdisciplinary applications of Artificial Intelligence in healthcare, sports analytics, and multimedia systems. His work bridges Machine Learning , AI Alignment , and Applied AI across clinical and real-world domains. Key Affiliations: Simula Research Laboratory (AI Department Head) Research Themes: Explainable AI in medicine, multimodal data analysis, and AI-driven health monitoring Research Interests include: Developing AI/ML algorithms for medical imaging (e.g., polyp detection, embryo analysis) Addressing missing data challenges in healthcare through novel imputation techniques Creating multimodal virtual avatars for investigative interview training Designing edge AI systems for sports analytics and sustainable fishing Recent Publications highlight collaborations with institutions in Norway and globally, with a focus on: Medical Applications: Polyp segmentation, ECG analysis, and explainable models for disease detection Sports Analytics: Athlete performance prediction and soccer video processing Data Infrastructure: Lifelogging datasets (ScopeSense), semantic representation frameworks Labs & Teams include leadership in Simula’s AI Department and participation in projects like Medico Multimedia Task , ImageCLEF , and MediaEval workshops. His work emphasizes responsible AI innovation in public sectors and privacy-preserving systems for edge environments.
Hanbyul Joo is an Assistant Professor in the Department of Computer Science and Engineering at Seoul National University (SNU). Prior to joining SNU, he was a Research Scientist at Facebook AI Research (FAIR) in Menlo Park. He completed his Ph.D. in the Robotics Institute at Carnegie Mellon University, working with Yaser Sheikh, and received his M.S. in Electrical Engineering and B.S. in Computer Science from KAIST, Korea. Dr. Joo's educational journey began at KAIST, where he earned both his Bachelor's and Master's degrees. He then pursued his Ph.D. at Carnegie Mellon University's Robotics Institute, completing his dissertation titled "Sensing, Measuring, and Modeling Social Signals in Nonverbal Communication." His doctoral work focused on developing the Panoptic Studio, a unique sensing system with over 500 synchronized cameras for capturing social interactions. Dr. Joo's research primarily focuses on endowing machines and robots with the ability to perceive and understand human behaviors in 3D . His goal is to build "social Artificial Intelligence" that can interact with humans using social signals (body languages). He pursues this direction using data-driven methods where data is collected by measuring the wide spectrum of social signals transmitted during interpersonal social interaction. His research spans computer vision, machine learning, computer graphics, and robotics , with particular emphasis on 3D human pose estimation, human-object interaction, and social signal processing. His recent publications demonstrate a clear trend toward leveraging diffusion models for 3D reconstruction and generation tasks, with a focus on human-centric applications. His work bridges the gap between 2D image understanding and 3D scene reconstruction, often utilizing pre-trained models to overcome data limitations. The research consistently addresses fundamental challenges in understanding human behavior, interaction with objects, and social dynamics in 3D space. Dr. Joo is a recipient of several prestigious awards including the Samsung Scholarship and the CVPR Best Student Paper Award in 2018 . His paper "Total Capture: A 3D Deformation Model for Tracking Faces, Hands, and Bodies" received this honor at CVPR 2018. His research has been widely recognized in top computer vision and AI conferences, with multiple oral presentations at venues like CVPR, ICCV, and ECCV. Dr. Joo actively mentors a large group of students, with approximately 15 current students working toward MS/PhD degrees under his supervision. His lab, the SNU VCLab, focuses on cutting-edge research in computer vision and graphics. He has secured significant research funding through his work, though specific grant details aren't provided on his website. Dr. Joo frequently serves as an area chair for major conferences including CVPR, ICCV, and NeurIPS, demonstrating his standing in the academic community. Dr. Joo leads the SNU VCLab, which has developed several notable datasets and tools including SNU ParaHome, FrankMocap, and the CMU Panoptic Studio Dataset. His lab maintains strong industry connections, with students interning at leading companies like Meta. The lab's research focuses on building the infrastructure and algorithms needed for social AI, with an emphasis on practical applications that can be deployed in real-world settings.
Fernando De la Torre is a Courtesy Professor in the Department of Electrical and Computer Engineering at Carnegie Mellon University, with an affiliation to the Robotics Institute where he has been a research faculty member since 2005. He holds a Ph.D. in Electronic Engineering from Ramon Llull University (2002). His research focuses on machine learning and computer vision, with applications in human health, augmented/virtual reality, generative models, and data-centric methodologies. He directs the Human Sensing Laboratory, which explores technologies for human behavior analysis and health monitoring. Notable contributions include founding FacioMetrics LLC (acquired by Meta), advancing facial recognition and 3D human digitization, and developing frameworks for robust visual models. His work bridges theory and practice, with over 225 peer-reviewed publications and editorial roles, including Associate Editor for IEEE Transactions on Pattern Analysis and Machine Intelligence. Recent research trends emphasize generative AI applications in satellite imagery analysis, VR/AR rendering optimizations (e.g., Gaussian splatting), and clinical motion recognition for healthcare. His projects often intersect with industry, addressing challenges in wearable health monitoring and immersive technologies. His lab collaborations span academia and industry, focusing on scalable solutions for 3D human modeling, adversarial robustness, and multimodal data fusion. Key achievements include pioneering work on 3D face animation from speech and garment reconstruction from single images.
Siyu Tang is an Assistant Professor in the Department of Computer Science at ETH Zürich, where she leads the Computer Vision and Learning Group (VLG) at the Institute of Visual Computing. Her research focuses on computational models for human perception and digitalization through computer vision and machine learning. Her educational background includes: PhD in Computer Science, Max Planck Institute for Informatics (2017), supervised by Prof. Bernt Schiele Master of Science in Media Informatics, RWTH Aachen University Bachelor of Science in Computer Science, Zhejiang University, China Dr. Tang specializes in human-centric computer vision, developing statistical models for motion analysis, pose estimation, and digital human creation. Her work integrates machine learning with optimization techniques to enable machines to interpret human activities from visual data, with applications spanning virtual reality, healthcare, and human-computer interaction. Key research thrusts include generative models for content creation, egocentric vision, and human motion synthesis. Her recent publications (2024-2025) demonstrate intense focus on 3D human modeling and neural rendering, with Gaussian splatting emerging as a dominant technique for efficient avatar creation and scene reconstruction. Significant themes include text-driven motion synthesis using diffusion models, relightable avatars, surgical training applications, and egocentric multimodal pretraining. This work bridges computer vision, graphics, and machine learning to advance human digitalization. No scientific awards were mentioned in the provided text. Dr. Tang leads the VLG research group at ETH Zürich, mentoring PhD and Master's students in human-centric AI. She previously secured an early career research grant from the Max Planck Institute for Intelligent Systems to establish her independent research program. Her group actively pursues funding for projects in human motion analysis, 3D reconstruction, and generative modeling, with strong industry and clinical collaborations. The Computer Vision and Learning Group (VLG) operates within ETH's Institute of Visual Computing, maintaining dedicated facilities for motion capture, 3D scanning, and high-performance computing. The team collaborates internationally with institutions like the Max Planck Society and focuses on scalable solutions for real-world human digitalization challenges, including surgical training systems and immersive virtual environments.