Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Jia Deng is a Professor of Computer Science at Princeton University and directs the Princeton Vision & Learning Lab. His research focuses on computer vision, machine learning, and robotics, with an emphasis on advancing 3D vision and synthetic data generation. Ph.D., Princeton University, 2012 B.Eng., Tsinghua University, Computer Science His work spans optical flow, depth estimation, and visual reasoning, leveraging procedural scene generation and robust neural architectures. Recent publications highlight advancements in multi-layer depth estimation, stereo matching, and simulation environments for embodied AI. Alfred P. Sloan Research Fellowship, 2018 NSF CAREER Award, 2020 ONR Young Investigator Award, 2020 Multiple Best Paper Awards (ECCV, ICCV, 3DV) Deng leads the Princeton Vision & Learning Lab, which develops foundational tools for computer vision and machine learning. His mentorship extends to advising students and collaborating on interdisciplinary projects.
Noah Snavely is a Professor of Computer Science at Cornell Tech and the Cornell Ann S. Bowers College of Computing and Information Science. His research spans computer vision and graphics with a focus on 3D scene understanding from image collections. He also works at Google DeepMind in NYC and leads the Cornell Graphics and Vision Group. His research interests include computer vision and computer graphics , particularly in recovering 3D structure from large photo collections, neural radiance fields, image-based rendering, and scene understanding. His work enables applications in mapping technologies, immersive VR experiences, and synthesizing 3D worlds from text prompts with implications for game design and filmmaking. Recent publications show a strong trend toward generative 3D modeling and neural rendering , with significant contributions in neural radiance fields, view synthesis, and dynamic scene reconstruction. His research group explores unstructured photo collections to develop new technology for 3D world modeling and image analysis. PECASE Microsoft New Faculty Fellowship Alfred P. Sloan Fellowship SIGGRAPH Significant New Researcher Award ACM Fellow (2023) IEEE Fellow NSF CAREER Award Helmholtz Prize Snavely has mentored numerous PhD students and postdocs including Ruojin Cai, Qianqian Wang, and Zhengqi Li. His teaching includes CS5670: Computer Vision across multiple semesters. He has received the Cornell College of Engineering Mr. and Mrs. Richard F. Tucker Teaching Excellence Award (2012). At Cornell Tech, his research group develops technology for modeling the world in 3D from online photo collections and analyzing images for computer graphics applications. His work has been featured in major news outlets including the New York Times, New Scientist, and MIT Technology Review.
Alexander Schwing is an Associate Professor in the Department of Electrical and Computer Engineering and Computer Science at the University of Illinois at Urbana-Champaign, affiliated with the Coordinated Science Laboratory. His research focuses on machine learning and computer vision with applications in 3D scene understanding, generative modeling, and multi-agent systems. Education: Diploma in Electrical Engineering and Information Technology, Technical University of Munich (TUM) PhD in Computer Science, ETH Zurich Postdoctoral Fellow, University of Toronto Research Interests: Structured prediction in deep learning Generative adversarial networks and stability Multi-modal vision-language models 3D scene reconstruction from single images Embodied agent collaboration Semantic segmentation with temporal coherence Recent Publications: Highlight trends in neural rendering, video object segmentation, and reinforcement learning with applications to 3D modeling and multi-agent systems. Notable innovations include SAIL-VOS dataset for amodal segmentation and NeRFDeformer for single-view scene transformation. Scientific Awards: NSF CAREER Award, 3M and Amazon research awards, multiple student recognition awards, ETH Zurich PhD medal, and best paper at Intelligent Tutoring Systems 2014. Teaching: Offers graduate courses in Pattern Recognition (ECE 544) and Machine Learning (CS 446/ECE 449). Previously taught at University of Toronto and ETH Zurich. Labs & Collaborations: Leads research at Coordinated Science Laboratory (UIUC) with collaborations across University of Toronto, ETH Zurich, and industry partners like Samsung SAIT and Amazon.
Olga Sorkine-Hornung is a Professor of Computer Science at ETH Zurich and head of the Institute of Visual Computing. She leads the Interactive Geometry Lab, focusing on theoretical and practical advancements in digital content creation, geometry processing, and shape modeling. Current position: ETH Zurich, Department of Computer Science Previous roles: Courant Institute (NYU), Technical University of Berlin Education: BSc and PhD from Tel Aviv University, postdoc at TU Berlin Her research spans shape representation, digital fabrication, computer animation, and fundamental geometry processing. Key contributions include Laplacian surface editing, as-rigid-as-possible deformation, and generalized winding numbers. She works on applications in VR, AR, and autonomous systems. Awards include Test of Time Awards (2024), ACM Fellow (2020), ERC Consolidator Grant (2020), and EUROGRAPHICS Young Researcher Award (2008). She has supervised numerous students and co-developed software libraries like libigl and Instant Meshes . Co-chair roles for SIGGRAPH, Eurographics, and Pacific Graphics Editorial board member for ACM Transactions on Graphics and other journals Keynote speaker at VMV, CVPR, and SIAM conferences Her work bridges mathematical rigor with practical implementation, advancing computer graphics and geometry processing through intuitive algorithms that maintain surface detail while enabling efficient computation.
Lourdes Agapito is a Professor of 3D Vision at the Department of Computer Science, University College London (UCL), within the Faculty of Engineering Sciences. She leads research in Non-Rigid Structure from Motion (NR-SFM) and 3D reconstruction from monocular video sequences. Her work addresses dynamic scenes, deformable objects, and articulated structures, with applications in robotics and computer vision. She holds an ERC Starting Grant (2008–2014) and led the EU Horizon 2020-funded Second Hands project (2014–2019), collaborating with institutions like EPFL and KIT to develop robots with 3D visual perception for maintenance tasks. Her research group focuses on dense optical flow estimation, video registration, and deformable tracking. Agapito’s research interests include monocular 3D reconstruction, non-rigid motion analysis, and neural approaches to 3D modeling. She has supervised multiple PhD students and postdocs, including notable researchers such as Ravi Garg and Marco Paladini (Sullivan Prize recipient). Her contributions to conferences include roles as Program Chair for CVPR 2016 and CVPR 2017, and she has authored influential papers on topics like Video-Popup (ECCV 2014) and Modal Space (CVPR 2017). Current projects involve advancing neural parametric models and real-time 3D reconstruction techniques. Awards include the ERC Starting Grant and recognition for her team’s work in non-rigid reconstruction. She actively mentors students and collaborates on grants, with recent openings for postdocs and PhD candidates in 3D vision and robotics.
Deva Ramanan is a Professor at the Robotics Institute of Carnegie Melllon University, where he leads research in computer vision and machine learning. His work focuses on modeling human visual perception, leveraging large-scale visual data, and developing systems for 3D understanding, neural rendering, and autonomous systems. He advises a large group of PhD students and has mentored numerous postdoctoral researchers now in leading roles across industry and academia. His research interests include computer vision, machine learning, human perception modeling, 3D scene understanding, neural rendering, autonomous driving, video understanding, and multimodal foundation models. These areas reflect his focus on both foundational models and their application to real-world problems in robotics and AI. The recent publications highlight a strong trend toward multimodal and 3D-aware models, with increasing use of diffusion models, neural fields, and large vision-language systems. Key themes include scene flow, 3D reconstruction from monocular video, autonomous driving perception, and robust evaluation of vision-language models. There is a clear emphasis on both methodological innovation and practical deployment in dynamic environments. Marr Prize, Honorable Mention (ICCV 2021) Best Paper, Honorable Mention (ECCV 2020) Best Paper Finalist (WACV 2024) Best Paper Award (WACV 2016) Best Industrial Paper, Honorable Mention (BMVC 2017) Marr Prize winner (ICCV 2009) Deva Ramanan has advised numerous PhD and master’s students, many of whom are now at top institutions and companies including Apple, Meta, Google, Nvidia, OpenAI, and Princeton. He has received substantial funding from IARPA, DARPA, NSF, Intel, Google, and Facebook for projects in video analytics, dispersed computing, visual cloud systems, and multi-task recognition. His group has developed influential datasets and benchmarks used widely in the community. He leads a vibrant research lab focused on advancing computer vision through deep learning and multimodal integration. His team works on core challenges in perception, including 3D reconstruction, motion modeling, object detection, and scene understanding, with applications in robotics and autonomous systems.
Haibin Ling is the SUNY Empire Innovation Professor in the Department of Computer Science at Stony Brook University, part of the College of Engineering and Applied Sciences. His research focuses on computer vision, medical image analysis, augmented reality, and AI applications in science. He holds a Ph.D. from the University of Maryland (2006) and prior degrees from Peking University. Previously, he worked at Temple University (2008–2019) and held roles at Siemens Corporate Research, UCLA, and Microsoft Research Asia. Professor Ling's work spans biomedical imaging, AI for science, and human-computer interaction. He leads the CV Lab and collaborates with the AI Institute at Stony Brook. Awards include the NSF CAREER Award (2014), Best Student Paper (ACM UIST 2003), and IEEE Fellow (2020). He serves on editorial boards for IEEE Trans. PAMI, Pattern Recognition, and CVIU, and chairs major conferences like CVPR. His research group includes over 50 students and alumni, with active projects in tracking benchmarks (LaSOT), Leafsnap, and medical imaging tools. Notable publications address OCTA flow estimation, backdoor attacks on vision models, and topology-guided medical learning. Collaborations involve institutions like Temple University and Stony Brook's Department of Applied Mathematics and Statistics.
Erik B. Sudderth is a Professor of Computer Science and Statistics and Chancellor's Fellow at the University of California, Irvine (UCI). He leads the Learning, Inference, & Vision Group and directs multiple research centers, including the UCI Center for Machine Learning and Intelligent Systems and the HPI Research Center in Machine Learning and Data Science. He previously served as an Associate Professor at Brown University. Education: B.S. (summa cum laude) in Electrical Engineering from UC San Diego (1999), M.S. and Ph.D. in EECS from MIT (2002, 2006). His research focuses on statistical methods for scalable machine learning, Bayesian nonparametrics, probabilistic graphical models, and applications in computer vision, AI, and environmental science. Key areas include nonparametric clustering, deep generative models, and particle-based inference algorithms. Research interests span diverse topics: advancing Bayesian nonparametric models for medical time series, scalable variational inference, and AI ethics. Notable contributions include the NET-VISA seismic monitoring system (ISBA Mitchell Prize, 2014), the BNPy toolbox (NSF CAREER Award), and work on diverse particle max-product algorithms for continuous inference. Scientific awards include the NSF CAREER Award, ISBA Mitchell Prize, and recognition as one of "AI's 10 to Watch" (IEEE). He has served as editor for top journals (JMLR, IEEE PAMI) and conference chairs (NeurIPS, CVPR). His work bridges theory and practice, with applications in robotics, climate science, and healthcare. Labs/Teams: UCI Learning, Inference, & Vision Group; UCI Center for Machine Learning; CREATE Technology Center. Grants include NSF funding for visually impaired collaboration tools and soil biogeochemical modeling.
Perla Maiolino serves as an Associate Professor in Engineering Science at the University of Oxford and Principal Investigator of the Soft Robotics Lab (SRL) within the Oxford Robotics Institute. Her academic foundation includes BEng, MEng, and PhD degrees in Robotics and Automation from the University of Genoa, where she pioneered CySkin technology for distributed tactile sensing in robots—later exhibited at the Science Museum in London. She expanded her expertise during a 2017-2018 postdoctoral fellowship at Cambridge University's Biologically Inspired Robotics Lab, focusing on soft robotics and tactile perception. Dr. Maiolino's research centers on developing artificial skin systems, soft robotic actuators, and distributed sensing architectures. Her work bridges biological inspiration with engineering innovation to create robots capable of safe human interaction and dexterous manipulation in unstructured environments. Key contributions include compliant beaded-string jamming mechanisms for anthropomorphic fingers, monolithic 3D-printed soft pneumatic arms (JAMMit!), and distributed time-of-flight sensor networks for robotic self-awareness. Recent publications (2024-2025) reveal a strong convergence of tactile sensing with machine learning, featuring optical flow for gesture recognition, diffusion models for artificial skin simulation, and zero-shot sim-to-real transfer techniques. Her team has made significant advances in multi-modal sensing integration, variable stiffness actuation, and scene flow estimation for robots operating in dynamic surroundings. Scientific Awards No specific awards were documented in the provided institutional materials. Advising and Grants While her leadership of the Soft Robotics Lab implies active student supervision and grant management, detailed information about advisees or funded projects was not included in the source documentation. Labs and Teams As Principal Investigator of the Soft Robotics Lab at Oxford Robotics Institute, Dr. Maiolino directs research on tactile perception systems, soft actuation mechanisms, and sensor-integrated robotic structures. The lab's work focuses on applications requiring safe physical interaction, including healthcare robotics and human-robot collaboration scenarios, with emphasis on multi-material 3D printing and embedded sensing technologies.
Michael J. Black is a Professor and Honorarprofessor at the University of Tübingen's Faculty of Science, Department of Computer Science, and a founding Director of the Max Planck Institute for Intelligent Systems, leading the Perceiving Systems department. He holds a B.Sc. from the University of British Columbia (1985), M.S. from Stanford (1989), and Ph.D. in Computer Science from Yale (1992). His research focuses on computer vision, 3D human modeling, motion capture, and AI-driven digital humans. Key contributions include the SMPL body model, optical flow algorithms, and datasets like Middlebury Flow and Sintel. He has received major awards such as the PAMI Distinguished Researcher Award, multiple Koenderink and Longuet-Higgins Prizes, and is a member of the German National Academy of Sciences Leopoldina and Royal Swedish Academy of Sciences. His commercial ventures include co-founding Body Labs (acquired by Amazon) and Meshcapade, advancing 3D human generation and interaction technologies. Recent work includes markerless motion capture systems (e.g., MAMMA, PICO), 3D hair and garment synthesis, and AI tools like ChatHuman for 3D human interaction analysis. His research bridges vision, graphics, and robotics, with applications in animation, healthcare, and robotics.
Ranjay Krishna is an Assistant Professor at the Paul G. Allen School of Computer Science & Engineering at the University of Washington, where he co-directs the RAIVN lab and leads the computer vision team at the Allen Institute for AI (Ai2). His research intersects computer vision , natural language processing , robotics , and human-computer interaction . PhD in Computer Science from Stanford University (2021) Bachelor's and Master's degrees from Stanford and Cornell His work has received best paper , outstanding paper , and orals at top conferences like CVPR, ACL, CSCW, NeurIPS, UIST, and ECCV. Media outlets including Science , Forbes , and PBS NOVA have covered his research. He has been supported by grants from Google , Apple , NFS , and others. Ranjay advises a diverse group of 15 PhD and postdoctoral researchers , including Jieyu Zhang, Benlin Liu, and Cheng-Yu Hsieh. His teams have developed benchmarks like MemoryBench and The Colosseum , and his PathFinder framework achieved 74% accuracy in skin melanoma diagnosis—surpassing human experts by 9%. Notable contributions include: Perception Tokens for visual reasoning in MLMs SAM2Act for robotic manipulation with memory Synthetic Visual Genome dataset with 5.6M relationships
Dylan Campbell is a Lecturer in Computing at the Australian National University (ANU), affiliated with the ANU College of Systems & Society. His research focuses on computer vision, optimization, and robotics, particularly in 3D vision and deep learning applications. He has held prior roles as a Research Fellow at the University of Oxford’s Visual Geometry Group and ANU’s Australian Centre for Robotic Vision. Campbell holds a PhD from ANU (2018) and a BE in Mechatronic Engineering from UNSW (2012). Research interests include geometric sensor alignment, neural radiance fields, and differentiable optimization layers. He actively supervises students (7 PhD/DPhil, 3 MEng, 9 honours) and teaches advanced courses in computer vision and robotics. Notable awards include the Marr Prize Honourable Mention (2017) and the IEEE Australia Council Postgraduate Student Paper Competition (2018). He has organized workshops at ECCV and CVPR, served as a reviewer for top conferences like CVPR/ICCV/ECCV, and contributed to datasets like SEED4D and RefRef. His work emphasizes efficient training of neural networks and leveraging symmetries in data for long-range connections.
Michael J. Black is a Professor and Director at the Max Planck Institute for Intelligent Systems in Tübingen, Germany, where he leads the Perceiving Systems department and serves as Managing Director . He is also an Honorarprofessor at the University of Tübingen 's Faculty of Science . His career spans roles at Brown University (2000-2010), Xerox PARC, and academic-industry collaborations with Amazon and Meshcapade.
Niels Henze is a Professor at the Chair of Media Informatics within the Faculty of Languages, Literature and Cultural Studies at the University of Regensburg, where he has been serving since May 2018. His research centers on human-computer interaction, with a strong focus on predictive models in interactive systems, mobile interaction, augmented and virtual reality, and attention-aware computing. His research interests include: Human-Computer Interaction (HCI) Predictive modeling for runtime adaptation Mobile and wearable interaction Augmented and Virtual Reality (AR/VR) Attention-aware and context-sensitive systems Sociocognitive aspects of interactive technologies The analysis of his recent publications reveals a consistent focus on leveraging user behavior and contextual cues to build intelligent, adaptive interfaces. His work integrates machine learning with interaction design, emphasizing real-time model updates, implicit feedback, and cognitive load awareness to improve usability and user experience across mobile and immersive platforms. Scientific awards: No awards mentioned in the provided text. Niels Henze leads research in adaptive interactive systems and advises students in the field of media informatics. He has previously held a junior professorship at the University of Stuttgart and completed his doctorate at the University of Oldenburg under Susanne Boll. He is actively involved in advancing the theoretical and practical foundations of predictive and attention-aware computing. Grants and funding sources are not specified in the text. He is associated with the Institute for Information and Media, Language and Culture (I:IMSK) at the University of Regensburg, contributing to a multidisciplinary environment that bridges informatics with cultural and linguistic studies.