Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Christian Theobalt is a Professor of Computer Science at Saarland University and Scientific Director of the Visual Computing and Artificial Intelligence Department at the Max Planck Institute for Informatics . He leads the Saarbruecken Center for Visual Computing as a strategic partnership between Google and MPI. PhD in Computer Science (2005) from MPI-INF/Saarland University Postdoctoral Researcher at MPI (2005-2007) Visiting Assistant Professor at Stanford (2007-2009) His research focuses on the intersection of Computer Graphics, Computer Vision, and Artificial Intelligence , with specializations in: 3D/4D Human Reconstruction Neural Rendering Performance Capture Geometric Deep Learning Volumetric Video Quantum Visual Computing Recent article trends show emphasis on: Quantum computing applications in visual reconstruction Neural rendering with radiance fields Human motion capture from egocentric views Gesture-language interaction modeling Multi-modal scene understanding Real-time free-viewpoint rendering Scientific Awards : Fellow of EUROGRAPHICS (2022) CVPR Best Student Paper Honorable Mention (2020) ERC Consolidator Grant (2017) Karl Heinz Beckurts Award (2017) Busy Beaver Teaching Award (2016) ERC Starting Grant (2013) Advising over 50 students and researchers including: Current researchers: Viktor Rudnev, Linjie Lyu, Mohit Mendiratta Postdocs: Kwang In Kim, Kiran Varanasi Alumni: Franziska Mueller (Google), Dushyant Mehta (Qualcomm), Ayush Tewari (MIT) Grants include ERC grants, Google Glass Research Award, and multiple industry partnerships. His lab maintains cutting-edge facilities with: Multi-camera capture systems Quantum annealing infrastructure HDR radiance field technology GPU clusters for AI research Time-of-flight imaging systems Event camera arrays
Professor Hutan Ashrafian is a clinician-scientist and surgeon at the University of Leeds Business School, holding dual roles as Professor of Research Impact and Senior Research Fellow at Imperial College London. His expertise spans AI in healthcare, metabolic surgery, and ancient history. He pioneered STARD-AI and QUADAS-AI guidelines for AI diagnostics, collaborated on AI-driven breast cancer screening with Google and NHS, and led global health policy initiatives. Ashrafian’s work bridges medicine, history, and philosophy, including his contributions to understanding historical figures’ medical conditions, such as King Tutankhamun’s epilepsy. He has authored over 550 publications, 12 books, and holds a PhD in computational physiology and an MBA with distinction. Awards include the Hunterian Prize and Wellcome Trust Fellowship. His research also explores AI ethics, quantum physics paradoxes, and art-based medical diagnostics. Education: PhD in Computational Physiology (Imperial College London) MBA (Warwick Business School) MD (University College London) Bachelor of Science in Immunology (University College London) Research Interests: Ashrafian’s work spans: AI in Medicine: Diagnostic algorithms, guideline development (STARD-AI/QUADAS-AI), and ethical frameworks. Metabolic Surgery: Bariatric interventions, gut microbiome, and obesity treatments. Ancient History: Medical analyses of historical figures (e.g., Tutankhamun, Julius Caesar) and art-based pathology identification. Philosophy: AI rights (AIONAI law), temporal paradoxes, and the Simulation Argument. Awards: Royal College of Surgeons Arris and Gale Lectureship Hunterian Prize Wellcome Trust Research Fellowship Advising & Innovation: Supervised >50 PhD students, co-founded Oxford Medical Products (weight loss tech), and serves as CSO at Flagship Pioneering’s Preemptive Health division. He advises on NHS digital transformation and global health policy. Labs/Teams: Leads AI initiatives at Imperial’s Institute of Global Health Innovation and collaborates with Google, NICE, and international institutions on health tech solutions.
Kristen Grauman is a Full Professor in the Department of Computer Science at the University of Texas at Austin, where she leads the UT Computer Vision Group. Her research focuses on computer vision and machine learning, with applications in visual recognition, video analysis, and multi-modal perception. She received her B.A. from Boston College and her Ph.D. from MIT. Her research interests span visual recognition, image and video search, video analysis, first-person vision, embodied and multi-modal perception, and interactive machine learning. She has made significant contributions to the field, particularly in developing algorithms for understanding visual content and human activities from video, including foundational work on the Pyramid Match Kernel and relative attributes. Her recent publications reveal a strong emphasis on egocentric (first-person) vision, audio-visual learning, and view-invariant representations. There is a clear trajectory toward multi-modal integration (vision, audio, language) and real-world applications in instructional videos, human activity understanding, and embodied AI systems. She has received numerous awards including: AAAI Fellow (2019) J. K. Aggarwal Prize, International Association for Pattern Recognition (2018) Helmholtz Prize (2017) UT Austin Academy of Distinguished Teachers (2017) Best Paper Award, Asian Conference on Computer Vision (2016) Presidential Early Career Award for Scientists and Engineers (2014) Computers and Thought Award, International Joint Conferences on Artificial Intelligence (2013) Pattern Analysis and Machine Intelligence Young Researcher Award (2013) Alfred P. Sloan Research Fellow (2012) Marr Prize (2011) Prof. Grauman serves as Associate Editor-in-Chief for the IEEE Transactions on Pattern Analysis and Machine Intelligence. She has secured substantial research funding including the Presidential Early Career Award, NSF grants, and industry partnerships. Her advising has produced numerous influential publications and students who are now leaders in computer vision. She leads the UT Computer Vision Group, which collaborates closely with the Electrical and Computer Engineering Department. The group is pioneering large-scale egocentric video research through projects like Ego4D and Ego-Exo4D, focusing on real-world applications in human activity understanding, audio-visual perception, and interactive systems.
Vineeth N Balasubramanian is a Professor in the Department of Computer Science & Engineering at the Indian Institute of Technology Hyderabad, with affiliate faculty status in the Department of Artificial Intelligence. His research focuses on the intersection of deep learning, machine learning, and computer vision, emphasizing explainability, robustness, and real-world applications. He leads Lab 1055, which investigates problems such as Explainable and robust AI/ML systems Lifelong learning in evolving environments Multimodal vision-language models Applications in agriculture, autonomous navigation, and human behavior analysis His recent work includes causal reasoning in transformers, vision-language model capabilities, and drone-based object detection. Funded by organizations like Google, Microsoft, Intel, and DST, he has received multiple awards including the World's Top 2% Scientists (2022-23), INSA/INAE Fellowships, and Best Paper recognitions. Lab 1055 collaborates with institutions like CMU, UBC, and Monash University, contributing to cutting-edge advancements in AI.
Chris Donahue is an Assistant Professor in the Computer Science Department at Carnegie Mellon University . He also serves as a part-time Research Scientist at Google DeepMind on the Magenta team. His work focuses on leveraging generative AI to enhance human creativity, particularly in music. Education: PhD in Computer Science (UC San Diego), Postdoctoral Scholar (Stanford University) His research spans controllable generative modeling of music and audio , with a focus on real-time interactive systems. Projects like Piano Genie , Beat Sage , and Copilot Arena demonstrate his commitment to real-world deployment. His Generative Creativity Lab (G-CLef) explores AI applications beyond music, including programming and natural language. Recent publications highlight advancements in multimodal music evaluation , real-time adaptation , and AI-driven sound morphing . He co-developed Magenta RealTime , an open-weight real-time music generation model, and MusicFX DJ Mode . Scientific Awards: Best Paper Award (top 1) at NAACL Student Research Workshop 2025 Best Paper Award (top 1% of submissions) at CHI 2025 Best Paper Runner-up at ISMIR 2021 He co-advises PhD students like Wayne Chi (NDSEG Fellow) and mentors Irmak Bukey . His lab receives support from the AIxArts incubator fund at CMU .
Nima Fazeli is an Assistant Professor of Robotics at the University of Michigan (2020–Present), holding courtesy appointments in Computer Science & Engineering (CSE) and Mechanical Engineering. He directs the Manipulation and Machine Intelligence (MMint) Lab, focusing on enabling dexterous robotic manipulation through multimodal representation learning, tactile sensing, and model-based reasoning. His work integrates mechanics, perception, controls, and planning to achieve autonomous interaction with uncertain environments. Education: PhD, MIT (2019); MSc, University of Maryland (2014); BSc, Amirkabir University of Technology (2011) Research interests emphasize embodied intelligence , including visuo-tactile fusion, contact dynamics modeling, and cross-modal learning. Recent work explores tactile shadows, deformable object manipulation, and language-guided robot control. His research is supported by the NSF CAREER grant and National Robotics Initiative, with applications in manufacturing, assistive robotics, and space systems. Publications span topics like tactile sensing hardware (e.g., GelSlim 4.0), visuo-tactile implicit representations (ViTaSCOPE), and failure recovery policies (Racer). His team’s work has been featured in outlets like The New York Times and BBC. Key Awards: NSF CAREER Grant (2024) Teaching includes Introduction to Robotic Manipulation . Collaborations involve cross-disciplinary projects with mechanical, electrical, and biomedical engineering groups.
Christian Rupprecht is an Associate Professor at the Department of Computer Science, University of Oxford, specializing in computer vision and machine learning. His research focuses on unsupervised learning, 3D reconstruction, and visual understanding. His work includes contributions to conferences such as GCPR'25, ICCV'25, and CVPR'25, with papers spanning topics like correspondence estimation, animal pose modeling, and synthetic data generation. He leads projects within the prestigious Visual Geometry Group (VGG). Notably, his paper VGGT received the Best Paper Award at CVPR'25. His research integrates deep learning and geometric modeling, emphasizing robustness and generalization in visual systems. Best Paper Award at CVPR'25
Alexander Schwing is an Associate Professor in the Department of Electrical and Computer Engineering and Computer Science at the University of Illinois at Urbana-Champaign, affiliated with the Coordinated Science Laboratory. His research focuses on machine learning and computer vision with applications in 3D scene understanding, generative modeling, and multi-agent systems. Education: Diploma in Electrical Engineering and Information Technology, Technical University of Munich (TUM) PhD in Computer Science, ETH Zurich Postdoctoral Fellow, University of Toronto Research Interests: Structured prediction in deep learning Generative adversarial networks and stability Multi-modal vision-language models 3D scene reconstruction from single images Embodied agent collaboration Semantic segmentation with temporal coherence Recent Publications: Highlight trends in neural rendering, video object segmentation, and reinforcement learning with applications to 3D modeling and multi-agent systems. Notable innovations include SAIL-VOS dataset for amodal segmentation and NeRFDeformer for single-view scene transformation. Scientific Awards: NSF CAREER Award, 3M and Amazon research awards, multiple student recognition awards, ETH Zurich PhD medal, and best paper at Intelligent Tutoring Systems 2014. Teaching: Offers graduate courses in Pattern Recognition (ECE 544) and Machine Learning (CS 446/ECE 449). Previously taught at University of Toronto and ETH Zurich. Labs & Collaborations: Leads research at Coordinated Science Laboratory (UIUC) with collaborations across University of Toronto, ETH Zurich, and industry partners like Samsung SAIT and Amazon.
Prof. Matthias Nießner is a Professor at the Technical University of Munich, leading the Visual Computing Lab. His research intersects computer graphics, vision, and AI, focusing on 3D reconstruction, semantic understanding, and AI-driven video synthesis. He holds a PhD from the University of Erlangen-Nuremberg (2013) and was a Visiting Assistant Professor at Stanford University (2013–2017). Notable awards include the ERC Starting Grant (2018), Nvidia Professorship Award, and Eurographics Young Researcher Award (2019). His work has been featured in mainstream media and led to startups like Synthesia Inc. Research spans Gaussian splatting, neural radiance fields, and generative AI for 3D avatars. Over 150 publications include SIGGRAPH, CVPR, and ECCV, with best paper awards. Projects like Face2Face and ScanNet have driven innovation in facial reenactment and 3D scene datasets. Education: PhD in Computer Science, University of Erlangen-Nuremberg (2013) Diploma in Computer Science, University of Erlangen-Nuremberg (2010) Research Interests: 3D digitization, neural rendering, generative AI, non-rigid reconstruction, and applications in AR/VR. Awards: ERC Starting Grant (2018) Nvidia Professorship Award (2018) Google Faculty Award (2018) SIGGRAPH Best Emerging Tech Award (2016) Grants: Over €1.5M from ERC and industry partnerships. Labs/Teams: Visual Computing Lab at TUM and Synthesia Inc. (co-founder). Key projects include ScanNet (large 3D indoor dataset), Face2Face (real-time facial reenactment), and Gaussian-based 3D avatars. Current work focuses on diffusion models, neural radiance fields, and AI-generated media detection.
Anthony Rowe is the Siewiorek and Walker Family Professor of Electrical and Computer Engineering at Carnegie Mellon University (CMU) and a Chief Scientist at Bosch Research. His primary affiliation is with the CyLab and the Wireless, Sensing and Embedded Systems (WiSE Lab) at CMU. He specializes in networked embedded systems, sensor networks, and extended reality (XR) technologies. His research emphasizes energy-efficient sensing, real-time localization, and XR integration with physical systems. Research Focus: His work spans XR systems (e.g., AR/VR edge networking in ARENA), mmWave radar for sensing (e.g., tire wear monitoring via Osprey), distributed edge computing (Silverline), and low-power wide-area networking (OpenChirp). Recent efforts include AI-integrated XR platforms (XaiR) and radar tomography (DART). Grants & Projects: Leads the CONIX Research Center ($27.5M NSF/DARPA grant), Bosch-funded edge computing projects, and DOE initiatives on microgrids. Notable projects include ARENA (XR edge architecture), GridBallast (smart grid control), and rural microgrid deployments in Haiti. Awards: Best Student Paper (ISMAR 2024), Best Paper (IPSN 2020), and the Steven J. Fenves Research Award (2015). Recognized for innovations in localization (MobiCom 2021), radar (ICRA 2023), and energy systems (BuildSys 2010). Teaching: Teaches courses on embedded systems (18-349/18-449), real-time systems, and mixed reality (18-453). Courses emphasize hands-on design and real-world applications. Labs & Teams: Directs the WiSE Lab, collaborating with Bosch Research and industry partners. The lab develops open-source frameworks like ARENA and OpenChirp, and contributes to standards for edge computing and sensing.
Haijun Xia is an Assistant Professor at the University of California, San Diego (UCSD), where he directs the Foundation Interface Lab and contributes to the Cognitive Science and Design Lab. His work focuses on Human-Computer Interaction (HCI) and Human-AI Collaboration , emphasizing the development of dynamic, malleable interfaces that integrate human cognition with intelligent systems to enhance thinking and working paradigms. Education: Bachelor's, Tsinghua University Master's and PhD, University of Toronto His research spans generative AI , interface design , and creative collaboration , with applications in data visualization, programming, and scholarly work. He advocates for treating information as a malleable material that can be flexibly manipulated by users and AI agents. Recent publications highlight advancements in: 2025 CHI Conference - Generative user interfaces, malleable overview-detail designs, and compositional structures for human-AI co-creation 2024 CHI Conference - Structured design space exploration with LLMs, ASCII diagram analysis, and adaptive presentation systems 2023 CHI/UIST Conferences - Data particle visualization, contextual logging tools, and metaphor generation for science communication Scientific recognitions include: Multiple Best Paper Awards and Honorable Mentions at CHI and UIST Hellman Fellowship and grants from NSF , Microsoft , Google , Meta , Apple , and Adobe He actively recruits undergraduate and graduate interns and has an upcoming postdoc position in early 2025 for researchers interested in foundational interface development.
Insup Lee is the Cecilia Fitler Moore Professor in the Department of Computer and Information Science and Director of the PRECISE Center at the University of Pennsylvania's School of Engineering and Applied Science. He holds a secondary appointment in the Department of Electrical and Systems Engineering and the Perelman School of Medicine’s Department of Biostatistics, Epidemiology, and Informatics. IEEE TCCPS Distinguished Leadership Award (2023) Fellow of the AAAS (2022) Test of Time Award, Runtime Verification (2019) Fellow of the ACM (2017) Best Paper Awards at IEEE ICPS, ACM/IEEE ICCPS, and MEMOCODE His research focuses on cyber-physical systems , real-time and embedded systems , safe autonomy , and internet of medical things , with applications in healthcare and connected systems. He advises PhD students including Eric Lu, Kaustubh Sridhar, Sooyong Jang, and Jean Park (co-advised with Kevin Johnson). Recent publications address safety monitoring for learning-enabled systems, model-free control synthesis using reinforcement learning, and multilingual toxicity guardrails for large language models. His team collaborates with institutions like Hillrom and Penn Nursing to optimize medical device usage in clinical settings.
Mathieu Salzmann is a Senior Scientist and Lecturer at École Polytechnique Fédérale de Lausanne (EPFL), affiliated with the Computer Vision Laboratory (CVLAB) in the School of Computer and Communication Sciences (IC). He also holds a courtesy appointment with the EPFL College of Humanities and serves as Deputy Chief Data Scientist at the Swiss Data Science Center (SDSC). He has held concurrent roles in teaching units including SIN, SODH, and SSC, reflecting his interdisciplinary engagement. His research focuses on the intersection of machine learning and computer vision, particularly in deep learning for 2D and 3D visual scene understanding, efficient and robust models, domain adaptation, and interpretable AI. These interests are evident across his extensive publication record in top-tier venues. His recent publications (2023–2024) show a consistent trend in advancing deep learning methods for visual recognition, with strong representation at CVPR, ICCV, ECCV, ICML, ICLR, and NeurIPS. Topics include domain generalization, 3D understanding, model robustness, and multimodal learning, often with applications in real-world systems. His editorial roles as Associate Editor for IEEE TPAMI and Action Editor for TMLR further highlight his leadership in the field. Area Chair: ICML 2023, CVPR 2023, ICCV 2023, NeurIPS 2023, AAAI 2024, ECCV 2024 Associate Editor: IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) Action Editor: Transactions on Machine Learning Research (TMLR) Mathieu Salzmann has supervised numerous PhD students at EPFL, both current and past, including Bouquet Yann Yanis, Javed Saqib, Li Shuangqi, and others. He has also been involved in research grants and collaborative projects, such as his work with S. Süsstrunk and R. Baroni on comics reconfiguration. His part-time role as Senior GNC Engineer at ClearSpace (2020–2024) illustrates his applied research engagement in aerospace systems. He is actively involved in EPFL’s data science and AI research ecosystem through SDSC and multiple labs.
Derek W Hoiem is a Professor in the Siebel School for Computing and Data Science at the University of Illinois Urbana-Champaign, where he has been a faculty member since 2009. His research focuses on computer vision and related areas, and he is also the co-founder and Chief Science Officer of Reconstruct, an AI-based construction technology company. His educational background includes: PhD in Robotics, Carnegie Mellon University (2007) Beckman Postdoctoral Fellowship (2008) Prof. Hoiem's research spans computer vision, with a focus on object recognition, scene understanding, and graphics. His work also extends to mobile robotics and 3D scene reconstruction. He has made significant contributions in areas such as visual recognition, 3D modeling, and the application of computer vision in construction monitoring. His recent publications (2023-2025) demonstrate a strong focus on advancing multimodal understanding, particularly in region-based representations, 3D vision, and neural radiance fields. There is a clear trend towards integrating language and vision, improving efficiency in neural networks, and applying computer vision to real-world problems such as construction progress monitoring. His scientific awards and honors are extensive and include: IEEE Fellow (2022) University Scholar (2022) Koendrink Prize (2022) Dean's Award for Excellence in Research, Associate Professor (2021) Campus Distinguished Promotion Award (2015) Best Paper Award: IEEE Winter Conference on Applications in Computer Vision (WACV) (2015) CW Gear Junior Faculty Award (2014) IEEE PAMI Young Researcher Award (2014) Dean's Award for Excellence in Research, Assistant Professor (2014) Sloan Research Fellowship (2013) Intel Early Career Faculty Honor Program Award (2012) NSF CAREER Award (2011) ACM Doctoral Dissertation Award, Honorable Mention (2008) Carnegie Mellon University SCS Distinguished Dissertation Award (2008) Best Paper Award: IEEE Computer Vision and Pattern Recognition (CVPR) (2006) Prof. Hoiem has secured significant research funding, including an NSF CAREER award and an Intel Early Career Faculty award. He is also actively involved in technology transfer, having co-founded Reconstruct where he serves as Chief Science Officer. His teaching excellence is reflected in multiple "List of Teachers Ranked as Excellent" awards spanning from 2010 to 2021. Prof. Hoiem leads a research group at UIUC focused on computer vision and 3D scene understanding. Additionally, he co-founded and serves as Chief Science Officer at Reconstruct, which develops AI-based solutions for construction monitoring.