Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Christian Theobalt is a Professor of Computer Science at Saarland University and Scientific Director of the Visual Computing and Artificial Intelligence Department at the Max Planck Institute for Informatics . He leads the Saarbruecken Center for Visual Computing as a strategic partnership between Google and MPI. PhD in Computer Science (2005) from MPI-INF/Saarland University Postdoctoral Researcher at MPI (2005-2007) Visiting Assistant Professor at Stanford (2007-2009) His research focuses on the intersection of Computer Graphics, Computer Vision, and Artificial Intelligence , with specializations in: 3D/4D Human Reconstruction Neural Rendering Performance Capture Geometric Deep Learning Volumetric Video Quantum Visual Computing Recent article trends show emphasis on: Quantum computing applications in visual reconstruction Neural rendering with radiance fields Human motion capture from egocentric views Gesture-language interaction modeling Multi-modal scene understanding Real-time free-viewpoint rendering Scientific Awards : Fellow of EUROGRAPHICS (2022) CVPR Best Student Paper Honorable Mention (2020) ERC Consolidator Grant (2017) Karl Heinz Beckurts Award (2017) Busy Beaver Teaching Award (2016) ERC Starting Grant (2013) Advising over 50 students and researchers including: Current researchers: Viktor Rudnev, Linjie Lyu, Mohit Mendiratta Postdocs: Kwang In Kim, Kiran Varanasi Alumni: Franziska Mueller (Google), Dushyant Mehta (Qualcomm), Ayush Tewari (MIT) Grants include ERC grants, Google Glass Research Award, and multiple industry partnerships. His lab maintains cutting-edge facilities with: Multi-camera capture systems Quantum annealing infrastructure HDR radiance field technology GPU clusters for AI research Time-of-flight imaging systems Event camera arrays
Chua Tat Seng is a Professor at the School of Computing, National University of Singapore (NUS), holding the KITHCT Chair Professorship since 2009. He serves as co-Director of the NExT++ Center, a joint research center between NUS and Tsinghua University focused on Extreme Search. His academic career spans over three decades at NUS, where he has held various leadership positions including Acting Dean of the School of Computing (1998-2000) and Acting Head of the Department of Information Systems & Computer Science (1996-1998). Professor Chua's research spans unstructured data analytics , multimedia information retrieval , recommendation and conversation systems , and emerging applications in e-commerce and fintech . He established the Lab for Media Search (LMS) at the School of Computing and has been instrumental in advancing multimodal learning and search technologies. His work bridges theoretical foundations with practical applications, particularly in developing trustable AI systems for real-world deployment. His recent publications demonstrate a strong focus on large language models for recommendation systems , multimodal learning , and generative AI applications . The research trends show increasing emphasis on LLM-based recommendation, multimodal understanding, and addressing fundamental challenges in AI reliability, fairness, and efficiency. His work spans theoretical advancements in representation learning to practical applications in e-commerce, finance, and healthcare domains. ACM SIGMM Technical Achievement Award 2015 Multiple Best Paper Awards across ACM Multimedia, IEEE Transactions, and MMM conferences (2007-2020) Professor Chua has supervised 37 PhD students since 2004, establishing himself as a dedicated mentor in the academic community. His research has been supported by substantial grants including NExT++ ($12 million), Base Metals Price Forecasting ($200,000), and Multilingual Multimodal Knowledge Graph ($500,000). He maintains active collaborations with Tsinghua University, University of Southampton, and industry partners like Four Elements Capital and Singapore Press Holdings. As co-Director of the NExT++ Center, he leads a major research initiative focused on Web Intelligence and User Empowerment. His visiting professorships at Tsinghua University (2017-present) and Zhejiang University (2021-present) reflect his international impact in the field of multimedia and AI research.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Kristen Grauman is a Full Professor in the Department of Computer Science at the University of Texas at Austin, where she leads the UT Computer Vision Group. Her research focuses on computer vision and machine learning, with applications in visual recognition, video analysis, and multi-modal perception. She received her B.A. from Boston College and her Ph.D. from MIT. Her research interests span visual recognition, image and video search, video analysis, first-person vision, embodied and multi-modal perception, and interactive machine learning. She has made significant contributions to the field, particularly in developing algorithms for understanding visual content and human activities from video, including foundational work on the Pyramid Match Kernel and relative attributes. Her recent publications reveal a strong emphasis on egocentric (first-person) vision, audio-visual learning, and view-invariant representations. There is a clear trajectory toward multi-modal integration (vision, audio, language) and real-world applications in instructional videos, human activity understanding, and embodied AI systems. She has received numerous awards including: AAAI Fellow (2019) J. K. Aggarwal Prize, International Association for Pattern Recognition (2018) Helmholtz Prize (2017) UT Austin Academy of Distinguished Teachers (2017) Best Paper Award, Asian Conference on Computer Vision (2016) Presidential Early Career Award for Scientists and Engineers (2014) Computers and Thought Award, International Joint Conferences on Artificial Intelligence (2013) Pattern Analysis and Machine Intelligence Young Researcher Award (2013) Alfred P. Sloan Research Fellow (2012) Marr Prize (2011) Prof. Grauman serves as Associate Editor-in-Chief for the IEEE Transactions on Pattern Analysis and Machine Intelligence. She has secured substantial research funding including the Presidential Early Career Award, NSF grants, and industry partnerships. Her advising has produced numerous influential publications and students who are now leaders in computer vision. She leads the UT Computer Vision Group, which collaborates closely with the Electrical and Computer Engineering Department. The group is pioneering large-scale egocentric video research through projects like Ego4D and Ego-Exo4D, focusing on real-world applications in human activity understanding, audio-visual perception, and interactive systems.
Jennifer Olsen, PhD, is an Assistant Professor of Computer Science at the University of San Diego since 2020. She holds a PhD, MS, and BS in Human-Computer Interaction and Cognitive Science from Carnegie Mellon University, followed by postdoctoral research at the Swiss Federal Institute of Technology (EPFL), Lausanne, Switzerland. Her research focuses on the intersection of human-computer interaction, cognition, and education, emphasizing collaborative learning and educational technology design from both learner and instructor perspectives. Education: PhD in Human-Computer Interaction, Carnegie Mellon University MS in Human-Computer Interaction, Carnegie Mellon University BS in Cognitive Science, Carnegie Mellon University Research Interests: Dr. Olsen explores how collaboration supports learning, designs technologies to enhance educational practices, and investigates gaze-based metrics for understanding collaborative problem-solving. Her work spans gamified robotics, AI-driven orchestration systems, and virtual reality applications in vocational training. She emphasizes learner-centered design and the integration of social robots and virtual agents in pedagogical settings. Grants/Advising: While no specific grants or advisees are listed, her prolific publication record indicates active involvement in educational technology research and development. Her work addresses challenges in classroom orchestration, multimodal data analysis, and accessibility in educational robotics. Labs/Teams: Collaborates with interdisciplinary teams focused on educational technology, human-robot interaction, and adaptive learning systems. Her research leverages tools like FROG orchestration graphs and eye-tracking technologies to develop practical classroom solutions.
Adrian Weller is a prominent researcher and academic at the University of Cambridge, serving as a Director of Research in Machine Learning within the Department of Engineering. He holds multiple significant leadership roles including Programme Director for Trust and Society at the Leverhulme Centre for the Future of Intelligence (CFI), and previously served as Programme Director for AI at The Alan Turing Institute, the UK national institute for data science and AI. His work bridges theoretical machine learning research with practical applications and societal implications of artificial intelligence. Weller's research interests span a broad spectrum of AI and machine learning topics with a particular focus on ensuring beneficial societal outcomes. His work encompasses explainability, fairness, robustness, scalability, privacy, safety, and ethics in AI systems. He has made significant contributions to trustworthy machine learning, including developing frameworks for AI governance, certification, and human-AI collaboration. His research group actively investigates neuro-symbolic approaches, privacy-preserving techniques, and methods for improving the reliability and interpretability of AI systems. His recent publications demonstrate a strong trend toward addressing the practical challenges of deploying AI systems in real-world contexts, particularly focusing on certification frameworks, governance mechanisms, and human-centered approaches. His work spans theoretical advances in machine learning architectures while maintaining a strong connection to societal impact, with publications appearing in top venues across AI, machine learning, and interdisciplinary applications. Scientific Awards: MBE for services to digital innovation (2022 Queen's Birthday Honours) Turing AI Fellowship for Trustworthy Machine Learning Weller actively supervises a large group of PhD students and postdocs, with current students including Juyeon Heo, Yanzhi Chen, Katie Collins, Isaac Reid, Yichao Liang, Herbie Bradley, and Shoaib Siddiqui. His former students have gone on to positions at leading institutions including Google DeepMind, ETH Zurich, NYU, and MPI-IS Tübingen. He has served on numerous advisory boards including the Centre for Data Ethics and Innovation, UNESCO's expert group on AI ethics, and the World Economic Forum's Global Future Council on AI. His research has been supported through his Turing AI Fellowship and various collaborative projects focused on safe and ethical AI development. Weller leads a vibrant research group focused on trustworthy machine learning, which actively organizes workshops and conferences including ICML 2024 (where he served as Program Chair), multiple workshops on responsible AI, and events through the ELLIS network. His group collaborates extensively across disciplines, working with researchers in computer science, social sciences, law, and policy to address the multifaceted challenges of developing beneficial AI systems.
Prof. Luke Zettlemoyer is an Adjunct Professor of Computer Science and Engineering at the University of Washington, with affiliations to the Department of Linguistics. He focuses on machine learning, natural language processing, and multimodal systems, contributing to advancements in large language models, ethical AI, and scalable architectures. His research addresses challenges in model alignment, generalization, and cross-domain integration. Key research interests include multimodal reward models, efficient tokenization strategies, and model optimization techniques. He has explored topics such as neural trajectories for robot learning, content-adaptive image processing, and ethical mitigation of verbatim data reproduction. His publications span 2023–2025, emphasizing practical applications of AI in robotics, vision-language systems, and scalable retrieval-based models. While no formal awards are listed, his work reflects significant contributions to foundational AI research.
Adriana Kovashka is an Associate Professor at the University of Pittsburgh , affiliated with the School of Computing and Information and serving as Department Chair . Her academic journey began with BA degrees in Computer Science and Media Studies from Pomona College (2008) and a PhD in Computer Science from The University of Texas at Austin (2014). Joined Pitt’s faculty in January 2015 NSF CAREER awardee (2021) Google Faculty Research Award recipient Dr. Kovashka’s research spans Computer Vision , Machine Learning , and Natural Language Processing , focusing on visual rhetoric, weak multimodal supervision, and domain adaptation. She pioneered techniques for analyzing political imagery, developing robust object detection frameworks, and exploring the intersection of visual and textual persuasion through large-scale annotated datasets. Her recent work emphasizes geographic diversity in vision-language systems, audio-visual fusion for domain generalization, and shape-texture bias mitigation in CNNs. Key publications include groundbreaking studies on symbolic reasoning, multimodal dialogue systems, and ethical AI applications in education. Scientific honors include: NSF CRII Award (2016) NSF CAREER Award (2021) Pitt CRDF Award (2016, 2018) Best Paper at ECV Workshop (2021) Google Faculty Research Award (2016, 2018) Dr. Kovashka actively mentors students in multimodal learning projects and collaborates with interdisciplinary teams on NSF-funded initiatives. She co-organizes workshops like the first CVPR workshop on advertisement understanding and leads research groups exploring human-AI co-learning systems.
Daniel Epstein is an Associate Professor in the Department of Informatics at the University of California, Irvine (UCI), where he leads the Personal Informatics Everyday (PIE) Lab. He holds a courtesy appointment in the Department of Computer Science and is affiliated with the Connected Learning Lab, Institute for Future Health, and Accessibility Research Collective. Epstein earned his Ph.D. in Computer Science & Engineering from the University of Washington and a B.S. in Computer Science from the University of Virginia. His research focuses on how personal tracking technology can better serve diverse users' needs, emphasizing ethical design, health equity, and long-term user engagement. Notable areas include pregnancy tracking, wearable technology for children and ADHD management, and AI-driven health chatbots. Epstein has received grants from NSF, Snap Inc., and UCI programs, totaling over $650k in funding. Epstein's work bridges HCI and health informatics, with over 60 peer-reviewed publications. He advises six Ph.D. students and has mentored numerous undergraduates. Awards include the ACM Senior Member distinction (2025), CHI Best Paper (2023), and NSF CAREER Award (2023). Professional service includes program committees for CHI, CSCW, and UbiComp, plus leadership roles in workshops like the Workgroup on Interactive Systems in Healthcare (WISH). Beyond academia, Epstein enjoys hiking, collecting turtle figures, and advocating for accessible technology.
Dr. Hassan Qudrat-Ullah is a Professor at the School of Administrative Studies, York University, and Coordinator of the Certificate in Logistics Management. He holds a PhD in Decision Sciences from NUS Business School and completed a post-doctoral fellowship at Carnegie Mellon University. His research focuses on dynamic decision making, system dynamics modeling, energy planning, and interactive learning environments. He teaches courses on quantitative methods, logistics, and decision analysis, informed by global industry experience across 20+ countries. Research interests include sustainability, climate change, systems thinking, and educational applications of decision sciences. He serves as Editor-in-Chief of the International Journal of Complexity in Applied Science and Technology and is a member of IEEE, DSI, and the International System Dynamics Society. His work has been published in Energy , Decision Support Systems , and others. Key projects include studies on 'structured-debriefing in dynamic decision making' and renewable energy policies in Africa. His recent articles (2023–2025) address AI integration in energy governance, system dynamics for supply chain resilience, and education for sustainability. Hassan advocates for systems thinking in K-12 education and enjoys traveling (visited 129 countries) and bird-watching.
Aniket (Niki) Kittur is a Professor in the Human-Computer Interaction Institute (HCII) at Carnegie Mellon University's School of Computer Science. His research focuses on AI-augmented cognition, exploring how humans and computational systems can partner to accelerate knowledge acquisition and innovation. With over 100 peer-reviewed publications and 17 best paper awards or honorable mentions, Dr. Kittur has established himself as a leader in human-computer interaction and social computing. Dr. Kittur's research spans multiple interconnected areas focused on enhancing human cognition through technology. His work on crowd-augmented cognition examines how crowds and computation can collectively enhance human intellect. In social computing, he investigates how sociotechnical architectures can motivate and coordinate large groups for complex tasks. His research in data visualization and sensemaking explores how people collect, organize, and make decisions with online information. Most recently, his work has shifted toward LLM-augmented cognition, investigating how humans and large language models can work together in partnership to achieve better decisions and greater creativity than either could alone. Dr. Kittur's publications reveal a consistent trajectory toward increasingly sophisticated human-AI collaboration systems. Early work focused on crowdsourcing complex tasks (CrowdForge, CrowdSynthesis), then evolved to knowledge acceleration systems (Knowledge Accelerator, Fuse), and has recently centered on LLM-integrated cognition (Selenite, InkSpire, BioSpark). His research consistently demonstrates practical impact, with systems influencing products used by millions and receiving coverage in major media outlets including Nature News, The Economist, and The Wall Street Journal. Kavli fellow NSF CAREER award recipient Allen Newell Award for Research Excellence Inductee into the CHI Academy 17 best paper awards or honorable mentions across his publications Dr. Kittur has advised numerous students, including Andrew Kuznetsov, and has secured major research funding from NSF, NIH, Google, Microsoft, Bosch, Toyota, and other organizations. His Skeema browser extension project, which addresses tab overload through innovative project-based organization, has shown remarkable user retention and represents a practical application of his research on knowledge acceleration. His work with industry partners has influenced products used by millions, including Semantic Reader, Google Shopping, and Wikipedia features.
Wei-Lun (Harry) Chao is an Associate Professor in the Department of Computer Science and Engineering at the Ohio State University (OSU), College of Engineering. Promoted to this role in May 2025, he is also an Innovation Scholar and Distinguished Assistant Professor of Engineering Inclusive Excellence. His work spans machine learning, computer vision, and their applications in autonomous driving, healthcare, biology, and natural language processing. Research Focus: Machine learning with imperfect data, interpretable and personalized learning, robust perception for autonomous systems, and visual recognition in real-world scenarios. Awards: 2025 OSU Early Career Distinguished Scholar Award, CVPR Best Student Paper Award (2024), CSE Faculty Teaching Award (2024), Lumley Research Award (2023). Grants: Funded by NSF, NIH, ONR, Cisco, AWS, and Google. Notable Research Trends: The 15 most recent articles highlight his work on vision foundation models, federated learning, diffusion models for biological species generation, interpretable vision transformers, and robust perception systems for autonomous driving. Key subfields include sparse autoencoders, 3D object detection, semi-supervised learning, and anomaly detection in scientific domains. Scientific Awards: 2025 Early Career Distinguished Scholar Award (OSU) CVPR Best Student Paper Award (2024) CSE Faculty Teaching Award (2024) Lumley Research Award (2023) Mentoring & Grants: As an advisor for the OSU Buckeye AutoDrive Team and AI Club, he mentors graduate and undergraduate students. His research is supported by major grants from NSF, NIH, ONR, and industry partners like Cisco and Google.
Prof. Stefan Leutenegger is a tenure-track Assistant Professor at Technische Universität München (TUM), leading the Machine Learning for Robotics group within the TUM School of Computation, Information, and Technology. Previously, he held roles as Senior Lecturer (2018–2021) and Lecturer (2014–2018) at Imperial College London's Dyson Robotics Lab, where he founded the Smart Robotics Lab. He earned his PhD (2014) from ETH Zurich, focusing on autonomous solar-powered aircraft navigation, and holds BSc (2006) and MSc (2009) in Mechanical Engineering from ETH Zurich. His research centers on mobile robotics, particularly enabling robots (e.g., drones) to perceive and navigate complex environments using machine learning and sensor data fusion. Key focus areas include SLAM, event-based vision, 3D reconstruction, and autonomous exploration. He has pioneered algorithms like BRISK (2011), OKVIS (2014), and ElasticFusion (2016), advancing real-time robotics perception. Notable Awards: Imperial College President's Award (2018), Best ECCV Paper (2016), ETH Medal for Dissertations (2015). Labs: TUM's Machine Learning for Robotics Group, Imperial's Smart Robotics Lab. Publications: Over 100 papers, including seminal works in CVPR, ECCV, and Robotics: Science and Systems. Current projects include DigiForests (forest inventory via robotics), aerial additive manufacturing, and object-centric semantic mapping. His work bridges theory and practice, with applications in autonomous drones, construction robotics, and human-robot interaction.
Lior Wolf is a Professor at the School of Computer Science, Tel Aviv University. Previously, he was a postdoctoral researcher at MIT's Center for Biological and Computational Learning (CBCL) under Prof. Tomaso Poggio and earned his PhD from Hebrew University of Jerusalem with Prof. Amnon Shashua. His educational background includes: PhD in Computer Science, Hebrew University of Jerusalem Postdoctoral Research, MIT CBCL Prof. Wolf's research centers on artificial intelligence with seminal contributions to deep learning, computer vision, and natural language processing. His work bridges theoretical foundations (e.g., attention mechanisms, transformer analysis) with practical applications in medical imaging, speech processing, and sign language technology. He pioneered methods for neural network interpretability, efficient sequence modeling, and multimodal fusion. Analysis of his 2023-2025 publications reveals dominant trends in large language model optimization (neuron pruning, attention analysis), efficient video generation, and cross-modal learning. His work increasingly integrates medical applications (fMRI/EEG analysis) while maintaining theoretical rigor in model architecture design. His scientific achievements include: Best paper award at EMNLP 2024 for 'Backward Lens: Projecting Language Model Gradients into the Vocabulary Space' Best paper award at SCIA 2023 for 'Gradient Adjusting Networks for Domain Inversion' Best paper award at FG 2021 for 'Generating Master Faces for Dictionary Attacks' Prof. Wolf mentors graduate students in the School of Computer Science and leads research at the ICRC building laboratory. His team collaborates with 'the friends of TAU' on projects spanning biometric security, medical imaging, and generative AI. Current work focuses on efficient transformers, neural network interpretability, and multimodal medical diagnostics.