Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Michael Kaess is an Associate Professor at the Robotics Institute, Carnegie Mellon University (CMU), within the School of Computer Science. He leads the Robot Perception Lab (RPL) and contributes to the Field Robotics Center (FRC) and Computer Vision Group (CV). His research focuses on efficient perception algorithms for mobile robots, particularly in 3D mapping, SLAM, and sensor fusion using vision, LiDAR, inertial, and sonar data. Kaess holds a PhD in Computer Science from Georgia Tech and was a postdoc at MIT's Marine Robotics Lab. Education: Georgia Institute of Technology, PhD in Computer Science (2008) MIT, Postdoctoral Associate (2008–2010) Research Interests: Kaess develops algorithms for robust and efficient inference in robotics, emphasizing factor graphs and linear algebra. His work spans underwater robotics, aerial systems, tactile SLAM, and multi-sensor integration. Key areas include SLAM with planes/lines, imaging sonar reconstruction, and neural field methods for LiDAR-visual fusion. Publications: Over 145 papers, including work on EDPLVO (visual odometry), HoloOcean (underwater simulation), and neural radiance fields with LiDAR. Recent trends focus on robust incremental smoothing, acoustic-optical fusion, and real-time volumetric mapping. Awards: Recognized with the RSS Test of Time Award (2020), Outstanding Associate Editor (2022), and paper awards at ICRA/ICRA. Active in conference organization (IROS/ICRA program committees). Advising & Grants: Supervises 10+ current PhD/MSc students, with past advisees contributing to CoRL/ICRA work. Manages grants in perception, autonomy, and marine robotics. Teaches courses like Robot Localization and Mapping (16-833). Labs/Teams: Directs RPL, collaborates with FRC on field robotics. Develops open-source tools like GTSAM (GNU Toolkit for Smoothing and Mapping).
Aarti Singh is a Professor in the Machine Learning Department at Carnegie Mellon University and Director of the NSF AI Institute for Societal Decision Making. She leads research at the intersection of machine learning, statistics, and decision making, with applications to scientific and societal domains. Her work focuses on designing principled interactive algorithms for learning and decision making under uncertainty. Education: Ph.D. in Electrical Engineering, University of Wisconsin-Madison (2008) M.S. in Electrical Engineering, University of Wisconsin-Madison (2003) B.E. in Electronics and Communication Engineering, University of Delhi (2001) Research Interests: Professor Singh's research centers on developing interactive machine learning algorithms that go beyond finding input-output associations to make higher-level decisions about the most informative data and actions. Her work spans autonomous decision making, including active sampling, stochastic optimization, bandits, and reinforcement learning that are statistically optimal, computationally tractable, and robust. She also investigates human factors in decision making, designing algorithms that model and leverage human feedback while accounting for bias, memory effects, and calibration. Her research has applications in material science, cosmology, and peer review systems. Research Trends: Professor Singh's recent publications demonstrate a strong focus on reinforcement learning, particularly in developing more efficient and robust algorithms for decision making under uncertainty. Her work bridges theoretical foundations with practical applications, spanning from fundamental algorithm development to real-world implementation in scientific domains. There's a clear trajectory toward integrating human factors into decision-making algorithms, with significant contributions to peer review systems and preference learning. Scientific Awards: NSF Career Award United States Air Force Young Investigator Award A. Nico Habermann Faculty Chair Award Harold A. Peterson Best Dissertation Award Multiple paper awards Advising and Grants: Professor Singh has advised numerous PhD and master's students, many of whom have gone on to faculty positions or research roles at leading institutions. Her research is supported by prestigious grants from ONR, Simons Foundation, AFRL, ARL, and NSF. She serves as General Chair (2025) and Program Chair (2020) for the International Conference on Machine Learning (ICML) and has held leadership roles in multiple professional organizations. Research Team: Professor Singh leads a vibrant research group within the Machine Learning Department at CMU, with current PhD students working on topics including reinforcement learning, human-AI collaboration, and decision making under uncertainty. She also directs the NSF AI Institute for Societal Decision Making, which brings together researchers from multiple disciplines to develop AI systems that support human decision making in societal contexts.
Furong Huang is an Associate Professor at the University of Maryland's Department of Computer Science, with affiliations at the Institute for Advanced Computer Studies, Center for Machine Learning, Maryland Robotics Center, and Applied Mathematics, Statistics, and Scientific Computation Program. Her research bridges trustworthy machine learning, sequential decision-making, and foundation models for robotics, emphasizing reliability, interpretability, and ethical standards. Research Interests: Trustworthy AI Generative AI Reinforcement Learning AI Security Algorithmic Fairness Foundation Models for Robotics Recent Publications span leading conferences (NeurIPS, ICML, ICLR, CVPR) and journals, focusing on: Robustness in Vision-Language Systems Trustworthy Generative AI Foundation Models for Sequential Decision-Making AI Security and Watermarking Scientific Awards MIT TR35 Innovator Under 35 (Asia Pacific 2022) Best Paper Award, AdvML Frontier Workshop, NeurIPS 2024 NSF NAIRR Pilot Awardee Microsoft Accelerate Foundation Models Research Award (2023) JP Morgan Faculty Research Awards (2019–2022) Advising and Grants : Her lab has graduated students to roles at OpenAI, Google, Meta, and Netflix. Research funded by DARPA, NSF, ONR, AFOSR, and industry partners like Microsoft, Adobe, and Capital One. Labs & Teams : Leads research groups focused on Trustworthy AI and Robotics at the University of Maryland, collaborating with the Maryland Robotics Center and Applied Mathematics Program.
Prof. Luke Zettlemoyer is an Adjunct Professor of Computer Science and Engineering at the University of Washington, with affiliations to the Department of Linguistics. He focuses on machine learning, natural language processing, and multimodal systems, contributing to advancements in large language models, ethical AI, and scalable architectures. His research addresses challenges in model alignment, generalization, and cross-domain integration. Key research interests include multimodal reward models, efficient tokenization strategies, and model optimization techniques. He has explored topics such as neural trajectories for robot learning, content-adaptive image processing, and ethical mitigation of verbatim data reproduction. His publications span 2023–2025, emphasizing practical applications of AI in robotics, vision-language systems, and scalable retrieval-based models. While no formal awards are listed, his work reflects significant contributions to foundational AI research.
Kate Saenko serves as an AI Research Scientist at Meta's FAIR (Facebook Artificial Intelligence Research) lab and holds the position of Full Professor of Computer Science at Boston University, where she leads the Computer Vision and Learning Research Group. Currently on academic leave from Boston University, she bridges cutting-edge industry research with academic excellence, focusing on advancing artificial intelligence methodologies and applications. Her educational background includes a PhD in Electrical Engineering and Computer Science (EECS) from the Massachusetts Institute of Technology (MIT), followed by postdoctoral training at the University of California, Berkeley and Harvard University. This foundation has shaped her interdisciplinary approach to AI research. Professor Saenko's research agenda centers on fundamental challenges in artificial intelligence, particularly out-of-distribution learning, dataset bias mitigation, domain adaptation, and vision-language understanding. Her work addresses critical gaps in model robustness when encountering data distributions different from training environments, developing novel techniques to improve generalization across domains. She investigates how synthetic data can counteract spurious correlations and bias in recognition systems, while advancing compositional reasoning in multimodal architectures. Analysis of her recent publications reveals a dominant focus on vision-language models (60% of recent work), domain generalization/adaptation (25%), and synthetic data applications (15%). Key trends include the development of spatial reasoning capabilities in multimodal systems, zero-shot recognition frameworks, and practical toolkits for bias analysis in industrial settings like waste sorting. Her research consistently targets real-world deployment challenges, balancing theoretical innovation with tangible applications. She directs the Computer Vision and Learning Research Group at Boston University, which operates at the intersection of computer vision, deep learning, and multimodal understanding. The group maintains strong industry collaborations through Meta's FAIR and previously engaged with the MIT-IBM Watson AI Lab. Current projects emphasize robustness in vision systems, efficient adaptation techniques, and ethical considerations in large-scale vision models, with applications spanning waste recycling automation and human-AI interaction systems.
Katerina Fragkiadaki is the JPMorgan Chase Associate Professor of Computer Science in the Machine Learning Department at Carnegie Mellon University. She works at the intersection of Artificial Intelligence, Computer Vision, Machine Learning, Language Understanding, and Robotics. PhD from GRASP Lab, University of Pennsylvania Postdoctoral researcher at UC Berkeley (with Jitendra Malik) and Google Research Recipient of NSF CAREER, DARPA Young Investigator, Amazon, Google, Sony, UPMC, and AFOSR awards Organizer of CoRL 2023 Workshop on Generalist Robots ICLR 2024 Program Chair, multiple area chair roles Her research group focuses on developing machines that autonomously improve world models through human-environment interactions, with specific emphasis on: Representation learning and video understanding 2D/3D unified vision-language models Generative simulation and reinforcement learning Real2Sim/Sim2Real robot learning Continual learning and spatial common sense 3D scene reconstruction and dynamics Recent publications highlight advancements in: 3D mesh generation with compositional transformers Unified 2D/3D perception frameworks Physics-aware generative models Diffusion-based robotic manipulation policies Embodied agents with memory prompting Awards include: 2024: DARPA Young Investigator Award 2023: Amazon Faculty Award 2022: Sony Faculty Research Award 2021: UPMC Faculty Research Award 2020: NSF CAREER Award 2019: Google Faculty Award Key collaborations span institutions including UC Berkeley, Google Research, Stanford, MIT, and University of Tsukuba. Her work bridges theoretical innovation with practical applications in: Autonomous robot manipulation 4D world modeling Language-grounded perception Visual dynamics prediction Embodied program synthesis Physics-based simulation engines
Arti Singh is an Assistant Professor in the Department of Agronomy at Iowa State University. Her research focuses on plant breeding, soybean diseases, genomics, and phenomics, with a strong emphasis on integrating artificial intelligence and high-throughput technologies into agricultural systems. She leads projects involving AI-driven disease identification, precision agriculture, and crop improvement strategies. Her expertise includes developing machine learning models for real-time weed and insect classification (e.g., WeedNet and InsectNet), deploying drones and ground robots for crop phenotyping, and leveraging genomic data to map traits like flowering time and disease resistance in legumes. Singh collaborates on initiatives like the AIIRA Institute for Resilient Agriculture and the BioTrove biodiversity dataset. Singh’s work spans plant stress phenotyping, digital twin technologies for plant sciences, and multi-sensor phenotyping for early disease detection. Her research bridges computational methods with traditional agronomy, aiming to enhance crop resilience and sustainability in the face of environmental challenges. Her recent projects include optimizing robotic navigation for precision agriculture, improving soybean yield estimation via video analysis, and dissecting genetic architectures of traits in mungbean and soybean using GWAS and genomic tools. She actively contributes to conferences and publishes in high-impact journals, advancing both foundational and applied aspects of agricultural science.
Alan Ritter is an Associate Professor at the School of Interactive Computing , Georgia Institute of Technology, with additional affiliation to the Machine Learning Center . His research focuses on Natural Language Processing , particularly robust models across domains/languages with fewer labels and efficient resource use, plus data-driven dialogue agents for open-topic conversations. Research Interests : Robust NLP models, cross-lingual transfer, resource-efficient learning, dialogue systems, cultural bias measurement, and privacy-aware language models Students : Mentors Ph.D. students in Georgia Tech's ML and CS programs, including Junmo Kang, Yang Chen, and Duong Minh Le. Alumni include Fan Bai (Ph.D. 2023), Yang Chen (Ph.D. 2024), and Andrew Li (M.S. 2024). Awards : NSF CAREER Award, Amazon Research Award, ACL 2024 Best Social Impact Paper, IUI 2009 Best Student Paper. Recent Work : Studies training budget allocation between supervised and preference-based finetuning, cross-lingual information extraction, cultural bias in LLMs, and privacy risk mitigation in social media disclosures. Service : Served as Program Chair for NAACL 2025, Area Chair for multiple top-tier conferences (COLM, EMNLP, ACL, EACL, AAAI). Email : alan.ritter@cc.gatech.edu
Hanjie Chen is an Assistant Professor in the Department of Computer Science at Rice University, affiliated with the Ken Kennedy Institute. She holds a Ph.D. from the University of Virginia and a Master's from the University of Science and Technology of China. Her research focuses on Natural Language Processing, Interpretable Machine Learning, and Trustworthy AI, emphasizing model explainability, alignment with human needs, and applications in healthcare, sports, and medicine. She has advised numerous students and led initiatives in AI ethics and education. Education: Ph.D. (Computer Science, UVA 2023), M.Sc. (USTC 2018), B.Sc. (Nanjing University of Aeronautics and Astronautics 2015). Awards include the Outstanding Doctoral Student Award (UVA 2023) and John A. Stankovic Research Award (UVA 2023). She has organized workshops like BlackboxNLP and served on program committees for ACL, NAACL, and EMNLP. Her recent work includes developing benchmarks like SPORTU for multimodal LLMs, evaluating medical question-answering systems, and advancing methods for robust rationale evaluation (RORA). She teaches courses on Natural Language Processing and Trustworthy NLP, emphasizing pedagogical innovation recognized by teaching awards at UVA. Research collaborations include internships at Microsoft Research, IBM, and the Allen Institute for AI. She mentors students in SURF programs and advocates for diversity in tech, serving as a mentor in UVA's CSGSG Council.
Olga Sorkine-Hornung is a Professor of Computer Science at ETH Zurich and head of the Institute of Visual Computing. She leads the Interactive Geometry Lab, focusing on theoretical and practical advancements in digital content creation, geometry processing, and shape modeling. Current position: ETH Zurich, Department of Computer Science Previous roles: Courant Institute (NYU), Technical University of Berlin Education: BSc and PhD from Tel Aviv University, postdoc at TU Berlin Her research spans shape representation, digital fabrication, computer animation, and fundamental geometry processing. Key contributions include Laplacian surface editing, as-rigid-as-possible deformation, and generalized winding numbers. She works on applications in VR, AR, and autonomous systems. Awards include Test of Time Awards (2024), ACM Fellow (2020), ERC Consolidator Grant (2020), and EUROGRAPHICS Young Researcher Award (2008). She has supervised numerous students and co-developed software libraries like libigl and Instant Meshes . Co-chair roles for SIGGRAPH, Eurographics, and Pacific Graphics Editorial board member for ACM Transactions on Graphics and other journals Keynote speaker at VMV, CVPR, and SIAM conferences Her work bridges mathematical rigor with practical implementation, advancing computer graphics and geometry processing through intuitive algorithms that maintain surface detail while enabling efficient computation.
Roberto Martin-Martin is an Assistant Professor of Computer Science at the University of Texas at Austin, where he leads the Robot Interactive Intelligence (RobIN) Lab. His research bridges robotics, computer vision, and machine learning to enable robots to operate autonomously in human-centric environments like homes and offices. Previously, he was a Postdoctoral Scholar at the Stanford Vision and Learning Lab working with Fei-Fei Li and Silvio Savarese, and an AI Researcher at Salesforce AI. Education: Ph.D. and M.Sc. in Robotics from Technische Universität Berlin (TUB), advised by Professor Oliver Brock B.Sc. from Universidad Politécnica de Madrid Dr. Martin-Martin's research focuses on developing AI algorithms that combine reinforcement learning and imitation learning with advanced planning and control to address core challenges in robot perception. His work spans mobile and whole-body manipulation, dexterous and contact-rich interactions, and long-horizon tasks in unstructured environments. He takes inspiration from human cognition through psychology and cognitive science to develop solutions for skills ranging from simple pick-and-place operations to complex tasks like cooking and furniture assembly. His recent publications demonstrate a strong trend toward enabling robots to learn from human demonstrations, particularly through video, and to safely adapt these demonstrations to their own morphology. There's significant emphasis on mobile manipulation, bimanual tasks, and developing hardware that supports robust robot learning through trial and error. His work shows increasing integration of large language models and vision-language models to enhance robot understanding and task execution. Scientific Awards: RSS Pioneer (2020) Winner of Amazon Picking Challenge (2015) RSS Best Systems Paper Award (2016) ICRA Best Paper Award IROS Best Mechanism Award Amazon Faculty Award AAAI Young Faculty IJCAI Early Faculty Nominated for Best Paper at IROS (2014, 2017) Dr. Martin-Martin advises PhD students including Arpit Bahety, who is working on mobile manipulation and learning. He serves as Chair of the IEEE Technical Committee on Mobile Manipulation and is a co-founder of QueerInRobotics. His research is supported by industry partnerships and academic funding sources that enable his lab to develop both hardware and software innovations in robotics. He directs the Robot Interactive Intelligence (RobIN) Lab at UT Austin, which takes a holistic approach to robot intelligence, developing both the hardware (like the BaRiFlex gripper) and software frameworks necessary for robots to learn from interaction. The lab's research addresses the full pipeline from perception to action, with particular emphasis on learning from human demonstrations, safe exploration, and adapting to novel objects and environments.
Alexei A. Efros is a Professor in the Department of Electrical Engineering and Computer Sciences (EECS) at UC Berkeley, where he holds the Howard Friesen Professorship and is affiliated with the Berkeley Artificial Intelligence Research (BAIR) Lab. He previously served on the faculty at the Robotics Institute of Carnegie Mellon University (CMU) and completed a postdoctoral fellowship at the University of Oxford. His research spans data-driven computer vision, self-supervised learning, computational photography, and applications to computer graphics and robotics. His research interests include: Data-Driven Computer Vision Self-Supervised and Unsupervised Learning Generative Models and Image Synthesis Visual Representation Learning Applications in Robotics and Human-Computer Interaction Intersections with Human Vision and the Humanities The recent publications highlight a strong trend toward self-supervised learning, visual reasoning, and generative modeling, particularly diffusion models and 3D scene understanding. His work increasingly bridges computer vision with language, robotics, and cognitive science, emphasizing interpretability and real-world applicability. There is a clear focus on leveraging unlabeled data and developing methods for robust, generalizable AI systems. His scientific awards and recognitions include: Berkeley Fellowship Google Fellowship Soros Fellowship NSF Fellowship SIGGRAPH Outstanding Doctoral Dissertation Award Facebook Fellowship Adobe Fellowship CMU School of Computer Science Distinguished Dissertation Award ACM Doctoral Dissertation Honorable Mention Alexei Efros has advised numerous PhD students and postdocs, many of whom have gone on to faculty positions at top institutions including CMU, Stanford, MIT, Columbia, NYU, and Georgia Tech. His lab has received research funding from major tech companies and federal agencies, though specific grants are not detailed in the text. He teaches core computer vision and machine learning courses at both undergraduate and graduate levels at UC Berkeley. His research group is highly active, with ongoing projects in 3D perception, generative modeling, and vision-language systems. He leads a vibrant research lab at UC Berkeley, part of the BAIR consortium, collaborating with leading researchers such as Jitendra Malik, Trevor Darrell, Pieter Abbeel, and Angjoo Kanazawa. His lab fosters strong interdisciplinary connections with institutions worldwide, including Oxford, INRIA, and École Normale Supérieure.
Andrew Zisserman is a Royal Society Research Professor at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). His research focuses on computer vision, artificial intelligence, and neural networks, with significant contributions to multimodal learning, video understanding, and 3D scene analysis. He leads projects exploring visual-language models, audio-visual synchronization, and clinical imaging applications. Key research areas include: Video analysis and temporal modeling Multimodal systems for sign language translation and action recognition 3D shape estimation and physical property inference Foundation models and cross-modal retrieval Recent work highlights: Developed Flamingo and Tapir models for video-language tasks Advancements in spinal MRI analysis and clinical imaging Leadership in EGO4D and VoxCeleb challenges Honors include Fellowship of the Royal Society (FRS) and the ISSLS Prize in Clinical Science 2023 for spinal analysis innovations. His lab collaborates globally, emphasizing real-world applications in healthcare and autonomous systems.