Wei Xu is an Associate Professor at Georgia Institute of Technology's College of Computing and School of Interactive Computing, with affiliations to the Machine Learning Center. Their research bridges machine learning, natural language processing, and social media with focus areas in large language models, cultural bias mitigation, multilingual capabilities, and human-AI collaboration in text evaluation. NSF CAREER and Google Academic Research Award recipient Director of NLP X Lab PhD from New York University, BSMS from Tsinghua University Research interests span: Multilingual Multicultural LLMs addressing representational gaps and cultural adaptation in language models (NAACL 2025, ACL 2024); Robustness and Reasoning through dynamic AGI evaluations (ACL 2024, EMNLP 2024); Interdisciplinary NLP applications in security, healthcare, and law (EMNLP 2024, ACL 2024). Recent publications focus on multilingual alignment (NAACL 2025), privacy risk estimation (arXiv 2025), cultural bias analysis (ACL 2024), and medical text simplification (EMNLP 2024). Key themes include bias mitigation, multimodal processing, and practical LLM evaluation. Scientific Awards : NSF CAREER, Google/Sony/Criteo research awards, ACL'24 Best Social Impact Award, COLING'18 Best Paper Advising 15 PhD/MS/BSMS students including Yao Dou (human-centered LLM evaluation), Tarek Naous (multilingual LLMs), and alumni like Chao Jiang (Apple AI/ML) and Yang Chen (NVIDIA research scientist). Teaches graduate courses on NLP and LLMs.
Jean Oh is a Researcher at the Robotics Institute of Carnegie Mellon University (CMU) , leading the interdisciplinary Bot Intelligence Group (BIG) . Her work focuses on developing persistent robots that co-exist and collaborate with humans in shared environments, emphasizing continuous improvement through training, exploration, and human interaction. Education: Ph.D. in Language and Information Technologies, CMU M.S. in Computer Science, Columbia University B.S. in Biotechnology, Yonsei University Oh's research integrates vision, language, and planning systems in robotics, with applications in human-robot teaming , self-driving cars , disaster response , eldercare , and creative robotics . She has pioneered projects like socially-compliant robot navigation in human crowds and AI-driven robotic painting systems. Recent publication trends highlight her work in vision-language planning , social navigation , computational creativity , and human-robot collaboration . Notable contributions include the StyleCLIPDraw algorithm for text-to-art generation and Social-PatteRNN for human-like trajectory prediction. Scientific Awards: Best Paper Award in Cognitive Robotics (ICRA'18, ICRA'15) Best Systems Paper Finalist (HRI'25) Best Oral Paper Finalist (Humanoids'24) Best Paper in Entertainment (IROS'24) Argoverse Challenge Winner (CVPR'24) Best Student Paper (AIAA'24) Best Demo Finalist (RoboSoft'24) Oh mentors a diverse team of PhD, MS, and undergraduate students from CMU departments including Robotics, Computer Science, and Mechanical Engineering. Her research is funded by US Army Research Lab , DiDi Chuxing , and DARPA , with collaborations across industry and academia .
Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Christian Theobalt is a Professor of Computer Science at Saarland University and Scientific Director of the Visual Computing and Artificial Intelligence Department at the Max Planck Institute for Informatics . He leads the Saarbruecken Center for Visual Computing as a strategic partnership between Google and MPI. PhD in Computer Science (2005) from MPI-INF/Saarland University Postdoctoral Researcher at MPI (2005-2007) Visiting Assistant Professor at Stanford (2007-2009) His research focuses on the intersection of Computer Graphics, Computer Vision, and Artificial Intelligence , with specializations in: 3D/4D Human Reconstruction Neural Rendering Performance Capture Geometric Deep Learning Volumetric Video Quantum Visual Computing Recent article trends show emphasis on: Quantum computing applications in visual reconstruction Neural rendering with radiance fields Human motion capture from egocentric views Gesture-language interaction modeling Multi-modal scene understanding Real-time free-viewpoint rendering Scientific Awards : Fellow of EUROGRAPHICS (2022) CVPR Best Student Paper Honorable Mention (2020) ERC Consolidator Grant (2017) Karl Heinz Beckurts Award (2017) Busy Beaver Teaching Award (2016) ERC Starting Grant (2013) Advising over 50 students and researchers including: Current researchers: Viktor Rudnev, Linjie Lyu, Mohit Mendiratta Postdocs: Kwang In Kim, Kiran Varanasi Alumni: Franziska Mueller (Google), Dushyant Mehta (Qualcomm), Ayush Tewari (MIT) Grants include ERC grants, Google Glass Research Award, and multiple industry partnerships. His lab maintains cutting-edge facilities with: Multi-camera capture systems Quantum annealing infrastructure HDR radiance field technology GPU clusters for AI research Time-of-flight imaging systems Event camera arrays
Justin Johnson is an Assistant Professor at the University of Michigan's College of Engineering, Department of Electrical Engineering and Computer Science, and a Research Scientist at Facebook AI Research (FAIR). His work bridges computer vision, machine learning, and deep learning, focusing on visual reasoning, vision-language tasks, image generation, and 3D reasoning using neural networks. PhD, Stanford University (advised by Fei-Fei Li) His research interests span visual reasoning , vision and language , 3D vision , and image generation , with a focus on innovative applications of deep neural networks. Recent publications highlight work on 3D consistency, self-supervised learning, and multimodal integration of vision and text. Notable contributions include PyTorch3D for 3D data processing, and foundational work in visual question answering , neural style transfer , and scene graph-based image generation . Publications span top conferences like ICCV, CVPR, and NeurIPS. He teaches courses including EECS 498/598: Deep Learning for Computer Vision and EECS 442: Computer Vision at University of Michigan, with prior involvement in Stanford's CS 231N in co-teaching roles. Software projects include open-source frameworks like fast-neural-style for real-time artistic style transfer, and PyTorch3D for efficient 3D deep learning. These tools demonstrate his commitment to practical implementations and community-driven research.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Yuki M. Asano is a full Professor at the University of Technology Nuremberg , leading the Fundamental AI (FunAI) Lab . Previously, he led the QUVA Lab at the University of Amsterdam and earned his PhD at the Visual Geometry Group (VGG) of the University of Oxford under Andrea Vedaldi and Christian Rupprecht. University of Technology Nuremberg (2024–present) University of Amsterdam (prior to 2024) University of Oxford (PhD, 2020) His research spans Artificial Intelligence , Machine Learning , and Computer Vision , with a focus on Causal Representation Learning , Self-Supervised Learning , and Efficient Model Adaptation . He pioneered techniques like BISCUIT (causal variable identification) and VeRA (parameter-efficient fine-tuning). His work extends to Medical Imaging and Environmental Monitoring through applications in fetal ultrasound analysis and marine debris detection. Recent publications (2023–2025) highlight advancements in Self-Supervised Learning , Vision-Language Models , and 3D Understanding . Notable papers include TWIST & SCOUT (multimodal LLM grounding), SIGMA (masked video modeling), and GeneralAD (anomaly detection). His ICCV 2023 work on Self-Ordering Point Clouds and MoSiC (optimal-transport motion trajectories) underscores his interdisciplinary approach. He received the JUPITER compute grant (2025) and an Outstanding Paper Award at ICLR 2024 . His collaborations span institutions like MIT-IBM Watson AI Lab, Qualcomm AI Research, and University of Amsterdam.
Kristen Grauman is a Full Professor in the Department of Computer Science at the University of Texas at Austin, where she leads the UT Computer Vision Group. Her research focuses on computer vision and machine learning, with applications in visual recognition, video analysis, and multi-modal perception. She received her B.A. from Boston College and her Ph.D. from MIT. Her research interests span visual recognition, image and video search, video analysis, first-person vision, embodied and multi-modal perception, and interactive machine learning. She has made significant contributions to the field, particularly in developing algorithms for understanding visual content and human activities from video, including foundational work on the Pyramid Match Kernel and relative attributes. Her recent publications reveal a strong emphasis on egocentric (first-person) vision, audio-visual learning, and view-invariant representations. There is a clear trajectory toward multi-modal integration (vision, audio, language) and real-world applications in instructional videos, human activity understanding, and embodied AI systems. She has received numerous awards including: AAAI Fellow (2019) J. K. Aggarwal Prize, International Association for Pattern Recognition (2018) Helmholtz Prize (2017) UT Austin Academy of Distinguished Teachers (2017) Best Paper Award, Asian Conference on Computer Vision (2016) Presidential Early Career Award for Scientists and Engineers (2014) Computers and Thought Award, International Joint Conferences on Artificial Intelligence (2013) Pattern Analysis and Machine Intelligence Young Researcher Award (2013) Alfred P. Sloan Research Fellow (2012) Marr Prize (2011) Prof. Grauman serves as Associate Editor-in-Chief for the IEEE Transactions on Pattern Analysis and Machine Intelligence. She has secured substantial research funding including the Presidential Early Career Award, NSF grants, and industry partnerships. Her advising has produced numerous influential publications and students who are now leaders in computer vision. She leads the UT Computer Vision Group, which collaborates closely with the Electrical and Computer Engineering Department. The group is pioneering large-scale egocentric video research through projects like Ego4D and Ego-Exo4D, focusing on real-world applications in human activity understanding, audio-visual perception, and interactive systems.
Hao Su is an Associate Professor in the Department of Computer Science and Engineering at University of California, San Diego . He serves as Chairman & CTO of Hillbot Inc , and leads the SU Lab which focuses on building autonomous systems that learn actively in physical environments. His affiliations include the Institute for Learning-enabled Optimization at Scale , Artificial Intelligence Group , Contextual Robotics Institute , Halicioğlu Data Science Institute , and Center for Visual Computing . As a researcher in Computer Vision, Robotics, and Neural Geometry , he has made significant contributions to 3D foundation models, reward-free world models, diffusion policy frameworks, and GPU-accelerated simulation environments. His 2024-2025 publications include advancements in hand-eye calibration, dynamic mesh reconstruction, and multi-stage robotic manipulation. His scientific awards include: Frontiers of Science Award (2025) TPAMI Young Research Award (2025) NSF CAREER Award (2023) ACM SIGGRAPH Best Doctorate Thesis Honorable Mention (2019) He has served as Program Chair for CVPR 2025 and Area Chair for ICLR 2022 and NeurIPS 2023 , while previously serving as Publication Chair for 3DV 2016 and Program Committee for SIGGRAPH Asia Workshops .
Luca Carlone is the Boeing Career Development Associate Professor in the Department of Aeronautics and Astronautics at MIT and a Principal Investigator at the Laboratory for Information & Decision Systems (LIDS) . He leads the SPARK Lab , focusing on developing certifiable perception algorithms for autonomous systems. PhD in Mechatronics (Polytechnic University of Turin, 2012) Research spans robotics, computer vision, and optimization Research Interests : Certifiable Perception algorithms for high-integrity systems High-level Perception (geometric, semantic, physical understanding) Efficient Perception methods for resource-constrained robots Scientific Contributions include: 2024 Outstanding Systems Paper Award (RSS) 2023 IEEE Transactions on Robotics King-Sun Fu Award 2021 NSF CAREER Award 2020 AIAA Advising Award 2019 Amazon Research Award Advising : Teaches graduate courses like Visual Navigation for Autonomous Vehicles and Robotics: Science and Systems . Collaborates with institutions including JPL, Caltech, and KAIST through the DARPA SubT Challenge.
Kenji Kawaguchi is the Presidential Young Professor in the Department of Computer Science at the National University of Singapore (NUS), where he leads the Deep Learning Lab and is a faculty affiliate at the NUS Institute of Data Science. His research bridges theoretical and applied machine learning, focusing on deep learning, large language models, and physics-informed neural networks. His educational background includes a Ph.D. and S.M. in Computer Science and Electrical Engineering from the Massachusetts Institute of Technology (MIT), advised by Leslie Pack Kaelbling, and a postdoctoral fellowship at Harvard University’s Center of Mathematical Sciences and Applications. Dr. Kawaguchi’s research interests center on the theoretical foundations of deep learning, optimization, generalization, and applications in areas such as molecular modeling, AI safety, and efficient training of large models. He has made significant contributions to understanding in-context learning, diffusion models, and neural operators for partial differential equations. His recent publications (2023–2025) reflect a strong trend toward improving the efficiency, robustness, and interpretability of large-scale models, particularly in language and scientific domains. Key themes include LLM alignment and safety, diffusion model optimization, and physics-informed learning for high-dimensional problems. Presidential Young Professor He has served as Area Chair and PC Member for top-tier conferences including NeurIPS, ICML, ICLR, AAAI, and UAI, and as reviewer for journals such as JMLR and Annals of Statistics. He has delivered invited talks at Harvard, MIT, Stanford, CMU, Brown, and Google Research, reflecting his international recognition. He actively mentors students and welcomes PhD candidates and postdocs to join his research group.
Noah Snavely is a Professor of Computer Science at Cornell Tech and the Cornell Ann S. Bowers College of Computing and Information Science. His research spans computer vision and graphics with a focus on 3D scene understanding from image collections. He also works at Google DeepMind in NYC and leads the Cornell Graphics and Vision Group. His research interests include computer vision and computer graphics , particularly in recovering 3D structure from large photo collections, neural radiance fields, image-based rendering, and scene understanding. His work enables applications in mapping technologies, immersive VR experiences, and synthesizing 3D worlds from text prompts with implications for game design and filmmaking. Recent publications show a strong trend toward generative 3D modeling and neural rendering , with significant contributions in neural radiance fields, view synthesis, and dynamic scene reconstruction. His research group explores unstructured photo collections to develop new technology for 3D world modeling and image analysis. PECASE Microsoft New Faculty Fellowship Alfred P. Sloan Fellowship SIGGRAPH Significant New Researcher Award ACM Fellow (2023) IEEE Fellow NSF CAREER Award Helmholtz Prize Snavely has mentored numerous PhD students and postdocs including Ruojin Cai, Qianqian Wang, and Zhengqi Li. His teaching includes CS5670: Computer Vision across multiple semesters. He has received the Cornell College of Engineering Mr. and Mrs. Richard F. Tucker Teaching Excellence Award (2012). At Cornell Tech, his research group develops technology for modeling the world in 3D from online photo collections and analyzing images for computer graphics applications. His work has been featured in major news outlets including the New York Times, New Scientist, and MIT Technology Review.
Kyros Kutulakos is a Professor in the Department of Computer Science at the University of Toronto, where he leads research in computational imaging and 3D sensing. His affiliations include the Toronto Computational Imaging Group, Computer Vision Group, Dynamic Graphics Project (DGP), and Vector Institute Group. He teaches graduate and undergraduate courses such as CSC320 (Introduction to Visual Computing) and CSC2530 (Computational Imaging & 3D Sensing). His research interests span computational imaging, non-line-of-sight imaging, single-photon detectors, 3D sensing, and neural rendering. Notable contributions include advancements in structured-light imaging, time-of-flight systems, and super-oscillatory microscopy. He has advised numerous PhD and MSc students, fostering cutting-edge research in imaging technologies. Kutulakos has received prestigious awards, including the Dean’s Research Excellence Award (2023) and multiple best paper prizes (e.g., Marr Prize at ICCV 2023). He has served as program chair for ICCV 2013, ICCP 2010, and CVPR 2003, contributing to academic leadership in computer vision. His work bridges optics, photonics, and computation, with applications in autonomous systems, medical imaging, and astronomy. Current research focuses on extreme imaging scenarios, such as imaging in pitch-black environments and around corners, leveraging novel sensor designs and computational techniques.
Adriana Kovashka is an Associate Professor at the University of Pittsburgh , affiliated with the School of Computing and Information and serving as Department Chair . Her academic journey began with BA degrees in Computer Science and Media Studies from Pomona College (2008) and a PhD in Computer Science from The University of Texas at Austin (2014). Joined Pitt’s faculty in January 2015 NSF CAREER awardee (2021) Google Faculty Research Award recipient Dr. Kovashka’s research spans Computer Vision , Machine Learning , and Natural Language Processing , focusing on visual rhetoric, weak multimodal supervision, and domain adaptation. She pioneered techniques for analyzing political imagery, developing robust object detection frameworks, and exploring the intersection of visual and textual persuasion through large-scale annotated datasets. Her recent work emphasizes geographic diversity in vision-language systems, audio-visual fusion for domain generalization, and shape-texture bias mitigation in CNNs. Key publications include groundbreaking studies on symbolic reasoning, multimodal dialogue systems, and ethical AI applications in education. Scientific honors include: NSF CRII Award (2016) NSF CAREER Award (2021) Pitt CRDF Award (2016, 2018) Best Paper at ECV Workshop (2021) Google Faculty Research Award (2016, 2018) Dr. Kovashka actively mentors students in multimodal learning projects and collaborates with interdisciplinary teams on NSF-funded initiatives. She co-organizes workshops like the first CVPR workshop on advertisement understanding and leads research groups exploring human-AI co-learning systems.