Mark Y. Liberman is the Christopher H. Browne Distinguished Professor of Linguistics and Trustee Professor at the University of Pennsylvania. He holds a joint appointment in the Department of Linguistics and the Department of Computer and Information Science. His roles include Director of the Linguistic Data Consortium (LDC), Faculty Director of Ware College House, and former Director of the Institute for Research in Cognitive Science. Education: A.B. in Linguistics and Applied Mathematics from Harvard University (1965–1969), M.S. (1972) and Ph.D. (1975) in Linguistics from MIT. Research focuses on corpus-based phonetics, clinical linguistics applications, tonal phonology, formal models for linguistic annotation, and computational linguistics. He explores speech production, prosody, and interdisciplinary topics like language evolution and neurobiology of speech. Recent articles highlight advancements in speech biomarkers for neurodegenerative diseases, autism analysis, and computational linguistics. Awards include Fellowships from the AAAS and Linguistic Society of America. He advises graduate students and leads large-scale language resource initiatives like LDC, contributing to open-access linguistic datasets. Labs/Teams: Linguistic Data Consortium (LDC), Institute for Research in Cognitive Science (IRCS), and collaborations in computational linguistics and neuroscience.
Andrea Vedaldi is a Professor of Computer Vision and Machine Learning at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). He specializes in unsupervised methods for understanding images and videos, focusing on 3D geometry and semantics. His research bridges foundational AI and practical applications, with contributions to generative models, neural fields, and self-supervised learning. Education: PhD in Computer Science (2008), University of California, Los Angeles MSc in Computer Science (2005), UCLA BSc in Information Engineering (2003), University of Padua Research Interests: Unsupervised learning, 3D perception, generative AI, neural rendering, and scalable vision systems. His work emphasizes ethical, responsible AI aligned with ERC-funded projects like UNION (ERC Consolidator Grant). Key Contributions: Co-developer of VLFeat and MatConvNet libraries Leader in 3D reconstruction and diffusion models (e.g., CatFree3D) Recipient of the PAMI Thomas S. Huang Prize and multiple best paper awards Grants & Service: Principal Investigator on £2.3M ERC Consolidator Grant (UNION) Co-organizer of major conferences (ECCV 2020 Program Chair, CVPR 2023 Area Chair) Reviewer for top journals/conferences (PAMI, CVPR, NeurIPS) Labs & Teams: VGG Group at Oxford, collaborating on projects like Meta 3D Gen and Common Objects in 3D (CO3D).
Kai-Wei Chang is an Associate Professor at the University of California, Los Angeles (UCLA) in the Department of Computer Science, part of the Henry Samueli School of Engineering. He is also an Amazon Scholar at Alexa AI, focusing on advancing trustworthy AI and multimodal foundation models. His research bridges NLP, machine learning, and ethical AI, with a focus on fairness, robustness, and bias mitigation in language and vision-language systems. Education: Ph.D. in Computer Science (UIUC, 2015), M.S. and B.S. in Computer Science and Electrical Engineering from National Taiwan University. Research Interests: Trustworthy NLP (fairness, robustness), Multimodal Foundation Models (e.g., VisualBERT, GLIP), Reasoning in LLMs, and mitigating societal biases in AI systems. Notable contributions include pioneering work on aligning NLP models with human values and developing SOTA multimodal models like DesCo and GLIP. Awards: Sloan Research Fellowship (2021), Okawa Grant (2018), EMNLP Best Paper (2017), KDD Best Paper (2010). His work is funded by NSF, IARPA, ONR, and industry partners like Amazon, Google, and Facebook. Service: VP-Elect of SIGDAT, Ethics Committee Chair (NAACL 2022), Organizer of Trustworthy NLP Workshops, and Senior Area Chair for top conferences (ACL, NeurIPS, AAAI). Labs/Teams: Leads the UCLA Natural Language Processing Group, fostering interdisciplinary research on ethical AI and multimodal systems.
Christian Theobalt is a Professor of Computer Science at Saarland University and Scientific Director of the Visual Computing and Artificial Intelligence Department at the Max Planck Institute for Informatics . He leads the Saarbruecken Center for Visual Computing as a strategic partnership between Google and MPI. PhD in Computer Science (2005) from MPI-INF/Saarland University Postdoctoral Researcher at MPI (2005-2007) Visiting Assistant Professor at Stanford (2007-2009) His research focuses on the intersection of Computer Graphics, Computer Vision, and Artificial Intelligence , with specializations in: 3D/4D Human Reconstruction Neural Rendering Performance Capture Geometric Deep Learning Volumetric Video Quantum Visual Computing Recent article trends show emphasis on: Quantum computing applications in visual reconstruction Neural rendering with radiance fields Human motion capture from egocentric views Gesture-language interaction modeling Multi-modal scene understanding Real-time free-viewpoint rendering Scientific Awards : Fellow of EUROGRAPHICS (2022) CVPR Best Student Paper Honorable Mention (2020) ERC Consolidator Grant (2017) Karl Heinz Beckurts Award (2017) Busy Beaver Teaching Award (2016) ERC Starting Grant (2013) Advising over 50 students and researchers including: Current researchers: Viktor Rudnev, Linjie Lyu, Mohit Mendiratta Postdocs: Kwang In Kim, Kiran Varanasi Alumni: Franziska Mueller (Google), Dushyant Mehta (Qualcomm), Ayush Tewari (MIT) Grants include ERC grants, Google Glass Research Award, and multiple industry partnerships. His lab maintains cutting-edge facilities with: Multi-camera capture systems Quantum annealing infrastructure HDR radiance field technology GPU clusters for AI research Time-of-flight imaging systems Event camera arrays
Russell Epstein is a Professor and Director of Graduate Studies in the Department of Psychology at the University of Pennsylvania. He is affiliated with the Center for Cognitive Neuroscience and Goddard Labs. His research focuses on neural mechanisms underlying visual scene perception, spatial navigation, and memory. Epstein holds a BA in Physics from the University of Chicago and a PhD in Applied Mathematics from Harvard University. Epstein’s research interests include high-level vision, spatial cognition, and the neural basis of environmental representations. His lab uses functional MRI and cognitive neuroscience techniques to study how scenes, objects, landmarks, and spaces are encoded in brain systems such as the parahippocampal place area and retrosplenial cortex. Recent work explores cognitive maps, grid-like neural representations, and the role of multisensory cues in navigation. His articles emphasize spatial navigation strategies, hierarchical cognitive maps, and the interplay between perception and memory. Notable contributions include investigations into hippocampal spatial metrics, olfactory navigation, and the neural underpinnings of environmental learning. Epstein teaches courses on cognitive neuroscience, including PSYC 149 and PSYC 600. He advises two graduate students in Psychology and has no listed scientific awards. His work is supported by grants (unspecified) and conducted within collaborative teams at the Center for Cognitive Neuroscience. Epstein’s research extends to labs focused on spatial cognition and neuroimaging, advancing understanding of how humans mentally map environments through visual and sensory integration.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Yuki M. Asano is a full Professor at the University of Technology Nuremberg , leading the Fundamental AI (FunAI) Lab . Previously, he led the QUVA Lab at the University of Amsterdam and earned his PhD at the Visual Geometry Group (VGG) of the University of Oxford under Andrea Vedaldi and Christian Rupprecht. University of Technology Nuremberg (2024–present) University of Amsterdam (prior to 2024) University of Oxford (PhD, 2020) His research spans Artificial Intelligence , Machine Learning , and Computer Vision , with a focus on Causal Representation Learning , Self-Supervised Learning , and Efficient Model Adaptation . He pioneered techniques like BISCUIT (causal variable identification) and VeRA (parameter-efficient fine-tuning). His work extends to Medical Imaging and Environmental Monitoring through applications in fetal ultrasound analysis and marine debris detection. Recent publications (2023–2025) highlight advancements in Self-Supervised Learning , Vision-Language Models , and 3D Understanding . Notable papers include TWIST & SCOUT (multimodal LLM grounding), SIGMA (masked video modeling), and GeneralAD (anomaly detection). His ICCV 2023 work on Self-Ordering Point Clouds and MoSiC (optimal-transport motion trajectories) underscores his interdisciplinary approach. He received the JUPITER compute grant (2025) and an Outstanding Paper Award at ICLR 2024 . His collaborations span institutions like MIT-IBM Watson AI Lab, Qualcomm AI Research, and University of Amsterdam.
David Bamman is an Associate Professor in the School of Information at UC Berkeley, specializing in applying Natural Language Processing (NLP) and machine learning to cultural and social science questions. He leads research in born-literary NLP, computational humanities, and cultural analytics, with affiliated roles in EECS, Linguistics, and Computational Precision Health. Bamman holds degrees from Carnegie Mellon (Ph.D., 2015), Boston University (M.A., 2006), and University of Wisconsin-Madison (B.A., 1998). His work is supported by NEH, NSF, and industry grants. Educations: Ph.D. in Computer Science (2015), Carnegie Mellon University M.A. in Applied Linguistics (2006), Boston University B.A. in Classics (1998), University of Wisconsin-Madison Research Interests: NLP for underserved domains (e.g., literature, social media), coreference resolution, cultural analytics, and computational methods for studying literature and culture. Projects include LitBank and BookNLP datasets. Grants & Awards: Hellman Fellow (2019), Amazon Research Award (2017), NSF CAREER Award, and NEH funding. Teaching: Courses include Natural Language Processing (Info 159/259), Computational Humanities (INFO 190), and Applied NLP (INFO 256). His research group explores topics like racial representation in high school literature, Hollywood diversity metrics, and the sociocultural implications of LLMs. Bamman advises multiple PhD students and collaborates on datasets like CMU Book Summaries and 11K Latin Books.
Luca Carlone is the Boeing Career Development Associate Professor in the Department of Aeronautics and Astronautics at MIT and a Principal Investigator at the Laboratory for Information & Decision Systems (LIDS) . He leads the SPARK Lab , focusing on developing certifiable perception algorithms for autonomous systems. PhD in Mechatronics (Polytechnic University of Turin, 2012) Research spans robotics, computer vision, and optimization Research Interests : Certifiable Perception algorithms for high-integrity systems High-level Perception (geometric, semantic, physical understanding) Efficient Perception methods for resource-constrained robots Scientific Contributions include: 2024 Outstanding Systems Paper Award (RSS) 2023 IEEE Transactions on Robotics King-Sun Fu Award 2021 NSF CAREER Award 2020 AIAA Advising Award 2019 Amazon Research Award Advising : Teaches graduate courses like Visual Navigation for Autonomous Vehicles and Robotics: Science and Systems . Collaborates with institutions including JPL, Caltech, and KAIST through the DARPA SubT Challenge.
Mark Steedman is a Professor in the School of Informatics at the University of Edinburgh, where he conducts research in Artificial Intelligence, Computational Cognitive and Social Science, and Natural Language and Speech Processing. He is affiliated with the Institute for Language, Cognition and Computation (ILCC), the Centre for Speech Technology Research (CSTR), and the Human Communications Research Center (HCRC). He also holds an adjunct professorship in Computer and Information Science at the University of Pennsylvania. His research focuses on Combinatory Categorial Grammar (CCG) , computational linguistics , prosody and intonation , temporal semantics , gesture in communication , and computational music analysis . He has authored foundational books including Surface Structure and Interpretation , The Syntactic Process , and Taking Scope . The recent publications reflect a strong trend toward integrating formal grammatical frameworks like CCG with modern neural and distributional models, particularly in semantic parsing, entailment reasoning, and cognitive modeling. His work bridges symbolic and statistical approaches in NLP, often focusing on robust, wide-coverage parsing and semantic interpretation. Best Paper Award at AACL/IJCNLP 2023 for 'Smoothing Entailment Graphs with Language Models' Best Paper Award at ACL 2023 for 'Extrinsic Evaluation of Machine Translation Metrics' Influential Paper Award 2017 from IFAAMAS for 'Animated Conversation' Mark Steedman has supervised numerous PhD students and collaborated widely across institutions. He leads research in formal grammar applications to cognitive modeling, dialogue, and multimodal communication. His lab contributes to CCG software and semantic parsing tools, and he continues to be actively involved in advancing the integration of symbolic and neural AI.
Professor Andrew Davison holds the position of Professor of Robot Vision at Imperial College London's Department of Computing. He leads the Dyson Robotics Laboratory and the Robot Vision Research Group, focusing on advancing SLAM (Simultaneous Localization and Mapping) and Spatial AI. His groundbreaking work includes the MonoSLAM algorithm (2003), enabling real-time 3D vision for robotics and AR/VR. Current research emphasizes scalable, semantic-rich Spatial AI systems, as outlined in his FutureMapping papers (2018–2019). Education: BA in Physics (Oxford, 1994), D.Phil. (Oxford, 1998). Postdoctoral work at AIST, Japan (1998–2000), followed by a lectureship at Imperial (2002–present). Industrial collaborations include SLAMcore, a Spatial AI startup, and Dyson Robotics Lab. Over 18 PhD students supervised, many now leading roles at Meta, NVIDIA, SLAMcore, and academia. Notable contributions include DTAM, KinectFusion, and Event Camera SLAM. Recognized for software tools like SceneLib and contributions to robotics benchmarks (SLAMBench). Active on Twitter (@AjdDavison) for research updates.
Lerrel Pinto is an Assistant Professor of Computer Science at the Courant Institute of Mathematical Sciences at New York University (NYU), where he leads the General-purpose Robotics and AI Lab (GRAIL) as part of the CILVR research group. His work bridges the gap between theoretical machine learning and practical robotics applications, with a focus on enabling robots to generalize and adapt in real-world environments. Dr. Pinto received his undergraduate degree from IIT Guwahati, followed by a PhD from the Robotics Institute at Carnegie Mellon University (CMU). He then completed a postdoctoral fellowship at the University of California, Berkeley before joining NYU as faculty. His research program centers on robot learning and decision making, with several key thrusts that demonstrate his innovative approach to robotics. Pinto's work emphasizes large-scale learning techniques that leverage both extensive data and sophisticated model architectures. A significant portion of his research focuses on representation learning for sensory data, particularly developing methods that enable robots to make sense of visual, tactile, and auditory inputs. His lab has made notable contributions to reinforcement learning algorithms that allow robots to adapt to new scenarios with minimal retraining. Pinto also champions open-source robotics , developing affordable robot platforms that democratize access to robotics research. Analysis of Pinto's recent publications reveals a strong trend toward multimodal perception in robotics, integrating visual, tactile, and auditory information to create more robust robot systems. His work increasingly focuses on zero-shot and few-shot learning capabilities, enabling robots to handle novel situations without extensive retraining. There's also a clear progression toward general-purpose robotics , moving away from task-specific solutions toward more flexible systems that can handle diverse real-world challenges. Dr. Pinto's scientific contributions have been recognized with several prestigious awards: Sloan Research Fellowship (2025) NSF CAREER Award (2024) RAL Early Career Award (2024) Best Student Paper Award at ICRA (2016) Outstanding Paper Award at MFM-EAI workshop at ICML (2024) Best Paper Award at NGSM workshop at ICML (2024) Best Student Paper Award at RSS (2023) As an advisor, Pinto has mentored numerous students who have gone on to impactful careers in both academia and industry. His former PhD student Denis Yarats co-founded Perplexity.AI, while Mahi Shafiullah became a postdoc at UC Berkeley and Meta AI. Many of his Masters students have pursued PhDs at top institutions like CMU, MIT, and Stanford, or joined leading robotics companies including 1X, Fauna Robotics, and NVIDIA. Pinto's lab has secured significant research funding, including the NSF CAREER award and likely other grants supporting his robotics research program. The General-purpose Robotics and AI Lab (GRAIL) that Pinto leads brings together a diverse team of researchers working on cutting-edge robotics challenges. The lab maintains strong collaborations with industry partners and other academic institutions, facilitating technology transfer and real-world impact. GRAIL's research spans multiple robotics platforms and focuses on developing algorithms that enable robots to learn from diverse experiences and generalize across environments.
Xingxing Zuo is an Assistant Professor (tenure-track) in the Robotics Department at MBZUAI. He holds a PhD from Zhejiang University (2021) and a Bachelor’s from UESTC (2016). Previously, he was a Postdoctoral Scholar at Caltech (2024–2025), a Postdoc at ETH Zurich (2019–2021), and held visiting roles at TU Munich, University of Delaware, and University of Technology Sydney. His research focuses on robotics, 3D computer vision, and embodied AI, with emphasis on robot-human collaboration, state estimation, and sensor fusion. Educations: PhD in Robotics, Zhejiang University (2021, with honors) Bachelor’s in Computer Science, University of Electronic Science and Technology of China (2016, with honors) Research Highlights: Develops novel methods for LiDAR-camera-inertial fusion, neural radiance fields, and radar-cameras systems Pioneered techniques like Flying Co-Stereo (long-range aerial mapping) and FMGS (vision-language embedded 3D splatting) Focuses on real-time SLAM, robust depth estimation, and photorealistic scene reconstruction Awards & Recognition: Best Paper Finalist at ICRA 2021 (CodeVIO) Oral Presentation at ICCV 2021 (MBA-VO) Recipient of Google Visiting Faculty Researcher (2023) Grants & Labs: Organized Thermal Infrared in Robotics workshop at ICRA 2025 Leads research on embodied AI and multi-sensor SLAM systems Develops open-source tools like LIC-Fusion and Coco-LIC frameworks
Adriana Kovashka is an Associate Professor at the University of Pittsburgh , affiliated with the School of Computing and Information and serving as Department Chair . Her academic journey began with BA degrees in Computer Science and Media Studies from Pomona College (2008) and a PhD in Computer Science from The University of Texas at Austin (2014). Joined Pitt’s faculty in January 2015 NSF CAREER awardee (2021) Google Faculty Research Award recipient Dr. Kovashka’s research spans Computer Vision , Machine Learning , and Natural Language Processing , focusing on visual rhetoric, weak multimodal supervision, and domain adaptation. She pioneered techniques for analyzing political imagery, developing robust object detection frameworks, and exploring the intersection of visual and textual persuasion through large-scale annotated datasets. Her recent work emphasizes geographic diversity in vision-language systems, audio-visual fusion for domain generalization, and shape-texture bias mitigation in CNNs. Key publications include groundbreaking studies on symbolic reasoning, multimodal dialogue systems, and ethical AI applications in education. Scientific honors include: NSF CRII Award (2016) NSF CAREER Award (2021) Pitt CRDF Award (2016, 2018) Best Paper at ECV Workshop (2021) Google Faculty Research Award (2016, 2018) Dr. Kovashka actively mentors students in multimodal learning projects and collaborates with interdisciplinary teams on NSF-funded initiatives. She co-organizes workshops like the first CVPR workshop on advertisement understanding and leads research groups exploring human-AI co-learning systems.
Katerina Fragkiadaki is the JPMorgan Chase Associate Professor of Computer Science in the Machine Learning Department at Carnegie Mellon University. She works at the intersection of Artificial Intelligence, Computer Vision, Machine Learning, Language Understanding, and Robotics. PhD from GRASP Lab, University of Pennsylvania Postdoctoral researcher at UC Berkeley (with Jitendra Malik) and Google Research Recipient of NSF CAREER, DARPA Young Investigator, Amazon, Google, Sony, UPMC, and AFOSR awards Organizer of CoRL 2023 Workshop on Generalist Robots ICLR 2024 Program Chair, multiple area chair roles Her research group focuses on developing machines that autonomously improve world models through human-environment interactions, with specific emphasis on: Representation learning and video understanding 2D/3D unified vision-language models Generative simulation and reinforcement learning Real2Sim/Sim2Real robot learning Continual learning and spatial common sense 3D scene reconstruction and dynamics Recent publications highlight advancements in: 3D mesh generation with compositional transformers Unified 2D/3D perception frameworks Physics-aware generative models Diffusion-based robotic manipulation policies Embodied agents with memory prompting Awards include: 2024: DARPA Young Investigator Award 2023: Amazon Faculty Award 2022: Sony Faculty Research Award 2021: UPMC Faculty Research Award 2020: NSF CAREER Award 2019: Google Faculty Award Key collaborations span institutions including UC Berkeley, Google Research, Stanford, MIT, and University of Tsukuba. Her work bridges theoretical innovation with practical applications in: Autonomous robot manipulation 4D world modeling Language-grounded perception Visual dynamics prediction Embodied program synthesis Physics-based simulation engines