Winnie Wong is a Professor of Rhetoric at the University of California, Berkeley. Her research focuses on art history, particularly exploring themes of fakes, forgeries, and authorship through interdisciplinary approaches. She examines the global artistic networks of Hong Kong, Guangzhou, and Shenzhen, challenging Western narratives of Chinese artistic practices. Wong holds a PhD from MIT's History, Theory + Criticism program and was a Harvard Society of Fellows Junior Fellow. Her major works include Van Gogh on Demand: China and the Readymade (2014, Joseph Levenson Prize winner), Learning from Shenzhen (2017), and a forthcoming book on Canton trade portraitists. She teaches courses like Theory of the Copy and Rhetoric of the Image , mentoring undergraduates through the Undergraduate Research Apprentices program and advising PhD students in art history and visual studies. Wong's research has been supported by grants from Mellon Foundation, ACLS, SSRC, and others. Her work has been translated into multiple languages and appears in journals like Current Anthropology and Law & Literature . She frequently discusses her work on podcasts and collaborates with museums including M+ and the Asian Art Museum.
Kaiming He is an Associate Professor with tenure in the Department of Electrical Engineering and Computer Science (EECS) at the Massachusetts Institute of Technology (MIT), holding the Douglas Ross (1954) Career Development Professor of Software Technology chair. He also works part-time as a Distinguished Scientist at Google DeepMind. Prior to joining MIT in 2024, he was a research scientist at Facebook AI Research (FAIR) from 2016 to 2024, and a researcher at Microsoft Research Asia (MSRA) from 2011 to 2016. Dr. He received his PhD from the Chinese University of Hong Kong in 2011 and his Bachelor of Science from Tsinghua University in 2007. His academic journey reflects a strong foundation in computer science and engineering that has led to transformative contributions in artificial intelligence. His research primarily focuses on computer vision and deep learning, with pioneering work on deep residual networks (ResNets), visual object detection and segmentation, and self-supervised learning. He is best known for his work on Deep Residual Networks (ResNets), recognized as the most-cited paper of the twenty-first century. The residual connections he pioneered are now fundamental components in modern deep learning architectures including Transformers, AlphaGo Zero, AlphaFold, and various generative AI models. His recent publications demonstrate continued innovation across generative models, transformer architectures, and cross-disciplinary AI applications. His work bridges theoretical advances in neural network design with practical implementations that address real-world challenges in physics, biology, and other scientific domains. PAMI Young Researcher Award (2018) Best Paper Award, CVPR (2009, 2016) Best Paper Award, ICCV (2017) Best Student Paper Award, ICCV (2017) Everingham Prize, ICCV (2021) Most-cited paper of the twenty-first century Dr. He advises graduate students including Jake Austin, Xingjian Bai, and Mingyang Deng, and teaches advanced courses such as "6.S978: Deep Generative Models" (Fall 2024) and "6.8300/6.8301: Advances in Computer Vision" (Spring 2024). His research group actively explores how AI can serve as a unifying framework across scientific disciplines, breaking down traditional barriers between fields through shared methodologies and tools.
Wei Xu is an Associate Professor at Georgia Institute of Technology's College of Computing and School of Interactive Computing, with affiliations to the Machine Learning Center. Their research bridges machine learning, natural language processing, and social media with focus areas in large language models, cultural bias mitigation, multilingual capabilities, and human-AI collaboration in text evaluation. NSF CAREER and Google Academic Research Award recipient Director of NLP X Lab PhD from New York University, BSMS from Tsinghua University Research interests span: Multilingual Multicultural LLMs addressing representational gaps and cultural adaptation in language models (NAACL 2025, ACL 2024); Robustness and Reasoning through dynamic AGI evaluations (ACL 2024, EMNLP 2024); Interdisciplinary NLP applications in security, healthcare, and law (EMNLP 2024, ACL 2024). Recent publications focus on multilingual alignment (NAACL 2025), privacy risk estimation (arXiv 2025), cultural bias analysis (ACL 2024), and medical text simplification (EMNLP 2024). Key themes include bias mitigation, multimodal processing, and practical LLM evaluation. Scientific Awards : NSF CAREER, Google/Sony/Criteo research awards, ACL'24 Best Social Impact Award, COLING'18 Best Paper Advising 15 PhD/MS/BSMS students including Yao Dou (human-centered LLM evaluation), Tarek Naous (multilingual LLMs), and alumni like Chao Jiang (Apple AI/ML) and Yang Chen (NVIDIA research scientist). Teaches graduate courses on NLP and LLMs.
Andrea Vedaldi is a Professor of Computer Vision and Machine Learning at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). He specializes in unsupervised methods for understanding images and videos, focusing on 3D geometry and semantics. His research bridges foundational AI and practical applications, with contributions to generative models, neural fields, and self-supervised learning. Education: PhD in Computer Science (2008), University of California, Los Angeles MSc in Computer Science (2005), UCLA BSc in Information Engineering (2003), University of Padua Research Interests: Unsupervised learning, 3D perception, generative AI, neural rendering, and scalable vision systems. His work emphasizes ethical, responsible AI aligned with ERC-funded projects like UNION (ERC Consolidator Grant). Key Contributions: Co-developer of VLFeat and MatConvNet libraries Leader in 3D reconstruction and diffusion models (e.g., CatFree3D) Recipient of the PAMI Thomas S. Huang Prize and multiple best paper awards Grants & Service: Principal Investigator on £2.3M ERC Consolidator Grant (UNION) Co-organizer of major conferences (ECCV 2020 Program Chair, CVPR 2023 Area Chair) Reviewer for top journals/conferences (PAMI, CVPR, NeurIPS) Labs & Teams: VGG Group at Oxford, collaborating on projects like Meta 3D Gen and Common Objects in 3D (CO3D).
Antoine Bosselut is a Tenure Track Assistant Professor at École Polytechnique Fédérale de Lausanne (EPFL) in the School of Computer and Communication Sciences, where he leads the EPFL NLP group. His research focuses on developing AI reasoning agents that can model, represent, and reason about human and world knowledge, with applications in health, education, and global fairness. His research interests span multiple critical areas in modern AI: LLM Representations of Knowledge: Understanding what language models know and how they represent that knowledge internally Reasoning Algorithms: Developing methods to improve LLMs' reasoning capabilities through symbolic systems, neuroscience, and cognitive science Large-scale AI Development: Creating open-source foundation models with multilingual capabilities AI Democratization: Ensuring equitable access to AI technologies across different cultural and regulatory contexts His recent publications reveal a strong focus on the intersection of language models and cognitive science, with multiple studies examining how LLMs align with human cognition. He's also deeply engaged in practical applications of NLP technology, particularly in education and multilingual settings, as evidenced by his work on evaluating AI's impact on higher education and developing culturally-aware language models. His scientific achievements have been recognized with several prestigious awards: Outstanding Paper Award at NAACL 2025 ELLIS Scholar designation in 2024 Outstanding Paper Award at ACL 2023 Inclusion in Forbes 30 Under 30 list for Science & Healthcare in 2021 Bosselut actively mentors numerous PhD students across diverse NLP research areas and has secured significant recognition for his work in both academic and media circles. His lab has been featured in major publications including Communications of the ACM, The Atlantic, and Quanta Magazine, highlighting the societal impact of his research on commonsense reasoning and AI capabilities.
Justin Johnson is an Assistant Professor at the University of Michigan's College of Engineering, Department of Electrical Engineering and Computer Science, and a Research Scientist at Facebook AI Research (FAIR). His work bridges computer vision, machine learning, and deep learning, focusing on visual reasoning, vision-language tasks, image generation, and 3D reasoning using neural networks. PhD, Stanford University (advised by Fei-Fei Li) His research interests span visual reasoning , vision and language , 3D vision , and image generation , with a focus on innovative applications of deep neural networks. Recent publications highlight work on 3D consistency, self-supervised learning, and multimodal integration of vision and text. Notable contributions include PyTorch3D for 3D data processing, and foundational work in visual question answering , neural style transfer , and scene graph-based image generation . Publications span top conferences like ICCV, CVPR, and NeurIPS. He teaches courses including EECS 498/598: Deep Learning for Computer Vision and EECS 442: Computer Vision at University of Michigan, with prior involvement in Stanford's CS 231N in co-teaching roles. Software projects include open-source frameworks like fast-neural-style for real-time artistic style transfer, and PyTorch3D for efficient 3D deep learning. These tools demonstrate his commitment to practical implementations and community-driven research.
Kai-Wei Chang is an Associate Professor at the University of California, Los Angeles (UCLA) in the Department of Computer Science, part of the Henry Samueli School of Engineering. He is also an Amazon Scholar at Alexa AI, focusing on advancing trustworthy AI and multimodal foundation models. His research bridges NLP, machine learning, and ethical AI, with a focus on fairness, robustness, and bias mitigation in language and vision-language systems. Education: Ph.D. in Computer Science (UIUC, 2015), M.S. and B.S. in Computer Science and Electrical Engineering from National Taiwan University. Research Interests: Trustworthy NLP (fairness, robustness), Multimodal Foundation Models (e.g., VisualBERT, GLIP), Reasoning in LLMs, and mitigating societal biases in AI systems. Notable contributions include pioneering work on aligning NLP models with human values and developing SOTA multimodal models like DesCo and GLIP. Awards: Sloan Research Fellowship (2021), Okawa Grant (2018), EMNLP Best Paper (2017), KDD Best Paper (2010). His work is funded by NSF, IARPA, ONR, and industry partners like Amazon, Google, and Facebook. Service: VP-Elect of SIGDAT, Ethics Committee Chair (NAACL 2022), Organizer of Trustworthy NLP Workshops, and Senior Area Chair for top conferences (ACL, NeurIPS, AAAI). Labs/Teams: Leads the UCLA Natural Language Processing Group, fostering interdisciplinary research on ethical AI and multimodal systems.
Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Christian Theobalt is a Professor of Computer Science at Saarland University and Scientific Director of the Visual Computing and Artificial Intelligence Department at the Max Planck Institute for Informatics . He leads the Saarbruecken Center for Visual Computing as a strategic partnership between Google and MPI. PhD in Computer Science (2005) from MPI-INF/Saarland University Postdoctoral Researcher at MPI (2005-2007) Visiting Assistant Professor at Stanford (2007-2009) His research focuses on the intersection of Computer Graphics, Computer Vision, and Artificial Intelligence , with specializations in: 3D/4D Human Reconstruction Neural Rendering Performance Capture Geometric Deep Learning Volumetric Video Quantum Visual Computing Recent article trends show emphasis on: Quantum computing applications in visual reconstruction Neural rendering with radiance fields Human motion capture from egocentric views Gesture-language interaction modeling Multi-modal scene understanding Real-time free-viewpoint rendering Scientific Awards : Fellow of EUROGRAPHICS (2022) CVPR Best Student Paper Honorable Mention (2020) ERC Consolidator Grant (2017) Karl Heinz Beckurts Award (2017) Busy Beaver Teaching Award (2016) ERC Starting Grant (2013) Advising over 50 students and researchers including: Current researchers: Viktor Rudnev, Linjie Lyu, Mohit Mendiratta Postdocs: Kwang In Kim, Kiran Varanasi Alumni: Franziska Mueller (Google), Dushyant Mehta (Qualcomm), Ayush Tewari (MIT) Grants include ERC grants, Google Glass Research Award, and multiple industry partnerships. His lab maintains cutting-edge facilities with: Multi-camera capture systems Quantum annealing infrastructure HDR radiance field technology GPU clusters for AI research Time-of-flight imaging systems Event camera arrays
Jerrica Ty Rowlett serves as an Assistant Professor of Communication and Language Studies within the College of Arts and Sciences at Bryant University, where she examines digital communication's societal implications through research and teaching in digital culture and social media dynamics. Education Ph.D., Florida State University M.A., Clemson University B.A., Georgetown College Her research program investigates identity formation, civic engagement, and digital inequity through critical analysis of social media platforms. She explores how digital environments shape interpersonal relationships, political discourse, and community building while addressing power dynamics in online spaces, with particular attention to abusive contexts, gender representation, and emotional experiences. Analysis of her publication record (2015-2024) reveals a consistent trajectory examining social media's evolving societal impact. Early work focused on educational applications and youth culture, progressing to sophisticated analyses of political meme culture, relationship dissolution in abusive contexts, and emotional contagion online. Her scholarship demonstrates increasing attention to social justice dimensions within digital communication. Scientific Awards No awards documented in current sources Advising and Grants Student advising details and research grant information are not specified in available materials, though her extensive publication record indicates active scholarly mentorship and research productivity. Labs and Teams No dedicated research laboratories or formal collaborative teams are mentioned in the provided information.
Russell Epstein is a Professor and Director of Graduate Studies in the Department of Psychology at the University of Pennsylvania. He is affiliated with the Center for Cognitive Neuroscience and Goddard Labs. His research focuses on neural mechanisms underlying visual scene perception, spatial navigation, and memory. Epstein holds a BA in Physics from the University of Chicago and a PhD in Applied Mathematics from Harvard University. Epstein’s research interests include high-level vision, spatial cognition, and the neural basis of environmental representations. His lab uses functional MRI and cognitive neuroscience techniques to study how scenes, objects, landmarks, and spaces are encoded in brain systems such as the parahippocampal place area and retrosplenial cortex. Recent work explores cognitive maps, grid-like neural representations, and the role of multisensory cues in navigation. His articles emphasize spatial navigation strategies, hierarchical cognitive maps, and the interplay between perception and memory. Notable contributions include investigations into hippocampal spatial metrics, olfactory navigation, and the neural underpinnings of environmental learning. Epstein teaches courses on cognitive neuroscience, including PSYC 149 and PSYC 600. He advises two graduate students in Psychology and has no listed scientific awards. His work is supported by grants (unspecified) and conducted within collaborative teams at the Center for Cognitive Neuroscience. Epstein’s research extends to labs focused on spatial cognition and neuroimaging, advancing understanding of how humans mentally map environments through visual and sensory integration.
Chua Tat Seng is a Professor at the School of Computing, National University of Singapore (NUS), holding the KITHCT Chair Professorship since 2009. He serves as co-Director of the NExT++ Center, a joint research center between NUS and Tsinghua University focused on Extreme Search. His academic career spans over three decades at NUS, where he has held various leadership positions including Acting Dean of the School of Computing (1998-2000) and Acting Head of the Department of Information Systems & Computer Science (1996-1998). Professor Chua's research spans unstructured data analytics , multimedia information retrieval , recommendation and conversation systems , and emerging applications in e-commerce and fintech . He established the Lab for Media Search (LMS) at the School of Computing and has been instrumental in advancing multimodal learning and search technologies. His work bridges theoretical foundations with practical applications, particularly in developing trustable AI systems for real-world deployment. His recent publications demonstrate a strong focus on large language models for recommendation systems , multimodal learning , and generative AI applications . The research trends show increasing emphasis on LLM-based recommendation, multimodal understanding, and addressing fundamental challenges in AI reliability, fairness, and efficiency. His work spans theoretical advancements in representation learning to practical applications in e-commerce, finance, and healthcare domains. ACM SIGMM Technical Achievement Award 2015 Multiple Best Paper Awards across ACM Multimedia, IEEE Transactions, and MMM conferences (2007-2020) Professor Chua has supervised 37 PhD students since 2004, establishing himself as a dedicated mentor in the academic community. His research has been supported by substantial grants including NExT++ ($12 million), Base Metals Price Forecasting ($200,000), and Multilingual Multimodal Knowledge Graph ($500,000). He maintains active collaborations with Tsinghua University, University of Southampton, and industry partners like Four Elements Capital and Singapore Press Holdings. As co-Director of the NExT++ Center, he leads a major research initiative focused on Web Intelligence and User Empowerment. His visiting professorships at Tsinghua University (2017-present) and Zhejiang University (2021-present) reflect his international impact in the field of multimedia and AI research.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Yuki M. Asano is a full Professor at the University of Technology Nuremberg , leading the Fundamental AI (FunAI) Lab . Previously, he led the QUVA Lab at the University of Amsterdam and earned his PhD at the Visual Geometry Group (VGG) of the University of Oxford under Andrea Vedaldi and Christian Rupprecht. University of Technology Nuremberg (2024–present) University of Amsterdam (prior to 2024) University of Oxford (PhD, 2020) His research spans Artificial Intelligence , Machine Learning , and Computer Vision , with a focus on Causal Representation Learning , Self-Supervised Learning , and Efficient Model Adaptation . He pioneered techniques like BISCUIT (causal variable identification) and VeRA (parameter-efficient fine-tuning). His work extends to Medical Imaging and Environmental Monitoring through applications in fetal ultrasound analysis and marine debris detection. Recent publications (2023–2025) highlight advancements in Self-Supervised Learning , Vision-Language Models , and 3D Understanding . Notable papers include TWIST & SCOUT (multimodal LLM grounding), SIGMA (masked video modeling), and GeneralAD (anomaly detection). His ICCV 2023 work on Self-Ordering Point Clouds and MoSiC (optimal-transport motion trajectories) underscores his interdisciplinary approach. He received the JUPITER compute grant (2025) and an Outstanding Paper Award at ICLR 2024 . His collaborations span institutions like MIT-IBM Watson AI Lab, Qualcomm AI Research, and University of Amsterdam.
Kristen Grauman is a Full Professor in the Department of Computer Science at the University of Texas at Austin, where she leads the UT Computer Vision Group. Her research focuses on computer vision and machine learning, with applications in visual recognition, video analysis, and multi-modal perception. She received her B.A. from Boston College and her Ph.D. from MIT. Her research interests span visual recognition, image and video search, video analysis, first-person vision, embodied and multi-modal perception, and interactive machine learning. She has made significant contributions to the field, particularly in developing algorithms for understanding visual content and human activities from video, including foundational work on the Pyramid Match Kernel and relative attributes. Her recent publications reveal a strong emphasis on egocentric (first-person) vision, audio-visual learning, and view-invariant representations. There is a clear trajectory toward multi-modal integration (vision, audio, language) and real-world applications in instructional videos, human activity understanding, and embodied AI systems. She has received numerous awards including: AAAI Fellow (2019) J. K. Aggarwal Prize, International Association for Pattern Recognition (2018) Helmholtz Prize (2017) UT Austin Academy of Distinguished Teachers (2017) Best Paper Award, Asian Conference on Computer Vision (2016) Presidential Early Career Award for Scientists and Engineers (2014) Computers and Thought Award, International Joint Conferences on Artificial Intelligence (2013) Pattern Analysis and Machine Intelligence Young Researcher Award (2013) Alfred P. Sloan Research Fellow (2012) Marr Prize (2011) Prof. Grauman serves as Associate Editor-in-Chief for the IEEE Transactions on Pattern Analysis and Machine Intelligence. She has secured substantial research funding including the Presidential Early Career Award, NSF grants, and industry partnerships. Her advising has produced numerous influential publications and students who are now leaders in computer vision. She leads the UT Computer Vision Group, which collaborates closely with the Electrical and Computer Engineering Department. The group is pioneering large-scale egocentric video research through projects like Ego4D and Ego-Exo4D, focusing on real-world applications in human activity understanding, audio-visual perception, and interactive systems.