Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Robert West is an Associate Professor at EPFL (École polytechnique fédérale de Lausanne) in the School of Computer and Communication Sciences , leading the Data Science Lab (dlab) . His research focuses on Natural Language Processing , Machine Learning , and Computational Social Science , analyzing human-generated data from the web, social media, and online platforms. Education : PhD in Computer Science (2016) - Stanford University MSc in Computer Science (2010) - McGill University BSc in Computer Science (2007) - Technische Universität München Research Interests : West develops algorithms for analyzing large-scale web data, with emphasis on multilingual NLP , social network analysis , and AI ethics . His work bridges machine learning with social science to understand digital human behavior. Scientific Awards : ICWSM’22 Adamic–Glance Distinguished Young Researcher Award Google Faculty Research Award Facebook Research Award Multiple Outstanding Paper Awards at ICWSM and WWW Advising & Grants : He advises 12 PhD students and has secured funding from the Swiss National Science Foundation , Swiss Data Science Center , and industry partners. His lab maintains collaborations with Microsoft Research and CROSS . Labs & Collaborations : West leads the Data Science Lab at EPFL, which focuses on web-scale data analysis , privacy-preserving machine learning , and AI for social good . The lab develops tools like Wikispeedia and Quotebank for public data exploration.
Prof. Konrad Schindler holds the position of Full Professor at the Department of Civil, Environmental and Geomatic Engineering at ETH Zürich. He is also the Head of the Institute of Geodesy and Photogrammetry (IGP), leading research and educational activities in geomatics and computer vision. His career spans roles as a Photogrammetric Engineer, scientific assistant, postdoc researcher, and academic faculty across institutions including Graz University of Technology, Monash University, and TU Darmstadt before joining ETH Zürich in 2010. Education: Undergraduate studies in Geodesy (1992–1995), Graz University of Technology, Austria MEng in Photogrammetry and Geoinformation (1995–1999), Vienna University of Technology, Austria PhD in Computer Science (2001–2003), Graz University of Technology, Austria Research focuses on Photogrammetry , Remote Sensing , Computer Vision , and Image Understanding with interdisciplinary applications in environmental monitoring, geospatial analysis, and disaster response. He develops computational methods for 3D reconstruction, fusion of multi-modal data, and AI-driven solutions for satellite imagery interpretation. His work bridges geomatic engineering and machine learning to address challenges in urban mapping, climate modeling, and biological systems analysis. Publications reflect expertise in geospatial AI, diffusion models, and benchmarking datasets for disaster resilience. Notable works include Marigold (image analysis adaptation) and BRIGHT (building damage assessment). His research emphasizes practicality and scalability, such as affordable depth estimation and global biomass datasets. He has received the 2013 Marr Prize Honourable Mention (IEEE) and the 2012 U.V. Helava Award (ISPRS), alongside several Best Presentation Awards. His contributions span technical leadership, editorial roles (ISPRS Journal), and service to Swiss remote sensing commissions. Advising and grants: While no specific advisee names or grant details are listed, his career trajectory includes mentoring postdocs and junior faculty. He teaches advanced courses in Photogrammetry , Image Interpretation , and Machine Vision , integrating cutting-edge AI techniques into curricula. His research group collaborates on global-scale projects like canopy height mapping and satellite-based climate variable assessments. Labs/Teams: As Institute Head, he oversees the IGP lab at ETH Zürich, with prior affiliations including the Digital Perception Lab (Monash University) and the Computer Vision Lab (ETH Zurich). His work often involves multi-institutional collaborations focused on geospatial AI and environmental science.
Shrikanth (Shri) Narayanan is a University Professor and holder of the Niki and Max Nikias Chair in Engineering at the University of Southern California (USC), serving as the inaugural Vice President for Presidential Initiatives. He leads the Signal Analysis and Interpretation Lab (SAIL) and holds joint appointments in Computer Science, Linguistics, Psychology, Neuroscience, Pediatrics, and Otolaryngology-Head and Neck Surgery. His research focuses on speech and audio processing, behavioral signal processing, and real-time MRI of speech production, with applications in healthcare, education, and technology. Education: B.E. in Electrical Engineering from College of Engineering, Guindy (Chennai, India, 1988); M.S., Engineer, and Ph.D. in Electrical Engineering from UCLA (1990, 1992, 1995). Research interests span computational linguistics, machine learning, and multimodal human behavior analysis. He pioneered technologies for speech biomarkers in mental health, real-time MRI of speech production, and wearable sensor systems for longitudinal health studies. His work in speech emotion recognition, forensic interviews, and clinical applications has been recognized through over 40 awards, including the IEEE Flanagan Award and ISCA Medal. He has published extensively in journals like Proceedings of the IEEE , Journal of the Acoustical Society of America , and PLOS One . Key Grants: NSF CAREER, Okawa Research, IBM Faculty, Google/Amazon awards. Labs/Teams: Signal Analysis & Interpretation Lab (SAIL), USC Information Sciences Institute (ISI), Google Visiting Faculty Researcher.
Dr. Diyi Yang is an Assistant Professor in the Computer Science Department at Stanford University. She leads the Social and Language Technologies (SALT) Lab, affiliated with the Stanford NLP Group, Stanford HCI Group, Stanford AI Lab (SAIL), and Stanford Human-Centered Artificial Intelligence (HAI). Her research focuses on socially aware natural language processing, large language models (LLMs), and human-AI interaction, aiming to improve human-human and human-computer communication through socially grounded AI systems. Education: Ph.D. in Language Technologies Institute, Carnegie Mellon University (2013–2019) B.S. in ACM Honored Class, Shanghai Jiao Tong University (2009–2013) Research Interests: Dr. Yang’s work bridges computational social science and NLP. She explores how AI can understand social contexts in language use and develop systems that respect cultural norms, ethical standards, and human values. Her lab’s projects include AI companions for skill training (e.g., Rehearsal Dialects), norm-aware LLMs (NormBank), and frameworks for human-AI collaboration (Co-Gym). Recent efforts address bias in AI, societal impacts of LLMs, and ethical evaluation of human-AI systems. Awards & Honors: 2024: Sloan Research Fellowship, ONR Young Investigator Award 2023: Adamic-Glance Young Distinguished Award (ICWSM), Kavli Fellow (NAS) 2022: NSF CAREER Award, Microsoft Research Faculty Fellow 2020: IEEE AI’s 10 to Watch Advising & Grants: Dr. Yang advises over 10 PhD students and postdocs, co-leading projects on LLM evaluation (SWE-bench/SWE-smith), human-AI ethics, and culturally aware NLP. Her research is supported by NSF, Amazon, DARPA, Google, and Stanford’s HAI initiative. Labs & Teams: The SALT Lab collaborates across disciplines, with projects spanning computer science, linguistics, and social sciences. Current initiatives include developing AI tools for mental health support (AI Partner & Mentor) and auditing societal impacts of LLMs.
Laurel MacKenzie is an Associate Professor in the Department of Linguistics at New York University (NYU), affiliated with the Faculty of Arts and Science. She specializes in variationist sociolinguistics, dialectology, and language change, with a focus on English and French varieties. Her work integrates quantitative analysis of speech data to explore intra-speaker variation and language evolution. She co-directs the NYU Sociolinguistics Lab and leads the NSF-funded NYC Individual Differences Corpus project, alongside the Our Dialects initiative, an online atlas of British English dialects. Education: PhD in Linguistics, University of Pennsylvania (2012) BA in Linguistics and French, University of California, Berkeley (2006) Research Interests: Morphological and syntactic variation Regional dialects of English and French Linguistic pedagogy and public engagement Recent Projects: Recent work includes publications on participle leveling in English, sociolinguistic replication studies, and grammatical variation analysis. She has also collaborated on dialect mapping tools and consulted for media projects on language change and accents. Awards: No awards explicitly listed in provided texts. Labs/Teams: NYU Sociolinguistics Lab (Co-Director) Our Dialects Project (Academic Lead)
Fanny Yang is an Assistant Professor in the Computer Science Department at ETH Zurich . She previously held postdoctoral positions at Stanford University and a Junior Fellowship at the Institute for Theoretical Studies at ETH Zurich, advised by Nicolai Meinshausen . Her PhD was completed at the EECS Department of UC Berkeley , supervised by Martin Wainwright . Research Interests : Theoretical foundations of machine learning and statistics , particularly focusing on overparameterized models and high-dimensional data . Developing trustworthy ML models with emphasis on distributional robustness , domain generalization , and interpretability . Applications in medical diagnostics and treatment effect analysis , aiming to address reliability issues in real-world domains. Recent Work Trends : 2025 Articles : Theoretical analysis of semi-supervised multi-objective learning , test-time scaling with verifiers, and foundation models for efficient randomized experiments . 2024 Articles : Studies on robust mixture learning , privacy-preserving data synthesis via optimal transport , confounding quantification in causal inference , and semi-private learning frameworks. 2023 Articles : Investigations into active vs. passive learning in high dimensions, inductive bias in noisy interpolation , and semi-supervised novelty detection using model ensembles .
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Yuki M. Asano is a full Professor at the University of Technology Nuremberg , leading the Fundamental AI (FunAI) Lab . Previously, he led the QUVA Lab at the University of Amsterdam and earned his PhD at the Visual Geometry Group (VGG) of the University of Oxford under Andrea Vedaldi and Christian Rupprecht. University of Technology Nuremberg (2024–present) University of Amsterdam (prior to 2024) University of Oxford (PhD, 2020) His research spans Artificial Intelligence , Machine Learning , and Computer Vision , with a focus on Causal Representation Learning , Self-Supervised Learning , and Efficient Model Adaptation . He pioneered techniques like BISCUIT (causal variable identification) and VeRA (parameter-efficient fine-tuning). His work extends to Medical Imaging and Environmental Monitoring through applications in fetal ultrasound analysis and marine debris detection. Recent publications (2023–2025) highlight advancements in Self-Supervised Learning , Vision-Language Models , and 3D Understanding . Notable papers include TWIST & SCOUT (multimodal LLM grounding), SIGMA (masked video modeling), and GeneralAD (anomaly detection). His ICCV 2023 work on Self-Ordering Point Clouds and MoSiC (optimal-transport motion trajectories) underscores his interdisciplinary approach. He received the JUPITER compute grant (2025) and an Outstanding Paper Award at ICLR 2024 . His collaborations span institutions like MIT-IBM Watson AI Lab, Qualcomm AI Research, and University of Amsterdam.
Kristen Grauman is a Full Professor in the Department of Computer Science at the University of Texas at Austin, where she leads the UT Computer Vision Group. Her research focuses on computer vision and machine learning, with applications in visual recognition, video analysis, and multi-modal perception. She received her B.A. from Boston College and her Ph.D. from MIT. Her research interests span visual recognition, image and video search, video analysis, first-person vision, embodied and multi-modal perception, and interactive machine learning. She has made significant contributions to the field, particularly in developing algorithms for understanding visual content and human activities from video, including foundational work on the Pyramid Match Kernel and relative attributes. Her recent publications reveal a strong emphasis on egocentric (first-person) vision, audio-visual learning, and view-invariant representations. There is a clear trajectory toward multi-modal integration (vision, audio, language) and real-world applications in instructional videos, human activity understanding, and embodied AI systems. She has received numerous awards including: AAAI Fellow (2019) J. K. Aggarwal Prize, International Association for Pattern Recognition (2018) Helmholtz Prize (2017) UT Austin Academy of Distinguished Teachers (2017) Best Paper Award, Asian Conference on Computer Vision (2016) Presidential Early Career Award for Scientists and Engineers (2014) Computers and Thought Award, International Joint Conferences on Artificial Intelligence (2013) Pattern Analysis and Machine Intelligence Young Researcher Award (2013) Alfred P. Sloan Research Fellow (2012) Marr Prize (2011) Prof. Grauman serves as Associate Editor-in-Chief for the IEEE Transactions on Pattern Analysis and Machine Intelligence. She has secured substantial research funding including the Presidential Early Career Award, NSF grants, and industry partnerships. Her advising has produced numerous influential publications and students who are now leaders in computer vision. She leads the UT Computer Vision Group, which collaborates closely with the Electrical and Computer Engineering Department. The group is pioneering large-scale egocentric video research through projects like Ego4D and Ego-Exo4D, focusing on real-world applications in human activity understanding, audio-visual perception, and interactive systems.
Jennifer Olsen, PhD, is an Assistant Professor of Computer Science at the University of San Diego since 2020. She holds a PhD, MS, and BS in Human-Computer Interaction and Cognitive Science from Carnegie Mellon University, followed by postdoctoral research at the Swiss Federal Institute of Technology (EPFL), Lausanne, Switzerland. Her research focuses on the intersection of human-computer interaction, cognition, and education, emphasizing collaborative learning and educational technology design from both learner and instructor perspectives. Education: PhD in Human-Computer Interaction, Carnegie Mellon University MS in Human-Computer Interaction, Carnegie Mellon University BS in Cognitive Science, Carnegie Mellon University Research Interests: Dr. Olsen explores how collaboration supports learning, designs technologies to enhance educational practices, and investigates gaze-based metrics for understanding collaborative problem-solving. Her work spans gamified robotics, AI-driven orchestration systems, and virtual reality applications in vocational training. She emphasizes learner-centered design and the integration of social robots and virtual agents in pedagogical settings. Grants/Advising: While no specific grants or advisees are listed, her prolific publication record indicates active involvement in educational technology research and development. Her work addresses challenges in classroom orchestration, multimodal data analysis, and accessibility in educational robotics. Labs/Teams: Collaborates with interdisciplinary teams focused on educational technology, human-robot interaction, and adaptive learning systems. Her research leverages tools like FROG orchestration graphs and eye-tracking technologies to develop practical classroom solutions.
Sara Ng is a Visiting Assistant Professor in the Department of Linguistics at Western Washington University. She holds a PhD in Linguistics from the University of Washington (2024) and an MS in Computational Linguistics (2023), alongside a B.A. in Linguistics and B.S. in Applied Mathematics from the University of Utah (2017). Her research focuses on computational models of prosody, speech perception, and their integration with speech technologies like automatic speech recognition (ASR). She investigates how prosody conveys pragmatic meaning, influences conversational dynamics, and impacts listeners with hearing impairments. Her work bridges computational methods and linguistic theory, addressing challenges in clinical and technological applications. Education PhD in Linguistics, University of Washington, 2024 MS in Computational Linguistics, University of Washington, 2023 B.A. in Linguistics (Honors), B.S. in Applied Mathematics, University of Utah, 2017 Research Interests Ng’s research explores computational linguistics, prosody modeling, speech technology, and hearing impairment studies. She develops methods to leverage prosodic cues for tasks like punctuation prediction and investigates the impact of hearing loss on speech perception. Grants & Awards 2024: Nominated for UW Excellence in Teaching Award 2023: Excellence in Linguistic Research Graduate Fellowship 2017: University of Utah Top Scholar Award Teaching Ng teaches courses in phonetics, computational linguistics, and linguistics for honors students. She emphasizes pedagogical innovation and inclusivity, with experience as an instructor of record and teaching assistant at both Western Washington University and the University of Washington. Labs & Collaborations She collaborates with the TIAL Lab (University of Washington), Phonetics Lab, Hearing Aid Laboratory (Northwestern University), and CLILLAC-ARP (Université Paris Cité). Her work intersects with clinical and engineering domains, addressing real-world applications of speech technology.
Nadia Figueroa is the Shalini and Rajeev Misra Presidential Assistant Professor in the Mechanical Engineering and Applied Mechanics (MEAM) Department at the University of Pennsylvania. She holds secondary appointments in Computer and Information Science (CIS) and Electrical and Systems Engineering (ESE), and is a core faculty member at the GRASP Lab. Prior to Penn, she was a Postdoctoral Associate at MIT’s CSAIL and earned her PhD in Robotics from EPFL under Prof. Aude Billard. Her research focuses on physical and perceptual adaptive intelligence for robots, enabling fluid collaboration with humans in dynamic environments. Key applications include robot learning from demonstration , human-robot co-manipulation , safe navigation in human-centric spaces , and rehabilitation robotics . Her work integrates machine learning control theory artificial intelligence biomechanics psychology with guarantees of stability, safety, and robustness . Recent publications highlight advancements in reactive collision avoidance dynamical system learning intent estimation EEG-driven assistive control origami-based reconfigurable robots across platforms like autonomous vehicles and humanoid robots. She has authored a 2022 textbook on dynamical systems for robot control and received the Presidential Assistant Professorship at Penn.
Adrian Weller is a prominent researcher and academic at the University of Cambridge, serving as a Director of Research in Machine Learning within the Department of Engineering. He holds multiple significant leadership roles including Programme Director for Trust and Society at the Leverhulme Centre for the Future of Intelligence (CFI), and previously served as Programme Director for AI at The Alan Turing Institute, the UK national institute for data science and AI. His work bridges theoretical machine learning research with practical applications and societal implications of artificial intelligence. Weller's research interests span a broad spectrum of AI and machine learning topics with a particular focus on ensuring beneficial societal outcomes. His work encompasses explainability, fairness, robustness, scalability, privacy, safety, and ethics in AI systems. He has made significant contributions to trustworthy machine learning, including developing frameworks for AI governance, certification, and human-AI collaboration. His research group actively investigates neuro-symbolic approaches, privacy-preserving techniques, and methods for improving the reliability and interpretability of AI systems. His recent publications demonstrate a strong trend toward addressing the practical challenges of deploying AI systems in real-world contexts, particularly focusing on certification frameworks, governance mechanisms, and human-centered approaches. His work spans theoretical advances in machine learning architectures while maintaining a strong connection to societal impact, with publications appearing in top venues across AI, machine learning, and interdisciplinary applications. Scientific Awards: MBE for services to digital innovation (2022 Queen's Birthday Honours) Turing AI Fellowship for Trustworthy Machine Learning Weller actively supervises a large group of PhD students and postdocs, with current students including Juyeon Heo, Yanzhi Chen, Katie Collins, Isaac Reid, Yichao Liang, Herbie Bradley, and Shoaib Siddiqui. His former students have gone on to positions at leading institutions including Google DeepMind, ETH Zurich, NYU, and MPI-IS Tübingen. He has served on numerous advisory boards including the Centre for Data Ethics and Innovation, UNESCO's expert group on AI ethics, and the World Economic Forum's Global Future Council on AI. His research has been supported through his Turing AI Fellowship and various collaborative projects focused on safe and ethical AI development. Weller leads a vibrant research group focused on trustworthy machine learning, which actively organizes workshops and conferences including ICML 2024 (where he served as Program Chair), multiple workshops on responsible AI, and events through the ELLIS network. His group collaborates extensively across disciplines, working with researchers in computer science, social sciences, law, and policy to address the multifaceted challenges of developing beneficial AI systems.
Manik Varma is a Distinguished Scientist and Vice President at Microsoft Research India, and an Adjunct Professor at the Indian Institute of Technology Delhi. He is a Fellow of the Indian Academies of Science (IASc, INSA, NASI), the Indian National Academy of Engineering (INAE), and the Association for Computing Machinery (ACM). He has received prestigious awards such as the Shanti Swarup Bhatnagar Prize and Microsoft Gold Star Award. Education : BSc in Physics from St. Stephen's College (David Raja Ram Prize) BA in Theoretical Physics from the University of Oxford (Rhodes Scholar) DPhil in Computer Vision and Machine Learning from the University of Oxford (University Scholar) Post-doctoral Fellow at the Mathematical Sciences Research Institute (MSRI), Berkeley Visiting Miller Professor at UC Berkeley His research focuses on Machine Learning (Extreme Classification, Resource-efficient ML, Supervised Learning), Information Retrieval (Computational Advertising, Dense Retrieval, Recommender Systems), and Computer Vision (Image Search, Object Recognition). Recent work includes graph-regularized encoders, label variance reduction, and multimodal classification frameworks. His publications span extreme classification algorithms like NGAME , SiameseXML , and DECAF , with applications in IoT, web search, and recommendation systems. He leads a research group at Microsoft Research India and advises PhD students at IIT Delhi. Scientific Awards : Shanti Swarup Bhatnagar Prize (Government of India) Microsoft Gold Star and Achievement Awards WSDM 2019 Best Paper Prize BuildSys 2019 Best Paper Runner-up Fellow of ACM, IASc, INSA, NASI, INAE He has supervised numerous PhD students, including Sonu Mehta and Suchith Prabhu, and collaborates with institutions like Microsoft Research India, IIT Delhi, and UC Berkeley. His research has led to scalable solutions for billion-label classification and resource-constrained IoT applications.