Wei Xu is an Associate Professor at Georgia Institute of Technology's College of Computing and School of Interactive Computing, with affiliations to the Machine Learning Center. Their research bridges machine learning, natural language processing, and social media with focus areas in large language models, cultural bias mitigation, multilingual capabilities, and human-AI collaboration in text evaluation. NSF CAREER and Google Academic Research Award recipient Director of NLP X Lab PhD from New York University, BSMS from Tsinghua University Research interests span: Multilingual Multicultural LLMs addressing representational gaps and cultural adaptation in language models (NAACL 2025, ACL 2024); Robustness and Reasoning through dynamic AGI evaluations (ACL 2024, EMNLP 2024); Interdisciplinary NLP applications in security, healthcare, and law (EMNLP 2024, ACL 2024). Recent publications focus on multilingual alignment (NAACL 2025), privacy risk estimation (arXiv 2025), cultural bias analysis (ACL 2024), and medical text simplification (EMNLP 2024). Key themes include bias mitigation, multimodal processing, and practical LLM evaluation. Scientific Awards : NSF CAREER, Google/Sony/Criteo research awards, ACL'24 Best Social Impact Award, COLING'18 Best Paper Advising 15 PhD/MS/BSMS students including Yao Dou (human-centered LLM evaluation), Tarek Naous (multilingual LLMs), and alumni like Chao Jiang (Apple AI/ML) and Yang Chen (NVIDIA research scientist). Teaches graduate courses on NLP and LLMs.
Stefano Ermon is an Associate Professor in the Department of Computer Science at Stanford University, affiliated with the Artificial Intelligence Laboratory and a Senior Fellow at the Woods Institute for the Environment. His research focuses on advancing machine learning and generative AI techniques to address societal and environmental challenges, including computational sustainability, geospatial analysis, and climate science. He holds a Ph.D. from Cornell University (2015). Education: Ph.D. in Computer Science, Cornell University (2015). Research Interests: Ermon’s work bridges foundational machine learning (e.g., diffusion models, generative AI, and optimization) with applications in sustainability, geospatial analysis (via satellite imagery), and earth observation systems. Notable contributions include predicting poverty using satellite data and developing scalable methods for molecule generation. Articles Trends: His recent work emphasizes diffusion models for generative tasks (e.g., text-to-image, molecule design), geospatial AI (e.g., environmental monitoring), and ethical AI (e.g., bias mitigation in LLMs). He also explores applications in robotics and scientific computing. Awards: He has received prestigious awards, including the ICML 2024 Best Paper Award, Sloan Research Fellowship, Microsoft Research Faculty Fellowship, and the IJCAI Computers and Thought Award. Advising and Grants: Ermon teaches courses like Probabilistic Graphical Models (CS228) and has secured grants from NSF, ONR, AFOSR, and private foundations. His lab develops tools for climate science and sustainable development. Labs/Teams: Leads the Stanford AI Lab group focused on computational sustainability and generative AI, collaborating with institutions like the Woods Institute for environmental applications.
Justin Johnson is an Assistant Professor at the University of Michigan's College of Engineering, Department of Electrical Engineering and Computer Science, and a Research Scientist at Facebook AI Research (FAIR). His work bridges computer vision, machine learning, and deep learning, focusing on visual reasoning, vision-language tasks, image generation, and 3D reasoning using neural networks. PhD, Stanford University (advised by Fei-Fei Li) His research interests span visual reasoning , vision and language , 3D vision , and image generation , with a focus on innovative applications of deep neural networks. Recent publications highlight work on 3D consistency, self-supervised learning, and multimodal integration of vision and text. Notable contributions include PyTorch3D for 3D data processing, and foundational work in visual question answering , neural style transfer , and scene graph-based image generation . Publications span top conferences like ICCV, CVPR, and NeurIPS. He teaches courses including EECS 498/598: Deep Learning for Computer Vision and EECS 442: Computer Vision at University of Michigan, with prior involvement in Stanford's CS 231N in co-teaching roles. Software projects include open-source frameworks like fast-neural-style for real-time artistic style transfer, and PyTorch3D for efficient 3D deep learning. These tools demonstrate his commitment to practical implementations and community-driven research.
Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Christian Theobalt is a Professor of Computer Science at Saarland University and Scientific Director of the Visual Computing and Artificial Intelligence Department at the Max Planck Institute for Informatics . He leads the Saarbruecken Center for Visual Computing as a strategic partnership between Google and MPI. PhD in Computer Science (2005) from MPI-INF/Saarland University Postdoctoral Researcher at MPI (2005-2007) Visiting Assistant Professor at Stanford (2007-2009) His research focuses on the intersection of Computer Graphics, Computer Vision, and Artificial Intelligence , with specializations in: 3D/4D Human Reconstruction Neural Rendering Performance Capture Geometric Deep Learning Volumetric Video Quantum Visual Computing Recent article trends show emphasis on: Quantum computing applications in visual reconstruction Neural rendering with radiance fields Human motion capture from egocentric views Gesture-language interaction modeling Multi-modal scene understanding Real-time free-viewpoint rendering Scientific Awards : Fellow of EUROGRAPHICS (2022) CVPR Best Student Paper Honorable Mention (2020) ERC Consolidator Grant (2017) Karl Heinz Beckurts Award (2017) Busy Beaver Teaching Award (2016) ERC Starting Grant (2013) Advising over 50 students and researchers including: Current researchers: Viktor Rudnev, Linjie Lyu, Mohit Mendiratta Postdocs: Kwang In Kim, Kiran Varanasi Alumni: Franziska Mueller (Google), Dushyant Mehta (Qualcomm), Ayush Tewari (MIT) Grants include ERC grants, Google Glass Research Award, and multiple industry partnerships. His lab maintains cutting-edge facilities with: Multi-camera capture systems Quantum annealing infrastructure HDR radiance field technology GPU clusters for AI research Time-of-flight imaging systems Event camera arrays
Prof. Konrad Schindler holds the position of Full Professor at the Department of Civil, Environmental and Geomatic Engineering at ETH Zürich. He is also the Head of the Institute of Geodesy and Photogrammetry (IGP), leading research and educational activities in geomatics and computer vision. His career spans roles as a Photogrammetric Engineer, scientific assistant, postdoc researcher, and academic faculty across institutions including Graz University of Technology, Monash University, and TU Darmstadt before joining ETH Zürich in 2010. Education: Undergraduate studies in Geodesy (1992–1995), Graz University of Technology, Austria MEng in Photogrammetry and Geoinformation (1995–1999), Vienna University of Technology, Austria PhD in Computer Science (2001–2003), Graz University of Technology, Austria Research focuses on Photogrammetry , Remote Sensing , Computer Vision , and Image Understanding with interdisciplinary applications in environmental monitoring, geospatial analysis, and disaster response. He develops computational methods for 3D reconstruction, fusion of multi-modal data, and AI-driven solutions for satellite imagery interpretation. His work bridges geomatic engineering and machine learning to address challenges in urban mapping, climate modeling, and biological systems analysis. Publications reflect expertise in geospatial AI, diffusion models, and benchmarking datasets for disaster resilience. Notable works include Marigold (image analysis adaptation) and BRIGHT (building damage assessment). His research emphasizes practicality and scalability, such as affordable depth estimation and global biomass datasets. He has received the 2013 Marr Prize Honourable Mention (IEEE) and the 2012 U.V. Helava Award (ISPRS), alongside several Best Presentation Awards. His contributions span technical leadership, editorial roles (ISPRS Journal), and service to Swiss remote sensing commissions. Advising and grants: While no specific advisee names or grant details are listed, his career trajectory includes mentoring postdocs and junior faculty. He teaches advanced courses in Photogrammetry , Image Interpretation , and Machine Vision , integrating cutting-edge AI techniques into curricula. His research group collaborates on global-scale projects like canopy height mapping and satellite-based climate variable assessments. Labs/Teams: As Institute Head, he oversees the IGP lab at ETH Zürich, with prior affiliations including the Digital Perception Lab (Monash University) and the Computer Vision Lab (ETH Zurich). His work often involves multi-institutional collaborations focused on geospatial AI and environmental science.
Russell Epstein is a Professor and Director of Graduate Studies in the Department of Psychology at the University of Pennsylvania. He is affiliated with the Center for Cognitive Neuroscience and Goddard Labs. His research focuses on neural mechanisms underlying visual scene perception, spatial navigation, and memory. Epstein holds a BA in Physics from the University of Chicago and a PhD in Applied Mathematics from Harvard University. Epstein’s research interests include high-level vision, spatial cognition, and the neural basis of environmental representations. His lab uses functional MRI and cognitive neuroscience techniques to study how scenes, objects, landmarks, and spaces are encoded in brain systems such as the parahippocampal place area and retrosplenial cortex. Recent work explores cognitive maps, grid-like neural representations, and the role of multisensory cues in navigation. His articles emphasize spatial navigation strategies, hierarchical cognitive maps, and the interplay between perception and memory. Notable contributions include investigations into hippocampal spatial metrics, olfactory navigation, and the neural underpinnings of environmental learning. Epstein teaches courses on cognitive neuroscience, including PSYC 149 and PSYC 600. He advises two graduate students in Psychology and has no listed scientific awards. His work is supported by grants (unspecified) and conducted within collaborative teams at the Center for Cognitive Neuroscience. Epstein’s research extends to labs focused on spatial cognition and neuroimaging, advancing understanding of how humans mentally map environments through visual and sensory integration.
Minh Q. Phan is an Associate Professor of Engineering at Dartmouth College's Thayer School of Engineering. His expertise spans system identification, iterative learning control, model predictive control, robotic swarm control, and intelligent control systems. He holds a BS from the University of California, Berkeley, and MS/M.Phil/PhD degrees from Columbia University in Mechanical Engineering. Dr. Phan has contributed to over 50 peer-reviewed publications and serves as an Associate Editor for the Journal of Guidance, Control, and Dynamics. His research focuses on advancing control theory applications in robotics, structural health monitoring, and sustainable construction materials. Key contributions include the development of OKID (Observer/Kalman Filter Identification) methods and bilinear system identification frameworks. Education History: Bachelor of Science in Mechanical Engineering, UC Berkeley, 1985 Master of Science in Mechanical Engineering, Columbia University, 1986 Master of Philosophy in Mechanical Engineering, Columbia University, 1988 Doctor of Philosophy in Mechanical Engineering, Columbia University, 1989 Research Interests: Advanced control methodologies for dynamic systems Model-based predictive control strategies Applications in robotics and aerospace engineering Structural health monitoring via system identification Machine learning for materials science Teaching Responsibilities include courses like ENGG 149 (Systems Identification), ENGS 145 (Modern Control Theory), and ENGG 148 (Structural Mechanics). His work bridges theoretical control systems with practical industrial applications, including automation in food processing and sustainable construction practices. Dr. Phan has collaborated on projects addressing viral epidemiology in Vietnam and coastal erosion mitigation strategies.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
David Bamman is an Associate Professor in the School of Information at UC Berkeley, specializing in applying Natural Language Processing (NLP) and machine learning to cultural and social science questions. He leads research in born-literary NLP, computational humanities, and cultural analytics, with affiliated roles in EECS, Linguistics, and Computational Precision Health. Bamman holds degrees from Carnegie Mellon (Ph.D., 2015), Boston University (M.A., 2006), and University of Wisconsin-Madison (B.A., 1998). His work is supported by NEH, NSF, and industry grants. Educations: Ph.D. in Computer Science (2015), Carnegie Mellon University M.A. in Applied Linguistics (2006), Boston University B.A. in Classics (1998), University of Wisconsin-Madison Research Interests: NLP for underserved domains (e.g., literature, social media), coreference resolution, cultural analytics, and computational methods for studying literature and culture. Projects include LitBank and BookNLP datasets. Grants & Awards: Hellman Fellow (2019), Amazon Research Award (2017), NSF CAREER Award, and NEH funding. Teaching: Courses include Natural Language Processing (Info 159/259), Computational Humanities (INFO 190), and Applied NLP (INFO 256). His research group explores topics like racial representation in high school literature, Hollywood diversity metrics, and the sociocultural implications of LLMs. Bamman advises multiple PhD students and collaborates on datasets like CMU Book Summaries and 11K Latin Books.
Adriana Kovashka is an Associate Professor at the University of Pittsburgh , affiliated with the School of Computing and Information and serving as Department Chair . Her academic journey began with BA degrees in Computer Science and Media Studies from Pomona College (2008) and a PhD in Computer Science from The University of Texas at Austin (2014). Joined Pitt’s faculty in January 2015 NSF CAREER awardee (2021) Google Faculty Research Award recipient Dr. Kovashka’s research spans Computer Vision , Machine Learning , and Natural Language Processing , focusing on visual rhetoric, weak multimodal supervision, and domain adaptation. She pioneered techniques for analyzing political imagery, developing robust object detection frameworks, and exploring the intersection of visual and textual persuasion through large-scale annotated datasets. Her recent work emphasizes geographic diversity in vision-language systems, audio-visual fusion for domain generalization, and shape-texture bias mitigation in CNNs. Key publications include groundbreaking studies on symbolic reasoning, multimodal dialogue systems, and ethical AI applications in education. Scientific honors include: NSF CRII Award (2016) NSF CAREER Award (2021) Pitt CRDF Award (2016, 2018) Best Paper at ECV Workshop (2021) Google Faculty Research Award (2016, 2018) Dr. Kovashka actively mentors students in multimodal learning projects and collaborates with interdisciplinary teams on NSF-funded initiatives. She co-organizes workshops like the first CVPR workshop on advertisement understanding and leads research groups exploring human-AI co-learning systems.
Greg Durrett is an Associate Professor in the Department of Computer Science at University of Texas at Austin, leading the TAUR Lab (Text Analysis, Understanding, and Reasoning ). His research focuses on advancing Large Language Models (LLMs) for knowledge-intensive tasks in medical information processing scientific discovery legal reasoning . He received his B.S. in Computer Science and Mathematics from MIT (2010) and Ph.D. in Computer Science from UC Berkeley (2016). His work develops techniques to train LLMs with new capabilities augment models for reliability assess model outputs improve reasoning frameworks . His 15 most recent publications (2021-2025) span knowledge propagation in LLMs chain-of-thought reasoning code generation benchmarks multi-modal reasoning fact verification discourse analysis . Scientific honors include NSF CAREER Award (2024) NSF grants (2018, 2024) Bloomberg Data Science Grant (2017) Facebook Fellowship (2014) Best Paper Finalist (EMNLP 2013) . Teaching: CS388: Natural Language Processing (graduate) CS371N: NLP (undergraduate) High school NLP module .
Nima Fazeli is an Assistant Professor of Robotics at the University of Michigan (2020–Present), holding courtesy appointments in Computer Science & Engineering (CSE) and Mechanical Engineering. He directs the Manipulation and Machine Intelligence (MMint) Lab, focusing on enabling dexterous robotic manipulation through multimodal representation learning, tactile sensing, and model-based reasoning. His work integrates mechanics, perception, controls, and planning to achieve autonomous interaction with uncertain environments. Education: PhD, MIT (2019); MSc, University of Maryland (2014); BSc, Amirkabir University of Technology (2011) Research interests emphasize embodied intelligence , including visuo-tactile fusion, contact dynamics modeling, and cross-modal learning. Recent work explores tactile shadows, deformable object manipulation, and language-guided robot control. His research is supported by the NSF CAREER grant and National Robotics Initiative, with applications in manufacturing, assistive robotics, and space systems. Publications span topics like tactile sensing hardware (e.g., GelSlim 4.0), visuo-tactile implicit representations (ViTaSCOPE), and failure recovery policies (Racer). His team’s work has been featured in outlets like The New York Times and BBC. Key Awards: NSF CAREER Grant (2024) Teaching includes Introduction to Robotic Manipulation . Collaborations involve cross-disciplinary projects with mechanical, electrical, and biomedical engineering groups.
Hanjie Chen is an Assistant Professor in the Department of Computer Science at Rice University, affiliated with the Ken Kennedy Institute. She holds a Ph.D. from the University of Virginia and a Master's from the University of Science and Technology of China. Her research focuses on Natural Language Processing, Interpretable Machine Learning, and Trustworthy AI, emphasizing model explainability, alignment with human needs, and applications in healthcare, sports, and medicine. She has advised numerous students and led initiatives in AI ethics and education. Education: Ph.D. (Computer Science, UVA 2023), M.Sc. (USTC 2018), B.Sc. (Nanjing University of Aeronautics and Astronautics 2015). Awards include the Outstanding Doctoral Student Award (UVA 2023) and John A. Stankovic Research Award (UVA 2023). She has organized workshops like BlackboxNLP and served on program committees for ACL, NAACL, and EMNLP. Her recent work includes developing benchmarks like SPORTU for multimodal LLMs, evaluating medical question-answering systems, and advancing methods for robust rationale evaluation (RORA). She teaches courses on Natural Language Processing and Trustworthy NLP, emphasizing pedagogical innovation recognized by teaching awards at UVA. Research collaborations include internships at Microsoft Research, IBM, and the Allen Institute for AI. She mentors students in SURF programs and advocates for diversity in tech, serving as a mentor in UVA's CSGSG Council.
Alexander Schwing is an Associate Professor in the Department of Electrical and Computer Engineering and Computer Science at the University of Illinois at Urbana-Champaign, affiliated with the Coordinated Science Laboratory. His research focuses on machine learning and computer vision with applications in 3D scene understanding, generative modeling, and multi-agent systems. Education: Diploma in Electrical Engineering and Information Technology, Technical University of Munich (TUM) PhD in Computer Science, ETH Zurich Postdoctoral Fellow, University of Toronto Research Interests: Structured prediction in deep learning Generative adversarial networks and stability Multi-modal vision-language models 3D scene reconstruction from single images Embodied agent collaboration Semantic segmentation with temporal coherence Recent Publications: Highlight trends in neural rendering, video object segmentation, and reinforcement learning with applications to 3D modeling and multi-agent systems. Notable innovations include SAIL-VOS dataset for amodal segmentation and NeRFDeformer for single-view scene transformation. Scientific Awards: NSF CAREER Award, 3M and Amazon research awards, multiple student recognition awards, ETH Zurich PhD medal, and best paper at Intelligent Tutoring Systems 2014. Teaching: Offers graduate courses in Pattern Recognition (ECE 544) and Machine Learning (CS 446/ECE 449). Previously taught at University of Toronto and ETH Zurich. Labs & Collaborations: Leads research at Coordinated Science Laboratory (UIUC) with collaborations across University of Toronto, ETH Zurich, and industry partners like Samsung SAIT and Amazon.