Wei Xu is an Associate Professor at Georgia Institute of Technology's College of Computing and School of Interactive Computing, with affiliations to the Machine Learning Center. Their research bridges machine learning, natural language processing, and social media with focus areas in large language models, cultural bias mitigation, multilingual capabilities, and human-AI collaboration in text evaluation. NSF CAREER and Google Academic Research Award recipient Director of NLP X Lab PhD from New York University, BSMS from Tsinghua University Research interests span: Multilingual Multicultural LLMs addressing representational gaps and cultural adaptation in language models (NAACL 2025, ACL 2024); Robustness and Reasoning through dynamic AGI evaluations (ACL 2024, EMNLP 2024); Interdisciplinary NLP applications in security, healthcare, and law (EMNLP 2024, ACL 2024). Recent publications focus on multilingual alignment (NAACL 2025), privacy risk estimation (arXiv 2025), cultural bias analysis (ACL 2024), and medical text simplification (EMNLP 2024). Key themes include bias mitigation, multimodal processing, and practical LLM evaluation. Scientific Awards : NSF CAREER, Google/Sony/Criteo research awards, ACL'24 Best Social Impact Award, COLING'18 Best Paper Advising 15 PhD/MS/BSMS students including Yao Dou (human-centered LLM evaluation), Tarek Naous (multilingual LLMs), and alumni like Chao Jiang (Apple AI/ML) and Yang Chen (NVIDIA research scientist). Teaches graduate courses on NLP and LLMs.
Jean Oh is a Researcher at the Robotics Institute of Carnegie Mellon University (CMU) , leading the interdisciplinary Bot Intelligence Group (BIG) . Her work focuses on developing persistent robots that co-exist and collaborate with humans in shared environments, emphasizing continuous improvement through training, exploration, and human interaction. Education: Ph.D. in Language and Information Technologies, CMU M.S. in Computer Science, Columbia University B.S. in Biotechnology, Yonsei University Oh's research integrates vision, language, and planning systems in robotics, with applications in human-robot teaming , self-driving cars , disaster response , eldercare , and creative robotics . She has pioneered projects like socially-compliant robot navigation in human crowds and AI-driven robotic painting systems. Recent publication trends highlight her work in vision-language planning , social navigation , computational creativity , and human-robot collaboration . Notable contributions include the StyleCLIPDraw algorithm for text-to-art generation and Social-PatteRNN for human-like trajectory prediction. Scientific Awards: Best Paper Award in Cognitive Robotics (ICRA'18, ICRA'15) Best Systems Paper Finalist (HRI'25) Best Oral Paper Finalist (Humanoids'24) Best Paper in Entertainment (IROS'24) Argoverse Challenge Winner (CVPR'24) Best Student Paper (AIAA'24) Best Demo Finalist (RoboSoft'24) Oh mentors a diverse team of PhD, MS, and undergraduate students from CMU departments including Robotics, Computer Science, and Mechanical Engineering. Her research is funded by US Army Research Lab , DiDi Chuxing , and DARPA , with collaborations across industry and academia .
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Kristen Grauman is a Full Professor in the Department of Computer Science at the University of Texas at Austin, where she leads the UT Computer Vision Group. Her research focuses on computer vision and machine learning, with applications in visual recognition, video analysis, and multi-modal perception. She received her B.A. from Boston College and her Ph.D. from MIT. Her research interests span visual recognition, image and video search, video analysis, first-person vision, embodied and multi-modal perception, and interactive machine learning. She has made significant contributions to the field, particularly in developing algorithms for understanding visual content and human activities from video, including foundational work on the Pyramid Match Kernel and relative attributes. Her recent publications reveal a strong emphasis on egocentric (first-person) vision, audio-visual learning, and view-invariant representations. There is a clear trajectory toward multi-modal integration (vision, audio, language) and real-world applications in instructional videos, human activity understanding, and embodied AI systems. She has received numerous awards including: AAAI Fellow (2019) J. K. Aggarwal Prize, International Association for Pattern Recognition (2018) Helmholtz Prize (2017) UT Austin Academy of Distinguished Teachers (2017) Best Paper Award, Asian Conference on Computer Vision (2016) Presidential Early Career Award for Scientists and Engineers (2014) Computers and Thought Award, International Joint Conferences on Artificial Intelligence (2013) Pattern Analysis and Machine Intelligence Young Researcher Award (2013) Alfred P. Sloan Research Fellow (2012) Marr Prize (2011) Prof. Grauman serves as Associate Editor-in-Chief for the IEEE Transactions on Pattern Analysis and Machine Intelligence. She has secured substantial research funding including the Presidential Early Career Award, NSF grants, and industry partnerships. Her advising has produced numerous influential publications and students who are now leaders in computer vision. She leads the UT Computer Vision Group, which collaborates closely with the Electrical and Computer Engineering Department. The group is pioneering large-scale egocentric video research through projects like Ego4D and Ego-Exo4D, focusing on real-world applications in human activity understanding, audio-visual perception, and interactive systems.
Polina Golland is a Professor in the Department of Electrical Engineering and Computer Science (EECS) at MIT and a Principal Investigator in the Computer Science and Artificial Intelligence Laboratory (CSAIL). Her research focuses on developing novel techniques for biomedical image analysis and understanding, particularly in medical vision, AI/ML, and health care applications. She leads the Medical Vision Group and collaborates with the Vision Group at CSAIL. Her work emphasizes statistical modeling of medical images, shape modeling, and predictive analytics for biological processes. Current projects include fetal MRI analysis, cardiac MRI segmentation, and quantitative assessment of pulmonary edema in chest X-rays. She has secured grants from NIH, MIT-IBM Watson AI Lab, and other institutions to support her research. Dr. Golland teaches courses on inference, probability, and probabilistic systems. She advises graduate students in MIT's EECS program and has mentored numerous postdocs and researchers. Her lab focuses on translating advanced imaging techniques into clinical workflows, with applications in neuroimaging, fetal health monitoring, and cardiovascular disease analysis. Notable collaborations include work with Harvard Medical School affiliates, Brigham and Women's Hospital, and the MIT Jameel Clinic. Her research aims to bridge computational methods with clinical needs, improving diagnostic tools and treatment planning through machine learning and medical imaging innovation.
Adriana Kovashka is an Associate Professor at the University of Pittsburgh , affiliated with the School of Computing and Information and serving as Department Chair . Her academic journey began with BA degrees in Computer Science and Media Studies from Pomona College (2008) and a PhD in Computer Science from The University of Texas at Austin (2014). Joined Pitt’s faculty in January 2015 NSF CAREER awardee (2021) Google Faculty Research Award recipient Dr. Kovashka’s research spans Computer Vision , Machine Learning , and Natural Language Processing , focusing on visual rhetoric, weak multimodal supervision, and domain adaptation. She pioneered techniques for analyzing political imagery, developing robust object detection frameworks, and exploring the intersection of visual and textual persuasion through large-scale annotated datasets. Her recent work emphasizes geographic diversity in vision-language systems, audio-visual fusion for domain generalization, and shape-texture bias mitigation in CNNs. Key publications include groundbreaking studies on symbolic reasoning, multimodal dialogue systems, and ethical AI applications in education. Scientific honors include: NSF CRII Award (2016) NSF CAREER Award (2021) Pitt CRDF Award (2016, 2018) Best Paper at ECV Workshop (2021) Google Faculty Research Award (2016, 2018) Dr. Kovashka actively mentors students in multimodal learning projects and collaborates with interdisciplinary teams on NSF-funded initiatives. She co-organizes workshops like the first CVPR workshop on advertisement understanding and leads research groups exploring human-AI co-learning systems.
Greg Durrett is an Associate Professor in the Department of Computer Science at University of Texas at Austin, leading the TAUR Lab (Text Analysis, Understanding, and Reasoning ). His research focuses on advancing Large Language Models (LLMs) for knowledge-intensive tasks in medical information processing scientific discovery legal reasoning . He received his B.S. in Computer Science and Mathematics from MIT (2010) and Ph.D. in Computer Science from UC Berkeley (2016). His work develops techniques to train LLMs with new capabilities augment models for reliability assess model outputs improve reasoning frameworks . His 15 most recent publications (2021-2025) span knowledge propagation in LLMs chain-of-thought reasoning code generation benchmarks multi-modal reasoning fact verification discourse analysis . Scientific honors include NSF CAREER Award (2024) NSF grants (2018, 2024) Bloomberg Data Science Grant (2017) Facebook Fellowship (2014) Best Paper Finalist (EMNLP 2013) . Teaching: CS388: Natural Language Processing (graduate) CS371N: NLP (undergraduate) High school NLP module .
Robert Rohling is a Professor at the University of British Columbia's Faculty of Applied Science, affiliated with the Department of Mechanical Engineering and holding a joint appointment with the Department of Electrical and Computer Engineering. As Director of the Institute of Computing, Information and Cognitive Systems (ICICS), his research focuses on biomedical engineering, medical imaging, robotics, and computational methods. B.A.Sc. (UBC) M.Eng. (McGill) Ph.D. (Cambridge) Rohling's work spans three primary research areas: medical imaging (3D ultrasound, spatial compounding, elasticity reconstruction), medical information systems (radiologist navigation tools for large image datasets), and robotic calibration for surgical applications. His multidisciplinary approach integrates mechanical and electrical engineering principles with clinical needs. Rohling's publications (2020-2022) reveal trends in advanced ultrasound techniques (e.g., shear wave vibro-elastography), AI-driven image processing (cycleGAN translation), and computational optimization for diagnostic accuracy. Keywords across his work include Medical Imaging, Biomedical Engineering, Robotics, and Computational Modeling. As director of the Robotics and Control Laboratory , Rohling leads interdisciplinary collaborations with industry and clinical partners to address practical challenges in medical diagnostics and surgical robotics. His research emphasizes translating engineering innovations into clinical practice.
Ole Winther is Professor in High dimensional biological data analysis/Machine learning at the Department of Biology, University of Copenhagen and Professor in Data science and complexity at DTU Compute, Technical University of Denmark. He serves as CRO and co-founder of raffle.ai, CTO and co-founder of FindZebra, Head of ELLIS Unit Copenhagen, and co-PI of the Machine Learning for Life Science Center. His research spans Bioinformatics , Machine Learning , and AI for Science , focusing on applying deep learning to biological sequence analysis, latent variable models, and medical NLP. Winther's work develops predictive and generative models for bioinformatics, with significant contributions to protein localization tools (SignalP, DeepLoc, DeepTMHMM), single-cell genomics, and novel deep learning architectures like variational autoencoders and diffusion models. Analysis of Winther's recent publications (2023-2025) reveals a strong trend toward integrating protein language models with traditional bioinformatics approaches and applying diffusion models to scientific problems. His work bridges theoretical machine learning advancements with practical applications in biology and medicine, particularly in protein sequence analysis, medical search engines, and scientific simulation acceleration. Winther currently supervises a diverse research group including Panagiotis Antoniadis, Rachael M. DeVries, Jun Wang, Beatrix M. G. Nielsen, Felix G. Teufel, Irene R. Rodriguez, Anders Christensen, and Christopher Heje Grønbech. His former students have established successful careers at institutions including Google, Apple, and various startups, with notable alumni like Casper Sønderby (Google Brain) and Søren Sønderby (Apple). He leads significant research initiatives including the ELLIS Unit Copenhagen and the Machine Learning for Life Science Center, while maintaining active industry partnerships through his co-founded companies raffle.ai (enterprise search using NLP) and FindZebra (search engine for rare diseases). His teaching includes Deep Learning courses at both DTU (02456) and University of Copenhagen (NDAK24002U).
Alane Suhr is an Assistant Professor at UC Berkeley's Electrical Engineering and Computer Sciences (EECS) department and a member of the Berkeley Artificial Intelligence Research Lab (BAIR). Her research focuses on natural language processing (NLP), machine learning, and computer vision, emphasizing systems that interact with humans through language. She designs models and datasets for language grounding (e.g., NLVR) and develops algorithms for learning through interaction. Education: PhD in Computer Science, Cornell University (2022), advised by Yoav Artzi Bachelor's in Computer Science and Engineering, Ohio State University (2016), with a Linguistics minor Research Interests: Interactive systems for collaborative language use (e.g., CerealBar) Language grounding in multimodal contexts Embodied agents and reinforcement learning Generalization in NLP and the role of language in learning Key Contributions: Developed the NLVR dataset for visual reasoning with natural language Pioneered work on SWE-Gym for training software engineering agents Contributed to research on language models' sensitivity and commonsense reasoning Awards: ACM Doctoral Dissertation Award (2022) Outstanding Paper Awards at ACL 2023 and EMNLP 2021 Professional Activities: Organized workshops at ICML, NeurIPS, and ACL on topics like agents, generalization, and theory of mind Active in academic outreach and conference speaking (e.g., ICML, NeurIPS, CVPR) Labs/Teams: Berkeley Artificial Intelligence Research (BAIR) Lab SWE-Gym research group
Prof. Dr. Claudine Moulin is a Full Professor of Historical Linguistics at the Department of German Studies, University of Trier. She co-directs the Trier Center for Digital Humanities and serves as Vertrauensdozentin der Deutschen Forschungsgemeinschaft. With academic roots in Brussels and Bamberg, she completed her PhD in 1990 and Habilitation in 1999 at the University of Bamberg. Her research spans medieval and early modern languages, manuscript studies, digital humanities, and Luxembourgish linguistics. As a leading scholar in historical linguistics, Moulin's work explores urban language history, phraseology, and cultural knowledge transmission through manuscripts. She has directed major digital humanities projects like LexicoLux and Cartul@rium, while pioneering digital codicology methods for medieval texts. Her research combines traditional philology with computational approaches. Her academic accolades include the Akademiepreis Rheinland-Pfalz (2010), Landesverdienstorden (2014), and the Verdienstkreuz am Bande (2025). She has held visiting professorships at Sorbonne, EHESS Paris, and EPHE, and received fellowships from the Humboldt Foundation and European Science Foundation. Notable Academic Roles: Co-founder and Chair of DHd (Digital Humanities in German-speaking countries) Founding member of Trier Center for Language and Communication Member of German Historical Institute Paris scientific board Advisory positions at Austrian Center for Digital Humanities and Herzog August Bibliothek Her editorial leadership includes chief editorship of Sprachwissenschaft journal and co-editorship of Germanistische Bibliothek monograph series. She has mentored over 40 graduate students in topics ranging from medieval marginalia to urban language studies.
Lior Wolf is a Professor at the School of Computer Science, Tel Aviv University. Previously, he was a postdoctoral researcher at MIT's Center for Biological and Computational Learning (CBCL) under Prof. Tomaso Poggio and earned his PhD from Hebrew University of Jerusalem with Prof. Amnon Shashua. His educational background includes: PhD in Computer Science, Hebrew University of Jerusalem Postdoctoral Research, MIT CBCL Prof. Wolf's research centers on artificial intelligence with seminal contributions to deep learning, computer vision, and natural language processing. His work bridges theoretical foundations (e.g., attention mechanisms, transformer analysis) with practical applications in medical imaging, speech processing, and sign language technology. He pioneered methods for neural network interpretability, efficient sequence modeling, and multimodal fusion. Analysis of his 2023-2025 publications reveals dominant trends in large language model optimization (neuron pruning, attention analysis), efficient video generation, and cross-modal learning. His work increasingly integrates medical applications (fMRI/EEG analysis) while maintaining theoretical rigor in model architecture design. His scientific achievements include: Best paper award at EMNLP 2024 for 'Backward Lens: Projecting Language Model Gradients into the Vocabulary Space' Best paper award at SCIA 2023 for 'Gradient Adjusting Networks for Domain Inversion' Best paper award at FG 2021 for 'Generating Master Faces for Dictionary Attacks' Prof. Wolf mentors graduate students in the School of Computer Science and leads research at the ICRC building laboratory. His team collaborates with 'the friends of TAU' on projects spanning biometric security, medical imaging, and generative AI. Current work focuses on efficient transformers, neural network interpretability, and multimodal medical diagnostics.
Professor Hongdong Li is a Tenured Professor at the School of Computing, Australian National University (ANU), within the College of Engineering and Computer Science. His research focuses on 3D Computer Vision, Machine Learning, and their applications in dynamic environments. He has held visiting roles at Carnegie Mellon University and has contributed to significant projects like the Australia Bionic Eyes initiative. Education: PhD (Electrical Engineering). Research Interests : 3D Computer Vision fundamentals and applied AI systems Learning-based 3D perception for plant sciences Robot navigation in unfamiliar environments Awards : Marr Prize Honourable Mention CVPR Best Paper Award Advising & Grants : Supervised 40+ PhD students, with funding from ARC, CSIRO, Microsoft, and firms like OPPO/Tencent. Active in projects such as bushfire detection via video analytics and sign language translation systems. Labs/Teams : Co-founder of the Australian Centre for Robotic Vision (ACRV). Collaborates globally on cross-view localization and autonomous systems.
Andrew Zisserman is a Royal Society Research Professor at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). His research focuses on computer vision, artificial intelligence, and neural networks, with significant contributions to multimodal learning, video understanding, and 3D scene analysis. He leads projects exploring visual-language models, audio-visual synchronization, and clinical imaging applications. Key research areas include: Video analysis and temporal modeling Multimodal systems for sign language translation and action recognition 3D shape estimation and physical property inference Foundation models and cross-modal retrieval Recent work highlights: Developed Flamingo and Tapir models for video-language tasks Advancements in spinal MRI analysis and clinical imaging Leadership in EGO4D and VoxCeleb challenges Honors include Fellowship of the Royal Society (FRS) and the ISSLS Prize in Clinical Science 2023 for spinal analysis innovations. His lab collaborates globally, emphasizing real-world applications in healthcare and autonomous systems.
Jiatao Gu is an Assistant Professor in the Department of Computer and Information Science (CIS) at the University of Pennsylvania, with a part-time role as Staff Research Scientist at Apple (MLR). He holds a Ph.D. in Electrical and Electronic Engineering from the University of Hong Kong (2018) and a B.Eng. in Electronic Engineering from Tsinghua University (2014). His research focuses on generative machine learning and AI agent interaction with the physical world, emphasizing multi-modal systems spanning language, images, videos, and 3D. Key themes include efficient modeling , flexible architecture design , and scalable decision-making frameworks . 2025: ICLR paper on DART framework 2024: TMLR work on GFlowNet alignment 2023: NeurIPS research on diffusion stability 2022: ACL papers on speech translation Recent publications explore diffusion models for text-to-image synthesis, 3D reconstruction, and efficient sampling techniques. His work addresses fundamental challenges in attention mechanisms, entropy collapse, and multi-stage distillation while advancing non-autoregressive translation and vision-language reasoning . Prospective students can apply through his recruitment process at UPenn. Prior affiliations include Meta AI (FAIR Labs) and academic collaborations with institutions like New York University's CILVR Lab.