Wei Xu is an Associate Professor at Georgia Institute of Technology's College of Computing and School of Interactive Computing, with affiliations to the Machine Learning Center. Their research bridges machine learning, natural language processing, and social media with focus areas in large language models, cultural bias mitigation, multilingual capabilities, and human-AI collaboration in text evaluation. NSF CAREER and Google Academic Research Award recipient Director of NLP X Lab PhD from New York University, BSMS from Tsinghua University Research interests span: Multilingual Multicultural LLMs addressing representational gaps and cultural adaptation in language models (NAACL 2025, ACL 2024); Robustness and Reasoning through dynamic AGI evaluations (ACL 2024, EMNLP 2024); Interdisciplinary NLP applications in security, healthcare, and law (EMNLP 2024, ACL 2024). Recent publications focus on multilingual alignment (NAACL 2025), privacy risk estimation (arXiv 2025), cultural bias analysis (ACL 2024), and medical text simplification (EMNLP 2024). Key themes include bias mitigation, multimodal processing, and practical LLM evaluation. Scientific Awards : NSF CAREER, Google/Sony/Criteo research awards, ACL'24 Best Social Impact Award, COLING'18 Best Paper Advising 15 PhD/MS/BSMS students including Yao Dou (human-centered LLM evaluation), Tarek Naous (multilingual LLMs), and alumni like Chao Jiang (Apple AI/ML) and Yang Chen (NVIDIA research scientist). Teaches graduate courses on NLP and LLMs.
Antoine Bosselut is a Tenure Track Assistant Professor at École Polytechnique Fédérale de Lausanne (EPFL) in the School of Computer and Communication Sciences, where he leads the EPFL NLP group. His research focuses on developing AI reasoning agents that can model, represent, and reason about human and world knowledge, with applications in health, education, and global fairness. His research interests span multiple critical areas in modern AI: LLM Representations of Knowledge: Understanding what language models know and how they represent that knowledge internally Reasoning Algorithms: Developing methods to improve LLMs' reasoning capabilities through symbolic systems, neuroscience, and cognitive science Large-scale AI Development: Creating open-source foundation models with multilingual capabilities AI Democratization: Ensuring equitable access to AI technologies across different cultural and regulatory contexts His recent publications reveal a strong focus on the intersection of language models and cognitive science, with multiple studies examining how LLMs align with human cognition. He's also deeply engaged in practical applications of NLP technology, particularly in education and multilingual settings, as evidenced by his work on evaluating AI's impact on higher education and developing culturally-aware language models. His scientific achievements have been recognized with several prestigious awards: Outstanding Paper Award at NAACL 2025 ELLIS Scholar designation in 2024 Outstanding Paper Award at ACL 2023 Inclusion in Forbes 30 Under 30 list for Science & Healthcare in 2021 Bosselut actively mentors numerous PhD students across diverse NLP research areas and has secured significant recognition for his work in both academic and media circles. His lab has been featured in major publications including Communications of the ACM, The Atlantic, and Quanta Magazine, highlighting the societal impact of his research on commonsense reasoning and AI capabilities.
Jean Oh is a Researcher at the Robotics Institute of Carnegie Mellon University (CMU) , leading the interdisciplinary Bot Intelligence Group (BIG) . Her work focuses on developing persistent robots that co-exist and collaborate with humans in shared environments, emphasizing continuous improvement through training, exploration, and human interaction. Education: Ph.D. in Language and Information Technologies, CMU M.S. in Computer Science, Columbia University B.S. in Biotechnology, Yonsei University Oh's research integrates vision, language, and planning systems in robotics, with applications in human-robot teaming , self-driving cars , disaster response , eldercare , and creative robotics . She has pioneered projects like socially-compliant robot navigation in human crowds and AI-driven robotic painting systems. Recent publication trends highlight her work in vision-language planning , social navigation , computational creativity , and human-robot collaboration . Notable contributions include the StyleCLIPDraw algorithm for text-to-art generation and Social-PatteRNN for human-like trajectory prediction. Scientific Awards: Best Paper Award in Cognitive Robotics (ICRA'18, ICRA'15) Best Systems Paper Finalist (HRI'25) Best Oral Paper Finalist (Humanoids'24) Best Paper in Entertainment (IROS'24) Argoverse Challenge Winner (CVPR'24) Best Student Paper (AIAA'24) Best Demo Finalist (RoboSoft'24) Oh mentors a diverse team of PhD, MS, and undergraduate students from CMU departments including Robotics, Computer Science, and Mechanical Engineering. Her research is funded by US Army Research Lab , DiDi Chuxing , and DARPA , with collaborations across industry and academia .
Chua Tat Seng is a Professor at the School of Computing, National University of Singapore (NUS), holding the KITHCT Chair Professorship since 2009. He serves as co-Director of the NExT++ Center, a joint research center between NUS and Tsinghua University focused on Extreme Search. His academic career spans over three decades at NUS, where he has held various leadership positions including Acting Dean of the School of Computing (1998-2000) and Acting Head of the Department of Information Systems & Computer Science (1996-1998). Professor Chua's research spans unstructured data analytics , multimedia information retrieval , recommendation and conversation systems , and emerging applications in e-commerce and fintech . He established the Lab for Media Search (LMS) at the School of Computing and has been instrumental in advancing multimodal learning and search technologies. His work bridges theoretical foundations with practical applications, particularly in developing trustable AI systems for real-world deployment. His recent publications demonstrate a strong focus on large language models for recommendation systems , multimodal learning , and generative AI applications . The research trends show increasing emphasis on LLM-based recommendation, multimodal understanding, and addressing fundamental challenges in AI reliability, fairness, and efficiency. His work spans theoretical advancements in representation learning to practical applications in e-commerce, finance, and healthcare domains. ACM SIGMM Technical Achievement Award 2015 Multiple Best Paper Awards across ACM Multimedia, IEEE Transactions, and MMM conferences (2007-2020) Professor Chua has supervised 37 PhD students since 2004, establishing himself as a dedicated mentor in the academic community. His research has been supported by substantial grants including NExT++ ($12 million), Base Metals Price Forecasting ($200,000), and Multilingual Multimodal Knowledge Graph ($500,000). He maintains active collaborations with Tsinghua University, University of Southampton, and industry partners like Four Elements Capital and Singapore Press Holdings. As co-Director of the NExT++ Center, he leads a major research initiative focused on Web Intelligence and User Empowerment. His visiting professorships at Tsinghua University (2017-present) and Zhejiang University (2021-present) reflect his international impact in the field of multimedia and AI research.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Jennifer Olsen, PhD, is an Assistant Professor of Computer Science at the University of San Diego since 2020. She holds a PhD, MS, and BS in Human-Computer Interaction and Cognitive Science from Carnegie Mellon University, followed by postdoctoral research at the Swiss Federal Institute of Technology (EPFL), Lausanne, Switzerland. Her research focuses on the intersection of human-computer interaction, cognition, and education, emphasizing collaborative learning and educational technology design from both learner and instructor perspectives. Education: PhD in Human-Computer Interaction, Carnegie Mellon University MS in Human-Computer Interaction, Carnegie Mellon University BS in Cognitive Science, Carnegie Mellon University Research Interests: Dr. Olsen explores how collaboration supports learning, designs technologies to enhance educational practices, and investigates gaze-based metrics for understanding collaborative problem-solving. Her work spans gamified robotics, AI-driven orchestration systems, and virtual reality applications in vocational training. She emphasizes learner-centered design and the integration of social robots and virtual agents in pedagogical settings. Grants/Advising: While no specific grants or advisees are listed, her prolific publication record indicates active involvement in educational technology research and development. Her work addresses challenges in classroom orchestration, multimodal data analysis, and accessibility in educational robotics. Labs/Teams: Collaborates with interdisciplinary teams focused on educational technology, human-robot interaction, and adaptive learning systems. Her research leverages tools like FROG orchestration graphs and eye-tracking technologies to develop practical classroom solutions.
Luca Carlone is the Boeing Career Development Associate Professor in the Department of Aeronautics and Astronautics at MIT and a Principal Investigator at the Laboratory for Information & Decision Systems (LIDS) . He leads the SPARK Lab , focusing on developing certifiable perception algorithms for autonomous systems. PhD in Mechatronics (Polytechnic University of Turin, 2012) Research spans robotics, computer vision, and optimization Research Interests : Certifiable Perception algorithms for high-integrity systems High-level Perception (geometric, semantic, physical understanding) Efficient Perception methods for resource-constrained robots Scientific Contributions include: 2024 Outstanding Systems Paper Award (RSS) 2023 IEEE Transactions on Robotics King-Sun Fu Award 2021 NSF CAREER Award 2020 AIAA Advising Award 2019 Amazon Research Award Advising : Teaches graduate courses like Visual Navigation for Autonomous Vehicles and Robotics: Science and Systems . Collaborates with institutions including JPL, Caltech, and KAIST through the DARPA SubT Challenge.
Michael Kaess is an Associate Professor at the Robotics Institute, Carnegie Mellon University (CMU), within the School of Computer Science. He leads the Robot Perception Lab (RPL) and contributes to the Field Robotics Center (FRC) and Computer Vision Group (CV). His research focuses on efficient perception algorithms for mobile robots, particularly in 3D mapping, SLAM, and sensor fusion using vision, LiDAR, inertial, and sonar data. Kaess holds a PhD in Computer Science from Georgia Tech and was a postdoc at MIT's Marine Robotics Lab. Education: Georgia Institute of Technology, PhD in Computer Science (2008) MIT, Postdoctoral Associate (2008–2010) Research Interests: Kaess develops algorithms for robust and efficient inference in robotics, emphasizing factor graphs and linear algebra. His work spans underwater robotics, aerial systems, tactile SLAM, and multi-sensor integration. Key areas include SLAM with planes/lines, imaging sonar reconstruction, and neural field methods for LiDAR-visual fusion. Publications: Over 145 papers, including work on EDPLVO (visual odometry), HoloOcean (underwater simulation), and neural radiance fields with LiDAR. Recent trends focus on robust incremental smoothing, acoustic-optical fusion, and real-time volumetric mapping. Awards: Recognized with the RSS Test of Time Award (2020), Outstanding Associate Editor (2022), and paper awards at ICRA/ICRA. Active in conference organization (IROS/ICRA program committees). Advising & Grants: Supervises 10+ current PhD/MSc students, with past advisees contributing to CoRL/ICRA work. Manages grants in perception, autonomy, and marine robotics. Teaches courses like Robot Localization and Mapping (16-833). Labs/Teams: Directs RPL, collaborates with FRC on field robotics. Develops open-source tools like GTSAM (GNU Toolkit for Smoothing and Mapping).
Professor Andrew Davison holds the position of Professor of Robot Vision at Imperial College London's Department of Computing. He leads the Dyson Robotics Laboratory and the Robot Vision Research Group, focusing on advancing SLAM (Simultaneous Localization and Mapping) and Spatial AI. His groundbreaking work includes the MonoSLAM algorithm (2003), enabling real-time 3D vision for robotics and AR/VR. Current research emphasizes scalable, semantic-rich Spatial AI systems, as outlined in his FutureMapping papers (2018–2019). Education: BA in Physics (Oxford, 1994), D.Phil. (Oxford, 1998). Postdoctoral work at AIST, Japan (1998–2000), followed by a lectureship at Imperial (2002–present). Industrial collaborations include SLAMcore, a Spatial AI startup, and Dyson Robotics Lab. Over 18 PhD students supervised, many now leading roles at Meta, NVIDIA, SLAMcore, and academia. Notable contributions include DTAM, KinectFusion, and Event Camera SLAM. Recognized for software tools like SceneLib and contributions to robotics benchmarks (SLAMBench). Active on Twitter (@AjdDavison) for research updates.
Furong Huang is an Associate Professor at the University of Maryland's Department of Computer Science, with affiliations at the Institute for Advanced Computer Studies, Center for Machine Learning, Maryland Robotics Center, and Applied Mathematics, Statistics, and Scientific Computation Program. Her research bridges trustworthy machine learning, sequential decision-making, and foundation models for robotics, emphasizing reliability, interpretability, and ethical standards. Research Interests: Trustworthy AI Generative AI Reinforcement Learning AI Security Algorithmic Fairness Foundation Models for Robotics Recent Publications span leading conferences (NeurIPS, ICML, ICLR, CVPR) and journals, focusing on: Robustness in Vision-Language Systems Trustworthy Generative AI Foundation Models for Sequential Decision-Making AI Security and Watermarking Scientific Awards MIT TR35 Innovator Under 35 (Asia Pacific 2022) Best Paper Award, AdvML Frontier Workshop, NeurIPS 2024 NSF NAIRR Pilot Awardee Microsoft Accelerate Foundation Models Research Award (2023) JP Morgan Faculty Research Awards (2019–2022) Advising and Grants : Her lab has graduated students to roles at OpenAI, Google, Meta, and Netflix. Research funded by DARPA, NSF, ONR, AFOSR, and industry partners like Microsoft, Adobe, and Capital One. Labs & Teams : Leads research groups focused on Trustworthy AI and Robotics at the University of Maryland, collaborating with the Maryland Robotics Center and Applied Mathematics Program.
Lerrel Pinto is an Assistant Professor of Computer Science at the Courant Institute of Mathematical Sciences at New York University (NYU), where he leads the General-purpose Robotics and AI Lab (GRAIL) as part of the CILVR research group. His work bridges the gap between theoretical machine learning and practical robotics applications, with a focus on enabling robots to generalize and adapt in real-world environments. Dr. Pinto received his undergraduate degree from IIT Guwahati, followed by a PhD from the Robotics Institute at Carnegie Mellon University (CMU). He then completed a postdoctoral fellowship at the University of California, Berkeley before joining NYU as faculty. His research program centers on robot learning and decision making, with several key thrusts that demonstrate his innovative approach to robotics. Pinto's work emphasizes large-scale learning techniques that leverage both extensive data and sophisticated model architectures. A significant portion of his research focuses on representation learning for sensory data, particularly developing methods that enable robots to make sense of visual, tactile, and auditory inputs. His lab has made notable contributions to reinforcement learning algorithms that allow robots to adapt to new scenarios with minimal retraining. Pinto also champions open-source robotics , developing affordable robot platforms that democratize access to robotics research. Analysis of Pinto's recent publications reveals a strong trend toward multimodal perception in robotics, integrating visual, tactile, and auditory information to create more robust robot systems. His work increasingly focuses on zero-shot and few-shot learning capabilities, enabling robots to handle novel situations without extensive retraining. There's also a clear progression toward general-purpose robotics , moving away from task-specific solutions toward more flexible systems that can handle diverse real-world challenges. Dr. Pinto's scientific contributions have been recognized with several prestigious awards: Sloan Research Fellowship (2025) NSF CAREER Award (2024) RAL Early Career Award (2024) Best Student Paper Award at ICRA (2016) Outstanding Paper Award at MFM-EAI workshop at ICML (2024) Best Paper Award at NGSM workshop at ICML (2024) Best Student Paper Award at RSS (2023) As an advisor, Pinto has mentored numerous students who have gone on to impactful careers in both academia and industry. His former PhD student Denis Yarats co-founded Perplexity.AI, while Mahi Shafiullah became a postdoc at UC Berkeley and Meta AI. Many of his Masters students have pursued PhDs at top institutions like CMU, MIT, and Stanford, or joined leading robotics companies including 1X, Fauna Robotics, and NVIDIA. Pinto's lab has secured significant research funding, including the NSF CAREER award and likely other grants supporting his robotics research program. The General-purpose Robotics and AI Lab (GRAIL) that Pinto leads brings together a diverse team of researchers working on cutting-edge robotics challenges. The lab maintains strong collaborations with industry partners and other academic institutions, facilitating technology transfer and real-world impact. GRAIL's research spans multiple robotics platforms and focuses on developing algorithms that enable robots to learn from diverse experiences and generalize across environments.
Chris Donahue is an Assistant Professor in the Computer Science Department at Carnegie Mellon University . He also serves as a part-time Research Scientist at Google DeepMind on the Magenta team. His work focuses on leveraging generative AI to enhance human creativity, particularly in music. Education: PhD in Computer Science (UC San Diego), Postdoctoral Scholar (Stanford University) His research spans controllable generative modeling of music and audio , with a focus on real-time interactive systems. Projects like Piano Genie , Beat Sage , and Copilot Arena demonstrate his commitment to real-world deployment. His Generative Creativity Lab (G-CLef) explores AI applications beyond music, including programming and natural language. Recent publications highlight advancements in multimodal music evaluation , real-time adaptation , and AI-driven sound morphing . He co-developed Magenta RealTime , an open-weight real-time music generation model, and MusicFX DJ Mode . Scientific Awards: Best Paper Award (top 1) at NAACL Student Research Workshop 2025 Best Paper Award (top 1% of submissions) at CHI 2025 Best Paper Runner-up at ISMIR 2021 He co-advises PhD students like Wayne Chi (NDSEG Fellow) and mentors Irmak Bukey . His lab receives support from the AIxArts incubator fund at CMU .
Adriana Kovashka is an Associate Professor at the University of Pittsburgh , affiliated with the School of Computing and Information and serving as Department Chair . Her academic journey began with BA degrees in Computer Science and Media Studies from Pomona College (2008) and a PhD in Computer Science from The University of Texas at Austin (2014). Joined Pitt’s faculty in January 2015 NSF CAREER awardee (2021) Google Faculty Research Award recipient Dr. Kovashka’s research spans Computer Vision , Machine Learning , and Natural Language Processing , focusing on visual rhetoric, weak multimodal supervision, and domain adaptation. She pioneered techniques for analyzing political imagery, developing robust object detection frameworks, and exploring the intersection of visual and textual persuasion through large-scale annotated datasets. Her recent work emphasizes geographic diversity in vision-language systems, audio-visual fusion for domain generalization, and shape-texture bias mitigation in CNNs. Key publications include groundbreaking studies on symbolic reasoning, multimodal dialogue systems, and ethical AI applications in education. Scientific honors include: NSF CRII Award (2016) NSF CAREER Award (2021) Pitt CRDF Award (2016, 2018) Best Paper at ECV Workshop (2021) Google Faculty Research Award (2016, 2018) Dr. Kovashka actively mentors students in multimodal learning projects and collaborates with interdisciplinary teams on NSF-funded initiatives. She co-organizes workshops like the first CVPR workshop on advertisement understanding and leads research groups exploring human-AI co-learning systems.
Katerina Fragkiadaki is the JPMorgan Chase Associate Professor of Computer Science in the Machine Learning Department at Carnegie Mellon University. She works at the intersection of Artificial Intelligence, Computer Vision, Machine Learning, Language Understanding, and Robotics. PhD from GRASP Lab, University of Pennsylvania Postdoctoral researcher at UC Berkeley (with Jitendra Malik) and Google Research Recipient of NSF CAREER, DARPA Young Investigator, Amazon, Google, Sony, UPMC, and AFOSR awards Organizer of CoRL 2023 Workshop on Generalist Robots ICLR 2024 Program Chair, multiple area chair roles Her research group focuses on developing machines that autonomously improve world models through human-environment interactions, with specific emphasis on: Representation learning and video understanding 2D/3D unified vision-language models Generative simulation and reinforcement learning Real2Sim/Sim2Real robot learning Continual learning and spatial common sense 3D scene reconstruction and dynamics Recent publications highlight advancements in: 3D mesh generation with compositional transformers Unified 2D/3D perception frameworks Physics-aware generative models Diffusion-based robotic manipulation policies Embodied agents with memory prompting Awards include: 2024: DARPA Young Investigator Award 2023: Amazon Faculty Award 2022: Sony Faculty Research Award 2021: UPMC Faculty Research Award 2020: NSF CAREER Award 2019: Google Faculty Award Key collaborations span institutions including UC Berkeley, Google Research, Stanford, MIT, and University of Tsukuba. Her work bridges theoretical innovation with practical applications in: Autonomous robot manipulation 4D world modeling Language-grounded perception Visual dynamics prediction Embodied program synthesis Physics-based simulation engines
Alane Suhr is an Assistant Professor at UC Berkeley's Electrical Engineering and Computer Sciences (EECS) department and a member of the Berkeley Artificial Intelligence Research Lab (BAIR). Her research focuses on natural language processing (NLP), machine learning, and computer vision, emphasizing systems that interact with humans through language. She designs models and datasets for language grounding (e.g., NLVR) and develops algorithms for learning through interaction. Education: PhD in Computer Science, Cornell University (2022), advised by Yoav Artzi Bachelor's in Computer Science and Engineering, Ohio State University (2016), with a Linguistics minor Research Interests: Interactive systems for collaborative language use (e.g., CerealBar) Language grounding in multimodal contexts Embodied agents and reinforcement learning Generalization in NLP and the role of language in learning Key Contributions: Developed the NLVR dataset for visual reasoning with natural language Pioneered work on SWE-Gym for training software engineering agents Contributed to research on language models' sensitivity and commonsense reasoning Awards: ACM Doctoral Dissertation Award (2022) Outstanding Paper Awards at ACL 2023 and EMNLP 2021 Professional Activities: Organized workshops at ICML, NeurIPS, and ACL on topics like agents, generalization, and theory of mind Active in academic outreach and conference speaking (e.g., ICML, NeurIPS, CVPR) Labs/Teams: Berkeley Artificial Intelligence Research (BAIR) Lab SWE-Gym research group