Wei Xu is an Associate Professor at Georgia Institute of Technology's College of Computing and School of Interactive Computing, with affiliations to the Machine Learning Center. Their research bridges machine learning, natural language processing, and social media with focus areas in large language models, cultural bias mitigation, multilingual capabilities, and human-AI collaboration in text evaluation. NSF CAREER and Google Academic Research Award recipient Director of NLP X Lab PhD from New York University, BSMS from Tsinghua University Research interests span: Multilingual Multicultural LLMs addressing representational gaps and cultural adaptation in language models (NAACL 2025, ACL 2024); Robustness and Reasoning through dynamic AGI evaluations (ACL 2024, EMNLP 2024); Interdisciplinary NLP applications in security, healthcare, and law (EMNLP 2024, ACL 2024). Recent publications focus on multilingual alignment (NAACL 2025), privacy risk estimation (arXiv 2025), cultural bias analysis (ACL 2024), and medical text simplification (EMNLP 2024). Key themes include bias mitigation, multimodal processing, and practical LLM evaluation. Scientific Awards : NSF CAREER, Google/Sony/Criteo research awards, ACL'24 Best Social Impact Award, COLING'18 Best Paper Advising 15 PhD/MS/BSMS students including Yao Dou (human-centered LLM evaluation), Tarek Naous (multilingual LLMs), and alumni like Chao Jiang (Apple AI/ML) and Yang Chen (NVIDIA research scientist). Teaches graduate courses on NLP and LLMs.
Dr. Chang Xu is an Associate Professor in Machine Learning and Computer Vision at the University of Sydney's School of Computer Science. He holds a Bachelor of Engineering from Tianjin University and a PhD from Peking University. His research focuses on machine learning, data mining, and their applications in AI and computer vision, including multi-view learning, visual search, and face recognition. He is an ARC Future Fellow and a member of the Sydney Southeast Asia Centre and The Net Zero Institute. Education: B.E. in Engineering (Tianjin University), Ph.D. in Computer Science (Peking University). His research interests emphasize handling heterogeneous data, exploring data variety, and developing algorithms for robust AI systems. His work includes adversarial robustness, neural architecture search, and efficient deep learning models. Research trends in his articles include adversarial robustness in neural architectures, efficient vision transformers, multimodal 3D style transfer, and underwater image restoration. Key contributions span image restoration, video super-resolution, and lightweight network design. He has advised multiple PhD and master's students on topics like diffusion models, radar image synthesis, and graph similarity. Awards: ARC Future Fellow. Collaborations focus on cross-domain data integration and AI applications. His labs and teams explore generative models, robust learning, and scalable robotics policies. Recent work includes diffusion models for action segmentation and robust vision-language systems.
Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Kai-Wei Chang is an Associate Professor at the University of California, Los Angeles (UCLA) in the Department of Computer Science, part of the Henry Samueli School of Engineering. He is also an Amazon Scholar at Alexa AI, focusing on advancing trustworthy AI and multimodal foundation models. His research bridges NLP, machine learning, and ethical AI, with a focus on fairness, robustness, and bias mitigation in language and vision-language systems. Education: Ph.D. in Computer Science (UIUC, 2015), M.S. and B.S. in Computer Science and Electrical Engineering from National Taiwan University. Research Interests: Trustworthy NLP (fairness, robustness), Multimodal Foundation Models (e.g., VisualBERT, GLIP), Reasoning in LLMs, and mitigating societal biases in AI systems. Notable contributions include pioneering work on aligning NLP models with human values and developing SOTA multimodal models like DesCo and GLIP. Awards: Sloan Research Fellowship (2021), Okawa Grant (2018), EMNLP Best Paper (2017), KDD Best Paper (2010). His work is funded by NSF, IARPA, ONR, and industry partners like Amazon, Google, and Facebook. Service: VP-Elect of SIGDAT, Ethics Committee Chair (NAACL 2022), Organizer of Trustworthy NLP Workshops, and Senior Area Chair for top conferences (ACL, NeurIPS, AAAI). Labs/Teams: Leads the UCLA Natural Language Processing Group, fostering interdisciplinary research on ethical AI and multimodal systems.
Carl Vondrick is a Professor in the Department of Computer Science at Columbia University. His research focuses on creating robust and versatile perception systems that leverage video and interaction with the natural world, with applications in 3D reconstruction, visual question answering, and robot manipulation. Former research scientist at Google Visiting researcher at Cruise Education: PhD (2017) from MIT, advised by Antonio Torralba BS (2011) from UC Irvine, advised by Deva Ramanan His research explores multimodal approaches for cross-task and cross-modal transfer, scene dynamics, audiovisual perception, interpretable models, and spatial awareness systems. The lab emphasizes zero-shot generalization and neuro-symbolic methods while addressing safety and robustness in AI systems. Key publication trends include: 2025: Video generation for robotics 2024: Differentiable rendering and cross-modal reasoning 2023: Robust perception and 3D modeling Scientific Awards: 2024 PAMI Young Researcher Award 2021 NSF CAREER Award Teaching Roles: Teaching Computer Vision II (2021-2025), Computer Vision I (2018-2019), and Representation Learning (2020-2022). Advising: Advises 8 current PhD students and has mentored 5 graduated students now at institutions like MBZUAI and UMD. The lab recruits 1-2 PhD students annually through Columbia’s PhD program. Grants and Collaborations: Funded by NSF, DARPA, Toyota Research Institute, Amazon Research, and Google.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Luca Carlone is the Boeing Career Development Associate Professor in the Department of Aeronautics and Astronautics at MIT and a Principal Investigator at the Laboratory for Information & Decision Systems (LIDS) . He leads the SPARK Lab , focusing on developing certifiable perception algorithms for autonomous systems. PhD in Mechatronics (Polytechnic University of Turin, 2012) Research spans robotics, computer vision, and optimization Research Interests : Certifiable Perception algorithms for high-integrity systems High-level Perception (geometric, semantic, physical understanding) Efficient Perception methods for resource-constrained robots Scientific Contributions include: 2024 Outstanding Systems Paper Award (RSS) 2023 IEEE Transactions on Robotics King-Sun Fu Award 2021 NSF CAREER Award 2020 AIAA Advising Award 2019 Amazon Research Award Advising : Teaches graduate courses like Visual Navigation for Autonomous Vehicles and Robotics: Science and Systems . Collaborates with institutions including JPL, Caltech, and KAIST through the DARPA SubT Challenge.
Michael Kaess is an Associate Professor at the Robotics Institute, Carnegie Mellon University (CMU), within the School of Computer Science. He leads the Robot Perception Lab (RPL) and contributes to the Field Robotics Center (FRC) and Computer Vision Group (CV). His research focuses on efficient perception algorithms for mobile robots, particularly in 3D mapping, SLAM, and sensor fusion using vision, LiDAR, inertial, and sonar data. Kaess holds a PhD in Computer Science from Georgia Tech and was a postdoc at MIT's Marine Robotics Lab. Education: Georgia Institute of Technology, PhD in Computer Science (2008) MIT, Postdoctoral Associate (2008–2010) Research Interests: Kaess develops algorithms for robust and efficient inference in robotics, emphasizing factor graphs and linear algebra. His work spans underwater robotics, aerial systems, tactile SLAM, and multi-sensor integration. Key areas include SLAM with planes/lines, imaging sonar reconstruction, and neural field methods for LiDAR-visual fusion. Publications: Over 145 papers, including work on EDPLVO (visual odometry), HoloOcean (underwater simulation), and neural radiance fields with LiDAR. Recent trends focus on robust incremental smoothing, acoustic-optical fusion, and real-time volumetric mapping. Awards: Recognized with the RSS Test of Time Award (2020), Outstanding Associate Editor (2022), and paper awards at ICRA/ICRA. Active in conference organization (IROS/ICRA program committees). Advising & Grants: Supervises 10+ current PhD/MSc students, with past advisees contributing to CoRL/ICRA work. Manages grants in perception, autonomy, and marine robotics. Teaches courses like Robot Localization and Mapping (16-833). Labs/Teams: Directs RPL, collaborates with FRC on field robotics. Develops open-source tools like GTSAM (GNU Toolkit for Smoothing and Mapping).
Shoudong Huang is a Professor at the School of Mechanical and Mechatronic Engineering , University of Technology Sydney, and Deputy Director of the UTS Robotics Institute. His research focuses on mobile robot navigation , SLAM , nonlinear state estimation , and surgical robotics . He has published over 200 papers and is recognized as one of the 100 Most Influential Scholars in Robotics (Aminer, 2018). PhD in Automatic Control, Northeastern University (China) Postdoctoral Research Fellow, University of Hong Kong (1998-2000) Research Fellow, Australian National University (2001-2003) Full-time academic roles at UTS since 2004 His work addresses challenges in robot localization across extreme environments (underwater, underground mining, surgical settings) and develops globally optimal SLAM algorithms with guaranteed performance. He has secured over $4 million AUD in external funding, including ARC Discovery grants and industry partnerships. Recent publications emphasize cross-modal calibration (camera-LiDAR), interval analysis for bounded noise , and template-based deformable surface reconstruction . These span applications in autonomous driving, surgical navigation, and UAV guidance. Chancellor’s Medal for Research Excellence (2020) Supervisor of the Year (2023) Best Paper Award (2016 ICARCV) Huang serves as Associate Editor for IEEE Transactions on Robotics and International Journal of Robotics Research , and has held leadership roles in top robotics conferences like IROS and RSS. His collaborations span MIT, USC, Zhejiang University, and industry partners including PMSW Research Pty Ltd and Multiplex Constructions Pty Ltd.
Vineeth N Balasubramanian is a Professor in the Department of Computer Science & Engineering at the Indian Institute of Technology Hyderabad, with affiliate faculty status in the Department of Artificial Intelligence. His research focuses on the intersection of deep learning, machine learning, and computer vision, emphasizing explainability, robustness, and real-world applications. He leads Lab 1055, which investigates problems such as Explainable and robust AI/ML systems Lifelong learning in evolving environments Multimodal vision-language models Applications in agriculture, autonomous navigation, and human behavior analysis His recent work includes causal reasoning in transformers, vision-language model capabilities, and drone-based object detection. Funded by organizations like Google, Microsoft, Intel, and DST, he has received multiple awards including the World's Top 2% Scientists (2022-23), INSA/INAE Fellowships, and Best Paper recognitions. Lab 1055 collaborates with institutions like CMU, UBC, and Monash University, contributing to cutting-edge advancements in AI.
Lerrel Pinto is an Assistant Professor of Computer Science at the Courant Institute of Mathematical Sciences at New York University (NYU), where he leads the General-purpose Robotics and AI Lab (GRAIL) as part of the CILVR research group. His work bridges the gap between theoretical machine learning and practical robotics applications, with a focus on enabling robots to generalize and adapt in real-world environments. Dr. Pinto received his undergraduate degree from IIT Guwahati, followed by a PhD from the Robotics Institute at Carnegie Mellon University (CMU). He then completed a postdoctoral fellowship at the University of California, Berkeley before joining NYU as faculty. His research program centers on robot learning and decision making, with several key thrusts that demonstrate his innovative approach to robotics. Pinto's work emphasizes large-scale learning techniques that leverage both extensive data and sophisticated model architectures. A significant portion of his research focuses on representation learning for sensory data, particularly developing methods that enable robots to make sense of visual, tactile, and auditory inputs. His lab has made notable contributions to reinforcement learning algorithms that allow robots to adapt to new scenarios with minimal retraining. Pinto also champions open-source robotics , developing affordable robot platforms that democratize access to robotics research. Analysis of Pinto's recent publications reveals a strong trend toward multimodal perception in robotics, integrating visual, tactile, and auditory information to create more robust robot systems. His work increasingly focuses on zero-shot and few-shot learning capabilities, enabling robots to handle novel situations without extensive retraining. There's also a clear progression toward general-purpose robotics , moving away from task-specific solutions toward more flexible systems that can handle diverse real-world challenges. Dr. Pinto's scientific contributions have been recognized with several prestigious awards: Sloan Research Fellowship (2025) NSF CAREER Award (2024) RAL Early Career Award (2024) Best Student Paper Award at ICRA (2016) Outstanding Paper Award at MFM-EAI workshop at ICML (2024) Best Paper Award at NGSM workshop at ICML (2024) Best Student Paper Award at RSS (2023) As an advisor, Pinto has mentored numerous students who have gone on to impactful careers in both academia and industry. His former PhD student Denis Yarats co-founded Perplexity.AI, while Mahi Shafiullah became a postdoc at UC Berkeley and Meta AI. Many of his Masters students have pursued PhDs at top institutions like CMU, MIT, and Stanford, or joined leading robotics companies including 1X, Fauna Robotics, and NVIDIA. Pinto's lab has secured significant research funding, including the NSF CAREER award and likely other grants supporting his robotics research program. The General-purpose Robotics and AI Lab (GRAIL) that Pinto leads brings together a diverse team of researchers working on cutting-edge robotics challenges. The lab maintains strong collaborations with industry partners and other academic institutions, facilitating technology transfer and real-world impact. GRAIL's research spans multiple robotics platforms and focuses on developing algorithms that enable robots to learn from diverse experiences and generalize across environments.
Xingxing Zuo is an Assistant Professor (tenure-track) in the Robotics Department at MBZUAI. He holds a PhD from Zhejiang University (2021) and a Bachelor’s from UESTC (2016). Previously, he was a Postdoctoral Scholar at Caltech (2024–2025), a Postdoc at ETH Zurich (2019–2021), and held visiting roles at TU Munich, University of Delaware, and University of Technology Sydney. His research focuses on robotics, 3D computer vision, and embodied AI, with emphasis on robot-human collaboration, state estimation, and sensor fusion. Educations: PhD in Robotics, Zhejiang University (2021, with honors) Bachelor’s in Computer Science, University of Electronic Science and Technology of China (2016, with honors) Research Highlights: Develops novel methods for LiDAR-camera-inertial fusion, neural radiance fields, and radar-cameras systems Pioneered techniques like Flying Co-Stereo (long-range aerial mapping) and FMGS (vision-language embedded 3D splatting) Focuses on real-time SLAM, robust depth estimation, and photorealistic scene reconstruction Awards & Recognition: Best Paper Finalist at ICRA 2021 (CodeVIO) Oral Presentation at ICCV 2021 (MBA-VO) Recipient of Google Visiting Faculty Researcher (2023) Grants & Labs: Organized Thermal Infrared in Robotics workshop at ICRA 2025 Leads research on embodied AI and multi-sensor SLAM systems Develops open-source tools like LIC-Fusion and Coco-LIC frameworks
Olga Sorkine-Hornung is a Professor of Computer Science at ETH Zurich and head of the Institute of Visual Computing. She leads the Interactive Geometry Lab, focusing on theoretical and practical advancements in digital content creation, geometry processing, and shape modeling. Current position: ETH Zurich, Department of Computer Science Previous roles: Courant Institute (NYU), Technical University of Berlin Education: BSc and PhD from Tel Aviv University, postdoc at TU Berlin Her research spans shape representation, digital fabrication, computer animation, and fundamental geometry processing. Key contributions include Laplacian surface editing, as-rigid-as-possible deformation, and generalized winding numbers. She works on applications in VR, AR, and autonomous systems. Awards include Test of Time Awards (2024), ACM Fellow (2020), ERC Consolidator Grant (2020), and EUROGRAPHICS Young Researcher Award (2008). She has supervised numerous students and co-developed software libraries like libigl and Instant Meshes . Co-chair roles for SIGGRAPH, Eurographics, and Pacific Graphics Editorial board member for ACM Transactions on Graphics and other journals Keynote speaker at VMV, CVPR, and SIAM conferences Her work bridges mathematical rigor with practical implementation, advancing computer graphics and geometry processing through intuitive algorithms that maintain surface detail while enabling efficient computation.
Alexei A. Efros is a Professor in the Department of Electrical Engineering and Computer Sciences (EECS) at UC Berkeley, where he holds the Howard Friesen Professorship and is affiliated with the Berkeley Artificial Intelligence Research (BAIR) Lab. He previously served on the faculty at the Robotics Institute of Carnegie Mellon University (CMU) and completed a postdoctoral fellowship at the University of Oxford. His research spans data-driven computer vision, self-supervised learning, computational photography, and applications to computer graphics and robotics. His research interests include: Data-Driven Computer Vision Self-Supervised and Unsupervised Learning Generative Models and Image Synthesis Visual Representation Learning Applications in Robotics and Human-Computer Interaction Intersections with Human Vision and the Humanities The recent publications highlight a strong trend toward self-supervised learning, visual reasoning, and generative modeling, particularly diffusion models and 3D scene understanding. His work increasingly bridges computer vision with language, robotics, and cognitive science, emphasizing interpretability and real-world applicability. There is a clear focus on leveraging unlabeled data and developing methods for robust, generalizable AI systems. His scientific awards and recognitions include: Berkeley Fellowship Google Fellowship Soros Fellowship NSF Fellowship SIGGRAPH Outstanding Doctoral Dissertation Award Facebook Fellowship Adobe Fellowship CMU School of Computer Science Distinguished Dissertation Award ACM Doctoral Dissertation Honorable Mention Alexei Efros has advised numerous PhD students and postdocs, many of whom have gone on to faculty positions at top institutions including CMU, Stanford, MIT, Columbia, NYU, and Georgia Tech. His lab has received research funding from major tech companies and federal agencies, though specific grants are not detailed in the text. He teaches core computer vision and machine learning courses at both undergraduate and graduate levels at UC Berkeley. His research group is highly active, with ongoing projects in 3D perception, generative modeling, and vision-language systems. He leads a vibrant research lab at UC Berkeley, part of the BAIR consortium, collaborating with leading researchers such as Jitendra Malik, Trevor Darrell, Pieter Abbeel, and Angjoo Kanazawa. His lab fosters strong interdisciplinary connections with institutions worldwide, including Oxford, INRIA, and École Normale Supérieure.
Wei-Lun (Harry) Chao is an Associate Professor in the Department of Computer Science and Engineering at the Ohio State University (OSU), College of Engineering. Promoted to this role in May 2025, he is also an Innovation Scholar and Distinguished Assistant Professor of Engineering Inclusive Excellence. His work spans machine learning, computer vision, and their applications in autonomous driving, healthcare, biology, and natural language processing. Research Focus: Machine learning with imperfect data, interpretable and personalized learning, robust perception for autonomous systems, and visual recognition in real-world scenarios. Awards: 2025 OSU Early Career Distinguished Scholar Award, CVPR Best Student Paper Award (2024), CSE Faculty Teaching Award (2024), Lumley Research Award (2023). Grants: Funded by NSF, NIH, ONR, Cisco, AWS, and Google. Notable Research Trends: The 15 most recent articles highlight his work on vision foundation models, federated learning, diffusion models for biological species generation, interpretable vision transformers, and robust perception systems for autonomous driving. Key subfields include sparse autoencoders, 3D object detection, semi-supervised learning, and anomaly detection in scientific domains. Scientific Awards: 2025 Early Career Distinguished Scholar Award (OSU) CVPR Best Student Paper Award (2024) CSE Faculty Teaching Award (2024) Lumley Research Award (2023) Mentoring & Grants: As an advisor for the OSU Buckeye AutoDrive Team and AI Club, he mentors graduate and undergraduate students. His research is supported by major grants from NSF, NIH, ONR, and industry partners like Cisco and Google.