Sebastian Riedel is a Professor at University College London (UCL) and a Researcher at DeepMind, leading the UCL NLP Lab. His work focuses on teaching machines to read, reason, and write, integrating Natural Language Processing (NLP) with Machine Learning. He holds an Allen Distinguished Investigator award and has held roles at FAIR, UMass Amherst, Tokyo University, and the University of Edinburgh. Education: PhD in Computer Science from the University of Edinburgh (advisor: Ewan Klein), postdoctoral research at UMass Amherst (advisor: Andrew McCallum), and research at Tokyo University (advisor: Tsujii Junichi). Research Interests: NLP, machine learning, information extraction, and multimodal models like Gemini. He develops tools such as UCLEED (BioNLP event extractor), frontlets (Scala map wrappers), and thebibbrag (BibTeX to HTML converter). Awards: Allen Distinguished Investigator. Software contributions include GitHub repositories for NLP, machine learning, and data tools. Contact: s.riedel@ucl.ac.uk | Office: 1st Floor, 90 High Holborn, London WC1V 6LJ | Office Hours: Mondays 11 AM–12 PM.
Yi Fang is an Associate Professor of Computer Engineering and an affiliated Associate Professor of Computer Science at New York University Abu Dhabi (NYUAD), and a Global Network Associate Professor at NYU Tandon. He is a core faculty member in the Division of Engineering, specializing in Electrical and Computer Engineering. His research is centered at the intersection of Embodied AI, Robotics, and AI-driven assistive technologies, with strong support from agencies such as the US NSF, UAE ADEK, and ASPIRE. PhD, Purdue University Yi Fang's research interests span 3D Computer Vision, Multimedia Processing, Machine Learning, Deep Learning, and Embodied AI . He focuses on AI-driven perception, learning, and real-world applications, particularly in engineering, medicine, and accessibility. His lab, the Embodied AI and Robotics (AIR) Lab, develops intelligent robotic systems that integrate perception, learning, and decision-making to solve complex societal challenges. His work emphasizes large-scale visual computing, deep visual learning, and cross-domain/multimodal foundation models , with recent innovations in assistive AI for the Deaf and Hard-of-Hearing community. The 15 most recent publications reflect a consistent focus on 3D vision, sketch-based 3D retrieval, point cloud learning, and assistive computer vision . His work leverages deep learning, adversarial training, metric learning, and generative models to bridge modalities such as sketches, depth images, and 3D models. There is a clear trend toward cross-modal understanding, unsupervised representation learning, and real-world assistive applications , especially for visually impaired individuals. Yi Fang actively contributes to the academic community as an Area Chair for top-tier conferences including CVPR, ECCV, ICCV, IJCAI, and IROS. He also serves in peer review and mentoring roles, shaping the future of AI and robotics research. As a dedicated educator, he teaches foundational and advanced courses such as Computer Vision, Applied Machine Learning, Data Structures, and Capstone Design . He mentors students through research seminars and honors projects, fostering innovation and technical excellence. His research is supported by major grants from US NSF, UAE ADEK, and ASPIRE, enabling high-impact interdisciplinary collaborations. He founded and directs the Embodied AI and Robotics (AIR) Lab at NYU Abu Dhabi, a dedicated research space for developing intelligent systems that seamlessly integrate perception, learning, and decision-making. The lab promotes interdisciplinary collaboration across engineering, medicine, and social sciences, advancing the frontiers of Embodied AI.
Georgia Gkioxari is an Assistant Professor in the Division of Computing and Mathematical Sciences at Caltech , with a part-time affiliation at Meta AI . Her work focuses on extending visual perception models through advanced 2D and 3D representation learning, spatial reasoning, and generative models. Education: Not explicitly mentioned in the text Research interests span 3D perception , spatial reasoning , and vision-language integration , with projects like Visual Agentic AI for Spatial Reasoning and Token-by-Token Multimodal Alignment . Her publications emphasize 3D object detection , reconstruction , and generative modeling techniques including diffusion models and transformers . Scientific recognition includes the Meta LLM Evaluation Research Grant , Okawa Research Grant , Google Faculty Scholar Award 2024 , and Amazon Research Award . She teaches courses like Large Language & Vision Models (EE/CS 148) and Learning & 3D (CS 101) at Caltech. Labs & Teams: Leads Glab with members including Ilona Demler, Ziqi Ma, and Damiano Marsili
Professor Winston Hsu is a distinguished faculty member in the Department of Computer Science and Information Engineering at National Taiwan University, where he has served as a full professor since 2015. He is the co-director of the Communications and Multimedia Laboratory (CMLab) and founder of the MiRA (Multimedia indexing, Retrieval, and Analysis) research group. Additionally, he serves as the Founding Director for NVIDIA AI Lab at NTU, the first such lab in Asia. Professor Hsu received his Ph.D. in Electrical Engineering from Columbia University in 2007 under the supervision of Professor Shih-Fu Chang. Prior to his academic career, he was a founding engineer and research manager at CyberLink Corp., now a public image/video software company. National Taiwan University (2007-Present): Professor (2015-Present), Assistant/Associate Professor (2007-2015) MobileDrive (2021-2024): CTO and Vice President (Joint Venture between Foxconn and Stellantis) IBM TJ Watson Research Center (2016-2017): Visiting Scientist Microsoft Research Redmond (2014): Visiting Researcher Columbia University (2007): Ph.D. in Electrical Engineering Professor Hsu's research focuses on machine learning, computer vision, large-scale image and video search and recognition, and embedded AI. His work spans from fundamental research in visual recognition to practical applications in automotive systems, medical imaging, and e-commerce. He has pioneered work in disguised face recognition, low-resolution face hallucination, 3D model search, and virtual try-on systems. His current research emphasizes Embodied AI, integrating perception, action, and learning technologies for applications in automotive and robotics domains. His research group has produced numerous influential publications, particularly in top computer vision and multimedia conferences like CVPR, where they won first place in the Disguised Face Recognition competition in 2018. Their work spans diverse application areas including security, medical diagnostics, automotive systems, and e-commerce solutions, demonstrating strong translation from academic research to real-world impact. IBM Research Pat Goldberg Memorial Best Paper Award (2018) First Place, IARPA Disguised Faces in the Wild Competition (CVPR 2018) Best Brave New Idea Paper Award, ACM Multimedia 2017 NVIDIA AI LAB Award (First in Asia, 2016) First Place, MSR-Bing Image Retrieval Challenge (2013) World's Top 2% Scientists (2023) Professor Hsu actively mentors students and researchers, with his group consistently recruiting PhD students, postdocs, and research assistants. He has successfully bridged academia and industry through multiple collaborations, including his role as CTO at MobileDrive (a joint venture between Foxconn and Stellantis) from 2021-2024. His research has been supported by significant industry partnerships with Microsoft, IBM, and NVIDIA, as well as government grants from Taiwan's Ministry of Science and Technology. His laboratory, the Communications and Multimedia Laboratory (CMLab), maintains strong industry connections and focuses on cutting-edge research in visual AI. The lab has produced numerous award-winning projects and maintains active collaborations with global technology companies, particularly in the automotive and consumer electronics sectors.
Jeannette Bohg is an Assistant Professor of Computer Science at Stanford University, directing the Interactive Perception and Robot Learning Lab. Previously, she was a group leader at the Autonomous Motion Department (AMD) of the MPI for Intelligent Systems (2012-2017). She holds a PhD from KTH Royal Institute of Technology (Stockholm) and degrees from Chalmers University and TU Dresden. Her research focuses on perception, learning, and real-time multi-modal methods for autonomous robotic manipulation and grasping, aiming to bridge principles of human sensorimotor coordination with robotic implementation. Education: PhD in Robotics (KTH), MSc in Art & Technology (Chalmers), Diploma in Computer Science (TU Dresden) Research interests include developing goal-directed, real-time robotic systems capable of meaningful feedback for execution and learning. Key areas are dexterous manipulation, imitation learning, and cross-embodiment policy transfer. Notable contributions include the TidyBot platform and work on force-aware surgical robotics. Awards include the 2019 IEEE ICRA Best Paper Award, 2019 IEEE RA Early Career Award, and 2020 RSS Early Career Award. Her lab explores intersections of robotics, ML, and computer vision. Advising: Actively mentoring students/postdocs in manipulation, perception, and learning. Grants and collaborations span NSF, Stanford AI Lab, and industry partnerships. Future work emphasizes robust real-world deployment and human-robot collaboration. Labs/Teams: Leads the Interactive Perception and Robot Learning Lab, contributing to Stanford’s AI ecosystem. Previously managed the MPI AMD group, fostering interdisciplinary research in autonomous systems.
Nima Mesgarani is an Associate Professor of Electrical Engineering at Columbia Engineering, Columbia University, affiliated with the Sense, Collect and Move Data Committee. His research bridges engineering and neuroscience through reverse-engineering neural signal processing mechanisms, leading to advancements in brain-machine interfaces, neural prosthetics, and speech processing algorithms. He received his PhD in Electrical Engineering from the University of Maryland and completed postdoctoral training at Johns Hopkins University's Center for Language and Speech Processing and UC San Francisco's Neurosurgery Department. Research Focus Professor Mesgarani's lab integrates computational neuroscience and engineering to study acoustic signal processing. Key areas include: Neural decoding of speech and auditory attention in multi-talker environments Development of brain-controlled hearing technologies Novel speech separation and synthesis algorithms inspired by cortical processing Cross-modal learning between auditory and visual systems Applications of large language models in neural signal interpretation Publication Trends Analysis of his 15 most recent articles (2025) reveals dominant themes: neural decoding techniques using intracranial EEG, brain-inspired speech separation models (e.g., Mamba architectures), applications of large language models in auditory neuroscience, cross-modal distillation methods, and clinical translation of audio processing algorithms. A strong emphasis emerges on real-time brain-computer interfaces and noise-robust speech processing. Laboratory and Collaborations Mesgarani directs an interdisciplinary lab developing neurotechnology for hearing restoration. His team collaborates with neurosurgery departments and speech processing centers, focusing on translating theoretical models into clinical brain-machine interfaces. The lab's work has yielded patents for brain-informed speech separation systems and attention-decoding frameworks.
Claus Lamm is a Full Professor of Biological Psychology at the University of Vienna , where he leads the Social, Cognitive and Affective Neuroscience Unit (SCAN-Unit) . He serves as Vice Dean for Research and Advancement of Early Career Researchers at the Faculty of Psychology and holds affiliations with the Vienna Cognitive Science Hub , Environment & Climate Change Hub , and Austrian Academy of Sciences . His academic career spans international collaborations and formative research experience abroad. Scientific Focus: Lamm investigates the neural underpinnings of empathy and prosocial behavior , employing multi-modal approaches combining neuroimaging, psychopharmacology, and psychoneuroendocrinology . His work extends to comparative studies with ravens and dogs, and explores environmental social neuroscience through climate change decision-making research. Recent publications show trends in cross-cultural psychology , machine learning applications , and neurobiological pathways related to social behavior. Awards & Grants: Recipient of the APS Mentor Award for his support of early career researchers. Funded by European Research Council , Austrian Science Fund , Vienna Science and Technology Fund , and intramural grants exceeding €10 million. Key projects include "Unravelling the opioid system in empathy" and "Comparative dog-human fMRI" . Media Engagement: A prominent public science communicator, Lamm has appeared in Nature , Science Magazine , and Austrian media outlets like Ö1 Mittagsjournal and ORF2 , discussing topics from pandemic psychology to social media effects . He maintains active outreach through Science TV and educational programs .
Xiaoming Liu is the Anil K. and Nandita Jain Endowed Professor of Engineering and MSU Foundation Professor in the Department of Computer Science and Engineering at Michigan State University . Holding a Ph.D. from Carnegie Mellon University (2004), he leads cutting-edge research in computer vision and machine learning. Research Interests : Computer Vision Pattern Recognition Image and Video Processing Machine Learning Medical Image Analysis Multimedia Retrieval Recent Research Trends : Focus on 3D object detection and depth estimation Development of robust biometric recognition systems Integration of radar-camera fusion for autonomous systems Advancements in self-supervised and multimodal learning Exploration of adversarial AI security Creation of interpretable forgery detection frameworks Teaching : Spring 2013: CSE891-006 Computer Vision Seminar Fall 2012-2015: CSE803 Computer Vision Spring 2014-2017: CSE 471 Media Processing and Multimedia Contact Information : Email: liuxm@cse.msu.edu Office: EB 3137, Michigan State University Phone: +1 (517) 355-2359
Dr. George Stamou is a Professor at the School of Electrical and Computer Engineering of the National Technical University of Athens (NTUA), serving as Director of the Artificial Intelligence and Learning Systems Laboratory (AILS). His expertise spans knowledge representation, machine learning, neural networks, and semantic technologies. He leads interdisciplinary initiatives such as the postgraduate program 'Data Science and Machine Learning' (2018–2022). Research Interests: Focuses on knowledge graphs, interpretable AI, semantic web applications, and multimodal learning. His work integrates formal logic systems (e.g., description logics) with modern deep learning techniques, addressing challenges in explainability, bias detection, and ethical AI applications. Publications: Over 150 articles in AI journals/conferences with an h-index of 34 (Google Scholar). Notable contributions include datasets like CHORDONOMICON (music analysis), GOSt-MT (gender bias in MT), and methodologies for counterfactual explanations in machine learning. Awards & Committees: Active in W3C and RuleML standardization bodies. Co-organized major AI conferences. Recognized for contributions to semantic interoperability and knowledge-based systems. Labs & Teams: Directs AILS-NTUA lab and collaborates with CISRI (Computer & Information Systems Research Institute). Engages in EU projects like CultureLabs (cultural heritage digitalization) andsmarty4covid (health data analysis).
Yonatan Bisk is an Assistant Professor at Carnegie Mellon University (CMU) in the School of Computer Science , with dual appointments in the Language Technologies Institute and Robotics Institute . His research bridges Natural Language Processing (NLP) with robotics, focusing on grounded and embodied language understanding. Assistant Professor, Language Technologies Institute, CMU (2021–Present) Courtesy Appointment, Robotics Institute, CMU Research Themes : Language as a social codification of embodied experience Interpretable multimodal model training Human-robot collaboration frameworks Embodied question-answering systems Selected Trends : His recent publications show increasing focus on cross-modal attention mechanisms (Vid2Robot), error detection in toolchains (Tools Fail), and theory-of-mind reasoning in language agents (SOTOPIA). Multimodal integration spans vision, audio, and robotic control contexts (ANAVI). Labs & Collaborations : Founder of CLAW Lab (Connecting Language to Action and the World) Collaborations with Microsoft Research, Meta Inc, and CMU's REAL (Robotics, Embodied AI, Learning) community
Kostas Bekris is a Professor in the Department of Computer Science at Rutgers University, specializing in Robotics and Artificial Intelligence. His research focuses on motion planning, autonomous manipulation, and robot control, with notable contributions to tensegrity robotics, perception-driven systems, and large-scale package handling. He leads a team conducting groundbreaking work in robotics, supported by grants from NSF, NASA, and industry collaborators like ExxonMobil. His group emphasizes interdisciplinary approaches, combining machine learning, topological methods, and differentiable physics modeling to advance robot capabilities in complex environments. Education details are not explicitly stated in the provided texts, but his academic career has included significant mentorship of PhD students and postdoctoral researchers. Key projects involve vision-driven manipulation pipelines, obstacle detection systems (PROBE), and resilient robot designs inspired by biological structures. He has been recognized for his work through prestigious awards including the NASA Early Career Grant and multiple NSF grants, as well as team achievements in robotics competitions like the Amazon Picking Challenge. Research interests span robotics subfields such as: Autonomous manipulation in cluttered environments Learning-based control for dynamic systems Topological data analysis for motion reasoning Tensegrity and soft robotics architectures Sim-to-real transfer in robotic tasks His team's work has produced open-source software tools and datasets, advancing benchmarks in manipulation and perception. Recent articles emphasize scalable solutions for industrial automation and robust navigation strategies in unstructured settings. Scientific achievements include: Development of PROBE for proprioceptive obstacle detection Advances in differentiable physics engines for tensegrity systems NSF-funded projects on robotic rearrangement and modular morphologies Advising contributions span over a decade, with current advisees focusing on topics like non-prehensile manipulation and large-scale storage optimization. Collaborations with industry (e.g., ExxonMobil) and academic partners (Yale University) reflect his commitment to applied robotics research. Labs and teams under his leadership include the Rutgers CS Robotics Group, contributing to projects like the ARIAC challenge platform and packing/industrial automation systems. Future work targets improved robot resilience in disaster scenarios and enhanced human-robot collaboration paradigms.
Raquel Fernández is Full Professor of Computational Linguistics and Dialogue Systems at the University of Amsterdam, where she leads the Dialogue Modelling Group at the Institute for Logic, Language & Computation (ILLC). As Vice-Director for Research at ILLC and a Fellow of the ELLIS Society, she bridges computational linguistics, cognitive science, and artificial intelligence through her research on language use in multimodal and conversational contexts. PhD in Computational Linguistics from King's College London Prior research positions at University of Potsdam and Stanford University's CSLI Her work explores how cognitive constraints, social interaction, and perception shape language use, with a focus on: Visually-grounded language processing Multimodal dialogue modeling Model uncertainty and calibration Language grounding in multimodal data Language learning and semantic change Dialogue reference resolution Recent publications analyze multimodal reasoning limitations, cross-lingual knowledge consistency, and uncertainty modeling in dialogue systems. She has received multiple accolades including an ERC Consolidator Grant , NWO VENI/VIDI/Aspasia fellowships , and EMNLP/GenBench awards . Outstanding Paper Award (EMNLP 2023) Best Data Award (GenBench Workshop 2023) ELLIS Society Fellow ERC Consolidator Grant #819455 recipient NWO VENI/VIDI/Aspasia awardee As a leader in academic service, she serves on the SIGDAT Executive Committee and chairs multiple conference committees. Her lab develops models for multimodal dialogue, visual storytelling, and grounded language understanding.
Ranjay Krishna is an Assistant Professor at the Paul G. Allen School of Computer Science & Engineering at the University of Washington, where he co-directs the RAIVN lab and leads the computer vision team at the Allen Institute for AI (Ai2). His research intersects computer vision , natural language processing , robotics , and human-computer interaction . PhD in Computer Science from Stanford University (2021) Bachelor's and Master's degrees from Stanford and Cornell His work has received best paper , outstanding paper , and orals at top conferences like CVPR, ACL, CSCW, NeurIPS, UIST, and ECCV. Media outlets including Science , Forbes , and PBS NOVA have covered his research. He has been supported by grants from Google , Apple , NFS , and others. Ranjay advises a diverse group of 15 PhD and postdoctoral researchers , including Jieyu Zhang, Benlin Liu, and Cheng-Yu Hsieh. His teams have developed benchmarks like MemoryBench and The Colosseum , and his PathFinder framework achieved 74% accuracy in skin melanoma diagnosis—surpassing human experts by 9%. Notable contributions include: Perception Tokens for visual reasoning in MLMs SAM2Act for robotic manipulation with memory Synthetic Visual Genome dataset with 5.6M relationships
Dr. Yunjie Yang is an Associate Professor at the University of Edinburgh's School of Engineering, with affiliations at the Edinburgh Futures Institute (EFI), the Edinburgh Generative AI Laboratory (GAIL), and the Edinburgh Centre for Robotics. He previously held the Chancellor's Fellow in Data Driven Innovation (2018-2023) and Bayes Innovation Fellow (2023-2024) positions. His research focuses on AI-powered sensing and imaging, machine learning, and soft sensors & electronics for robotics. Yang received his PhD in Engineering Electronics from the University of Edinburgh, MSc in Control Science & Engineering from Tsinghua University, and BEng in Measurement & Control Engineering from Anhui University. After his PhD, he worked as a Postdoctoral Research Associate in Chemical Species Tomography before securing his lectureship. His research interests center on developing intelligent sensing systems that replicate human perception capabilities for robotics and intelligent systems. He pioneers flexible sensing and imaging technologies across various scales through innovative multi-modal sensors, soft electronics, and their modeling using machine learning approaches. His work aims to enable autonomous physical artificial intelligence by bridging the gap between robotic systems and human-like perception. Analysis of his recent publications reveals a strong focus on soft robotics perception, particularly through electrical impedance tomography (EIT) and transformer-based architectures. His research spans medical imaging applications, digital twin modeling for industrial processes, and machine learning approaches for sensor data interpretation. The trend shows increasing integration of physics-informed deep learning with traditional tomographic techniques to achieve higher accuracy and efficiency. European Research Council (ERC) Starting Grant (2024) IEEE J. Barry Oakes Advancement Award (2024) IEEE I&M Society Graduate Fellowship Award (2015) Multiple Best Paper Awards Senior Member of IEEE Fellow of the International Society for Industrial Process Tomography Fellow of the Higher Education Academy ESI highly cited papers Dr. Yang serves as Associate Editor for IEEE Transactions on Instrumentation and Measurement and holds editorial positions with Scientific Reports and IEEE Sensors Journal. His research has been licensed to overseas research institutes and industry partners and received wide media coverage including BBC, EFE, USA Today, and STV. He has secured significant grant funding including the prestigious ERC Starting Grant. He leads the Edinburgh SMART Lab (Sensing/imaging + Machine Learning + Robotics), which aims to replicate human perception capabilities for robotics and advance flexible sensing technologies through innovative multi-modal sensors and machine learning approaches. The lab focuses on enabling autonomous physical artificial intelligence with applications spanning medical diagnostics, industrial monitoring, and advanced robotics systems.
Patricio Vela is a Professor at the School of Electrical and Computer Engineering , Georgia Institute of Technology , specializing in geometric perspectives for control theory and computer vision. His research focuses on computer vision integration for semi-autonomous systems, nonlinear control of robotic systems, and biologically inspired mechanics. Education: B.S. (1998) and Ph.D. (2003) from Caltech Research Areas: Autonomy, Robotics, Computer Vision, Control Theory Key Contributions: Geometry-based control systems, visual navigation frameworks, SLAM benchmarking Recent publications highlight advances in vision-based motion planning , 6D pose tracking , and safe navigation policies for autonomous robots. His work bridges geometric mechanics with deep learning for robust perception and control in dynamic environments. Awards: HENAAC Most Promising Engineer (2005) Contact: pvela@gatech.edu | Office: TSRB 441 | Phone: 404.894.8749