David Lindlbauer is an Assistant Professor at the Human-Computer Interaction Institute (HCII) of Carnegie Mellon University, where he leads the Augmented Perception Lab and co-directs the CMU Extended Reality Technology Center. His research bridges human perception, extended reality (AR/VR), and computational interaction techniques, focusing on developing systems that dynamically adapt interface elements based on environmental context, user cognition, and task requirements. He completed his PhD at TU Berlin under Prof. Marc Alexa and held a postdoctoral position at ETH Zurich's Advanced Interactive Technologies lab. His work has been published extensively at top venues including ACM CHI, UIST, and IEEE VR, with research themes spanning gaze tracking, spatial audio optimization, haptic feedback, and multimodal notification systems. Media outlets like MIT Technology Review and Fast Company Design have featured his innovations. Dr. Lindlbauer has received prestigious grants from Meta, NSF, and ETH Zurich, and serves on program committees for CHI, UIST, and ISMAR. He has been recognized with Best Paper awards at ISS 2023 and CHI 2016, and his lab develops tools like MineXR for personalized XR interfaces and RealityReplay for temporal change visualization in mixed reality environments.
Yonatan Bisk is an Assistant Professor at Carnegie Mellon University within the Language Technologies Institute (with courtesy appointment in Robotics Institute). His research bridges Natural Language Processing , Robotics , and Embodied AI , focusing on language grounding, theory of mind, and multimodal interaction. Education : Ph.D. in Computer Science from University of Illinois at Urbana-Champaign Postdoctoral Experience : USC ISI, University of Washington, Allen Institute for AI Industry Appointments : Microsoft Research, Meta AI His research emphasizes embodied language systems and social intelligence in AI . Recent projects include WebArena for autonomous agents, SOTOPIA for social reasoning, and HomeRobot for open-vocabulary manipulation. He leads the REAL Center (Robotics, Embodied AI, and Learning) to foster interdisciplinary collaboration. Key scientific awards include selection for the DARPA ISAT Study Group (2024). He teaches courses like "Talking to Robots" and "Multimodal Machine Learning" while serving as area chair/editor across NLP, Robotics, and ML communities.
Dinesh Jayaraman is an Assistant Professor at the University of Pennsylvania, with primary and secondary appointments in the Department of Computer and Information Science (CIS) and Electrical and Systems Engineering (ESE), respectively. He leads the Perception, Action, and Learning (PennPAL) Research Group at the GRASP Laboratory, focusing on interdisciplinary research at the intersection of robotics, machine learning, and computer vision. Research Interests: Robotics, computer vision, reinforcement learning, and autonomous systems. Recent Publications: His work explores vision-language models for robotic tool use, symmetry-based control acceleration, articulated object modeling, and in-context learning frameworks. Awards: Recipient of the 2022 NSF CAREER Award for innovative contributions to robotics and AI. Teaching: Co-teaching a robot-learning seminar (CIS 7000/ESE 6800) with Antonio Loquercio in Spring 2025. Students: Advising PhD candidates including Edward Hu, Arjun Krishna, and co-advised students with Osbert Bastani, Vijay Kumar, and Rajeev Alur.
Danica Kragic is a Professor of Computer Science at the School of Electrical Engineering and Computer Science at the Royal Institute of Technology (KTH) in Stockholm, Sweden. She serves as the Director of the Centre for Autonomous Systems and leads the Robotics, Perception and Learning Lab at KTH. Her research focuses on advancing robotics capabilities through computer vision and machine learning approaches. MSc in Mechanical Engineering from the Technical University of Rijeka, Croatia (1995) PhD in Computer Science from KTH (2001) Professor Kragic's research primarily centers on robotics, computer vision, and machine learning, with particular emphasis on robotic manipulation, grasp planning, and human-robot interaction. Her work bridges theoretical foundations with practical applications, exploring how robots can understand and interact with objects in complex environments. She investigates how visual and tactile sensing can be integrated to improve robotic perception and manipulation capabilities, with applications ranging from industrial automation to assistive robotics. Her recent publications demonstrate a strong focus on advanced grasp planning techniques, tactile sensing for manipulation, and mathematical representations for robotic control. Kragic's research shows increasing integration of machine learning approaches with traditional robotics frameworks, particularly in the areas of grasp synthesis, object recognition, and human-robot collaboration. Her work spans theoretical contributions in mathematical representations of grasps to practical implementations of robotic systems capable of adapting to novel objects and situations. 2007 IEEE Robotics and Automation Society Early Academic Career Award IEEE Fellow ERC Starting Grant (2012) Member of The Royal Swedish Academy of Sciences Member of The Royal Swedish Academy of Engineering Sciences Honorary Doctorate from Lappeenranta University of Technology Professor Kragic's research has been supported by major funding bodies including the EU, Knut and Alice Wallenberg Foundation, Swedish Foundation for Strategic Research, and Swedish Research Council. While specific student names aren't listed in the provided information, her publication record suggests extensive mentorship of PhD students and postdoctoral researchers in robotics and computer vision. Her lab, the Robotics, Perception and Learning Lab, serves as a hub for interdisciplinary research connecting computer science, engineering, and cognitive science perspectives on robotic systems. As Director of the Centre for Autonomous Systems at KTH, Kragic oversees a major research initiative focused on advancing autonomous technologies. Her Robotics, Perception and Learning Lab brings together researchers working on visual perception, machine learning, and robotic manipulation, with particular emphasis on developing systems that can understand and interact with objects in unstructured environments. The lab's work spans theoretical foundations of robotic manipulation to practical implementations of systems capable of learning from experience.
Sebastian Riedel is a Professor at University College London (UCL) and a Researcher at DeepMind, leading the UCL NLP Lab. His work focuses on teaching machines to read, reason, and write, integrating Natural Language Processing (NLP) with Machine Learning. He holds an Allen Distinguished Investigator award and has held roles at FAIR, UMass Amherst, Tokyo University, and the University of Edinburgh. Education: PhD in Computer Science from the University of Edinburgh (advisor: Ewan Klein), postdoctoral research at UMass Amherst (advisor: Andrew McCallum), and research at Tokyo University (advisor: Tsujii Junichi). Research Interests: NLP, machine learning, information extraction, and multimodal models like Gemini. He develops tools such as UCLEED (BioNLP event extractor), frontlets (Scala map wrappers), and thebibbrag (BibTeX to HTML converter). Awards: Allen Distinguished Investigator. Software contributions include GitHub repositories for NLP, machine learning, and data tools. Contact: s.riedel@ucl.ac.uk | Office: 1st Floor, 90 High Holborn, London WC1V 6LJ | Office Hours: Mondays 11 AM–12 PM.
Yi Fang is an Associate Professor of Computer Engineering and an affiliated Associate Professor of Computer Science at New York University Abu Dhabi (NYUAD), and a Global Network Associate Professor at NYU Tandon. He is a core faculty member in the Division of Engineering, specializing in Electrical and Computer Engineering. His research is centered at the intersection of Embodied AI, Robotics, and AI-driven assistive technologies, with strong support from agencies such as the US NSF, UAE ADEK, and ASPIRE. PhD, Purdue University Yi Fang's research interests span 3D Computer Vision, Multimedia Processing, Machine Learning, Deep Learning, and Embodied AI . He focuses on AI-driven perception, learning, and real-world applications, particularly in engineering, medicine, and accessibility. His lab, the Embodied AI and Robotics (AIR) Lab, develops intelligent robotic systems that integrate perception, learning, and decision-making to solve complex societal challenges. His work emphasizes large-scale visual computing, deep visual learning, and cross-domain/multimodal foundation models , with recent innovations in assistive AI for the Deaf and Hard-of-Hearing community. The 15 most recent publications reflect a consistent focus on 3D vision, sketch-based 3D retrieval, point cloud learning, and assistive computer vision . His work leverages deep learning, adversarial training, metric learning, and generative models to bridge modalities such as sketches, depth images, and 3D models. There is a clear trend toward cross-modal understanding, unsupervised representation learning, and real-world assistive applications , especially for visually impaired individuals. Yi Fang actively contributes to the academic community as an Area Chair for top-tier conferences including CVPR, ECCV, ICCV, IJCAI, and IROS. He also serves in peer review and mentoring roles, shaping the future of AI and robotics research. As a dedicated educator, he teaches foundational and advanced courses such as Computer Vision, Applied Machine Learning, Data Structures, and Capstone Design . He mentors students through research seminars and honors projects, fostering innovation and technical excellence. His research is supported by major grants from US NSF, UAE ADEK, and ASPIRE, enabling high-impact interdisciplinary collaborations. He founded and directs the Embodied AI and Robotics (AIR) Lab at NYU Abu Dhabi, a dedicated research space for developing intelligent systems that seamlessly integrate perception, learning, and decision-making. The lab promotes interdisciplinary collaboration across engineering, medicine, and social sciences, advancing the frontiers of Embodied AI.
Georgia Gkioxari is an Assistant Professor in the Division of Computing and Mathematical Sciences at Caltech , with a part-time affiliation at Meta AI . Her work focuses on extending visual perception models through advanced 2D and 3D representation learning, spatial reasoning, and generative models. Education: Not explicitly mentioned in the text Research interests span 3D perception , spatial reasoning , and vision-language integration , with projects like Visual Agentic AI for Spatial Reasoning and Token-by-Token Multimodal Alignment . Her publications emphasize 3D object detection , reconstruction , and generative modeling techniques including diffusion models and transformers . Scientific recognition includes the Meta LLM Evaluation Research Grant , Okawa Research Grant , Google Faculty Scholar Award 2024 , and Amazon Research Award . She teaches courses like Large Language & Vision Models (EE/CS 148) and Learning & 3D (CS 101) at Caltech. Labs & Teams: Leads Glab with members including Ilona Demler, Ziqi Ma, and Damiano Marsili
Marcia O’Malley is the Thomas Michael Panos Family Professor in Mechanical Engineering, Computer Science, Electrical and Computer Engineering, and Bioengineering at Rice University’s George R. Brown School of Engineering. She chairs the Department of Mechanical Engineering and directs the Mechatronics and Haptic Interfaces (MAHI) Lab. Her research focuses on haptics and robotic rehabilitation, particularly wearable robotic systems for training and rehabilitation in virtual environments. She holds adjunct roles at Baylor College of Medicine and the University of Texas Medical School. Educated at Purdue University (B.S., 1996) and Vanderbilt University (M.S./Ph.D., 1999/2001), Dr. O’Malley has been recognized with prestigious awards, including the ONR Young Investigator Award, NSF CAREER Award, and multiple fellowships. She has twice won Rice’s George R. Brown Award for Superior Teaching. Her work bridges engineering and medicine, addressing human-robot interaction challenges in surgical training, workforce safety, and neurorehabilitation. The MAHI Lab develops devices like the hBracelet and Rice Haptic Rocker to enhance human-robot collaboration. She co-founded Houston Medical Robotics, Inc., applying her innovations to real-world medical applications. Research Interests: Haptics, wearable robotics, neural interfaces, surgical training metrics, and rehabilitation robotics. Labs/Teams: MAHI Lab (Biosciences Research Collaborative), collaborations with medical institutions. Grants/Awards: Extensive funding from NSF, ONR, and industry partnerships; leadership in editorial roles for IEEE Transactions on Haptics.
Prof. Dr. Otmar Hilliges is a Full Professor at the Department of Computer Science at ETH Zurich. He leads the AIT lab and serves as the head of the Institute of Intelligent Interactive Systems. His research focuses on spatio-temporal understanding of human movement and interaction, leveraging algorithms and representations from videos, images, and sensor data for applications in Augmented Reality (AR), Virtual Reality (VR), and Human-Robot Interaction. Education: Diplom (MSc) in Computer Science, Technical University of Munich (TUM), Germany PhD in Computer Science, Ludwig Maximilian University of Munich (LMU), Germany (2009) Research Interests: Hilliges' work spans computer vision, robotics, and human-computer interaction. He develops methods for 3D human pose estimation, generative models for realistic avatar creation, and physically plausible simulation of human-object interactions. His research emphasizes practical applications in AR/VR and assistive robotics, aiming to bridge the gap between perception and action. Grants & Contributions: ERC Consolidator Grant (2022-2027): 'AI-Perceive: Robust Human-Centric Computer Vision for Advanced AI-Agents' Google Research Agreement (2020-2025): 'Generative Modelling of Humans' Microsoft Research Grants: Focus on human-centric robotics and interactive technologies Labs & Teams: Leads the AIT Lab at ETH Zurich, which pioneers research in intelligent interactive systems, emphasizing human-centric AI and robotics. The lab collaborates on projects ranging from drone cinematography to haptic feedback systems.
Professor Winston Hsu is a distinguished faculty member in the Department of Computer Science and Information Engineering at National Taiwan University, where he has served as a full professor since 2015. He is the co-director of the Communications and Multimedia Laboratory (CMLab) and founder of the MiRA (Multimedia indexing, Retrieval, and Analysis) research group. Additionally, he serves as the Founding Director for NVIDIA AI Lab at NTU, the first such lab in Asia. Professor Hsu received his Ph.D. in Electrical Engineering from Columbia University in 2007 under the supervision of Professor Shih-Fu Chang. Prior to his academic career, he was a founding engineer and research manager at CyberLink Corp., now a public image/video software company. National Taiwan University (2007-Present): Professor (2015-Present), Assistant/Associate Professor (2007-2015) MobileDrive (2021-2024): CTO and Vice President (Joint Venture between Foxconn and Stellantis) IBM TJ Watson Research Center (2016-2017): Visiting Scientist Microsoft Research Redmond (2014): Visiting Researcher Columbia University (2007): Ph.D. in Electrical Engineering Professor Hsu's research focuses on machine learning, computer vision, large-scale image and video search and recognition, and embedded AI. His work spans from fundamental research in visual recognition to practical applications in automotive systems, medical imaging, and e-commerce. He has pioneered work in disguised face recognition, low-resolution face hallucination, 3D model search, and virtual try-on systems. His current research emphasizes Embodied AI, integrating perception, action, and learning technologies for applications in automotive and robotics domains. His research group has produced numerous influential publications, particularly in top computer vision and multimedia conferences like CVPR, where they won first place in the Disguised Face Recognition competition in 2018. Their work spans diverse application areas including security, medical diagnostics, automotive systems, and e-commerce solutions, demonstrating strong translation from academic research to real-world impact. IBM Research Pat Goldberg Memorial Best Paper Award (2018) First Place, IARPA Disguised Faces in the Wild Competition (CVPR 2018) Best Brave New Idea Paper Award, ACM Multimedia 2017 NVIDIA AI LAB Award (First in Asia, 2016) First Place, MSR-Bing Image Retrieval Challenge (2013) World's Top 2% Scientists (2023) Professor Hsu actively mentors students and researchers, with his group consistently recruiting PhD students, postdocs, and research assistants. He has successfully bridged academia and industry through multiple collaborations, including his role as CTO at MobileDrive (a joint venture between Foxconn and Stellantis) from 2021-2024. His research has been supported by significant industry partnerships with Microsoft, IBM, and NVIDIA, as well as government grants from Taiwan's Ministry of Science and Technology. His laboratory, the Communications and Multimedia Laboratory (CMLab), maintains strong industry connections and focuses on cutting-edge research in visual AI. The lab has produced numerous award-winning projects and maintains active collaborations with global technology companies, particularly in the automotive and consumer electronics sectors.
Mark Yatskar is an Assistant Professor in the Department of Computer and Information Science at the University of Pennsylvania. His research focuses on the intersection of natural language processing, computer vision, and fairness in machine learning. He earned his PhD from the University of Washington under advisors Luke Zettlemoyer and Ali Farhadi, and previously worked as a Young Investigator at the Allen Institute for Artificial Intelligence. Education: PhD in Computer Science, University of Washington (Advisor: Luke Zettlemoyer & Ali Farhadi) Research Interests: Yatskar's work explores how language can structure visual perception and mitigate human biases in machine learning systems. Key themes include: Natural language as a scaffold for visual intelligence Bias characterization and control in machine learning systems His lab currently investigates projects like language-guided bottlenecks, annotator cognitive heuristics, and gender bias amplification. Teaching: CIS 5300: Computational Linguistics (2021-2024) CIS 7000: Language and Vision (2020) CIS 6300: Efficient NLP (2023, 2025) Awards: Best Paper Award at EMNLP (Gender Bias Amplification Research) Advising & Grants: Yatskar advises a team of PhD/Master's students and actively seeks motivated researchers. His group has explored funding in areas like interpretable AI, multimodal reasoning, and dataset bias mitigation. Labs/Teams: Leads the Penn NLP & Vision Lab, focusing on projects like MolMo/PixMo open models, ViUniT visual unit tests, and bias mitigation frameworks.
Dr. Heather Inwood serves as University Associate Professor in Modern Chinese Literature and Culture within the Faculty of Asian and Middle Eastern Studies at the University of Cambridge. She is also a Fellow and Director of Studies at Trinity Hall, maintaining dual institutional affiliations that support her teaching and research activities across undergraduate and postgraduate programs. Her educational background includes a BA in Chinese Studies from the University of Cambridge (Trinity Hall), followed by advanced studies at Tsinghua University and Peking University, culminating in a PhD in modern Chinese literature from SOAS, University of London in 2008. Prior academic appointments include Assistant Professor at The Ohio State University (2008-2013) and Lecturer in Chinese Cultural Studies at the University of Manchester before returning to Cambridge in 2016. Professor Inwood's research program investigates the dynamic interactions between digital media and literary production in contemporary China, with particular focus on poetry communities, genre fiction, and transmedia storytelling. Her work bridges traditional literary analysis with digital humanities approaches, examining how online platforms transform authorship, reception, and cultural value. She has pioneered scholarship on internet poetry scenes, demonstrating their vitality despite public perceptions of poetry's decline in modern China. Her publication trajectory reveals a clear evolution from early work on poetry communities ( Verse Going Viral: China's New Media Scenes , 2014) toward broader investigations of digital narrative forms, game studies, and Sinophone cyberspace. Recent publications increasingly engage with science fiction, transmedia storytelling, and the algorithmic structures shaping literary production, reflecting both the changing digital landscape and her expanding methodological toolkit. Professor Inwood actively supervises five PhD students working on diverse projects spanning Misty poetry, queer narratives, gender studies, and poetic animality, demonstrating her commitment to nurturing next-generation scholarship in Chinese literary studies. Her research has been supported by British Academy Small Grants and similar funding mechanisms that enable sustained fieldwork and publication. Her teaching responsibilities include undergraduate courses on East Asian media and popular culture, Modern Chinese texts, and Modern Chinese literature, connecting classroom instruction directly to her research expertise. She also maintains public engagement through Chinese-language columns for BBC China and other media outlets, bridging academic and public discourse on contemporary Chinese culture.
Jeannette Bohg is an Assistant Professor of Computer Science at Stanford University, directing the Interactive Perception and Robot Learning Lab. Previously, she was a group leader at the Autonomous Motion Department (AMD) of the MPI for Intelligent Systems (2012-2017). She holds a PhD from KTH Royal Institute of Technology (Stockholm) and degrees from Chalmers University and TU Dresden. Her research focuses on perception, learning, and real-time multi-modal methods for autonomous robotic manipulation and grasping, aiming to bridge principles of human sensorimotor coordination with robotic implementation. Education: PhD in Robotics (KTH), MSc in Art & Technology (Chalmers), Diploma in Computer Science (TU Dresden) Research interests include developing goal-directed, real-time robotic systems capable of meaningful feedback for execution and learning. Key areas are dexterous manipulation, imitation learning, and cross-embodiment policy transfer. Notable contributions include the TidyBot platform and work on force-aware surgical robotics. Awards include the 2019 IEEE ICRA Best Paper Award, 2019 IEEE RA Early Career Award, and 2020 RSS Early Career Award. Her lab explores intersections of robotics, ML, and computer vision. Advising: Actively mentoring students/postdocs in manipulation, perception, and learning. Grants and collaborations span NSF, Stanford AI Lab, and industry partnerships. Future work emphasizes robust real-world deployment and human-robot collaboration. Labs/Teams: Leads the Interactive Perception and Robot Learning Lab, contributing to Stanford’s AI ecosystem. Previously managed the MPI AMD group, fostering interdisciplinary research in autonomous systems.
Nima Mesgarani is an Associate Professor of Electrical Engineering at Columbia Engineering, Columbia University, affiliated with the Sense, Collect and Move Data Committee. His research bridges engineering and neuroscience through reverse-engineering neural signal processing mechanisms, leading to advancements in brain-machine interfaces, neural prosthetics, and speech processing algorithms. He received his PhD in Electrical Engineering from the University of Maryland and completed postdoctoral training at Johns Hopkins University's Center for Language and Speech Processing and UC San Francisco's Neurosurgery Department. Research Focus Professor Mesgarani's lab integrates computational neuroscience and engineering to study acoustic signal processing. Key areas include: Neural decoding of speech and auditory attention in multi-talker environments Development of brain-controlled hearing technologies Novel speech separation and synthesis algorithms inspired by cortical processing Cross-modal learning between auditory and visual systems Applications of large language models in neural signal interpretation Publication Trends Analysis of his 15 most recent articles (2025) reveals dominant themes: neural decoding techniques using intracranial EEG, brain-inspired speech separation models (e.g., Mamba architectures), applications of large language models in auditory neuroscience, cross-modal distillation methods, and clinical translation of audio processing algorithms. A strong emphasis emerges on real-time brain-computer interfaces and noise-robust speech processing. Laboratory and Collaborations Mesgarani directs an interdisciplinary lab developing neurotechnology for hearing restoration. His team collaborates with neurosurgery departments and speech processing centers, focusing on translating theoretical models into clinical brain-machine interfaces. The lab's work has yielded patents for brain-informed speech separation systems and attention-decoding frameworks.
Professor Dorit Abusch is a faculty member in the Department of Linguistics and Philosophy at Cornell University's College of Arts & Sciences. Her research focuses on semantics, pragmatics, and their applications to visual narratives. She explores topics like tense semantics, presupposition triggering, modal logic, and the interplay between language and visual media. Current work extends linguistic methodologies to analyze art forms such as comics, cave paintings, and temple sculptures. Her research interests include formal semantics applied to visual narratives, dynamic semantics, possible world theory, and multimodal discourse representation. She investigates how visual elements like sequential art and pictorial sequences convey temporal progression, aspectual distinctions, and free perception constructions through semiotic frameworks. Recent presentations include talks on applying semantics to film and picturebooks at institutions like MIT and the University of Padua. Her publications emphasize cross-media analysis, with key works published in Linguistics & Philosophy and Sinn und Bedeutung . Abusch has received grants for projects studying visual narratives in Indian art and wall paintings of Rajasthan. These include a 2012-2013 Humanities Research Grant and a Cornell Institute for Social Sciences award. Her work bridges linguistics with philosophy and visual studies, offering innovative frameworks for understanding non-linguistic communication through formal semantic tools.