David Lindlbauer is an Assistant Professor at the Human-Computer Interaction Institute (HCII) of Carnegie Mellon University, where he leads the Augmented Perception Lab and co-directs the CMU Extended Reality Technology Center. His research bridges human perception, extended reality (AR/VR), and computational interaction techniques, focusing on developing systems that dynamically adapt interface elements based on environmental context, user cognition, and task requirements. He completed his PhD at TU Berlin under Prof. Marc Alexa and held a postdoctoral position at ETH Zurich's Advanced Interactive Technologies lab. His work has been published extensively at top venues including ACM CHI, UIST, and IEEE VR, with research themes spanning gaze tracking, spatial audio optimization, haptic feedback, and multimodal notification systems. Media outlets like MIT Technology Review and Fast Company Design have featured his innovations. Dr. Lindlbauer has received prestigious grants from Meta, NSF, and ETH Zurich, and serves on program committees for CHI, UIST, and ISMAR. He has been recognized with Best Paper awards at ISS 2023 and CHI 2016, and his lab develops tools like MineXR for personalized XR interfaces and RealityReplay for temporal change visualization in mixed reality environments.
Yonatan Bisk is an Assistant Professor at Carnegie Mellon University within the Language Technologies Institute (with courtesy appointment in Robotics Institute). His research bridges Natural Language Processing , Robotics , and Embodied AI , focusing on language grounding, theory of mind, and multimodal interaction. Education : Ph.D. in Computer Science from University of Illinois at Urbana-Champaign Postdoctoral Experience : USC ISI, University of Washington, Allen Institute for AI Industry Appointments : Microsoft Research, Meta AI His research emphasizes embodied language systems and social intelligence in AI . Recent projects include WebArena for autonomous agents, SOTOPIA for social reasoning, and HomeRobot for open-vocabulary manipulation. He leads the REAL Center (Robotics, Embodied AI, and Learning) to foster interdisciplinary collaboration. Key scientific awards include selection for the DARPA ISAT Study Group (2024). He teaches courses like "Talking to Robots" and "Multimodal Machine Learning" while serving as area chair/editor across NLP, Robotics, and ML communities.
Georgia Gkioxari is an Assistant Professor in the Division of Computing and Mathematical Sciences at Caltech , with a part-time affiliation at Meta AI . Her work focuses on extending visual perception models through advanced 2D and 3D representation learning, spatial reasoning, and generative models. Education: Not explicitly mentioned in the text Research interests span 3D perception , spatial reasoning , and vision-language integration , with projects like Visual Agentic AI for Spatial Reasoning and Token-by-Token Multimodal Alignment . Her publications emphasize 3D object detection , reconstruction , and generative modeling techniques including diffusion models and transformers . Scientific recognition includes the Meta LLM Evaluation Research Grant , Okawa Research Grant , Google Faculty Scholar Award 2024 , and Amazon Research Award . She teaches courses like Large Language & Vision Models (EE/CS 148) and Learning & 3D (CS 101) at Caltech. Labs & Teams: Leads Glab with members including Ilona Demler, Ziqi Ma, and Damiano Marsili
Professor Winston Hsu is a distinguished faculty member in the Department of Computer Science and Information Engineering at National Taiwan University, where he has served as a full professor since 2015. He is the co-director of the Communications and Multimedia Laboratory (CMLab) and founder of the MiRA (Multimedia indexing, Retrieval, and Analysis) research group. Additionally, he serves as the Founding Director for NVIDIA AI Lab at NTU, the first such lab in Asia. Professor Hsu received his Ph.D. in Electrical Engineering from Columbia University in 2007 under the supervision of Professor Shih-Fu Chang. Prior to his academic career, he was a founding engineer and research manager at CyberLink Corp., now a public image/video software company. National Taiwan University (2007-Present): Professor (2015-Present), Assistant/Associate Professor (2007-2015) MobileDrive (2021-2024): CTO and Vice President (Joint Venture between Foxconn and Stellantis) IBM TJ Watson Research Center (2016-2017): Visiting Scientist Microsoft Research Redmond (2014): Visiting Researcher Columbia University (2007): Ph.D. in Electrical Engineering Professor Hsu's research focuses on machine learning, computer vision, large-scale image and video search and recognition, and embedded AI. His work spans from fundamental research in visual recognition to practical applications in automotive systems, medical imaging, and e-commerce. He has pioneered work in disguised face recognition, low-resolution face hallucination, 3D model search, and virtual try-on systems. His current research emphasizes Embodied AI, integrating perception, action, and learning technologies for applications in automotive and robotics domains. His research group has produced numerous influential publications, particularly in top computer vision and multimedia conferences like CVPR, where they won first place in the Disguised Face Recognition competition in 2018. Their work spans diverse application areas including security, medical diagnostics, automotive systems, and e-commerce solutions, demonstrating strong translation from academic research to real-world impact. IBM Research Pat Goldberg Memorial Best Paper Award (2018) First Place, IARPA Disguised Faces in the Wild Competition (CVPR 2018) Best Brave New Idea Paper Award, ACM Multimedia 2017 NVIDIA AI LAB Award (First in Asia, 2016) First Place, MSR-Bing Image Retrieval Challenge (2013) World's Top 2% Scientists (2023) Professor Hsu actively mentors students and researchers, with his group consistently recruiting PhD students, postdocs, and research assistants. He has successfully bridged academia and industry through multiple collaborations, including his role as CTO at MobileDrive (a joint venture between Foxconn and Stellantis) from 2021-2024. His research has been supported by significant industry partnerships with Microsoft, IBM, and NVIDIA, as well as government grants from Taiwan's Ministry of Science and Technology. His laboratory, the Communications and Multimedia Laboratory (CMLab), maintains strong industry connections and focuses on cutting-edge research in visual AI. The lab has produced numerous award-winning projects and maintains active collaborations with global technology companies, particularly in the automotive and consumer electronics sectors.
Sanjay Purushotham is an Assistant Professor in the Department of Information Systems at the University of Maryland Baltimore County (UMBC), with a PhD in Electrical Engineering from the University of Southern California (USC) and a postdoctoral background in Computer Science at USC's Integrated Media Systems Center (IMSC). His research focuses on machine learning, data mining, and their applications in biomedical informatics, social network analysis, and multimedia data mining. Key contributions include survival analysis models using pseudo values and federated learning frameworks for healthcare data. He has received awards including the Best Paper Award at SIGSPATIAL 2014 and a Best Poster Runnerup at SCMLS 2016. Education: PhD in Electrical Engineering (USC), Postdoc in Computer Science (USC) His work spans interdisciplinary areas such as domain adaptation for remote sensing, thermal face translation, and interpretable neural networks for medical applications. Recent projects include federated survival analysis models and climate-informatics frameworks for cloud property retrieval. He teaches courses in artificial intelligence, healthcare informatics, and statistical learning at UMBC. Research highlights include developing MedFuseNet for multimodal medical question answering and VDAM for multi-sensor cloud data analysis. His work on fair survival analysis models addresses algorithmic bias in healthcare predictions. Current grants include a NSF CAREER award for trustworthy federated learning in computational healthcare.
John MacFarlane is a Professor of Philosophy at the University of California, Berkeley , with affiliations to the Group in Logic and the Methodology of Science and co-organizer roles in the Meaning Sciences Club and Townsend Center Working Group . His research spans philosophy of language, philosophical logic, metaphysics, epistemology, philosophy of mathematics, history of logic , and ancient philosophy , particularly Aristotle . He has held visiting positions at institutions like École des Hautes Études en Sciences Sociales (EHESS), Paris . Education : Ph.D. in Philosophy (2000, University of Pittsburgh), M.A. in Classics (1997) and Philosophy (1994), A.B. summa cum laude in Philosophy (1991, Harvard). Research Interests focus on semantic relativism , particularly assessment sensitivity as explored in his book Assessment Sensitivity: Relative Truth and Its Applications (Oxford, 2014). He investigates how epistemic modals and vagueness interact with truth evaluation. His work in philosophical logic includes a textbook Philosophical Logic: A Contemporary Introduction (Routledge, 2021). Secondary interests include ancient philosophy (Aristotle, Plato), formal logic , and fiddlosophy . Article Trends reveal contributions to relativism , epistemic modals , vagueness , and logicism . His recent work (2025) examines Plato’s Protagoras and belief utility , while 2022–2020 articles address semantic indecision and probabilistic knowledge . Earlier works (2008–2000) include foundational studies on Frege, Kant, and Aristotle . Scientific Awards include a 2023 Residential Fellowship at Wissenschaftskolleg, Berlin (declined), 2016–17 Fellow at IEA Paris , and 2015 American Academy of Arts and Sciences membership . He has received multiple Humanities Research Fellowships at Berkeley (2016, 2008, 2003) and Research Enabling Grants (2000–2009). Advising includes supervising 15+ PhD students in areas like logic, epistemology, and philosophy of language . Notable advisees include Richard Lawrence (2017) and Sophie Dandelet (2021). He also guided honors theses at Berkeley, such as Benjamin Logan’s work on epistemic modals (2018). Labs and Teams involve co-organizing the Townsend Center Working Group in History and Philosophy of Logic, Mathematics, and Science and contributing to the Group in Logic and Methodology of Science . His interdisciplinary approach integrates philosophy, logic, and computational tools , including creating Pandoc for document conversion.
Nima Mesgarani is an Associate Professor of Electrical Engineering at Columbia Engineering, Columbia University, affiliated with the Sense, Collect and Move Data Committee. His research bridges engineering and neuroscience through reverse-engineering neural signal processing mechanisms, leading to advancements in brain-machine interfaces, neural prosthetics, and speech processing algorithms. He received his PhD in Electrical Engineering from the University of Maryland and completed postdoctoral training at Johns Hopkins University's Center for Language and Speech Processing and UC San Francisco's Neurosurgery Department. Research Focus Professor Mesgarani's lab integrates computational neuroscience and engineering to study acoustic signal processing. Key areas include: Neural decoding of speech and auditory attention in multi-talker environments Development of brain-controlled hearing technologies Novel speech separation and synthesis algorithms inspired by cortical processing Cross-modal learning between auditory and visual systems Applications of large language models in neural signal interpretation Publication Trends Analysis of his 15 most recent articles (2025) reveals dominant themes: neural decoding techniques using intracranial EEG, brain-inspired speech separation models (e.g., Mamba architectures), applications of large language models in auditory neuroscience, cross-modal distillation methods, and clinical translation of audio processing algorithms. A strong emphasis emerges on real-time brain-computer interfaces and noise-robust speech processing. Laboratory and Collaborations Mesgarani directs an interdisciplinary lab developing neurotechnology for hearing restoration. His team collaborates with neurosurgery departments and speech processing centers, focusing on translating theoretical models into clinical brain-machine interfaces. The lab's work has yielded patents for brain-informed speech separation systems and attention-decoding frameworks.
Professor Dorit Abusch is a faculty member in the Department of Linguistics and Philosophy at Cornell University's College of Arts & Sciences. Her research focuses on semantics, pragmatics, and their applications to visual narratives. She explores topics like tense semantics, presupposition triggering, modal logic, and the interplay between language and visual media. Current work extends linguistic methodologies to analyze art forms such as comics, cave paintings, and temple sculptures. Her research interests include formal semantics applied to visual narratives, dynamic semantics, possible world theory, and multimodal discourse representation. She investigates how visual elements like sequential art and pictorial sequences convey temporal progression, aspectual distinctions, and free perception constructions through semiotic frameworks. Recent presentations include talks on applying semantics to film and picturebooks at institutions like MIT and the University of Padua. Her publications emphasize cross-media analysis, with key works published in Linguistics & Philosophy and Sinn und Bedeutung . Abusch has received grants for projects studying visual narratives in Indian art and wall paintings of Rajasthan. These include a 2012-2013 Humanities Research Grant and a Cornell Institute for Social Sciences award. Her work bridges linguistics with philosophy and visual studies, offering innovative frameworks for understanding non-linguistic communication through formal semantic tools.
Professor Gabriel Brostow is a faculty member in the Department of Computer Science at University College London (UCL), where he leads research in Computer Vision and Human-Computer Interaction. He also serves as Chief Research Scientist and Senior Director of the R&D Team at Niantic, the company behind Pokémon GO. His work bridges academic research and industry applications, focusing on developing AI systems that enhance human capabilities through what he terms 'Human in the Loop AI'—now commonly referred to as Human-Centered AI. Brostow completed his BS in Electrical Engineering at UT Austin, followed by a PhD with Irfan Essa at Georgia Tech. He then pursued postdoctoral research with Roberto Cipolla's Computer Vision & Robotics Group at Cambridge University as a Marshall Sherfield Fellow, and with Marc Pollefeys in ETH Zurich's CVG Group. His research explores how AI, particularly Computer Vision, can serve as 'super-tools' for professionals across various domains including filmmaking, architecture, robotics, and scientific research. Specific interests include assistive technology for everyday life, authoring systems that maximize user effort, 3D reconstruction, depth estimation, and vision-language models. His work often involves creating systems that are validated through real-world human interaction to ensure practical utility. Analysis of his recent publications reveals a strong focus on practical applications of Computer Vision that directly interact with humans. His research spans 3D scene understanding, depth estimation, sketch-based interfaces, and multimodal AI systems. There's a clear emphasis on creating benchmarks and tools that facilitate human-AI collaboration, with applications in assistive technology, urban planning, filmmaking, and biodiversity monitoring. His work frequently appears at top conferences including CVPR, NeurIPS, ECCV, and CHI. Marshall Sherfield Fellowship Brostow actively mentors PhD students, with current advisees including Ross Murphy, Skanda Koppula, Gizem Unlu, Omiros Pantazis, and Jamie Watson. His alumni include numerous PhD graduates and MSc students who have gone on to successful careers in academia and industry. He emphasizes selecting students based on passion and potential rather than just academic credentials, valuing traits like helpfulness, drive, and hunger to learn. His research is supported through collaborations with major institutions and companies including DeepMind, MIT, and the University of Edinburgh. He leads a research group at UCL that collaborates closely with Niantic's R&D team, creating a unique bridge between academic research and industry application. His team's work frequently involves developing novel Computer Vision techniques that are validated through real-world human interaction, ensuring practical utility alongside technical innovation. The group explores blue-sky research problems with applications ranging from assistive technology to professional tools for filmmakers, architects, and scientists studying diverse environments.
Takako Fujioka is an Associate Professor of Music at Stanford University, affiliated with the Center for Computer Research in Music and Acoustics (CCRMA). Her research focuses on the neural mechanisms underlying auditory perception, auditory-motor coupling, and music-supported therapy for neurorehabilitation. She holds a Ph.D. in Physiology from the Graduate University for Advanced Studies, Japan, and M.Sc./B.Eng. degrees in Electrical Engineering from Waseda University. Her work combines neurophysiological techniques such as MEG and EEG to study brain plasticity in development, aging, and stroke recovery. Notable contributions include investigating how music influences motor and cognitive recovery in stroke patients, as well as exploring the neural basis of musical perception through rhythmic synchronization and pitch discrimination studies. Supported by awards from the Canadian Institutes of Health Research during her postdoctoral work at the Rotman Research Institute, her research bridges clinical neuroscience and music cognition. Dr. Fujioka’s expertise spans auditory neuroscience, neurorehabilitation, and technology-assisted music therapy. She has pioneered studies on tactile mapping for cochlear implant users and networked music performance systems, emphasizing cross-modal perception and human-technology interaction. Her findings contribute to both theoretical understanding of auditory processing and practical applications in medical and educational settings. Awards: Canadian Institutes of Health Research Awards (postdoctoral phase) Labs/Teams: CCRMA, Stanford Music Perception Laboratory, Rotman Research Institute collaborations Key Themes: Neuroplasticity, Music-Mediated Rehabilitation, Auditory-Motor Integration, Multisensory Processing Her recent work examines aging-related changes in binaural hearing and the role of beta/gamma oscillations in rhythmic processing. She advocates for translational research that connects neural mechanisms with real-world therapeutic interventions.
Rachel Rudinger is an Assistant Professor at the University of Maryland, affiliated with the Department of Computer Science and the University of Maryland Institute for Advanced Computer Studies (UMIACS). Her research focuses on Natural Language Processing (NLP), Machine Learning, and AI ethics, particularly addressing sociocultural biases and fairness in large language models (LLMs). She holds a PhD from Johns Hopkins University (2019) and a B.S. from Yale University (2013). Rudinger's work explores equitable cultural alignment in AI systems, common ground misalignment in dialog systems, and the mutual influence of gender and occupation in LLMs. She received the NSF CAREER Award in 2024 for her project on robust, fair, and culturally aware commonsense reasoning. Her recent publications investigate empathy gaps in LLMs, synthetic data effectiveness in disaster response, and bias measurement techniques across domains. As an advisor, she guides seven PhD students including Christabel Acquaye and Haozhe An. Her research spans diverse topics from legal language analysis to maternal health question answering, reflecting her commitment to interdisciplinary AI ethics. She actively contributes to workshops on commonsense representation and serves as a reviewer for top conferences in NLP and AI.
Raquel Fernández is Full Professor of Computational Linguistics and Dialogue Systems at the University of Amsterdam, where she leads the Dialogue Modelling Group at the Institute for Logic, Language & Computation (ILLC). As Vice-Director for Research at ILLC and a Fellow of the ELLIS Society, she bridges computational linguistics, cognitive science, and artificial intelligence through her research on language use in multimodal and conversational contexts. PhD in Computational Linguistics from King's College London Prior research positions at University of Potsdam and Stanford University's CSLI Her work explores how cognitive constraints, social interaction, and perception shape language use, with a focus on: Visually-grounded language processing Multimodal dialogue modeling Model uncertainty and calibration Language grounding in multimodal data Language learning and semantic change Dialogue reference resolution Recent publications analyze multimodal reasoning limitations, cross-lingual knowledge consistency, and uncertainty modeling in dialogue systems. She has received multiple accolades including an ERC Consolidator Grant , NWO VENI/VIDI/Aspasia fellowships , and EMNLP/GenBench awards . Outstanding Paper Award (EMNLP 2023) Best Data Award (GenBench Workshop 2023) ELLIS Society Fellow ERC Consolidator Grant #819455 recipient NWO VENI/VIDI/Aspasia awardee As a leader in academic service, she serves on the SIGDAT Executive Committee and chairs multiple conference committees. Her lab develops models for multimodal dialogue, visual storytelling, and grounded language understanding.
Dr. Sirui Li is a Lecturer at Murdoch University's School of Information Technology within the College of Science, Technology, Engineering and Mathematics. Her research focuses on Artificial Intelligence, Natural Language Processing (NLP), Machine Learning, Knowledge Graphs, Data Analysis, Temporal Data, and Multi-modal Models, with applications in medicine, agriculture, and mining. She collaborates with industry partners like BHP and has published in journals such as Food Chemistry and Knowledge and Information Systems , as well as conferences like ICSME and IJCNN. Education: Bachelor of Advanced Computing (Honours) in Computer Science at Australian National University Master of Computing (Specialising in AI) at ANU Ph.D. in Information Technology (AI) at Murdoch University Research interests include interdisciplinary applications of AI, such as clinical coding privacy solutions, disease spread modeling, and drug repurposing for pandemics. Her work emphasizes practical industry integration, demonstrated through awards like the 2024 EMNLP Best Demo Award and the 2023 Iron Ore Circuit Hackathon innovation prize. Professional roles include IEEE Western Australia Section committee membership, conference chair positions, and peer review for top journals. She actively mentors students pursuing Honours, Master's, or PhD projects in her areas of expertise.
Greg Keoleian is the Peter M. Wege Endowed Professor of Sustainable Systems at the University of Michigan. He co-founded and directs the Center for Sustainable Systems and co-leads the MI Hydrogen initiative. His research focuses on advancing life cycle assessment (LCA) methodologies to evaluate the sustainability of technologies, products, and systems. Key areas include renewable energy systems (e.g., wind, solar, bioenergy), transportation decarbonization, circular economy frameworks, and food/agricultural systems. He teaches interdisciplinary graduate courses on Sustainable Energy Systems and Industrial Ecology, and co-directs the Engineering Sustainable Systems Dual Degree Program and the Rackham Graduate Certificate Program in Industry Ecology. Research interests emphasize integrating environmental, economic, and social metrics into sustainability assessments. Notable contributions include pioneering LCA models for hydrogen energy systems, vehicle lifecycle optimization, and circular economy pathways for automotive materials. His work spans global and regional scales, addressing challenges such as reducing greenhouse gas emissions, optimizing resource efficiency, and promoting sustainable urban infrastructure. Keoleian’s recent work highlights synergies between emerging technologies (e.g., wireless charging, autonomous vehicles) and sustainability goals. He collaborates with policymakers and industry leaders to translate research into actionable strategies for deep decarbonization and circular economy adoption. Initiatives like the State of Michigan Hydrogen Roadmap exemplify his focus on bridging academic research with practical implementation. His interdisciplinary approach ensures comprehensive analysis of energy systems, industrial processes, and consumer behaviors. Keoleian’s teaching and mentorship emphasize systems thinking and cross-sector collaboration. He has shaped curricula that prepare students for leadership roles in sustainable systems design, policy development, and corporate sustainability management. Ongoing projects include optimizing reusable packaging systems, evaluating food waste impacts, and modeling urban energy justice in electric vehicle deployment.
John Kingston is a Professor of Linguistics and Director of the Phonetics Lab at the University of Massachusetts Amherst, where he has been since 1990. He holds a BA and MA from the University of Chicago (1976–1977) and a PhD from UC Berkeley (1985). His research focuses on the interplay between phonetics and phonology, particularly speech perception and its influence on phonological representations. He co-founded the Laboratory Phonology Conference series in 1987 and has conducted fieldwork on Otomanguean languages. His work emphasizes experimental methods to study phonological questions, including studies on vowel perception, tone systems, and cross-linguistic phonetic patterns. Kingston’s academic journey includes roles at the University of Texas, Austin (1984–1986) and Cornell University (1986–1990). His research explores how auditory processing and linguistic knowledge shape speech perception, with notable contributions to understanding tonogenesis, perceptual contrast effects, and vowel category learning in second languages. He collaborates on grants examining Ganong effects and phonological inventories, advocating for theories that bridge perceptual and structural aspects of language. His lab, the Phonetics Lab, supports experimental work on speech perception and production. Kingston is also the Honors Program Coordinator, mentoring students in linguistics and related fields. Despite no explicit awards listed, his extensive publications and conference leadership reflect his scholarly impact.