Yonatan Bisk is an Assistant Professor at Carnegie Mellon University within the Language Technologies Institute (with courtesy appointment in Robotics Institute). His research bridges Natural Language Processing , Robotics , and Embodied AI , focusing on language grounding, theory of mind, and multimodal interaction. Education : Ph.D. in Computer Science from University of Illinois at Urbana-Champaign Postdoctoral Experience : USC ISI, University of Washington, Allen Institute for AI Industry Appointments : Microsoft Research, Meta AI His research emphasizes embodied language systems and social intelligence in AI . Recent projects include WebArena for autonomous agents, SOTOPIA for social reasoning, and HomeRobot for open-vocabulary manipulation. He leads the REAL Center (Robotics, Embodied AI, and Learning) to foster interdisciplinary collaboration. Key scientific awards include selection for the DARPA ISAT Study Group (2024). He teaches courses like "Talking to Robots" and "Multimodal Machine Learning" while serving as area chair/editor across NLP, Robotics, and ML communities.
Joan Bruna is a Full Professor of Computer Science, Data Science, and affiliated Mathematics at New York University's Courant Institute and Center for Data Science. He leads the CILVR group and co-founded the MaD group. His research focuses on mathematical foundations of machine learning, deep learning, signal processing, and their applications in computational science, climate modeling, and geophysics. He holds a Ph.D. in Applied Mathematics from École Polytechnique and has held roles at UC Berkeley, the Institute for Advanced Study, and the Flatiron Institute. Education: Ph.D. (2013) and M.Sc. (2005) in Applied Mathematics, École Polytechnique and ENS Cachan; dual B.Sc./M.Sc. in Telecommunications and Mathematics from UPC Barcelona (2002–2004). Awards include the NSF CAREER Award (2019) and Alfred P. Sloan Fellowship (2018). Research interests span theoretical aspects of deep learning, optimization, and statistical methods, with applications to climate science and inverse problems. He has advised numerous PhD students and postdocs, contributing to seminal work on scattering networks, geometric deep learning, and neural ODEs. Professional service includes editorial roles at TMLR, JMLR, and TPAMI, plus program chairs for major conferences. His work bridges mathematics and machine learning, emphasizing rigorous analysis and real-world impact.
Georgia Gkioxari is an Assistant Professor in the Division of Computing and Mathematical Sciences at Caltech , with a part-time affiliation at Meta AI . Her work focuses on extending visual perception models through advanced 2D and 3D representation learning, spatial reasoning, and generative models. Education: Not explicitly mentioned in the text Research interests span 3D perception , spatial reasoning , and vision-language integration , with projects like Visual Agentic AI for Spatial Reasoning and Token-by-Token Multimodal Alignment . Her publications emphasize 3D object detection , reconstruction , and generative modeling techniques including diffusion models and transformers . Scientific recognition includes the Meta LLM Evaluation Research Grant , Okawa Research Grant , Google Faculty Scholar Award 2024 , and Amazon Research Award . She teaches courses like Large Language & Vision Models (EE/CS 148) and Learning & 3D (CS 101) at Caltech. Labs & Teams: Leads Glab with members including Ilona Demler, Ziqi Ma, and Damiano Marsili
Yao Qin is an Assistant Professor in the Department of Electrical and Computer Engineering at the University of California, Santa Barbara (UCSB), with dual affiliation in the Department of Computer Science. She concurrently serves as Co-Director of the REAL AI Initiative at UCSB and holds a Senior Research Scientist position at Google DeepMind, where she contributes to the Gemini Multimodal project. Her academic credentials include a PhD in Computer Science and Engineering from the University of California, San Diego (advised by Prof. Garrison W. Cottrell) and a BS in Electrical Engineering from Dalian University of Technology. During her doctoral studies, she completed internships with pioneering researchers Geoffrey Hinton and Ian Goodfellow. Dr. Qin's research program centers on machine learning robustness, with emphasis on adversarial robustness, out-of-distribution generalization, and fairness. She develops reliable AI systems specifically for healthcare applications, with diabetes management as a primary focus. Her lab explores critical themes including AI safety in multimodal models and diabetes-specific AI solutions, particularly exercise metabolism modeling and glycemic effect prediction. Recent publications reveal a strong trajectory in robust machine learning with cross-domain applications. Her work consistently bridges theoretical robustness concepts with practical healthcare implementations, particularly in diabetes care. Key publication venues include CVPR, ICML, NeurIPS, and ICLR, with notable contributions to out-of-distribution detection, adversarial transfer learning, and multimodal AI safety. Her distinguished recognition includes: EECS Rising Star at MIT (2021) UCSB Regents' Junior Faculty Fellowship Award Helmsley Charitable Trust award for Type 1 diabetes research UCSB Faculty Research Grant American Diabetes Association Abstract Award (ADA-2025) Dr. Qin actively mentors four PhD students—Mehak Dhaliwal, Andong Hua, Kenan Tang, and Youngseok Yoon—on projects spanning LLMs for diabetes, multimodal robustness, and generative time-series modeling. Her research is funded by the Helmsley Charitable Trust and UCSB, with recent grants supporting exercise-specific AID algorithms for diabetes management. As Co-Director of the REAL AI Initiative, she leads a research ecosystem focused on developing reliable artificial intelligence. Current lab activities include organizing workshops at NeurIPS-2024 (AdvML-Frontiers and AIM-FM) and developing next-generation diabetes management tools through collaborations with medical institutions.
Vicente Ordóñez-Román is an Associate Professor in the Department of Computer Science at Rice University, part of the George R. Brown School of Engineering. His research focuses on the intersection of computer vision, natural language processing, and machine learning, with an emphasis on fair, transparent, and interpretable AI. He leads the Vision, Language, and Learning Lab and contributes to the Ken Kennedy Institute's Closed-loop Computer Vision research cluster. Education: PhD in Computer Science (UNC Chapel Hill, 2015), MS in Computer Science (Stony Brook University), and Engineering (Escuela Superior Politécnica del Litoral, Ecuador). Prior roles include Assistant Professor at the University of Virginia (2016-2021) and visiting positions at Adobe Research, the Allen Institute for AI, and Amazon. Research Interests : Developing multimodal AI systems that integrate visual and textual data, mitigating biases in AI, and advancing generative models. His work emphasizes ethical AI and societal impact, as seen in his contributions to the whitepaper advocating for federal regulation of facial recognition technologies. Awards & Recognition : NSF CAREER Award (2021), Marr Prize (ICCV 2013), Best Paper at EMNLP 2017, and multiple industry grants from Google, Amazon, and Facebook. His research has been featured in media outlets like WIRED, The New York Times, and Bloomberg News. Advising & Grants : Supervises a diverse research group spanning PhD, MS, and undergraduate students. Secured over $1.8 million in external funding, including NSF grants, Amazon FAI awards, and Google Cloud credits. Leads initiatives on bias mitigation, AI ethics, and multimodal learning. Labs & Collaborations : Directs the Vision, Language, and Learning Lab (vislang.ai), collaborating with industry partners like Adobe, Amazon, and SAP. Engages in interdisciplinary projects at the Ken Kennedy Institute, focusing on closed-loop computer vision systems.
Professor Winston Hsu is a distinguished faculty member in the Department of Computer Science and Information Engineering at National Taiwan University, where he has served as a full professor since 2015. He is the co-director of the Communications and Multimedia Laboratory (CMLab) and founder of the MiRA (Multimedia indexing, Retrieval, and Analysis) research group. Additionally, he serves as the Founding Director for NVIDIA AI Lab at NTU, the first such lab in Asia. Professor Hsu received his Ph.D. in Electrical Engineering from Columbia University in 2007 under the supervision of Professor Shih-Fu Chang. Prior to his academic career, he was a founding engineer and research manager at CyberLink Corp., now a public image/video software company. National Taiwan University (2007-Present): Professor (2015-Present), Assistant/Associate Professor (2007-2015) MobileDrive (2021-2024): CTO and Vice President (Joint Venture between Foxconn and Stellantis) IBM TJ Watson Research Center (2016-2017): Visiting Scientist Microsoft Research Redmond (2014): Visiting Researcher Columbia University (2007): Ph.D. in Electrical Engineering Professor Hsu's research focuses on machine learning, computer vision, large-scale image and video search and recognition, and embedded AI. His work spans from fundamental research in visual recognition to practical applications in automotive systems, medical imaging, and e-commerce. He has pioneered work in disguised face recognition, low-resolution face hallucination, 3D model search, and virtual try-on systems. His current research emphasizes Embodied AI, integrating perception, action, and learning technologies for applications in automotive and robotics domains. His research group has produced numerous influential publications, particularly in top computer vision and multimedia conferences like CVPR, where they won first place in the Disguised Face Recognition competition in 2018. Their work spans diverse application areas including security, medical diagnostics, automotive systems, and e-commerce solutions, demonstrating strong translation from academic research to real-world impact. IBM Research Pat Goldberg Memorial Best Paper Award (2018) First Place, IARPA Disguised Faces in the Wild Competition (CVPR 2018) Best Brave New Idea Paper Award, ACM Multimedia 2017 NVIDIA AI LAB Award (First in Asia, 2016) First Place, MSR-Bing Image Retrieval Challenge (2013) World's Top 2% Scientists (2023) Professor Hsu actively mentors students and researchers, with his group consistently recruiting PhD students, postdocs, and research assistants. He has successfully bridged academia and industry through multiple collaborations, including his role as CTO at MobileDrive (a joint venture between Foxconn and Stellantis) from 2021-2024. His research has been supported by significant industry partnerships with Microsoft, IBM, and NVIDIA, as well as government grants from Taiwan's Ministry of Science and Technology. His laboratory, the Communications and Multimedia Laboratory (CMLab), maintains strong industry connections and focuses on cutting-edge research in visual AI. The lab has produced numerous award-winning projects and maintains active collaborations with global technology companies, particularly in the automotive and consumer electronics sectors.
Mark Yatskar is an Assistant Professor in the Department of Computer and Information Science at the University of Pennsylvania. His research focuses on the intersection of natural language processing, computer vision, and fairness in machine learning. He earned his PhD from the University of Washington under advisors Luke Zettlemoyer and Ali Farhadi, and previously worked as a Young Investigator at the Allen Institute for Artificial Intelligence. Education: PhD in Computer Science, University of Washington (Advisor: Luke Zettlemoyer & Ali Farhadi) Research Interests: Yatskar's work explores how language can structure visual perception and mitigate human biases in machine learning systems. Key themes include: Natural language as a scaffold for visual intelligence Bias characterization and control in machine learning systems His lab currently investigates projects like language-guided bottlenecks, annotator cognitive heuristics, and gender bias amplification. Teaching: CIS 5300: Computational Linguistics (2021-2024) CIS 7000: Language and Vision (2020) CIS 6300: Efficient NLP (2023, 2025) Awards: Best Paper Award at EMNLP (Gender Bias Amplification Research) Advising & Grants: Yatskar advises a team of PhD/Master's students and actively seeks motivated researchers. His group has explored funding in areas like interpretable AI, multimodal reasoning, and dataset bias mitigation. Labs/Teams: Leads the Penn NLP & Vision Lab, focusing on projects like MolMo/PixMo open models, ViUniT visual unit tests, and bias mitigation frameworks.
Enamul Hoque Prince is an Associate Professor and Director of the School of Information Technology at York University. He leads the Intelligent Visualization Lab, funded by the Canada Foundation for Innovation (CFI) and Ontario Research Funds (ORF). He holds a PhD in Computer Science from the University of British Columbia and completed postdoctoral work at Stanford University. His research integrates information visualization, human-computer interaction (HCI), and natural language processing (NLP) to address information overload challenges. Dr. Prince's educational background includes a PhD from UBC, an MSc from Memorial University of Newfoundland, and a BSc from Chittagong University of Engineering & Technology. He has conducted research at institutions like Tableau Software and the Qatar Computing Research Institute and serves on committees for top conferences like ACL and IEEE Vis. His work is supported by grants from NSERC, CFI, and others. Research interests focus on NLP-driven visual analytics, user-adaptive visualization, and accessible interfaces. Notable projects include Evizeon (natural language interfaces for visual analytics), ConVisIT (topic modeling for online conversations), and CIDER (concept-based image search). Publications highlight trends in multimodal systems, chart comprehension, and accessibility. Key awards include the NSERC Discovery Grant (2019) and a Best Paper Honorable Mention at DIS 2021. He supervises graduate and undergraduate students in areas like visualization, NLP, and HCI. Labs and collaborations emphasize interdisciplinary approaches, with the Intelligent Visualization Lab advancing tools for data exploration and user-centered design. Teaching includes courses on design principles and information visualization.
Yang Zhou is an Associate Professor in the Department of Computer Science and Software Engineering at Auburn University, part of the Samuel Ginn College of Engineering. His research focuses on big data algorithms, machine learning, data mining, and distributed computing. He has contributed to advancements in federated learning frameworks, graph mining tools, and spatial machine learning for environmental applications like flood mapping. Education includes a Ph.D. in Computer Science from Georgia Tech (2021), M.E. in Computer Application Technology from Chongqing University (2016), and B.E. in Engineering from Jiangnan University (2014). His work emphasizes scalable algorithms for large-scale systems, with tools like DirDense for dense subgraph mining and FedASMU for federated learning optimization. Recent publications explore adversarial robustness, blockchain strategies in IoT, and curriculum-based learning for large language models. He advises on interdisciplinary projects at the intersection of AI and environmental science.
Jeannette Bohg is an Assistant Professor of Computer Science at Stanford University, directing the Interactive Perception and Robot Learning Lab. Previously, she was a group leader at the Autonomous Motion Department (AMD) of the MPI for Intelligent Systems (2012-2017). She holds a PhD from KTH Royal Institute of Technology (Stockholm) and degrees from Chalmers University and TU Dresden. Her research focuses on perception, learning, and real-time multi-modal methods for autonomous robotic manipulation and grasping, aiming to bridge principles of human sensorimotor coordination with robotic implementation. Education: PhD in Robotics (KTH), MSc in Art & Technology (Chalmers), Diploma in Computer Science (TU Dresden) Research interests include developing goal-directed, real-time robotic systems capable of meaningful feedback for execution and learning. Key areas are dexterous manipulation, imitation learning, and cross-embodiment policy transfer. Notable contributions include the TidyBot platform and work on force-aware surgical robotics. Awards include the 2019 IEEE ICRA Best Paper Award, 2019 IEEE RA Early Career Award, and 2020 RSS Early Career Award. Her lab explores intersections of robotics, ML, and computer vision. Advising: Actively mentoring students/postdocs in manipulation, perception, and learning. Grants and collaborations span NSF, Stanford AI Lab, and industry partnerships. Future work emphasizes robust real-world deployment and human-robot collaboration. Labs/Teams: Leads the Interactive Perception and Robot Learning Lab, contributing to Stanford’s AI ecosystem. Previously managed the MPI AMD group, fostering interdisciplinary research in autonomous systems.
Dr. Emily Wall is an Assistant Professor in the Department of Computer Science at Emory University, leading the CAV Lab (Cognition and Visualization). She earned her Ph.D. in Computer Science from Georgia Tech (2020) and completed a postdoctoral fellowship at Northwestern University before joining Emory. Ph.D., Computer Science, Georgia Tech (2020) Postdoctoral Researcher, Northwestern University Current: Assistant Professor, Emory University Her research focuses on enhancing human decision-making through data visualization and visual analytics by addressing cognitive limitations and biases. Key areas include metacognition promotion, computational strategies for bias mitigation, and human-AI collaboration in data analysis. She combines cognitive psychology principles with technical implementation using tools like D3.js and Tableau. Recent publications highlight her work on LLM applications in visual analytics, cognitive bias interventions, and trust calibration in AI systems. Current students like Mengyu Chen, Shiyao Li, and Zhongzheng Xu are presenting at top conferences such as CHI and IEEE VIS. She also explores educational applications through interactive visualization pedagogy and innovative assessment methods. Dr. Wall actively engages in academic service through graduate mentorship, conference judging, and diversity initiatives. Her lab emphasizes both technical proficiency and social justice implications in visualization design, reflected in course projects that merge artistic expression with data-driven storytelling.
Nima Mesgarani is an Associate Professor of Electrical Engineering at Columbia Engineering, Columbia University, affiliated with the Sense, Collect and Move Data Committee. His research bridges engineering and neuroscience through reverse-engineering neural signal processing mechanisms, leading to advancements in brain-machine interfaces, neural prosthetics, and speech processing algorithms. He received his PhD in Electrical Engineering from the University of Maryland and completed postdoctoral training at Johns Hopkins University's Center for Language and Speech Processing and UC San Francisco's Neurosurgery Department. Research Focus Professor Mesgarani's lab integrates computational neuroscience and engineering to study acoustic signal processing. Key areas include: Neural decoding of speech and auditory attention in multi-talker environments Development of brain-controlled hearing technologies Novel speech separation and synthesis algorithms inspired by cortical processing Cross-modal learning between auditory and visual systems Applications of large language models in neural signal interpretation Publication Trends Analysis of his 15 most recent articles (2025) reveals dominant themes: neural decoding techniques using intracranial EEG, brain-inspired speech separation models (e.g., Mamba architectures), applications of large language models in auditory neuroscience, cross-modal distillation methods, and clinical translation of audio processing algorithms. A strong emphasis emerges on real-time brain-computer interfaces and noise-robust speech processing. Laboratory and Collaborations Mesgarani directs an interdisciplinary lab developing neurotechnology for hearing restoration. His team collaborates with neurosurgery departments and speech processing centers, focusing on translating theoretical models into clinical brain-machine interfaces. The lab's work has yielded patents for brain-informed speech separation systems and attention-decoding frameworks.
Prof. Mark Heitmann is a Professor of Marketing & Customer Insight at the University of Hamburg Business School. He holds a visiting appointment at Nova School of Business and Economics. His research focuses on AI-driven marketing strategies, social media analytics, brand equity, and consumer decision-making. He has held academic roles at institutions including the University of St. Gallen and Christian-Albrechts-University in Kiel before joining Hamburg in 2011. Academic Career: Holds a PhD (2004) and post-doctoral Habilitation (2007) from the University of St. Gallen. Visiting researcher at Columbia University and the Max Planck Institute for Human Development. Research Interests: Applications of artificial intelligence in marketing Social media and digital transformation Ethical consumer behavior Brand self-expression and visual design Marketing technology (MarTech) innovations Key Achievements: Over 20 peer-reviewed publications in top journals like Journal of Marketing Research and Harvard Business Review . Award-winning research includes the MSI/H. Paul Root Award and IJRM-EMAC Best Article Award. Active in translating research into scalable commercial solutions through software-as-a-service ventures. Team & Collaborations: Leads a research team including M.Sc. candidates Sammar Rath, Julia Rosada, Maximilian Witte, Tijmen Jansen, Claus Hegmann-Napp, and Magdalena Heynicke. Engages in interdisciplinary projects with industry partners.
Danai Koutra is an Associate Professor in Computer Science and Engineering at the University of Michigan, Ann Arbor, and an Amazon Scholar. Her research focuses on large-scale graph mining, graph neural networks, and interpretable machine learning methods for understanding complex networks. Key roles include leading the GEMS Lab and contributing to projects like DeltaCon (graph similarity) and VoG (graph summarization). She holds a PhD from Carnegie Mellon University and has authored over 80 publications in top venues like KDD, SDM, and NeurIPS. Educations: PhD in Computer Science, Carnegie Mellon University (2015) MS in Computer Science, Carnegie Mellon University (2015) Diploma in Electrical & Computer Engineering, National Technical University of Athens (2010) Research Interests: Her work spans graph mining, anomaly detection, knowledge graph completion, and applications in neuroscience, healthcare, and social networks. Recent projects include MAGNET (multi-agent graph networks) and GT2VEC (multimodal graph-text encoders). Grants & Awards: Recipient of the 2025 PECASE award, NSF CAREER Award (2019), and the 2016 ACM SIGKDD Dissertation Award. Active in organizing conferences like KDD and ECML/PKDD. Labs & Teams: Directs the GEMS Lab, collaborating on projects like FIDDLE (clinical data preprocessing) and SpecGreedy (dense subgraph detection). Engages in interdisciplinary efforts, including M-DICE (urban mobility analysis with Detroit).
Claus Lamm is a Full Professor of Biological Psychology at the University of Vienna , where he leads the Social, Cognitive and Affective Neuroscience Unit (SCAN-Unit) . He serves as Vice Dean for Research and Advancement of Early Career Researchers at the Faculty of Psychology and holds affiliations with the Vienna Cognitive Science Hub , Environment & Climate Change Hub , and Austrian Academy of Sciences . His academic career spans international collaborations and formative research experience abroad. Scientific Focus: Lamm investigates the neural underpinnings of empathy and prosocial behavior , employing multi-modal approaches combining neuroimaging, psychopharmacology, and psychoneuroendocrinology . His work extends to comparative studies with ravens and dogs, and explores environmental social neuroscience through climate change decision-making research. Recent publications show trends in cross-cultural psychology , machine learning applications , and neurobiological pathways related to social behavior. Awards & Grants: Recipient of the APS Mentor Award for his support of early career researchers. Funded by European Research Council , Austrian Science Fund , Vienna Science and Technology Fund , and intramural grants exceeding €10 million. Key projects include "Unravelling the opioid system in empathy" and "Comparative dog-human fMRI" . Media Engagement: A prominent public science communicator, Lamm has appeared in Nature , Science Magazine , and Austrian media outlets like Ö1 Mittagsjournal and ORF2 , discussing topics from pandemic psychology to social media effects . He maintains active outreach through Science TV and educational programs .