Georgia Gkioxari is an Assistant Professor in the Division of Computing and Mathematical Sciences at Caltech , with a part-time affiliation at Meta AI . Her work focuses on extending visual perception models through advanced 2D and 3D representation learning, spatial reasoning, and generative models. Education: Not explicitly mentioned in the text Research interests span 3D perception , spatial reasoning , and vision-language integration , with projects like Visual Agentic AI for Spatial Reasoning and Token-by-Token Multimodal Alignment . Her publications emphasize 3D object detection , reconstruction , and generative modeling techniques including diffusion models and transformers . Scientific recognition includes the Meta LLM Evaluation Research Grant , Okawa Research Grant , Google Faculty Scholar Award 2024 , and Amazon Research Award . She teaches courses like Large Language & Vision Models (EE/CS 148) and Learning & 3D (CS 101) at Caltech. Labs & Teams: Leads Glab with members including Ilona Demler, Ziqi Ma, and Damiano Marsili
Jacky Bourgeois is an Assistant Professor of Data-Centric Design at Delft University of Technology's Faculty of Industrial Design Engineering (IDE), leading the Data-Centric Delft Design Lab. His work bridges Human-Computer Interaction, Data Science, and Participatory Design, focusing on methods like data donation to empower collaborative data exploration. He holds a computer science background, previously researching ubiquitous technologies for domestic energy practices of 'energy farmers' (solar households). Key projects include the PROMISE study on patient-led oncological care and the Data Donation initiative promoting ethical data practices. Education: Background in Computer Science; academic focus on sustainable design engineering and IoT. Teaching includes courses like 'Data-Centric Design for...' (2024) and 'Digital Product Development' (2023). Research Interests: Data as subjective inquiry material, data intimacy, data commons, and human-data interaction in home contexts. Notable contributions include PAIRcolator for collaborative data reflection and Tangi for 360° video insights. Awards include Best Paper at CHI 2025 and DIS '24 Best Paper. Grants & Collaborations: NWO-funded research on hospital sounds (2024), IoT Rapid Proto Labs for industry-academia partnerships, and ancillary role as a scientific employee at The Open University (2025–2027). Labs & Teams: Director of the Data-Centric Delft Design Lab; collaborates with the HCI Data & Design SIG. Current projects address ethical frameworks for data donation and patient-centered healthcare systems.
Mark Yatskar is an Assistant Professor in the Department of Computer and Information Science at the University of Pennsylvania. His research focuses on the intersection of natural language processing, computer vision, and fairness in machine learning. He earned his PhD from the University of Washington under advisors Luke Zettlemoyer and Ali Farhadi, and previously worked as a Young Investigator at the Allen Institute for Artificial Intelligence. Education: PhD in Computer Science, University of Washington (Advisor: Luke Zettlemoyer & Ali Farhadi) Research Interests: Yatskar's work explores how language can structure visual perception and mitigate human biases in machine learning systems. Key themes include: Natural language as a scaffold for visual intelligence Bias characterization and control in machine learning systems His lab currently investigates projects like language-guided bottlenecks, annotator cognitive heuristics, and gender bias amplification. Teaching: CIS 5300: Computational Linguistics (2021-2024) CIS 7000: Language and Vision (2020) CIS 6300: Efficient NLP (2023, 2025) Awards: Best Paper Award at EMNLP (Gender Bias Amplification Research) Advising & Grants: Yatskar advises a team of PhD/Master's students and actively seeks motivated researchers. His group has explored funding in areas like interpretable AI, multimodal reasoning, and dataset bias mitigation. Labs/Teams: Leads the Penn NLP & Vision Lab, focusing on projects like MolMo/PixMo open models, ViUniT visual unit tests, and bias mitigation frameworks.
Enamul Hoque Prince is an Associate Professor and Director of the School of Information Technology at York University. He leads the Intelligent Visualization Lab, funded by the Canada Foundation for Innovation (CFI) and Ontario Research Funds (ORF). He holds a PhD in Computer Science from the University of British Columbia and completed postdoctoral work at Stanford University. His research integrates information visualization, human-computer interaction (HCI), and natural language processing (NLP) to address information overload challenges. Dr. Prince's educational background includes a PhD from UBC, an MSc from Memorial University of Newfoundland, and a BSc from Chittagong University of Engineering & Technology. He has conducted research at institutions like Tableau Software and the Qatar Computing Research Institute and serves on committees for top conferences like ACL and IEEE Vis. His work is supported by grants from NSERC, CFI, and others. Research interests focus on NLP-driven visual analytics, user-adaptive visualization, and accessible interfaces. Notable projects include Evizeon (natural language interfaces for visual analytics), ConVisIT (topic modeling for online conversations), and CIDER (concept-based image search). Publications highlight trends in multimodal systems, chart comprehension, and accessibility. Key awards include the NSERC Discovery Grant (2019) and a Best Paper Honorable Mention at DIS 2021. He supervises graduate and undergraduate students in areas like visualization, NLP, and HCI. Labs and collaborations emphasize interdisciplinary approaches, with the Intelligent Visualization Lab advancing tools for data exploration and user-centered design. Teaching includes courses on design principles and information visualization.
Professor Winston Hsu is a distinguished faculty member in the Department of Computer Science and Information Engineering at National Taiwan University, where he has served as a full professor since 2015. He is the co-director of the Communications and Multimedia Laboratory (CMLab) and founder of the MiRA (Multimedia indexing, Retrieval, and Analysis) research group. Additionally, he serves as the Founding Director for NVIDIA AI Lab at NTU, the first such lab in Asia. Professor Hsu received his Ph.D. in Electrical Engineering from Columbia University in 2007 under the supervision of Professor Shih-Fu Chang. Prior to his academic career, he was a founding engineer and research manager at CyberLink Corp., now a public image/video software company. National Taiwan University (2007-Present): Professor (2015-Present), Assistant/Associate Professor (2007-2015) MobileDrive (2021-2024): CTO and Vice President (Joint Venture between Foxconn and Stellantis) IBM TJ Watson Research Center (2016-2017): Visiting Scientist Microsoft Research Redmond (2014): Visiting Researcher Columbia University (2007): Ph.D. in Electrical Engineering Professor Hsu's research focuses on machine learning, computer vision, large-scale image and video search and recognition, and embedded AI. His work spans from fundamental research in visual recognition to practical applications in automotive systems, medical imaging, and e-commerce. He has pioneered work in disguised face recognition, low-resolution face hallucination, 3D model search, and virtual try-on systems. His current research emphasizes Embodied AI, integrating perception, action, and learning technologies for applications in automotive and robotics domains. His research group has produced numerous influential publications, particularly in top computer vision and multimedia conferences like CVPR, where they won first place in the Disguised Face Recognition competition in 2018. Their work spans diverse application areas including security, medical diagnostics, automotive systems, and e-commerce solutions, demonstrating strong translation from academic research to real-world impact. IBM Research Pat Goldberg Memorial Best Paper Award (2018) First Place, IARPA Disguised Faces in the Wild Competition (CVPR 2018) Best Brave New Idea Paper Award, ACM Multimedia 2017 NVIDIA AI LAB Award (First in Asia, 2016) First Place, MSR-Bing Image Retrieval Challenge (2013) World's Top 2% Scientists (2023) Professor Hsu actively mentors students and researchers, with his group consistently recruiting PhD students, postdocs, and research assistants. He has successfully bridged academia and industry through multiple collaborations, including his role as CTO at MobileDrive (a joint venture between Foxconn and Stellantis) from 2021-2024. His research has been supported by significant industry partnerships with Microsoft, IBM, and NVIDIA, as well as government grants from Taiwan's Ministry of Science and Technology. His laboratory, the Communications and Multimedia Laboratory (CMLab), maintains strong industry connections and focuses on cutting-edge research in visual AI. The lab has produced numerous award-winning projects and maintains active collaborations with global technology companies, particularly in the automotive and consumer electronics sectors.
Sanjay Purushotham is an Assistant Professor in the Department of Information Systems at the University of Maryland Baltimore County (UMBC), with a PhD in Electrical Engineering from the University of Southern California (USC) and a postdoctoral background in Computer Science at USC's Integrated Media Systems Center (IMSC). His research focuses on machine learning, data mining, and their applications in biomedical informatics, social network analysis, and multimedia data mining. Key contributions include survival analysis models using pseudo values and federated learning frameworks for healthcare data. He has received awards including the Best Paper Award at SIGSPATIAL 2014 and a Best Poster Runnerup at SCMLS 2016. Education: PhD in Electrical Engineering (USC), Postdoc in Computer Science (USC) His work spans interdisciplinary areas such as domain adaptation for remote sensing, thermal face translation, and interpretable neural networks for medical applications. Recent projects include federated survival analysis models and climate-informatics frameworks for cloud property retrieval. He teaches courses in artificial intelligence, healthcare informatics, and statistical learning at UMBC. Research highlights include developing MedFuseNet for multimodal medical question answering and VDAM for multi-sensor cloud data analysis. His work on fair survival analysis models addresses algorithmic bias in healthcare predictions. Current grants include a NSF CAREER award for trustworthy federated learning in computational healthcare.
Olga Russakovsky is an Associate Professor in the Computer Science Department at Princeton University, with additional affiliations at the Center for Statistics and Machine Learning and the Center for Information Technology Policy. She serves as Associate Director of the Princeton AI Lab and chairs the Board of Directors for AI4ALL, the nonprofit she co-founded in 2017. Her research integrates computer vision with machine learning, human-computer interaction, and fairness/transparency in AI systems. She leads the Visual AI Lab, focusing on developing intelligent systems that understand the visual world while addressing societal impacts of technology. Dr. Russakovsky is a prominent advocate for diversity in AI. She co-founded multiple educational initiatives including: Stanford AI4ALL (originally SAILORS, 2015) for high school girls Princeton AI4ALL (2018) for underrepresented racial/ethnic groups AI4ALL nonprofit (2017) to cultivate diverse AI leadership AI4ALL Open Learning (free K-12 AI curriculum) College Pathways program for undergraduate inclusion She has influenced policy through congressional engagement and media outreach, with featured coverage in MIT Technology Review, TheAtlantic, Wired, and EdWeek addressing AI's diversity crisis.
Nima Mesgarani is an Associate Professor of Electrical Engineering at Columbia Engineering, Columbia University, affiliated with the Sense, Collect and Move Data Committee. His research bridges engineering and neuroscience through reverse-engineering neural signal processing mechanisms, leading to advancements in brain-machine interfaces, neural prosthetics, and speech processing algorithms. He received his PhD in Electrical Engineering from the University of Maryland and completed postdoctoral training at Johns Hopkins University's Center for Language and Speech Processing and UC San Francisco's Neurosurgery Department. Research Focus Professor Mesgarani's lab integrates computational neuroscience and engineering to study acoustic signal processing. Key areas include: Neural decoding of speech and auditory attention in multi-talker environments Development of brain-controlled hearing technologies Novel speech separation and synthesis algorithms inspired by cortical processing Cross-modal learning between auditory and visual systems Applications of large language models in neural signal interpretation Publication Trends Analysis of his 15 most recent articles (2025) reveals dominant themes: neural decoding techniques using intracranial EEG, brain-inspired speech separation models (e.g., Mamba architectures), applications of large language models in auditory neuroscience, cross-modal distillation methods, and clinical translation of audio processing algorithms. A strong emphasis emerges on real-time brain-computer interfaces and noise-robust speech processing. Laboratory and Collaborations Mesgarani directs an interdisciplinary lab developing neurotechnology for hearing restoration. His team collaborates with neurosurgery departments and speech processing centers, focusing on translating theoretical models into clinical brain-machine interfaces. The lab's work has yielded patents for brain-informed speech separation systems and attention-decoding frameworks.
Clio Andris is an Associate Professor at Georgia Tech, jointly appointed in the School of City and Regional Planning and the School of Interactive Computing. She directs the Friendly Cities Lab, focusing on mathematical models of social networks applied to urban planning, transportation, and geography. Her work integrates spatial analysis with visualization, emphasizing interdisciplinary collaboration. Education and Career: Andris earned a PhD in Urban Information Systems from MIT (2011), where she was an NDSEG Fellow. She held postdoctoral positions at Singapore-MIT Alliance for Research and Technology and the Santa Fe Institute. Prior to Georgia Tech, she was a faculty member at Penn State’s Department of Geography, affiliated with the GeoVISTA Center. Research Focus: Her research bridges social networks, geovisualization, and urban informatics. Key areas include spatial social network analysis, GIS applications for urban policy, and the impact of digital tools on civic engagement. She has developed innovative visual analytics tools like SNoMaN and ROBIN to democratize spatial data exploration. Awards: She received the NSF CAREER Award (2021) and NDSEG Fellowship (2011). Her lab is affiliated with the Center for Spatial Planning Analytics and Visualization (CSPAV) and the Information Visualization Lab. Labs and Collaborations: The Friendly Cities Lab focuses on socially just urban design through computational methods. Her work addresses issues like food security networks, pandemic impacts on biodiversity, and community mapping for activism. Grants and Outreach: Her NSF-funded projects emphasize public good applications, such as real-time pandemic risk communication and educational tools for migration data. She actively collaborates with non-profits and policymakers to translate research into actionable urban strategies.
Dr. Jeewanie Jayasinghe Arachchige is a Lecturer in the Department of Computer Science at Vrije Universiteit Amsterdam, Faculty of Science. She teaches undergraduate courses including Bachelor Project Computer Science, Professional Development, and Software Engineering Processes for the academic year 2024–2025. Her research focuses on process mining , healthcare informatics , and data security . She applies process mining to analyze healthcare pathways and subpopulation treatment variations, develops explainable AI frameworks for predictive analytics, and examines data governance in emerging architectures like Data Lakehouses. Her work intersects legal informatics, particularly formalizing Sri Lankan civil court processes using ontology engineering. Recent publications highlight trends in balancing simplicity and complexity in process modeling, Industry 4.0 healthcare applications, and cybersecurity in model-driven web development. She has contributed to over 20 peer-reviewed articles since 2006, spanning topics from service-oriented architectures to value network analysis. Her teaching and research emphasize practical applications of IT in healthcare, legal systems, and enterprise environments. No ancillary activities are currently recorded.
Prof. Bernt Schiele is a Max Planck Director at the Max Planck Institute for Informatics and holds a Professorship at Saarland University. His research focuses on understanding multimodal sensor data, with key areas in computer vision, 3D object recognition, and machine learning. He leads the Computer Vision and Machine Learning group, addressing challenges in sensor fusion, scene understanding, and human activity recognition. Schiele has held academic roles at TU Darmstadt, ETH Zurich, and MIT, and contributes to top journals like IEEE Transactions on PAMI and conferences like ECCV. His work emphasizes robust models, interpretability, and domain adaptation for real-world applications. Education: PhD (1997, Grenoble), MSc (1994 Karlsruhe/1993 Grenoble) Key Positions: MIT (1997-2000), ETH Zurich (1999-2004), TU Darmstadt (2004-2010) Research interests span 3D scene understanding, multimodal sensor processing, and machine learning techniques for large-scale data. His recent work advances robust object detection, explainable AI, and domain-invariant training methods. He also chairs major conferences like ECCV 2018 and co-chairs ICCV 2011. Publications highlight innovations in interpretable vision transformers, certified explanations, and test-time adaptation. Despite no listed awards, his contributions shape foundational areas of computer vision and multimodal AI.
Wim Gevers is a faculty member at the Université libre de Bruxelles (ULB) and leads the CS4S – Cognitive Control & Sleep laboratory within the CRCN research centre. His work bridges cognitive psychology, neuroscience and sleep research to understand how the brain exerts control over thoughts and actions and how sleep contributes to these processes. Research Interests Cognitive Control & Metacognition: Investigating how subjective experiences such as confidence and the "urge-to-err" guide strategic adjustments in behaviour. Working Memory & Ordinal Cognition: Examining how order information is maintained and manipulated, and how these processes relate to mathematical competence. Sleep, Memory & Decision Making: Exploring how sleep-dependent consolidation influences motor learning and decision strategies. Across his 2022–2025 publications a clear trend emerges: a focus on metacognitive monitoring —how humans evaluate their own cognitive states—and the role of emotional and temporal context in shaping those evaluations. Studies range from reaction-time introspection and confidence judgements in perceptual tasks to the impact of aging and depression on metacognitive accuracy. Doctoral Supervision & Mentoring Whitney Stee (PhD 2024) – Sleep-dependent structural brain reorganization & motor learning Gaia Corlazzoli (PhD 2024) – Subjective experience in decision-making Myrtille Dewulf (PhD 2023) – Ordinal coding mechanisms in working memory Rebeca Sifuentes-Ortega (PhD 2023) – REM sleep and memory reactivation All dissertations were defended at ULB, Faculté des Sciences psychologiques et de l’éducation, with Wim Gevers formally listed as Promotor . Laboratory & Collaborative Networks As head of CS4S, Gevers coordinates a multidisciplinary team that combines behavioural experimentation, EEG/MEG, computational modelling and sleep polysomnography. The lab is embedded in the larger CRCN ecosystem, fostering collaborations with groups such as CO3 (consciousness), LCLD (language & deafness), and UR2NF (neurofunctional imaging).
Joydeep Biswas is an Associate Professor in the Computer Science Department at the University of Texas at Austin, where he serves as the Director of the Autonomous Mobile Robotics Laboratory (AMRL). He is also affiliated with Texas Robotics, the UT Machine Learning Laboratory, and UT Good Systems. Previously, he was an Assistant Professor in the College of Information and Computer Sciences at the University of Massachusetts Amherst. Dr. Biswas earned his PhD in Robotics from Carnegie Mellon University in 2014 and his B.Tech in Engineering Physics from the Indian Institute of Technology Bombay in 2008. His educational background has provided him with a strong foundation in both theoretical and applied aspects of robotics and artificial intelligence. Dr. Biswas's research focuses on enabling long-term autonomy for mobile robots operating in human environments. His work spans robot perception, motion planning, control systems, and AI, with the ultimate goal of creating self-sufficient autonomous mobile robots that can perform tasks accurately and robustly in real-world settings. He is particularly interested in perception, planning, and failure recovery for autonomous mobile robots, which supports his vision of having autonomous service mobile robots deployed at campus-to-city scale, both indoors and outdoors, performing assistive tasks over deployments spanning years. His IJCAI 2019 Early Career Spotlight talk summarizes much of his research to date and ongoing interests. His recent research has shown a strong trend toward social navigation, human-robot interaction, and the application of machine learning techniques to robotics problems. There's a clear progression from fundamental robotics research toward more complex, real-world applications that require robots to understand and navigate human social spaces effectively. His work increasingly integrates large language models and other advanced AI techniques with traditional robotics approaches, as evidenced by his recent publications on topics like preference-conditioned navigation, social navigation benchmarks, and instruction-following navigation systems. Dr. Biswas has received numerous prestigious awards including the NSF CAREER Award (2021), J.P. Morgan Faculty Research Award (2019), Amazon Research Award (2019), and a grant from Northrop Grumman Mission Systems (2018). These awards recognize his innovative contributions to the field of robotics and autonomous systems. As a dedicated educator and mentor, Dr. Biswas actively supervises PhD and master's students, with his PhD student Sadegh Rabiee winning the student poster award at the Northrop Grumman University Symposium 2019. He has secured significant grant funding from the National Science Foundation for projects including 'Introspective Perception and Planning for Long-Term Autonomy' and 'Interactive Synthesis and Repair For Robot Programs,' demonstrating his ability to secure competitive research funding and his commitment to advancing the field. Dr. Biswas leads the Autonomous Mobile Robotics Laboratory (AMRL), which serves as a hub for interdisciplinary research in mobile robotics. The lab has developed notable resources such as the UT Campus Object Dataset (CODA) for 3D perception research and SOCIALGYM, a framework for benchmarking social robot navigation. His team regularly deploys robots on the UT Austin campus and in urban environments to test and refine their approaches in realistic settings, bridging the gap between simulation and real-world application.
Huamin Qu is a Chair Professor in the Department of Computer Science and Engineering at the Hong Kong University of Science and Technology (HKUST). He serves as the Founding Dean of the Academy of Interdisciplinary Studies (AIS), Founding Head of the Division of Emerging Interdisciplinary Areas (EMIA), and was the Founding Acting Head of Computational Media and Arts (CMA) at HKUST(GZ). Qu directs the VisLab and coordinates the Human-Computer Interaction (HCI) group. He obtained his BS in Mathematics from Xi'an Jiaotong University and MS/PhD in Computer Science from Stony Brook University. Qu's research integrates Data Visualization , Human-Computer Interaction , and Human-Centered AI , with applications in urban informatics, social networks, and explainable AI. His work focuses on developing interactive systems for big data analytics, visual storytelling, and AI-driven decision support. Research extends to multimodal communication, fintech, and augmented reality applications. His publications emphasize visual analytics for complex datasets (mobility, social media, financial), interaction techniques for immersive environments, and AI-enhanced visualization tools. Recent works explore explainable AI interfaces and large-scale data communication frameworks. IEEE Visualization Academy (2020) IEEE VGTC Technical Achievement Award AI 2000 Most Influential Scholar (2019, 2023, 2024) 21 Best Paper/Honorable Mention awards IBM Faculty Award (2009) APICTA Merit Award (2015) Yelp Dataset Grand Prize (2018) Qu has advised 48 PhD graduates (21 now faculty at institutions like UC Davis, University of Minnesota, Texas A&M) and 30 MPhil students. He secured major grants including RGC theme-based projects (digital citizenship, air pollution), UGC AoE (slope safety), and China's 973 Program. As VisLab director, he leads 20+ researchers in visualization/HCI projects adopted by Microsoft, IBM, Huawei, and Tencent.
Professor Gabriel Brostow is a faculty member in the Department of Computer Science at University College London (UCL), where he leads research in Computer Vision and Human-Computer Interaction. He also serves as Chief Research Scientist and Senior Director of the R&D Team at Niantic, the company behind Pokémon GO. His work bridges academic research and industry applications, focusing on developing AI systems that enhance human capabilities through what he terms 'Human in the Loop AI'—now commonly referred to as Human-Centered AI. Brostow completed his BS in Electrical Engineering at UT Austin, followed by a PhD with Irfan Essa at Georgia Tech. He then pursued postdoctoral research with Roberto Cipolla's Computer Vision & Robotics Group at Cambridge University as a Marshall Sherfield Fellow, and with Marc Pollefeys in ETH Zurich's CVG Group. His research explores how AI, particularly Computer Vision, can serve as 'super-tools' for professionals across various domains including filmmaking, architecture, robotics, and scientific research. Specific interests include assistive technology for everyday life, authoring systems that maximize user effort, 3D reconstruction, depth estimation, and vision-language models. His work often involves creating systems that are validated through real-world human interaction to ensure practical utility. Analysis of his recent publications reveals a strong focus on practical applications of Computer Vision that directly interact with humans. His research spans 3D scene understanding, depth estimation, sketch-based interfaces, and multimodal AI systems. There's a clear emphasis on creating benchmarks and tools that facilitate human-AI collaboration, with applications in assistive technology, urban planning, filmmaking, and biodiversity monitoring. His work frequently appears at top conferences including CVPR, NeurIPS, ECCV, and CHI. Marshall Sherfield Fellowship Brostow actively mentors PhD students, with current advisees including Ross Murphy, Skanda Koppula, Gizem Unlu, Omiros Pantazis, and Jamie Watson. His alumni include numerous PhD graduates and MSc students who have gone on to successful careers in academia and industry. He emphasizes selecting students based on passion and potential rather than just academic credentials, valuing traits like helpfulness, drive, and hunger to learn. His research is supported through collaborations with major institutions and companies including DeepMind, MIT, and the University of Edinburgh. He leads a research group at UCL that collaborates closely with Niantic's R&D team, creating a unique bridge between academic research and industry application. His team's work frequently involves developing novel Computer Vision techniques that are validated through real-world human interaction, ensuring practical utility alongside technical innovation. The group explores blue-sky research problems with applications ranging from assistive technology to professional tools for filmmakers, architects, and scientists studying diverse environments.