David A. Smith is an Associate Professor at the Khoury College of Computer Sciences, Northeastern University. His research focuses on Natural Language Processing (NLP) and computational linguistics, with applications in machine translation, information retrieval, digital humanities, and social sciences. He is a founding member of the NULab for Texts, Maps, and Networks, a research center focused on digital humanities and computational social sciences. Smith's work has been funded by grants from the Mellon Foundation, NEH, and IMLS, supporting projects such as the Viral Texts initiative analyzing 19th-century newspaper networks and the Oceanic Exchanges project tracking transnational information flows. He has contributed to advancements in OCR for historical texts, text reuse detection, and computational analysis of classical languages. He has advised numerous PhD students, including Shijia Liu, Si Wu, and Ryan Muther, and teaches courses like Natural Language Processing and Information Retrieval. His research has been featured in outlets like Wired and the Economist .
Jiatao Gu is an Assistant Professor in the Department of Computer and Information Science (CIS) at the University of Pennsylvania, with a part-time role as Staff Research Scientist at Apple (MLR). He holds a Ph.D. in Electrical and Electronic Engineering from the University of Hong Kong (2018) and a B.Eng. in Electronic Engineering from Tsinghua University (2014). His research focuses on generative machine learning and AI agent interaction with the physical world, emphasizing multi-modal systems spanning language, images, videos, and 3D. Key themes include efficient modeling , flexible architecture design , and scalable decision-making frameworks . 2025: ICLR paper on DART framework 2024: TMLR work on GFlowNet alignment 2023: NeurIPS research on diffusion stability 2022: ACL papers on speech translation Recent publications explore diffusion models for text-to-image synthesis, 3D reconstruction, and efficient sampling techniques. His work addresses fundamental challenges in attention mechanisms, entropy collapse, and multi-stage distillation while advancing non-autoregressive translation and vision-language reasoning . Prospective students can apply through his recruitment process at UPenn. Prior affiliations include Meta AI (FAIR Labs) and academic collaborations with institutions like New York University's CILVR Lab.
Prof. Laura Leal-Taixé is a Professor at the Technical University of Munich (TUM) in the Department of Informatics, leading the Dynamic Vision and Learning group. She holds the Rudolf Mößbauer Tenure Track Assistant Professorship, promoted to W3 Associate Professorship in 2022. Her research focuses on computer vision and machine learning, particularly video analysis, multi-object tracking, and autonomous driving applications. Education: B.Sc./M.Sc. in Telecommunications Engineering (Technical University of Catalonia, UPC) Ph.D. in Information Processing (Leibniz University of Hannover, Germany) Postdoctoral Research at ETH Zürich (Switzerland) and TUM Research Interests: Laura’s work addresses challenges in video analysis, including motion analysis, semantic segmentation, and integrating social dynamics into urban traffic modeling. Her project socialMaps , funded by the Sofja Kovalevskaja Award, explores decoupling vehicle and pedestrian traffic using dynamic maps. Her research combines optimization techniques, deep learning, and sensor data for real-world applications like autonomous systems. Recent Trends: Her publications highlight advancements in multi-object tracking, 3D LiDAR segmentation, and trajectory forecasting, emphasizing neural networks and geometric approaches. Collaborations with industry and academia underscore her work’s practical impact. Awards: 2017 Sofja Kovalevskaja Award (€1.65M, Humboldt Foundation) 2017 DAAD Australia-German Joint Research Grant Travel grants from Women in CV, CVPR Doctoral Consortium, and Vodafone Foundation Labs & Projects: Leads the Dynamic Vision and Learning group at TUM, focusing on vision-based AI for autonomous systems and environmental monitoring. Active in developing datasets like DynamicEarthNet for semantic change analysis.
Alexei A. Efros is the Howard Friesen Professor in the EECS Department at the University of California, Berkeley, and a core member of the Berkeley Artificial Intelligence Research (BAIR) Lab. Previously, he spent a decade at Carnegie Mellon University’s Robotics Institute. His research focuses on data-driven computer vision, self-supervised learning, computational photography, and generative models. He has pioneered advancements in visual representation learning, including seminal work on neural radiance fields and generative adversarial networks. Education Background: Efros holds a PhD in Computer Science from MIT, though specific details of his academic journey are not explicitly provided in the text. His career includes postdoctoral research at the University of Oxford with Andrew Zisserman and collaborative work with Team WILLOW at INRIA Paris. Research Interests: Efros explores how vast uncurated visual data can be leveraged for understanding and synthesizing the visual world. Key areas include self-supervised learning, generative models, and applications in robotics and art. His lab has contributed influential techniques such as Style Transfer, GAN-based image synthesis, and neural scene representation learning. Recent work emphasizes real-time adaptation (Test-Time Training), 3D perception models, and ethical AI implications of generative systems. Publications: Over 150+ publications span topics like Generative Adversarial Networks (GANs), unsupervised learning, and visual-linguistic models. Notable works include Unpaired Image-to-Image Translation (CUT/GAU), Style Transfer , and Swapping Autoencoder . His research has significant industry impact, with techniques adopted in Adobe’s software and generative AI applications. Grants & Collaborations: Efros has secured major funding from NSF, DARPA, and industry partnerships (e.g., Adobe, NVIDIA). He co-leads projects on scalable vision models, ethical AI, and real-world perception systems. Current collaborations include work with MIT, NYU, and INRIA Paris. Labs & Teams: Leads the BAIR Vision Group at Berkeley, fostering interdisciplinary research between computer vision, graphics, and robotics. The group emphasizes Slow Science principles, prioritizing deep exploration over rapid publication.
Almut Sophia Koepke is a junior research group leader at the Technical University of Munich and University of Tübingen, focusing on multimodal learning problems integrating sound, vision, and text. Her work bridges foundational research in audio-visual understanding with practical applications in few-shot learning, zero-shot translation, and cross-modal attention mechanisms.
Min Yen Kan is an Associate Professor and Vice Dean of Undergraduate Studies at the National University of Singapore's School of Computing, Department of Computer Science. With a PhD from Columbia University (2002), he leads the Web Information Retrieval / Natural Language Processing Group (WING.NUS) and serves as ACL Ethics Committee co-chair. His research spans Natural Language Processing , Large Language Models , Digital Libraries , and Information Retrieval , with specific focus on scientific discourse analysis, fact verification, and multimodal systems. Current projects include Scholarly Document Information Extraction (TRL 6), Task-Oriented Dialogue Systems (TRL 4), and Recommendation Systems (TRL 5). Recent publications reveal strong trends in LLM limitations (bias, hallucination, evaluation), conversational recommendation systems , and misinformation detection . His work consistently bridges theoretical NLP with real-world applications in digital libraries and scientific communication. Award highlights include: CIKM 2019 Best Paper Award ACL Distinguished Service Awards Vannevar Bush Best Paper Award (JCDL 2012) ACM Distinguished Speaker designation Kan mentors PhD students with placements at Google and USTC, and serves as associate editor for Information Retrieval and survey editor for Journal of AI Research . His lab WING.NUS develops practical tools like SciWING for scientific document processing and FANG for fake news detection. Media engagements include commentary on AI regulations in Southeast Asia and workforce implications in the AI era.
Lexing Xie is a Professor in Computer Science at the Australian National University. He leads the ANU Computational Media Lab ( http://cm.cecs.anu.edu.au ) and the ANU Integrated AI Network. His work focuses on the intersection of machine learning, social media analysis, and multimedia understanding. Dr. Xie's research broadly focuses on innovative design and use of machine learning algorithms, especially on large-scale graph data and collective behaviour. His recent work spans several key areas: Popularity in social media -- understanding, predicting, and optimization Multimedia knowledge graphs, vision and language integration Humanising machine intelligence through better understanding of social dynamics His publications reveal a strong trend toward understanding information diffusion patterns in social media, particularly through visual content. He has made significant contributions to the study of visual memes, popularity prediction using point processes, and multimodal learning that connects vision with language. His work often bridges theoretical machine learning with practical applications in social media analysis. Dr. Xie has received recognition for his research, including an Honourable Mention at CSCW 2019 for his work on attention flow in online video networks. His research has been supported by collaborations with major institutions including IBM Research and Columbia University. As an advisor, Dr. Xie has mentored numerous students who have gone on to contribute significantly to publications in top-tier conferences. His lab, the ANU Computational Media Lab, serves as a hub for interdisciplinary research connecting computer science with social sciences.
Matthias Hein is a Professor at the Department of Computer Science, Faculty of Mathematics and Natural Sciences, University of Tübingen. His research focuses on Machine Learning , Adversarial Robustness , and Out-of-Distribution Detection , with applications in computer vision and medical imaging. He has received notable recognition including the Best Paper Honorable Mention Prize at ICLR 2021 and Outstanding Paper Award at CVPR 2021. His work includes developing benchmarks like RobustBench and Spurious ImageNet , and frameworks such as Sparse-RS and DIG-IN . His recent publications emphasize adversarial robustness across multiple domains (vision, text), counterfactual explanations for classifiers, and improved OOD detection methods . Collaborators include prominent researchers like Francesco Croce, Julian Bitterwolf, and Alexander Meinke. Scientific Awards : Best Paper Honorable Mention (ICLR 2021) CVPR 2021 Outstanding Paper Award Key Research Areas : Adversarial Robustness Vision-Language Models Medical Imaging AI Neural Network Calibration
Roles & Affiliations: Manolis Savva is an Associate Professor at Simon Fraser University's School of Computing Science and holds a Canada Research Chair in Computer Graphics. He leads research in 3D scene understanding, with applications in graphics, vision, and robotics. Previously, he was a researcher at Facebook AI and Princeton University. Education: Ph.D. in Computer Science (2016), Stanford University, advised by Pat Hanrahan B.A. in Physics and Computer Science (2009), Cornell University Research Interests: His work focuses on analyzing, organizing, and generating 3D content, particularly for holistic scene understanding. Key areas include articulated objects, embodied AI, and datasets like ScanNet , Matterport3D , and Habitat . His methods drive applications in robotics, autonomous agents, and virtual environments. Publications Trends: Recent work emphasizes generative models (e.g., SINGAPO for articulated object parts), embodied AI benchmarks (Habitat), and multimodal scene analysis. His papers often address challenges in scalability, realism, and cross-modal fusion for 3D environments. Awards: CHCCS Early Career Researcher Award (2022) ICLR 2023 Outstanding Paper Award ICCV 2019 Best Paper Nomination (Habitat) Advising & Grants: Supervised over 15 graduate students, many advancing to top PhD programs and tech companies. Active in grants for embodied AI, scene understanding, and robotics. Labs/Teams: Leads the 3DLG (3D Learning Group) and GrUVi (Graphics and Vision) groups at SFU. Collaborates extensively with industry (e.g., Meta, NVIDIA) on AI-driven 3D research.
Professor Carlo Harvey is a creative technologist at the School of Digital Arts (SODA), Manchester Metropolitan University. His interdisciplinary research merges games , machine learning , virtual production , and cultural heritage reinterpretation . He leads industry collaborations with entities like Jaguar Land Rover and Epic Games, focusing on AI-driven interactive audio, real-time visualization, and accessibility solutions. Award-winning projects : TIGA, Innovate UK, and Epic Games MegaGrant for Accession Industry partnerships : Automotive sector, cultural institutions His research spans human-computer interaction , multisensory virtual environments , and acoustic-visual cross-modal perception . Recent publications address robotic simulations, motion alignment, and haptic feedback systems. Scientific recognition : TIGA Award, Innovate UK Funding, Epic Games MegaGrant Advocacy : Digital inclusion, creative collaboration, social impact of technology
Barbara Caputo is a Full Professor at Politecnico di Torino, leading the VANDAL Laboratory and directing the AI@PoliTo Interdepartmental Lab. She holds a double affiliation with the Italian Institute of Technology (IIT) and has held roles at Idiap-EPFL and Sapienza University. Her research focuses on AI, computer vision, domain adaptation, and federated learning. She contributes to national AI policy, including the Italian Strategy on AI and the National PhD on AI for Industry 4.0. She is an ERC Laureate and ELLIS Fellow, co-founding ELLIS society. Her work spans visual place recognition, action recognition, and cross-domain learning. Education: PhD in Computer Science from KTH Royal Institute of Technology (2005). Major roles include Rector’s Advisor on AI at PoliTo, Board Member of ELLIS, and coordinator of the AI & Industry 4.0 vertical in the National PhD program. Awards include ERC Laureate (2017), ELLIS Fellow (2019), and Inspiring Fifty Italy (2018). Her research emphasizes federated learning, domain adaptation, and AI ethics. Recent articles explore domain generalization, resource-efficient federated models, and AI-environment interactions. She collaborates with institutions like MUR, CNR, and the European Commission on AI policy and tech initiatives.
Prof. Matthias Nießner is a Professor at the Technical University of Munich , where he leads the Visual Computing Lab . Prior to this, he held a Visiting Assistant Professor position at Stanford University . His work bridges computer vision , graphics , and machine learning , focusing on 3D reconstruction , semantic scene understanding , and AI-driven video synthesis . Prof. Nießner has published over 150 works in top venues like SIGGRAPH , CVPR , and ECCV , with several receiving best paper awards (SIGCHI’14, HPG’15, SPG’18, SIGGRAPH’16 Emerging Tech). His research has garnered international media attention, including features in the New York Times , Wall Street Journal , and MIT Technological Review , as well as TV demonstrations (e.g., Jimmy Kimmel Live for Face2Face technology). Awards : TUM-IAS Rudolph Moessbauer Fellowship (2017–ongoing) Google Faculty Award (2017) Nvidia Professor Partnership Award (2018) ERC Starting Grant (2018, €1.5M) Eurographics Young Researcher Award (2019) Research Trends : 3D Gaussian Splatting for real-time rendering Neural Radiance Fields (NeRF) with mesh supervision Audio-driven facial animation via diffusion models Latent space diffusion for 3D scenes Self-supervised and zero-shot methods for 3D and image analysis As a co-founder and director of Synthesia Inc. , he drives democratization of synthetic media. His YouTube channel has over 5 million views, reflecting his impact beyond academia.
Xavier Serra is a Full Professor at the Department of Engineering at Universitat Pompeu Fabra (UPF), Barcelona. He is the founder and director of the Music Technology Group (MTG), and leads the UPF-BMAT Chair on AI and Music. He also coordinates the Master in Sound and Music Computing and serves as President of the Phonos Foundation. His research focuses on audio signal processing, sound and music computing, and computational musicology, emphasizing open science and open innovation. Education: BSc in Biology, University of Barcelona (1981) Master in Music, Florida State University (1983) PhD in Computer Music, Stanford University (1989) Research Interests: Audio Signal Processing Data-Driven and Knowledge-Driven Methodologies Music Information Retrieval Cultural Music Analysis (e.g., Carnatic/Turkish/Andalusian Music) Music Education Technology Notable Projects: CompMusic (ERC Advanced Grant, 2010-2017): Multicultural computational music analysis Open datasets: Freesound, Saraga, FSD50K Technologies: Reactable, Vocaloid, Essentia API Recent Trends in Articles: Focus on AI-driven audio processing (neural fingerprints, generative models), cross-cultural music analysis, and explainable music difficulty estimation. Awards: ERC Advanced Grant (2010) for CompMusic Project. Labs/Teams: Director of MTG, Phonos Foundation, and UPF-BMAT Chair. Active in open-source projects and international collaborations.
Haibin Ling is the SUNY Empire Innovation Professor in the Department of Computer Science at Stony Brook University, part of the College of Engineering and Applied Sciences. His research focuses on computer vision, medical image analysis, augmented reality, and AI applications in science. He holds a Ph.D. from the University of Maryland (2006) and prior degrees from Peking University. Previously, he worked at Temple University (2008–2019) and held roles at Siemens Corporate Research, UCLA, and Microsoft Research Asia. Professor Ling's work spans biomedical imaging, AI for science, and human-computer interaction. He leads the CV Lab and collaborates with the AI Institute at Stony Brook. Awards include the NSF CAREER Award (2014), Best Student Paper (ACM UIST 2003), and IEEE Fellow (2020). He serves on editorial boards for IEEE Trans. PAMI, Pattern Recognition, and CVIU, and chairs major conferences like CVPR. His research group includes over 50 students and alumni, with active projects in tracking benchmarks (LaSOT), Leafsnap, and medical imaging tools. Notable publications address OCTA flow estimation, backdoor attacks on vision models, and topology-guided medical learning. Collaborations involve institutions like Temple University and Stony Brook's Department of Applied Mathematics and Statistics.
Anders Søgaard is a Professor at the University of Copenhagen , affiliated with both the Department of Computer Science and the Department of Communication. His research bridges Natural Language Processing and Machine Learning with a focus on AI ethics , explainability , and human-AI interaction . Primary Affiliation: Department of Computer Science, University of Copenhagen Secondary Affiliation: Department of Communication, University of Copenhagen Email: soegaard@di.ku.dk, soegaard@hum.ku.dk Research Interests His work spans Natural Language Processing , Machine Learning , and AI ethics , with recent studies addressing: Trustworthiness in AI systems Explainable AI (XAI) frameworks Multilingual model fairness and alignment Human-AI collaboration in reasoning tasks Ethical implications of social robots Mental health analytics using ML Recent Publications His 2025 output highlights trends in: AI ethics (e.g., fairness metrics, trustworthy systems) Multilingual model analysis (knowledge retention, cross-lingual transfer) Human-centric AI (gaze data, cultural considerations) Applications in healthcare and social good