Andrew Zisserman is a Royal Society Research Professor at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). His research focuses on computer vision, artificial intelligence, and neural networks, with significant contributions to multimodal learning, video understanding, and 3D scene analysis. He leads projects exploring visual-language models, audio-visual synchronization, and clinical imaging applications. Key research areas include: Video analysis and temporal modeling Multimodal systems for sign language translation and action recognition 3D shape estimation and physical property inference Foundation models and cross-modal retrieval Recent work highlights: Developed Flamingo and Tapir models for video-language tasks Advancements in spinal MRI analysis and clinical imaging Leadership in EGO4D and VoxCeleb challenges Honors include Fellowship of the Royal Society (FRS) and the ISSLS Prize in Clinical Science 2023 for spinal analysis innovations. His lab collaborates globally, emphasizing real-world applications in healthcare and autonomous systems.
Almut Sophia Koepke is a junior research group leader at the Technical University of Munich and University of Tübingen, focusing on multimodal learning problems integrating sound, vision, and text. Her work bridges foundational research in audio-visual understanding with practical applications in few-shot learning, zero-shot translation, and cross-modal attention mechanisms.
Timothy M. Hospedales is a Professor of Artificial Intelligence at the Institute of Perception, Action and Behaviour within the School of Informatics at the University of Edinburgh . He also serves as VP AI and Head of Samsung AI Research Centre Europe . His research focuses on efficient and robust AI , emphasizing meta-learning , lifelong transfer-learning , and domain adaptation in both probabilistic and deep learning frameworks. Applications span computer vision , vision and language , reinforcement learning for robotics , and finance . Professor at University of Edinburgh (2020–present) ELLIS Fellow (2021) Head of Samsung AI Research Europe (2020–present) Founding Director of Applied Machine Learning Lab at QMUL (2012–2016) His work includes pioneering contributions to meta-learning , few-shot learning , and self-supervised methods , with notable awards such as the Best Paper Prize at ICML AutoML 2018 and Best Student Paper at ICPR 2018 . He has co-authored 15+ recent papers on topics like Vision-Language Models , Medical AI Fairness , and Diffusion Model Optimization . He served as Program Co-Chair for BMVC 2018 and AAAI 2022 , and authored a book on Visual Adaptation in the Deep Learning Era (2022). Co-Chair, BMVC 2018 Guest Editor, IET CV Special Issue (2016) Keynote Speaker at TASK-CV Workshop (ECCV 2016) Special Issue on Fewer Labels (IEEE PAMI 2020) His leadership extends to organizing workshops like the Learning-to-Learn Workshop at ICLR 2021 , Meta-Learning Workshop at NeurIPS 2020 , and Domain Generalisation Workshop at ICLR 2023 . Current projects include Meta-Omnium (CVPR 2023) for general-purpose meta-learning and MetaAudio (ICANN 2022) for few-shot audio classification benchmarks.
Professor Carlo Harvey is a creative technologist at the School of Digital Arts (SODA), Manchester Metropolitan University. His interdisciplinary research merges games , machine learning , virtual production , and cultural heritage reinterpretation . He leads industry collaborations with entities like Jaguar Land Rover and Epic Games, focusing on AI-driven interactive audio, real-time visualization, and accessibility solutions. Award-winning projects : TIGA, Innovate UK, and Epic Games MegaGrant for Accession Industry partnerships : Automotive sector, cultural institutions His research spans human-computer interaction , multisensory virtual environments , and acoustic-visual cross-modal perception . Recent publications address robotic simulations, motion alignment, and haptic feedback systems. Scientific recognition : TIGA Award, Innovate UK Funding, Epic Games MegaGrant Advocacy : Digital inclusion, creative collaboration, social impact of technology
Dr A. I. Shihab is a Senior Lecturer at Kingston University's Faculty of Engineering, Computing and the Environment, Department of Networks and Digital Media. He teaches programming languages (C++/Java), data structures, web development, and AI/machine learning. His research focuses on affective computing and machine learning applications including: Acoustic event detection in sports environments Audio signal analysis for tennis match modeling Multi-camera visual surveillance systems Medical imaging analysis using fuzzy clustering techniques Publications demonstrate expertise in combining audio/video modalities for sports analytics (tennis rallies, court-shots) and developing Markov models for sound event sequence analysis. Contact: a.shihab@kingston.ac.uk
Andrew Markham is a Professor of Computer Science at the University of Oxford , affiliated with Kellogg College . He leads a research group focusing on Cyber Physical Systems (CPS) , specializing in sensors, signal processing, and machine learning to enable machines to better perceive the physical world. His work emphasizes cross-disciplinary collaboration, notably in wildlife tracking and indoor positioning systems. He has held roles as a Postdoctoral Fellow (2008-2012), Associate Professor (2013), and Full Professor (2021). Education : PhD in Electrical Engineering (University of Cape Town, 2008), BSc (Hons) in Electrical Engineering (2004). Research Interests : Tracking and localization in GPS-denied environments (e.g., underground, indoors), magneto-inductive systems, physics-informed machine learning, and data-driven approaches for noisy sensor data. His projects include wildlife monitoring via wireless sensor networks and mmWave radar for human motion capture. Key Projects : CARACAL acoustic monitoring system, mmPoint dense human tracking, and RandLA-Net for large-scale point cloud segmentation. His work spans robotics, environmental sensing, and biomedical applications. Advising & Grants : Supervises over 30 students and collaborates with industrial partners. Research teams include Cyber Physical Systems, Autonomous Ubiquitous Sensing, and Wildlife Monitoring initiatives. Labs/Teams : Leads the CPS research group, focusing on sensor networks, inertial navigation, and multimodal fusion systems. Collaborates with zoology and earth science disciplines on applied projects.
Cecilia Mascolo is a Professor of Mobile Systems at the University of Cambridge , specifically in the Department of Computer Science and Technology . She co-directs the Centre for Mobile, Wearable System and Augmented Intelligence and is a Fellow of Jesus College, Cambridge . Her research focuses on mobile systems , machine learning for mobile health , and earable technology . She has been awarded prestigious grants such as the ERC Advanced Research Grant (2019-2025) and the EPSRC Open Research Fellowship (2025-2030). Currently on sabbatical at Harvard University , her work bridges systems and machine learning for health applications. Education: PhD in Computer Science from the University of Bologna, Italy. Previous Affiliation: Faculty at University College London before 2008. Her research spans mobile and wearable systems for health and behavior monitoring, focusing on on-device machine learning , uncertainty-aware models , and audio-based diagnostics . Key areas include federated learning , edge computing , and respiratory disease progression analysis via wearables. She explores earable technology for physiological monitoring, gait analysis, and even toothbrushing tracking using in-ear sensors. Her recent publications highlight advancements in earable-based health monitoring , including heart rate estimation , respiratory rate detection , and ECG analysis using machine learning. She emphasizes longitudinal health data from consumer devices, advocating for scalable diagnostics beyond traditional clinical standards. Scientific Awards: ERC Advanced Research Grant EPSRC Open Research Fellowship Best Paper Award - IEEE Percom 10-Year Impact Award - ACM Ubicomp Computer Laboratory Ring Hall of Fame Best Paper Award Student: Andrea Ferlini - ACM SIGMOBILE Doctoral Dissertation Runner-up She leads the Mobile Systems Research Laboratory , mentoring a team of 15 researchers (postdocs and PhD students), and has graduated over 25 PhD students. Her teaching includes Mobile Health courses at the University of Cambridge, and she serves as Director of Studies for Computer Science at Jesus College.
Iain Murray is Professor of Machine Learning and Inference at the School of Informatics, University of Edinburgh. His research focuses on developing flexible probabilistic models applicable across diverse domains including cosmology, neuroscience, perception, speech, sports, and text. Program Chair for ICLR (2018) Publications Chair for ICML (2017, 2018) Area Chair for AISTATS, ICLR, ICML, NeurIPS, and UAI Amazon Scholar (2018-2024), first appointed in Europe Murray's research interests center on probabilistic reasoning using machine learning, with specific expertise in density estimation and Markov chain Monte Carlo methods. His work spans theoretical foundations and practical applications, with significant contributions to neural autoregressive distribution estimation (NADE), real-valued NADE (RNADE), and pseudo-marginal slice sampling techniques. His research has enabled advances in flexible probabilistic modeling across multiple domains. His publications show consistent focus on advancing probabilistic modeling techniques, with recent work emphasizing neural autoregressive models, density estimation methods, and efficient sampling algorithms. The research trajectory demonstrates progression from foundational work on NADE to increasingly sophisticated deep learning approaches for density estimation and inference. Notable Paper Award for NADE work Amazon Scholar (2018-2024) Murray has supervised numerous PhD students who have gone on to prominent positions at Google DeepMind, NYU, stability.ai, and other leading institutions. His teaching responsibilities include the Machine Learning and Pattern Recognition course and project supervision. His research group focuses on developing tractable probabilistic models with applications across multiple scientific domains.
Dr. João F. Henriques is a Research Fellow at the Royal Academy of Engineering and a core member of the Visual Geometry Group (VGG) at the University of Oxford. His work spans the intersection of machine learning , deep learning , and computer vision , with notable contributions to visual tracking , 3D reconstruction , and robotics . He actively mentors DPhil students and collaborates across disciplines including AI safety , NeRFs , and optimisation . Current Students: Marian Longa, Tim Franzmeyer, Dominik Kloepfer, Yash Bhalgat, Shivani Mall, Lorenza Prospero, Mark Eid Graduated Students: Xu Ji, Mandela Patrick, Shu Ishida, Andreea Oncescu Research Trends from his recent work include advances in 3D scene reconstruction (e.g., Flash3D, GST), robotic adaptation (Rapid Motor Adaptation), and multimodal learning (Text2Loc, SCENES). His publications frequently address theoretical guarantees in unsupervised detection and reinforcement learning for POMDP environments. Scientific Recognition includes: Research Fellow, Royal Academy of Engineering CVPR Best Paper Finalist (2012) for Kernelized Correlation Filters (KCF) SIGBOVIK 2020 Most Timely Paper Award for Deep Industrial Espionage He also develops open-source tools like OverBoard , a Python dashboard for deep learning experiment monitoring, and advocates for preregistration workshops to improve machine learning research transparency.
Dr. Lin Wang is a Lecturer in Applied Data Science and Signal Processing at Queen Mary University of London (QMUL), affiliated with the School of Electronic Engineering and Computer Science. He leads the Machine Listening Lab and is a member of the Centre for Multimodal AI, Centre for Intelligent Sensing (CIS), and Institute of Coding (IoC). His research focuses on audio-visual signal processing, robotic perception, and machine learning, with applications in healthcare, drone-based sensing, and human activity recognition. Dr. Wang holds a PhD from Dalian University of Technology and has held postdoctoral positions at QMUL, the University of Sussex, and the Alexander von Humboldt Foundation in Germany. Education and Roles: PhD in Signal Processing, Dalian University of Technology (2010) Postdoc at Queen Mary University of London (2014–2017) Postdoc at University of Sussex (2017–2018) Alexander von Humboldt Fellow at University of Oldenburg (2011–2013) Fellow of the Higher Education Academy (UK) Research Interests: Audio-visual signal processing for drones and wearable devices Machine listening and robotic perception Machine learning for healthcare and environmental monitoring Human activity recognition using multimodal sensors Awards and Grants: Early Career Champion on AI&Data, UK Acoustics Network Outstanding Article Award, Frontiers in Computer Science (2022) EPSRC grant: Bioacoustic Monitoring Using Drones (£46,821, 2022–2023) Innovate UK grant: Music Source Separation (£48,144, 2024–2025) Teaching and Students: Dr. Wang teaches Applied Statistics , Website Design and Authoring , and Machine Learning for Visual Data Analysis . He supervises PhD students including Ashish Alex (speech separation), Michael Clayton (drone audition), and Dmitrii Mukhutdinov (audio-visual processing). Labs and Teams: He co-leads the Machine Listening Lab and is part of the Centre for Multimodal AI, focusing on interdisciplinary projects in robotics, acoustics, and AI.
Professor Maja Pantic is a Professor of Affective & Behavioural Computing at the Department of Computing, Faculty of Engineering, Imperial College London. Her research focuses on artificial intelligence, image processing, and audio-visual speech recognition. She leads projects in multimodal systems, including facial analysis, emotion recognition, and speech-driven animation. Affiliations include the AI for Healthcare initiative, the Artificial Intelligence Network, and the Machine Learning Network. Her work addresses challenges in real-time speech enhancement, cross-modal learning, and synthetic data generation. Recent publications emphasize advancements in audiovisual speech synthesis, lip-reading, and emotion-aware systems. She has contributed to datasets like KAN-AV and SEWA DB, advancing research in face analysis and affective computing.
Professor Adrian Hilton is a distinguished faculty member at the University of Surrey, serving as Director of the Centre for Vision, Speech and Signal Processing (CVSSP) and Director of the Surrey Institute for People-Centred AI. He is affiliated with the School of Computer Science and Electronic Engineering and leads the Visual Media Research Lab (V-Lab). His research focuses on pioneering next-generation 4D computer vision technologies that enable machines to understand and model dynamic real-world scenes. Key areas include 3D/4D shape capture, computer vision, machine learning, graphics, and animation for applications in sports analysis, film/TV production, virtual reality, and medical imaging. His work bridges the gap between real and computer-generated imagery, with notable contributions in volumetric capture, motion capture, and free-viewpoint video. Hilton's recent publications demonstrate a strong trend toward multimodal integration, particularly combining audio and visual processing for spatial audio applications, while advancing 4D reconstruction techniques for human performance capture. His work increasingly incorporates transformer architectures and neural rendering techniques for improved illumination estimation, shadow modeling, and multi-view consistency. Scientific Awards and Recognition Two EU IST Innovation Prizes Manufacturing Industry Achievement Award Royal Society Industry Fellowship (2008-2011) Royal Society Wolfson Research Merit Award in 4D Vision (2013-2018) Fellow of the Royal Academy of Engineering (FREng) Fellow of the International Association for Pattern Recognition (FIAPR) Fellow of the Institution of Engineering and Technology (FIET) Hilton actively mentors PhD and post-doctoral researchers through his leadership of CVSSP, which has a grant portfolio exceeding £31M and comprises 170 researchers. He has successfully commercialized several technologies, including systems used by the BBC for sports commentary visualization. His research collaborations span major industry partners including BBC, BT, Sony, Framestore, and The Foundry. He co-founded the G3 Games forum and the CVMP Conference on Visual Media Production, demonstrating strong engagement with the creative industries. Current research projects include the S3A Programme Grant in Future Spatial Audio and InnovateUK's ALIVE project for 360 video reconstruction.
Kevin Chetty is a Professor of Wireless Sensing at University College London (UCL), leading the Urban Wireless Sensing Lab within the Department of Security and Crime Science. His work bridges radar technology, machine learning, and healthcare applications, with a focus on passive sensing systems. Education: PhD in Medical Ultrasound Physics (Imperial College London, 2004-2007), MRes in Image and X-Ray Physics (King's College London, 2003), BSc in Physics (King's College London, 1999) Research spans radar micro-Doppler signature analysis for human behavior classification, software-defined radar development, and integrated communication-sensing systems, with applications in security, healthcare, and smart environments. Recent work emphasizes privacy-preserving technologies and edge processing for real-time operations. Scientific awards include the 2022 IET Radar Systems Best Paper Runner-Up, 2022 IEEE Radar Conference 2nd Place, and 2015 National Instruments Engineering Impact Award. He has received funding from government and industry sectors in telecommunications, IoT, security, and healthcare. Teaching roles: Programme Convener for MSc Crime Science and IEP Minor in Crime and Security Engineering; Module Convener for Security Technologies and Crime Mapping & Spatial Analysis Consultancy: Huawei Technologies (2020-2022), Metropolitan Police Service (2019)
Dr. Armin Mustafa is an Associate Professor in Computer Vision and AI at the University of Surrey, where he holds a prestigious Royal Academy of Engineering Research Fellow position. He is affiliated with the Centre for Vision, Speech and Signal Processing (CVSSP), the School of Computer Science and Electronic Engineering, and the Surrey Institute for People-Centred Artificial Intelligence (PAI). His research focuses on developing AI systems for visual understanding of complex dynamic scenes, with applications in entertainment, autonomous systems, and augmented/virtual reality. Dr. Mustafa completed his PhD in general dynamic scene reconstruction from multi-view videos in 2016 from the University of Surrey under the supervision of Prof. Adrian Hilton. Prior to his doctoral studies, he worked for three years (2010-2013) at Samsung Research Institute in Bangalore, India, in the field of Computer Vision. His research expertise spans Computer Vision, Scene Understanding, 3D/4D Vision, Virtual Reality, Light Fields, Machine Learning, Video Captioning, Augmented Reality, Artificial Intelligence, and Audio-visual Video Understanding. Dr. Mustafa has pioneered advances in 4D vision, NLP, and Scene Understanding over the past decade, with a particular focus on enabling machines to model and interpret real-world environments for socially beneficial applications. His work bridges theoretical advances in computer vision with practical applications in media production, virtual reality, and autonomous systems. Analysis of Dr. Mustafa's recent publications reveals a strong focus on multimodal learning, particularly the integration of audio and visual information for scene understanding. His work spans diverse areas including shadow detection and removal, audio event classification, video captioning, person image generation, and dynamic scene reconstruction. A notable trend is his exploration of transformer architectures for both vision and audio tasks, as well as the application of self-supervised learning techniques to reduce dependency on labeled data. Dr. Mustafa has received numerous prestigious awards: 2018 - Research Fellowship, The Royal Academy of Engineering, UK 2017 - Young Researcher award, CVPR 2016 - Doctoral Consortium grant, CVPR 2015 - BMVA travel grant for ICCV 2014 - Set-Squared Research to Innovator grant 2013 - Overseas Research Scholarship, FEPS, The University of Surrey 2010 - Cadence Silver Medal, Indian Institute of Technology, Kanpur As a dedicated mentor, Dr. Mustafa supervises several PhD students working on cutting-edge topics including multi-person reconstruction, audio-visual scene understanding, and automatic storyboard generation. His research is supported by significant grants including a £15 million UKRI Prosperity Partnership with the BBC (AI4ME), a 5-year Royal Academy of Engineering fellowship (4D Vision for Perceptive Machines), and multiple projects with industry partners such as Figment Productions and Foundry. Dr. Mustafa is an active member of the Centre for Vision, Speech and Signal Processing (CVSSP), one of the world's leading research centers in vision, speech, and signal processing. He also contributes to the Surrey Institute for People-Centred Artificial Intelligence (PAI), where he serves as a Surrey AI Fellow. His work often involves collaboration with industry partners and other academic institutions across Europe.
Dr. Swati Chandna is a Senior Lecturer at the School of Computing and Mathematical Sciences, Birkbeck, University of London. She holds an honorary position as an Honorary Lecturer in Statistics at University College London (UCL) from January 2023 to January 2026. She earned her PhD in Statistics from Imperial College London in 2013. Her research focuses on statistical modeling, network analysis, and bioinformatics, with notable contributions to stochastic networks, single-cell genomic data analysis, and complex-valued signal processing. Teaching responsibilities include modules such as Bayesian Methods, Analysing Data, Statistical Analysis, and Project Applied Statistics. She serves as Admissions Tutor for Graduate Certificate and Diploma in Statistics for Data Science and as School Ethics Lead at Birkbeck. Her work bridges theoretical statistics with practical applications in genomics, environmental modeling, and biomedical research. Dr. Chandna’s recent research explores topics like covariate-driven network estimation, stochastic modeling of genomic data, and bootstrap techniques in source separation. Her publications reflect interdisciplinary collaboration across statistics, computer science, and life sciences.