Andrew Zisserman is a Royal Society Research Professor at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). His research focuses on computer vision, artificial intelligence, and neural networks, with significant contributions to multimodal learning, video understanding, and 3D scene analysis. He leads projects exploring visual-language models, audio-visual synchronization, and clinical imaging applications. Key research areas include: Video analysis and temporal modeling Multimodal systems for sign language translation and action recognition 3D shape estimation and physical property inference Foundation models and cross-modal retrieval Recent work highlights: Developed Flamingo and Tapir models for video-language tasks Advancements in spinal MRI analysis and clinical imaging Leadership in EGO4D and VoxCeleb challenges Honors include Fellowship of the Royal Society (FRS) and the ISSLS Prize in Clinical Science 2023 for spinal analysis innovations. His lab collaborates globally, emphasizing real-world applications in healthcare and autonomous systems.
Dr A. I. Shihab is a Senior Lecturer at Kingston University's Faculty of Engineering, Computing and the Environment, Department of Networks and Digital Media. He teaches programming languages (C++/Java), data structures, web development, and AI/machine learning. His research focuses on affective computing and machine learning applications including: Acoustic event detection in sports environments Audio signal analysis for tennis match modeling Multi-camera visual surveillance systems Medical imaging analysis using fuzzy clustering techniques Publications demonstrate expertise in combining audio/video modalities for sports analytics (tennis rallies, court-shots) and developing Markov models for sound event sequence analysis. Contact: a.shihab@kingston.ac.uk
Andrew Markham is a Professor of Computer Science at the University of Oxford , affiliated with Kellogg College . He leads a research group focusing on Cyber Physical Systems (CPS) , specializing in sensors, signal processing, and machine learning to enable machines to better perceive the physical world. His work emphasizes cross-disciplinary collaboration, notably in wildlife tracking and indoor positioning systems. He has held roles as a Postdoctoral Fellow (2008-2012), Associate Professor (2013), and Full Professor (2021). Education : PhD in Electrical Engineering (University of Cape Town, 2008), BSc (Hons) in Electrical Engineering (2004). Research Interests : Tracking and localization in GPS-denied environments (e.g., underground, indoors), magneto-inductive systems, physics-informed machine learning, and data-driven approaches for noisy sensor data. His projects include wildlife monitoring via wireless sensor networks and mmWave radar for human motion capture. Key Projects : CARACAL acoustic monitoring system, mmPoint dense human tracking, and RandLA-Net for large-scale point cloud segmentation. His work spans robotics, environmental sensing, and biomedical applications. Advising & Grants : Supervises over 30 students and collaborates with industrial partners. Research teams include Cyber Physical Systems, Autonomous Ubiquitous Sensing, and Wildlife Monitoring initiatives. Labs/Teams : Leads the CPS research group, focusing on sensor networks, inertial navigation, and multimodal fusion systems. Collaborates with zoology and earth science disciplines on applied projects.
Cecilia Mascolo is a Professor of Mobile Systems at the University of Cambridge , specifically in the Department of Computer Science and Technology . She co-directs the Centre for Mobile, Wearable System and Augmented Intelligence and is a Fellow of Jesus College, Cambridge . Her research focuses on mobile systems , machine learning for mobile health , and earable technology . She has been awarded prestigious grants such as the ERC Advanced Research Grant (2019-2025) and the EPSRC Open Research Fellowship (2025-2030). Currently on sabbatical at Harvard University , her work bridges systems and machine learning for health applications. Education: PhD in Computer Science from the University of Bologna, Italy. Previous Affiliation: Faculty at University College London before 2008. Her research spans mobile and wearable systems for health and behavior monitoring, focusing on on-device machine learning , uncertainty-aware models , and audio-based diagnostics . Key areas include federated learning , edge computing , and respiratory disease progression analysis via wearables. She explores earable technology for physiological monitoring, gait analysis, and even toothbrushing tracking using in-ear sensors. Her recent publications highlight advancements in earable-based health monitoring , including heart rate estimation , respiratory rate detection , and ECG analysis using machine learning. She emphasizes longitudinal health data from consumer devices, advocating for scalable diagnostics beyond traditional clinical standards. Scientific Awards: ERC Advanced Research Grant EPSRC Open Research Fellowship Best Paper Award - IEEE Percom 10-Year Impact Award - ACM Ubicomp Computer Laboratory Ring Hall of Fame Best Paper Award Student: Andrea Ferlini - ACM SIGMOBILE Doctoral Dissertation Runner-up She leads the Mobile Systems Research Laboratory , mentoring a team of 15 researchers (postdocs and PhD students), and has graduated over 25 PhD students. Her teaching includes Mobile Health courses at the University of Cambridge, and she serves as Director of Studies for Computer Science at Jesus College.
Dr. Lin Wang is a Lecturer in Applied Data Science and Signal Processing at Queen Mary University of London (QMUL), affiliated with the School of Electronic Engineering and Computer Science. He leads the Machine Listening Lab and is a member of the Centre for Multimodal AI, Centre for Intelligent Sensing (CIS), and Institute of Coding (IoC). His research focuses on audio-visual signal processing, robotic perception, and machine learning, with applications in healthcare, drone-based sensing, and human activity recognition. Dr. Wang holds a PhD from Dalian University of Technology and has held postdoctoral positions at QMUL, the University of Sussex, and the Alexander von Humboldt Foundation in Germany. Education and Roles: PhD in Signal Processing, Dalian University of Technology (2010) Postdoc at Queen Mary University of London (2014–2017) Postdoc at University of Sussex (2017–2018) Alexander von Humboldt Fellow at University of Oldenburg (2011–2013) Fellow of the Higher Education Academy (UK) Research Interests: Audio-visual signal processing for drones and wearable devices Machine listening and robotic perception Machine learning for healthcare and environmental monitoring Human activity recognition using multimodal sensors Awards and Grants: Early Career Champion on AI&Data, UK Acoustics Network Outstanding Article Award, Frontiers in Computer Science (2022) EPSRC grant: Bioacoustic Monitoring Using Drones (£46,821, 2022–2023) Innovate UK grant: Music Source Separation (£48,144, 2024–2025) Teaching and Students: Dr. Wang teaches Applied Statistics , Website Design and Authoring , and Machine Learning for Visual Data Analysis . He supervises PhD students including Ashish Alex (speech separation), Michael Clayton (drone audition), and Dmitrii Mukhutdinov (audio-visual processing). Labs and Teams: He co-leads the Machine Listening Lab and is part of the Centre for Multimodal AI, focusing on interdisciplinary projects in robotics, acoustics, and AI.
Professor Annie Mahtani is a Professor of Electroacoustic Composition and Practice at the University of Birmingham's Department of Music. She holds a BMus (Hons), MA, MPhil, and PhD from the University of Birmingham and Birmingham City University. Her work focuses on electroacoustic music, acousmatic composition, multichannel audio, spatialisation, and field recordings, with significant contributions to site-specific installations and cross-disciplinary collaborations. She co-directs the SOUNDkitchen collective and serves on key boards such as the British ElectroAcoustic Network (BEAN) and the British Section of ISCM. Her research explores the sonic identity of environmental sounds, large-scale multichannel composition, and ambisonic audio techniques. As a performer with BEAST (Birmingham Electroacoustic Sound Theatre), she develops live acousmatic performances and soundwalks. Recent projects include Becoming Tree (an immersive audio retelling of a literary work) and Minimum Monument , performed at Birmingham Hippodrome. She actively collaborates with dance and theatre companies like Rosie Kay Dance Company. Professor Mahtani teaches undergraduate and postgraduate modules in electroacoustic composition and is involved in admissions and curriculum leadership. Her works have been performed globally in festivals like Klang! Électroacoustique and Sound + Environment, showcasing her innovative contributions to contemporary music and sound art.
Professor Maja Pantic is a Professor of Affective & Behavioural Computing at the Department of Computing, Faculty of Engineering, Imperial College London. Her research focuses on artificial intelligence, image processing, and audio-visual speech recognition. She leads projects in multimodal systems, including facial analysis, emotion recognition, and speech-driven animation. Affiliations include the AI for Healthcare initiative, the Artificial Intelligence Network, and the Machine Learning Network. Her work addresses challenges in real-time speech enhancement, cross-modal learning, and synthetic data generation. Recent publications emphasize advancements in audiovisual speech synthesis, lip-reading, and emotion-aware systems. She has contributed to datasets like KAN-AV and SEWA DB, advancing research in face analysis and affective computing.
Professor Adrian Hilton is a distinguished faculty member at the University of Surrey, serving as Director of the Centre for Vision, Speech and Signal Processing (CVSSP) and Director of the Surrey Institute for People-Centred AI. He is affiliated with the School of Computer Science and Electronic Engineering and leads the Visual Media Research Lab (V-Lab). His research focuses on pioneering next-generation 4D computer vision technologies that enable machines to understand and model dynamic real-world scenes. Key areas include 3D/4D shape capture, computer vision, machine learning, graphics, and animation for applications in sports analysis, film/TV production, virtual reality, and medical imaging. His work bridges the gap between real and computer-generated imagery, with notable contributions in volumetric capture, motion capture, and free-viewpoint video. Hilton's recent publications demonstrate a strong trend toward multimodal integration, particularly combining audio and visual processing for spatial audio applications, while advancing 4D reconstruction techniques for human performance capture. His work increasingly incorporates transformer architectures and neural rendering techniques for improved illumination estimation, shadow modeling, and multi-view consistency. Scientific Awards and Recognition Two EU IST Innovation Prizes Manufacturing Industry Achievement Award Royal Society Industry Fellowship (2008-2011) Royal Society Wolfson Research Merit Award in 4D Vision (2013-2018) Fellow of the Royal Academy of Engineering (FREng) Fellow of the International Association for Pattern Recognition (FIAPR) Fellow of the Institution of Engineering and Technology (FIET) Hilton actively mentors PhD and post-doctoral researchers through his leadership of CVSSP, which has a grant portfolio exceeding £31M and comprises 170 researchers. He has successfully commercialized several technologies, including systems used by the BBC for sports commentary visualization. His research collaborations span major industry partners including BBC, BT, Sony, Framestore, and The Foundry. He co-founded the G3 Games forum and the CVMP Conference on Visual Media Production, demonstrating strong engagement with the creative industries. Current research projects include the S3A Programme Grant in Future Spatial Audio and InnovateUK's ALIVE project for 360 video reconstruction.
Kevin Chetty is a Professor of Wireless Sensing at University College London (UCL), leading the Urban Wireless Sensing Lab within the Department of Security and Crime Science. His work bridges radar technology, machine learning, and healthcare applications, with a focus on passive sensing systems. Education: PhD in Medical Ultrasound Physics (Imperial College London, 2004-2007), MRes in Image and X-Ray Physics (King's College London, 2003), BSc in Physics (King's College London, 1999) Research spans radar micro-Doppler signature analysis for human behavior classification, software-defined radar development, and integrated communication-sensing systems, with applications in security, healthcare, and smart environments. Recent work emphasizes privacy-preserving technologies and edge processing for real-time operations. Scientific awards include the 2022 IET Radar Systems Best Paper Runner-Up, 2022 IEEE Radar Conference 2nd Place, and 2015 National Instruments Engineering Impact Award. He has received funding from government and industry sectors in telecommunications, IoT, security, and healthcare. Teaching roles: Programme Convener for MSc Crime Science and IEP Minor in Crime and Security Engineering; Module Convener for Security Technologies and Crime Mapping & Spatial Analysis Consultancy: Huawei Technologies (2020-2022), Metropolitan Police Service (2019)
Dr. Armin Mustafa is an Associate Professor in Computer Vision and AI at the University of Surrey, where he holds a prestigious Royal Academy of Engineering Research Fellow position. He is affiliated with the Centre for Vision, Speech and Signal Processing (CVSSP), the School of Computer Science and Electronic Engineering, and the Surrey Institute for People-Centred Artificial Intelligence (PAI). His research focuses on developing AI systems for visual understanding of complex dynamic scenes, with applications in entertainment, autonomous systems, and augmented/virtual reality. Dr. Mustafa completed his PhD in general dynamic scene reconstruction from multi-view videos in 2016 from the University of Surrey under the supervision of Prof. Adrian Hilton. Prior to his doctoral studies, he worked for three years (2010-2013) at Samsung Research Institute in Bangalore, India, in the field of Computer Vision. His research expertise spans Computer Vision, Scene Understanding, 3D/4D Vision, Virtual Reality, Light Fields, Machine Learning, Video Captioning, Augmented Reality, Artificial Intelligence, and Audio-visual Video Understanding. Dr. Mustafa has pioneered advances in 4D vision, NLP, and Scene Understanding over the past decade, with a particular focus on enabling machines to model and interpret real-world environments for socially beneficial applications. His work bridges theoretical advances in computer vision with practical applications in media production, virtual reality, and autonomous systems. Analysis of Dr. Mustafa's recent publications reveals a strong focus on multimodal learning, particularly the integration of audio and visual information for scene understanding. His work spans diverse areas including shadow detection and removal, audio event classification, video captioning, person image generation, and dynamic scene reconstruction. A notable trend is his exploration of transformer architectures for both vision and audio tasks, as well as the application of self-supervised learning techniques to reduce dependency on labeled data. Dr. Mustafa has received numerous prestigious awards: 2018 - Research Fellowship, The Royal Academy of Engineering, UK 2017 - Young Researcher award, CVPR 2016 - Doctoral Consortium grant, CVPR 2015 - BMVA travel grant for ICCV 2014 - Set-Squared Research to Innovator grant 2013 - Overseas Research Scholarship, FEPS, The University of Surrey 2010 - Cadence Silver Medal, Indian Institute of Technology, Kanpur As a dedicated mentor, Dr. Mustafa supervises several PhD students working on cutting-edge topics including multi-person reconstruction, audio-visual scene understanding, and automatic storyboard generation. His research is supported by significant grants including a £15 million UKRI Prosperity Partnership with the BBC (AI4ME), a 5-year Royal Academy of Engineering fellowship (4D Vision for Perceptive Machines), and multiple projects with industry partners such as Figment Productions and Foundry. Dr. Mustafa is an active member of the Centre for Vision, Speech and Signal Processing (CVSSP), one of the world's leading research centers in vision, speech, and signal processing. He also contributes to the Surrey Institute for People-Centred Artificial Intelligence (PAI), where he serves as a Surrey AI Fellow. His work often involves collaboration with industry partners and other academic institutions across Europe.
Dr. Swati Chandna is a Senior Lecturer at the School of Computing and Mathematical Sciences, Birkbeck, University of London. She holds an honorary position as an Honorary Lecturer in Statistics at University College London (UCL) from January 2023 to January 2026. She earned her PhD in Statistics from Imperial College London in 2013. Her research focuses on statistical modeling, network analysis, and bioinformatics, with notable contributions to stochastic networks, single-cell genomic data analysis, and complex-valued signal processing. Teaching responsibilities include modules such as Bayesian Methods, Analysing Data, Statistical Analysis, and Project Applied Statistics. She serves as Admissions Tutor for Graduate Certificate and Diploma in Statistics for Data Science and as School Ethics Lead at Birkbeck. Her work bridges theoretical statistics with practical applications in genomics, environmental modeling, and biomedical research. Dr. Chandna’s recent research explores topics like covariate-driven network estimation, stochastic modeling of genomic data, and bootstrap techniques in source separation. Her publications reflect interdisciplinary collaboration across statistics, computer science, and life sciences.
Professor Stefan Bleeck is a Professor of Hearing Science and Technology at the University of Southampton, leading the Hearing and Balance Centre and directing the Institute of Sound and Vibration Research (ISVR). His research focuses on the intersection of hearing science, audiology, and signal processing, with specialties in bio-inspired auditory modeling, speech intelligibility in noise, cochlear implants, and auditory evoked potentials. He holds a PhD in computational neuroscience and has held roles including Head of the Hearing and Balance Centre. Awards include Vice-Chancellor's Teaching Awards (2009) and Google Research Awards (2012). Education: Diploma in Physics (University of Darmstadt, 1995), PhD in Computational Neuroscience (University of Darmstadt, 2000). Research spans experimental, computational, and clinical approaches to improve hearing aids and cochlear implants. Active projects include developing speech enhancement algorithms, antiphasic speech tests for hidden hearing loss, and neural-space speech processing. Supervises multiple PhD students in engineering and computer science. Publications highlight advancements in speech enhancement, bio-inspired models, and cross-linguistic hearing tests. Collaborates with institutions like Google and the European Union on projects funded by EPSRC, Cancer Research UK, and others. His work aims to enhance speech understanding for hearing-impaired individuals through innovative signal processing and auditory modeling.
Professor Guy Brown is Chair of Computer Science at the University of Sheffield's School of Computer Science. He holds a BSc in Applied Science (1984), PhD in Computer Science (1992), and MEd in Teaching and Learning (1997). His research focuses on Computational Auditory Scene Analysis (CASA), noise-robust speech recognition, auditory modeling, and binaural processing. Research interests include: Machine hearing systems for sound source separation Reverberation-robust speech processing Auditory scene analysis models for normal/impaired hearing Applications in robotics and healthcare technologies Publication trends show recent focus on deep learning approaches for biomedical applications including sleep apnea detection, respiratory sound analysis, and multimodal health monitoring systems using neural networks. Honors include: University Senate Award for Excellence in Teaching (2014) Microsoft Software Engineering Innovation Award (2013) He leads doctoral supervision for 15+ students and has secured research funding from EPSRC, Innovate UK, EU FP7, and AHRC. Manages the Speech and Hearing research group and has held visiting positions at international institutions including LIMSI-CNRS and ATR Japan.
Professor Jon Barker is a faculty member at the University of Sheffield , where he holds a Personal Chair in the School of Computer Science . He leads the Speech and Hearing (SpandH) research group and co-founded the CHiME international workshop series on robust speech recognition. Education : PhD in Computer Science (University of Sheffield, 1999); BA in Electrical and Information Sciences (Cambridge University). Research Focus : His work bridges machine listening and human auditory perception , with key contributions to noise-robust speech recognition , speech intelligibility prediction , and hearing aid signal processing for speech and music. Recent projects include the Clarity Challenges and Cadenza Challenges , large-scale machine learning initiatives to improve accessibility for hearing-impaired users. Publication Trends : Recent articles emphasize machine learning for hearing aid optimization , dysarthric speech recognition , audio-visual integration , and music demixing algorithms . Collaborations span speech processing, psychoacoustics, and biomedical engineering. Scientific Awards : EURASIP Best Paper Award (2009) ISCA Best Paper Award (2008) Grants and Leadership : He has secured major EPSRC grants including EnhanceMusic (2022-2026) and Challenges to Revolutionise Hearing Device Processing (2019-2025). He co-led the TAPAS Marie Curie Training Network (2017-2022) and led projects like AV-COGHEAR (2015-2018) and CHiME (2009-2012). Labs and Teams : Barker collaborates closely with the Speech and Hearing Research Group and contributes to international initiatives like the CHiME Workshop . His lab develops open datasets such as the Clarity Speech Corpus and Audio-Visual Lombard Corpus .
Dr. Ke Chen is a Senior Lecturer in the School of Computer Science at The University of Manchester, leading the Machine Learning and Perception (MLP@UoM) Lab. His research focuses on machine learning, deep learning, reinforcement learning, and their applications in intelligent systems, computer vision, and audio/speech processing. He has supervised over 50 PhD students and holds editorial roles in journals like Neural Networks and IEEE Transactions on Neural Networks . Dr. Chen has received awards such as the NSFC Distinguished Principal Young Investigator Award (2001) and JSPS Research Award (1993). Education: PhD in Computer Science (1990), with academic positions at institutions including Peking University, The Ohio State University, and Kyushu Institute of Technology. His professional activities include roles in IEEE Computational Intelligence Society committees and external examiner roles at universities like Essex. Research interests span computational cognitive systems, biometric authentication, and video game AI. Key contributions include advancements in deep architectures, reinforcement learning, and speaker-specific feature extraction. His lab, MLP@UoM, explores topics like explainable AI and transfer learning. Recent publications (2024) include work on goal-conditioned reinforcement learning and bias-resilient algorithms. He is actively involved in international conferences, serving as a keynote speaker and program committee member.