Oliver Kroemer is an Associate Professor at Carnegie Mellon University's Robotics Institute (RI), affiliated with the Intelligent Autonomous Manipulation (IAM) Lab. His research focuses on enabling robots to learn versatile manipulation skills through lifelong frameworks, with applications in elder care, environmental maintenance, and hazardous operations. Developed methods for robot learning via physical interaction and reinforcement learning Created representations for contact states and motor primitives to improve skill generalization Research Interests: Spanning robot learning, tactile sensing, force-velocity control, and lifelong skill acquisition. Projects include Agile and Dynamic Interactions for Mobile Manipulation and Integrated Planning and Learning (Pillar project). Scientific Awards: Finalist, Georges Giralt Ph.D. Award (2015) Education: Masters & Bachelors in Engineering, University of Cambridge (2008) Ph.D., Technische Universitaet Darmstadt (2014) Students & Affiliates: Current PhD: Mark Lee, Sarvesh Patil, Saumya Saxena, Yunus Seker, Zilin Si Past PhD: Alex LaGrassa, Tabitha Lee, Qiao Liang, Shivam Vats, Kevin Zhang
WANG Ye is an Associate Professor in the Department of Computer Science at the School of Computing, National University of Singapore (NUS). He holds a PhD in Information Technology from Tampere University of Technology, Finland, and has been a tenured faculty member at NUS since 2002, following his industry research role at Nokia Research Center. He is the director of the Sound and Music Computing Lab at NUS, leading cutting-edge research in AI-driven music and health technologies. PhD, Information Technology, Tampere University of Technology, Finland (2002) MSc, Telecommunications, Braunschweig University of Technology, Germany (1993) BSc, Telecommunications, South China University of Technology, China (1983) His research is centered on Sound and Music Computing for Human Health and Potential (SMC4HHP) , with a focus on eHealth, eLearning, mobile/wearable computing, and music information retrieval. His work spans AI for stroke rehabilitation, language learning through singing, singing voice synthesis, and automatic music transcription. He has pioneered systems like SLIONS (language learning via karaoke), CocoLyricist (AI co-creation for stroke recovery), and SinTechSVS (expressive singing voice synthesis). The latest articles highlight a strong trend in AI-driven music and health technologies , particularly in controllable lyric generation, singing voice synthesis, automatic pronunciation assessment, and multimodal music transcription. The research increasingly integrates large language models, explainable AI, fairness, and real-world deployment, reflecting a shift from theoretical exploration to practical, human-centered applications in healthcare and education. Dr. Wang has received numerous scientific honors, including: Best Paper Awards at ACM MM, ISMIR, IEEE ISM, and CHI First Prize, Asia Pacific Assistive, Rehabilitative, and Therapeutic Technologies Challenge (2015) Faculty Teaching Excellence Award, NUS School of Computing (2024) Top Paper Award, ACM Multimedia 2022 AI in Medicine Collaborative Grant for CocoLyricist project He has supervised over 11 PhD and 20 MComp students and is currently guiding six PhD candidates. His grants come from MOE, NRF, A*STAR, Nokia, and Smule. He has served as General Chair of ISMIR2017 and TPC Co-Chair of ICOT2017, and is on the editorial boards of IEEE Transactions on Multimedia and Journal of New Music Research. He has also developed and taught the first course on Sound and Music Computing in Singapore. Dr. Wang leads the Sound and Music Computing Lab (SMC Lab) , a multidisciplinary team exploring the synergy of music computing, AI, mobile technology, and cloud systems for health and education. The lab actively collaborates with medical institutions such as NUS Yong Loo Lin School of Medicine, Singapore General Hospital, and Harvard Medical School, and is currently working on projects in AI-supported language learning, stroke rehabilitation, and intelligent music interfaces.
Juan Pablo Bello is Professor of Music Technology and Computer Science & Engineering at NYU. He directs the Music and Audio Research Lab (MARL) and the Center for Urban Science and Progress (CUSP). He holds a Ph.D. in Electronic Engineering from Queen Mary University of London (2003) and a B.Eng. from Universidad Simón Bolívar (1998). Research integrates signal processing, machine learning and acoustics to analyze music/audio. Current projects include urban sound mapping, music structure discovery, and audio event detection. Funded by NSF and industry partners, with 100+ publications in IEEE/ACM venues. Honors include an NSF CAREER Award and Fulbright scholarship. He teaches graduate courses in music informatics and fosters industry-academia partnerships in audio AI.
Professor Carlo Harvey is a creative technologist at the School of Digital Arts (SODA), Manchester Metropolitan University. His interdisciplinary research merges games , machine learning , virtual production , and cultural heritage reinterpretation . He leads industry collaborations with entities like Jaguar Land Rover and Epic Games, focusing on AI-driven interactive audio, real-time visualization, and accessibility solutions. Award-winning projects : TIGA, Innovate UK, and Epic Games MegaGrant for Accession Industry partnerships : Automotive sector, cultural institutions His research spans human-computer interaction , multisensory virtual environments , and acoustic-visual cross-modal perception . Recent publications address robotic simulations, motion alignment, and haptic feedback systems. Scientific recognition : TIGA Award, Innovate UK Funding, Epic Games MegaGrant Advocacy : Digital inclusion, creative collaboration, social impact of technology
Professor Simon Godsill MA PhD FIET FIEEE is a University Professor of Statistical Signal Processing in the Department of Engineering at the University of Cambridge. He heads a research team specializing in statistical signal processing, digital audio restoration, and Bayesian inference. His work addresses the processing and analysis of digital speech, audio, tracking systems, and financial datasets, with a focus on probabilistic modeling and computational methods. Research interests include statistical signal processing , degraded signal restoration , and Bayesian computational methods . Recent publications emphasize Gaussian processes, variational inference, and multi-object tracking for applications in audio enhancement and financial data analysis. He co-founded the audio remastering company CEDAR Audio Ltd in 1988. Scientific awards: Fellow of the Institution of Engineering and Technology (FIET) Fellow of the Institute of Electrical and Electronics Engineers (FIEEE) Outside academia, he enjoys singing, cricket, piano/organ playing, and running. His team at Cambridge's Engineering department focuses on robust tracking algorithms and signal enhancement techniques.
University of Illinois Urbana-ChampaignUnited States
Pasquale Bottalico serves as Associate Professor in the Department of Speech and Hearing Science at the University of Illinois, with dual appointments as Associate Professor at the Center for Latin American and Caribbean Studies and Affiliate Faculty in the School of Music. His unique interdisciplinary profile bridges engineering, music performance, and speech science, reflecting his dual academic training and professional artistry. His educational foundation includes: Bachelor's in Telecommunications Engineering from Univeristà Mediterranea di Reggio Calabria, Italy Concurrent Opera Singing degree from F. Cilea Music Academy, Reggio Calabria Master's in Telecommunications Engineering from Politecnico di Torino, Italy Ph.D. in Metrology specializing in acoustics measurement uncertainty and classroom acoustics Dr. Bottalico's research centers on vocal load quantification and professional voice techniques , with significant contributions to understanding vocal fatigue in teachers and singers. His work spans Speech Intelligibility in educational environments, Room Acoustics for performance and learning spaces, and Musical Acoustics of historical vocal styles. A distinctive thread throughout his research examines how acoustic conditions modulate voice production and perception, increasingly incorporating virtual reality and bone conduction technologies for innovative assessment and intervention approaches. His Colombian vocal health study demonstrates cross-cultural applications of his work. Analysis of his 2023-2025 publications reveals three dominant research trajectories: (1) The impact of noise and dysphonia on children's speech processing in educational settings, using multimodal assessment including EEG; (2) Virtual reality applications for voice production research and therapeutic intervention; (3) Cross-cultural validation of vocal fatigue metrics and development of biofeedback systems. His work consistently bridges engineering precision with clinical applicability, particularly for professional voice users in challenging acoustic environments. No scientific awards were documented in the available information. While specific advising relationships aren't detailed, his research collaborations span international institutions including Colombian and Italian universities, suggesting graduate mentorship in interdisciplinary projects. No grant information was provided, though his systematic reviews and cross-cultural studies imply externally funded research activities. Though no dedicated laboratory is specified, his virtual reality voice studies and acoustic parameter assessments suggest affiliations with audio engineering facilities and voice clinics, likely through the Speech and Hearing Science department's research infrastructure.
Prof. Roger Wattenhofer is a Full Professor at the Department of Information Technology and Electrical Engineering, ETH Zurich, Switzerland, and Deputy head of the Computer Engineering and Networks Lab. He holds a doctorate in Computer Science from ETH Zurich (1998) and has held positions at Brown University and Microsoft Research before returning to ETH. His research focuses on distributed computing, wireless networks, and algorithmic systems design, with contributions to Byzantine agreement protocols, blockchain technologies, and neural network architectures. He teaches courses such as Distributed Systems and Computational Thinking. Education: Ph.D. in Computer Science (ETH Zurich, 1998). Research interests include distributed systems, network algorithms, and the intersection of machine learning with distributed computing. His work spans both theoretical foundations and practical implementations, addressing challenges in fault tolerance, consensus mechanisms, and algorithmic efficiency. Recent publications explore topics like adversarial robustness in voting systems, privacy in reinforcement learning, and generative music models. He actively contributes to open-source frameworks and benchmarks for neural algorithmic reasoning.
CHAN Mun Choon is a Professor at the School of Computing, National University of Singapore (NUS) , where he directs the NUS-NCS Joint Laboratory for Cyber Security . He previously worked at Bell Labs (1997-2003) and holds a PhD from Columbia University (1997). His research spans systems and networking with specific interests in mobile computing, software-defined networking, and cyber-physical systems . PhD, Electrical Engineering (1997), Columbia University M.Phil., Electrical Engineering (1993), Columbia University MS, Electrical Engineering (1993), Columbia University BS, Computer & Electrical Engineering (1990), Purdue University His recent work focuses on 5G network architecture , data center fault debugging , and energy-efficient mobile sensing . He has published over 100 papers and holds 7 US patents , including cache-based compaction techniques with 210+ citations. His projects include fronthaul slicing for 5G, network-wide packet history frameworks, and participatory indoor localization. Scientific recognition includes: Best Paper Awards: IEEE ICNP 2019, ACM SOSR 2019, ICDCN 2016 Best Demo: IPSN 2016 Distinguished Member, INFOCOM TPC (2016, 2020, 2021) He serves as Vice-Dean, Graduate Studies and Vice-Dean, Academic Affairs at NUS Computing, and has graduated 21 PhD students . His lab develops solutions for network security , latency-sensitive applications , and mobile sensing .
Najim Dehak is an Associate Professor in the Department of Electrical and Computer Engineering at Johns Hopkins University, part of the Whiting School of Engineering. His research focuses on machine learning applied to speech processing, audio classification, and health applications. He is renowned for developing the I-vector representation for speaker recognition, introduced in 2008 during a workshop at Johns Hopkins’ Center for Language and Speech Processing. Prior to this role, he was a research scientist at MIT’s Computer Science and Artificial Intelligence Laboratory. Dehak holds a PhD from the School of Advanced Technology in Montreal (2009). He is a Senior Member of IEEE and contributes to the IEEE Speech and Language Technical Committee. His work bridges AI, healthcare, and signal processing, with notable contributions to neurodegenerative disease detection via speech and handwriting analysis. Research interests include adversarial attacks on speech systems, multimodal biomarker discovery, and robust speech processing across demographics. His lab’s tools, like the Hermespeech Recorder, enable scalable data collection for clinical and research applications. Education: PhD in Advanced Technology (2009), Montreal Affiliations: Johns Hopkins University, IEEE Labs/Teams: Center for Language and Speech Processing (CLSP) His recent work explores AI’s role in aging research, including Alzheimer’s and Parkinson’s disease detection through speech, eye tracking, and handwriting analysis. Ongoing projects address fairness in speaker verification and robustness against adversarial attacks in ASR systems.
Nasir Memon is a Professor of Computer Science and Engineering at the New York University Tandon School of Engineering and concurrently serves as the Dean of Engineering at NYU Shanghai. He has been a faculty member at NYU Tandon since September 1998. Memon is a co-founder of NYU's Center for Cyber Security (CCS) and NYU Abu Dhabi, and the founder of key initiatives such as the OSIRIS Lab, CyberSecurity Awareness Week (CSAW), the NYU Tandon Bridge program, and the Cyber Fellows program. His work focuses on advancing cybersecurity education and addressing systemic biases in AI-driven systems. Education: Ph.D., Computer Science, University of Nebraska Master of Science, Mathematics, Birla Institute of Technology and Science (BITS), Pilani Bachelor of Engineering, Chemical Engineering, BITS, Pilani Research Interests: Media Forensics and Authentication Biometric Security and Privacy Data Compression and Privacy-Preserving Techniques Network Security and Incident Response AI Ethics and Fairness in Machine Learning Cybersecurity Education and Workforce Development Awards and Honors: IEEE Fellow (2010) SPIE Fellow (2014) Jacobs Excellence in Education Award (2002) NSF CAREER Award (1997) Advising and Grants: Advises Ph.D. students like Anubhav Jain and Govind Mittal, and mentors Master’s students such as Rishit Dholakia. Recipient of NSF grants and funding from NYU Abu Dhabi and Indiana University Bloomington for projects like computational tools for fact-checking and AI-driven bias mitigation. Labs and Teams: Directs the OSIRIS Lab, a leading research group in cybersecurity and AI. Leads initiatives at the NYU Center for Cybersecurity (CCS) and collaborates with NYU Abu Dhabi’s Center for Cyber Security.
Xavier Serra is a Full Professor at the Department of Engineering at Universitat Pompeu Fabra (UPF), Barcelona. He is the founder and director of the Music Technology Group (MTG), and leads the UPF-BMAT Chair on AI and Music. He also coordinates the Master in Sound and Music Computing and serves as President of the Phonos Foundation. His research focuses on audio signal processing, sound and music computing, and computational musicology, emphasizing open science and open innovation. Education: BSc in Biology, University of Barcelona (1981) Master in Music, Florida State University (1983) PhD in Computer Music, Stanford University (1989) Research Interests: Audio Signal Processing Data-Driven and Knowledge-Driven Methodologies Music Information Retrieval Cultural Music Analysis (e.g., Carnatic/Turkish/Andalusian Music) Music Education Technology Notable Projects: CompMusic (ERC Advanced Grant, 2010-2017): Multicultural computational music analysis Open datasets: Freesound, Saraga, FSD50K Technologies: Reactable, Vocaloid, Essentia API Recent Trends in Articles: Focus on AI-driven audio processing (neural fingerprints, generative models), cross-cultural music analysis, and explainable music difficulty estimation. Awards: ERC Advanced Grant (2010) for CompMusic Project. Labs/Teams: Director of MTG, Phonos Foundation, and UPF-BMAT Chair. Active in open-source projects and international collaborations.
Changxi Zheng is an Associate Professor in the Department of Computer Science at Columbia University's School of Engineering and Applied Science (SEAS). He directs Columbia's Computer Graphics Group (C2G2) within the Columbia Vision and Graphics Center (CVGC). After receiving his PhD from Cornell University, he joined the faculty of Computer Science Department at Columbia, where he has established himself as a leading researcher in computer graphics and scientific computing. Dr. Zheng's research spans multiple areas of applied computer science with a particular focus on computer graphics and scientific computing. His work centers around developing numerical models for simulating physical phenomena involving complex motions such as fluids, bubbles, and thin rods, along with their resulting acoustic waves. Leveraging computational insights from these models, he devises methods for improving tangible object creation, enabling novel human-computer interactions, and developing software tools for acoustic and photonic devices. His research has attracted significant public interest and media coverage, including projects like FontCode, AirCode, and Computational Metallophone Design. His recent publications reveal a strong interdisciplinary approach, bridging computer graphics, physics simulation, machine learning, and hardware design. His work demonstrates consistent innovation in computational methods for simulating physical phenomena and applying these techniques to practical problems in 3D printing, acoustic modeling, and interactive systems. The breadth of his research spans from fundamental physics-based simulations to practical applications in industry. Columbia SEAS Dean's Fellow (for advised students) NSF Graduate Research Fellow (for Ruilin Xu) Snap Research Fellow (for Rundi Wu) CKGSB Fellow (for Yun Fei) Adobe Research Fellow (for Gabriel Cirio) Marie Sklodowska-Curie Individual Fellow (for Rundi Wu) Best Paper Award at ACM International Conference on Multimedia (ACMMM), 2019 Dr. Zheng actively mentors a diverse group of students, including current PhD candidates and postdoctoral researchers. His research group has received support from various sources that enable their innovative work in computational graphics and physics-based simulation. He has supervised numerous successful students who have gone on to positions at leading technology companies including Adobe, Tencent, Facebook, and academic institutions. As director of Columbia's Computer Graphics Group (C2G2) within the Columbia Vision and Graphics Center (CVGC), Dr. Zheng leads a vibrant research team focused on advancing the state of the art in computer graphics, physics-based simulation, and their applications. The group maintains strong collaborations with industry partners and academic institutions worldwide, fostering an environment of innovation and practical application of theoretical concepts.
Elena Maria Baralis is a Full Professor at the Department of Control and Computer Science (DAUIN) at the Polytechnic University of Turin. She serves as Pro-Rector, member of the Board of Directors (without voting rights), member of the Academic Senate (without voting rights), and coordinator of the University's Permanent Observatory for Monitoring the Academic Sector. She chairs the Control and Computer Engineering Department and previously chaired the Computer Engineering School from October 2012 to October 2018. Her research interests focus on database systems and data mining, specifically explainable AI, bias detection in data analytics, and machine learning algorithms for big data. Her work spans various application domains including predictive maintenance, Industry 4.0, and healthcare. Recent publications demonstrate her expertise in speech processing, bias mitigation, and innovative neural network architectures like Kolmogorov-Arnold Networks. Her research output shows a clear trend toward addressing fairness and explainability in AI systems while exploring novel approaches to speech and language understanding. Professor Baralis has received significant recognition including becoming a Fellow of the Academy of Sciences of Turin in 2017. She has served as Editor-in-Chief for IEEE Internet of Things Journal (2016-2019) and Knowledge and Information Systems (2014-present). She actively mentors doctoral students including Claudio Savelli (researching Machine Unlearning), Eleonora Poeta, Giuseppe Gallipoli, Alkis Koudounas, and others. Her research is supported by numerous projects including AI4CTI (Artificial Intelligence for Cyber Threat Intelligence, 2025-2028), Smart manufacturing driven by Machine Learning in Industry 4.0 (2019-2020), and I-REACT (2016-2019).
Mark Gales is Professor of Information Engineering at the University of Cambridge and an Official Fellow at Emmanuel College. He is currently on sabbatical leave for the 2024/25 academic year. Prior to his academic career, he worked as a consultant at Roke Manor Research Ltd, developing radar systems, before transitioning to speech and language processing. PhD in 'Model-Based Techniques for Robust Speech Recognition' (University of Cambridge, 1995) BA in Electrical and Information Sciences (University of Cambridge, 1988) His research focuses on speech and language processing , particularly in automated language assessment and low-resource speech technology . He leads the Automated Language Teaching and Assessment (ALTA) Institute , which collaborates with Cambridge University Press & Assessment (CUP&A) to develop commercial tools like Linguaskill and Speak & Improve . These platforms provide automated spoken/written assessment for millions of users globally. Recent publications highlight his work in LLM-driven speech processing , including adversarial attacks on foundation models, end-to-end spoken error correction, and uncertainty estimation frameworks. His team's research spans multilingual capabilities, with deployments in languages ranging from Dholuo to Tok Pisin . Awards : IEEE Fellow, ISCA Fellow Leadership : Fellows' Steward at Emmanuel College Mark has contributed extensively to Hidden Markov Model (HMM) applications in speech recognition, which underpinned early automatic speech systems. His work now bridges LLM-based language assessment with cross-lingual transfer learning and robustness testing for real-world deployments.
University of California , Santa Barbara (UCSB)United States
Yao Qin is an Assistant Professor in the Department of Electrical and Computer Engineering at the University of California, Santa Barbara (UCSB), with dual affiliation in the Department of Computer Science. She concurrently serves as Co-Director of the REAL AI Initiative at UCSB and holds a Senior Research Scientist position at Google DeepMind, where she contributes to the Gemini Multimodal project. Her academic credentials include a PhD in Computer Science and Engineering from the University of California, San Diego (advised by Prof. Garrison W. Cottrell) and a BS in Electrical Engineering from Dalian University of Technology. During her doctoral studies, she completed internships with pioneering researchers Geoffrey Hinton and Ian Goodfellow. Dr. Qin's research program centers on machine learning robustness, with emphasis on adversarial robustness, out-of-distribution generalization, and fairness. She develops reliable AI systems specifically for healthcare applications, with diabetes management as a primary focus. Her lab explores critical themes including AI safety in multimodal models and diabetes-specific AI solutions, particularly exercise metabolism modeling and glycemic effect prediction. Recent publications reveal a strong trajectory in robust machine learning with cross-domain applications. Her work consistently bridges theoretical robustness concepts with practical healthcare implementations, particularly in diabetes care. Key publication venues include CVPR, ICML, NeurIPS, and ICLR, with notable contributions to out-of-distribution detection, adversarial transfer learning, and multimodal AI safety. Her distinguished recognition includes: EECS Rising Star at MIT (2021) UCSB Regents' Junior Faculty Fellowship Award Helmsley Charitable Trust award for Type 1 diabetes research UCSB Faculty Research Grant American Diabetes Association Abstract Award (ADA-2025) Dr. Qin actively mentors four PhD students—Mehak Dhaliwal, Andong Hua, Kenan Tang, and Youngseok Yoon—on projects spanning LLMs for diabetes, multimodal robustness, and generative time-series modeling. Her research is funded by the Helmsley Charitable Trust and UCSB, with recent grants supporting exercise-specific AID algorithms for diabetes management. As Co-Director of the REAL AI Initiative, she leads a research ecosystem focused on developing reliable artificial intelligence. Current lab activities include organizing workshops at NeurIPS-2024 (AdvML-Frontiers and AIM-FM) and developing next-generation diabetes management tools through collaborations with medical institutions.