Aykut Erdem is an Associate Professor of Computer Engineering at Koç University, affiliated with the KUIS AI Center. He earned his PhD, MSc, and BSc in Computer Engineering from Middle East Technical University (METU), with visiting researcher experiences at Virginia Tech (2004) and MIT (2007). His research focuses on learning-based approaches for visual data understanding, including image editing, visual saliency estimation, and vision-language integration. Research Interests: Vision and Graphics, Machine Learning, Artificial Intelligence, Computer Vision, Natural Language Processing, Generative Artificial Intelligence. Recent work includes text-guided image/video editing, diffusion models for object removal, and GAN-based frameworks for domain adaptation. Scientific Awards: Young Scientist Award (BAGEP 2021) by Science Academy in Computer Engineering Best Paper Award at 5th Multimodal Learning and Applications Workshop (2022) Collaborations and Funding: Principal investigator for TUBITAK 1001 project on generative AI for visual data (2021-2024), Adobe Research Gift (2023), and co-investigator for multiple TUBITAK grants. Serves as Associate Editor for IEEE Transactions on Image Processing (2022-present).
Berrak Sisman is an Assistant Professor in the Department of Electrical and Computer Engineering at Johns Hopkins University, affiliated with the Data Science and AI Institute and the Center for Language and Speech Processing (CLSP). She leads the Speech & Machine Learning Lab (SmILe Lab), focusing on AI-driven speech technologies. She received her PhD from the National University of Singapore in 2020 and was previously a tenure-track faculty member at the University of Texas at Dallas (2022–2024). Research Interests: Her work spans artificial intelligence, speech synthesis, voice conversion, emotion analysis in speech, medical speech applications, and secure speech technology. She develops neural models for expressive and adaptive speech processing. Publications: Her recent articles (2024–2025) emphasize speech emotion recognition, zero-shot prosody control, accent conversion, and disentangled representations in TTS, reflecting a focus on cross-modal learning, robustness, and real-world applications. Awards & Grants: NSF CAREER Award (2024) Amazon Faculty Research Award (2022) Singapore Ministry of Education Award (2021) A*STAR Singapore International Graduate Award (2016–2020) Leadership: She directs the SmILe Lab, recruiting PhD/Master’s students for projects in neural speech modeling. Her grants include NSF and Amazon funding for voice conversion and emotion synthesis research.
Erkut Erdem is a Professor in the Department of Computer Engineering at Hacettepe University, where he leads the Computer Vision Laboratory (HUCVL). His research focuses on computer vision and machine learning, particularly on incorporating different kinds of context (spatial, temporal and cross-modal) into visual processing across all levels from low to high-level vision. He received his Ph.D. (2008), M.Sc. (2003), and B.Sc. (2001) from Middle East Technical University. Prior to joining Hacettepe University in 2010, he completed a post-doctoral fellowship at Ecole Nationale Supérieure des Télécommunications (2009-2010) and held visiting researcher positions at UCLA (2007) and Virginia Tech (2004). His current research interests include Visual Saliency Prediction, Automatic Image Description, Video/Photoset Summarization, Image Filtering, and Image Editing. Recent work has focused on multimodal learning with video-language models, diffusion-based image editing, and event-based vision for low-light conditions. His research has been published in top venues including NeurIPS, ICLR, ICCV, SIGGRAPH, and ACL. He has received significant recognition including The Young Researcher Award from Turkish Academy of Sciences and being named a 2022 Outstanding Associate Editor of IEEE Transactions on Multimedia. He has secured multiple research projects funded by TUBITAK and received gift funds from Adobe Research for text-guided image synthesis work. Current Teaching: BBM202: Algorithms, AIN434/BBM444: Fundamentals of Computational Photography Graduate Supervision: 6 current Ph.D. students, numerous recent graduates including Burak Ercan (2024) and Aysun Kocak (2023) Professional Affiliations: Co-affiliated with Koç University and İş Bank AI Center (KUIS AI)
Professor David Clifton is the Royal Academy of Engineering Chair of Clinical Machine Learning at the University of Oxford’s Institute of Biomedical Engineering. He leads the Computational Health Informatics (CHI) Lab, focusing on AI-driven healthcare solutions with a strong emphasis on translational research in low- and middle-income countries (LMICs). His work spans digital health technologies, medical imaging analysis, and AI ethics. Clifton holds multiple fellowships, including from the Alan Turing Institute and Fudan University. Key affiliations include co-directorship of the Oxford-CityU Centre for Cardiovascular Engineering and involvement in the Wellcome Trust’s Flagship Centre in Vietnam. His research has been commercialized through spinouts like OBS Medical and Oxehealth. Notable projects include AI tools for non-invasive vital sign monitoring and pandemic response strategies using audio-based health data. Clifton’s awards include the IEEE Early Career Award (2022) and the Vice-Chancellor’s Innovation Prize. His lab’s Suzhou branch focuses on open-source digital health research using public datasets. Current research themes include multimodal data integration, generative AI in healthcare, and equitable AI deployment across global health systems.
Shuran Song is an Assistant Professor of Electrical Engineering at Stanford University, with a courtesy appointment in Computer Science. Previously, she was faculty at Columbia University. She holds a Ph.D. in Computer Science from Princeton University and a BEng from HKUST. Her research focuses on the intersection of computer vision and robotics, particularly in embodied AI, robot manipulation, and sensorimotor learning. Song's work emphasizes learning from physical interactions to enable robots to perform complex tasks autonomously. She leads the Robotics and Embodied AI Lab (REAL@Stanford) and has received prestigious awards, including the NSF Career Award, Sloan Fellowship, and Microsoft Faculty Fellowship. Education: Ph.D., Computer Science, Princeton University; BEng, HKUST Affiliations: Stanford School of Engineering, Department of Electrical Engineering Research interests include deformable object manipulation, visuomotor policy learning, and generalizable robot skills. Her lab develops algorithms for robots to learn through interaction, with applications in household assistance (e.g., TidyBot) and industrial automation. Notable contributions include the TossingBot and Diffusion Policy frameworks. Publications span robotics, computer vision, and AI conferences (RSS, ICRA, CVPR), focusing on policy learning, deformable object handling, and embodied intelligence. Awards highlight her impact in advancing robot learning and perception. Advises doctoral and master's students in robotics and AI, and collaborates on grants from NSF, DoD, and industry partners. Teaches courses on robot perception and embodied AI at Stanford.
Stephen Bach is an Assistant Professor in the Computer Science Department at Brown University, where he leads the BATS (Bach's Awesome Team of Students) research group. His research focuses on improving how humans teach computers through programmatic weak supervision and methods for learning from fewer examples like zero-shot and few-shot learning. His primary research interests include weak supervision, data programming, probabilistic soft logic (PSL), statistical relational learning (SRL), information extraction, zero-shot learning, and few-shot learning. Bach's work often focuses on exploiting high-level, symbolic or semantically meaningful domain knowledge, with applications in information extraction, image understanding, scientific discovery, and data science. Bach's recent publications show a strong focus on language models, weak supervision techniques, and multimodal learning, particularly examining the capabilities and limitations of models like CLIP. His research has increasingly emphasized practical applications in low-resource settings and cross-lingual scenarios. Best Paper Award at NeurIPS Workshop on Socially Responsible Language Modelling Research (SoLaR) 2023 Larry S. Davis Doctoral Dissertation Award Selected for oral presentation at ICLR 2024 Best of VLDB 2018 paper selection Bach advises numerous Ph.D., Master's, and undergraduate students, many of whom have gone on to positions at leading tech companies, research institutions, and graduate programs. His research group has developed several influential frameworks including Snorkel (for weak supervision), PSL (Probabilistic Soft Logic), T0 (for zero-shot task generalization), ZSL-KG (for zero-shot learning with knowledge graphs), TAGLETS (for semi-supervised learning with auxiliary data), and WISER (for programmatic weak supervision in sequence tagging).
Dr. Benjamin Busam is a Senior Research Scientist at the Technical University of Munich , affiliated with the Chair for Computer Science Applications in Medicine under Prof. Nassir Navab. Starting September 2025, he will hold the Professorship for Photogrammetry and Remote Sensing at TUM. His career includes leadership roles at FRAMOS Imaging Systems and Huawei Research in London. Education: Mathematics (TUM), Mathematics and Physics (ParisTech, University of Melbourne), PhD in Mathematics (TUM, 2014) His research focuses on 3D computer vision , multi-modal sensor fusion , and their applications in collaborative robotics and augmented reality . He specializes in projective geometry , 6D pose estimation , and neural radiance fields , with a particular emphasis on photometrically challenging environments. Recent publications highlight advancements in 3D scene understanding , neural rendering , and medical imaging , often leveraging machine learning and vision-language models . His work has been recognized through awards like the EMVA Young Professional Award (2015) and Innovation Pioneer of the Year (2019) , along with multiple Outstanding Reviewer distinctions at leading conferences. Dr. Busam has supervised numerous PhD and MSc students on topics including 6D pose estimation , medical augmented reality , and robotic ultrasound , collaborating with institutions like MIT , École Polytechnique , and University of Padova .
Dr. Merry Mani is an Associate Professor in Radiology and Imaging Sciences and Biomedical Engineering, specializing in biomedical imaging and signal processing. Her work focuses on advancing MRI-based imaging technologies to study neurological disorders such as Alzheimer's, Autism, and Epilepsy. She holds a Ph.D. in Electrical and Computer Engineering from the University of Rochester (2014) and completed a postdoctoral fellowship at the University of Iowa School of Medicine (2018). Her research combines biophysical modeling with machine learning to explore brain microstructures. Key achievements include the NNARSAD Young Investigator Grant and NIH-funded projects like 'Fast Multi-dimensional Diffusion MRI with Sparse Sampling'. Her lab develops cutting-edge reconstruction methods like qModeL and MUSSELS, prioritizing high spatio-temporal resolution imaging. Major contributions span diffusion MRI acquisition, model-based deep learning, and clinical applications in neurodegenerative diseases. Notable grants include NIH R01EB031169 for Alzheimer’s neurodegeneration studies and projects on rTMS for depression. Her work bridges imaging innovation with clinical impact, aiming to improve diagnosis and treatment through advanced imaging biomarkers.
Joakim Nivre is a Professor at Uppsala University's Department of Linguistics and Philology. He is a leading researcher in computational linguistics, with a focus on dependency parsing, Universal Dependencies (UD) framework development, and multilingual NLP applications. His recent work explores LLMs in climate change discourse analysis, pharmacovigilance explainability, and historical text processing. Key research areas: Dependency parsing theory, Universal Dependencies standardization, LLM evaluation Collaborations: SweSAT-1.0 benchmark development, ClimateEval project, PARSEME integration His 2025-2023 publications demonstrate expertise in explainable AI for healthcare, synthetic data generation for idioms, and multilingual benchmark design. Notably, he co-developed SweSAT-1.0 to evaluate Swedish LLMs and contributed to typology-informed UD revisions. Despite extensive work in NLP, no scientific awards are mentioned in available texts.
Nicholas Antipa is an Assistant Professor at the University of California San Diego's Jacobs School of Engineering, in the Electrical and Computer Engineering department. His research focuses on the co-design of optical systems and algorithms to develop advanced computational imaging systems, leveraging innovations in 3D printing, sensors, machine learning, and AI. He holds a PhD in Computational Imaging from UC Berkeley and previously worked at the Lawrence Livermore National Lab on optical metrology for the National Ignition Facility. His work includes pioneering projects like the DiffuserCam and Miniscope3D, which enable high-dimensional optical signal capture and 3D microscopy. Education: PhD in Computational Imaging, UC Berkeley (2020) MS in Optics, University of Rochester Institute of Optics BS in Optical Science and Engineering, UC Davis Research Interests: Computational imaging systems, single-shot high-dimensional optical capture, lensless imaging, and applications in neuroscience and marine science. His lab explores novel optical designs, compressed sensing, and AI-driven imaging techniques to push the boundaries of conventional systems. Scientific Awards: Best Paper at ICCP 2019, 2016 Best Demo at ICCP 2017 No. 2 in Optica 15 Top-Cited Articles (2020) Affiliations: Director of the Computational Imaging Systems Lab at UCSD. Collaborates with institutions like Lawrence Livermore National Lab and the Scripps Institution of Oceanography for projects in marine sediment mapping and underwater object detection. His lab emphasizes open-source tools, such as the DiffuserCam Raspberry Pi tutorial.
Cecilia Mascolo is a Professor of Mobile Systems at the University of Cambridge , specifically in the Department of Computer Science and Technology . She co-directs the Centre for Mobile, Wearable System and Augmented Intelligence and is a Fellow of Jesus College, Cambridge . Her research focuses on mobile systems , machine learning for mobile health , and earable technology . She has been awarded prestigious grants such as the ERC Advanced Research Grant (2019-2025) and the EPSRC Open Research Fellowship (2025-2030). Currently on sabbatical at Harvard University , her work bridges systems and machine learning for health applications. Education: PhD in Computer Science from the University of Bologna, Italy. Previous Affiliation: Faculty at University College London before 2008. Her research spans mobile and wearable systems for health and behavior monitoring, focusing on on-device machine learning , uncertainty-aware models , and audio-based diagnostics . Key areas include federated learning , edge computing , and respiratory disease progression analysis via wearables. She explores earable technology for physiological monitoring, gait analysis, and even toothbrushing tracking using in-ear sensors. Her recent publications highlight advancements in earable-based health monitoring , including heart rate estimation , respiratory rate detection , and ECG analysis using machine learning. She emphasizes longitudinal health data from consumer devices, advocating for scalable diagnostics beyond traditional clinical standards. Scientific Awards: ERC Advanced Research Grant EPSRC Open Research Fellowship Best Paper Award - IEEE Percom 10-Year Impact Award - ACM Ubicomp Computer Laboratory Ring Hall of Fame Best Paper Award Student: Andrea Ferlini - ACM SIGMOBILE Doctoral Dissertation Runner-up She leads the Mobile Systems Research Laboratory , mentoring a team of 15 researchers (postdocs and PhD students), and has graduated over 25 PhD students. Her teaching includes Mobile Health courses at the University of Cambridge, and she serves as Director of Studies for Computer Science at Jesus College.
Aswin Sankaranarayanan is a Professor in the Department of Electrical and Computer Engineering at Carnegie Mellon University (CMU) , where he leads the Image Science Lab . His research focuses on computational photography , 3D shape estimation , and novel imaging system design . He earned his Ph.D. in Electrical and Computer Engineering (2009) from the University of Maryland and completed a postdoctoral fellowship at Rice University (2012) . Research Themes: Developing imaging systems that exploit low-dimensional signal models to overcome traditional sensing limitations Co-design of optics and processing algorithms for efficient sensing Application of non-linear signal models to high-dimensional data Advancing compressed sensing and big data processing techniques Scientific Recognition: SIGGRAPH 2023 Best Paper Award (Split-Lohmann Multifocal Displays) CVPR 2019 Best Paper Award (Fermat Paths for NLOS Reconstruction) NSF CAREER Award (2017) Dean’s Early Career Fellowship (2018-2021) Herschel Rich Invention Award (2016) Technical Contributions: His recent publications reveal expertise in non-line-of-sight shape reconstruction , VR/AR display systems , and biomedical imaging . Collaborations span institutions like University College London and University of Toronto.
Desmond Elliott is an Associate Professor in the Natural Language Processing section at the Department of Computer Science, University of Copenhagen (UCPH). His research focuses on multimodal and multilingual models with specific emphasis on vision-language integration and tokenization-free NLP approaches. He teaches Bachelor and Master's level courses including Advanced Topics in Natural Language Processing (since 2019), Grundlæggende Data Science (since 2023), and previously Data Science (2021-2023). His research interests center on building and understanding multimodal and multilingual models , particularly exploring vision and language interactions through billion-parameter systems. Current work investigates cultural representation disparities in vision-language models, parameter-efficient captioning, and multimodal distributional semantics across diverse domains including food culture and medical imaging. His methodology emphasizes real-world applicability in non-English contexts and ethical considerations in multimodal systems. Elliott's recent publications (2025) demonstrate leadership in multimodal NLP, with significant contributions to vision-language pretraining, multilingual evaluation frameworks, and clinical NLP applications. His work spans theoretical advancements in model architectures and practical implementations addressing challenges in low-resource languages and domain adaptation. Best Long Paper Award at EMNLP 2021 Best Poster Award at COLING 2019 As an active educator, Elliott contributes to courses on Fair and Transparent Machine Learning and previously taught Information Retrieval. His research collaborations span international institutions with particular focus on European and non-English language contexts, reflecting UCPH's recognition as Europe's #1 institution for HCI research over the past decade.
Mikko Kurimo is a Full Professor at Aalto University's Department of Information and Communications Engineering, School of Electrical Engineering. He earned his M.Sc., Lic.Tech., and D.Sc.(Tech.) from Helsinki University of Technology (1992, 1994, 1997) and pioneered neural networks for automatic speech recognition (ASR) in his PhD thesis. After research roles at IDIAP (Swiss AI center) and visiting positions at University of Colorado, Edinburgh, SRI, ICSI, and Nitech, he leads Aalto's ASR group since 2000. His work focuses on unsupervised subword modeling for morphologically complex languages (Finnish, Estonian, Turkish, Arabic) and large speech foundation models. PhD in Neural ASR (Helsinki University of Technology, 1997) Research Scientist at IDIAP (Switzerland) Visiting Fellow at University of Colorado, Edinburgh, SRI, ICSI, Nitech Head of Aalto ASR Group (2000-present) His research spans deep learning for ASR, spoken language modeling , and low-resource language solutions . Recent work explores continued pre-training of self-supervised models, multimodal emotion recognition, and pronunciation assessment using LLMs. He led the winning team in the 2017 Multi-Genre Broadcast challenge and secured competitive funding in Tekes Challenge Finland and EC's H2020-ICT-2017. Key article trends include: Advancements in children's speech recognition and dysarthric speech processing Integration of generative AI for language learning feedback Specialization in low-resource Uralic languages (Finnish, Northern Sámi) Development of robust ASR systems for complex phonetic environments Scientific Awards ACM Multimedia 2023 Computational Paralinguistics Challenge Prize First place in MGB3 2017 Arabic ASR Challenge ISCA Best Student Paper Award (2011) Professeur Invité at Université de Saint-Etienne (2005-2006) Royal Society International Short Visit Fellowship (2004) Professor Kurimo leads the Speech Recognition Group at Aalto, collaborating with COIN (Centre of Excellence in Computational Inference) and AIRC (Adaptive Informatics Research Centre). His projects like CaptainA mobile app demonstrate practical applications of ASR in language education. He has supervised numerous publications with co-authors in domains spanning bandwidth extension, stuttering detection, and speech sound disorder assessment.
Jesper Rindom Jensen is an Associate Professor in the Department of Electronic Systems at Aalborg University, Denmark, under the Technical Faculty of IT and Design. He is the Head of the Audio Analysis Lab, a leading research group in audio signal processing, since 2023. His work bridges theoretical signal processing and practical applications in artificial intelligence and audio systems. Full Name: Jesper Rindom Jensen Institution: Aalborg University School: The Technical Faculty of IT and Design Department: Department of Electronic Systems Research Lab: Audio Analysis Lab Email: jrj@es.aau.dk Office: Fredrik Bajers Vej 7B, B5-206, 9220 Aalborg Øst, Denmark Education: M.Sc. in Electronic Systems, Aalborg University (cum laude, 2009) Ph.D. in Signal Processing, Aalborg University (2012) Research Interests: Jesper Rindom Jensen's research centers on audio signal processing, with a strong emphasis on artificial intelligence, speech enhancement, noise reduction, beamforming, and multichannel systems. His work applies to diverse domains including robot and drone audition, spatial audio, and active noise control. He develops novel filtering techniques, including variable span linear filters and harmonic beamformers, to improve speech quality and intelligibility in noisy and reverberant environments. Publication Trends: His recent publications (2023–2025) show a strong trend toward integrating deep learning with classical signal processing, particularly in direction-of-arrival estimation, underwater acoustics, and robust multichannel systems. There is a clear focus on real-world applications, including sound zone control, active noise control, and limited-data scenarios using knowledge distillation. His work consistently emphasizes robustness, efficiency, and practical deployment. Scientific Awards and Recognition: AAU Talent for emerging research leaders Recipient of a competitive postdoc grant from the Danish Independent Research Council Advising and Grants: Jesper has supervised multiple PhD and master’s students, including Nørholm, Karimian-Azari, Zhang, and Wang. He has led significant research projects such as 'Sound Processing for Robots and Drones' (2018–2020) and participated in others related to joint audio-visual tracking and speech enhancement. His research has been supported by national funding bodies, reflecting its innovation and impact. Labs and Teams: He is a founding and core member of the Audio Analysis Lab at Aalborg University, which focuses on cutting-edge audio signal processing and AI-driven solutions. The lab fosters interdisciplinary collaboration and has produced numerous publications, datasets, and real-world applications. Jensen’s leadership since 2023 underscores his pivotal role in shaping the lab’s research direction.