Slim Essid is a Full Professor at Télécom Paris, leading the Audio Data Analysis and Signal Processing (ADASP) group. He holds a Doctorat (Ph.D.) and Habilitation from Université Pierre et Marie Curie (UPMC). With 15+ years of research experience, he has advised 15 PhD graduates and currently co-advises 10 others. His work focuses on machine learning, signal processing, and multimodal systems, publishing over 150 peer-reviewed papers. He serves as a reviewer for top journals/conferences (e.g., IEEE Transactions) and research funding agencies. Education: State Engineering Degree, École Nationale d’Ingénieurs de Tunis (2001) M.Sc. (D.E.A.) in Digital Communication Systems, École Nationale Supérieure des Télécommunications, Paris (2002) Ph.D., Université Pierre et Marie Curie (2005) Habilitation (HDR), UPMC (2015) Research Interests: Multimodal learning, self-supervised representations, audio-visual segmentation, music structure analysis, domain generalization, and speech enhancement. Recent publications highlight innovations like TACO (training-free sound-prompted segmentation) and CLOUDS (domain-generalized semantic segmentation framework using foundation models). His work bridges audio processing with vision and language models, emphasizing unsupervised/zero-shot approaches. Key achievements include state-of-the-art methods in sound event detection, speaker diarization, and music segmentation. He collaborates with 14 post-docs and leads projects funded by French/EU agencies.









