Haibin Ling is the SUNY Empire Innovation Professor in the Department of Computer Science at Stony Brook University, part of the College of Engineering and Applied Sciences. His research focuses on computer vision, medical image analysis, augmented reality, and AI applications in science. He holds a Ph.D. from the University of Maryland (2006) and prior degrees from Peking University. Previously, he worked at Temple University (2008–2019) and held roles at Siemens Corporate Research, UCLA, and Microsoft Research Asia. Professor Ling's work spans biomedical imaging, AI for science, and human-computer interaction. He leads the CV Lab and collaborates with the AI Institute at Stony Brook. Awards include the NSF CAREER Award (2014), Best Student Paper (ACM UIST 2003), and IEEE Fellow (2020). He serves on editorial boards for IEEE Trans. PAMI, Pattern Recognition, and CVIU, and chairs major conferences like CVPR. His research group includes over 50 students and alumni, with active projects in tracking benchmarks (LaSOT), Leafsnap, and medical imaging tools. Notable publications address OCTA flow estimation, backdoor attacks on vision models, and topology-guided medical learning. Collaborations involve institutions like Temple University and Stony Brook's Department of Applied Mathematics and Statistics.
Angel Xuan Chang is an Associate Professor at Simon Fraser University's School of Computing Science, affiliated with labs including 3DLG, GrUVi, SFU NatLang, SFU AI/ML, and VINCI. He holds a Canada CIFAR AI Chair and was a TUM-IAS Hans Fischer Fellow (2018-2022). His research bridges natural language processing (NLP), 3D scene understanding, and embodied AI, focusing on language-grounded 3D generation and biodiversity monitoring via DNA barcodes. Recent work includes NuiScene (unbounded outdoor scene generation), ViGiL3D (3D visual grounding dataset), and CLIBD (vision-genomics biodiversity analysis). He advises students in projects like BIOSCAN-5M insect dataset and embodied AI navigation. His 2025 highlights include multiple ICCV and ICLR papers, workshops at ICML and CVPR, and a CRV invited talk. Education: Ph.D. in Computer Science from Stanford University (2014), advised by Chris Manning. Previous roles include visiting research scientist at Facebook AI Research and researcher at Eloquent Labs.
Prof. Anya Belz is Full Professor of Computer Science at Dublin City University's School of Computing and Science Lead at ADAPT Research Centre. A leading NLP researcher with PhD-level expertise, she specializes in natural language generation, evaluation methodologies, and multimodal systems. Recipient of multiple best paper awards and NAACL Test of Time Award nomination. Research innovations include foundational work on statistical language generation (deployed in weather forecasting systems), comparative evaluation frameworks, vision-language integration, and reproducibility quantification. Current EPSRC-funded ReproHum project coordinates 20 global labs studying evaluation consistency. Achievements : Developed industry-deployed generation systems for accessibility applications Pioneered cross-modal alignment techniques for image description Authored 100+ publications spanning generation, evaluation, and reproducibility
Tiancheng Zhao is a principal researcher at the Binjiang Institute of Zhejiang University and founder of the Om Artificial Intelligence Laboratory (Om AI Lab), dedicated to frontier open multimodal AGI research for building next-generation agents that transform work and life through advanced human-machine interaction. His academic credentials include: Ph.D. in Computer Science from Carnegie Mellon University (2016-2019) under Prof. Maxine Eskenazi, Prof. Louis-Philippe Morency, Prof. William W. Cohen, and Dr. Dilek Hakkani-Tur, with pioneering dissertation “Learning to Converse With Latent Actions” in end-to-end generative conversational models M.S. in Computer Science from Carnegie Mellon University (2014-2016) B.S. in Electrical Engineering from UCLA (2010-2014) with Summa Cum Laude, focusing on speech signal processing under Prof. Abeer Alwan Dr. Zhao’s research centers on multimodal foundation models and agents, tackling three core challenges: Multimodal Models for cross-modal representation learning in high-dimensional data, Learning to Learn for effective skill acquisition from diverse signals (supervised labels, rewards, meta-learning), and AI Agents for open-world understanding and complex decision-making. His work bridges computer vision, natural language processing, and real-world applications including healthcare analytics and remote sensing. Analysis of his 50+ publications reveals accelerating innovation in multimodal large language models (2024-2025), with emphasis on stable vision-language architectures (VLM-R1), agent orchestration frameworks, and domain-specific applications in geospatial analysis and healthcare. Key trends include solving long-tail distribution challenges in satellite imagery, developing human-like zooming capabilities for multimodal LLMs, and creating unified benchmarks for autonomous GUI testing. His scientific recognition includes: National Breakthrough Technology Award by Ministry of Science and Technology (2021) Microsoft Research Best & Brightest PhD (2018) BEST PAPER AWARD at SIGDIAL 2018 Best Paper Nomination at SIGDIAL 2016 Top 1 Outstanding Bachelor of Science Award at UCLA (2014) As Om AI Lab founder, Dr. Zhao leads research teams developing computational building blocks for human-AI collaboration. While specific student mentorship details aren’t public, his extensive publication record with junior co-authors indicates active research supervision. Current projects focus on practical system implementations for real-world multimodal agent deployment across diverse domains.
Eugenia Rho is an Assistant Professor in the Department of Computer Science at Virginia Polytechnic Institute and State University (Virginia Tech), part of the College of Engineering. Her research focuses on data analytics, machine learning, natural language processing, and human-computer interaction, with particular emphasis on social media discourse, online identity dynamics, and ethical AI applications. Education includes a Ph.D. in Information and Computer Sciences from the University of California, Irvine (2020), and a B.A. in Political Science from Columbia University (2011). Her interdisciplinary background bridges computer science and social sciences. Research interests span AI-assisted communication tools, counterspeech strategies for online hate mitigation, and neurodivergent perspectives in technology design. Her work often integrates computational methods with social science theories to address real-world challenges such as bias detection, mental health support, and ethical AI deployment. Recent publications highlight themes like AI collaboration in writing, identity-driven online interactions, and the efficacy of counterspeech. Her projects frequently involve designing human-centered technologies that prioritize accessibility and ethical considerations. No scientific awards are explicitly mentioned in the provided text. She maintains an active Google Scholar profile and a personal homepage (URLs not provided in the text).
Shinji Watanabe is an Associate Professor at Carnegie Mellon University's Language Technologies Institute and a Courtesy Professor in the Electrical and Computer Engineering department. He holds a Ph.D. (Dr. Eng.) from Waseda University, Japan, and has held research roles at NTT Communication Science Laboratories, Mitsubishi Electric Research Laboratories (MERL), and Johns Hopkins University. His research focuses on automatic speech recognition, speech enhancement, and machine learning for speech processing. Watanabe has published over 300 peer-reviewed papers and received the Best Paper Award at IEEE ASRU 2019. His work emphasizes robust speech processing in challenging environments, multilingual models, and neural audio codecs. He leads the ESPnet toolkit development for end-to-end speech processing systems and contributes to technical committees like IEEE SLTC and APSIPA SLA. Recent research trends include streaming speech systems, universal speech enhancement (URGENT challenges), and fusion of discrete speech units with self-supervised representations. He explores scalable speech foundation models through benchmarks like ML-SUPERB 2.0 and investigates cross-modal audio-visual processing in challenges like MISP 2025. Education : B.S., M.S., Ph.D. (Waseda University) Affiliations : CMU Language Technologies Institute, CMU ECE, Former roles at MERL and Johns Hopkins Key Projects : ESPnet, OpenWhisper-Style Models, URGENT Challenge Frameworks
Berrak Sisman is an Assistant Professor in the Department of Electrical and Computer Engineering at Johns Hopkins University, affiliated with the Data Science and AI Institute and the Center for Language and Speech Processing (CLSP). She leads the Speech & Machine Learning Lab (SmILe Lab), focusing on AI-driven speech technologies. She received her PhD from the National University of Singapore in 2020 and was previously a tenure-track faculty member at the University of Texas at Dallas (2022–2024). Research Interests: Her work spans artificial intelligence, speech synthesis, voice conversion, emotion analysis in speech, medical speech applications, and secure speech technology. She develops neural models for expressive and adaptive speech processing. Publications: Her recent articles (2024–2025) emphasize speech emotion recognition, zero-shot prosody control, accent conversion, and disentangled representations in TTS, reflecting a focus on cross-modal learning, robustness, and real-world applications. Awards & Grants: NSF CAREER Award (2024) Amazon Faculty Research Award (2022) Singapore Ministry of Education Award (2021) A*STAR Singapore International Graduate Award (2016–2020) Leadership: She directs the SmILe Lab, recruiting PhD/Master’s students for projects in neural speech modeling. Her grants include NSF and Amazon funding for voice conversion and emotion synthesis research.
Erkut Erdem is a Professor in the Department of Computer Engineering at Hacettepe University, where he leads the Computer Vision Laboratory (HUCVL). His research focuses on computer vision and machine learning, particularly on incorporating different kinds of context (spatial, temporal and cross-modal) into visual processing across all levels from low to high-level vision. He received his Ph.D. (2008), M.Sc. (2003), and B.Sc. (2001) from Middle East Technical University. Prior to joining Hacettepe University in 2010, he completed a post-doctoral fellowship at Ecole Nationale Supérieure des Télécommunications (2009-2010) and held visiting researcher positions at UCLA (2007) and Virginia Tech (2004). His current research interests include Visual Saliency Prediction, Automatic Image Description, Video/Photoset Summarization, Image Filtering, and Image Editing. Recent work has focused on multimodal learning with video-language models, diffusion-based image editing, and event-based vision for low-light conditions. His research has been published in top venues including NeurIPS, ICLR, ICCV, SIGGRAPH, and ACL. He has received significant recognition including The Young Researcher Award from Turkish Academy of Sciences and being named a 2022 Outstanding Associate Editor of IEEE Transactions on Multimedia. He has secured multiple research projects funded by TUBITAK and received gift funds from Adobe Research for text-guided image synthesis work. Current Teaching: BBM202: Algorithms, AIN434/BBM444: Fundamentals of Computational Photography Graduate Supervision: 6 current Ph.D. students, numerous recent graduates including Burak Ercan (2024) and Aysun Kocak (2023) Professional Affiliations: Co-affiliated with Koç University and İş Bank AI Center (KUIS AI)
Alexei A. Efros is the Howard Friesen Professor in the EECS Department at UC Berkeley, affiliated with the Berkeley Artificial Intelligence Research (BAIR) Lab. Previously, he spent a decade at CMU's Robotics Institute and held a postdoc at the University of Oxford under Andrew Zisserman. He collaborates with INRIA/École Normale Supérieure in Paris. His research focuses on self-supervised learning, generative models, and visual data mining, with applications to robotics, computational photography, and art. Education & Academic Roles: Postdoc at Oxford (with Andrew Zisserman), faculty at CMU (2005–2015), currently at UC Berkeley. Teaches courses like CS 180/280A (Computer Vision) and CS 280 (Graduate Computer Vision). Research Interests: Self-supervised learning, generative models (e.g., diffusion models, inpainting), visual commonsense, and cross-modal reasoning. His work bridges computer vision and graphics, emphasizing data-driven approaches. Recent projects include Visual Jenga, Diffusion Models as Data Mining Tools, and Prioritized Generative Replay. Grants & Labs: Leads the Efros Research Group, advised over 40 PhD students (e.g., Jun-Yan Zhu, Tinghui Zhou). Collaborates with institutions like INRIA and NVIDIA. Active in grants related to AI, vision, and robotics. Labs/Teams: BAIR Lab (UC Berkeley), former affiliations with CMU Robotics Institute and Willow Team (INRIA/ENS Paris). Current lab focuses on generative AI, 3D perception, and visual reasoning.
Xuezhe Ma is an Assistant Professor in the Department of Computer Science at the University of Southern California's Viterbi School of Engineering. Previously, he was a Ph.D. student at Carnegie Mellon University's Language Technologies Institute, where he worked under the supervision of Professor Eduard Hovy. His academic journey includes a Master's degree from Shanghai Jiao Tong University's Center for Brain-like Computing and Machine Intelligence and a Bachelor's degree in Computer Science from the same institution. Ph.D. in Computer Science, Carnegie Mellon University (completed ~2020) M.S. in Brain-like Computing, Shanghai Jiao Tong University B.S. in Computer Science, Shanghai Jiao Tong University Dr. Ma's research spans multiple areas at the intersection of Natural Language Processing and Machine Learning, with particular focus on structured prediction, syntactic and semantic parsing, machine translation, language generation, and deep generative models. His recent work has expanded into vision-language models, large language model architectures, and applications across computer vision tasks. His research combines theoretical foundations with practical implementations, as evidenced by his development of tools like NeuroNLP2 and MaxParser. His publication record shows a clear trajectory from foundational NLP work during his PhD (including papers on dependency parsing and sequence labeling) to more recent contributions in generative models and large language systems. The 15 most recent publications reveal a strong focus on addressing fundamental challenges in generative modeling, context handling, and multimodal integration, with applications spanning literary translation, medical imaging, and news diffusion analysis. AI2 Outstanding Intern Award (2018) Dr. Ma has secured research funding supporting his work in generative models and language technologies, with projects focusing on improving the efficiency and capabilities of large language models. His research group at USC is actively working on next-generation language understanding and generation systems, with particular emphasis on context-aware modeling and multimodal integration. He has established collaborations with industry partners including the Allen Institute for AI and has contributed to open-source projects like Texar. At USC, Dr. Ma leads research in the Information Sciences Institute, directing projects on efficient large language model architectures and multimodal reasoning systems. His lab focuses on developing novel approaches to context handling, model efficiency, and multimodal integration, with applications across diverse domains including healthcare, literary analysis, and news media.
Wenping Wang is a Professor in the Department of Computer Science & Engineering at Texas A&M University, part of the College of Engineering. His research focuses on computer graphics, computer vision, geometric modeling, and visualization. He holds Fellowships from ACM and IEEE, and has received notable awards including the 2021 AsiaGraphics Outstanding Technical Contributions Award and the 2017 John Gregory Memorial Award. Wang's educational background includes a Ph.D. from the University of Alberta and M.Eng. and B.Sc. degrees from Shandong University. His work spans advancements in neural implicit surfaces, 3D reconstruction, and medical imaging applications such as orthodontic treatment prediction. He has authored numerous influential papers in top-tier conferences like SIGGRAPH and journals like ACM Transactions on Graphics. His research interests emphasize bridging geometric modeling with machine learning, particularly in neural rendering, surface parameterization, and medical visualization. Recent projects include developing frameworks for automatic tooth alignment and high-fidelity 3D geometry generation. Wang's contributions have significantly impacted both theoretical foundations and practical applications in computer graphics.
Professor Emma Moore is a leading sociolinguist at the School of English, University of Sheffield . Specializing in the social dimensions of language, her work integrates methodologies from anthropology and sociology to investigate how linguistic variation constructs identity and social affiliations. Research focus: Sociolinguistic style, dialect contact, youth language, identity negotiation, and community-based fieldwork Key projects: Adolescence and grammar (1999-), Isles of Scilly dialect (2008-), language and inequality (2014-), language perception software (2016-) Educational background : PhD in Sociolinguistics, University of Manchester Stanford University (USA) during PhD research Research trends across her 2021-2025 publications reveal interdisciplinary exploration of: Socio-syntactic variation Clinical linguistics (2025 MDS-UPDRS studies) Digital discourse post-pandemic Indexicality and cognitive representations Dialect contact in insular communities Educational policy implications Scientific recognition : British Academy Mid-Career Fellow (2019-2020) Elected Fellow, Royal Anthropological Institute (2014) AHRC/BA funded research projects (1999-2015) Research group leadership : Mentored 9 PhD students (2004-2024) and supervised 3 undergraduate SURE projects, including analyses of: Gender in rap performance Language style and workplace inequality Oral history digitization Professional contributions : Editorial Board: Language in Society (2015-), Gender and Language (2011-2017) External Examiner: Lancaster University MA in English Language (2013-2017) Collaborative software development for dialect perception testing
Bryan Pardo is a Professor of Computer Science at Northwestern University and head of the Interactive Audio Lab. He co-directs the Northwestern Center for Human Computer Interaction + Design and chairs the Computer Science Diversity Committee. He teaches courses in Deep Learning, Machine Learning, Generative Modeling, and Digital Music Instrument Design. PhD in Computer Science and Engineering, University of Michigan MMus in Jazz and Improvisation, University of Michigan MS in Computer Science, Ohio State University BMus in Jazz Composition, Ohio State University His research focuses on machine understanding and manipulation of sound, particularly in music and speech domains. Key areas include Machine Learning (e.g., automated gradient clipping), Signal Processing (e.g., Multi-scale Common-fate Transform), and Human Computer Interaction. Applications involve inclusive audio interfaces, audio search engines, source separation, natural language-controlled audio effects, privacy-preserving adversarial attacks on voice recognition, and music co-creation tools. Recent publications highlight advancements in neural watermarking (MaskMark), masked acoustic modeling (VampNet), and real-time adversarial privacy systems for speech. His lab's work has been applied in Adobe's AI-powered audio editor and Lexie B2 hearing aids. Scientific Awards: $1.8 million NSF Future of Work award $440K NSF grant for accessible music programming $200K Toyota grant $100K Sony grant TorchCrepe pitch tracker: 20 million+ downloads Bryan Pardo advises PhD student Max Morrison and collaborates with researchers like Patrick O'Reilly, Zeyu Jin, and Prem Seetharaman. His lab develops technologies for blind and visually impaired audio creators, including HaptEQ and Eyes-free tools.
Karen Panetta is a Professor at Tufts University School of Engineering with appointments in Electrical and Computer Engineering, Computer Science, Mechanical Engineering, and Academic Services. She currently serves as Dean of Graduate Education for the School of Engineering and holds the title of Distinguished Professor. Ph.D. in Electrical Engineering, Northeastern University M.S. in Electrical Engineering, Northeastern University B.S. in Computer Engineering, Boston University Dr. Panetta's research focuses on developing efficient algorithms for simulation, modeling, and signal and image processing for security and biomedical applications. Her work brings together artificial intelligence, machine learning, and visual sensing systems to create solutions for robot vision and biomedical imaging. She develops algorithms inspired by the human visual system to enable machines to 'see' like humans, with applications in homeland security, biomedicine, facial recognition, and search and rescue operations. Her research has significant humanitarian applications, addressing global challenges facing women and children. Dr. Panetta has received numerous prestigious awards including induction into the National Academy of Engineering (2023), the Presidential Award for Science and Engineering Education and Mentoring (2011), and the IEEE Award for Distinguished Ethical Practices (2013). She is a fellow of multiple prestigious academies including the National Academy of Inventors, European Academy of Sciences and the Arts, and IEEE. Member, National Academy of Engineering (2023) Presidential Award for Science and Engineering Education and Mentoring (2011) IEEE Award for Distinguished Ethical Practices (2013) Fellow, National Academy of Inventors Fellow, European Academy of Sciences and the Arts Fellow, Asia-Pacific Artificial Intelligence Association As an educator and mentor, Dr. Panetta founded the nationally acclaimed Nerd Girls program to promote engineering to young students, particularly women. She previously served as worldwide director for IEEE Women in Engineering and editor-in-chief of the IEEE Women in Engineering magazine. Her approach to graduate education emphasizes the importance of building strong collaborative relationships between faculty and students, with a focus on proactive communication and documentation of research progress. Dr. Panetta's humanitarian research applies engineering solutions to global challenges, including developing technology to help doctors find cancerous tumors, security screeners find concealed weapons, and law enforcement agencies find criminals and missing children. Her work demonstrates a commitment to 'Doing The Right Thing' by addressing issues affecting populations with limited resources or 'voice' in society.
Prof. Margret Keuper is a Professor of Machine Learning at the University of Mannheim's School of Business Informatics and Mathematics, leading the Data and Web Science Group. She is also affiliated with the Max-Planck-Institute for Informatics and ELLIS (fellow since 2024). Her research focuses on robust deep learning, neural architecture search, and computer vision tasks like motion segmentation and adversarial defense. She holds a PhD from the University of Freiburg and previously held positions at the University of Siegen and the University of Mannheim. Her work spans projects funded by DFG and BMBF, including Climate Visions for social media analysis and TrackOpt for motion tracking. She teaches courses on computer vision, generative models, and reinforcement learning. She actively serves on program committees for top conferences like CVPR, ECCV, and NeurIPS, and is an associate editor for IEEE TPAMI and JAIR. Education: PhD in Computer Science from University of Freiburg (advisor: Thomas Brox) Research Projects: Learning to Sense (DFG), Climate Visions (BMBF), TrackOpt (BMBF) Key Roles: Head of Mannheim Master in Data Science Examination Board, Member of MSc Business Informatics Board Her research emphasizes robustness in AI systems, with contributions to adversarial attacks, domain generalization, and efficient solvers for large-scale problems. She advises over 15 PhD students across academic and industry partnerships.