Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Prof. Konrad Schindler holds the position of Full Professor at the Department of Civil, Environmental and Geomatic Engineering at ETH Zürich. He is also the Head of the Institute of Geodesy and Photogrammetry (IGP), leading research and educational activities in geomatics and computer vision. His career spans roles as a Photogrammetric Engineer, scientific assistant, postdoc researcher, and academic faculty across institutions including Graz University of Technology, Monash University, and TU Darmstadt before joining ETH Zürich in 2010. Education: Undergraduate studies in Geodesy (1992–1995), Graz University of Technology, Austria MEng in Photogrammetry and Geoinformation (1995–1999), Vienna University of Technology, Austria PhD in Computer Science (2001–2003), Graz University of Technology, Austria Research focuses on Photogrammetry , Remote Sensing , Computer Vision , and Image Understanding with interdisciplinary applications in environmental monitoring, geospatial analysis, and disaster response. He develops computational methods for 3D reconstruction, fusion of multi-modal data, and AI-driven solutions for satellite imagery interpretation. His work bridges geomatic engineering and machine learning to address challenges in urban mapping, climate modeling, and biological systems analysis. Publications reflect expertise in geospatial AI, diffusion models, and benchmarking datasets for disaster resilience. Notable works include Marigold (image analysis adaptation) and BRIGHT (building damage assessment). His research emphasizes practicality and scalability, such as affordable depth estimation and global biomass datasets. He has received the 2013 Marr Prize Honourable Mention (IEEE) and the 2012 U.V. Helava Award (ISPRS), alongside several Best Presentation Awards. His contributions span technical leadership, editorial roles (ISPRS Journal), and service to Swiss remote sensing commissions. Advising and grants: While no specific advisee names or grant details are listed, his career trajectory includes mentoring postdocs and junior faculty. He teaches advanced courses in Photogrammetry , Image Interpretation , and Machine Vision , integrating cutting-edge AI techniques into curricula. His research group collaborates on global-scale projects like canopy height mapping and satellite-based climate variable assessments. Labs/Teams: As Institute Head, he oversees the IGP lab at ETH Zürich, with prior affiliations including the Digital Perception Lab (Monash University) and the Computer Vision Lab (ETH Zurich). His work often involves multi-institutional collaborations focused on geospatial AI and environmental science.
Yuki M. Asano is a full Professor at the University of Technology Nuremberg , leading the Fundamental AI (FunAI) Lab . Previously, he led the QUVA Lab at the University of Amsterdam and earned his PhD at the Visual Geometry Group (VGG) of the University of Oxford under Andrea Vedaldi and Christian Rupprecht. University of Technology Nuremberg (2024–present) University of Amsterdam (prior to 2024) University of Oxford (PhD, 2020) His research spans Artificial Intelligence , Machine Learning , and Computer Vision , with a focus on Causal Representation Learning , Self-Supervised Learning , and Efficient Model Adaptation . He pioneered techniques like BISCUIT (causal variable identification) and VeRA (parameter-efficient fine-tuning). His work extends to Medical Imaging and Environmental Monitoring through applications in fetal ultrasound analysis and marine debris detection. Recent publications (2023–2025) highlight advancements in Self-Supervised Learning , Vision-Language Models , and 3D Understanding . Notable papers include TWIST & SCOUT (multimodal LLM grounding), SIGMA (masked video modeling), and GeneralAD (anomaly detection). His ICCV 2023 work on Self-Ordering Point Clouds and MoSiC (optimal-transport motion trajectories) underscores his interdisciplinary approach. He received the JUPITER compute grant (2025) and an Outstanding Paper Award at ICLR 2024 . His collaborations span institutions like MIT-IBM Watson AI Lab, Qualcomm AI Research, and University of Amsterdam.
Kenji Kawaguchi is the Presidential Young Professor in the Department of Computer Science at the National University of Singapore (NUS), where he leads the Deep Learning Lab and is a faculty affiliate at the NUS Institute of Data Science. His research bridges theoretical and applied machine learning, focusing on deep learning, large language models, and physics-informed neural networks. His educational background includes a Ph.D. and S.M. in Computer Science and Electrical Engineering from the Massachusetts Institute of Technology (MIT), advised by Leslie Pack Kaelbling, and a postdoctoral fellowship at Harvard University’s Center of Mathematical Sciences and Applications. Dr. Kawaguchi’s research interests center on the theoretical foundations of deep learning, optimization, generalization, and applications in areas such as molecular modeling, AI safety, and efficient training of large models. He has made significant contributions to understanding in-context learning, diffusion models, and neural operators for partial differential equations. His recent publications (2023–2025) reflect a strong trend toward improving the efficiency, robustness, and interpretability of large-scale models, particularly in language and scientific domains. Key themes include LLM alignment and safety, diffusion model optimization, and physics-informed learning for high-dimensional problems. Presidential Young Professor He has served as Area Chair and PC Member for top-tier conferences including NeurIPS, ICML, ICLR, AAAI, and UAI, and as reviewer for journals such as JMLR and Annals of Statistics. He has delivered invited talks at Harvard, MIT, Stanford, CMU, Brown, and Google Research, reflecting his international recognition. He actively mentors students and welcomes PhD candidates and postdocs to join his research group.
Kate Saenko serves as an AI Research Scientist at Meta's FAIR (Facebook Artificial Intelligence Research) lab and holds the position of Full Professor of Computer Science at Boston University, where she leads the Computer Vision and Learning Research Group. Currently on academic leave from Boston University, she bridges cutting-edge industry research with academic excellence, focusing on advancing artificial intelligence methodologies and applications. Her educational background includes a PhD in Electrical Engineering and Computer Science (EECS) from the Massachusetts Institute of Technology (MIT), followed by postdoctoral training at the University of California, Berkeley and Harvard University. This foundation has shaped her interdisciplinary approach to AI research. Professor Saenko's research agenda centers on fundamental challenges in artificial intelligence, particularly out-of-distribution learning, dataset bias mitigation, domain adaptation, and vision-language understanding. Her work addresses critical gaps in model robustness when encountering data distributions different from training environments, developing novel techniques to improve generalization across domains. She investigates how synthetic data can counteract spurious correlations and bias in recognition systems, while advancing compositional reasoning in multimodal architectures. Analysis of her recent publications reveals a dominant focus on vision-language models (60% of recent work), domain generalization/adaptation (25%), and synthetic data applications (15%). Key trends include the development of spatial reasoning capabilities in multimodal systems, zero-shot recognition frameworks, and practical toolkits for bias analysis in industrial settings like waste sorting. Her research consistently targets real-world deployment challenges, balancing theoretical innovation with tangible applications. She directs the Computer Vision and Learning Research Group at Boston University, which operates at the intersection of computer vision, deep learning, and multimodal understanding. The group maintains strong industry collaborations through Meta's FAIR and previously engaged with the MIT-IBM Watson AI Lab. Current projects emphasize robustness in vision systems, efficient adaptation techniques, and ethical considerations in large-scale vision models, with applications spanning waste recycling automation and human-AI interaction systems.
Alane Suhr is an Assistant Professor at UC Berkeley's Electrical Engineering and Computer Sciences (EECS) department and a member of the Berkeley Artificial Intelligence Research Lab (BAIR). Her research focuses on natural language processing (NLP), machine learning, and computer vision, emphasizing systems that interact with humans through language. She designs models and datasets for language grounding (e.g., NLVR) and develops algorithms for learning through interaction. Education: PhD in Computer Science, Cornell University (2022), advised by Yoav Artzi Bachelor's in Computer Science and Engineering, Ohio State University (2016), with a Linguistics minor Research Interests: Interactive systems for collaborative language use (e.g., CerealBar) Language grounding in multimodal contexts Embodied agents and reinforcement learning Generalization in NLP and the role of language in learning Key Contributions: Developed the NLVR dataset for visual reasoning with natural language Pioneered work on SWE-Gym for training software engineering agents Contributed to research on language models' sensitivity and commonsense reasoning Awards: ACM Doctoral Dissertation Award (2022) Outstanding Paper Awards at ACL 2023 and EMNLP 2021 Professional Activities: Organized workshops at ICML, NeurIPS, and ACL on topics like agents, generalization, and theory of mind Active in academic outreach and conference speaking (e.g., ICML, NeurIPS, CVPR) Labs/Teams: Berkeley Artificial Intelligence Research (BAIR) Lab SWE-Gym research group
Wei-Lun (Harry) Chao is an Associate Professor in the Department of Computer Science and Engineering at the Ohio State University (OSU), College of Engineering. Promoted to this role in May 2025, he is also an Innovation Scholar and Distinguished Assistant Professor of Engineering Inclusive Excellence. His work spans machine learning, computer vision, and their applications in autonomous driving, healthcare, biology, and natural language processing. Research Focus: Machine learning with imperfect data, interpretable and personalized learning, robust perception for autonomous systems, and visual recognition in real-world scenarios. Awards: 2025 OSU Early Career Distinguished Scholar Award, CVPR Best Student Paper Award (2024), CSE Faculty Teaching Award (2024), Lumley Research Award (2023). Grants: Funded by NSF, NIH, ONR, Cisco, AWS, and Google. Notable Research Trends: The 15 most recent articles highlight his work on vision foundation models, federated learning, diffusion models for biological species generation, interpretable vision transformers, and robust perception systems for autonomous driving. Key subfields include sparse autoencoders, 3D object detection, semi-supervised learning, and anomaly detection in scientific domains. Scientific Awards: 2025 Early Career Distinguished Scholar Award (OSU) CVPR Best Student Paper Award (2024) CSE Faculty Teaching Award (2024) Lumley Research Award (2023) Mentoring & Grants: As an advisor for the OSU Buckeye AutoDrive Team and AI Club, he mentors graduate and undergraduate students. His research is supported by major grants from NSF, NIH, ONR, and industry partners like Cisco and Google.
Olga Russakovsky is an Associate Professor of Computer Science at Princeton University and Associate Director of the Princeton Laboratory for Artificial Intelligence. Her research focuses on computer vision , machine learning , human-computer interaction , and fairness, accountability, and transparency in AI systems. Princeton University faculty member since 2025 Affiliated with Princeton's Center for Statistics and Machine Learning and Center for Information Technology Policy Scientific Recognition: Presidential Early Career Award for Scientists and Engineers (2025) PAMI Young Researcher Award (2022) AnitaB.org Emerging Leader Abie Award (2020) CRA-WP Anita Borg Early Career Award (2020) MIT Technology Review 35-under-35 Innovator (2017) PAMI Everingham Prize (2016) As a co-founder and current Board Chair of AI4ALL , she drives initiatives to expand diversity in AI. Her recent publications demonstrate expertise in vision-language models , deepfake detection , and ethical AI systems .
Aishwarya Agrawal is an Assistant Professor at Université de Montréal in the Department of Computer Science and Operations Research (DIRO), affiliated with Mila – Quebec Institute of Artificial Intelligence and a Canada CIFAR AI Chair. She also serves as a research scientist at Google DeepMind, spending one day weekly there. Education: B.E. in Electrical Engineering (IIT Gandhinagar, 2014), Ph.D. in Computer Science (Georgia Tech, 2019). Her research focuses on multimodal learning , deep learning , natural language processing , and computer vision , particularly in developing AI systems that 'see' and 'communicate' effectively. Grants & Awards: Canada CIFAR AI Chair, 2020 Sigma Xi Best PhD Thesis Award, NVIDIA Fellowship (2018–2019), and multiple fellowships from Google and Facebook. She leads projects like Advancing Multimodal Vision-Language Learning (CRSNG-funded) and StarDoc: Document Structure Extraction (MITACS). Research Contributions: Pioneered benchmarks like CulturalVQA and UI-Vision , and frameworks such as PROGRESS for efficient VLM training. Her work emphasizes cross-modal alignment, robust evaluation, and cultural understanding in AI systems. Labs/Teams: Active in Mila’s core academic group and collaborates with Google DeepMind on multimodal and vision-language research. Supervises a dynamic team of PhD and master’s students in Montreal.
Bo Li is an Associate Professor at the University of Illinois at Urbana-Champaign, affiliated with the Siebel School of Computing and Data Science. Her research focuses on trustworthy machine learning, emphasizing robustness, privacy, and security in AI systems. She leads the Secure Learning Lab (SL²), exploring adversarial attacks and defenses across digital and physical domains. Key contributions include foundational work on adversarial examples, robust learning frameworks, and privacy-preserving techniques. Her academic roles include advisory board positions at the Center for Artificial Intelligence Innovation (CAII) and membership in the Information Trust Institute (ITI). She collaborates with institutions like the Advanced Digital Science Center (ADSC) and the Quantum Information Science and Technology Center (IQUIST). Notable recognitions include the IJCAI Computers and Thought Award (2022), MIT Technology Review's 35 Innovators Under 35 (2020), and multiple best paper awards. Recent work addresses AI safety through frameworks like ShieldAgent and AutoRedTeamer, aiming to enhance system resilience against adversarial threats. Her research spans theoretical guarantees, practical defenses, and ethical AI deployment. Students advised include Chulin Xie, Linyi Li, and Boxin Wang, who have received prestigious fellowships such as the IBM PhD Fellowship and Rising Stars in ML Awards. Advising: Guides PhD students in adversarial ML, privacy, and security. Labs/Teams: Secure Learning Lab (SL²), collaboration with ALERT program. Grants/Funding: NSF CAREER Award, Amazon/Google Faculty Awards, and industry partnerships.
Maksym Andriushchenko is an incoming faculty member at the ELLIS Institute Tübingen and Max Planck Institute for Intelligent Systems , where he will lead the AI Safety and Alignment group as a principal investigator. Currently, he is a Doctoral Assistant at the Theory of Machine Learning Laboratory (TML) within the Department of Computer Science (IINFCOM) at Swiss Federal Institute of Technology in Lausanne (EPFL) . His work focuses on AI safety, adversarial robustness, and alignment of large language models (LLMs) with societal values.
Elisa Ricci is a Full Professor at the Department of Information Engineering and Computer Science (DISI) at the University of Trento and serves as Head of the Research Unit Deep Visual Learning at Fondazione Bruno Kessler. She coordinates the Doctoral Program in Information Engineering and Computer Science at the University of Trento and holds prestigious fellowships from ELLIS and IAPR. Her research focuses on advancing computer vision and deep learning systems capable of operating in open-world environments. Key interests include domain adaptation, continual learning, and self-supervised learning for visual and multi-modal data processing. Her work addresses critical challenges in enabling machines to adapt to new domains without forgetting prior knowledge, with applications spanning robotics perception, medical imaging, and privacy-preserving AI systems. Recent publications reveal a dominant trend toward leveraging vision-language models for open-vocabulary tasks, training-free adaptation methods, and federated learning architectures. Significant research thrusts include machine unlearning for privacy, robustness against bias in visual classifiers, and novel class discovery using foundation models—particularly evident in 2025 publications addressing medical imaging, 3D segmentation, and collaborative generative systems. Her major recognitions include: ELLIS Fellow IAPR Fellow As Doctoral Program Coordinator at the University of Trento, she oversees PhD training while leading the Deep Visual Learning unit at Fondazione Bruno Kessler. Her research group secures competitive grants in AI-driven perception systems, though specific funding sources aren't detailed in the source material. The Deep Visual Learning Research Unit specializes in open-world computer vision challenges, developing frameworks for domain adaptation, continual learning, and multi-modal perception. Current projects integrate generative models with robotics applications while addressing privacy concerns in vision-language systems through unlearning techniques.
Raquel Fernández is Full Professor of Computational Linguistics and Dialogue Systems at the University of Amsterdam, where she leads the Dialogue Modelling Group at the Institute for Logic, Language & Computation (ILLC). As Vice-Director for Research at ILLC and a Fellow of the ELLIS Society, she bridges computational linguistics, cognitive science, and artificial intelligence through her research on language use in multimodal and conversational contexts. PhD in Computational Linguistics from King's College London Prior research positions at University of Potsdam and Stanford University's CSLI Her work explores how cognitive constraints, social interaction, and perception shape language use, with a focus on: Visually-grounded language processing Multimodal dialogue modeling Model uncertainty and calibration Language grounding in multimodal data Language learning and semantic change Dialogue reference resolution Recent publications analyze multimodal reasoning limitations, cross-lingual knowledge consistency, and uncertainty modeling in dialogue systems. She has received multiple accolades including an ERC Consolidator Grant , NWO VENI/VIDI/Aspasia fellowships , and EMNLP/GenBench awards . Outstanding Paper Award (EMNLP 2023) Best Data Award (GenBench Workshop 2023) ELLIS Society Fellow ERC Consolidator Grant #819455 recipient NWO VENI/VIDI/Aspasia awardee As a leader in academic service, she serves on the SIGDAT Executive Committee and chairs multiple conference committees. Her lab develops models for multimodal dialogue, visual storytelling, and grounded language understanding.
James Glass is a Senior Research Scientist at the Massachusetts Institute of Technology (MIT) and heads the Spoken Language Systems Group within MIT's Computer Science and Artificial Intelligence Laboratory (CSAIL). He is also affiliated with the Harvard-MIT Division of Health Sciences and Technology. His research spans automatic speech recognition, multimodal learning, and spoken language understanding, with applications in healthcare and video analysis. Education: SM and PhD in Electrical Engineering and Computer Science from MIT His work focuses on paralinguistic speech analysis, health markers in speech, and the intersection of speech and natural language processing. Recent trends emphasize audio-visual alignment, recursive reasoning, and AI applications in cognitive disorder diagnosis. Scientific awards include IEEE Fellow, ISCA Fellow, and Associate Editor for IEEE Transactions on Pattern Analysis and Machine Intelligence. His group explores unsupervised learning, speaker verification, and social text analysis. James leads the Spoken Language Systems Group at CSAIL, collaborating with institutions like IBM and Harvard-MIT Division of Health Sciences and Technology. His research integrates vision-language models, neural audio codecs, and self-supervised frameworks.
Shiyu Chang is an Associate Professor in the Department of Computer Science at the University of California, Santa Barbara , focusing on machine learning with applications in natural language processing and computer vision . He previously worked as a research scientist at the MIT-IBM Watson AI Lab alongside Prof. Regina Barzilay and Prof. Tommi Jaakkola, and earned both his B.S. and Ph.D. in Computer Science from the University of Illinois at Urbana-Champaign , advised by Prof. Thomas S. Huang. Education : PhD, University of Illinois at Urbana-Champaign BS, University of Illinois at Urbana-Champaign His research centers on enhancing AI systems through human-AI interaction , aiming to improve interpretability , transferability , and adversarial robustness in LLMs. Recent work includes LLM watermarking defense , uncertainty decomposition , and self-denoised smoothing for model robustness. His publications span premier venues like ICML , NeurIPS , CVPR , and ACL , with recurring themes in diffusion models , LLM optimization , and ethical AI (e.g., hallucination detection, unlearning frameworks). He actively mentors students, several of whom are marked as advisees (☆) in his publications.