Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Vineeth N Balasubramanian is a Professor in the Department of Computer Science & Engineering at the Indian Institute of Technology Hyderabad, with affiliate faculty status in the Department of Artificial Intelligence. His research focuses on the intersection of deep learning, machine learning, and computer vision, emphasizing explainability, robustness, and real-world applications. He leads Lab 1055, which investigates problems such as Explainable and robust AI/ML systems Lifelong learning in evolving environments Multimodal vision-language models Applications in agriculture, autonomous navigation, and human behavior analysis His recent work includes causal reasoning in transformers, vision-language model capabilities, and drone-based object detection. Funded by organizations like Google, Microsoft, Intel, and DST, he has received multiple awards including the World's Top 2% Scientists (2022-23), INSA/INAE Fellowships, and Best Paper recognitions. Lab 1055 collaborates with institutions like CMU, UBC, and Monash University, contributing to cutting-edge advancements in AI.
Lerrel Pinto is an Assistant Professor of Computer Science at the Courant Institute of Mathematical Sciences at New York University (NYU), where he leads the General-purpose Robotics and AI Lab (GRAIL) as part of the CILVR research group. His work bridges the gap between theoretical machine learning and practical robotics applications, with a focus on enabling robots to generalize and adapt in real-world environments. Dr. Pinto received his undergraduate degree from IIT Guwahati, followed by a PhD from the Robotics Institute at Carnegie Mellon University (CMU). He then completed a postdoctoral fellowship at the University of California, Berkeley before joining NYU as faculty. His research program centers on robot learning and decision making, with several key thrusts that demonstrate his innovative approach to robotics. Pinto's work emphasizes large-scale learning techniques that leverage both extensive data and sophisticated model architectures. A significant portion of his research focuses on representation learning for sensory data, particularly developing methods that enable robots to make sense of visual, tactile, and auditory inputs. His lab has made notable contributions to reinforcement learning algorithms that allow robots to adapt to new scenarios with minimal retraining. Pinto also champions open-source robotics , developing affordable robot platforms that democratize access to robotics research. Analysis of Pinto's recent publications reveals a strong trend toward multimodal perception in robotics, integrating visual, tactile, and auditory information to create more robust robot systems. His work increasingly focuses on zero-shot and few-shot learning capabilities, enabling robots to handle novel situations without extensive retraining. There's also a clear progression toward general-purpose robotics , moving away from task-specific solutions toward more flexible systems that can handle diverse real-world challenges. Dr. Pinto's scientific contributions have been recognized with several prestigious awards: Sloan Research Fellowship (2025) NSF CAREER Award (2024) RAL Early Career Award (2024) Best Student Paper Award at ICRA (2016) Outstanding Paper Award at MFM-EAI workshop at ICML (2024) Best Paper Award at NGSM workshop at ICML (2024) Best Student Paper Award at RSS (2023) As an advisor, Pinto has mentored numerous students who have gone on to impactful careers in both academia and industry. His former PhD student Denis Yarats co-founded Perplexity.AI, while Mahi Shafiullah became a postdoc at UC Berkeley and Meta AI. Many of his Masters students have pursued PhDs at top institutions like CMU, MIT, and Stanford, or joined leading robotics companies including 1X, Fauna Robotics, and NVIDIA. Pinto's lab has secured significant research funding, including the NSF CAREER award and likely other grants supporting his robotics research program. The General-purpose Robotics and AI Lab (GRAIL) that Pinto leads brings together a diverse team of researchers working on cutting-edge robotics challenges. The lab maintains strong collaborations with industry partners and other academic institutions, facilitating technology transfer and real-world impact. GRAIL's research spans multiple robotics platforms and focuses on developing algorithms that enable robots to learn from diverse experiences and generalize across environments.
Heng Huang is the Brendan Iribe Endowed Professor in the Department of Computer Science and the Department of Electrical and Computer Engineering at the University of Maryland, College Park . He earned his Ph.D. in Computer Science from Dartmouth College and holds prior degrees from Shanghai Jiao Tong University. His research focuses on advancing the foundations and applications of artificial intelligence, particularly in machine learning, data mining, natural language processing, computer vision, and biomedical informatics . His work integrates large-scale optimization, fairness, and robustness in deep learning systems. Heng Huang’s recent publications demonstrate a strong trend in large language models, federated learning, model watermarking, continual learning, and medical image analysis . His work appears consistently in top venues like NeurIPS, ICML, CVPR, ICLR, and MICCAI, reflecting a broad impact across theoretical and applied AI. He actively mentors students and postdocs, seeking highly motivated researchers in machine learning and related domains. His work has significant implications for healthcare, privacy, and trustworthy AI.
Katerina Fragkiadaki is the JPMorgan Chase Associate Professor of Computer Science in the Machine Learning Department at Carnegie Mellon University. She works at the intersection of Artificial Intelligence, Computer Vision, Machine Learning, Language Understanding, and Robotics. PhD from GRASP Lab, University of Pennsylvania Postdoctoral researcher at UC Berkeley (with Jitendra Malik) and Google Research Recipient of NSF CAREER, DARPA Young Investigator, Amazon, Google, Sony, UPMC, and AFOSR awards Organizer of CoRL 2023 Workshop on Generalist Robots ICLR 2024 Program Chair, multiple area chair roles Her research group focuses on developing machines that autonomously improve world models through human-environment interactions, with specific emphasis on: Representation learning and video understanding 2D/3D unified vision-language models Generative simulation and reinforcement learning Real2Sim/Sim2Real robot learning Continual learning and spatial common sense 3D scene reconstruction and dynamics Recent publications highlight advancements in: 3D mesh generation with compositional transformers Unified 2D/3D perception frameworks Physics-aware generative models Diffusion-based robotic manipulation policies Embodied agents with memory prompting Awards include: 2024: DARPA Young Investigator Award 2023: Amazon Faculty Award 2022: Sony Faculty Research Award 2021: UPMC Faculty Research Award 2020: NSF CAREER Award 2019: Google Faculty Award Key collaborations span institutions including UC Berkeley, Google Research, Stanford, MIT, and University of Tsukuba. Her work bridges theoretical innovation with practical applications in: Autonomous robot manipulation 4D world modeling Language-grounded perception Visual dynamics prediction Embodied program synthesis Physics-based simulation engines
Arti Singh is an Assistant Professor in the Department of Agronomy at Iowa State University. Her research focuses on plant breeding, soybean diseases, genomics, and phenomics, with a strong emphasis on integrating artificial intelligence and high-throughput technologies into agricultural systems. She leads projects involving AI-driven disease identification, precision agriculture, and crop improvement strategies. Her expertise includes developing machine learning models for real-time weed and insect classification (e.g., WeedNet and InsectNet), deploying drones and ground robots for crop phenotyping, and leveraging genomic data to map traits like flowering time and disease resistance in legumes. Singh collaborates on initiatives like the AIIRA Institute for Resilient Agriculture and the BioTrove biodiversity dataset. Singh’s work spans plant stress phenotyping, digital twin technologies for plant sciences, and multi-sensor phenotyping for early disease detection. Her research bridges computational methods with traditional agronomy, aiming to enhance crop resilience and sustainability in the face of environmental challenges. Her recent projects include optimizing robotic navigation for precision agriculture, improving soybean yield estimation via video analysis, and dissecting genetic architectures of traits in mungbean and soybean using GWAS and genomic tools. She actively contributes to conferences and publishes in high-impact journals, advancing both foundational and applied aspects of agricultural science.
Lior Wolf is a Professor at the School of Computer Science, Tel Aviv University. Previously, he was a postdoctoral researcher at MIT's Center for Biological and Computational Learning (CBCL) under Prof. Tomaso Poggio and earned his PhD from Hebrew University of Jerusalem with Prof. Amnon Shashua. His educational background includes: PhD in Computer Science, Hebrew University of Jerusalem Postdoctoral Research, MIT CBCL Prof. Wolf's research centers on artificial intelligence with seminal contributions to deep learning, computer vision, and natural language processing. His work bridges theoretical foundations (e.g., attention mechanisms, transformer analysis) with practical applications in medical imaging, speech processing, and sign language technology. He pioneered methods for neural network interpretability, efficient sequence modeling, and multimodal fusion. Analysis of his 2023-2025 publications reveals dominant trends in large language model optimization (neuron pruning, attention analysis), efficient video generation, and cross-modal learning. His work increasingly integrates medical applications (fMRI/EEG analysis) while maintaining theoretical rigor in model architecture design. His scientific achievements include: Best paper award at EMNLP 2024 for 'Backward Lens: Projecting Language Model Gradients into the Vocabulary Space' Best paper award at SCIA 2023 for 'Gradient Adjusting Networks for Domain Inversion' Best paper award at FG 2021 for 'Generating Master Faces for Dictionary Attacks' Prof. Wolf mentors graduate students in the School of Computer Science and leads research at the ICRC building laboratory. His team collaborates with 'the friends of TAU' on projects spanning biometric security, medical imaging, and generative AI. Current work focuses on efficient transformers, neural network interpretability, and multimodal medical diagnostics.
Aishwarya Agrawal is an Assistant Professor at Université de Montréal in the Department of Computer Science and Operations Research (DIRO), affiliated with Mila – Quebec Institute of Artificial Intelligence and a Canada CIFAR AI Chair. She also serves as a research scientist at Google DeepMind, spending one day weekly there. Education: B.E. in Electrical Engineering (IIT Gandhinagar, 2014), Ph.D. in Computer Science (Georgia Tech, 2019). Her research focuses on multimodal learning , deep learning , natural language processing , and computer vision , particularly in developing AI systems that 'see' and 'communicate' effectively. Grants & Awards: Canada CIFAR AI Chair, 2020 Sigma Xi Best PhD Thesis Award, NVIDIA Fellowship (2018–2019), and multiple fellowships from Google and Facebook. She leads projects like Advancing Multimodal Vision-Language Learning (CRSNG-funded) and StarDoc: Document Structure Extraction (MITACS). Research Contributions: Pioneered benchmarks like CulturalVQA and UI-Vision , and frameworks such as PROGRESS for efficient VLM training. Her work emphasizes cross-modal alignment, robust evaluation, and cultural understanding in AI systems. Labs/Teams: Active in Mila’s core academic group and collaborates with Google DeepMind on multimodal and vision-language research. Supervises a dynamic team of PhD and master’s students in Montreal.
Almut Sophia Koepke is a junior research group leader at the Technical University of Munich and University of Tübingen, focusing on multimodal learning problems integrating sound, vision, and text. Her work bridges foundational research in audio-visual understanding with practical applications in few-shot learning, zero-shot translation, and cross-modal attention mechanisms.
Deva Ramanan is a Professor at the Robotics Institute of Carnegie Melllon University, where he leads research in computer vision and machine learning. His work focuses on modeling human visual perception, leveraging large-scale visual data, and developing systems for 3D understanding, neural rendering, and autonomous systems. He advises a large group of PhD students and has mentored numerous postdoctoral researchers now in leading roles across industry and academia. His research interests include computer vision, machine learning, human perception modeling, 3D scene understanding, neural rendering, autonomous driving, video understanding, and multimodal foundation models. These areas reflect his focus on both foundational models and their application to real-world problems in robotics and AI. The recent publications highlight a strong trend toward multimodal and 3D-aware models, with increasing use of diffusion models, neural fields, and large vision-language systems. Key themes include scene flow, 3D reconstruction from monocular video, autonomous driving perception, and robust evaluation of vision-language models. There is a clear emphasis on both methodological innovation and practical deployment in dynamic environments. Marr Prize, Honorable Mention (ICCV 2021) Best Paper, Honorable Mention (ECCV 2020) Best Paper Finalist (WACV 2024) Best Paper Award (WACV 2016) Best Industrial Paper, Honorable Mention (BMVC 2017) Marr Prize winner (ICCV 2009) Deva Ramanan has advised numerous PhD and master’s students, many of whom are now at top institutions and companies including Apple, Meta, Google, Nvidia, OpenAI, and Princeton. He has received substantial funding from IARPA, DARPA, NSF, Intel, Google, and Facebook for projects in video analytics, dispersed computing, visual cloud systems, and multi-task recognition. His group has developed influential datasets and benchmarks used widely in the community. He leads a vibrant research lab focused on advancing computer vision through deep learning and multimodal integration. His team works on core challenges in perception, including 3D reconstruction, motion modeling, object detection, and scene understanding, with applications in robotics and autonomous systems.
Dai Zhongxiang is an Assistant Professor and Presidential Young Fellow at the School of Data Science, The Chinese University of Hong Kong, Shenzhen (CUHKSZ), where he joined in August 2024. Previously, he was a Postdoctoral Associate at MIT's Laboratory for Information and Decision Systems (January-June 2024) and a Postdoctoral Fellow at the National University of Singapore's Department of Computer Science (April 2021-December 2023). He completed his Ph.D. in Artificial Intelligence at NUS under the supervision of Bryan Kian Hsiang Low and Patrick Jaillet. Dr. Dai's research focuses on the intersection of theoretical and practical AI, with particular emphasis on large language models (LLMs) and optimization techniques. His work spans both theoretical foundations of multi-armed bandits and Bayesian optimization, as well as practical applications in LLM inference, including prompt optimization, in-context learning, personalization of LLMs, LLM-based agents, and scaling up test-time computation of LLMs. His research approach often bridges theoretical principles with real-world applications, particularly in AI4Science problems. His recent publications demonstrate a clear trend toward advancing LLM capabilities through optimization techniques, with increasing focus on practical deployment challenges. The research spans both theoretical contributions to optimization theory and applied work on enhancing LLM performance in real-world scenarios. His work on dueling bandits, neural bandits, and zeroth-order optimization has been consistently published in top-tier venues including NeurIPS, ICML, ICLR, and ACL. Presidential Young Fellow, CUHKSZ (2024) Dean's Graduate Research Excellence Award, NUS (2021) Research Achievement Award × 2, NUS (2019 & 2020) Singapore-MIT Alliance Graduate Fellowship (2017) Dr. Dai actively mentors multiple Ph.D. students and research assistants, with several of his students' papers accepted to top conferences. His research has received significant attention, with invitations to serve as Area Chair for NeurIPS 2025 and ICLR 2025, reflecting his growing influence in the machine learning community. His work bridges theoretical machine learning with practical applications in large-scale AI systems.
WANG Ye is an Associate Professor in the Department of Computer Science at the School of Computing, National University of Singapore (NUS). He holds a PhD in Information Technology from Tampere University of Technology, Finland, and has been a tenured faculty member at NUS since 2002, following his industry research role at Nokia Research Center. He is the director of the Sound and Music Computing Lab at NUS, leading cutting-edge research in AI-driven music and health technologies. PhD, Information Technology, Tampere University of Technology, Finland (2002) MSc, Telecommunications, Braunschweig University of Technology, Germany (1993) BSc, Telecommunications, South China University of Technology, China (1983) His research is centered on Sound and Music Computing for Human Health and Potential (SMC4HHP) , with a focus on eHealth, eLearning, mobile/wearable computing, and music information retrieval. His work spans AI for stroke rehabilitation, language learning through singing, singing voice synthesis, and automatic music transcription. He has pioneered systems like SLIONS (language learning via karaoke), CocoLyricist (AI co-creation for stroke recovery), and SinTechSVS (expressive singing voice synthesis). The latest articles highlight a strong trend in AI-driven music and health technologies , particularly in controllable lyric generation, singing voice synthesis, automatic pronunciation assessment, and multimodal music transcription. The research increasingly integrates large language models, explainable AI, fairness, and real-world deployment, reflecting a shift from theoretical exploration to practical, human-centered applications in healthcare and education. Dr. Wang has received numerous scientific honors, including: Best Paper Awards at ACM MM, ISMIR, IEEE ISM, and CHI First Prize, Asia Pacific Assistive, Rehabilitative, and Therapeutic Technologies Challenge (2015) Faculty Teaching Excellence Award, NUS School of Computing (2024) Top Paper Award, ACM Multimedia 2022 AI in Medicine Collaborative Grant for CocoLyricist project He has supervised over 11 PhD and 20 MComp students and is currently guiding six PhD candidates. His grants come from MOE, NRF, A*STAR, Nokia, and Smule. He has served as General Chair of ISMIR2017 and TPC Co-Chair of ICOT2017, and is on the editorial boards of IEEE Transactions on Multimedia and Journal of New Music Research. He has also developed and taught the first course on Sound and Music Computing in Singapore. Dr. Wang leads the Sound and Music Computing Lab (SMC Lab) , a multidisciplinary team exploring the synergy of music computing, AI, mobile technology, and cloud systems for health and education. The lab actively collaborates with medical institutions such as NUS Yong Loo Lin School of Medicine, Singapore General Hospital, and Harvard Medical School, and is currently working on projects in AI-supported language learning, stroke rehabilitation, and intelligent music interfaces.
Timothy M. Hospedales is a Professor of Artificial Intelligence at the Institute of Perception, Action and Behaviour within the School of Informatics at the University of Edinburgh . He also serves as VP AI and Head of Samsung AI Research Centre Europe . His research focuses on efficient and robust AI , emphasizing meta-learning , lifelong transfer-learning , and domain adaptation in both probabilistic and deep learning frameworks. Applications span computer vision , vision and language , reinforcement learning for robotics , and finance . Professor at University of Edinburgh (2020–present) ELLIS Fellow (2021) Head of Samsung AI Research Europe (2020–present) Founding Director of Applied Machine Learning Lab at QMUL (2012–2016) His work includes pioneering contributions to meta-learning , few-shot learning , and self-supervised methods , with notable awards such as the Best Paper Prize at ICML AutoML 2018 and Best Student Paper at ICPR 2018 . He has co-authored 15+ recent papers on topics like Vision-Language Models , Medical AI Fairness , and Diffusion Model Optimization . He served as Program Co-Chair for BMVC 2018 and AAAI 2022 , and authored a book on Visual Adaptation in the Deep Learning Era (2022). Co-Chair, BMVC 2018 Guest Editor, IET CV Special Issue (2016) Keynote Speaker at TASK-CV Workshop (ECCV 2016) Special Issue on Fewer Labels (IEEE PAMI 2020) His leadership extends to organizing workshops like the Learning-to-Learn Workshop at ICLR 2021 , Meta-Learning Workshop at NeurIPS 2020 , and Domain Generalisation Workshop at ICLR 2023 . Current projects include Meta-Omnium (CVPR 2023) for general-purpose meta-learning and MetaAudio (ICANN 2022) for few-shot audio classification benchmarks.
Tim G. J. Rudner is an Assistant Professor in the Department of Statistical Sciences at the University of Toronto, a Faculty Member at the Vector Institute, and a Title A Fellow at Trinity College, University of Cambridge. He was previously an Assistant Professor and Faculty Fellow at New York University. University: University of Toronto School: Faculty of Arts and Science Department: Department of Statistical Sciences Affiliation: Vector Institute, Trinity College (Cambridge) He holds a PhD in Computer Science and an MSc in Statistics from the University of Oxford, where he was advised by Yee Whye Teh and Yarin Gal, and a BS in Applied Mathematics and Economics from Yale University. PhD: Computer Science, University of Oxford MSc: Statistics, University of Oxford BS: Applied Mathematics and Economics, Yale University His research focuses on building robust, transparent, and trustworthy machine learning systems, particularly for high-stakes applications. He develops probabilistic models that improve generalization under distribution shifts, provide reliable uncertainty estimates, and enable fair and interpretable predictions. His work spans generative models, large language models, healthcare, and biomedical discovery. The recent publications highlight a strong trend toward function-space modeling, Bayesian regularization, and AI safety. Tim's work emphasizes principled uncertainty quantification, robustness to subpopulation and semantic shifts, and the development of frameworks for AI governance and specification. His research bridges theoretical advances with real-world applications, especially in safety-critical domains like medicine and defense. Tim has received numerous accolades including being named a Rhodes Scholar, Qualcomm Innovation Fellow, and 2024 Rising Star in Generative AI. He was awarded a $700,000 Foundational Research Grant and a $30,000 Apple Seed Grant for improving LLM trustworthiness. Rhodes Scholar Qualcomm Innovation Fellow AISTATS Notable Paper Award (2024) Outstanding Paper Award, ICLR GenAI4DM Workshop (2024) Apple Seed Grant ($30,000) Foundational Research Grant ($700,000) NeurIPS Spotlight Talk 2024 Rising Star in Generative AI He actively mentors students, particularly first-generation and low-income scholars, and has contributed to major policy frameworks including the OECD AI Classification Framework and a series of CSET issue briefs on AI safety. His work demonstrates a strong commitment to responsible AI development, combining technical rigor with societal impact. Tim leads research efforts at the intersection of machine learning theory and practical deployment, with ongoing projects in generative modeling, reliable LLMs, and AI governance. His lab produces high-impact work regularly published at top-tier conferences such as NeurIPS, ICML, and AISTATS.
Prof. Michael Moor is a tenure-track Assistant Professor for Medical AI at ETH Zurich's Department of Biosystems Science and Engineering in Basel. Previously, he conducted postdoctoral research at Stanford University under Prof. Jure Leskovec, focusing on medical foundation models. His work spans causal learning, multimodal AI, and sepsis prediction. Moor holds an MD from the University of Basel and a PhD from ETH Zurich's Machine Learning and Computational Biology Lab under Prof. Karsten Borgwardt. Research Interests: Generalist medical AI models Multimodal medical reasoning Zero-shot and few-shot learning Clinical causal inference Retrieval-augmented language models Sepsis prediction systems Key Achievements: Published foundational work in Nature (2023) on generalist medical AI Developed Med-Flamingo multimodal model (2023) Zero-shot causal learning framework accepted to NeurIPS 2023 (Spotlight) International sepsis prediction study with 156k ICU patients (2023) Lab Activities: Leads ETH's Medical Foundation Models group, collaborating with Stanford HAI and NASA Ames. Active in creating medical AI benchmarks like AgentClinic and developing retrieval-augmented systems like Almanac.