Dr. Chang Xu is an Associate Professor in Machine Learning and Computer Vision at the University of Sydney's School of Computer Science. He holds a Bachelor of Engineering from Tianjin University and a PhD from Peking University. His research focuses on machine learning, data mining, and their applications in AI and computer vision, including multi-view learning, visual search, and face recognition. He is an ARC Future Fellow and a member of the Sydney Southeast Asia Centre and The Net Zero Institute. Education: B.E. in Engineering (Tianjin University), Ph.D. in Computer Science (Peking University). His research interests emphasize handling heterogeneous data, exploring data variety, and developing algorithms for robust AI systems. His work includes adversarial robustness, neural architecture search, and efficient deep learning models. Research trends in his articles include adversarial robustness in neural architectures, efficient vision transformers, multimodal 3D style transfer, and underwater image restoration. Key contributions span image restoration, video super-resolution, and lightweight network design. He has advised multiple PhD and master's students on topics like diffusion models, radar image synthesis, and graph similarity. Awards: ARC Future Fellow. Collaborations focus on cross-domain data integration and AI applications. His labs and teams explore generative models, robust learning, and scalable robotics policies. Recent work includes diffusion models for action segmentation and robust vision-language systems.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Kristen Grauman is a Full Professor in the Department of Computer Science at the University of Texas at Austin, where she leads the UT Computer Vision Group. Her research focuses on computer vision and machine learning, with applications in visual recognition, video analysis, and multi-modal perception. She received her B.A. from Boston College and her Ph.D. from MIT. Her research interests span visual recognition, image and video search, video analysis, first-person vision, embodied and multi-modal perception, and interactive machine learning. She has made significant contributions to the field, particularly in developing algorithms for understanding visual content and human activities from video, including foundational work on the Pyramid Match Kernel and relative attributes. Her recent publications reveal a strong emphasis on egocentric (first-person) vision, audio-visual learning, and view-invariant representations. There is a clear trajectory toward multi-modal integration (vision, audio, language) and real-world applications in instructional videos, human activity understanding, and embodied AI systems. She has received numerous awards including: AAAI Fellow (2019) J. K. Aggarwal Prize, International Association for Pattern Recognition (2018) Helmholtz Prize (2017) UT Austin Academy of Distinguished Teachers (2017) Best Paper Award, Asian Conference on Computer Vision (2016) Presidential Early Career Award for Scientists and Engineers (2014) Computers and Thought Award, International Joint Conferences on Artificial Intelligence (2013) Pattern Analysis and Machine Intelligence Young Researcher Award (2013) Alfred P. Sloan Research Fellow (2012) Marr Prize (2011) Prof. Grauman serves as Associate Editor-in-Chief for the IEEE Transactions on Pattern Analysis and Machine Intelligence. She has secured substantial research funding including the Presidential Early Career Award, NSF grants, and industry partnerships. Her advising has produced numerous influential publications and students who are now leaders in computer vision. She leads the UT Computer Vision Group, which collaborates closely with the Electrical and Computer Engineering Department. The group is pioneering large-scale egocentric video research through projects like Ego4D and Ego-Exo4D, focusing on real-world applications in human activity understanding, audio-visual perception, and interactive systems.
Noah Snavely is a Professor of Computer Science at Cornell Tech and the Cornell Ann S. Bowers College of Computing and Information Science. His research spans computer vision and graphics with a focus on 3D scene understanding from image collections. He also works at Google DeepMind in NYC and leads the Cornell Graphics and Vision Group. His research interests include computer vision and computer graphics , particularly in recovering 3D structure from large photo collections, neural radiance fields, image-based rendering, and scene understanding. His work enables applications in mapping technologies, immersive VR experiences, and synthesizing 3D worlds from text prompts with implications for game design and filmmaking. Recent publications show a strong trend toward generative 3D modeling and neural rendering , with significant contributions in neural radiance fields, view synthesis, and dynamic scene reconstruction. His research group explores unstructured photo collections to develop new technology for 3D world modeling and image analysis. PECASE Microsoft New Faculty Fellowship Alfred P. Sloan Fellowship SIGGRAPH Significant New Researcher Award ACM Fellow (2023) IEEE Fellow NSF CAREER Award Helmholtz Prize Snavely has mentored numerous PhD students and postdocs including Ruojin Cai, Qianqian Wang, and Zhengqi Li. His teaching includes CS5670: Computer Vision across multiple semesters. He has received the Cornell College of Engineering Mr. and Mrs. Richard F. Tucker Teaching Excellence Award (2012). At Cornell Tech, his research group develops technology for modeling the world in 3D from online photo collections and analyzing images for computer graphics applications. His work has been featured in major news outlets including the New York Times, New Scientist, and MIT Technology Review.
Adriana Kovashka is an Associate Professor at the University of Pittsburgh , affiliated with the School of Computing and Information and serving as Department Chair . Her academic journey began with BA degrees in Computer Science and Media Studies from Pomona College (2008) and a PhD in Computer Science from The University of Texas at Austin (2014). Joined Pitt’s faculty in January 2015 NSF CAREER awardee (2021) Google Faculty Research Award recipient Dr. Kovashka’s research spans Computer Vision , Machine Learning , and Natural Language Processing , focusing on visual rhetoric, weak multimodal supervision, and domain adaptation. She pioneered techniques for analyzing political imagery, developing robust object detection frameworks, and exploring the intersection of visual and textual persuasion through large-scale annotated datasets. Her recent work emphasizes geographic diversity in vision-language systems, audio-visual fusion for domain generalization, and shape-texture bias mitigation in CNNs. Key publications include groundbreaking studies on symbolic reasoning, multimodal dialogue systems, and ethical AI applications in education. Scientific honors include: NSF CRII Award (2016) NSF CAREER Award (2021) Pitt CRDF Award (2016, 2018) Best Paper at ECV Workshop (2021) Google Faculty Research Award (2016, 2018) Dr. Kovashka actively mentors students in multimodal learning projects and collaborates with interdisciplinary teams on NSF-funded initiatives. She co-organizes workshops like the first CVPR workshop on advertisement understanding and leads research groups exploring human-AI co-learning systems.
Olga Sorkine-Hornung is a Professor of Computer Science at ETH Zurich and head of the Institute of Visual Computing. She leads the Interactive Geometry Lab, focusing on theoretical and practical advancements in digital content creation, geometry processing, and shape modeling. Current position: ETH Zurich, Department of Computer Science Previous roles: Courant Institute (NYU), Technical University of Berlin Education: BSc and PhD from Tel Aviv University, postdoc at TU Berlin Her research spans shape representation, digital fabrication, computer animation, and fundamental geometry processing. Key contributions include Laplacian surface editing, as-rigid-as-possible deformation, and generalized winding numbers. She works on applications in VR, AR, and autonomous systems. Awards include Test of Time Awards (2024), ACM Fellow (2020), ERC Consolidator Grant (2020), and EUROGRAPHICS Young Researcher Award (2008). She has supervised numerous students and co-developed software libraries like libigl and Instant Meshes . Co-chair roles for SIGGRAPH, Eurographics, and Pacific Graphics Editorial board member for ACM Transactions on Graphics and other journals Keynote speaker at VMV, CVPR, and SIAM conferences Her work bridges mathematical rigor with practical implementation, advancing computer graphics and geometry processing through intuitive algorithms that maintain surface detail while enabling efficient computation.
Mathieu Salzmann is a Senior Scientist and Lecturer at École Polytechnique Fédérale de Lausanne (EPFL), affiliated with the Computer Vision Laboratory (CVLAB) in the School of Computer and Communication Sciences (IC). He also holds a courtesy appointment with the EPFL College of Humanities and serves as Deputy Chief Data Scientist at the Swiss Data Science Center (SDSC). He has held concurrent roles in teaching units including SIN, SODH, and SSC, reflecting his interdisciplinary engagement. His research focuses on the intersection of machine learning and computer vision, particularly in deep learning for 2D and 3D visual scene understanding, efficient and robust models, domain adaptation, and interpretable AI. These interests are evident across his extensive publication record in top-tier venues. His recent publications (2023–2024) show a consistent trend in advancing deep learning methods for visual recognition, with strong representation at CVPR, ICCV, ECCV, ICML, ICLR, and NeurIPS. Topics include domain generalization, 3D understanding, model robustness, and multimodal learning, often with applications in real-world systems. His editorial roles as Associate Editor for IEEE TPAMI and Action Editor for TMLR further highlight his leadership in the field. Area Chair: ICML 2023, CVPR 2023, ICCV 2023, NeurIPS 2023, AAAI 2024, ECCV 2024 Associate Editor: IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) Action Editor: Transactions on Machine Learning Research (TMLR) Mathieu Salzmann has supervised numerous PhD students at EPFL, both current and past, including Bouquet Yann Yanis, Javed Saqib, Li Shuangqi, and others. He has also been involved in research grants and collaborative projects, such as his work with S. Süsstrunk and R. Baroni on comics reconfiguration. His part-time role as Senior GNC Engineer at ClearSpace (2020–2024) illustrates his applied research engagement in aerospace systems. He is actively involved in EPFL’s data science and AI research ecosystem through SDSC and multiple labs.
Lior Wolf is a Professor at the School of Computer Science, Tel Aviv University. Previously, he was a postdoctoral researcher at MIT's Center for Biological and Computational Learning (CBCL) under Prof. Tomaso Poggio and earned his PhD from Hebrew University of Jerusalem with Prof. Amnon Shashua. His educational background includes: PhD in Computer Science, Hebrew University of Jerusalem Postdoctoral Research, MIT CBCL Prof. Wolf's research centers on artificial intelligence with seminal contributions to deep learning, computer vision, and natural language processing. His work bridges theoretical foundations (e.g., attention mechanisms, transformer analysis) with practical applications in medical imaging, speech processing, and sign language technology. He pioneered methods for neural network interpretability, efficient sequence modeling, and multimodal fusion. Analysis of his 2023-2025 publications reveals dominant trends in large language model optimization (neuron pruning, attention analysis), efficient video generation, and cross-modal learning. His work increasingly integrates medical applications (fMRI/EEG analysis) while maintaining theoretical rigor in model architecture design. His scientific achievements include: Best paper award at EMNLP 2024 for 'Backward Lens: Projecting Language Model Gradients into the Vocabulary Space' Best paper award at SCIA 2023 for 'Gradient Adjusting Networks for Domain Inversion' Best paper award at FG 2021 for 'Generating Master Faces for Dictionary Attacks' Prof. Wolf mentors graduate students in the School of Computer Science and leads research at the ICRC building laboratory. His team collaborates with 'the friends of TAU' on projects spanning biometric security, medical imaging, and generative AI. Current work focuses on efficient transformers, neural network interpretability, and multimodal medical diagnostics.
Adam Runions is a researcher in the Department of Computer Science at the University of Calgary, leading the MPG Partner Group in computational analysis of leaf development through collaborative work with Miltos Tsiantis. His group is embedded in the Graphics Cluster, focusing on interdisciplinary problems at the intersection of computer science and developmental biology. University of Calgary - Department of Computer Science MPG Partner Group (2022) Graphics Cluster affiliation His research explores computational modeling and analysis of plant form and development across multiple scales, integrating geometric modeling, physically-based simulation, and computer-aided design. Key themes include plant morphogenesis, self-organization of natural forms, and cross-disciplinary applications in computer graphics and animation. Recent publications emphasize plant development (leaf shape, bark patterning), mathematical modeling (auxin-driven patterning), and geometric techniques (subdivision surfaces, PUPs). Collaborations span institutions like the Max Planck Institute for Plant Breeding Research. Scientific Awards Marie Sklodowska-Curie Fellowship Best Paper Award (International Conference on Cyberworlds 2015) Best Student Paper Award (Computer Graphics International 2011) The group actively recruits BSc, MSc, and PhD students with backgrounds in computer science and mathematics for projects on plant form simulation and digital content creation. Research integrates evolutionary biology, biomechanical modeling, and computational techniques.
Jiatao Gu is an Assistant Professor in the Department of Computer and Information Science (CIS) at the University of Pennsylvania, with a part-time role as Staff Research Scientist at Apple (MLR). He holds a Ph.D. in Electrical and Electronic Engineering from the University of Hong Kong (2018) and a B.Eng. in Electronic Engineering from Tsinghua University (2014). His research focuses on generative machine learning and AI agent interaction with the physical world, emphasizing multi-modal systems spanning language, images, videos, and 3D. Key themes include efficient modeling , flexible architecture design , and scalable decision-making frameworks . 2025: ICLR paper on DART framework 2024: TMLR work on GFlowNet alignment 2023: NeurIPS research on diffusion stability 2022: ACL papers on speech translation Recent publications explore diffusion models for text-to-image synthesis, 3D reconstruction, and efficient sampling techniques. His work addresses fundamental challenges in attention mechanisms, entropy collapse, and multi-stage distillation while advancing non-autoregressive translation and vision-language reasoning . Prospective students can apply through his recruitment process at UPenn. Prior affiliations include Meta AI (FAIR Labs) and academic collaborations with institutions like New York University's CILVR Lab.
Xingang Pan is an Assistant Professor in the College of Computing and Data Science at Nanyang Technological University (NTU), leading the MMLab@NTU. His research focuses on generative AI and visual content creation, particularly in generative models, 3D vision, computer graphics, and computer vision. Prior to NTU, he was a postdoc at the Max Planck Institute for Informatics and earned his Ph.D. from the Chinese University of Hong Kong (2021) and B.Sc. from Tsinghua University (2016). His work emphasizes generative intelligence, exploring long-term world simulation, diffusion models, and multi-scale 3D generation. Notable contributions include WORLDMEM (2025), Alias-free Latent Diffusion (2025), and SAR3D (2025). His research has been published in top venues like CVPR, ICCV, and SIGGRAPH. Xingang Pan oversees the MMLab@NTU, which actively recruits students globally without nationality constraints. The lab’s projects include GAN2Shape (unsupervised 3D reconstruction from 2D GANs) and LN3Diff (scalable 3D generation).
Maria Gorlatova is an Associate Professor of Electrical and Computer Engineering at Duke University's Pratt School of Engineering, where she leads the Intelligent Interactive Internet of Things (I3T) Lab. She also holds a secondary affiliation as Faculty Network Member of the Duke Institute for Brain Sciences and has previously served as Assistant Professor of Computer Science. Dr. Gorlatova earned her Ph.D. in Electrical Engineering from Columbia University (2013), following M.Sc. and B.Sc. (Summa Cum Laude) degrees in Electrical Engineering from University of Ottawa, Canada. Prior to joining Duke, she was an Associate Research Scholar in the Electrical Engineering Department and Associate Director of the Princeton EDGE Lab at Princeton University (2016-2018). She also has industry experience with Telcordia Technologies, IBM, and D. E. Shaw Research. Her research focuses on advancing intelligent behavior in Internet of Things systems and applications, particularly in mobile pervasive systems and the Internet of Things. Her work crosses traditional discipline boundaries, requiring thinking across multiple layers of system and protocol stacks. Current research themes include breaking barriers for technologies that enable fundamentally new deployments and experiences, such as energy harvesting, artificial intelligence adapted to IoT constraints, and augmented reality. Her lab specifically develops edge- and IoT-enabled intelligent augmented reality platforms, with applications in healthcare and human-robot collaboration. Analyzing her recent publications reveals a strong focus on augmented reality systems, particularly for medical applications. Her work spans computer vision for AR, spatial tracking, SLAM systems, vision-language models for AR security, and VR/AR applications in neurosurgery and rehabilitation. A significant portion of her recent work addresses challenges in mixed reality for medical procedures, demonstrating the translational impact of her research. Google Anita Borg USA Fellowship Canadian Graduate Scholar CGS NSERC Fellowships Columbia University Presidential Fellowship Columbia University Jury Award for Outstanding Achievement in Communications ACM SenSys Best Student Demonstration Award IEEE Communications Society Young Author Best Paper Award IEEE Communications Society Award for Advances in Communications Best Research Artifact Award, IEEE IPSN (2020) N2 Women Rising Star, Networking Networking Women (N2Women) (2019) Dr. Gorlatova's research has been supported by various funding sources that enable her work on edge computing for augmented reality, IoT systems, and medical applications. She actively mentors graduate students who frequently appear as first authors on her publications, indicating strong student involvement in her research. Her I3T Lab at Duke focuses on creating human-facing pervasive mobile computing platforms that enable transformative applications, with recent emphasis on creating advanced augmented reality platforms that integrate edge computing and IoT technologies. The I3T Lab is developing next-generation AR systems with capabilities in edge AI, collaborative spatial awareness, AR user cognitive context sensing, and AR QoS/QoE evaluation. Current projects include applications in healthcare (particularly neurosurgery guidance and rehabilitation) and human-robot collaboration scenarios, demonstrating the lab's focus on real-world impact of pervasive computing technologies.
Prof. Olga Sorkine Hornung is a Full Professor of Computer Science at ETH Zürich, leading the Interactive Geometry Lab. She holds a BSc and PhD from Tel Aviv University (2000 and 2006) and conducted postdoctoral research at Technical University Berlin. Her research focuses on computer graphics, geometric modeling, and geometry processing, with applications in shape editing, digital fabrication, and animation. She has received numerous accolades, including the ACM Fellowship (2020), ERC Consolidator Grant (2020), and the Golden Owl Teaching Award (2021). Her work bridges theoretical foundations and practical algorithms, addressing challenges in parameterization, surface compression, and interactive design tools. Her research interests span: Computer Graphics & Visualization Geometric Modeling & Processing 3D Content Creation & Digital Fabrication Garment Design & Simulation Human Motion Analysis & Animation Awards and grants include: 2024: Best Paper Honorable Mention (EUROGRAPHICS) 2023: Member of Swiss Academy of Engineering Sciences (SATW) 2020: ERC Consolidator Grant 2017: Rössler Prize (ETH Zurich) Her lab focuses on developing novel methods for interactive geometry processing, with recent advancements in garment modeling (e.g., AIpparel, Rags2Riches) and motion retargeting systems like WalkTheDog. She actively collaborates on interdisciplinary projects, including biomedical applications and sustainable fashion technology.
Yu-Ru Lin is an Associate Professor at the School of Computing and Information, University of Pittsburgh, and serves as Research & Academic Director at the Institute for Cyber Law, Policy and Security (Pitt Cyber). She leads the Pitt Computational Social Dynamics Lab (PICSO Lab) and holds secondary appointments in Political Science, Computer Science, and the Intelligent Systems Program. PhD in Computer Science from Arizona State University Postdoctoral research at Harvard University and Northeastern University Her research focuses on computational approaches for: Networked social dynamics High-dimensional social information summarization Trust and distrust propagation Misinformation detection Policy diffusion analysis Recent publications span 2014-2020, covering: Social media crisis response Graph visualization techniques Policy diffusion patterns Misinformation mitigation Temporal topic modeling Scientific support includes: National Science Foundation (NSF) awards Minerva/ONR funding AFOSR grant for distrust modeling DARPA Understanding Group Biases program As director of PICSO Lab, she leads interdisciplinary research teams on: Digital accountability Urban mobility patterns Health informatics via crowdsourcing Trust-influence dynamics
Roles & Affiliations: Manolis Savva is an Associate Professor at Simon Fraser University's School of Computing Science and holds a Canada Research Chair in Computer Graphics. He leads research in 3D scene understanding, with applications in graphics, vision, and robotics. Previously, he was a researcher at Facebook AI and Princeton University. Education: Ph.D. in Computer Science (2016), Stanford University, advised by Pat Hanrahan B.A. in Physics and Computer Science (2009), Cornell University Research Interests: His work focuses on analyzing, organizing, and generating 3D content, particularly for holistic scene understanding. Key areas include articulated objects, embodied AI, and datasets like ScanNet , Matterport3D , and Habitat . His methods drive applications in robotics, autonomous agents, and virtual environments. Publications Trends: Recent work emphasizes generative models (e.g., SINGAPO for articulated object parts), embodied AI benchmarks (Habitat), and multimodal scene analysis. His papers often address challenges in scalability, realism, and cross-modal fusion for 3D environments. Awards: CHCCS Early Career Researcher Award (2022) ICLR 2023 Outstanding Paper Award ICCV 2019 Best Paper Nomination (Habitat) Advising & Grants: Supervised over 15 graduate students, many advancing to top PhD programs and tech companies. Active in grants for embodied AI, scene understanding, and robotics. Labs/Teams: Leads the 3DLG (3D Learning Group) and GrUVi (Graphics and Vision) groups at SFU. Collaborates extensively with industry (e.g., Meta, NVIDIA) on AI-driven 3D research.