Prof. Matthias Nießner is a Professor at the Technical University of Munich , where he leads the Visual Computing Lab . Prior to this, he held a Visiting Assistant Professor position at Stanford University . His work bridges computer vision , graphics , and machine learning , focusing on 3D reconstruction , semantic scene understanding , and AI-driven video synthesis . Prof. Nießner has published over 150 works in top venues like SIGGRAPH , CVPR , and ECCV , with several receiving best paper awards (SIGCHI’14, HPG’15, SPG’18, SIGGRAPH’16 Emerging Tech). His research has garnered international media attention, including features in the New York Times , Wall Street Journal , and MIT Technological Review , as well as TV demonstrations (e.g., Jimmy Kimmel Live for Face2Face technology). Awards : TUM-IAS Rudolph Moessbauer Fellowship (2017–ongoing) Google Faculty Award (2017) Nvidia Professor Partnership Award (2018) ERC Starting Grant (2018, €1.5M) Eurographics Young Researcher Award (2019) Research Trends : 3D Gaussian Splatting for real-time rendering Neural Radiance Fields (NeRF) with mesh supervision Audio-driven facial animation via diffusion models Latent space diffusion for 3D scenes Self-supervised and zero-shot methods for 3D and image analysis As a co-founder and director of Synthesia Inc. , he drives democratization of synthetic media. His YouTube channel has over 5 million views, reflecting his impact beyond academia.
Changxi Zheng is an Associate Professor in the Department of Computer Science at Columbia University's School of Engineering and Applied Science (SEAS). He directs Columbia's Computer Graphics Group (C2G2) within the Columbia Vision and Graphics Center (CVGC). After receiving his PhD from Cornell University, he joined the faculty of Computer Science Department at Columbia, where he has established himself as a leading researcher in computer graphics and scientific computing. Dr. Zheng's research spans multiple areas of applied computer science with a particular focus on computer graphics and scientific computing. His work centers around developing numerical models for simulating physical phenomena involving complex motions such as fluids, bubbles, and thin rods, along with their resulting acoustic waves. Leveraging computational insights from these models, he devises methods for improving tangible object creation, enabling novel human-computer interactions, and developing software tools for acoustic and photonic devices. His research has attracted significant public interest and media coverage, including projects like FontCode, AirCode, and Computational Metallophone Design. His recent publications reveal a strong interdisciplinary approach, bridging computer graphics, physics simulation, machine learning, and hardware design. His work demonstrates consistent innovation in computational methods for simulating physical phenomena and applying these techniques to practical problems in 3D printing, acoustic modeling, and interactive systems. The breadth of his research spans from fundamental physics-based simulations to practical applications in industry. Columbia SEAS Dean's Fellow (for advised students) NSF Graduate Research Fellow (for Ruilin Xu) Snap Research Fellow (for Rundi Wu) CKGSB Fellow (for Yun Fei) Adobe Research Fellow (for Gabriel Cirio) Marie Sklodowska-Curie Individual Fellow (for Rundi Wu) Best Paper Award at ACM International Conference on Multimedia (ACMMM), 2019 Dr. Zheng actively mentors a diverse group of students, including current PhD candidates and postdoctoral researchers. His research group has received support from various sources that enable their innovative work in computational graphics and physics-based simulation. He has supervised numerous successful students who have gone on to positions at leading technology companies including Adobe, Tencent, Facebook, and academic institutions. As director of Columbia's Computer Graphics Group (C2G2) within the Columbia Vision and Graphics Center (CVGC), Dr. Zheng leads a vibrant research team focused on advancing the state of the art in computer graphics, physics-based simulation, and their applications. The group maintains strong collaborations with industry partners and academic institutions worldwide, fostering an environment of innovation and practical application of theoretical concepts.
Mark Yatskar is an Assistant Professor in the Department of Computer and Information Science at the University of Pennsylvania. His research focuses on the intersection of natural language processing, computer vision, and fairness in machine learning. He earned his PhD from the University of Washington under advisors Luke Zettlemoyer and Ali Farhadi, and previously worked as a Young Investigator at the Allen Institute for Artificial Intelligence. Education: PhD in Computer Science, University of Washington (Advisor: Luke Zettlemoyer & Ali Farhadi) Research Interests: Yatskar's work explores how language can structure visual perception and mitigate human biases in machine learning systems. Key themes include: Natural language as a scaffold for visual intelligence Bias characterization and control in machine learning systems His lab currently investigates projects like language-guided bottlenecks, annotator cognitive heuristics, and gender bias amplification. Teaching: CIS 5300: Computational Linguistics (2021-2024) CIS 7000: Language and Vision (2020) CIS 6300: Efficient NLP (2023, 2025) Awards: Best Paper Award at EMNLP (Gender Bias Amplification Research) Advising & Grants: Yatskar advises a team of PhD/Master's students and actively seeks motivated researchers. His group has explored funding in areas like interpretable AI, multimodal reasoning, and dataset bias mitigation. Labs/Teams: Leads the Penn NLP & Vision Lab, focusing on projects like MolMo/PixMo open models, ViUniT visual unit tests, and bias mitigation frameworks.
Andrea Tagliasacchi is an Associate Professor at Simon Fraser University's School of Computing Science, holding the Visual Computing Research Chair. He is also a part-time (20%) staff research scientist at Google DeepMind (Toronto) and an associate professor (status-only) at the University of Toronto's computer science department. His research focuses on 3D visual perception at the intersection of computer vision, graphics, and machine learning. Education: EPFL – Postdoc Simon Fraser University – PhD (NSERC Alexander Graham Bell Fellow) Politecnico di Milano – MSc (Gold Medalist) Research Interests: His work emphasizes 3D reconstruction, neural fields, and applications in robotics, autonomous systems, and augmented reality. Recent advancements include scalable 3D Gaussian splatting, robust neural rendering techniques, and diffusion models for 4D generation. Notable Articles: Recent work spans real-time differentiable ray tracing, stochastic rasterization for 3D Gaussian splats, and generative image composition using neural fields. His publications often blend theoretical contributions with practical applications in CVPR, SIGGRAPH, and NeurIPS. Awards: 2015 SGP Best Paper Award 2020 CVPR Best Student Paper Award 2024 CVPR Best Paper Honorable Mention Advising & Grants: Advised 14+ PhD/MSc students (e.g., Baptiste Angles, Sara Sabour) and co-advised with notable figures like Geoffrey Hinton. Active in grants involving neural field compression, robotic perception, and generative AI. Labs & Teams: Leads a lab at SFU focused on 3D vision and neural fields, collaborating with industry partners like Google Brain and Samsung Research.
Professor Gabriel Brostow is a faculty member in the Department of Computer Science at University College London (UCL), where he leads research in Computer Vision and Human-Computer Interaction. He also serves as Chief Research Scientist and Senior Director of the R&D Team at Niantic, the company behind Pokémon GO. His work bridges academic research and industry applications, focusing on developing AI systems that enhance human capabilities through what he terms 'Human in the Loop AI'—now commonly referred to as Human-Centered AI. Brostow completed his BS in Electrical Engineering at UT Austin, followed by a PhD with Irfan Essa at Georgia Tech. He then pursued postdoctoral research with Roberto Cipolla's Computer Vision & Robotics Group at Cambridge University as a Marshall Sherfield Fellow, and with Marc Pollefeys in ETH Zurich's CVG Group. His research explores how AI, particularly Computer Vision, can serve as 'super-tools' for professionals across various domains including filmmaking, architecture, robotics, and scientific research. Specific interests include assistive technology for everyday life, authoring systems that maximize user effort, 3D reconstruction, depth estimation, and vision-language models. His work often involves creating systems that are validated through real-world human interaction to ensure practical utility. Analysis of his recent publications reveals a strong focus on practical applications of Computer Vision that directly interact with humans. His research spans 3D scene understanding, depth estimation, sketch-based interfaces, and multimodal AI systems. There's a clear emphasis on creating benchmarks and tools that facilitate human-AI collaboration, with applications in assistive technology, urban planning, filmmaking, and biodiversity monitoring. His work frequently appears at top conferences including CVPR, NeurIPS, ECCV, and CHI. Marshall Sherfield Fellowship Brostow actively mentors PhD students, with current advisees including Ross Murphy, Skanda Koppula, Gizem Unlu, Omiros Pantazis, and Jamie Watson. His alumni include numerous PhD graduates and MSc students who have gone on to successful careers in academia and industry. He emphasizes selecting students based on passion and potential rather than just academic credentials, valuing traits like helpfulness, drive, and hunger to learn. His research is supported through collaborations with major institutions and companies including DeepMind, MIT, and the University of Edinburgh. He leads a research group at UCL that collaborates closely with Niantic's R&D team, creating a unique bridge between academic research and industry application. His team's work frequently involves developing novel Computer Vision techniques that are validated through real-world human interaction, ensuring practical utility alongside technical innovation. The group explores blue-sky research problems with applications ranging from assistive technology to professional tools for filmmakers, architects, and scientists studying diverse environments.
Xiaoming Liu is the Anil K. and Nandita Jain Endowed Professor of Engineering and MSU Foundation Professor in the Department of Computer Science and Engineering at Michigan State University . Holding a Ph.D. from Carnegie Mellon University (2004), he leads cutting-edge research in computer vision and machine learning. Research Interests : Computer Vision Pattern Recognition Image and Video Processing Machine Learning Medical Image Analysis Multimedia Retrieval Recent Research Trends : Focus on 3D object detection and depth estimation Development of robust biometric recognition systems Integration of radar-camera fusion for autonomous systems Advancements in self-supervised and multimodal learning Exploration of adversarial AI security Creation of interpretable forgery detection frameworks Teaching : Spring 2013: CSE891-006 Computer Vision Seminar Fall 2012-2015: CSE803 Computer Vision Spring 2014-2017: CSE 471 Media Processing and Multimedia Contact Information : Email: liuxm@cse.msu.edu Office: EB 3137, Michigan State University Phone: +1 (517) 355-2359
Michael J. Black is a Professor and Honorarprofessor at the University of Tübingen's Faculty of Science, Department of Computer Science, and a founding Director of the Max Planck Institute for Intelligent Systems, leading the Perceiving Systems department. He holds a B.Sc. from the University of British Columbia (1985), M.S. from Stanford (1989), and Ph.D. in Computer Science from Yale (1992). His research focuses on computer vision, 3D human modeling, motion capture, and AI-driven digital humans. Key contributions include the SMPL body model, optical flow algorithms, and datasets like Middlebury Flow and Sintel. He has received major awards such as the PAMI Distinguished Researcher Award, multiple Koenderink and Longuet-Higgins Prizes, and is a member of the German National Academy of Sciences Leopoldina and Royal Swedish Academy of Sciences. His commercial ventures include co-founding Body Labs (acquired by Amazon) and Meshcapade, advancing 3D human generation and interaction technologies. Recent work includes markerless motion capture systems (e.g., MAMMA, PICO), 3D hair and garment synthesis, and AI tools like ChatHuman for 3D human interaction analysis. His research bridges vision, graphics, and robotics, with applications in animation, healthcare, and robotics.
Francesco Pilati is an Associate Professor at the Department of Industrial Engineering, University of Trento, where he serves as local coordinator for the scientific field ING-IND/17 (Industrial Plants and Logistic Systems). He chairs the research group on Industrial Plants, Production Systems, and Logistics, and teaches courses in Industrial Plants and Design of Digital Production and Assembly Systems. As coordinator of the Master's program in Management and Industrial Systems Engineering and University Coordinator for the EIT double degree in Zero-Defect Manufacture, Pilati bridges academic leadership with advanced manufacturing research. He has also served as Invited Lecturer at universities in Vienna and Göttingen. His research focuses on integrating environmental sustainability with technical-economic criteria through multi-objective optimization and impact assessment. Key areas include: Distribution networks and warehousing systems Manufacturing and assembly line design Hybrid energy production systems Digitization of manual production processes using depth cameras Recent publications highlight applications of Industry 4.0 technologies to pandemic safety, logistics optimization, and smart manufacturing. Pilati has received significant recognition including the Philip Morris Italia Empowering Research Award (2016) and Autostrade per l'Italia academic recognition. His editorial contributions include guest editing special issues on Digital Twins and Smart Factories in Q1 journals.
Ira Kemelmacher-Shlizerman is a Full Professor of Computer Science at the Paul G. Allen School of Computer Science & Engineering at the University of Washington and Director of the UW Reality Lab. She also serves as a Principal Scientist at Google, where she leads the Shopping Gen AI visuals teams focusing on Virtual Try-On, 3D, and product videos. Her research spans computer vision, computer graphics, and Generative AI, with particular contributions to virtual try-on technology, 3D modeling, and augmented reality applications. Professor Kemelmacher-Shlizerman's research interests focus on Generative AI applications in visual computing. Her work bridges the gap between theoretical computer vision and practical applications, particularly in e-commerce and virtual reality. She has made significant contributions to virtual try-on technology, 3D editing with generative models, and AI applications for shopping experiences. Her research combines deep learning with traditional computer vision techniques to solve challenging problems in image and video synthesis. Her recent publications demonstrate a strong trend toward Generative AI applications for visual shopping experiences, virtual try-on technology, and 3D content creation. The work spans multiple top conferences including CVPR, SIGGRAPH, and ICCV, with a focus on practical applications of computer vision and graphics. Her research has evolved from foundational work in face reconstruction and aging to current applications in virtual shopping and 3D content generation. Google faculty award Madrona prize GeekWire Innovation of the Year Award Covers of CACM and SIGGRAPH Best student paper honorable mention at CVPR'21 Best demo runner up MobiSys'22 Senior member of IEEE Distinguished Member of ACM Professor Kemelmacher-Shlizerman has successfully tech-transferred multiple research projects to industry. She founded Dreambit, a startup acquired by Meta, and previously built and launched the Face Movies feature at Google. She currently leads Google's Shopping Gen AI visuals teams, focusing on 10x improvements to shopping journeys. Her UW Reality Lab serves as a hub for AR/VR research with industry partnerships. She has mentored numerous PhD students who have become researchers in both academia and industry, with several publications featuring student co-authors receiving recognition at top conferences. Professor Kemelmacher-Shlizerman leads the Graphics and Imaging Laboratory (GRAIL) and the UW Reality Lab, which focuses on augmented and virtual reality research with industry partnerships including Google. The labs work on cutting-edge projects in virtual try-on, 3D modeling, and immersive experiences, bridging academic research with real-world applications.
Dr Theodor Dunkelgrün is a Senior Postdoctoral Researcher at Trinity College, University of Cambridge , and Academic Co-Ordinator of the Andrew W. Mellon Foundation-funded project Religious Diversity and the Secular University (2017–2022) at CRASSH. He was previously a postdoctoral fellow at St John’s College (2013–2018) and has taught across History, Divinity, and Classics at Cambridge, Bryn Mawr, UPenn, and the University of Antwerp. Education : Christelijk Gymnasium Sorghvliet (The Hague), Leiden University, University of Chicago (PhD, 2012). Visiting graduate work at Princeton (2005–06) and Oxford (2008). Research Focus : Intersections of biblical scholarship, history of the book, and interfaith intellectual exchange (1450–1900), particularly in the Low Countries and Mediterranean World. Examines how Jewish, Christian, and Muslim scholars engaged with the Hebrew Bible through technological shifts (print to photography) and institutional contexts. Publications include edited volumes like The Jewish Bookshop of the World (2020) and Bastards and Believers (2020), alongside articles on topics ranging from the Antwerp Polyglot Bible to Jewish converts in early modern Europe. Collaborations span institutions in Taiwan, Oxford, and Leiden, with ongoing work on global philological practices and the Hebrew Bible’s material history. Professional Engagement : Co-founded the Cambridge Seminar in Early Modern Scholarship and Religion (2017), and actively promotes academic freedom through refugee scholar initiatives. Affiliated with the Cambridge Forum for Jewish Studies, Centre for Material Texts, and international networks like EMoDiR and Ideas in Motion.
Devi Parikh is an Associate Professor at the School of Interactive Computing, Georgia Institute of Technology, and a Research Director at Meta’s FAIR lab. Her research focuses on generative models, AI for creativity, computer vision, and natural language processing. Education: B.S. in Electrical and Computer Engineering from Rowan University (2005), M.S. and Ph.D. in Electrical and Computer Engineering from Carnegie Mellon University (2007, 2009). Research interests include embodied AI, human-AI collaboration, and creative applications of AI. She has held visiting positions at Cornell, MIT, CMU, and others. Awards include NSF CAREER Award, IJCAI Computers and Thought Award, and multiple fellowships. Led development of Habitat , a platform for embodied AI research, and contributed to the Open Catalyst Project for renewable energy storage.
Dr. Iro Armeni is Assistant Professor of Civil and Environmental Engineering at Stanford University, leading the Gradient Spaces research group. Her interdisciplinary research bridges architecture, civil engineering, and computer vision to develop data-driven methods for sustainable and adaptive built environments. Professor Armeni's work focuses on creating gradient environments that blend physical and digital realities through mixed reality technologies. She develops computational methods for 3D scene understanding, generative design, and adaptive spaces that respond to human needs. Her research integrates AI with architectural design to improve sustainability, inclusivity, and reusability of built spaces. Current projects include 3D scene graph representations, automated BIM modeling from visual data, and neuro-symbolic approaches for design optimization. She has developed tools like HoloLabel (AR semantic labeling) and SemSpray (VR annotation) for construction information management. Professor Armeni holds a PhD from Stanford University, supported by a Google PhD Fellowship, and completed postdoctoral research at ETH Zurich with an ETH Fellowship. She teaches courses on Computer Vision for the Built Environment and Mixed Reality applications.
Aykut Erdem is an Associate Professor of Computer Engineering at Koç University, affiliated with the KUIS AI Center. He earned his PhD, MSc, and BSc in Computer Engineering from Middle East Technical University (METU), with visiting researcher experiences at Virginia Tech (2004) and MIT (2007). His research focuses on learning-based approaches for visual data understanding, including image editing, visual saliency estimation, and vision-language integration. Research Interests: Vision and Graphics, Machine Learning, Artificial Intelligence, Computer Vision, Natural Language Processing, Generative Artificial Intelligence. Recent work includes text-guided image/video editing, diffusion models for object removal, and GAN-based frameworks for domain adaptation. Scientific Awards: Young Scientist Award (BAGEP 2021) by Science Academy in Computer Engineering Best Paper Award at 5th Multimodal Learning and Applications Workshop (2022) Collaborations and Funding: Principal investigator for TUBITAK 1001 project on generative AI for visual data (2021-2024), Adobe Research Gift (2023), and co-investigator for multiple TUBITAK grants. Serves as Associate Editor for IEEE Transactions on Image Processing (2022-present).
James T. Enns is a Professor and Distinguished University Scholar in the Department of Psychology within the Faculty of Arts at the University of British Columbia. His research primarily focuses on the role of attention in human vision, with secondary interests in developmental psychology and human-machine interaction. Dr. Enns' research interests span perception, attention, vision, cognition, development, and human-machine interaction. His work explores how the human mind selects information, with particular emphasis on visual attention mechanisms. His laboratory research investigates the fundamental processes of visual perception and how attention modulates these processes across different contexts and developmental stages. Analysis of Dr. Enns' publication record reveals consistent engagement with visual perception, attentional mechanisms, and cognitive processing. His research demonstrates expertise in both theoretical frameworks and experimental methodologies related to visual attention, with applications spanning basic cognitive science to potential implementations in human-computer interaction systems. His work often bridges theoretical cognitive psychology with practical applications. Canadian Society for Brain, Behaviour and Cognitive Science Donald O. Hebb Distinguished Contribution Award (2013) Distinguished University Scholar, UBC (2004) Robert E. Knox Master Teaching Award (2004) Royal Society of Canada Fellow (2002) Killam Faculty Research Prize (1994) Killam Faculty Research Fellowship (1993) Society of Experimental Psychology Fellow Dr. Enns has served as Editor for the Journal of Experimental Psychology: Human Perception and Performance, and as Associate Editor for Psychological Science, Consciousness and Cognition, and Visual Cognition. His research has been supported by grants from NSERC, the Canadian Foundation for Innovation, the Australian Research Council, BC Health, and Nissan. He has authored textbooks on perception, edited research volumes on the Development of Attention, and published numerous scientific articles on vision, attention, and cognitive science. Dr. Enns is currently accepting graduate students and continues to mentor the next generation of cognitive scientists. Dr. Enns leads the UBC Vision Lab, which focuses on how the human mind selects information. The lab conducts research on visual attention, perception, and cognitive processes using a variety of experimental methodologies.
Alexei A. Efros is the Howard Friesen Professor in the EECS Department at UC Berkeley, affiliated with the Berkeley Artificial Intelligence Research (BAIR) Lab. Previously, he spent a decade at CMU's Robotics Institute and held a postdoc at the University of Oxford under Andrew Zisserman. He collaborates with INRIA/École Normale Supérieure in Paris. His research focuses on self-supervised learning, generative models, and visual data mining, with applications to robotics, computational photography, and art. Education & Academic Roles: Postdoc at Oxford (with Andrew Zisserman), faculty at CMU (2005–2015), currently at UC Berkeley. Teaches courses like CS 180/280A (Computer Vision) and CS 280 (Graduate Computer Vision). Research Interests: Self-supervised learning, generative models (e.g., diffusion models, inpainting), visual commonsense, and cross-modal reasoning. His work bridges computer vision and graphics, emphasizing data-driven approaches. Recent projects include Visual Jenga, Diffusion Models as Data Mining Tools, and Prioritized Generative Replay. Grants & Labs: Leads the Efros Research Group, advised over 40 PhD students (e.g., Jun-Yan Zhu, Tinghui Zhou). Collaborates with institutions like INRIA and NVIDIA. Active in grants related to AI, vision, and robotics. Labs/Teams: BAIR Lab (UC Berkeley), former affiliations with CMU Robotics Institute and Willow Team (INRIA/ENS Paris). Current lab focuses on generative AI, 3D perception, and visual reasoning.