Dr. Chang Xu is an Associate Professor in Machine Learning and Computer Vision at the University of Sydney's School of Computer Science. He holds a Bachelor of Engineering from Tianjin University and a PhD from Peking University. His research focuses on machine learning, data mining, and their applications in AI and computer vision, including multi-view learning, visual search, and face recognition. He is an ARC Future Fellow and a member of the Sydney Southeast Asia Centre and The Net Zero Institute. Education: B.E. in Engineering (Tianjin University), Ph.D. in Computer Science (Peking University). His research interests emphasize handling heterogeneous data, exploring data variety, and developing algorithms for robust AI systems. His work includes adversarial robustness, neural architecture search, and efficient deep learning models. Research trends in his articles include adversarial robustness in neural architectures, efficient vision transformers, multimodal 3D style transfer, and underwater image restoration. Key contributions span image restoration, video super-resolution, and lightweight network design. He has advised multiple PhD and master's students on topics like diffusion models, radar image synthesis, and graph similarity. Awards: ARC Future Fellow. Collaborations focus on cross-domain data integration and AI applications. His labs and teams explore generative models, robust learning, and scalable robotics policies. Recent work includes diffusion models for action segmentation and robust vision-language systems.
Deva Kannan Ramanan is a Professor at the Robotics Institute of Carnegie Mellon University , focusing on computer vision , machine learning , and human-centered robotics . His work bridges neurorobotics and visual perception , with applications in autonomous driving and 4D reconstruction . Research Topics Computer Vision 3-D Vision and Recognition Visual Servoing Neurorobotics Human-Centered Robotics Graphics & Creative Tools His recent publications in CVPR , ICRA , and ICCV emphasize 4D human reconstruction , neural rendering , and vision-language models for autonomous systems. He serves as General Chair of CVPR 2027 and Program Chair of CVPR 2018 , with IARPA funding for aerial-ground rendering (2023-2027). Current students include PhD candidates Sally Chen, Kangle Deng, and Zhiqiu Lin, while past advisees like Arun Vasudevan and Olga Russakovsky now hold positions at Amazon and Meta respectively.
Tien Tsin Wong is a Professor in the Department of Data Science & AI at Monash University, Australia. Previously, he served as a Professor at the Chinese University of Hong Kong (1999–2024) and held a Visiting Assistant Professor position at the Hong Kong University of Science and Technology (1998–1999). His research focuses on Generative AI, Computer Graphics, Computer Vision, and Computational Manga, with significant contributions to GPU techniques, image-based rendering, and multimedia compression. Education: He earned a B.Sc. (1992), MPhil (1994), and PhD (1998) in Computer Science from the Chinese University of Hong Kong. Research Interests: His work bridges computational techniques with artistic applications, particularly in manga and animation. Notable areas include generative models, diffusion-based video synthesis, and physically plausible scene generation. His research aligns with UN Sustainable Development Goals through innovations in education and digital accessibility. Awards : He has received the 2004 Young Researcher Award, 2005 IEEE Transactions on Multimedia Prize Paper Award, and two international invention medals (Geneva 2018, Asia Hong Kong 2019). Editorial Roles : He serves as an Associate Editor for Computer Graphics Forum , IEEE Transactions on Visualization and Computer Graphics , and Computational Visual Media . His editorial work underscores his influence in advancing visualization and graphics research. Labs/Teams : While not explicitly named, his collaborations span global institutions, focusing on computational manga, generative AI, and GPU-optimized techniques. His work often involves interdisciplinary teams addressing challenges in digital media and AI.
Derek W Hoiem is a Professor in the Siebel School for Computing and Data Science at the University of Illinois Urbana-Champaign, where he has been a faculty member since 2009. His research focuses on computer vision and related areas, and he is also the co-founder and Chief Science Officer of Reconstruct, an AI-based construction technology company. His educational background includes: PhD in Robotics, Carnegie Mellon University (2007) Beckman Postdoctoral Fellowship (2008) Prof. Hoiem's research spans computer vision, with a focus on object recognition, scene understanding, and graphics. His work also extends to mobile robotics and 3D scene reconstruction. He has made significant contributions in areas such as visual recognition, 3D modeling, and the application of computer vision in construction monitoring. His recent publications (2023-2025) demonstrate a strong focus on advancing multimodal understanding, particularly in region-based representations, 3D vision, and neural radiance fields. There is a clear trend towards integrating language and vision, improving efficiency in neural networks, and applying computer vision to real-world problems such as construction progress monitoring. His scientific awards and honors are extensive and include: IEEE Fellow (2022) University Scholar (2022) Koendrink Prize (2022) Dean's Award for Excellence in Research, Associate Professor (2021) Campus Distinguished Promotion Award (2015) Best Paper Award: IEEE Winter Conference on Applications in Computer Vision (WACV) (2015) CW Gear Junior Faculty Award (2014) IEEE PAMI Young Researcher Award (2014) Dean's Award for Excellence in Research, Assistant Professor (2014) Sloan Research Fellowship (2013) Intel Early Career Faculty Honor Program Award (2012) NSF CAREER Award (2011) ACM Doctoral Dissertation Award, Honorable Mention (2008) Carnegie Mellon University SCS Distinguished Dissertation Award (2008) Best Paper Award: IEEE Computer Vision and Pattern Recognition (CVPR) (2006) Prof. Hoiem has secured significant research funding, including an NSF CAREER award and an Intel Early Career Faculty award. He is also actively involved in technology transfer, having co-founded Reconstruct where he serves as Chief Science Officer. His teaching excellence is reflected in multiple "List of Teachers Ranked as Excellent" awards spanning from 2010 to 2021. Prof. Hoiem leads a research group at UIUC focused on computer vision and 3D scene understanding. Additionally, he co-founded and serves as Chief Science Officer at Reconstruct, which develops AI-based solutions for construction monitoring.
Jiatao Gu is an Assistant Professor in the Department of Computer and Information Science (CIS) at the University of Pennsylvania, with a part-time role as Staff Research Scientist at Apple (MLR). He holds a Ph.D. in Electrical and Electronic Engineering from the University of Hong Kong (2018) and a B.Eng. in Electronic Engineering from Tsinghua University (2014). His research focuses on generative machine learning and AI agent interaction with the physical world, emphasizing multi-modal systems spanning language, images, videos, and 3D. Key themes include efficient modeling , flexible architecture design , and scalable decision-making frameworks . 2025: ICLR paper on DART framework 2024: TMLR work on GFlowNet alignment 2023: NeurIPS research on diffusion stability 2022: ACL papers on speech translation Recent publications explore diffusion models for text-to-image synthesis, 3D reconstruction, and efficient sampling techniques. His work addresses fundamental challenges in attention mechanisms, entropy collapse, and multi-stage distillation while advancing non-autoregressive translation and vision-language reasoning . Prospective students can apply through his recruitment process at UPenn. Prior affiliations include Meta AI (FAIR Labs) and academic collaborations with institutions like New York University's CILVR Lab.
Jonathan T. Barron is a Researcher at Google DeepMind in San Francisco, specializing in Computer Vision , Neural Rendering , and 3D Scene Reconstruction . He earned his PhD at UC Berkeley under Jitendra Malik and has pioneered advancements in NeRF (Neural Radiance Fields) and diffusion-based 3D generation. Research Interests : Computer Vision, Deep Learning, Generative AI, Image Processing, and 3D Reconstruction via Radiance Fields. His work includes Bolt3D for rapid 3D scene generation, CAT3D/CAT4D for text-to-3D/4D, and Zip-NeRF for anti-aliased radiance fields. He has also developed real-time rendering frameworks like SMERF and NeRF-Casting for reflections. Scientific awards: PAMI Young Researcher Award He has served as Area Chair for CVPR, ICCV, and NeurIPS, and his research is widely adopted in applications like Google's Lens Blur , Portrait Mode , and Jump VR .
Georgia Gkioxari is an Assistant Professor in the Division of Computing and Mathematical Sciences at Caltech , with a part-time affiliation at Meta AI . Her work focuses on extending visual perception models through advanced 2D and 3D representation learning, spatial reasoning, and generative models. Education: Not explicitly mentioned in the text Research interests span 3D perception , spatial reasoning , and vision-language integration , with projects like Visual Agentic AI for Spatial Reasoning and Token-by-Token Multimodal Alignment . Her publications emphasize 3D object detection , reconstruction , and generative modeling techniques including diffusion models and transformers . Scientific recognition includes the Meta LLM Evaluation Research Grant , Okawa Research Grant , Google Faculty Scholar Award 2024 , and Amazon Research Award . She teaches courses like Large Language & Vision Models (EE/CS 148) and Learning & 3D (CS 101) at Caltech. Labs & Teams: Leads Glab with members including Ilona Demler, Ziqi Ma, and Damiano Marsili
Takeshi Ikenaga is a Professor at Waseda University’s School of Fundamental Science and Engineering and Graduate School of Information, Production and Systems . He earned his Ph.D. in Information & Computer Science from Waseda University in 2001, following B.E. and M.E. degrees in Electrical Engineering (1988–1990). His career spans roles at NTT LSI Laboratories (1990–2002), Kitakyushu Foundation for Advancement of Industry, Science and Technology (FAIS) (1999–2002), and visiting researcher at the University of Massachusetts (1999–2000). Research Interests : Application-specific SoCs for video/image processing, including compression (H.264/AVC, H.265/HEVC), filters (super-resolution, noise reduction), recognition systems (feature detection, object tracking), and communication (UWB, LDPC). He also works on many-core processor design, ultra-low-delay vision systems, and sports analytics (volleyball, figure skating) with real-time 3D pose estimation and ball tracking. Awards : Recipient of the Furukawa Sansui Award (Waseda University, 1988) IEICE Research Encouragement Award (1992) Multiple Best Paper/Presentation Awards (2006–2022) at conferences including DAC/ISSCC, LSI IP Design, ISOCC, ISPACS, and CVIT APSIPA Distinguished Lecturer Certificate (2015) Waseda University Presidential Teaching Award (2020)
Yanzhi Wang is a Professor in the Department of Electrical and Computer Engineering at Northeastern University , affiliated with the Institute for Experiential AI and the Institute for the Wireless Internet of Things . He holds a PhD from the University of Southern California (2014). His research focuses on real-time AI systems, deep neural network compression, neuromorphic computing, and non-von Neumann architectures. Notable projects include NSF-funded initiatives on age-inclusive urban design, superconducting computing (DISCoVER), and edge device optimization (PatDNN). He has received prestigious awards such as the Army Research Office Young Investigator Award and the Constantinos Mavroidis Translational Research Award. His work emphasizes algorithm-hardware co-design for energy efficiency, with grants from NSF, ARO, and industry partners like Google. Recent research trends reflect his focus on accelerating vision transformers, diffusion models, and large language models for edge computing. He has pioneered methods like AutoViT and Fastcar, addressing latency and resource constraints in mobile platforms. Collaborations span academia and industry, driving innovations in superconducting circuits and neuromorphic systems.
Rozenn Dahyot is a Professor of Computer Science at Maynooth University within the Faculty of Science & Engineering. She previously held roles as Assistant and Associate Professor in Statistics at Trinity College Dublin (2008-2021) and Lecturer in Computer Science (2005-2008). Her research interests bridge Digital Signal Processing, Computer Vision, Machine Learning, and Statistical Analysis. She organized the European Signal Processing Conference (EUSIPCO2021) in Dublin and served as President of the Irish Pattern Recognition and Classification Society (IPRCS) from 2014-2020. Her work spans topics like semantic scene understanding, CNN compression, and medical image segmentation. Key contributions include advancements in graph-based image analysis, reinforcement learning optimization, and AI-driven systems for disaster management. Dahyot is a member of IEEE, ACM, and EURASIP, contributing to both academic and industrial collaborations.
Aswin Sankaranarayanan is a Professor in the Department of Electrical and Computer Engineering at Carnegie Mellon University (CMU) , where he leads the Image Science Lab . His research focuses on computational photography , 3D shape estimation , and novel imaging system design . He earned his Ph.D. in Electrical and Computer Engineering (2009) from the University of Maryland and completed a postdoctoral fellowship at Rice University (2012) . Research Themes: Developing imaging systems that exploit low-dimensional signal models to overcome traditional sensing limitations Co-design of optics and processing algorithms for efficient sensing Application of non-linear signal models to high-dimensional data Advancing compressed sensing and big data processing techniques Scientific Recognition: SIGGRAPH 2023 Best Paper Award (Split-Lohmann Multifocal Displays) CVPR 2019 Best Paper Award (Fermat Paths for NLOS Reconstruction) NSF CAREER Award (2017) Dean’s Early Career Fellowship (2018-2021) Herschel Rich Invention Award (2016) Technical Contributions: His recent publications reveal expertise in non-line-of-sight shape reconstruction , VR/AR display systems , and biomedical imaging . Collaborations span institutions like University College London and University of Toronto.
Svetlana Lazebnik is a Full Professor and Willett Faculty Scholar in the Department of Computer Science at the University of Illinois at Urbana-Champaign (UIUC), part of the Grainger College of Engineering. She holds a Ph.D. from UIUC (2006) and previously served as an Assistant Professor at the University of North Carolina at Chapel Hill (2007–2011). Her research focuses on computer vision, including generative models for virtual try-on, image stylization, scene understanding, and joint modeling of images and language. She has advised numerous Ph.D. students and postdocs, many of whom now hold prominent academic and industry roles. Education: Ph.D. in Computer Science, UIUC (2006); supervised by Jean Ponce. Research Interests: Her work spans generative adversarial networks (GANs), diffusion models, virtual try-on systems (e.g., Dressing-in-Order, Street Try-On), exemplar-based stylization, and large-scale photo analysis. She has pioneered spatial pyramid matching and contributed to binary code learning for image retrieval. Key Awards: NSF CAREER Award (2008), Microsoft Research Faculty Fellow (2009), Sloan Research Fellow (2013), IEEE Fellow (2021), and the Longuet-Higgins Prize (2016) for her CVPR 2006 paper. Teaching: Recent courses include CS 444 (Deep Learning for Computer Vision), CS 543 (Computer Vision), and a Ph.D. Job Search Seminar. She has also taught at UNC Chapel Hill. Grants & Funding: Supported by NSF, Amazon, AWS, Microsoft, Sloan Foundation, Google, ARO, and Adobe. Notable grants include CCF 2348624 and IIS 1718221. Labs/Groups: Leader in the Illinois CS Vision Group, contributing to collaborative projects on embodied AI, multi-agent systems, and visual-semantic reasoning.
Dr. Frederick Li is an Associate Professor in the Department of Computer Science at Durham University, UK. He holds editorial roles as Associate Editor of Frontiers in Education (Digital Education) and Editorial Board Member of Virtual Reality & Intelligent Hardware. His research focuses on Computer Graphics, Machine Learning, Geometric Modelling, Collaborative Virtual Environments, Visual Aesthetics, and Educational Technologies. He earned his B.A. (Hons) and M.Phil. from The Hong Kong Polytechnic University and his Ph.D. in Computer Graphics from City University of Hong Kong. Prior roles include Assistant Professor at HK PolyU and project manager of a Hong Kong Government ITF-funded project. **Education**: B.A. (Computing Studies) and M.Phil. from HK PolyU; Ph.D. in Computer Graphics (CityU Hong Kong). **Research Interests**: His work spans mesh saliency detection, human-object interaction recognition, cloud modeling, face beautification, and educational technology. Recent achievements include awards for papers (e.g., Best Paper at ITiCSE 2014) and recognition such as EPSRC Peer Review College membership. He leads Durham's Undergraduate Board of Examiners and has been an external examiner at Northumbria University. **Awards**: Best Paper (ACM ITiCSE 2014), Outstanding Paper (ICALT 2013), EPSRC Peer Review College (2024), Outstanding BMVC 2024 Reviewer. **Grants & Labs**: His research is supported by grants from EPSRC and others. He collaborates with the Centre for Vision and Visual Cognition, VIViD, and AIHS group at Durham.
Professor Moncef Gabbouj is a distinguished academic and researcher currently serving as Professor of Signal Processing at the Department of Computing Sciences, Faculty of Information Technology and Communication Sciences, Tampere University, Finland. Previously, he held the same position at Tampere University of Technology before the merger in 2019. He has also held visiting professorships at prestigious institutions including Hong Kong University of Technology and Science, University of Southern California, and Purdue University. Ph.D. and MSc. in Electrical Engineering from Purdue University, USA (1989 and 1986) B.Sc. in Electrical Engineering from Oklahoma State University, USA (1985) Prof. Gabbouj's research spans multiple domains within signal and image processing, with a strong focus on machine learning applications. His primary research interests include artificial intelligence, machine learning, Big Data analytics, multimedia content-based analysis, indexing and retrieval, nonlinear signal and image processing, voice conversion, and video processing and coding. His work bridges theoretical advancements with practical applications across various industries, particularly in multimedia communications and biomedical applications. His extensive publication record demonstrates a clear evolution from traditional signal processing techniques toward more sophisticated machine learning and deep learning approaches. Recent work shows increasing focus on convolutional neural networks for various applications including ECG classification, video processing, financial time-series analysis, and image recognition tasks, reflecting the broader trend in the field toward deep learning methodologies while maintaining strong foundations in signal processing theory. IEEE Fellow (2011) Member, Finnish Academy of Science and Letters (2014) Knight, First Class, of the Order of the White Rose of Finland (2006) Nokia Foundation Recognition Award (2005) Nokia Foundation Visiting Professor Award (2012) Finnish Cultural Foundation for Art and Science Award (2017) TUT Foundation Grand Award (2015) Prof. Gabbouj has supervised 64 doctoral and 72 Master's theses, demonstrating his significant contribution to academic mentoring. His research has been supported by substantial funding, including research grants totaling 8.5 million Euro (2001-2015). He has served as Academy of Finland Professor during 2011-2015 and has been involved in numerous EU research projects including Horizon, ESPRIT, HCM, IST, COST, Tempus and Erasmus programs. As Editor, Guest Editor or member of the Editorial Board of 6 international scientific journals, he has significantly influenced the academic discourse in his field. He leads the Signal Analysis and Machine Intelligence (SAMI) research group at Tampere University and serves as the Finland Site Director of the NSF IUCRC funded Center for Visual and Decision Informatics. His research unit focuses on applying advanced machine learning techniques to solve complex problems in signal processing, computer vision, and multimedia analytics, with applications ranging from healthcare to multimedia communications and financial analysis.
Sabine Süsstrunk is a Full Professor and Director of the Images and Visual Representation Laboratory (IVRL) at EPFL's School of Computer and Communication Sciences. She holds a BS in Scientific Photography from ETH Zürich, MS from Rochester Institute of Technology, and PhD from University of East Anglia. Her career includes positions at Hewlett-Packard Labs and Corbis Corporation. Her research explores computational imaging, computational photography, color processing, computer vision, and image quality. Key interests include near-infrared applications, multispectral imaging, and computational aesthetics. Her work bridges hardware and software solutions for imaging challenges. Publications demonstrate consistent focus on advancing generative models (diffusion models, neural cellular automata), 3D reconstruction (NeRF variants), and media integrity (DeepFake detection). Recent trends show increased emphasis on 3D vision, robustness in generative AI, and video analysis. Awards & Honors: IS&T/SPIE Electronic Imaging Scientist of the Year (2013) Raymond C. Bowman Teaching Award (2018) EPFL AGEPoly IC Polysphere Award (2020) 8 Best Paper/Demo Awards Fellowships: IEEE, IS&T, ELLIS, AIIA She leads the IVRL lab and advises PhD candidates while serving as President of the Swiss Science Council. Research is supported through competitive grants and industry collaborations.