Haibin Ling is the SUNY Empire Innovation Professor in the Department of Computer Science at Stony Brook University, part of the College of Engineering and Applied Sciences. His research focuses on computer vision, medical image analysis, augmented reality, and AI applications in science. He holds a Ph.D. from the University of Maryland (2006) and prior degrees from Peking University. Previously, he worked at Temple University (2008–2019) and held roles at Siemens Corporate Research, UCLA, and Microsoft Research Asia. Professor Ling's work spans biomedical imaging, AI for science, and human-computer interaction. He leads the CV Lab and collaborates with the AI Institute at Stony Brook. Awards include the NSF CAREER Award (2014), Best Student Paper (ACM UIST 2003), and IEEE Fellow (2020). He serves on editorial boards for IEEE Trans. PAMI, Pattern Recognition, and CVIU, and chairs major conferences like CVPR. His research group includes over 50 students and alumni, with active projects in tracking benchmarks (LaSOT), Leafsnap, and medical imaging tools. Notable publications address OCTA flow estimation, backdoor attacks on vision models, and topology-guided medical learning. Collaborations involve institutions like Temple University and Stony Brook's Department of Applied Mathematics and Statistics.
David B. Lindell is an Assistant Professor in the Department of Computer Science at the University of Toronto, with affiliations to the Vector Institute and AXL. He is a founding member of the Toronto Computational Imaging Group. His research focuses on physically based intelligent sensing, integrating physical models, signal processing, and AI to advance sensing systems. Notable projects include imaging around corners, through scattering media, and developing machine learning algorithms for 3D scene reconstruction. Education: Ph.D. in Computational Imaging from Stanford University (advisor: Gordon Wetzstein). Awards include the 2024 Ontario Early Researcher Award and the Best Student Paper at CVPR 2025. His work combines computational imaging with applications in computer graphics and autonomous systems. Research interests span non-line-of-sight imaging, single-photon sensing, and neural representations. Key contributions include the Light-Cone Transform (Nature 2018), confocal diffuse tomography (Nature Communications 2020), and AutoInt (CVPR 2021). His lab develops systems for 3D reconstruction, transient imaging, and photon-efficient sensors. Selected grants and support: NSF CAREER Award, DARPA REVEAL program, and KAUST Visual Computing Center funding. Active collaborations with industry and academic institutions on autonomous driving and medical imaging applications.
Ravi Ramamoorthi is the Ronald L. Graham Professor of Computer Science and Director of the UC San Diego Center for Visual Computing. He holds a faculty position in the Department of Computer Science and Engineering (CSE) and is an affiliate of the Department of Electrical and Computer Engineering (ECE). He joined UC San Diego in 2014, previously at UC Berkeley and Columbia University. He also holds a part-time appointment as a Distinguished Research Scientist at NVIDIA. His research focuses on visual computing, including rendering, computer vision, light field cameras, and physics-based modeling. Notable contributions include foundational work on spherical harmonic lighting, neural radiance fields (NeRF), and Monte Carlo rendering techniques. His work bridges graphics, vision, and signal processing with applications in sparse reconstruction, importance sampling, and real-time rendering. He teaches courses like CSE 167 (Computer Graphics), CSE 168 (Rendering), and advanced topics in computer graphics. Awards include ACM and IEEE Fellowships, the Okawa Foundation Grant, and multiple Frontiers of Science Awards. His research is supported by NSF, ONR, and industry collaborators including Adobe, Sony, and Qualcomm. Key projects include the Center for Visual Computing, Light Field research, and educational initiatives like edX MOOCs on computer graphics and rendering. His work has influenced industry tools (e.g., Pixar, RenderMan) and modern real-time rendering pipelines with denoising techniques.
Jonathan T. Barron is a Researcher at Google DeepMind in San Francisco, specializing in Computer Vision , Neural Rendering , and 3D Scene Reconstruction . He earned his PhD at UC Berkeley under Jitendra Malik and has pioneered advancements in NeRF (Neural Radiance Fields) and diffusion-based 3D generation. Research Interests : Computer Vision, Deep Learning, Generative AI, Image Processing, and 3D Reconstruction via Radiance Fields. His work includes Bolt3D for rapid 3D scene generation, CAT3D/CAT4D for text-to-3D/4D, and Zip-NeRF for anti-aliased radiance fields. He has also developed real-time rendering frameworks like SMERF and NeRF-Casting for reflections. Scientific awards: PAMI Young Researcher Award He has served as Area Chair for CVPR, ICCV, and NeurIPS, and his research is widely adopted in applications like Google's Lens Blur , Portrait Mode , and Jump VR .
Andreas Geiger is a Professor and Head of the Department of Computer Science at the University of Tübingen, Germany. He leads the Autonomous Vision Group (AVG) within CyberValley and is a core faculty member of the Tübingen AI Center. His roles also include PI in the ML in Science Excellence Cluster and the CRC Robust Vision, as well as ELLIS Fellow and coordinator of the ELLIS PhD program. He specializes in machine learning models for computer vision, robotics, and autonomous systems, with applications in self-driving cars, VR/AR, and scientific document analysis. Educational background: While not explicitly detailed, his positions imply a Ph.D. in Computer Science or related field. His work spans interdisciplinary collaborations with institutions like ETH Zürich, Microsoft, and the University of Bonn. Research focuses on 3D scene understanding, Gaussian splatting, generative models, and reliable autonomous systems. Notable contributions include the KITTI dataset and foundational work in neural radiance fields. Awards include the Sage 10-Year Impact Award (2024), ERC Starting Grant (2019), and IEEE PAMI Young Researcher Award (2018). Key projects include the Scholar Inbox paper recommender platform, ReSim (reliable world simulation), and advancements in 3D scene generation (e.g., UrbanCAD, PrITTI). His lab maintains a strong focus on open-source tools and datasets, such as the CARLA Route Generator. Grants and funding include support from Vector Stiftung (MINT innovation program) and EU initiatives like the ML in Science Cluster. His team collaborates internationally, with recent work presented at CVPR, SIGGRAPH, and NeurIPS.
Angel Xuan Chang is an Associate Professor at Simon Fraser University's School of Computing Science, affiliated with labs including 3DLG, GrUVi, SFU NatLang, SFU AI/ML, and VINCI. He holds a Canada CIFAR AI Chair and was a TUM-IAS Hans Fischer Fellow (2018-2022). His research bridges natural language processing (NLP), 3D scene understanding, and embodied AI, focusing on language-grounded 3D generation and biodiversity monitoring via DNA barcodes. Recent work includes NuiScene (unbounded outdoor scene generation), ViGiL3D (3D visual grounding dataset), and CLIBD (vision-genomics biodiversity analysis). He advises students in projects like BIOSCAN-5M insect dataset and embodied AI navigation. His 2025 highlights include multiple ICCV and ICLR papers, workshops at ICML and CVPR, and a CRV invited talk. Education: Ph.D. in Computer Science from Stanford University (2014), advised by Chris Manning. Previous roles include visiting research scientist at Facebook AI Research and researcher at Eloquent Labs.
Prof. Dr. Otmar Hilliges is a Full Professor at the Department of Computer Science at ETH Zurich. He leads the AIT lab and serves as the head of the Institute of Intelligent Interactive Systems. His research focuses on spatio-temporal understanding of human movement and interaction, leveraging algorithms and representations from videos, images, and sensor data for applications in Augmented Reality (AR), Virtual Reality (VR), and Human-Robot Interaction. Education: Diplom (MSc) in Computer Science, Technical University of Munich (TUM), Germany PhD in Computer Science, Ludwig Maximilian University of Munich (LMU), Germany (2009) Research Interests: Hilliges' work spans computer vision, robotics, and human-computer interaction. He develops methods for 3D human pose estimation, generative models for realistic avatar creation, and physically plausible simulation of human-object interactions. His research emphasizes practical applications in AR/VR and assistive robotics, aiming to bridge the gap between perception and action. Grants & Contributions: ERC Consolidator Grant (2022-2027): 'AI-Perceive: Robust Human-Centric Computer Vision for Advanced AI-Agents' Google Research Agreement (2020-2025): 'Generative Modelling of Humans' Microsoft Research Grants: Focus on human-centric robotics and interactive technologies Labs & Teams: Leads the AIT Lab at ETH Zurich, which pioneers research in intelligent interactive systems, emphasizing human-centric AI and robotics. The lab collaborates on projects ranging from drone cinematography to haptic feedback systems.
Georgia Gkioxari is an Assistant Professor in the Division of Computing and Mathematical Sciences at Caltech , with a part-time affiliation at Meta AI . Her work focuses on extending visual perception models through advanced 2D and 3D representation learning, spatial reasoning, and generative models. Education: Not explicitly mentioned in the text Research interests span 3D perception , spatial reasoning , and vision-language integration , with projects like Visual Agentic AI for Spatial Reasoning and Token-by-Token Multimodal Alignment . Her publications emphasize 3D object detection , reconstruction , and generative modeling techniques including diffusion models and transformers . Scientific recognition includes the Meta LLM Evaluation Research Grant , Okawa Research Grant , Google Faculty Scholar Award 2024 , and Amazon Research Award . She teaches courses like Large Language & Vision Models (EE/CS 148) and Learning & 3D (CS 101) at Caltech. Labs & Teams: Leads Glab with members including Ilona Demler, Ziqi Ma, and Damiano Marsili
Professor Winston Hsu is a distinguished faculty member in the Department of Computer Science and Information Engineering at National Taiwan University, where he has served as a full professor since 2015. He is the co-director of the Communications and Multimedia Laboratory (CMLab) and founder of the MiRA (Multimedia indexing, Retrieval, and Analysis) research group. Additionally, he serves as the Founding Director for NVIDIA AI Lab at NTU, the first such lab in Asia. Professor Hsu received his Ph.D. in Electrical Engineering from Columbia University in 2007 under the supervision of Professor Shih-Fu Chang. Prior to his academic career, he was a founding engineer and research manager at CyberLink Corp., now a public image/video software company. National Taiwan University (2007-Present): Professor (2015-Present), Assistant/Associate Professor (2007-2015) MobileDrive (2021-2024): CTO and Vice President (Joint Venture between Foxconn and Stellantis) IBM TJ Watson Research Center (2016-2017): Visiting Scientist Microsoft Research Redmond (2014): Visiting Researcher Columbia University (2007): Ph.D. in Electrical Engineering Professor Hsu's research focuses on machine learning, computer vision, large-scale image and video search and recognition, and embedded AI. His work spans from fundamental research in visual recognition to practical applications in automotive systems, medical imaging, and e-commerce. He has pioneered work in disguised face recognition, low-resolution face hallucination, 3D model search, and virtual try-on systems. His current research emphasizes Embodied AI, integrating perception, action, and learning technologies for applications in automotive and robotics domains. His research group has produced numerous influential publications, particularly in top computer vision and multimedia conferences like CVPR, where they won first place in the Disguised Face Recognition competition in 2018. Their work spans diverse application areas including security, medical diagnostics, automotive systems, and e-commerce solutions, demonstrating strong translation from academic research to real-world impact. IBM Research Pat Goldberg Memorial Best Paper Award (2018) First Place, IARPA Disguised Faces in the Wild Competition (CVPR 2018) Best Brave New Idea Paper Award, ACM Multimedia 2017 NVIDIA AI LAB Award (First in Asia, 2016) First Place, MSR-Bing Image Retrieval Challenge (2013) World's Top 2% Scientists (2023) Professor Hsu actively mentors students and researchers, with his group consistently recruiting PhD students, postdocs, and research assistants. He has successfully bridged academia and industry through multiple collaborations, including his role as CTO at MobileDrive (a joint venture between Foxconn and Stellantis) from 2021-2024. His research has been supported by significant industry partnerships with Microsoft, IBM, and NVIDIA, as well as government grants from Taiwan's Ministry of Science and Technology. His laboratory, the Communications and Multimedia Laboratory (CMLab), maintains strong industry connections and focuses on cutting-edge research in visual AI. The lab has produced numerous award-winning projects and maintains active collaborations with global technology companies, particularly in the automotive and consumer electronics sectors.
Prof. Bernt Schiele is a Max Planck Director at the Max Planck Institute for Informatics and holds a Professorship at Saarland University. His research focuses on understanding multimodal sensor data, with key areas in computer vision, 3D object recognition, and machine learning. He leads the Computer Vision and Machine Learning group, addressing challenges in sensor fusion, scene understanding, and human activity recognition. Schiele has held academic roles at TU Darmstadt, ETH Zurich, and MIT, and contributes to top journals like IEEE Transactions on PAMI and conferences like ECCV. His work emphasizes robust models, interpretability, and domain adaptation for real-world applications. Education: PhD (1997, Grenoble), MSc (1994 Karlsruhe/1993 Grenoble) Key Positions: MIT (1997-2000), ETH Zurich (1999-2004), TU Darmstadt (2004-2010) Research interests span 3D scene understanding, multimodal sensor processing, and machine learning techniques for large-scale data. His recent work advances robust object detection, explainable AI, and domain-invariant training methods. He also chairs major conferences like ECCV 2018 and co-chairs ICCV 2011. Publications highlight innovations in interpretable vision transformers, certified explanations, and test-time adaptation. Despite no listed awards, his contributions shape foundational areas of computer vision and multimodal AI.
Xiaoming Liu is the Anil K. and Nandita Jain Endowed Professor of Engineering and MSU Foundation Professor in the Department of Computer Science and Engineering at Michigan State University . Holding a Ph.D. from Carnegie Mellon University (2004), he leads cutting-edge research in computer vision and machine learning. Research Interests : Computer Vision Pattern Recognition Image and Video Processing Machine Learning Medical Image Analysis Multimedia Retrieval Recent Research Trends : Focus on 3D object detection and depth estimation Development of robust biometric recognition systems Integration of radar-camera fusion for autonomous systems Advancements in self-supervised and multimodal learning Exploration of adversarial AI security Creation of interpretable forgery detection frameworks Teaching : Spring 2013: CSE891-006 Computer Vision Seminar Fall 2012-2015: CSE803 Computer Vision Spring 2014-2017: CSE 471 Media Processing and Multimedia Contact Information : Email: liuxm@cse.msu.edu Office: EB 3137, Michigan State University Phone: +1 (517) 355-2359
Takeshi Ikenaga is a Professor at Waseda University’s School of Fundamental Science and Engineering and Graduate School of Information, Production and Systems . He earned his Ph.D. in Information & Computer Science from Waseda University in 2001, following B.E. and M.E. degrees in Electrical Engineering (1988–1990). His career spans roles at NTT LSI Laboratories (1990–2002), Kitakyushu Foundation for Advancement of Industry, Science and Technology (FAIS) (1999–2002), and visiting researcher at the University of Massachusetts (1999–2000). Research Interests : Application-specific SoCs for video/image processing, including compression (H.264/AVC, H.265/HEVC), filters (super-resolution, noise reduction), recognition systems (feature detection, object tracking), and communication (UWB, LDPC). He also works on many-core processor design, ultra-low-delay vision systems, and sports analytics (volleyball, figure skating) with real-time 3D pose estimation and ball tracking. Awards : Recipient of the Furukawa Sansui Award (Waseda University, 1988) IEICE Research Encouragement Award (1992) Multiple Best Paper/Presentation Awards (2006–2022) at conferences including DAC/ISSCC, LSI IP Design, ISOCC, ISPACS, and CVIT APSIPA Distinguished Lecturer Certificate (2015) Waseda University Presidential Teaching Award (2020)
Yanzhi Wang is a Professor in the Department of Electrical and Computer Engineering at Northeastern University , affiliated with the Institute for Experiential AI and the Institute for the Wireless Internet of Things . He holds a PhD from the University of Southern California (2014). His research focuses on real-time AI systems, deep neural network compression, neuromorphic computing, and non-von Neumann architectures. Notable projects include NSF-funded initiatives on age-inclusive urban design, superconducting computing (DISCoVER), and edge device optimization (PatDNN). He has received prestigious awards such as the Army Research Office Young Investigator Award and the Constantinos Mavroidis Translational Research Award. His work emphasizes algorithm-hardware co-design for energy efficiency, with grants from NSF, ARO, and industry partners like Google. Recent research trends reflect his focus on accelerating vision transformers, diffusion models, and large language models for edge computing. He has pioneered methods like AutoViT and Fastcar, addressing latency and resource constraints in mobile platforms. Collaborations span academia and industry, driving innovations in superconducting circuits and neuromorphic systems.
Michael J. Black is a Professor and Honorarprofessor at the University of Tübingen's Faculty of Science, Department of Computer Science, and a founding Director of the Max Planck Institute for Intelligent Systems, leading the Perceiving Systems department. He holds a B.Sc. from the University of British Columbia (1985), M.S. from Stanford (1989), and Ph.D. in Computer Science from Yale (1992). His research focuses on computer vision, 3D human modeling, motion capture, and AI-driven digital humans. Key contributions include the SMPL body model, optical flow algorithms, and datasets like Middlebury Flow and Sintel. He has received major awards such as the PAMI Distinguished Researcher Award, multiple Koenderink and Longuet-Higgins Prizes, and is a member of the German National Academy of Sciences Leopoldina and Royal Swedish Academy of Sciences. His commercial ventures include co-founding Body Labs (acquired by Amazon) and Meshcapade, advancing 3D human generation and interaction technologies. Recent work includes markerless motion capture systems (e.g., MAMMA, PICO), 3D hair and garment synthesis, and AI tools like ChatHuman for 3D human interaction analysis. His research bridges vision, graphics, and robotics, with applications in animation, healthcare, and robotics.
Olga Fink is a Tenure Track Assistant Professor at the École Polytechnique Fédérale de Lausanne (EPFL), affiliated with the Department of Intelligent Maintenance and Operations Systems (IMOS) within the School of Architecture, Civil and Environmental Engineering (ENAC). She also holds roles in PhD program committees for Civil and Environmental Engineering (EDCE) and Robotics, Control, and Intelligent Systems (EDRS). Her research focuses on machine learning for infrastructure monitoring, predictive maintenance, and physics-informed AI models. She teaches courses on machine learning, data science for infrastructure, and advanced deep learning topics. Fink advises multiple PhD students and is involved in interdisciplinary projects such as ThermoNeRF (multimodal 3D thermal modeling) and physics-informed neural networks for fault diagnostics. Her work bridges AI and engineering with applications in smart infrastructure, energy systems, and industrial IoT. Education: PhD in Engineering (inferred from role) Affiliations: IMOS Lab, ENAC-SGC, EPFL PhD Committees (EDCE, EDRS) Key Research Themes: Explainable AI, Digital Twins, Structural Health Monitoring, Domain Adaptation Her publications (2023–2025) emphasize robust AI for industrial systems, including fault detection in high-voltage equipment, multimodal data fusion, and physics-consistent models. She collaborates on EU and industry-funded projects, focusing on real-world applications like predictive maintenance and energy efficiency.