Geoffrey E. Hinton is a distinguished Professor in the Department of Computer Science at the University of Toronto. He is renowned for his foundational contributions to machine learning, particularly in the development of deep learning and neural networks. His research focuses on understanding learning processes in both artificial and biological systems, with key contributions including Boltzmann machines, backpropagation, and deep belief networks. He teaches advanced machine learning courses such as CSC2535, emphasizing topics like graphical models, variational inference, and deep learning architectures. His work has been published extensively in top journals and conferences, with recent papers exploring forward-forward algorithms, analog diffusion models, and scalable neural network training methods. Hinton has advised numerous PhD and master's students and collaborates with institutions like Vector Institute. He is a central figure in the global AI community, regularly presenting at conferences (e.g., 2023 talks on CBS 60 Minutes, BBC, and PBS). His lab focuses on advancing machine learning theory and applications, addressing challenges in vision, language, and generative models.
Greg Durrett is an Associate Professor in the Department of Computer Science at University of Texas at Austin, leading the TAUR Lab (Text Analysis, Understanding, and Reasoning ). His research focuses on advancing Large Language Models (LLMs) for knowledge-intensive tasks in medical information processing scientific discovery legal reasoning . He received his B.S. in Computer Science and Mathematics from MIT (2010) and Ph.D. in Computer Science from UC Berkeley (2016). His work develops techniques to train LLMs with new capabilities augment models for reliability assess model outputs improve reasoning frameworks . His 15 most recent publications (2021-2025) span knowledge propagation in LLMs chain-of-thought reasoning code generation benchmarks multi-modal reasoning fact verification discourse analysis . Scientific honors include NSF CAREER Award (2024) NSF grants (2018, 2024) Bloomberg Data Science Grant (2017) Facebook Fellowship (2014) Best Paper Finalist (EMNLP 2013) . Teaching: CS388: Natural Language Processing (graduate) CS371N: NLP (undergraduate) High school NLP module .
Olga Sorkine-Hornung is a Professor of Computer Science at ETH Zurich and head of the Institute of Visual Computing. She leads the Interactive Geometry Lab, focusing on theoretical and practical advancements in digital content creation, geometry processing, and shape modeling. Current position: ETH Zurich, Department of Computer Science Previous roles: Courant Institute (NYU), Technical University of Berlin Education: BSc and PhD from Tel Aviv University, postdoc at TU Berlin Her research spans shape representation, digital fabrication, computer animation, and fundamental geometry processing. Key contributions include Laplacian surface editing, as-rigid-as-possible deformation, and generalized winding numbers. She works on applications in VR, AR, and autonomous systems. Awards include Test of Time Awards (2024), ACM Fellow (2020), ERC Consolidator Grant (2020), and EUROGRAPHICS Young Researcher Award (2008). She has supervised numerous students and co-developed software libraries like libigl and Instant Meshes . Co-chair roles for SIGGRAPH, Eurographics, and Pacific Graphics Editorial board member for ACM Transactions on Graphics and other journals Keynote speaker at VMV, CVPR, and SIAM conferences Her work bridges mathematical rigor with practical implementation, advancing computer graphics and geometry processing through intuitive algorithms that maintain surface detail while enabling efficient computation.
Aniket (Niki) Kittur is a Professor in the Human-Computer Interaction Institute (HCII) at Carnegie Mellon University's School of Computer Science. His research focuses on AI-augmented cognition, exploring how humans and computational systems can partner to accelerate knowledge acquisition and innovation. With over 100 peer-reviewed publications and 17 best paper awards or honorable mentions, Dr. Kittur has established himself as a leader in human-computer interaction and social computing. Dr. Kittur's research spans multiple interconnected areas focused on enhancing human cognition through technology. His work on crowd-augmented cognition examines how crowds and computation can collectively enhance human intellect. In social computing, he investigates how sociotechnical architectures can motivate and coordinate large groups for complex tasks. His research in data visualization and sensemaking explores how people collect, organize, and make decisions with online information. Most recently, his work has shifted toward LLM-augmented cognition, investigating how humans and large language models can work together in partnership to achieve better decisions and greater creativity than either could alone. Dr. Kittur's publications reveal a consistent trajectory toward increasingly sophisticated human-AI collaboration systems. Early work focused on crowdsourcing complex tasks (CrowdForge, CrowdSynthesis), then evolved to knowledge acceleration systems (Knowledge Accelerator, Fuse), and has recently centered on LLM-integrated cognition (Selenite, InkSpire, BioSpark). His research consistently demonstrates practical impact, with systems influencing products used by millions and receiving coverage in major media outlets including Nature News, The Economist, and The Wall Street Journal. Kavli fellow NSF CAREER award recipient Allen Newell Award for Research Excellence Inductee into the CHI Academy 17 best paper awards or honorable mentions across his publications Dr. Kittur has advised numerous students, including Andrew Kuznetsov, and has secured major research funding from NSF, NIH, Google, Microsoft, Bosch, Toyota, and other organizations. His Skeema browser extension project, which addresses tab overload through innovative project-based organization, has shown remarkable user retention and represents a practical application of his research on knowledge acceleration. His work with industry partners has influenced products used by millions, including Semantic Reader, Google Shopping, and Wikipedia features.
Alane Suhr is an Assistant Professor at UC Berkeley's Electrical Engineering and Computer Sciences (EECS) department and a member of the Berkeley Artificial Intelligence Research Lab (BAIR). Her research focuses on natural language processing (NLP), machine learning, and computer vision, emphasizing systems that interact with humans through language. She designs models and datasets for language grounding (e.g., NLVR) and develops algorithms for learning through interaction. Education: PhD in Computer Science, Cornell University (2022), advised by Yoav Artzi Bachelor's in Computer Science and Engineering, Ohio State University (2016), with a Linguistics minor Research Interests: Interactive systems for collaborative language use (e.g., CerealBar) Language grounding in multimodal contexts Embodied agents and reinforcement learning Generalization in NLP and the role of language in learning Key Contributions: Developed the NLVR dataset for visual reasoning with natural language Pioneered work on SWE-Gym for training software engineering agents Contributed to research on language models' sensitivity and commonsense reasoning Awards: ACM Doctoral Dissertation Award (2022) Outstanding Paper Awards at ACL 2023 and EMNLP 2021 Professional Activities: Organized workshops at ICML, NeurIPS, and ACL on topics like agents, generalization, and theory of mind Active in academic outreach and conference speaking (e.g., ICML, NeurIPS, CVPR) Labs/Teams: Berkeley Artificial Intelligence Research (BAIR) Lab SWE-Gym research group
Alexander Schwing is an Associate Professor in the Department of Electrical and Computer Engineering and Computer Science at the University of Illinois at Urbana-Champaign, affiliated with the Coordinated Science Laboratory. His research focuses on machine learning and computer vision with applications in 3D scene understanding, generative modeling, and multi-agent systems. Education: Diploma in Electrical Engineering and Information Technology, Technical University of Munich (TUM) PhD in Computer Science, ETH Zurich Postdoctoral Fellow, University of Toronto Research Interests: Structured prediction in deep learning Generative adversarial networks and stability Multi-modal vision-language models 3D scene reconstruction from single images Embodied agent collaboration Semantic segmentation with temporal coherence Recent Publications: Highlight trends in neural rendering, video object segmentation, and reinforcement learning with applications to 3D modeling and multi-agent systems. Notable innovations include SAIL-VOS dataset for amodal segmentation and NeRFDeformer for single-view scene transformation. Scientific Awards: NSF CAREER Award, 3M and Amazon research awards, multiple student recognition awards, ETH Zurich PhD medal, and best paper at Intelligent Tutoring Systems 2014. Teaching: Offers graduate courses in Pattern Recognition (ECE 544) and Machine Learning (CS 446/ECE 449). Previously taught at University of Toronto and ETH Zurich. Labs & Collaborations: Leads research at Coordinated Science Laboratory (UIUC) with collaborations across University of Toronto, ETH Zurich, and industry partners like Samsung SAIT and Amazon.
Prof. Matthias Nießner is a Professor at the Technical University of Munich, leading the Visual Computing Lab. His research intersects computer graphics, vision, and AI, focusing on 3D reconstruction, semantic understanding, and AI-driven video synthesis. He holds a PhD from the University of Erlangen-Nuremberg (2013) and was a Visiting Assistant Professor at Stanford University (2013–2017). Notable awards include the ERC Starting Grant (2018), Nvidia Professorship Award, and Eurographics Young Researcher Award (2019). His work has been featured in mainstream media and led to startups like Synthesia Inc. Research spans Gaussian splatting, neural radiance fields, and generative AI for 3D avatars. Over 150 publications include SIGGRAPH, CVPR, and ECCV, with best paper awards. Projects like Face2Face and ScanNet have driven innovation in facial reenactment and 3D scene datasets. Education: PhD in Computer Science, University of Erlangen-Nuremberg (2013) Diploma in Computer Science, University of Erlangen-Nuremberg (2010) Research Interests: 3D digitization, neural rendering, generative AI, non-rigid reconstruction, and applications in AR/VR. Awards: ERC Starting Grant (2018) Nvidia Professorship Award (2018) Google Faculty Award (2018) SIGGRAPH Best Emerging Tech Award (2016) Grants: Over €1.5M from ERC and industry partnerships. Labs/Teams: Visual Computing Lab at TUM and Synthesia Inc. (co-founder). Key projects include ScanNet (large 3D indoor dataset), Face2Face (real-time facial reenactment), and Gaussian-based 3D avatars. Current work focuses on diffusion models, neural radiance fields, and AI-generated media detection.
Alexei A. Efros is a Professor in the Department of Electrical Engineering and Computer Sciences (EECS) at UC Berkeley, where he holds the Howard Friesen Professorship and is affiliated with the Berkeley Artificial Intelligence Research (BAIR) Lab. He previously served on the faculty at the Robotics Institute of Carnegie Mellon University (CMU) and completed a postdoctoral fellowship at the University of Oxford. His research spans data-driven computer vision, self-supervised learning, computational photography, and applications to computer graphics and robotics. His research interests include: Data-Driven Computer Vision Self-Supervised and Unsupervised Learning Generative Models and Image Synthesis Visual Representation Learning Applications in Robotics and Human-Computer Interaction Intersections with Human Vision and the Humanities The recent publications highlight a strong trend toward self-supervised learning, visual reasoning, and generative modeling, particularly diffusion models and 3D scene understanding. His work increasingly bridges computer vision with language, robotics, and cognitive science, emphasizing interpretability and real-world applicability. There is a clear focus on leveraging unlabeled data and developing methods for robust, generalizable AI systems. His scientific awards and recognitions include: Berkeley Fellowship Google Fellowship Soros Fellowship NSF Fellowship SIGGRAPH Outstanding Doctoral Dissertation Award Facebook Fellowship Adobe Fellowship CMU School of Computer Science Distinguished Dissertation Award ACM Doctoral Dissertation Honorable Mention Alexei Efros has advised numerous PhD students and postdocs, many of whom have gone on to faculty positions at top institutions including CMU, Stanford, MIT, Columbia, NYU, and Georgia Tech. His lab has received research funding from major tech companies and federal agencies, though specific grants are not detailed in the text. He teaches core computer vision and machine learning courses at both undergraduate and graduate levels at UC Berkeley. His research group is highly active, with ongoing projects in 3D perception, generative modeling, and vision-language systems. He leads a vibrant research lab at UC Berkeley, part of the BAIR consortium, collaborating with leading researchers such as Jitendra Malik, Trevor Darrell, Pieter Abbeel, and Angjoo Kanazawa. His lab fosters strong interdisciplinary connections with institutions worldwide, including Oxford, INRIA, and École Normale Supérieure.
Derek W Hoiem is a Professor in the Siebel School for Computing and Data Science at the University of Illinois Urbana-Champaign, where he has been a faculty member since 2009. His research focuses on computer vision and related areas, and he is also the co-founder and Chief Science Officer of Reconstruct, an AI-based construction technology company. His educational background includes: PhD in Robotics, Carnegie Mellon University (2007) Beckman Postdoctoral Fellowship (2008) Prof. Hoiem's research spans computer vision, with a focus on object recognition, scene understanding, and graphics. His work also extends to mobile robotics and 3D scene reconstruction. He has made significant contributions in areas such as visual recognition, 3D modeling, and the application of computer vision in construction monitoring. His recent publications (2023-2025) demonstrate a strong focus on advancing multimodal understanding, particularly in region-based representations, 3D vision, and neural radiance fields. There is a clear trend towards integrating language and vision, improving efficiency in neural networks, and applying computer vision to real-world problems such as construction progress monitoring. His scientific awards and honors are extensive and include: IEEE Fellow (2022) University Scholar (2022) Koendrink Prize (2022) Dean's Award for Excellence in Research, Associate Professor (2021) Campus Distinguished Promotion Award (2015) Best Paper Award: IEEE Winter Conference on Applications in Computer Vision (WACV) (2015) CW Gear Junior Faculty Award (2014) IEEE PAMI Young Researcher Award (2014) Dean's Award for Excellence in Research, Assistant Professor (2014) Sloan Research Fellowship (2013) Intel Early Career Faculty Honor Program Award (2012) NSF CAREER Award (2011) ACM Doctoral Dissertation Award, Honorable Mention (2008) Carnegie Mellon University SCS Distinguished Dissertation Award (2008) Best Paper Award: IEEE Computer Vision and Pattern Recognition (CVPR) (2006) Prof. Hoiem has secured significant research funding, including an NSF CAREER award and an Intel Early Career Faculty award. He is also actively involved in technology transfer, having co-founded Reconstruct where he serves as Chief Science Officer. His teaching excellence is reflected in multiple "List of Teachers Ranked as Excellent" awards spanning from 2010 to 2021. Prof. Hoiem leads a research group at UIUC focused on computer vision and 3D scene understanding. Additionally, he co-founded and serves as Chief Science Officer at Reconstruct, which develops AI-based solutions for construction monitoring.
Lior Wolf is a Professor at the School of Computer Science, Tel Aviv University. Previously, he was a postdoctoral researcher at MIT's Center for Biological and Computational Learning (CBCL) under Prof. Tomaso Poggio and earned his PhD from Hebrew University of Jerusalem with Prof. Amnon Shashua. His educational background includes: PhD in Computer Science, Hebrew University of Jerusalem Postdoctoral Research, MIT CBCL Prof. Wolf's research centers on artificial intelligence with seminal contributions to deep learning, computer vision, and natural language processing. His work bridges theoretical foundations (e.g., attention mechanisms, transformer analysis) with practical applications in medical imaging, speech processing, and sign language technology. He pioneered methods for neural network interpretability, efficient sequence modeling, and multimodal fusion. Analysis of his 2023-2025 publications reveals dominant trends in large language model optimization (neuron pruning, attention analysis), efficient video generation, and cross-modal learning. His work increasingly integrates medical applications (fMRI/EEG analysis) while maintaining theoretical rigor in model architecture design. His scientific achievements include: Best paper award at EMNLP 2024 for 'Backward Lens: Projecting Language Model Gradients into the Vocabulary Space' Best paper award at SCIA 2023 for 'Gradient Adjusting Networks for Domain Inversion' Best paper award at FG 2021 for 'Generating Master Faces for Dictionary Attacks' Prof. Wolf mentors graduate students in the School of Computer Science and leads research at the ICRC building laboratory. His team collaborates with 'the friends of TAU' on projects spanning biometric security, medical imaging, and generative AI. Current work focuses on efficient transformers, neural network interpretability, and multimodal medical diagnostics.
Lourdes Agapito is a Professor of 3D Vision at the Department of Computer Science, University College London (UCL), within the Faculty of Engineering Sciences. She leads research in Non-Rigid Structure from Motion (NR-SFM) and 3D reconstruction from monocular video sequences. Her work addresses dynamic scenes, deformable objects, and articulated structures, with applications in robotics and computer vision. She holds an ERC Starting Grant (2008–2014) and led the EU Horizon 2020-funded Second Hands project (2014–2019), collaborating with institutions like EPFL and KIT to develop robots with 3D visual perception for maintenance tasks. Her research group focuses on dense optical flow estimation, video registration, and deformable tracking. Agapito’s research interests include monocular 3D reconstruction, non-rigid motion analysis, and neural approaches to 3D modeling. She has supervised multiple PhD students and postdocs, including notable researchers such as Ravi Garg and Marco Paladini (Sullivan Prize recipient). Her contributions to conferences include roles as Program Chair for CVPR 2016 and CVPR 2017, and she has authored influential papers on topics like Video-Popup (ECCV 2014) and Modal Space (CVPR 2017). Current projects involve advancing neural parametric models and real-time 3D reconstruction techniques. Awards include the ERC Starting Grant and recognition for her team’s work in non-rigid reconstruction. She actively mentors students and collaborates on grants, with recent openings for postdocs and PhD candidates in 3D vision and robotics.
Andrea Tagliasacchi is an Associate Professor in the School of Computing Science at Simon Fraser University (SFU), where he holds the Visual Computing Research Chair. He is also a part-time (20%) Staff Research Scientist at Google DeepMind in Toronto and holds an associate professor (status only) appointment in the Department of Computer Science at the University of Toronto. Education: PhD in Computing Science – Simon Fraser University (NSERC Alexander Graham Bell Fellow) Postdoctoral Research – École Polytechnique Fédérale de Lausanne (EPFL) MSc in Computer Science – Politecnico di Milano (Gold Medalist) His research lies at the intersection of computer vision, computer graphics, and machine learning, with a focus on 3D visual perception. Key areas include neural radiance fields (NeRF), 3D Gaussian splatting, inverse rendering, and geometric deep learning, with applications in robotics, augmented reality, and autonomous systems. His work emphasizes robust and efficient scene understanding and reconstruction from visual data. His recent publications, appearing in top venues like CVPR, SIGGRAPH, NeurIPS, and ECCV, demonstrate a strong emphasis on neural fields, 3D reconstruction, and generative modeling. Trends include improving rendering efficiency, enhancing robustness to noise and distractors, and enabling controllable and 3D-aware generation. His group has made significant contributions to Gaussian splatting, NeRF optimization, and diffusion-based 3D/4D synthesis. Scientific Awards: 2024 CVPR Best Paper Award (Honorable Mention) 2020 CVPR Best Student Paper Award 2015 SGP Best Paper Award NSERC Alexander Graham Bell Canada Graduate Scholarship MITACS Best Paper Award (SIGGRAPH Asia 2009) NSF Best Poster Award (SGP 2012) He has advised numerous PhD and MSc students, many of whom are now researchers at leading institutions and companies. His research has been supported through collaborations with Google, Intel, and academic partners. He serves the community as a Senior Area Chair for CVPR 2025, Associate Editor for IEEE TPAMI (2024–2026), Guest Editor for IEEE TPAMI on 3D GenAI, and Program Chair for 3DV 2024. He leads a vibrant research lab at SFU focused on pushing the boundaries of 3D scene understanding with machine learning.
Andrew Zisserman is a Royal Society Research Professor at the University of Oxford's Department of Engineering Science, affiliated with the Visual Geometry Group (VGG). His research focuses on computer vision, artificial intelligence, and neural networks, with significant contributions to multimodal learning, video understanding, and 3D scene analysis. He leads projects exploring visual-language models, audio-visual synchronization, and clinical imaging applications. Key research areas include: Video analysis and temporal modeling Multimodal systems for sign language translation and action recognition 3D shape estimation and physical property inference Foundation models and cross-modal retrieval Recent work highlights: Developed Flamingo and Tapir models for video-language tasks Advancements in spinal MRI analysis and clinical imaging Leadership in EGO4D and VoxCeleb challenges Honors include Fellowship of the Royal Society (FRS) and the ISSLS Prize in Clinical Science 2023 for spinal analysis innovations. His lab collaborates globally, emphasizing real-world applications in healthcare and autonomous systems.
Jiatao Gu is an Assistant Professor in the Department of Computer and Information Science (CIS) at the University of Pennsylvania, with a part-time role as Staff Research Scientist at Apple (MLR). He holds a Ph.D. in Electrical and Electronic Engineering from the University of Hong Kong (2018) and a B.Eng. in Electronic Engineering from Tsinghua University (2014). His research focuses on generative machine learning and AI agent interaction with the physical world, emphasizing multi-modal systems spanning language, images, videos, and 3D. Key themes include efficient modeling , flexible architecture design , and scalable decision-making frameworks . 2025: ICLR paper on DART framework 2024: TMLR work on GFlowNet alignment 2023: NeurIPS research on diffusion stability 2022: ACL papers on speech translation Recent publications explore diffusion models for text-to-image synthesis, 3D reconstruction, and efficient sampling techniques. His work addresses fundamental challenges in attention mechanisms, entropy collapse, and multi-stage distillation while advancing non-autoregressive translation and vision-language reasoning . Prospective students can apply through his recruitment process at UPenn. Prior affiliations include Meta AI (FAIR Labs) and academic collaborations with institutions like New York University's CILVR Lab.
Xingang Pan is an Assistant Professor in the College of Computing and Data Science at Nanyang Technological University (NTU), leading the MMLab@NTU. His research focuses on generative AI and visual content creation, particularly in generative models, 3D vision, computer graphics, and computer vision. Prior to NTU, he was a postdoc at the Max Planck Institute for Informatics and earned his Ph.D. from the Chinese University of Hong Kong (2021) and B.Sc. from Tsinghua University (2016). His work emphasizes generative intelligence, exploring long-term world simulation, diffusion models, and multi-scale 3D generation. Notable contributions include WORLDMEM (2025), Alias-free Latent Diffusion (2025), and SAR3D (2025). His research has been published in top venues like CVPR, ICCV, and SIGGRAPH. Xingang Pan oversees the MMLab@NTU, which actively recruits students globally without nationality constraints. The lab’s projects include GAN2Shape (unsupervised 3D reconstruction from 2D GANs) and LN3Diff (scalable 3D generation).