Jan Alber is a Professor of Modern English and American Literature at the Department of English and American Literature and Culture, Justus Liebig University Giessen . He previously held positions at RWTH Aachen, Aarhus University, TU Darmstadt, and the University of Freiburg. His work bridges narrative theory, cognitive studies, and postcolonial literature, with a focus on unnatural narratives and empirical approaches to literary analysis. Fields of Interest: Unnatural Narratology, Cognitive Literary Studies, Climate Fiction, Postcolonial Literature Key Themes: Storyworlds, Ideology, Empirical Methods, Narrative Beginnings Alber's recent publications explore narrative theory's intersections with cognition and ideology, including analyses of digital post-postmodernist novels and climate change narratives. His empirical studies examine how readers process experimental narratives, particularly unnatural narrators and metaleptic structures. He has received prestigious awards such as the Habilitation Prize from the German Association of Anglists (2013) and the FRIAS junior group competition (2008). His research is funded by the AHRC, DFG, and Alexander von Humboldt Foundation, notably supporting interdisciplinary work on post-truth political narratives and Ukrainian scholars. Alber co-founded the Aachen Center for Cognitive and Empirical Literary Studies and serves as President of the International Society for the Study of Narrative. He organizes international workshops on cognitive narratology and teaches courses spanning literary history, theory, and postmodern fiction.
Dr. Diyi Yang is an Assistant Professor in the Computer Science Department at Stanford University. She leads the Social and Language Technologies (SALT) Lab, affiliated with the Stanford NLP Group, Stanford HCI Group, Stanford AI Lab (SAIL), and Stanford Human-Centered Artificial Intelligence (HAI). Her research focuses on socially aware natural language processing, large language models (LLMs), and human-AI interaction, aiming to improve human-human and human-computer communication through socially grounded AI systems. Education: Ph.D. in Language Technologies Institute, Carnegie Mellon University (2013–2019) B.S. in ACM Honored Class, Shanghai Jiao Tong University (2009–2013) Research Interests: Dr. Yang’s work bridges computational social science and NLP. She explores how AI can understand social contexts in language use and develop systems that respect cultural norms, ethical standards, and human values. Her lab’s projects include AI companions for skill training (e.g., Rehearsal Dialects), norm-aware LLMs (NormBank), and frameworks for human-AI collaboration (Co-Gym). Recent efforts address bias in AI, societal impacts of LLMs, and ethical evaluation of human-AI systems. Awards & Honors: 2024: Sloan Research Fellowship, ONR Young Investigator Award 2023: Adamic-Glance Young Distinguished Award (ICWSM), Kavli Fellow (NAS) 2022: NSF CAREER Award, Microsoft Research Faculty Fellow 2020: IEEE AI’s 10 to Watch Advising & Grants: Dr. Yang advises over 10 PhD students and postdocs, co-leading projects on LLM evaluation (SWE-bench/SWE-smith), human-AI ethics, and culturally aware NLP. Her research is supported by NSF, Amazon, DARPA, Google, and Stanford’s HAI initiative. Labs & Teams: The SALT Lab collaborates across disciplines, with projects spanning computer science, linguistics, and social sciences. Current initiatives include developing AI tools for mental health support (AI Partner & Mentor) and auditing societal impacts of LLMs.
Benjamin Van Roy is a Professor at Stanford University since 1998, affiliated with the Departments of Electrical Engineering and Management Science and Engineering, and the Institute for Computational and Mathematical Engineering. He leads the Efficient Agent Team at Google DeepMind and previously held leadership roles at Morgan Stanley, Unica, and Enuvis. He holds SB, SM, and PhD degrees in Computer Science and Electrical Engineering from MIT, advised by John Tsitsiklis. His research focuses on reinforcement learning, alignment, and information theory, with contributions to machine learning foundations, optimization, and finance. He has authored over 150 publications, including influential works on Thompson Sampling, approximate dynamic programming, and exploration strategies. His honors include INFORMS and IEEE Fellowships and the INFORMS Lanchester Prize. Van Roy advises doctoral students across academia and industry, with graduates at top institutions and companies like Meta, Tesla, and Citadel. He teaches courses on reinforcement learning, stochastic control, and optimization. His open-source projects include Epistemic Neural Networks and the Neural Testbed for evaluating machine learning models. Key contributions include foundational work in reinforcement learning theory, scalable methods for recommendation systems, and applications in finance and resource allocation. His research bridges theoretical insights with practical applications, emphasizing alignment and safety of AI systems.
Amir Zamir is a tenure-track Assistant Professor of Computer Science at the Swiss Federal Institute of Technology Lausanne (EPFL) in the School of Computer & Communication Sciences. Previously, he worked at UC Berkeley, Stanford, and UCF with prominent researchers including Silvio Savarese, Jitendra Malik, Mubarak Shah, Rahul Sukthankar, and Leonidas Guibas. He currently leads the Visual Intelligence & Learning Lab at EPFL and serves as chief scientist of Duranta, having previously been the CVML chief scientist of Aurora Solar (a Forbes AI 50 company valued at $4B in 2022) from 2015 to 2022. Dr. Zamir's research spans computer vision, machine learning, and artificial intelligence, with a focus on developing general multi-modal/multi-task vision systems that operate as active agents in the real world. His work emphasizes slow science principles, seeking fundamental understanding over quick publications. Key research areas include embodied vision, multimodal foundation models, computational imaging, and vision-language integration. His notable projects include 4M, Taskonomy, Gibson Environment, Omnidata, and MultiMAE, which have significantly influenced the field of computer vision and embodied AI. Zamir has made substantial contributions to the computer vision community through numerous high-impact publications and leadership roles. His work demonstrates a consistent focus on creating vision systems that go beyond narrow and passive methods toward more general, active, and embodied approaches. The trajectory of his research shows increasing sophistication in handling multiple modalities and tasks within unified frameworks, culminating in recent work on multimodal foundation models that can handle diverse vision tasks. Dr. Zamir has received numerous prestigious awards including the Young Researcher Award 2022 from ECCV, the PAMI Mark Everingham Prize 2022, SIGGRAPH 2022 Best Paper Award for CLIPasso, CVPR 2018 Best Paper Award for Taskonomy, and CVPR 2016 Best Student Paper Award. He is also an ELLIS Faculty Scholar and received the NVIDIA Pioneering Research Award in 2018 for the Gibson Environment. As an advisor, Dr. Zamir has mentored numerous PhD students including Roman Bachmann, Andrei Atanov, Rishubh Singh, Jason Toskov, Kunal Pratap Singh, Zhitong Gao, Mingqiao Ye, and Muhammad Uzair Khattak. His former PhD students include Oguzhan Kar (now at Apple), Alexander Sasha Sax (co-advised with Jitendra Malik, now at Meta FAIR), and Teresa Yeo (now at MIT-Singapore Alliance). Dr. Zamir teaches several advanced courses including CS-503 Visual Intelligence, CS-500 AI Product Management, COM-304 Intelligent Systems, and ENG-615 Topics in Autonomous Robotics. Dr. Zamir leads the Visual Intelligence & Learning Lab at EPFL, which focuses on developing fundamental methods for visual intelligence that can operate effectively in real-world environments. The lab takes an interdisciplinary approach combining computer vision, machine learning, robotics, and cognitive science to create systems that can perceive, understand, and interact with the world. Current research directions include multimodal foundation models, computational imaging, embodied vision, and personalization of generative models.
Jennifer Olsen, PhD, is an Assistant Professor of Computer Science at the University of San Diego since 2020. She holds a PhD, MS, and BS in Human-Computer Interaction and Cognitive Science from Carnegie Mellon University, followed by postdoctoral research at the Swiss Federal Institute of Technology (EPFL), Lausanne, Switzerland. Her research focuses on the intersection of human-computer interaction, cognition, and education, emphasizing collaborative learning and educational technology design from both learner and instructor perspectives. Education: PhD in Human-Computer Interaction, Carnegie Mellon University MS in Human-Computer Interaction, Carnegie Mellon University BS in Cognitive Science, Carnegie Mellon University Research Interests: Dr. Olsen explores how collaboration supports learning, designs technologies to enhance educational practices, and investigates gaze-based metrics for understanding collaborative problem-solving. Her work spans gamified robotics, AI-driven orchestration systems, and virtual reality applications in vocational training. She emphasizes learner-centered design and the integration of social robots and virtual agents in pedagogical settings. Grants/Advising: While no specific grants or advisees are listed, her prolific publication record indicates active involvement in educational technology research and development. Her work addresses challenges in classroom orchestration, multimodal data analysis, and accessibility in educational robotics. Labs/Teams: Collaborates with interdisciplinary teams focused on educational technology, human-robot interaction, and adaptive learning systems. Her research leverages tools like FROG orchestration graphs and eye-tracking technologies to develop practical classroom solutions.
Adrian Weller is a prominent researcher and academic at the University of Cambridge, serving as a Director of Research in Machine Learning within the Department of Engineering. He holds multiple significant leadership roles including Programme Director for Trust and Society at the Leverhulme Centre for the Future of Intelligence (CFI), and previously served as Programme Director for AI at The Alan Turing Institute, the UK national institute for data science and AI. His work bridges theoretical machine learning research with practical applications and societal implications of artificial intelligence. Weller's research interests span a broad spectrum of AI and machine learning topics with a particular focus on ensuring beneficial societal outcomes. His work encompasses explainability, fairness, robustness, scalability, privacy, safety, and ethics in AI systems. He has made significant contributions to trustworthy machine learning, including developing frameworks for AI governance, certification, and human-AI collaboration. His research group actively investigates neuro-symbolic approaches, privacy-preserving techniques, and methods for improving the reliability and interpretability of AI systems. His recent publications demonstrate a strong trend toward addressing the practical challenges of deploying AI systems in real-world contexts, particularly focusing on certification frameworks, governance mechanisms, and human-centered approaches. His work spans theoretical advances in machine learning architectures while maintaining a strong connection to societal impact, with publications appearing in top venues across AI, machine learning, and interdisciplinary applications. Scientific Awards: MBE for services to digital innovation (2022 Queen's Birthday Honours) Turing AI Fellowship for Trustworthy Machine Learning Weller actively supervises a large group of PhD students and postdocs, with current students including Juyeon Heo, Yanzhi Chen, Katie Collins, Isaac Reid, Yichao Liang, Herbie Bradley, and Shoaib Siddiqui. His former students have gone on to positions at leading institutions including Google DeepMind, ETH Zurich, NYU, and MPI-IS Tübingen. He has served on numerous advisory boards including the Centre for Data Ethics and Innovation, UNESCO's expert group on AI ethics, and the World Economic Forum's Global Future Council on AI. His research has been supported through his Turing AI Fellowship and various collaborative projects focused on safe and ethical AI development. Weller leads a vibrant research group focused on trustworthy machine learning, which actively organizes workshops and conferences including ICML 2024 (where he served as Program Chair), multiple workshops on responsible AI, and events through the ELLIS network. His group collaborates extensively across disciplines, working with researchers in computer science, social sciences, law, and policy to address the multifaceted challenges of developing beneficial AI systems.
Kenji Kawaguchi is the Presidential Young Professor in the Department of Computer Science at the National University of Singapore (NUS), where he leads the Deep Learning Lab and is a faculty affiliate at the NUS Institute of Data Science. His research bridges theoretical and applied machine learning, focusing on deep learning, large language models, and physics-informed neural networks. His educational background includes a Ph.D. and S.M. in Computer Science and Electrical Engineering from the Massachusetts Institute of Technology (MIT), advised by Leslie Pack Kaelbling, and a postdoctoral fellowship at Harvard University’s Center of Mathematical Sciences and Applications. Dr. Kawaguchi’s research interests center on the theoretical foundations of deep learning, optimization, generalization, and applications in areas such as molecular modeling, AI safety, and efficient training of large models. He has made significant contributions to understanding in-context learning, diffusion models, and neural operators for partial differential equations. His recent publications (2023–2025) reflect a strong trend toward improving the efficiency, robustness, and interpretability of large-scale models, particularly in language and scientific domains. Key themes include LLM alignment and safety, diffusion model optimization, and physics-informed learning for high-dimensional problems. Presidential Young Professor He has served as Area Chair and PC Member for top-tier conferences including NeurIPS, ICML, ICLR, AAAI, and UAI, and as reviewer for journals such as JMLR and Annals of Statistics. He has delivered invited talks at Harvard, MIT, Stanford, CMU, Brown, and Google Research, reflecting his international recognition. He actively mentors students and welcomes PhD candidates and postdocs to join his research group.
Mark Steedman is a Professor in the School of Informatics at the University of Edinburgh, where he conducts research in Artificial Intelligence, Computational Cognitive and Social Science, and Natural Language and Speech Processing. He is affiliated with the Institute for Language, Cognition and Computation (ILCC), the Centre for Speech Technology Research (CSTR), and the Human Communications Research Center (HCRC). He also holds an adjunct professorship in Computer and Information Science at the University of Pennsylvania. His research focuses on Combinatory Categorial Grammar (CCG) , computational linguistics , prosody and intonation , temporal semantics , gesture in communication , and computational music analysis . He has authored foundational books including Surface Structure and Interpretation , The Syntactic Process , and Taking Scope . The recent publications reflect a strong trend toward integrating formal grammatical frameworks like CCG with modern neural and distributional models, particularly in semantic parsing, entailment reasoning, and cognitive modeling. His work bridges symbolic and statistical approaches in NLP, often focusing on robust, wide-coverage parsing and semantic interpretation. Best Paper Award at AACL/IJCNLP 2023 for 'Smoothing Entailment Graphs with Language Models' Best Paper Award at ACL 2023 for 'Extrinsic Evaluation of Machine Translation Metrics' Influential Paper Award 2017 from IFAAMAS for 'Animated Conversation' Mark Steedman has supervised numerous PhD students and collaborated widely across institutions. He leads research in formal grammar applications to cognitive modeling, dialogue, and multimodal communication. His lab contributes to CCG software and semantic parsing tools, and he continues to be actively involved in advancing the integration of symbolic and neural AI.
Furong Huang is an Associate Professor at the University of Maryland's Department of Computer Science, with affiliations at the Institute for Advanced Computer Studies, Center for Machine Learning, Maryland Robotics Center, and Applied Mathematics, Statistics, and Scientific Computation Program. Her research bridges trustworthy machine learning, sequential decision-making, and foundation models for robotics, emphasizing reliability, interpretability, and ethical standards. Research Interests: Trustworthy AI Generative AI Reinforcement Learning AI Security Algorithmic Fairness Foundation Models for Robotics Recent Publications span leading conferences (NeurIPS, ICML, ICLR, CVPR) and journals, focusing on: Robustness in Vision-Language Systems Trustworthy Generative AI Foundation Models for Sequential Decision-Making AI Security and Watermarking Scientific Awards MIT TR35 Innovator Under 35 (Asia Pacific 2022) Best Paper Award, AdvML Frontier Workshop, NeurIPS 2024 NSF NAIRR Pilot Awardee Microsoft Accelerate Foundation Models Research Award (2023) JP Morgan Faculty Research Awards (2019–2022) Advising and Grants : Her lab has graduated students to roles at OpenAI, Google, Meta, and Netflix. Research funded by DARPA, NSF, ONR, AFOSR, and industry partners like Microsoft, Adobe, and Capital One. Labs & Teams : Leads research groups focused on Trustworthy AI and Robotics at the University of Maryland, collaborating with the Maryland Robotics Center and Applied Mathematics Program.
Dr. Carolyn Emery is a Professor in the Faculty of Kinesiology at the University of Calgary, with joint appointments in Pediatrics and Community Health Sciences at the Cumming School of Medicine. She holds leadership roles as Chair of the Sport Injury Prevention Research Centre (an International Olympic Committee Research Centre) and the Alberta Children’s Hospital Research Institute. Her affiliations include the McCaig Institute for Bone and Joint Health, Hotchkiss Brain Institute, and O’Brien Institute for Public Health. Dr. Emery’s education includes a BScPT from Queen’s University (1988), MSc in Epidemiology from the University of Calgary (1999), and PhD in Epidemiology from the University of Alberta (2004). Her research focuses on injury prevention in youth sports, concussion, and pediatric rehabilitation, aiming to reduce long-term health consequences such as osteoarthritis and post-concussion syndrome. She leads major grants including CIHR’s “SHRed Injuries” and NFL-funded concussion surveillance programs. Her research interests emphasize practical interventions for injury prevention, including neuromuscular training and policy changes. Recent work explores concussion outcomes in adolescents, biomechanical injury mechanisms, and global sport safety initiatives. She has published extensively on topics like sport-related musculoskeletal injuries, concussion management, and the epidemiology of youth sports. Dr. Emery has received over 20 awards, including the Royal Society of Canada Fellowship, Killam Professorship, and CIHR Canada Research Chair. Her work bridges clinical practice, policy, and community engagement to enhance public health outcomes in sports and youth recreation. Her lab, the Sport Injury Prevention Research Centre, collaborates internationally to develop evidence-based strategies for injury prevention. Current projects include evaluating the impact of tackle laws in rugby, head acceleration in wrestling, and long-term consequences of pediatric concussions.
David Williamson Shaffer is the Sears Bascom Professor of Learning Analytics and Vilas Distinguished Achievement Professor of Learning Sciences at the University of Wisconsin-Madison's Department of Educational Psychology. He is also a Data Philosopher at the Wisconsin Center for Education Research. His work focuses on merging statistical and qualitative methods to model complex human collaboration. Shaffer holds an AB in History and East Asian Studies from Harvard University (1987), and MS/PhD in Media Arts and Sciences from MIT (1996/1998). Before academia, he was a teacher, curriculum developer, and game designer. His research emphasizes Quantitative Ethnography —a methodology combining big data analysis with cultural interpretation. Key contributions include Epistemic Network Analysis (ENA) and over 250 publications, including influential books like How Computer Games Help Children Learn (2006) and Quantitative Ethnography (2017). Awardees of the EU Marie Curie Fellowship (2008) and Spencer Foundation Fellowship (2003), Shaffer's work has been presented globally at conferences like Learning Analytics & Knowledge and Computer Supported Collaborative Learning. His interdisciplinary approach bridges education, data science, and humanities, addressing challenges in authentic STEM experiences, collaborative learning analytics, and culturally responsive pedagogy.
Arman Cohan is an Assistant Professor in the Department of Computer Science at Yale University, where he leads the Yale NLP Lab since its founding in January 2023. His research spans natural language processing and machine learning with emphasis on language modeling, representation learning, retrieval systems, and specialized domain applications including scientific discovery and AI for science. His primary research interests include: Natural Language Processing Machine Learning Large Language Models Information Retrieval AI for Science Scientific Problem-Solving Recent publications (2025) demonstrate intense focus on evaluating and advancing LLM capabilities across multimodal reasoning, scientific claim verification, financial domain applications, and biological modeling. The lab consistently produces high-impact work accepted at top-tier conferences including ACL, EMNLP, and ICLR, with 11 papers at ACL 2025 alone. Scientific awards include: Best Paper Award at AI4Research Workshop (IJCAI 2024) Outstanding Paper Award at EACL 2023 Best Paper Award at ACL 2024 for Olmo language model research Professor Cohan actively advises PhD students including Kaili Liu, Jacob Dunefsky, Alan Li, Yilun Zhao, and has graduated researchers such as Linyong Nan (now at Zoom) and Ansong Ni (now at Meta). The lab maintains strong industry partnerships and receives substantial research funding as evidenced by its prolific output and conference presence. The Yale NLP Lab hosts the annual New England NLP Workshop and regularly features speakers from leading institutions including Meta AI, Allen Institute for AI, and DeepMind, fostering a collaborative environment for advancing NLP research.
Elaine Francis is a Professor of English and Linguistics at Purdue University's College of Liberal Arts, where she serves as Associate Head of the Department of English and directs the Experimental Linguistics Lab. She holds additional affiliations as an Affiliate Faculty Member in the Department of Linguistics and a Courtesy Faculty Member in the Department of Speech, Language, and Hearing Sciences. Ph.D. in Linguistics, University of Chicago (1999) B.A. in Linguistics, College of William and Mary (1993) MA in Linguistics, University of Chicago (1995) Former Assistant Professor at the University of Hong Kong (1999-2002) Professor Francis's research focuses on syntax and its interfaces with semantics, discourse information structure, and language processing in production and comprehension. She employs experimental methods to investigate syntactic, semantic, discourse-pragmatic, and cognitive factors underlying complex sentence structures. Her primary research areas include word order alternations, filler-gap dependencies, resumptive pronouns, and grammatical acceptability. She examines how various factors contribute to grammatical alternations in language use and explores how processing pressures in production and comprehension contribute to grammatical conventions. Her recent publication record shows consistent output across multiple linguistic subfields, with a strong emphasis on experimental syntax and psycholinguistics. Her 2022 book, Gradient Acceptability and Linguistic Theory , represents a significant contribution to understanding acceptability judgment tasks in relation to syntactic theory. Her work spans theoretical syntax, experimental linguistics, second language acquisition, and cross-linguistic comparison, often bridging the gap between theoretical frameworks and empirical evidence. Regularly teaches short courses at Linguistic Society of America Linguistic Institutes Member of Linguistic Society of America Ethics Committee Editorial board member of Glossa Psycholinguistics Co-editor of Polymorphous Linguistics: Jim McCawley's Legacy (2005) Co-editor of Mismatch: Form-Function Incongruity and the Architecture of Grammar (2003) Professor Francis actively supervises graduate students in linguistics, though she has indicated she will be unable to take on new graduate students for the 2025-2026 academic year. She has received research support for her experimental work, particularly for investigating syntactic phenomena through controlled experiments and corpus analysis. Her collaborative work extends across multiple institutions and disciplines, including collaborations with researchers in computational linguistics, cognitive science, and clinical linguistics. She directs the Experimental Linguistics Lab at Purdue University, which conducts research on syntactic processing using various experimental methodologies. The lab investigates how speakers produce and comprehend complex sentence structures, with particular attention to how grammatical knowledge interacts with processing constraints. Current projects include research on relative clauses, syntactic islands, and acceptability gradient phenomena across multiple languages.
Chris Donahue is an Assistant Professor in the Computer Science Department at Carnegie Mellon University . He also serves as a part-time Research Scientist at Google DeepMind on the Magenta team. His work focuses on leveraging generative AI to enhance human creativity, particularly in music. Education: PhD in Computer Science (UC San Diego), Postdoctoral Scholar (Stanford University) His research spans controllable generative modeling of music and audio , with a focus on real-time interactive systems. Projects like Piano Genie , Beat Sage , and Copilot Arena demonstrate his commitment to real-world deployment. His Generative Creativity Lab (G-CLef) explores AI applications beyond music, including programming and natural language. Recent publications highlight advancements in multimodal music evaluation , real-time adaptation , and AI-driven sound morphing . He co-developed Magenta RealTime , an open-weight real-time music generation model, and MusicFX DJ Mode . Scientific Awards: Best Paper Award (top 1) at NAACL Student Research Workshop 2025 Best Paper Award (top 1% of submissions) at CHI 2025 Best Paper Runner-up at ISMIR 2021 He co-advises PhD students like Wayne Chi (NDSEG Fellow) and mentors Irmak Bukey . His lab receives support from the AIxArts incubator fund at CMU .
Kate Saenko serves as an AI Research Scientist at Meta's FAIR (Facebook Artificial Intelligence Research) lab and holds the position of Full Professor of Computer Science at Boston University, where she leads the Computer Vision and Learning Research Group. Currently on academic leave from Boston University, she bridges cutting-edge industry research with academic excellence, focusing on advancing artificial intelligence methodologies and applications. Her educational background includes a PhD in Electrical Engineering and Computer Science (EECS) from the Massachusetts Institute of Technology (MIT), followed by postdoctoral training at the University of California, Berkeley and Harvard University. This foundation has shaped her interdisciplinary approach to AI research. Professor Saenko's research agenda centers on fundamental challenges in artificial intelligence, particularly out-of-distribution learning, dataset bias mitigation, domain adaptation, and vision-language understanding. Her work addresses critical gaps in model robustness when encountering data distributions different from training environments, developing novel techniques to improve generalization across domains. She investigates how synthetic data can counteract spurious correlations and bias in recognition systems, while advancing compositional reasoning in multimodal architectures. Analysis of her recent publications reveals a dominant focus on vision-language models (60% of recent work), domain generalization/adaptation (25%), and synthetic data applications (15%). Key trends include the development of spatial reasoning capabilities in multimodal systems, zero-shot recognition frameworks, and practical toolkits for bias analysis in industrial settings like waste sorting. Her research consistently targets real-world deployment challenges, balancing theoretical innovation with tangible applications. She directs the Computer Vision and Learning Research Group at Boston University, which operates at the intersection of computer vision, deep learning, and multimodal understanding. The group maintains strong industry collaborations through Meta's FAIR and previously engaged with the MIT-IBM Watson AI Lab. Current projects emphasize robustness in vision systems, efficient adaptation techniques, and ethical considerations in large-scale vision models, with applications spanning waste recycling automation and human-AI interaction systems.