https://arxiv.org/api/3RAVMEdQhZppVR6xzTlBr+EG18s 2026-09-11T19:05:14Z 32695 15 15 http://arxiv.org/abs/2609.11224v1 AI Soccer Analyst: Stage-Aware and Verifiable Human-AI Collaboration for Soccer Data Analysis 2026-09-10T08:25:09Z Sports data analysts translate domain questions into insights by combining computation with sport-specific domain expertise. Large language models ease programming, but prompt-to-report workflows may obscure decisions and evidence. We present AI Soccer Analyst, a mixed-initiative system with revisable stages: Data Understanding, Problem Definition, Structured Planning, Execution, Evidence-Grounded Reporting, and Interaction and Refinement. A formative study with five analysts first informed design goals for automation, verifiability, human control, and accessibility. Subsequently, a task-based evaluation with 16 participants combined system logs, retained artifacts, ratings, and open responses; 33 of 48 tasks met the operational completion criteria. Exploratory tests supported favorable participant perceptions of completed-task output quality, task achievement, reliability, and verifiability after Holm correction. Interaction records showed domain knowledge emerging through clarification, planning, and refinement. These findings position stage-aware human-AI collaboration as a practical approach for producing inspectable, revisable, and verifiable analyses while retaining domain-expert involvement in consequential decisions. 2026-09-10T08:25:09Z Calvin Yeung Keisuke Fujii http://arxiv.org/abs/2609.11198v1 (Whose defaults?) Is artificial intelligence reorienting archaeological methods? 2026-09-10T08:08:15Z Generative AI and the practice of "vibe coding" are changing how archaeologists carry out computational research, but their effects on the discipline's range of methods is still understudied. In this paper, we evaluate whether large language models (LLMs) are narrowing the variety of methods archaeologists use. We first analysed approximately 119,000 archaeology abstracts from Scopus, covering publications from 2010 to 2025. Using a locally run LLM, we identified the computational methods reported in each abstract and organised them into 25 broad categories (L2) and 241 finer clusters (L3). A Bayesian Dirichlet-multinomial model of method composition within sub-disciplines found a small but credible shift in method use after 2023. However, this shift was smaller than the variation already present across the full study period. No individual technique showed a significant change, and overall methodological diversity increased rather than declined. We then ran a controlled experiment to see whether LLMs recommend a narrower set of methods than archaeologists have used in practice. Two different open-weight models were asked to suggest methods for 28 archaeological research problems, with prompts providing three levels of methodological guidance: novice, intermediate, and expert. Recommendation diversity was much lower than in the published literature, particularly without methodological guidance. The models also tended to favour methods that were widely used before 2023, and their recommendations more closely resembled the post-2023 literature. Taken together, these results are consistent with LLMs pushing methodological choice towards convergence, although our study cannot establish a causal effect. They raise a broader question: how can archaeology retain methodological diversity as LLMs become more involved in research? 2026-09-10T08:08:15Z Lorenzo Cardarelli Roberto Ragno http://arxiv.org/abs/2609.11193v1 Conceptualising an Initial Design Space for Guidance in Digital Physical Activity Support 2026-09-10T08:02:40Z Providing guidance is frequently referenced as a key capability of digital health interventions targeting physical activity, yet the term remains poorly defined and inconsistently applied. Existing work often conflates guidance with related constructs such as personalisation, feedback, or persuasion, limiting both theoretical clarity and design progress. This paper conceptualises an initial design space of guidance in the context of digital physical activity support. We define guidance for physical activity as situated, action-oriented support that scaffolds users' embodied engagement in physical activity. Drawing on literature from behaviour change, human-computer interaction, embodied cognition, and digital health, we outline a design space that characterises guidance along multiple dimensions: scope, purpose, timing, context, modality, embodiment, adaptivity, autonomy, and affective quality. By offering a structured vocabulary and conceptual foundation, this work aims to support more coherent research, comparisons, and responsible design of digital health interventions featuring guidance for physical activity support. 2026-09-10T08:02:40Z Accepted for publication at NordiCHI 2026 Faith Young Markus Tatzgern Alexander Meschtscherjakov Jan Smeddinck http://arxiv.org/abs/2604.05519v2 Active noise cancellation on open-ear smart glasses 2026-09-10T07:13:51Z Active noise cancellation (ANC) is widely deployed on consumer headphones and earbuds to suppress environmental noise. However, existing ANC systems require an error microphone at the user's ear canal to measure residual sound, preventing deployment on emerging open-ear wearable devices such as smart glasses and VR headsets, which leave the ear unoccluded. Here we present an ANC system for open-ear wearables that suppresses environmental noise using only microphones and miniaturized open-ear speakers embedded within the frame of the wearables, removing the need for an in-ear error microphone. Our low-latency computational pipeline uses a neural network to estimate the noise at the ear from an array of eight microphones distributed around the wearable's frame and generates an anti-noise signal in real-time. This mapping generalizes to unseen users and acoustic environments without prior acoustic measurement. We develop a custom glasses prototype and evaluate across eleven unseen users and eight unseen environments under mobility in the 100 to 1000 Hz frequency range, where environmental noise is concentrated. We achieve a mean noise reduction of 9.6 dB without any calibration, and 11.2 dB with a brief user-specific calibration. Further, we demonstrate that our approach extends to the broader class of open-ear wearables including VR headsets and headbands. 2026-04-07T07:17:40Z Kuang Yuan Freddy Yifei Liu Tong Xiao Yiwen Song Chengyi Shen Saksham Bhutani Justin Chan Swarun Kumar http://arxiv.org/abs/2609.11109v1 How AI Coders Discuss, Disagree, and Reach Consensus: Challenges and Opportunities for LLM-Based Qualitative Coding 2026-09-10T05:34:47Z The utility of AI in multi-coder qualitative coding has been widely discussed, yet little empirical evidence exists to delineate the contexts in which it performs reliably. We address this gap by quantifying the effectiveness of multi-agent LLM coding across varied qualitative datasets, revealing key contextual and structural factors that mediate coding outcomes. We developed a literature-informed baseline pipeline that enables AI agents to independently code, debate, and reconcile disagreements. Results revealed that coding accuracy depends on factors such as codebook length, qualitative data similarity, and agent disagreement. Notably, intense and unresolved debates between agents led to higher accuracy. Our analysis showed that while LLMs emulate many human discussion behaviors, they lack adaptive responsiveness to context. From these findings, we offer design recommendations for building automated coding systems. Our open-source AI discussion dataset and methodological framework lay the groundwork for advancing the design of AI-mediated automated thematic analysis. 2026-09-10T05:34:47Z 34 pages, 7 tables Jeongyeon Kim John Mitchell http://arxiv.org/abs/2508.16076v3 Prompting with Sign Parameters for Low-resource Sign Language Instruction Generation 2026-09-10T05:32:11Z Sign Language (SL) enables two-way communication for the deaf and hard-of-hearing community, yet many sign languages remain under-resourced in the AI space. Sign Language Instruction Generation (SLIG) produces step-by-step textual instructions that enable non-SL users to imitate and learn SL gestures, promoting two-way interaction. We introduce BdSLIG, the first Bengali SLIG dataset, used to evaluate Vision Language Models (VLMs) (i) on under-resourced SLIG tasks, and (ii) on long-tail visual concepts, as Bengali SL is unlikely to appear in the VLM pre-training data. To enhance zero-shot performance, we introduce Sign Parameter-Infused (SPI) prompting, which integrates standard SL parameters, like hand shape, motion, and orientation, directly into the textual prompts. Subsuming standard sign parameters into the prompt makes the instructions more structured and reproducible than free-form natural text from vanilla prompting. We envision that our work would promote inclusivity and advancement in SL learning systems for the under-resourced communities. 2025-08-22T04:11:28Z Accepted at the ICCV 2025 Workshop on Vision Foundation Models and Generative AI for Accessibility (CV4A11y). OpenReview: https://openreview.net/pdf?id=KkVMBkjbra Md Tariquzzaman Md Farhan Ishmam Saiyma Sittul Muna Md Kamrul Hasan Hasan Mahmud http://arxiv.org/abs/2607.00445v2 Gaze-Informed Proactive AI Assistance for Children's Picture Exploration 2026-09-10T05:32:06Z Proactive assistance with large language models (LLMs) has received growing attention in the human computer interaction (HCI) community. However, most past work on proactive LLMs' assistance has focused on adult users and task-oriented settings, leaving open how such systems could support children, whose interests and needs are often expressed through gaze and other nonverbal behaviors rather than explicit requests. In this study, we focus on two key challenges of proactive assistance in children's picture exploration: when to provide assistance and what assistance to provide based on children's nonverbal behaviors. To address these challenges, we introduce Ollie, a gaze-informed proactive artificial intelligence (AI) assistant that offers short narrative descriptions based on where a child is looking. Ollie uses children's gaze to estimate their attention, identify their current visual focus, and select a related picture region for the LLM to verbally describe. In a within-subject experiment, we compared gaze-informed assistance with random assistance. Results show that gaze-informed assistance kept children's attention on their current focus for a longer period of time, and guided them more effectively to related picture regions. Children, parents, and a participating kindergarten teacher viewed Ollie positively and consider that it better matched children's interests when compared with the random assistance. This work shows the feasibility of using gaze as an implicit input for proactive AI assistance for children and provides design implications for future child-centered AI systems. 2026-07-01T04:58:08Z Zekun Wu Man Su Huiyong Li Tomohiro Nagashima Anna Maria Feit http://arxiv.org/abs/2608.16334v2 Transfer Learning of Keystroke Dynamics for Cross-Device User Authentication 2026-09-10T05:26:16Z Keystroke dynamics (typing patterns) can be used as a behavioural biometric modality for user authentication, with applications such as fraud prevention. While the modality has been shown to work well for single device authentication, its application to cross-device scenarios is more challenging. Dynamics learned on one device (eg., phone) may not be directly applicable to authentication on a secondary device with a different form factor (eg., tablet) due to changes in typing patterns that can lead to distribution drifts. To address this, we propose a cross-device user authentication system based on inductive transfer learning, where keystroke dynamics learned on one device are adapted to a secondary device. The adapted data is then combined with necessarily limited training data for the secondary device, which is used to robustly train a binary classifier. Furthermore, an extended set of keystroke features is used to better capture discriminative dynamics. Experiments on the BBMAS dataset show that proposed system achieves an equal error rate of 14.2% for the cross-device scenario, surpassing previous methods. 2026-08-17T09:40:43Z Nuwan Kaluarachchi Sevvandi Kandanaarachchi Kristen Moore Arathi Arakala Conrad Sanderson http://arxiv.org/abs/2609.11088v1 Visual-Motion-Induced Modulation of Pedestrian Trajectories Using Spatially Distributed Multi-Display Signage in Public Spaces 2026-09-10T04:55:05Z Multi-display signage (MDS), now ubiquitous in urban environments, has the potential to influence human behavior and experience in public spaces. However, despite its unique capability to present spatially distributed dynamic visual stimuli, its current use is mainly limited to advertising. In this study, we propose a perception-based approach for laterally modulating pedestrian trajectories as a nonverbal means of guiding pedestrians in public spaces. The approach is motivated by vection, the illusion of self-motion, and uses laterally moving monochrome stripes, a standard stimulus in vection research, presented across spatially distributed displays to elicit postural responses that may bias pedestrian trajectories. We evaluated the approach through a controlled laboratory experiment and a real-world field deployment involving actual pedestrian flows in a national museum. The laboratory experiment examined whether the MDS setup induced trajectory shifts in the direction predicted by prior research on the behavioral effects of vection. The field deployment investigated whether comparable effects would emerge in aggregate pedestrian behavior during unconstrained movement under conditions closer to those of urban public spaces. In the laboratory, full-screen motion significantly biased walking trajectories in the direction of visual motion, whereas partial-stripe motion produced no significant directional effect. In the field deployment, opposing full-screen motion conditions produced direction-consistent differences in aggregate pedestrian positions. The field results, observed despite the substantial variability in real-world pedestrian flows, extend the controlled laboratory findings and provide ecologically valid evidence supporting practical MDS-based pedestrian modulation in public settings. The results further suggest that sufficient visual-motion coverage may be important. 2026-09-10T04:55:05Z Yuri Mikawa Taiki Fukiage Yuki Kubota Takumi Yokosaka Maki Ogawa Kazushi Maruya http://arxiv.org/abs/2609.11077v1 X-Hinges: 3D Printing Self-Sensing Compliant Mechanisms for Continuous and Multi-DOF Motion Sensing 2026-09-10T04:33:07Z We present X-Hinges, a design and fabrication method for self-sensing compliant mechanisms based on multi-material FDM 3D printing. By co-printing two conductive filaments of different conductivities within a compliant body, we embed resistive sensing elements directly during fabrication without post-assembly, enabling continuous motion sensing across multiple degrees of freedom in a single print. The structure supports three degrees of freedom, each equipped with a dedicated sensing element configuration for multi-DOF motion estimation. We develop a precision data acquisition system and data-driven regression models that enable continuous, real-time motion sensing. We also introduce an interactive design tool for customizing the geometry, mechanical properties, degrees of freedom, and sensing configurations of X-Hinges. The tool also supports augmenting existing 3D models with self-sensing structures, endowing ordinary objects with continuous multi-DOF sensing capabilities. Finally, we present a set of application examples demonstrating the capability of X-Hinges for fabricating personalized interactive interfaces. 2026-09-10T04:33:07Z Accepted at ACM UIST 2026. 15 pages, 24 figures. Author's accepted manuscript Xiang Chang Haiyang Yan Stefanie Mueller Jiaji Li 10.1145/3830398.3830598 http://arxiv.org/abs/2609.09379v2 Agentic Web Accessibility Auditing: A Criterion-Specific Framework for Translating WCAG Requirements into Assessments 2026-09-10T03:47:39Z Web accessibility auditing requires interpreting diverse requirements and examining interface behavior. Rule-based checks and noninteractive model assessments can miss barriers requiring contextual or interactive evidence. We present an agentic framework that assigns a vision-language agent to each accessibility requirement. Guided by tailored instructions, agents inspect webpages, operate controls, and record evidence supporting their findings. We implement the framework for 40 requirements from the Web Content Accessibility Guidelines (WCAG). To compare detection and cost, we construct a dataset of 250 page-criterion records derived from expert audits across 11 scholarly platforms. Agents recover 67 of 78 reported positive cases (86% recall), compared with 36% for axe-core, a rule-based checker, and 67% for an uncued, noninteractive vision-language model, at lower precision (56%). They recover nine of ten Keyboard and No Keyboard Trap cases missed by both baselines. Together, the framework, implementations, and dataset support automated accessibility auditing grounded in inspectable evidence. 2026-09-08T19:27:12Z Preprint. 26 pages, 7 figures. Revised title, abstract, conclusion, and exposition; clarified methods, dataset construction, and interpretation of findings. Supplementary material and supporting research records included as ancillary files Arjun Mishra Pranav Karthik Byungjun Bae Dongwook Yoon http://arxiv.org/abs/2609.11000v1 ShellVis: Sandboxed Live Programming for Shell Scripts 2026-09-10T02:20:53Z Live programming provides visibility to programmers by running and tracing programs as they are edited. However, for programs with potentially harmful side effects, liveness can turn mistakes into disasters. We propose enabling live programming in environments with side effects via sandboxing: confining effects to a simulation of the true environment. We apply sandboxed live programming in the challenging context of shell scripting: a ubiquitous and powerful---yet notoriously opaque and error-prone---tool. ShellVis provides line-by-line feedback on a shell script's run-time behavior, with file operations sandboxed via a safe overlay of the file system. A qualitative user evaluation finds ShellVis to be helpful to participants, replacing tedious existing practices and instilling confidence. Participant responses also reveal areas for future research, particularly bridging the gulf of execution alongside the gulf of evaluation. ShellVis serves as a case study of how sandboxing can bring live-programming techniques into the many real-world programming contexts where side effects are important. 2026-09-10T02:20:53Z Joshua Horowitz Jeffrey Heer http://arxiv.org/abs/2609.10942v1 "Coder first, advocate second, college student third": The Liminality of Going to College as a Blind Computing Student 2026-09-10T01:03:17Z Blind or low vision (BLV) students are less likely to graduate from college, particularly in computing. Prior work documents accessibility challenges in high school and college, but we lack understanding of the transition process that produces this "leaky pipeline." To address this, we interviewed ten BLV college students about going to college to study computing. We analyzed our data through the lens of life transition, specifically Intersecting Liminality. Our findings reveal that some BLV students face such immense digital accessibility and college acclimation barriers that the only way forward as coders is to take on a "second job" as a blind advocate or drop out of the computing major. We argue that the college transition is a critical point for analysis and technological intervention, and further, that Intersecting Liminality provides a useful lens for HCI scholars to unpack the compounding challenges that prevent some BLV students from completing computing degrees. 2026-09-10T01:03:17Z To be published in The 28th International ACM SIGACCESS Conference on Computers and Accessibility (ASSETS 2026). 16 pages, 1 figure, 1 table Isabela Figueira Josahandi M. Cisneros Stacy M. Branham 10.1145/3797867.3829050 http://arxiv.org/abs/2609.10939v1 Evaluating Scaffolding-Oriented Multi-Agent Large Language Model System for Clinical Interview Training 2026-09-10T00:56:03Z Clinical education must prepare medical students to conduct safe and coherent patient interviews under conditions of uncertainty. Traditional standardized patient (SP) training is resource-intensive and difficult to scale. We developed a scaffolding-oriented multi-agent Large Language Model (LLM) AI Standardized Patient (AI-SP) training platform1. The system includes a patient agent for simulated dialog, a tutor agent providing Socratic prompts without disclosing diagnostic information, and a turn-level evaluator agent that monitors clinical progress without revealing summative scores. In a randomized controlled study (N = 100 medical students), participants were assigned to either a multi-agent (MA) scaffolding condition or a control condition. All students completed two learning sessions under their assigned condition followed by an examination conducted in a patient only environment. Performance was assessed using a standardized Objective Structured Clinical Examination (OSCE) based rubric. While no significant difference was observed in final diagnostic accuracy between groups, the multi-agent AI standardized patient system improved final examination scores compared to the control group utilizing structured progressive information disclosure; the most substantial and consistent improvements were observed in communication, the expression of empathy, and specific history-taking behaviors. These findings suggest that specialized LLM agents enhance the process quality of simulated clinical interviews without artificially inflating examination outcomes. To support future research, we release a multi-expert annotated dataset comprising transcripts, checklist annotations, turn-level evaluations, and OSCE-aligned scoring outcomes. This resource aims to facilitate the development of pedagogically grounded AI-SP systems and advance research on AI-supported clinical reasoning training. 2026-09-10T00:56:03Z Luming Yang Haoxian Liu Siqing Li Rong Jia Yue Xiao Guanhua Chen Li Lu http://arxiv.org/abs/2511.03673v2 OriFeel: Origami-Inspired Tactile Feedback via Surface Folding 2026-09-09T21:13:04Z People passively interact with ambient surfaces such as tables, chair backs, and armrests throughout daily life, making them natural candidates for ubiquitous tactile interfaces. However, transforming these everyday surfaces into practical haptic interfaces remains challenging. Existing solutions typically rely on dense arrays of actuators, resulting in bulky hardware and high power consumption that limit their integration into ambient objects. We present OriFeel, a structure-driven tactile interface that leverages a Miura-ori folding mechanism to distribute actuation across multiple interconnected folding units. Rather than mapping actuators to individual contact points, OriFeel uses embedded cables to exploit structural interconnections, distributing actuation across the surface and enabling spatially controllable tactile output across multiple folding units. We implement prototypes using rigid and soft materials and characterize their ability to distribute actuation across interconnected units. Our results demonstrate the feasibility of structure-driven tactile feedback via coordinated folding of compliant origami surface structures, providing a practical foundation for compact, scalable ambient haptic interfaces. 2025-11-05T17:42:40Z Shubham Rohal University of California, Merced Dong Yoon Lee University of California, Merced Shijia Pan University of California, Merced