https://arxiv.org/api/xD063qLhWdUHtb4MZZlrZtYeWl8 2026-09-11T20:47:49Z 32695 45 15 http://arxiv.org/abs/2609.09960v1 Somatosensory Activation and Attentional States in Creative Making 2026-09-09T09:48:43Z The methods for capturing the creative process come with associated tensions around memory recall, articulation, and communication during the act of making, as well as how to record these considerations. This paper has a twofold purpose: first, to offer an example of a mixed methodology, drawn from dance anthropology, sensory ethnography, and design, that applies embodied methods as an alternative for documenting creative making. Specifically, this incorporates the researcher-as-participant and the collation of fieldnotes, embodied knowledge/movement recall, with notation forms, and participant interviews. These are existing methods in dance anthropology; however, using them alongside exploratory prototyping and workshop approaches broadened this work into transdisciplinary practice. Second, it discusses the activation of somatosensory systems through wearable technology and the facilitation of heightened sensory awareness for the practitioner, leading to a subsequent ability to focus on creative decisions linked to reflection and metacognition. 2026-09-09T09:48:43Z In Proceedings of The First Reflection in Creative Experience (RiCE) Workshop (RiCE W1) arXiv:2607.24558 Katherine Rees http://arxiv.org/abs/2609.04355v2 VLA-Precision: Asymmetric Co-Bootstrapping for Efficient Real-World Online RL of Vision-Language-Action Models 2026-09-09T09:42:55Z Pretrained vision-language-action (VLA) models enable broad manipulation but remain unreliable in tasks demanding precision and repeatability. Applying real-world online reinforcement learning (RL) to VLA post-training enables autonomous trial-and-error improvement beyond demonstrations alone, but exposes two bottlenecks: 1) unreliable value signals can induce policy drift; 2) large-VLA overhead constrains throughput and sample efficiency. To address these challenges, we present VLA-Precision, an efficient real-world online RL framework featuring the Asymmetric Co-Bootstrapping (ACoB) algorithm and the ACoB-Stream architecture. Specifically, ACoB establishes asymmetric co-bootstrapping across timescales: early intervention-guided behavioral learning rapidly improves policy performance while enhancing online experience quality. As autonomous experience accumulates, global return propagation and local preference ranking progressively calibrate value estimates, yielding relative action advantages for reference-regularized policy improvement while suppressing drift. To enable ACoB on large VLAs, we develop ACoB-Stream, a closed-loop experience--policy architecture that establishes invariant-state decoupling and on-demand streaming as design principles, delivering up to 10.9$\times$ improvements in throughput and computational efficiency. Extensive evaluations on nine high-precision chemistry tasks across four categories and four robot embodiments show that VLA-Precision achieves 98.3\% mean success rate in 45.8 min/task, with 27.6 s episodes running at 1.2$\times$ and 1.8$\times$ the speeds of VLA and RL baselines. Resources are available at https://vla-precision.github.io. 2026-09-03T18:19:36Z 17 pages, 14 figures Chenyu Su Zhaolong Shen Yuan Qian Chen Qian Rui Zhang Feng Yan Weixing Chen Fei Zhang Jiamin Wang Shuang Cong Weiwei Shang http://arxiv.org/abs/2602.06759v3 "Tab, Tab, Bug": Security Pitfalls of Next Edit Suggestions in AI-Integrated IDEs 2026-09-09T07:36:24Z Modern AI-integrated IDEs are shifting from passive code completion to proactive Next Edit Suggestions (NES). Unlike traditional autocompletion, NES is designed to construct a richer context from both recent user interactions and the broader codebase to suggest multi-line, cross-line, or even cross-file modifications. This evolution significantly streamlines the programming workflow into a tab-by-tab interaction and enhances developer productivity. Consequently, NES introduces a more complex context retrieval mechanism and sophisticated interaction patterns. However, existing studies focus almost exclusively on the security implications of standalone LLM-based code generation, ignoring the potential attack vectors posed by NES in modern AI-integrated IDEs. The underlying mechanisms of NES remain under-explored, and their security implications are not yet fully understood. In this paper, we conduct the first systematic security study of NES systems. First, we perform an in-depth dissection of the NES mechanisms to understand the newly introduced threat vectors. It is found that NES retrieves a significantly expanded context, including inputs from imperceptible user actions and global codebase retrieval, which increases the attack surfaces. Second, we conduct a comprehensive in-lab study to evaluate the security implications of NES. The evaluation results reveal that NES is susceptible to context poisoning and is sensitive to transactional edits and human-IDE interactions. Third, we perform a large-scale online survey involving over 200 professional developers to assess the perceptions of NES security risks in real-world development workflows. The survey results indicate a general lack of awareness regarding the potential security pitfalls associated with NES, highlighting the need for increased education and improved security countermeasures in AI-integrated IDEs. 2026-02-06T15:06:36Z To appear in ACM CCS 2026 Yunlong Lyu Yixuan Tang Peng Chen Tian Dong Xinyu Wang Zhiqiang Dong Hao Chen 10.1145/3830454.3846699 http://arxiv.org/abs/2609.09789v1 Pairit: A Platform for Live Experiments on Human-AI Collaboration 2026-09-09T06:41:50Z Organizational design in the era of artificial intelligence requires experimental methods that can test how human-AI groups coordinate, delegate, and make decisions. Programmable platforms coordinate live human-to-human sessions or real-time human-AI chat, but researchers cannot easily declare experiment protocols in which AI participants both communicate and act on shared work within one auditable configuration. Here we introduce Pairit, an online platform that facilitates the design, testing, and deployment of experiments that test human-AI organizational designs and interventions. Through a single YAML configuration file, researchers declare an executable experiment graph (pages, routing, randomization, matchmaking, chat, shared workspaces, server-hosted agents, surveys, timers, and custom HTML components) and combine any number of humans and AI agents in live sessions. We have validated the feasibility of the platform through multiple live deployments, including peer-reviewed published studies, capturing high-resolution process traces of communication, negotiation, and collaborative work in live human-AI dyads. By representing complex interactive protocols as standardized, auditable configuration files, Pairit provides reusable infrastructure for specifying, deploying, and sharing live human-AI organizational experiments. 2026-09-09T06:41:50Z 15 pages, 3 figures Harang Ju Sinan Aral http://arxiv.org/abs/2609.09713v1 How Far Do Capability Cues Travel? Anthropomorphism and Differentiated Trust in a Platform-Embedded AI Assistant 2026-09-09T04:55:33Z Visible AI capabilities need not translate into broader judgments of trustworthiness. In a randomized 2 x 2 experiment with 270 U.S.-based Reddit users, an embedded assistant displayed one or three functions, with or without a brief rationale. Displaying three functions increased perceived multifunctionality; no other randomized main effect survived correction across the six outcomes. Rationale availability did not reliably increase perceived intelligence. Exploratory analysis indicated stronger uptake of the functional display at higher objective AI literacy. Among concurrently measured judgments, perceived multifunctionality was associated with perceived intelligence, which was associated with anthropomorphism and all three trust dimensions. After accounting for perceived intelligence, anthropomorphism was positively associated with benevolence, but not reliably with integrity or ability. These findings separate interface effects from relationships among users' perceptions and show why ability, integrity, and benevolence should be evaluated separately. 2026-09-09T04:55:33Z Chenchen Mao Hanjing Shi Haiyan Jia Dominic DiFranzo http://arxiv.org/abs/2609.09700v1 AppetiteCheck: Feasibility of Momentary Vagus Nerve Stimulation as an Implicit Intervention for Eating Behavior 2026-09-09T04:34:42Z Overeating and emotional eating are common health issues that affect people even without an eating disorder. The vagus nerve plays a critical role in the gut-brain axis, and implanted vagus nerve stimulators have been associated with reduced appetite. In this paper, we propose transcutaneous cervical vagus nerve stimulation (tcVNS) as a ubiquitous system to provide immediate, low-effort intervention during an eating episode. In a study with 24 participants, we evaluated a mobile, handheld tcVNS device during a single episode of distracted snacking. We found that participants ate 9.6% less and 23.6% more slowly during vagus nerve stimulation than during sham stimulation. Post-snacking satiety was the same in both conditions, while heart rate was lower during vagus nerve stimulation. The stimulation was described as subtle and barely noticeable. Overall, these results provide evidence for the feasibility of non-invasive vagus nerve stimulation as a low-attention intervention for managing eating behavior -- one that can be packaged inside ubiquitous interactive systems. As such, we extend the design space of implicit interfaces toward physiological intervention, motivating future ubiquitous systems that pair eating-related sensing with low-attention interventions. 2026-09-09T04:34:42Z 28 pages, 9 figures. Published in Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies, Volume 10, Issue 3, 2026 Proc. ACM Interact. Mob. Wearable Ubiquitous Technol. 10, 3, Article 101 (September 2026), 28 pages Tan Gemicioglu Jas Brooks Pedro Lopes Tanzeem Choudhury 10.1145/3831630 http://arxiv.org/abs/2601.15295v2 Elsewise: Authoring Open-ended Interactive Narrative with Possibility Space Visualization 2026-09-09T03:07:57Z Interactive narrative (IN) authors craft spaces of divergent narrative possibilities for players to explore, with the player's input determining which narrative possibilities they actually experience. Generative AI can enable new forms of IN by improvisationally expanding on pre-authored content in response to open-ended player input. However, this extrapolation risks widening the gap between author-envisioned and player-experienced stories, potentially limiting the strength of plot progression and the communication of the author's narrative intent. To bridge the gap, we introduce Elsewise: an authoring tool for LLM-based INs that implements a novel Bundled Storyline concept to enhance author's perception and understanding of the narrative possibility space, allowing authors to explore similarities and differences between possible playthroughs of their IN in terms of open-ended, user-configurable narrative dimensions. A user study (n=12) shows that our approach improves author anticipation of player-experienced narrative, leading to more effective control and exploration of the narrative possibility spaces. 2025-12-21T02:06:03Z Yi Wang John Joon Young Chung Melissa Roemmele Yuqian Sun Tiffany Wang Shm Garanganao Almeda Brett A. Halperin Yuwen Lu Max Kreminski http://arxiv.org/abs/2609.07001v2 Adaptive Complementarity in Human-AI Systems: Architecture as a State-Shaping Choice 2026-09-09T02:32:22Z Human-AI interaction can improve current performance while changing the capabilities and relationships on which future performance depends. We develop adaptive complementarity, a framework for choosing interaction architecture with these state consequences in view. Access, information exposure, task allocation, timing, and communication can alter which arrangement will be valuable later; their settings can often be reset faster than the capabilities, search patterns, or conventions they create. Three mechanisms organize the argument: information exposure and collective search, delegation and capability evolution, and strategic interdependence and information governance. Their integration yields cross-mechanism implications, including conditions under which a loss of expertise heterogeneity increases the information differentiation required to preserve independent search. We distinguish strong human-AI complementarity from advantage over another workflow and from advantage over an evolving reference policy. A knowledge-coverage illustration shows how earlier workflows create expertise profiles that can reverse current workflow rankings even at equal human competence. It distinguishes current fit from investment in future learning and shows how a stable, forward-looking workflow can capture nearly all the baseline value available to a fully informed policy. The framework directs evaluation toward the states present interaction creates, their consequences for later architectural fit, and the conditions under which observing and responding to them is worthwhile. 2026-09-07T03:47:11Z Babak Heydari http://arxiv.org/abs/2609.09575v1 Beyond Top Words: MonoTM for Topic Modeling with Interpretable Monosemantic Features 2026-09-09T00:54:46Z Topic models summarize large text corpora, but top-ranked words often provide only a limited representation of topic semantics. Sparse autoencoders (SAEs) offer a way to move beyond word-level descriptors by extracting interpretable features from dense representations, yet how feature interpretability relates to topic-inference quality remains unclear. We introduce \textbf{MonoTM}, an interpretable topic modeling framework that decouples these roles. Across three benchmark corpora, we show that document--topic mixture estimation and semantic interpretation favor different SAE configurations and feature subsets. MonoTM estimates mixtures from the full SAE bag-of-features representation and, with them fixed, learns topic descriptors over a separate vocabulary of corpus-grounded semantic features. This design preserves global topic structure while representing topics with semantic units more meaningful than individual words, making them more useful for downstream corpus analysis. 2026-09-09T00:54:46Z Accepted to appear in the Proceedings of AACL-IJCNLP 2026 Una Joh Bei Yu http://arxiv.org/abs/2609.09496v1 The Mutations of Machine Speech 2026-09-08T22:23:38Z Algorithmic outputs now populate the digital environments through which contemporary life is organized. The role of law in facilitating and constituting (rather than merely responding to) these processes is gaining increasing traction across scholarly accounts. This inquiry traces the evolution of algorithmic outputs attending to their legal underpinnings and social implications, surfacing the mutations of machine speech. The first mutation redefined speech as data to be queried: search engines transformed the web from a space of information retrieval into an economic regime of algorithmic visibility. The second mutation reframed speech as engagement: social media platforms fused moderation with amplification, turning expression into a metric of attention, governed by corporate architectures. The third mutation emerges in conversational systems and interfaces, where generative text displaces information retrieval, bringing with it dense technolegal entanglements and profound epistemic consequences. Scholars of freedom of expression, informational privacy, and communication studies have long grappled with these dynamics, yet their implications for broader legal thought have also become urgent. This piece seeks to organize and clarify the evolving debate around algorithmic speech, making this critical but often fragmented discourse more accessible to wider legal and interdisciplinary audiences. In doing so, it bridges the gap between observing technological transformation and critically assessing the constitutive role of law within it, offering a conceptual resource for researchers, students, policymakers, and practitioners navigating and contesting this evolving landscape. 2026-09-08T22:23:38Z Mauricio Figueroa 10.1007/978-3-031-87993-7_188-1 http://arxiv.org/abs/2502.03682v3 Towards On-Device Evidence Gathering for Intimate Partner Infiltration: A Feasibility Study for Joint Identity-Action Detection 2026-09-08T22:14:21Z Intimate Partner Infiltration (IPI) refers to phone-side privacy infiltration in intimate or close relationships, often enabled by physical access to a person's smartphone and discussed in technology-facilitated Intimate Partner Violence (IPV) contexts. Unlike conventional cyberattackers, IPI perpetrators leverage proximity and personal knowledge to circumvent standard protection, underscoring the need for targeted interventions, motivating device-side tools that surface such risk evidence for later review. While prior works have extensively studied IPV, and some have provided tailored and effective solutions such as security clinics, they are necessarily episodic and human-expert-intensive, and offer limited automated visibility into what happens on a smartphone between support sessions. Guided by a formative interview with experts (n=5), we take the first exploration into gathering IPI-risk evidence from a mobile system perspective and present AID, Automated IPI Detection, a data-driven system that continuously logs unauthorized access and suspicious behaviors on smartphones. In a controlled 27-participant study, AID achieves an end-to-end F1 score of 0.928 with a 7.0% false positive rate for Top-1 phone-side risk flagging; when preserving top-3 candidate action categories as report context, AID achieves an F1 score of 0.981 and a false positive rate of 1.6%. These findings demonstrate AID's potential as an evidence-support tool that complements current clinic-based interpretation and safety-planning. 2025-02-06T00:07:08Z Accepted to ACM IMWUT 2026 Weisi Yang Shinan Liu Feng Xiao Nick Feamster Stephen Xia http://arxiv.org/abs/2609.09483v1 Integrating Multi-Source Feedback in Computational Design 2026-09-08T21:58:07Z In real-world design practice, evaluations rarely rely on a single source of judgment. Designers routinely combine expert opinions, empirical studies, and computational models, each with distinct strengths and limitations. While machine learning offers methods to integrate multiple feedback sources, these approaches remain largely inaccessible to designers without technical expertise. In this paper, we explore how to integrate multiple feedback sources, primarily through: (1) a practical approach for multi-source integration, and (2) its implementation in MUSE, a no-code tool that allows designers to combine and balance diverse sources. Our technical findings show that independent modeling of multiple evaluation sources enables exploration across heterogeneous feedback, accommodates different evaluation speeds, surfaces disagreements between sources, and supports an adaptable evaluation setup that designers can reconfigure during their process. In a visualization design study, participants navigated their own judgments alongside simulator feedback, reporting a perception of enhanced confidence and flexibility. Our results highlight the viability of multi-source integration to support computational design, offering a step toward bridging the gap between advanced optimization methods and design practice. 2026-09-08T21:58:07Z Accepted for publication in ACM Transactions on Interactive Intelligent Systems (TiiS) Francisco Erivaldo Fernandes Junior Thomas Langerak Mira Keränen Danqing Shi Ardak Alipova Antti Oulasvirta http://arxiv.org/abs/2609.09472v1 Exploring 3D Glyph Physicalizations for Public Engagement through River Health 2026-09-08T21:41:52Z Introduction: In this paper, we present the preliminary design of a toolkit for making glyph-based physicalizations for public engagement. We use London river health data as a case study: a data set of significance to urban issues related to climate change and of interest to draw public attention, as part of the Greater London Authority's strategies. Design: We present the components of a 3D glyph-making toolkit, its encodings, and a step-by-step process for crafting a physicalization of a river's water quality using recycled materials. We reason about how users can use the template to learn about a data set while reflecting on the data's significance to their personal experience and self-mapping onto the physicalization. Reflection: We reflect on the opportunities that extending the design space of glyphs to 3D physicalization offers for supporting public engagement with complex, multi-dimensional data sets, scaffolding cognitive processes, and self-reflection, thereby bringing crucial environmental data to life. Conclusion: Future implementation of the 3D glyph template will enable the public of all abilities to explore river health data, physicalize complexity, and realize its relevance. We hope that its use in public engagement workshops will help raise awareness, invite care, and foster a sense of belonging. 2026-09-08T21:41:52Z 6 pages, 2 figures, 1 table. Accepted for publication in the Visualising Climate 2026 Proceedings. Conference: 4-6 Nov 2026, Bologna, Italy Maria Teresa Ortoleva Min Chen Rita Borgo Alfie Abdul-Rahman http://arxiv.org/abs/2609.09443v1 "It's Like Drinking from a Fire Hose": Understanding and Characterizing Video Learning Experiences for Individuals with ADHD 2026-09-08T20:51:32Z Video lectures have become increasingly prevalent for education and professional development, yet their static visuals, dense information, and long duration pose attentional challenges for individuals with ADHD. While adaptive learning offers opportunities towards ADHD-accessible video learning, little is known about how to suitably adapt such videos: What components in multimodal video lectures are challenging for ADHD viewers? How do these experiences surface in behavioral signals to trigger an adaptation? What presentations do they prefer? To answer these questions, we conducted an eye-tracking-based retrospective think-aloud study with 16 participants with ADHD, who watched and reflected on a curated set of video lecture segments. Our study uncovered video design elements that hindered learning and revealed participants' coping strategies along with their limitations. By jointly analyzing behavioral signals and retrospective reflections, we characterized how these experiences manifested in behavioral patterns. We further surfaced participants' practices for addressing learning needs beyond the video watching process, and derived design implications for future ADHD-friendly adaptive video learning systems. 2026-09-08T20:51:32Z Hanxiu 'Hazel' Zhu Weiyu Zhang Ru Wang Yuhang Zhao http://arxiv.org/abs/2609.09365v1 Echoes in the Algorithm: Analyzing the Fidelity of User Preferences Against Realized Platform Reach 2026-09-08T18:59:11Z What does popular content look like when platforms withhold the usual cues? On TikTok, users still form impressions about which videos are taking off even when likes and view counts are hidden, delayed, or pushed to the margins of the interface. We study this problem through TokOrNot, a web-based game in which participants compared pairs of TikTok videos and reported (i) which one they preferred and (ii) which one they believed had reached a larger audience. We benchmark these judgments against verified public view counts, which we use as a bounded proxy for realized platform reach. Across 3,513 judgments from 363 participants, participants identified the higher-reach video only modestly above chance (56.75%, 95% CI: 56.01-58.55). Preference aligned with the higher-view video at a similar rate, while preference and prediction matched in 83.48% of trials (95% CI: 83.12-85.95). Performance also varied across content categories. Taken together, these results do not suggest that users can reliably read platform success from content alone. Instead, they point to a looser and more uncertain interpretive process in which reach judgments often track personal taste or other weak heuristics when explicit popularity cues are absent. We discuss the implications for algorithmic literacy and for interface designs that reduce visible metrics without leaving users to infer reach from uneven or idiosyncratic cues alone. 2026-09-08T18:59:11Z Emelia Hughes Tim Weninger