https://arxiv.org/api/VI/V/uHICg5zJQK0tpAPKZdmZ/0 2026-09-11T21:04:13Z 32695 60 15 http://arxiv.org/abs/2609.09333v1 Where Does the Human End? Creative Agency with Generative AI across Five Years of Chinese Digital Painting 2026-09-08T18:18:38Z As generative AI enters creative work, practitioners must decide where AI assistance ends and human authorship begins. Human-agent interaction (HAI) research has examined AI as a tool, collaborator, consultant, and competitor. The longitudinal problem is how these roles are revised as systems become more capable, public, and economically embedded. We report a five-year interview study with 17 Chinese digital painters, based on annual semi-structured interviews from 2021 to 2025. Participants described recurring but non-uniform patterns of protective resistance, pragmatic task delegation, and, for some, reflective agency repartitioning. Early resistance protected observation, originality, signature, and ownership from AI. Later delegation placed AI in bounded tasks such as references, backgrounds, rough sketches, and client-facing drafts. By 2025, some participants built hybrid workflows around human-only zones, while others described fatigue, precarity, or difficulty locating a remaining human role. Peer norms, emotional climates, and production pressures shaped which delegations felt useful, acceptable, or exhausting. Copyright, authorship, and creative labor remained recurring limits on what participants were willing to delegate. We frame these accounts as longitudinal agency partitioning, the situated work of deciding which stages, responsibilities, values, and claims remain human in creative human-agent interaction. We discuss design implications for revisable agency-boundary controls, provenance scaffolds, and community-facing authorship norms. 2026-09-08T18:18:38Z Yibo Meng Ruiqi Chen Shuheng Cao Weijia Zhang Chengxi Zang http://arxiv.org/abs/2609.09332v1 Early Epistemic Settlement in AI-Assisted Writing 2026-09-08T18:18:16Z A language model can resolve a writer's current organizing problem while the construction needed for her own resolution remains unfinished. I call this early epistemic settlement. The supplied organization meets every demand then governing the passage, yet proceeding from it can displace work through which the writer would have changed those demands or become able to form further organizations. I distinguish the coordination needed to complete an already formable organization from construction that changes which organizations are formable in the first place. Settlement in the first case changes the relative work still required to bring available organizations to sufficiency. In the second, it can remove the need for the work through which another organization would become formable. Even when supplied resolution and continued construction leave the same visible qualification, different dependencies in the writer's inquiry may support different later organizations. Model suggestions can also contribute to this development when writers work through them while the problem remains unresolved. In theoretical and exploratory writing, the relations developed in reaching local adequacy help determine what the writer can later defend or develop. A sound judgment that the present passage is sufficient can therefore make further inquiry dispensable before that generative work has occurred. 2026-09-08T18:18:16Z Han-yu Wang http://arxiv.org/abs/2609.09331v1 Ephemeral Feeds and Enduring Rituals: RushTok and the Formation of Event-Based Algorithmic Communities 2026-09-08T18:17:35Z Each August, TikTok's For You page turns the University of Alabama's sorority recruitment into RushTok. We examine RushTok as an event-based algorithmic community: a collective assembled around a bounded offline ritual and sustained by recommendation. Using a mixed-methods survey (n=71) and a reflexive account of creator outreach, we ask who participates, how, and with what stakes. Findings show an ambiguous and entertainment based throughline; many called it a community (51/71) but few claimed membership (11/71). Affiliation centered on creators rather than shared practices, with parasocial attention clustering around a small set of potential new members (PNMs) and returning figures. Higher content exposure tracked with self-identification as a community member; those members commented, followed creators, and engaged across videos. Attempts to interview creators were met with silence or refusals, reflecting community boundary-work despite viral visibility. We outline implications for platform governance, including time-bounded context, graduated visibility, and aftercare. 2026-09-08T18:17:35Z Emelia Hughes Tim Weninger http://arxiv.org/abs/2609.09321v1 Endorsement Without New Evidence: How Sequential Voting Inflates Mandates in Online Community Governance 2026-09-08T18:06:15Z Online communities often treat large support margins in public elections as strong mandates. We argue that such margins can overstate the independent scrutiny behind a decision. Using 198,275 free-text rationales from Wikipedia admin elections, we introduce vote-text divergence, a measure that flags a decisive vote paired with a thin, deferential rationale. Divergence rises as voters arrive later, even after controlling for voter and election fixed effects. The pattern is consistent with information saturation: once prior text is accounted for, arrival order no longer predicts divergence, while accumulated prior evidence does. The effect is strongest among peripheral voters in the co-voting network. Yet divergence does not predict worse post-promotion outcomes, such as administrative activity or survival. Public tallies can therefore weaken the scrutiny signal even while selecting capable administrators: a margin may appear to reflect more consensus and support than it actually contains. 2026-09-08T18:06:15Z 11 pages, 2 figures, 5 tables. Under review Zihan Chen Lei Nico Zheng Di Zhu http://arxiv.org/abs/2605.05682v3 PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI 2026-09-08T17:58:26Z Recent developments in AI safety research have called for red-teaming methods that effectively surface potential risks posed by generative AI models, with growing emphasis on how red-teamers' backgrounds and perspectives shape their strategies and the risks they uncover. While automated red-teaming approaches promise to complement human red-teaming through larger-scale exploration, existing automated approaches do not account for human identities and rarely incorporate human inputs. In this work, we explore persona-driven red-teaming to advance both automated red-teaming and human-AI collaboration. We first develop PersonaTeaming Workflow, which incorporates personas into the adversarial prompt generation process to explore a wider spectrum of adversarial strategies. Compared to RainbowPlus, a state-of-the-art automated red-teaming method, PersonaTeaming Workflow achieves higher attack success rates while maintaining prompt diversity. However, since automated personas only approximate real human perspectives, we further instantiate PersonaTeaming Workflow as PersonaTeaming Playground, a user-facing interface that enables red-teamers to author their own personas and collaborate with AI to mutate and refine prompts. In a user study with 11 industry practitioners, we found that PersonaTeaming Playground enabled diverse red-teaming strategies and outputs that practitioners perceived as useful, and that AI-generated suggestions in the PersonaTeaming Playground encouraged out-of-the-box thinking even when practitioners did not follow them strictly. Together, our work advances both automated and human-in-the-loop approaches to red-teaming, while shedding light on interaction patterns and design insights for supporting human-AI collaboration in generative AI red-teaming. 2026-05-07T05:19:51Z Accepted to The ACM Symposium on User Interface Software and Technology (UIST) 2026 Wesley Hanwen Deng Mingxi Yan Sunnie S. Y. Kim Akshita Jha Lauren Wilcox Kenneth Holstein Motahhare Eslami Leon A. Gatys http://arxiv.org/abs/2609.09112v1 Travel Package Booking Application with API Bot 2026-09-08T17:45:03Z These days we are witnessing many mobile applications based on the recommended systems, which have become a great technology which is been used by the various mobile applications according to the situation. Recommendation provided by the mobile application is a key element for the person who is traveling to several places. For any tourist information application contextual information is much needed to guide the user on his interests this can be achieved by the Context-aware computing. Which provides the user most interactive system with the suggestions provided by it based on the input from the user in a certain location, here context includes the user's mental, social, physical environments. To achieve this contextual information, we will design and implement the context-aware user interface based on the user for which we have to study the user and design a rich user interface. The final outcome for which users have the satisfaction when using context-aware functionality will be much better than non-context-aware application. 2026-09-08T17:45:03Z 6 pages. Originally published in the International Journal of Recent Technology and Engineering (IJRTE), Volume 8, Issue 4, November 2019 International Journal of Recent Technology and Engineering (IJRTE), Vol. 8, Issue 4, November 2019, pp. 10199-10204 K Sai Karthik CH Naveen Aaditya Ravi Kiran Swarnalatha P 10.35940/ijrte.D8706.118419 http://arxiv.org/abs/2608.23137v5 Simulate, record, verify: A language-portable framework for muscle-grounded articulatory QA (extended version) 2026-09-08T17:33:46Z Articulatory corpora from real-time MRI and electromagnetic articulography capture tongue motion but carry no traceable labels for the muscle-driven process behind each configuration, and authoring such supervision by hand, separately for every language, does not scale. We present a simulator-based framework that turns controlled biomechanical inputs into verifiable, language-portable QA supervision. Each simulated configuration is stored with its generating input as a structured fact record; deterministic generators derive gold answers from records alone; and naturalization changes only surface form, with every output checked against its record. A new language therefore needs only a renderer and a lexicon, and new question types need no re-simulation. Instantiated as 3DTongueQA on the ArtiSynth Badin tongue model, 295,115 valid meshes yield 891,156 record-checked QA per language in English and Korean (87.2\% and 88.6\% first-pass verification); a Spanish renderer authored in about 20 minutes reaches 94.1\%, and the checker detects 97--99\% of injected corruptions. The generated supervision is domain-specific: zero-shot GPT-5 Pro reaches 7.2 Muscle EM, whereas a SpiralNet++--Qwen3-8B model trained on it reaches $62.9\pm9.2$ (2.2 with shuffled meshes) and task-specific readouts reach $88.7\pm0.7$. Code and templates: https://github.com/esh0504/muscle-grounded-qa. 2026-08-24T11:43:30Z 16 pages, 5 figures, 15 tables Seungho Eum Unsang Park http://arxiv.org/abs/2609.09070v1 Performance of Clinical AI System and Physicians and Frontier Language Models in primary care diagnostics 2026-09-08T17:25:47Z Clinical AI evaluation should encompass diagnosis and management after adaptive information gathering. We compared Doctorina, eight physicians and four standalone frontier language models in 150 synthetic Polish-language primary-care consultations. Doctorina achieved 82.0% Top-1 concordance versus 57.0% for physicians (difference, 25.0 percentage points; 95% confidence interval, 17.7-32.7) and 97.3% versus 85.0% primary-or-reference-differential concordance. Across 149 case pairs, normalized workup and treatment scores were 89.4 versus 66.9 and 83.7 versus 61.2. Doctorina had the highest diagnostic point estimates among all six groups; Kimi K3 ranked next, while Claude Opus 5 led the closely spaced management estimates of Opus, Doctorina and Kimi. A second Doctorina execution reproduced the advantages over physicians across all outcomes. Doctorina's advantage over physicians therefore extended from primary-diagnosis selection to higher-rated diagnostic workup and initial treatment after adaptive consultation. 2026-09-08T17:25:47Z Andy Nkansah Hanna Plotnitskaya Stanislau Salavei Anna Kozlova Piotr Gibas Julian Milek Viktar Harbachou Aleksey Ropan Pavel Satalkin http://arxiv.org/abs/2609.09061v1 Location-Independent Robot-Assisted Finishing Using Digital Twins and Extended Reality 2026-09-08T17:16:46Z This paper presents a cyber-physical system (CPS) for location-independent programming, supervision, training, and teleoperation of a Robot-Assisted Finishing (RAF) system used to post-process metal additive-manufactured (AM) components. A digital twin (DT) built in Unity is delivered to the operator as a WebGL application that supports both desktop and immersive modes through WebXR-compatible devices. Moreover, it exchanges robot state and pose commands with a collaborative robot through a Message Queuing Telemetry Transport (MQTT) broker. The DT enforces kinematic and collision constraints before a pose is released to the physical robot, and augments the virtual component with a color map of the surface topography that supports operator decisions on part repositioning or process termination. The architecture was validated on a specially designed physical RAF system. A steady-state joint synchronization error of 0.12 deg and a mean round-trip latency of 563 ms were measured, which is adequate for supervisory programming and intermittent teleoperation. 2026-09-08T17:16:46Z 6 pages, 5 figures, 2027 IEEE/SICE International Symposium on System Integration Jose Outeiro Jia Holt Tero Kaarlela Khalil Chakal http://arxiv.org/abs/2609.09038v1 Do Reasoning Representations Help Humans Evaluate LLM Outputs? 2026-09-08T17:03:57Z Reasoning representations are increasingly used as explanations for large language model outputs. Yet they are typically evaluated with model-centric criteria, such as answer accuracy and faithfulness, leaving it unclear whether they help people evaluate model responses. In this work, we study reasoning representations as human-facing interfaces rather than proxies for model reasoning ability. We conduct a controlled human study of six reasoning formats across tasks of varying complexity, supported by a web-based framework that randomizes task domains, problem instances, and representation order. The study collects fine-grained judgments of structural understanding, error detection and localization, and trust calibration. Our study shows a mismatch between perceived preference and support for human evaluation. Participants prefer planning- and decomposition-based representations, but simpler chain-of-thought traces better support verification, trust, and interpretability. Preferred representations also introduce calibration risks, with more false alarms on correct traces and high trust despite low willingness to verify. 2026-09-08T17:03:57Z 19 pages. EMNLP 2026 (Findings) Jaewoo Lim Sungbok Shin Sanghyun Hong http://arxiv.org/abs/2609.08982v1 Embedded Human-Centered Data Science in a Graduate Programming Course: A Framework and Case Study 2026-09-08T16:26:39Z As AI and data-driven systems pervade practice, there is an imperative for instructors to embed societal impact and ethics content into computing courses. In response, we present the Human-Centered Education for Learning in Information and eXplainable Computing (HELIX) framework for information science programs, organized around three iterative pillars - knowledge building, decision-making, and empowerment - with concrete actions for instructors and students. We applied the framework in a graduate, introductory programming course using readings, algorithmic design activities, and scenario-based reflections. We present a pilot implementation of this framework to examine changes in students' (n=22) knowledge acquisition, decision-making processes, and self-reflection regarding human-centered perspectives in data science. We release an anonymized materials kit (survey, assignments, analysis code) to support adoption. We discuss design tensions (workload, assessment, relevance to diverse information science learners) and provide guidelines for integrating human-centered content without overwhelming technical outcomes. Findings suggest that the HELIX Framework is feasible in information science contexts and future work should use comparative survey assessment to strengthen causal inferences. 2026-09-08T16:26:39Z Victoria Chui Kelly McConvey Daniel Chui Malayna Bernstein Shion Guha http://arxiv.org/abs/2609.08909v1 To Stop or Not to Stop: Exploring the Intention-Behavior Gaps in Smartphone Usage 2026-09-08T15:42:13Z As smartphones become integral to daily life, researchers have sought to identify when the use becomes problematic. Previous studies have operationalized problematic smartphone usage (PSU) from either an intention or a behavior perspective. Both risk delivering interventions not welcomed by users. We propose a novel approach to operationalizing PSU as the intention-behavior gap (IBG). We collected self-reported data on intentions to stop phone usage, alongside usage behavior data, from 37 participants over two weeks. We calculated IBG, examined effects of demographic and contextual variables, and developed machine learning models to predict IBG in real time. We found that IBG was explained by gender, time, app, and input interactions, among other factors. Intention was predicted most accurately with only personal data, whereas behavior and IBG were predicted most accurately with both personal and global data. Our findings can inform the design of future intervention tools optimized for timing and adaptive intensity. 2026-09-08T15:42:13Z MobileHCI 2026 Jian Zheng Eun Kyoung Choe 10.1145/3821659 http://arxiv.org/abs/2609.08890v1 Healthcare Utilization, Chronic Condition Management, and Workplace Functioning Among Users of a Purpose-Built Mental Health AI (Ash): Cross-Sectional Study 2026-09-08T15:27:30Z Mental health challenges can exacerbate physical symptoms and complicate management of chronic conditions. Purpose-built artificial intelligence (AI) tools may offer scalable support for co-occurring mental and physical health concerns. This cross-sectional study compared past-6-month healthcare utilization, chronic condition management, physical health behaviors, mental health change, and workplace functioning between active (n = 169) and non-users (n = 73) of a mental health AI (Ash). Participants had at least one chronic condition (e.g. hypertension, chronic pain). Binary outcomes were modeled as adjusted risk differences (RDs) using linear probability models and continuous outcomes were modeled with linear regression; all models were adjusted for hypertension. Relative to non-users, active users were more likely to report improved mental health (61.4% vs. 34.3%; RD = 0.27), higher medication adherence (91.7% vs. 76.4%, RD = 0.15), fewer skipped or delayed chronic-condition care activities (b = -0.44), and were less likely to report repeat urgent care visits (9.5% vs. 23.3%; RD = -0.15) and monthly-or-more absenteeism (24.2% vs. 45.2%; RD = -0.20, all ps < .05). Findings provide preliminary evidence that use of purpose-built AI may be associated with positive symptom-based and utilization outcomes for those managing co-occurring mental and physical concerns. 2026-09-08T15:27:30Z Kristen M. Van Swearingen Thomas D. Hull Jeffrey Swigert Caitlin A. Stamatis http://arxiv.org/abs/2609.08806v1 ArmPoser: Real-Time, Calibration-Free Arm Pose Estimation from Smartwatch IMU 2026-09-08T14:33:30Z Arm pose estimation enables applications in fitness, extended reality input, rehabilitation, and life logging. Prior smartwatch-based approaches rely on calibration poses and preprocessing pipelines that transform raw IMU measurements into standardized training formats. These steps hinder deployment in everyday settings and introduce errors due to imperfect calibration and sensor drift. We present ArmPoser, a calibration-free arm pose estimation system using a single smartwatch IMU. Our central contribution is training models directly in the reference frame native to consumer smartwatches, aligning learning with how IMU data is produced by deployed devices. By operating on device-native axes, ArmPoser removes the need for coordinate transformations, explicit alignment, and bone-offset calibration used in prior work. We further augment training with physically grounded variations in watch placement and arm morphology to account for user-specific variability. ArmPoser also includes a wear-configuration module that infers anterior or posterior forearm placement and crown orientation. We evaluate pose estimation on public benchmarks and on a 10-participant, 30-activity study using watchOS and Android smartwatches, where ArmPoser matches or exceeds calibrated baselines without any user calibration. 2026-09-08T14:33:30Z Bishnu Dev Vasco Xu Xi-Aan Loh Chenfeng Gao Henry Hoffmann Karan Ahuja http://arxiv.org/abs/2504.02109v2 A Systematic Review of Security Communication Strategies: Guidelines and Open Challenges 2026-09-08T14:19:09Z Cybersecurity incidents such as data breaches have become increasingly common, affecting millions of users and organizations worldwide. The complexity of cybersecurity threats challenges the effectiveness of existing security communication strategies. Through a systematic review of over 3,400 papers, we identify specific user difficulties including information overload, technical jargon comprehension, and balancing security awareness with comfort. Our findings reveal consistent communication paradoxes: users require technical details for credibility yet struggle with jargon and need risk awareness without experiencing anxiety. We propose seven evidence-based guidelines to improve security communication and identify critical research gaps including limited studies with older adults, children, and non-US populations, insufficient longitudinal research, and limited protocol sharing for reproducibility. Our guidelines emphasize user-centric communication adapted to cultural and demographic differences while ensuring security advice remains actionable. This work contributes to more effective security communication practices that enable users to recognize and respond to cybersecurity threats appropriately. 2025-04-02T20:18:38Z Carolina Carreira Alexandra Mendes João F. Ferreira Nicolas Christin