https://arxiv.org/api/pw2gaD2QxabNXv6UL9jXnC1ZmfY2026-09-10T19:13:47Z256613015http://arxiv.org/abs/2605.10382v2DREAMS: Modelling Support for Research into Engineering and Artistic Design2026-09-06T19:01:56ZDesign Research Methodology (DRM) supports systematic design research through representations such as Reference Models and Impact Models. However, the practical construction and maintenance of these models often remains manual, requiring repeated redrawing, layout adjustment, and separate handling of assumptions, references, and supporting evidence. This can make DRM modelling time-consuming, visually cluttered, and difficult to revise as models increase in complexity. This paper presents DREAMS, an early-stage prototype modelling environment developed to support the creation and maintenance of DRM Reference Models and Impact Models. The tool enables users to construct typed causal models using DRM-relevant elements, define signed causal relationships, and attach assumptions, experiential inputs, and references directly to causal links. It also provides layout support and search functions to improve readability, modifiability, and retrieval of supporting information. A preliminary comparative evaluation with four DRM users was conducted against manual modelling practice. The results indicate reductions in model creation time, revision time, repositioning effort, edge crossings, and evidence retrieval time when using DREAMS. These findings are interpreted as early evidence of practical potential rather than full validation. The contribution of the paper lies in identifying requirements for DRM-aligned modelling support, presenting the design and implementation of DREAMS, and demonstrating its potential to reduce modelling effort and improve traceability in DRM-based research.2026-05-11T11:24:05ZAccepted for presentation to the 11th International Conference on Research Into DesignApala Chakrabartihttp://arxiv.org/abs/2609.06582v1Blind directions of physical learning networks: where to measure and what to measure2026-09-06T13:00:08ZA physical learning network is read at a few accessible nodes. Every task and learning rule that uses only the steady voltages and currents there, at a fixed operating point, acts through the boundary response map. Parameter changes in the kernel of its Jacobian are blind to first order. A walk along one fiber of that map had not terminated at a 600-step cap, one edge then at 15.2 times its start. A decomposition theorem splits the response Jacobian over the hidden components. The blind dimension adds over components whose surviving slots are disjoint. A boundary-to-boundary edge adds one parameter and deletes one slot. Exposing a hidden node changes only its component. For one hidden node the component's contribution counts the bipartite components of the non-adjacency graph of its neighbors. For a pocket of h hidden nodes a factor-analysis bound caps what outside electrodes can expose. It is attained when every pocket node meets every neighbor and no edge joins two of those neighbors, so past a threshold further electrodes outside such a pocket expose nothing. An electrode inside such a pocket, when its hidden nodes form a clique, is worth h-1 directions where one outside is worth none. Reading the accessible nodes as vector displacements rather than potentials left no deficit beyond counting in all 90 spring networks tested, each a pocket fully joined to five or more accessible nodes, and nothing blind at all in 88. What limits the reading is the quantity measured as much as the number of contacts. Maximum-weight spanning forests give a proved upper bound on the blind dimension. The matching equality is proved for one hidden node and conjectured beyond. The forest count matched the numerical blind dimension in 1,451 of 1,500 held-out networks, and no shortfall separates an incomplete search from a false equality. A self-learning circuit shows the split on hardware.2026-09-06T13:00:08Z18 pages, 6 figures, 2 tables. Code, results and data: doi:10.5281/zenodo.22537350Quoc-Bao NguyenThai-Son Vuhttp://arxiv.org/abs/2511.15872v2AI-Assisted Writing Is Growing Fastest Among Less Established Scientists in Non-English-Speaking Countries2026-09-06T03:12:56ZThe recent emergence of AI-assisted writing raises an important question: how is this new technology being adopted across the scientific community, and how does adoption vary across linguistic and professional contexts? We analyze over two million full-text biomedical publications from PubMed Central from 2021 to 2024 using a distribution-based framework to estimate AI-generated content. We found that, in biomedical publications, AI-generated content increased substantially after ChatGPT, with larger increases in publications from countries with lower English proficiency. Increases were also greater among scientists with fewer publications and citations, those at earlier career stages, and those at lower-ranked institutions. Prior AI research experience was associated with greater increases in AI-assisted writing, which were also modestly associated with greater increases in publication productivity. These findings show that AI-assisted writing is growing fastest among biomedical scientists who may have historically faced barriers, a pattern with potentially positive implications for equity in science.2025-11-19T21:00:18ZJialin LiuYongyuan HeZhihan ZhengYi BuChaoqun Nihttp://arxiv.org/abs/2609.01620v2Quantized mereology2026-09-05T20:26:53ZThis paper investigates quantized mereology, which merges quantum information theory with classical mereology, the formal analysis of the part-whole relationship. It emphasizes that information about crisp mereological relations is captured by classical bits, whereas mereological vagueness involves the use of quantum bits, or qubits, to capture mereological information. Transitioning from classical bits to qubits offers a new understanding of part-whole relations among entities that are subject to mereological vagueness. This paper methodically studies the quantized mereology that arises from the classical RCC5 formalism, one of the most widely studied formalisms in Qualitative Reasoning (QR), a subfield of Artificial Intelligence. The argument is made that this method can be extended to quantize more sophisticated QR formalisms, mirroring how strategies are used in physics to transition from classical to quantum systems. The study explores: (i) the conversion of bits to qubits while understanding that these (qu)bits convey mereological information; (ii) the representation of qubit states via set partitions, acknowledging that set theory aligns more closely with classical logic than with the complex vector spaces employed in quantum physics; (iii) demonstrating that classical QR formalisms can be seen as special cases within quantized formalisms; and (iv) showing that qubits more effectively encapsulate information about vague phenomena.2026-07-15T22:40:07ZThomas Bittnerhttp://arxiv.org/abs/2607.25677v2Open-ended innovation arm-race in zero-sum games2026-09-05T15:00:05ZThis note discusses zero-sum games with open-ended innovation, whereby each player may introduce new strategies. The innovation process is modelled as a draw of new strategies form a distribution. It is argued that, when the cost of innovation is vanishingly small, this setting can lead to an everlasting innovation arm-race. In particular, the introduction of new technologies of the advanced player increases the marginal utility for technological innovation of the backward one, whereas innovation of the backward player disincentivizes the more advanced one to innovate.2026-07-28T12:54:39Z5 pages, 1 figuresMatteo Marsilihttp://arxiv.org/abs/2609.06005v1Price Dislocations, News Citations, and Epistemic Leverage on Polymarket2026-09-05T10:14:50ZPrediction-market probabilities increasingly appear in news coverage, yet little is known about which market movements become news or how much trading money sits behind the numbers journalists quote. Unlike a poll, a market price can be moved by anyone willing to trade, so the cost of manufacturing a number that circulates as news bears directly on the information environment. We link 173.7 million signed Polymarket trades to news coverage from 2024-2025. From 6,990 articles mentioning prediction-market venues, an LLM-based, human-validated matcher extracts 1,582 sentences quoting market odds and attributes 918 to the specific market whose price they cite. We then detect 44,976 price dislocations, movements of at least five percentage points backed by concentrated one-sided trading, and ask whether a market is cited more often afterward. In the days after a dislocation, a market's citation rate is about 33% higher than its matched baseline (log citation-rate ratio $τ_{\mathrm{cite}}=0.283$, permutation $p=0.001$), robust to binary and Poisson count outcomes. Yet move size is not the strongest predictor of citation: prominence dominates (standardized $β=0.610$ vs. $β=0.159$ for move size). Finally, we combine the dollar flow behind a given price change with observed citation rates into a metric we call epistemic leverage, the dollars needed to move a market five points and have the move cited. It stays near \$0.7-1.0 million across prominence quintiles, because cheaper-to-move markets are proportionally less likely to be cited. The implied threat model centers not on the long tail of cheaply moved markets but on the few prominent markets newsrooms treat as informational infrastructure, where a seven-figure price of influence sits within the budgets of actors with a large stake in the quoted number. We release aggregate event-study data and validation materials.2026-09-05T10:14:50Z17 pages, 3 figuresHazem IbrahimYasir Zakihttp://arxiv.org/abs/2505.13803v4GenAI Models Capture Urban Science but Oversimplify Complexity2026-09-05T04:45:48ZGenerative artificial intelligence (GenAI) models are increasingly used for scientific data generation, yet their alignment with empirical knowledge in urban science remains unclear. We therefore ask whether generated urban data reproduce empirical regularities and support repeatable experiments. We introduce AI4US, a framework for evaluating data synthesis and conditional intervention across text and image modalities. Four cases spanning urban systems, within-city structure, neighbourhood vitality and streetscape perception are evaluated against published parameters, observed urban data and human judgements. Generated outputs recovered recognizable scaling and distance-decay patterns and yielded measurable associations between neighbourhood morphology indicators and pedestrian activity, while model judgements showed positive agreement with sampled human choices. Controlled changes to urban conditions produced repeatable output responses across all four cases. However, generated data often compressed empirical numerical coverage, local variation and visual diversity. In an inspectable GenAI model, intermediate activations linearly distinguished some urban relationships, while activation edits changed only selected output scores. AI4US provides an empirically grounded approach for assessing GenAI as the virtual urban laboratory in urban data synthesis and controlled model experiments, while revealing its tendency to simplify empirical complexity.2025-05-20T01:32:05Z47 pages, 17 figuresYecheng ZhangRong ZhaoZimu HuangXinyu WangYue MaYing Longhttp://arxiv.org/abs/2609.05754v1The collective dynamics of online harassment2026-09-04T22:33:19ZFringe online message boards are often studied in the context of the extreme ideology that they produce. So far, however, not much of this research has focused on direct real-world harm in the all-too-common form of collective harassment. We directly analyze the complex dynamics of KiwiFarms, an online message board dedicated largely to the harassment of individuals from vulnerable communities. We conduct exploratory analyses of the hyperlink structure of the platform and linguistic changes over time, and prospective modeling of thread size. After establishing this broader picture of the complex traits of the system, we observe the temporal evolution of community-specific vocabulary, finding that the community's framing of their harassment targets persistently evokes more danger in the early 2020s than the late 2010s. We lastly find that early thread-growth behavior is predictive of longer-term thread virality. We discuss the implications for broader understanding of toxic online behavior and threat assessment, and make the case for studying fringe platforms as complex systems with significant societal impact.2026-09-04T22:33:19ZBenjamin Freixas EmeryBrian C. Keeganhttp://arxiv.org/abs/2609.05591v1WolfSociety: Understanding Collective Risk from Harmful-Agent Scaling in Financial Agent Societies2026-09-04T17:36:31ZSafety evaluations typically focus on individual agents, but interacting agents can spread harmful information and influence the environment in which later decisions are made. We study how collective failure changes with harmful-agent fraction and society size in a controlled financial agent society, where agents communicate over a social network and trade in a shared market. In the primary financial scenario, collective failure requires broad harmful diffusion together with severe price dislocation or liquidity stress. Across all tested society sizes, failure remains rare at low harmful fractions but rises sharply over a narrow range. As society size grows from N=100 to N=2000, the harmful fraction associated with a 50% failure probability decreases from 4.7% to 2.2%, while the corresponding number of harmful agents increases from approximately 5 to 44. In contrast, when the number of harmful agents is held fixed, their impact becomes weaker as the society grows. Controlled interventions further show that broader network reach shifts the collapse boundary toward lower harmful fractions, whereas stronger conformity alone has little effect. To characterize these effects, we introduce Agent Society Dynamics, a finite-size framework for relating harmful-agent fraction, society size, and interaction structure to collective failure. Overall, our results reveal a nonlinear, size-dependent collapse transition in financial agent societies, showing that collective failure depends not only on the prevalence of harmful agents but also on the size and interaction structure of the surrounding society. Code is available at https://github.com/SAIL-Research-Lab/WolfSociety.2026-09-04T17:36:31ZLejun ZhangSarah Lu-LiangXin JiangMuning WenWeinan ZhangShangding Guhttp://arxiv.org/abs/2609.05342v1Mitigating Disease Spread by Design in Refugee and IDP Camps2026-09-04T16:45:24ZDisease spread represents an increasing challenge in refugee and internally displaced person (IDP) settlements. The movement and interaction of people within camps is influenced by their layout, which therefore has the potential to significantly affect disease spread. This work aims at creating a methodology to explore the potential effects of different camp layouts as mitigating factors in the spread of diseases within settlements. We showcase proof-of-concept experiments by leveraging the JUNE agent-based epidemic model, discuss the kind of operational insights this methodology can facilitate, and provide a framework for future investigations.2026-09-04T16:45:24Z9 pages, 9 figures2023 ICRL First Workshop on Machine Learning and Global HealthGiulia ZarpellonJoseph Aylett-BullockFrank KraussMiguel Luengo-Orozhttp://arxiv.org/abs/2609.05584v1One equation allocates playing time in competitive team sports: the rulebook sets the arithmetic and the league sets the habit2026-09-04T16:00:10ZEvery team sport that lets a team replace players rations the replacements, and the ration settles who plays and for how long. Across eight codes, three in both sexes, we measure how playing time is allocated and find one equation behind it, complete in the one-way codes, its terms separating what the rulebook fixes from what a competition chooses. Under scarcity every league behaves alike: loaded starters are kept on at the same rate on four continents. Under freedom each expresses a preference, and the size of a league's departure is priced near one-for-one by what its own withdrawal minute prices (lambda = 0.87, 95% interval 0.54 to 1.21, twelve leagues), a law confirmed in four out-of-sample predictions and refused in one. The ordering of leagues by appetite persists across a rule change, and club wealth does not price it. Rules are the geometry of a competition; culture is its material.2026-09-04T16:00:10Z36 pages, 5 figures, 1 table, 5 Extended Data figures and 4 Extended Data tables; 75-page Supplementary Information included as an ancillary file. Replication record (manuscript, pipeline, tests and derived tables as one object): https://doi.org/10.5281/zenodo.22304792Gustavo Pedro Ricouhttp://arxiv.org/abs/2609.01633v2Omega-N: Interpretable Structural Node Descriptors and Their Applicability Domain2026-09-04T14:36:10ZA composite structural index summarises a network in one number, and for a triangle-based index it is spectrally redundant: Tr(A^3) is the third moment of the adjacency spectrum. The non-redundant content sits one level down, in diag(A^3), which depends on eigenvectors and is not spectrally determined. A corollary in the theory paper predicted that the global scalar should tie sharpened spectral baselines rather than beat them, while the node-wise attribution should do better where the number of structural epicentres is unknown.
We construct Omega-N by localizing each of the four factors. The direct localization is badly conditioned; two corrections from published practice fix it, a configuration-null excess per factor and a personalized-PageRank neighbourhood at several scales, giving ten interpretable features per node, with no attributes, training or embeddings.
Against a recursive feature engine at five levels of recursion, Omega-N wins on three and ties on two of the six in-domain evaluations, the sixth a declared null where every arm returns chance, with ten features against its 28 to 252 before pruning. Two statistics from the graph and labels, not from performance, partition the eight benchmarks without error, and the two they exclude are the two on which it loses.
The strongest application is drug-target prioritisation on protein interaction networks: +0.032 to +0.103 AUPRC over a six-feature centrality battery and +0.084 to +0.208 over the four-feature one, across three constructions, replicated on an independent AP-MS network and label source (degree-matched: +0.0723 on STRING, +0.0560 on BioPlex, p=0.00195). Adding Omega-N to centralities plus Node2Vec changes nothing. The claim is narrow and it is the point: ten named features, computed without training, match or beat hand-crafted centralities and a recursive engine, and do not touch learned representations.2026-08-21T16:16:30Z18 pages, 3 figures. Reference implementation, notebooks and data-preparation scripts at https://github.com/BiomeMakers/OmegaNAlberto Acedohttp://arxiv.org/abs/2609.04988v1A Strictly Proper Scoring-Rule Theory for Calibrating Stochastic Car-Following Models2026-09-04T10:46:01ZProblem definition: Fixed parameters and inputs in a stochastic simulator induce a distribution over complete trajectories, not one trajectory. Calibration must assess this distribution, including variability and temporal dependence, against observations. Yet stochastic car-following models are commonly calibrated with trajectory-error objectives inherited from deterministic modelling. Methodology/results: We establish a scoring-rule theory of stochastic calibration. Strict propriety requires the data-generating distribution to uniquely minimise expected score. MRMean-I, the average run-wise error, drives separable stochastic spread to zero; MRMean-II, the error of the ensemble-mean trajectory, cannot identify a parameter that changes only spread; and MRMin, the error of the closest simulated run, has a population target that changes with ensemble size. These results are confirmed for stochastic Intelligent Driver Model extensions with additive acceleration noise and random desired headway. We recommend exact maximum likelihood when the correct transition density is available; otherwise, an unbiased simulation-based estimator of a strictly proper score. The energy score meets this requirement and gives the best held-out distributional prediction among the evaluated simulation-based objectives, although both models retain too-narrow bands and miss persistent disturbances. Implications:Strict propriety separates a valid calibration target from parameter identifiability and model adequacy. The theory applies to vector-valued outputs from stochastic transportation simulators; the car-following experiments illustrate its scope.2026-09-04T10:46:01ZShirui ZhouShiteng ZhengJunzhe DingRui JiangJunfang Tianhttp://arxiv.org/abs/2609.04692v1Low-Dimensional Phase Diagram of Higher-Order Networked Systems2026-09-04T03:48:29ZHigher-order networks exhibit rich critical phenomena that cannot be captured by traditional pairwise models. Here, we develop an analytical dimension-reduction framework that maps higher-order networked dynamics onto an effective low-dimensional system, allowing accurate prediction of tipping boundaries, bistability regions, and the nature of phase transitions. We demonstrate the power of this framework across a range of dynamical processes, revealing distinct effects of higher-order interactions on transition continuity and hysteresis. Furthermore, we find that system resilience exhibits a profound dependence on the alignment between pairwise and higher-order connectivity, with assortative mixing enhancing tipping toward active states. Our findings establish a general theory for understanding the critical transitions in higher-order networks, offering new insights for anticipating and managing systemic risk in complex systems.2026-09-04T03:48:29ZJia-Jie QinJack Murdoch MooreXiaozhu ZhangGang Yanhttp://arxiv.org/abs/2609.04520v1Crowding controls the scaling of bus frequency with demand2026-09-03T22:18:56ZCities must allocate limited resources to maintain mobility, with uncertainties about the resulting state of the system. Analyzing roughly 3,000 bus routes with more than 4 billion yearly riders across 19 metropolitan areas worldwide, we uncover a robust scaling law of the form $f \sim (d/t)^α$ with exponent $α\in [1/2,\,2/3]$, linking the service frequency $f$ to passenger demand $d$ and route duration $t$. We show that this scaling emerges from a simple optimization principle: cities implicitly minimize total passenger waiting time under a fixed operational budget when both schedule frequency and crowding are taken into account. This mechanism produces two universal regimes: a frequency-dominated regime with $α= 1/2$ when crowding is negligible, and a capacity-dominated regime with $α= 2/3$ when most routes are overloaded. Intermediate exponents arise when only part of the network operates near capacity. Furthermore, we find that the benefits of additional investment are highly uneven across systems. For instance, our model suggests that a $20\%$ budget increase yields nearly a 5-minute reduction in daily waiting time per passenger in Boston, compared to only about 1 minute in Paris. These findings place urban transit within a broader class of constrained capacity-allocation problems, while highlighting a distinct regime in which prescribed route demands shape the allocation of limited service resources. The resulting scaling laws show how simple optimization principles can generate systematic exponents in complex transport systems, beyond the dissipation-based frameworks usually considered in physical and biological flow networks.2026-09-03T22:18:56ZProc. Natl. Acad. Sci. U.S.A. 123 (29) e2535998123 (2026)Siddharth PatwardhanŞirag ErkolFilippo RadicchiMarc Barthelemy10.1073/pnas.2535998123