https://arxiv.org/api/XGpZkAbbERs2fnZCstrJ4BcoP5E2026-07-21T08:30:33Z236656015http://arxiv.org/abs/2607.08546v1Rumour Spreading In Community Based Networks2026-07-09T14:40:35ZMany real-world networks have the characteristic that they are comprised of distinct groups or communities whose members contain many links within the community but with fewer connections to others. It is important to accurately model these types of networks to correctly predict the outcome of important spreading processes such as disease transmission, or the flow of information etc. Our motivating example is a network of traders within several investment institutions such as hedge funds. We assume an idealised scenario where traders within the same institution have many contacts and can share information quickly and easily but have fewer contacts to traders in other institutions, relying on personal networks, allowing for information to flow easily within a community and less-so between communities. In this paper we investigate a particular spreading process, the spread of a rumour, on a community based network that is characterised by two parameters; the within-group connectivity, and the between-group connectivity. We show that such networks have different characteristics to small-world or random networks that are often used to model the types of systems and that the network topology has a small but not insignificant effect on the spread of rumours on the network.2026-07-09T14:40:35Z11 pages, 8 figuresZhaoxi CuiAnthony O'Harehttp://arxiv.org/abs/2607.08520v1Elitism in the Aisle: A Long-Run Surname Measure of Legislative Elite Composition in Chile, 1834-20202026-07-09T14:15:54ZThe link between descriptive and substantive representation is well established in the literature but is hard to trace historically, where class records are thin. We introduce a replicable enduring-elite surname measure, pairing a contemporary socioeconomic criterion with historical elite registers, and apply it across the Chilean Congress, 1834-2020. Against a dynamic population reference built from 22.65 million birth registrations, the enduring-elite share of Congress falls from about half in the 1860s to about 12% in the 2010s, with a sharp drop of 11 to 13 points around the 1925 constitutional reform. In 1910-1950, composition co-moves with the legislative agenda, net of party: common-surname legislators emphasize labor foremost, elite legislators a statecraft agenda of defense, foreign affairs, and administration. Across this window, who sits in Congress moves together with what Congress attends to.2026-07-09T14:15:54ZMain text (4 figures, 1 table) plus online appendix; 56 pages totalNaim BroJuan Pablo Lunahttp://arxiv.org/abs/2607.08512v1The geopolitics of knowledge: tipping points, national fingerprints, and the unequal globalization of science2026-07-09T14:07:11ZScience is often portrayed as a universal and self-contained system, driven solely by the internal logic of knowledge accumulation and isolated from the turbulences of the socio-political world. In this paper, we challenge this narrative by providing systematic quantitative evidence that the global scientific ecosystem is deeply shaped by geopolitical transformations. Using a large-scale dataset of scientific publications drawn from the OpenAlex database, spanning over five decades and covering virtually all countries and disciplinary areas, we track the evolution of national research profiles and show that geopolitical dynamics shape scientific agendas at multiple scales. At the global level, intrinsic scientific change is slow and cumulative, but exogenous shocks, such as Chernobyl, September 11, and COVID-19, produce rapid disruptions that synchronously reconfigure the priorities of many countries at once. At the country level, we document a broad globalization of knowledge, yet deeply heterogeneous: while Global North countries converge toward a shared international agenda, Global South countries display strong dependence on international resources alongside locally distinctive research interests. Among emerging Southern economies, scientific power is increasingly asserted through specialized and independent agendas. Finally, we observe a reorganization of global scientific influence toward a more polycentric structure, with the emergence of a Southern cluster gravitating around Brazil and Indonesia as new regional hubs.2026-07-09T14:07:11ZIrina VorobevaMaxime LenormandGermana BerlantiniFloriana Gargiulohttp://arxiv.org/abs/2607.08374v1Large-Language-Models-as-a-Judge in Theory-Agnostic Adaptive Metric-Alignment for Prototypical Networks in Personality Recognition2026-07-09T11:49:14ZPersonality recognition has traditionally been constrained by theory-dependent formulations, where models are trained to fit predefined psychological taxonomies rather than uncovering shared underlying behavioral structure. This limits generalization, as personality itself is better understood as theory-invariant, while existing annotations reflect only partial and sometimes inconsistent views of the same latent traits. In this work, we introduce JAM ((J)udge for (A)daptive (M)etric-Alignment), a theory-agnostic framework that shifts learning from adapting to predefined personality theories toward discovering unified latent pseudo-facets that capture shared psychological structure. Rather than constraining the model to any personality taxonomy during training or inference, the framework learns generalizable psychological representations and can infer an individual's latent psychological profile directly from the textual samples, without requiring theory-specific labels. JAM achieves this through an Attention-Pooled Graph Prototypical Network that learns structured representations via clustering in embedding space, together with a Cross-Theory Harmonization (CTH) approach that integrates (i) Human-Guided Linkage and (ii) Machine-Induced Consensus to unify heterogeneous datasets without relying on predefined labels. To further improve robustness and data quality, we incorporate an LLM-as-a-Judge mechanism operating in two configurations, (i) LLM-before-the-loop and (ii) LLM-in-the-loop which identifies ambiguous samples to guide adaptive metric learning. Experiments show that JAM improves cross-framework generalization and performance, establishing a strong step toward theory-agnostic personality inference and supporting low-resource personality theories. The related code repository, model weights, and artifacts are available at https://research.jingjietan.com/JAM2026-07-09T11:49:14ZIEEE Transactions on Affective Computing (2026)Jing Jie TanBan-Hoe KwanDanny Wee-Kiat NgYan-Chai HumShih-Yu LoPo-An ChenNoriyuki KawarazakiKosuke TakanoAnissa Mokraoui10.1109/TAFFC.2026.3712379http://arxiv.org/abs/2606.28225v2Estimation-Prediction Tradeoff in Causal Probabilistic Temporal Graphs2026-07-08T23:17:43ZTemporal link prediction (TLP) is typically evaluated by predictive performance on unseen edges, but this criterion can conflate predictive accuracy with recovery of the underlying causal mechanism. In stochastic models, Fisher information governs the Cramér--Rao (CR) bound on parameter estimation error: higher Fisher information permits more accurate parameter recovery. We show that, under comonotonicity conditions between Fisher information and entropy, binary logistic models exhibit an estimation--prediction tradeoff: regimes with higher Fisher information, and hence smaller CR bounds, also have higher irreducible predictive entropy. To study this tradeoff in TLP, we introduce a probabilistic causal generator for temporal graphs with transient edges and known ground-truth causal structure, and validate the phenomenon empirically.2026-06-26T16:13:48Z8 pages, 2 figures (preliminary work)Aniq Ur Rahmanhttp://arxiv.org/abs/2512.02362v2Reconstructing Large Scale Production Networks2026-07-08T17:43:00ZFirm-to-firm production networks matter for aggregate propagation, but they are rarely observed. This paper reconstructs national-scale, weighted firm-to-firm networks from two public objects: a sectoral input--output table and the distribution of firm sizes by sector. The algorithm first draws a binary buyer-seller backbone from a sector-aware gravity model and then assigns weights by a minimum-energy program. A Markov closure makes the reconstructed network primitive, so it has a unique stationary distribution. The weighting program keeps one-step firm balances and sectoral flows close to the data; the stationary money vector is then checked ex post and remains close in aggregate. For the United States we reconstruct a network with about 6.5 million firms and 340 million links in roughly four hours on a single workstation. We also reconstruct the networks of Japan, the United Kingdom, Australia, Finland, and Denmark. The Japanese reconstruction, built without any link data, reproduces the heavy-tailed degree regime documented in the country's observed production network. The reconstructed networks exhibit customer tails heavier than supplier tails, though the algorithm treats the two sides symmetrically. We also run computational experiments on the reconstructed networks to assess the systemic risk posed by the failure of individual firms. These experiments show that neither firm size nor degree nor sectoral position is a good proxy for the aggregate losses generated by a firm's failure. For such questions, there is no good substitute for the complete weighted buyer-seller network that we reconstruct. We release the reconstruction code, the generated networks, a Python library, and a graphical2025-12-02T03:12:12ZAshwin BhattathiripadVipin P Veetilhttp://arxiv.org/abs/2601.18544v3The Cost of Inflation2026-07-08T17:33:22ZEmpirical evidence suggests that there is little to no correlation between the rate of inflation and the size of price change. Economists have hitherto taken this to mean that monetary shocks do not generate much deviation in relative prices and therefore inflation does not hurt the economy by impeding the workings of the price system. This paper presents a production network model of inflationary dynamics in which it is well possible for inflation to have near-zero correlation with the size of price change yet cause significant distortion of relative prices. The relative price distortion caused by inflation critically depends on the spectral gap, degree distribution, and assortativity of the production network.2026-01-26T14:49:11ZVipin P Veetilhttp://arxiv.org/abs/2607.07768v1Cascading Effects of the COVID-19 Pandemic on Barangays in the Philippines2026-07-08T16:11:06ZThe COVID-19 pandemic disrupted socio-economic and healthcare systems in the Philippines, significantly affecting barangays. This study analyzes the cascading effects of the COVID-19 pandemic on key aspects of a barangay, namely mobility, accessibility of public services, economic and financial health, food security, educational engagement, and physical health. It focuses on data from 2,122 Filipino households collected during May to June 2021 as part of the World Bank COVID-19 Households Survey. A Bayesian network model was constructed to programmatically map the conditional dependencies among these variables, utilizing Python libraries. Survey responses were grouped into common variables based on shared characteristics and standardized through z-score normalization to serve as nodes in the Bayesian network. By extending the Bayesian network into an influence diagram, the results will help identify interventions to guide local government units (LGUs) and policymakers in crafting tailored recovery programs and strategies that address impacts on physical health, economic and financial health, food security, public service access, mobility, and educational engagement. These efforts ultimately aim to enhance barangay resilience and preparedness for future public health crises. The results indicate that interventions aimed at boosting food production, stabilizing market prices, and expanding income opportunities are the most effective in improving community outcomes. This highlights the vital role of targeted economic and food security measures in mitigating the socio-economic impacts of the pandemic and offers valuable insights for shaping future response and recovery efforts.2026-07-08T16:11:06ZNaomi Ashley AmparoJohn Frederick MujiPaul James MontecilloJaymar SorianoVena Pearl Bongolanhttp://arxiv.org/abs/2607.07760v1Adversarial Social Epistemology for Assemblies of Humans and Large Language Models2026-07-08T15:09:49ZWe outline an adversarial social epistemology (ASE) for densely interactive communicative landscapes in which public assertions are scaffolded by chains of testimony, inference, institutional certification, and tacit trust. In such landscapes, agents have incentives and affordances to distort, color, omit, fabricate, or strategically under-specify information for private, reputational, rhetorical, or material gains. We argue that these phenomena are not adequately captured by familiar descriptions of epistemic bubbles, echo chambers, or misinformation diffusion. What requires explanation is how communicative agents exploit the commitments and entitlements that normally make scaffolded assertions trustworthy. We provide language that delivers the requisite analysis, outline mechanisms that subvert trust in scaffolded public communications, and outline machinery for auditing and redressing trust breaches arising from subverting the auditability of inferential chains, drawing on epistemic networks, enriched with an inferentialist semantics for interpreting assertions.2026-07-08T15:09:49Z50 pagesMihnea C. MoldoveanuJoel A. C. Baumhttp://arxiv.org/abs/2504.06318v4The Schwurbelarchiv: a German Language Telegram dataset for the Study of Conspiracy Theories2026-07-08T13:50:18ZSociality borne by language, as is the predominant digital trace on text-based social media platforms, harbours the raw material for exploring a multitude of social phenomena. Distinctively, the messaging service Telegram provides functionalities that allow for socially interactive as well as one-to-many communication. Our Telegram dataset contains over 5,800 groups and channels and 63 million messages, originating from a data-hoarding initiative named the ``Schwurbelarchiv'' (from German schwurbeln: speaking nonsense). Uniquely, it includes the transcriptions of over 3 million audio and video files. While the raw data was previously archived on the Internet Archive by an anonymous data hoarder, it was stored in a format that is difficult to process and largely inaccessible for systematic research. Our contribution consists of parsing, cleaning, and validating this raw archive, pseudonymising user data, and transcribing roughly 126,000 hours of audio and video content, thereby transforming this data hoard into a structured, research-ready dataset. This dataset publication details the structure, scope, and methodological specifics of the Schwurbelarchiv, emphasising its relevance for further research on the German-language conspiracy-theory-related discourse. We validate its predominantly German origin by linguistic and temporal markers and situate it within the context of similar datasets. We describe process and extent of the transcription of multimedia files. Thanks to this effort the dataset uniquely supports analysis of text from originally multimodal sources like voice messages and videos to investigate online social dynamics and content dissemination. Researchers can employ this resource to explore societal dynamics related to misinformation, political extremism, opinion adaptation, and social network structures.2025-04-08T09:11:46ZThis paper is 20 pages, 2 figures, 4 tables, and one datasetMathias AngermaierElisabeth HoeldrichJana LasserJoao Pinheiro Netohttp://arxiv.org/abs/2607.07399v1Forced condensation and anti-condensation on heavy-tailed networks2026-07-08T13:33:41ZWe study a driven selection mechanism on a fixed heavy-tailed network. At each step fresh mass is injected, its direction is recomputed from the current mass profile by a power-normalization rule, and the combined mass is transported by a primitive mixing matrix. The exponent $θ$ controls the feedback. Positive values give more weight to larger coordinates, while negative values favor smaller ones. When $θ=0$, the injected mass is distributed uniformly. After deterministic growth of the total mass is scaled out, the long-run injection profile is characterized by a nonlinear Perron-Frobenius fixed point on the simplex. Hilbert's projective metric gives a simple way to understand the stability of the system. The discounted network response brings positive profiles closer together, while the escort map scales their projective distance by $|θ|$. On heavy-tailed networks, this fixed point separates three effects that are often conflated: response or degree tilt, anomalous inverse-participation-ratio scaling, and genuine few-node localization. Positive feedback selects high-response nodes and, when response follows degree, a hub-directed branch. Negative feedback selects low-response nodes and typically produces a broad peripheral cloud unless the lower tail of the response field is itself thin. Numerical experiments on finite power-law networks support these results. They show convergence, illustrate when the forcing rate becomes unimportant because mixing is sufficiently fast, and confirm both the sign law and the crossover in the participation ratio. This mechanism is different from both conserved-mass condensation and graph growth. Instead, feedback selects a non-equilibrium profile on a fixed, heterogeneous network.2026-07-08T13:33:41Z30-page main article with 13 pages of supplementary material; 8 figuresAshwin BhattathiripadVipin P. Veetilhttp://arxiv.org/abs/2607.07387v1A Large Language Model-Driven Agent-Based Modeling Framework with Multi-Round Communication for Simulating Vaccine Opinion Dynamics2026-07-08T13:19:47ZRecently, Large Language Models (LLMs) have been utilized in various applications of computational social science and provide the possibility to integrate such models into agent-based modeling to explore the cognitive processes. However, how specific cognitive modules drive individual decisions and macro-level opinion dynamics remains unclear. Therefore, this study introduces a framework that integrates an LLM (Qwen3-8B) into agent-based modeling to investigate this problem, using vaccination opinion dynamics as a case study. We utilize this framework to simulate opinion dynamics among agents with heterogeneous profiles and social networks, evaluating scenarios by enabling different cognitive modules: a memory module and a prompt diversity module. The simulation results reveal that different cognitive modules have opposite impacts on our emergent opinion. Furthermore, the framework reproduces the non-linear behavior patterns of social influence observed in existing research, demonstrating our framework's validity and potential to reach the level 3 validation of agent-based models.2026-07-08T13:19:47Z11 pages, 5 figuresBo ZhangNa Jianghttp://arxiv.org/abs/2606.00893v2Hypergraph backboning2026-07-08T09:19:07ZHypergraphs provide a natural framework for describing complex networked systems with higher-order, non-dyadic interactions. Due to their high dimensionality and often redundant structure, a key challenge is to develop methods that simplify hypergraph representations while preserving the essential structure of interactions. Here we present a principled, efficient, and non-parametric information-theoretic method for pruning nested and/or redundant structures in hypergraphs, enabling a minimal representation of higher-order interactions in the presence of local heterogeneity. Our approach naturally extends to weighted hypergraphs, where higher-order topology and hyperedge weights combine to identify the system's structural backbone. We validate the method on controlled synthetic hypergraphs and apply it to empirical datasets from diverse domains, demonstrating substantial sparsification without loss of core structural information.2026-05-30T20:57:05ZAlec KirkleyHelcio FelippeFederico MaliziaFederico Battistonhttp://arxiv.org/abs/2607.06984v1Modeling Misinformation as a Commons Problem2026-07-08T04:11:48ZMisinformation often harms society not just by spreading a single false belief, but by breaking down the shared trust people rely on to evaluate what is true. This paper presents an agent-based simulation that frames trust as a collective resource and attention as a scarce private budget: when aggregate attention shifts toward low credibility content, the trust environment degrades, making credible information harder to process and correct. Across experiments, the model produces four recurring modes: credible stability, misinformation dominance, polarization, and a mixed baseline, with distinct signatures in trust trajectories and network structure. The results separate two control problems that matter for simulation-based policy exploration: the balance of trust repair versus harm largely determines whether the system recovers or collapses, while homophily and rewiring determine whether disagreement remains integrated or separates into persistent clusters. This foundation provides a transparent testbed for comparative experiments on interventions that must address both trust restoration and structural conditions for cross-cutting exposure.2026-07-08T04:11:48Z13 pages, Accepted at the Annual Modeling and Simulation Conference 2026Proc. of the 2026 Annual Modeling and Simulation Conference (ANNSIM'26)Vrinda Malhotrahttp://arxiv.org/abs/2607.06528v1Trust-Aware Citation Cartel Ranking in Scholarly Knowledge Graphs2026-07-07T17:33:12ZCitation-based systems usually treat each citation as an equal signal of scholarly influence, although citations can express very different relationships: direct method use, result comparison, broad background, or weak ceremonial acknowledgement. This distinction is crucial for citation-cartel analysis because dense internal citation alone is not suspicious; legitimate research communities are also densely connected. We present a trust-aware pipeline that combines citation graph structure with semantic citation intent to rank suspicious paper-level communities for audit. On a DBLP-derived graph with 500,000 papers and 4.87M citation edges, we use an LLM teacher to label 205,897 citation pairs, train a SciBERT student, and scale citation-intent typing to 2.04M unique graph edges. We then compute a Composite Cartel Index (CCI) that integrates internal density, citation inflation, reciprocity, semantic superficiality, degree assortativity, and trust-weighted PageRank shift. The highest-ranked community contains 1,079 papers and 8,603 internal citations, with 254.3x more internal citations than expected and 64.2% of them superficial. Comparisons against density-only, inflation-only, semantic-only, and random baselines show that CCI cannot be reduced to a single heuristic. Edge excision validation further shows that CCI-selected communities behave differently from matched random removals. The result is a reproducible, curator-facing ranking framework for prioritising communities that warrant closer inspection.2026-07-07T17:33:12ZPratyush GuptaVikranth UdandaraoSyam Sai Santosh Bandi