https://arxiv.org/api/GdwZvjXWNsCW7TYn3ai87LP6yC42026-07-20T04:37:16Z236583015http://arxiv.org/abs/2607.12086v1CityBehavEx: A Scalable and Empirically Validated LLM-Assisted Urban Simulation Platform2026-07-13T19:03:25ZRecent LLM-based multi-agent urban simulators can generate semantically rich city routines, but they remain costly to scale and are often weakly validated against empirical mobility patterns. We present CityBehavEx, an interactive LLM-assisted urban simulation platform that scales to city-size populations, exposes agent behavior for inspection, supports empirical validation, and generates mobility patterns that better match real-world spatial, temporal, and semantic distributions. Instead of invoking large language models for every agent action, CityBehavEx combines established human mobility models with fine-tuned cross-encoders that estimate semantic alignment between agent profiles, schedules, and activity transitions. This design enables large-scale simulations, as demonstrated in a case study of 100,000 agents over 75 days in under one hour on a single consumer GPU. The platform allows users to define simulation regions, launch experiments, inspect trajectories and activity traces, debug unrealistic behaviors, and validate generated routines against real-world mobility, time-use, and semantic metrics.2026-07-13T19:03:25Z10 pages, 3 figuresGustavo H. SantosAline VianaThiago H Silvahttp://arxiv.org/abs/2411.18614v2Optimal root recovery for uniform attachment trees and $d$-regular growing trees2026-07-13T18:02:06ZWe consider root-finding algorithms for random rooted trees grown by uniform attachment. Given an unlabeled copy of the tree and a target accuracy $\varepsilon > 0$, such an algorithm outputs a set of nodes that contains the root with probability at least $1 - \varepsilon$. We focus on the algorithm introduced by Bubeck, Devroye and Lugosi (2017) and proved to be optimal by Crane and Xu (2021). We prove that, for the optimal algorithm, an output set of size $\exp(O(\log^{1/2}(1/\varepsilon)))$ suffices; this bound is sharp and answers a question of Bubeck, Devroye and Lugosi (2017). We prove similar bounds for random regular trees that grow by uniform attachment, strengthening a result of Khim and Loh (2017).2024-11-27T18:57:40Z31 pages. Annals of Applied Probability, to appearLouigi Addario-BerryCatherine FontaineRobin KhanfirLouis-Roy LangevinSimone Têtuhttp://arxiv.org/abs/2607.11803v1"We are all in big trouble! *Shock Emoji": Personal Narratives in Expressing Emotions, Opinions, and Data Regarding Climate Change in TikTok Short Videos2026-07-13T16:54:22ZClimate change is a source of anxiety about the future. Understanding how people express themselves about climate change enables us to address such concerns. To study climate change expression on social media, we analyzed 200 TikTok videos tagged with #climatechange, identifying four categories of content: expression-feelings, views-appeals, news-information, and trend-hijacking. We found that creators use humor to package sharp critiques, avoiding direct confrontation. They replace complex discussions with life stories, such as adopting a vegetarian lifestyle or deleting emails. They borrow from news media to present fragmented information as scientific interpretations, creating a perception of scientific credibility, balancing scientific accuracy with emotionality. Analysis of viewer responses showed they engaged empathetically, reshaping interpretations of videos. These interactions risk reinforcing existing views but help build community on TikTok, which lacks community structure. This study reveals how creators may retell news on science using personal narratives, highlighting how short-form videos enable climate communication.2026-07-13T16:54:22Z44 pages, 33 figures, presenting this paper at CSCW 2026Chu ZhangSimai HuangShaohua WuYihuan ChenRay LC10.1145/3817034http://arxiv.org/abs/2607.11632v1Reproducing human biases in route choice using large language models: Toward scalable behavioral modeling2026-07-13T14:48:23ZHuman choice behavior, including route choice, exhibits systematic behavioral biases that deviate from the assumptions of full rationality. Cumulative prospect theory (CPT) has been widely recognized as an effective framework for characterizing such behavioral patterns. However, its large-scale application, particularly in simulation and agent-based modeling, critically depends on specifying individual-level CPT parameters, which remain a major bottleneck. Conventional approaches typically rely on surveys and controlled experiments to calibrate CPT parameters, yet these methods are difficult to generalize and often fail to capture the full diversity of human decision-making. To address this challenge, this paper investigates whether large language models (LLMs) can reproduce human behavioral biases in choice-making without explicit specification of prospect-theoretic parameters. Using route choice as a representative scenario, we design a behavioral evaluation framework and systematically compare LLM-generated decisions with established human behavioral patterns predicted by CPT. Experimental results demonstrate that LLMs are capable of reproducing non-rational human choice biases and can exhibit decision behaviors consistent with prospect-theoretic effects under uncertainty. These findings suggest that generative AI models may provide a scalable alternative for modeling human decision processes and offer a promising foundation for next-generation large-scale agent-based simulation and AI-driven behavioral research.2026-07-13T14:48:23ZJiangtao HanShoufeng MaShuxian XuGeng LiShuai LingNing JiaZhengbing Hehttp://arxiv.org/abs/2506.21426v2Evolution and determinants of firm-level systemic risk in local production networks2026-07-13T14:39:16ZRecent crises like the Covid-19 pandemic and geopolitical tensions have exposed vulnerabilities and caused disruptions of supply chains, leading to product shortages, increased costs, and economic instability. This has prompted growing efforts to assess systemic risk, namely the effects of firm disruptions on entire economies. However, the ability of firms to react to crises by rewiring their supply links has been largely overlooked, limiting our understanding of production networks resilience. Here, we study dynamics and determinants of firm-level systemic risk in the Hungarian economy from 2015 to 2022. We benchmark our results to a heuristic maximum entropy null model that generates randomized production networks while preserving the total input (demand) and output (supply) of each firm at the sector level. We show that the fairly stable set of firms with highest systemic risk undergoes a structural change during Covid-19, as those enabling economic exchanges become key players in the economy -- a pattern not reproduced by the null model. Although empirical systemic risk closely matches the null value prior to the pandemic, it becomes significantly lower afterwards, reflecting the emergence of a more resilient economy driven by firms' adaptive behavior. Furthermore, firms' international trade volume (being itself a channel of potential disruption) becomes a significant predictor of their systemic risk. However, international linkages alone cannot fully explain the observed trends, as imports and exports exert opposing effects on local systemic risk through the supply and demand channels.2025-06-26T16:08:22Z15 pages, 4 figuresPNAS Nexus, pgag233 (2026)Anna ManciniBalázs LengyelRiccardo Di ClementeGiulio Cimini10.1093/pnasnexus/pgag233http://arxiv.org/abs/2607.11379v1Hyperbolic embeddings for graph compression2026-07-13T10:41:48ZNetwork theoreticians hypothesize that the structure of real-world networks has a geometric origin. Especially, hyperbolic geometry was proven insightful in representing and modeling of scale-free networks. Embedders are algorithms used to find a geometric representation of a network. In this study, we introduce a fast lossless graph compression algorithm based on modern hyperbolic embedders. Experimental validation on real-world and generated networks shows that our algorithm beats state-of-the-art by up to 42% on real-world graphs.2026-07-13T10:41:48ZDorota Celinska-KopczynskaEryk Kopczynskihttp://arxiv.org/abs/2303.11452v3A Cheeger Inequality for Size-Specific Conductance2026-07-13T03:49:53ZThe $μ$-conductance measure proposed by Lovász and Simonovits is a size-specific conductance score that identifies the set with smallest conductance while disregarding those sets with volume smaller than a $μ$ fraction of the whole graph. Using $μ$-conductance enables us to study the network structures in new ways. In this manuscript we study a modified spectral cut for $μ$-conductance that is a natural relaxation of the integer program of $μ$-conductance and show that the optimum of this program has a two-sided Cheeger inequality with $μ$-conductance.2023-03-20T21:00:15ZAccepted by Discrete MathematicsYufan HuangDavid F. Gleichhttp://arxiv.org/abs/2605.21510v2Community-Aware Vertex Ordering for Reference-Based Graph Compression: A Cross-Encoder Empirical Study2026-07-13T02:28:54ZReference-based graph compression encodes each vertex's neighbor list as differences from a nearby encoded list. WebGraph's BVGraph fixes a single encoding pipeline and relies on a separately chosen vertex ordering -- typically URL-lexicographic or Layered Label Propagation (LLP). Their interaction is rarely measured. We propose a two-stage Leiden+LLP ordering: global LLP seeds labels, Leiden detects communities, and a final LLP pass reorders each community internally. We study how it interacts with reference-based compression, using BVGraph and three encoders we contribute -- BG, CS, and CG -- each picking, per vertex, the cheapest of up to 28 candidate decompositions. On graphs with poor initial vertex order, reordering with Leiden+LLP improves compression for every encoder measured, saving 0.9 to 4.6 bits per edge (bpe) over the original order on SNAP-style graphs delivered in vertex-ID order. The gain barely depends on the encoder: on four of five weakly ordered datasets, the four encoders agree on the Leiden+LLP-vs-plain-LLP gain within about +/- 0.04 bpe. On URL-ordered web crawls, where the ordering already encodes locality, BG and CS still benefit, while residual-sensitive configurations (BVGraph default, BV-HC, and CG relying on community contiguity) regress. The transfer holds across two encoder generations (Fibonacci-coded and fully context-adaptive range-coded) and three ordering seeds. With every structural bit entropy-coded, the best of our three encoders beats the strongest published baseline in each regime (Zuckerli, and the stronger of BV-HC / BVGraph default) on all seven datasets in every whole-graph and random-access comparison -- 28/28 cells, +0.3 to +35% over Zuckerli -- with the encoder-level gain consistently smaller than the ordering-level gain. All algorithms, the ordering pipeline, and generators are released as the Adjacently Julia library.2026-05-13T10:38:31Z33 pages, 4 figures, 13 tables. Open-source implementation and reproduction drivers at https://github.com/jimbotonic/Adjacently.jl. Preprint; comments welcomeJimmy Dubuissonhttp://arxiv.org/abs/2607.10780v1Return of the solo author: The changing division of labor in science in the age of generative AI2026-07-12T14:17:39ZModern science has experienced a long shift from individual work to team production. Generative artificial intelligence (AI) might appear to extend this trajectory by lowering research costs and enabling larger-scale collaboration. Yet if tasks once performed by coauthors can be delegated to AI, the same technology may also weaken the need for collaboration in parts of the research process. Here, we examine this tension by moving beyond average team size and focusing on the solo-authored tail of the author-count distribution. Analyzing over 300 million works across 26 fields, we find that the decades-long decline in solo authorship halted and partially reversed with ChatGPT's public release in late 2022. We also reveal that this is an uneven phenomenon: it is strongest in fields where coauthors' work is more readily replaceable, and weak or absent in fields that depend on physical collaboration. At the individual level, the recovery is not explained by the entry of new researchers or by changes in field composition. Instead, the break appears among authors who had written only with others, including those with no prior solo publications, and among long-established authors as well as newcomers. Their solo papers stay close to their own coauthored work while narrowing in scope and shifting toward computational topics. Because a solo paper is work without credited human coauthors, this study offers an empirical probe of how generative AI can substitute for scientific labor, and evidence of a reconfiguration of cognitive labor within papers rather than of team size.2026-07-12T14:17:39Z37 pages, 12 figuresAkira Matsuihttp://arxiv.org/abs/2607.10645v1MafiaScope: Non-Invasive, Time-Resolved Belief Probing for LLM Agents in Social Deduction Games2026-07-12T08:20:52ZAn LLM agent's public behaviour reveals little about its social reasoning: an agent that votes correctly may be guessing, and an agent that lies well leaves no trace of what it actually believes. We present MafiaScope, an open testbed that turns the social deduction game Mafia into a measurement instrument for machine Theory of Mind. After every public utterance, every agent privately answers a configurable set of structured probe questions; the answers never re-enter the game and are scored automatically against the ground truth the engine knows. An interactive visualizer renders the belief trajectories: impersonate mode shows the game as one agent sees it, panels chart timeline-aligned accuracy and calibration, and counterfactual replay forks any recorded step. In a 32-game DeepSeek case study with 13{,}815 parsed probe answers, stated confidence is poorly calibrated, with expected calibration error 0.17, agents over-predict being suspected 1.5 times, and a 30-fork replay experiment walks the counterfactual replay workflow end to end. Engine, viewer and a corpus of 200+ cross-model games are released under an open licence; live demo: https://karpovilia.github.io/mafiascope/; screencast: https://vimeo.com/1208920221.2026-07-12T08:20:52ZIlia Karpovhttp://arxiv.org/abs/2607.10402v1Large Language Models in Misinformation Ecosystems: Misuse, Defense, and Vulnerability2026-07-11T17:04:17ZLarge language models (LLMs) have transformed misinformation from a primarily content-centric problem into a broader ecosystem-level security challenge. When misused, LLMs create risks beyond false content generation, enabling attacks on the social contexts, evidence sources, retrieval corpora, and verification workflows that misinformation defense depends on. In this paper, we introduce a role-layer framework to unify these risks and defenses. The role dimension characterizes LLMs as attackers, defenders, and vulnerable components of verification systems, while the layer dimension covers content, social contexts, evidence environments, and verification workflows. Guided by this framework, we organize LLM-enabled attacks, investigate LLM-based detection and verification methods, analyze vulnerabilities in LLM-centric detection paradigms, and discuss existing countermeasures against LLM-enabled attacks. Building on this synthesis, we identify three key open challenges: moving from static detection accuracy to budgeted ecosystem-level risk evaluation, hardening LLM-centered verification pipelines against adversarial manipulation, and deploying auditable human-in-the-loop verification systems for trustworthy real-world misinformation defense.2026-07-11T17:04:17Z35 pages, 8 figuresLingwei WeiDou HuWei ZhouSonglin HuPhilip S. Yuhttp://arxiv.org/abs/2607.09990v1Political Power in International Trade2026-07-10T21:34:13ZEconomic power in international trade is the capacity of one country to impose loss on another by withdrawing from a trading relationship. This paper measures it. A model of the short run represents each trade restriction as a pattern of barred entries in the world matrix of input shares and maps it into a vector of losses by country and sector. The asymmetry between the two countries' losses under the same severance is the measure of power: a gap in substitution, since a buyer's dependence turns on how easily it finds another source and a seller's on how easily it finds another market. When a relationship is barred, buyers lean on alternative suppliers and barred suppliers on alternative buyers already present in the benchmark network, under the ceiling that no producer exceeds its pre-shock scale. The reallocation is a RAS balancing of the disrupted matrix that lets \emph{both sides} adjust together. Across 9,480 counterfactual severances on the 2022 world input--output network, mutual trade dependence is anything but mutual. The average bilateral asymmetry is 0.6 on a scale that runs from balance at zero to complete lopsidedness at one. The United States holds the favorable side in all its relationships, China in all but one. The same tilt runs far down the hierarchy: a severance with Russia would cost Belarus more than a tenth of its economic activity, but Russia only half of one percent. The asymmetry bears only a weak relation to bilateral trade imbalance but closely tracks whether a country sits at the core or the periphery. Power is a property of network position, not deficits.2026-07-10T21:34:13ZAshwin BhattathiripadVipin P Veetilhttp://arxiv.org/abs/2607.09528v1TSAI-MetaFraud: A Benchmark Dataset for Financial Fraud Transaction and Behavioral Risk Detection in Metaverse Ecosystems2026-07-10T15:36:13ZThe emergence of metaverse platforms has created virtual economies that introduce new challenges related to fraud, bot activity, and illicit financial behavior. Despite growing interest in trustworthy metaverse analytics, existing datasets typically focus on user behavior, authentication, or financial transactions in isolation, limiting the development and reproducible evaluation of multimodal fraud detection methods. To address this gap, we present TSAI-MetaFraud, a multimodal, multi-task benchmark dataset for fraud analytics in virtual economies. TSAI-MetaFraud integrates behavioral, transactional, and graph-structured information while incorporating realistic fraud and automated bot scenarios. We define benchmark tasks including transaction fraud detection, cross-modal node classification, temporal link prediction, and weakly supervised fraud detection, and provide baseline evaluations using machine learning models and graph neural networks. By jointly capturing behavioral activity, financial interactions, and relational structure within a unified virtual economy, TSAI-MetaFraud provides a benchmark for advancing multimodal learning, graph mining, fraud analytics, and trustworthy AI in emerging metaverse ecosystems.2026-07-10T15:36:13ZRefat Ishrak HemelEhsan HallajiRoozbeh Razavi-Farhttp://arxiv.org/abs/2505.24631v2Cascades on Networks with Functional Structure2026-07-10T11:18:13ZWe consider a version of the Watts threshold model on directed multiplex configuration model networks, and present a detailed analysis of the cascade size, single-seed cascade probability and cascade condition. We then introduce a smaller class of network models that we call "constrained multiplex networks", which is designed to represent networks with so-called "functional" or "complementary" structure. We find that the particular choice of functional structure affects the phase transitions of the cascade model in a variety of ways.2025-05-30T14:24:56ZChristian KlugeChristian Kuehnhttp://arxiv.org/abs/2502.18497v2LD-Leiden: Local Parallel Community Detection in Large Dynamic Networks2026-07-10T09:14:24ZDynamic community detection must update high-quality modularity partitions after edge batches, yet full Leiden reruns make small changes scale with the whole snapshot. Existing dynamic methods reduce work but often alter Leiden refinement, keep limited hierarchy state, or restrict graph support. This paper presents LD-Leiden, a local dynamic Leiden method for weighted directed and undirected graphs that preserves the move-refine-aggregate pipeline and updates only repaired affected regions. Its novelty is the combination of an affected-frontier rule after statistic repair, exact subtract-add aggregate repair, and conflict-filtered parallel local moves; together these mechanisms bound update cost by the visited frontier rather than the full graph. On real streams and streamed static graphs with up to 214M vertices and 3.30B edges, LD-Leiden is 48.77x faster than warm-started Leidenalg in 100-batch runs while preserving a 0.996 final modularity ratio. On the common undirected benchmark set, it is 6.94x faster than DF-Leiden and 9.73x faster than NetworKit while obtaining higher final modularity; synthetic sequences support the predicted local edge-volume scaling.2025-02-20T09:53:13Z10 pages, 5 figuresGrigoriy BokovAleksandr KonovalovAnna UporovaStanislav MoiseevIvan SafonovAlexander Radionov