https://arxiv.org/api/2qw/Xr8x5p1/ZBdEfvp4eRMIQIk2026-07-20T22:42:49Z24643015http://arxiv.org/abs/2605.09704v1Coexistence of trapped and flow-transported nuclei enables fast pigeon post communication across multinucleated cell2026-05-10T18:57:36ZMulti-nucleated cells exist in all domains of life, ranging from animals, plants and fungi to single-celled organisms such as the slime mold Physarum polycephalum. The large cell size, in the case of Physarum reaching centimeters and more, challenges the coordination of nuclei activity as signals need to cross large distances. In search for a mechanism for fast long-ranged communication among nuclei, we quantify nuclei dynamics and cytoplasmic flows in Physarum's tubular network. We observe nuclei in two interchangeable, dynamic states: mobile, flowing within the cytoplasmic shuttle flow, or trapped in the tube's porous cell cortex. As we find nuclei to accumulate at the tube's inner fluid-porous interface we theoretically explore and confirm, with physiological parameters, that slowing down of mobile nuclei during flow is sufficient for diffusible signal exchange between mobile and trapped nuclei. We analytically derive that communication akin to pigeon-post with mobile nuclei serving as pigeons shuttling between trapped nuclei acting as waypoints, gives rise to signaling velocities that account for the rapid intracellular reorganization observed in Physarum. Since signal transfer by flow-transported nuclei outcompetes the mere diffusion of signals encoded in cytosolic proteins, pigeon-post communication surpasses alternative signaling mechanisms, even diffusive relay signaling up to twenty-fold in velocity. The key ingredients of pigeon-post communication, namely alternating flows and waypoints, exist in other multi-nucleated cells and may also be generalized beyond intracellular signaling.2026-05-10T18:57:36ZProc. Natl. Acad. Sci. U.S.A. 122 (50) e2411101122 (2025)Johnny TongKaspar WachingerFabian K. HennNico SchrammaSiyu ChenKaren Alim10.1073/pnas.2411101122http://arxiv.org/abs/2605.07007v1Essential Role of Extrinsic Noise in Models of E. coli Division Control2026-05-07T22:42:35ZOur understanding of cell division control in bacteria still relies largely on interpreting correlations between phenomenological variables, with limited connection to the underlying molecular mechanisms.
Here, we analytically solve a stochastic threshold-accumulation model in which a size-dependent divisor protein triggers division upon reaching a noisy, autocorrelated threshold, quantifying within a unified framework the combined effects of intrinsic and extrinsic noise and key mechanistic parameters such as protein reset and threshold memory. We show that incorporating these elements yields behavior far richer than the commonly assumed adder, spanning a continuum of division strategies from timer to sizer while modulating size fluctuations in a nontrivial fashion. Comparison with single-cell E. coli data shows that extrinsic noise and additional mechanistic ingredients are required to account for the observed size fluctuations. The adder emerges when threshold correlations balance protein reset, generalizing the hypothesis that full reset is necessary to maintain adder control.
Our results establish a unified analytical framework linking stochastic molecular processes to emergent division laws, to be used in more complex bacterial cell-cycle models.2026-05-07T22:42:35Z20 pages, 5 figures, 1 tableMattia CoriglianoIFOM-ETS, The AIRC Institute of Molecular Oncology, Milan, ItalyDepartment of Physics "Aldo Pontremoli", University of Milan, Milan, ItalyKuheli BiswasDepartment of Physics of Complex Systems, Weizmann Institute of Science, Rehovot, IsraelMatteo BocchiolaDepartment of Physics "Aldo Pontremoli", University of Milan, Milan, ItalyDaniele MontagnaniDepartment of Physics "Aldo Pontremoli", University of Milan, Milan, ItalyAriel AmirDepartment of Physics of Complex Systems, Weizmann Institute of Science, Rehovot, IsraelMarco Cosentino LagomarsinoIFOM-ETS, The AIRC Institute of Molecular Oncology, Milan, ItalyDepartment of Physics "Aldo Pontremoli", University of Milan, Milan, ItalyIstituto Nazionale di Fisica Nucleare, Sezione di Milano, Milan, Italyhttp://arxiv.org/abs/2605.06598v1Mathematical Modeling of Early Embryonic Cell Cycles of Drosophila melanogaster2026-05-07T17:22:18ZIn the early stages of development, Drosophila melanogaster embryos possess very fast and well-coordinated cell cycles. In the cell cycle, CDK activity is essentially regulated by binding CDK and CycB to form an active complex and by phosphorylating CDK via CDC25 and dephosphorylating it via Wee1. We develop a mathematical model for the embryonic cell cycle which is biochemically sound and which can be rigorously analysed after a model reduction. We show that there exists a region in the parameter space where the model describes oscillations. We then focus on the role of two parameters: the CycB synthesis and the activation coefficient of APC. Our main biological hypothesis is that the first one is responsible for the period lengthening over the first 14 cycles which can be experimentally observed and this hypothesis is supported by numerical simulations of our model: if the CycB synthesis is made time-dependent with a prescribed dynamics, then our simulations show qualitatively a very similar behavior to experimental data reported in the literature.2026-05-07T17:22:18Z20 pages, 7 figuresMeskerem Abebaw MebratieBenedikt DrebesKatja KappArno MüllerWerner M. Seilerhttp://arxiv.org/abs/2605.06728v1OmicsLM: A Multimodal Large Language Model for Multi-Sample Omics Reasoning2026-05-07T11:27:11ZInterpreting transcriptomic data is one of the most common analytical tasks in modern biology. Yet most current models either consume expression profiles without producing natural-language biological explanations, or reason in language without direct access to quantitative omics measurements. We introduce OmicsLM, a multimodal LLM that connects quantitative omics profiles with natural-language biological tasks. OmicsLM represents each transcriptomic profile as a compact continuous representation within the LLM context. This interface preserves quantitative expression signal while allowing natural-language instructions, explicit gene mentions, and multiple interleaved biological samples to be processed together in one model context. We train OmicsLM on more than 5.5 million instruction-following examples spanning over 70 task types, combining continuous transcriptomic inputs, experimental data rendered through diverse language templates, and free-text biological knowledge and question-answering data. This mixture covers cell type annotation, perturbation prediction, clinical prediction, pathway reasoning, and open-ended biological question answering. Existing benchmarks evaluate either profile-level prediction or text-only biological QA, leaving language-guided, multi-sample reasoning over real expression profiles unmeasured. To close this gap, we introduce GEO-OmicsQA, a benchmark for multi-sample biological question answering built from real Gene Expression Omnibus (GEO) studies. We demonstrate that OmicsLM can use expression profiles directly and perform comparably to specialized omics models on profile-level tasks, while outperforming both omics-specialized models and general LLMs on language-guided biological reasoning over expression data.2026-05-07T11:27:11Z13 pages (main text), 14 pages (appendix), 1 figure, 10 tablesMaciej SypetkowskiJoanna KrawczykŁukasz SmolińskiRemigiusz KinasPrzemysław PietrzakTomasz JetkaRafał Powalskihttp://arxiv.org/abs/2605.04762v1TCRTransBench: A Comprehensive Benchmark for Bidirectional TCR-Peptide Sequence Generation2026-05-06T11:09:19ZT-cell receptor (TCR) interactions with antigenic peptides underpin adaptive immunity and are pivotal for personalized immunotherapy and vaccine development. Despite recent progress, computational modeling of TCR-peptide specificity remains challenging due to data scarcity, complex sequence dependencies, and the absence of standardized evaluation frameworks. To systematically address these issues, we introduce TCRTransBench, a comprehensive benchmark for bidirectional TCR-peptide sequence generation tasks. Specifically, we define two sequence-to-sequence (seq2seq) tasks: generating antigenic peptides from TCR sequences (TCR2PEP) and generating TCR sequences from antigenic peptides (PEP2TCR). Our framework provides a rigorously curated, MHC-free dataset comprising tens of thousands of validated TCR-peptide pairs, along with diverse evaluation metrics that integrate computational efficiency, sequence accuracy, and biological plausibility. Extensive benchmarking across representative neural architectures, including recurrent, convolutional, and transformer-based models, reveals key trade-offs among performance metrics, highlighting the effectiveness of transformers in capturing intricate biological interactions and the necessity of biologically informed evaluation criteria. TCRTransBench establishes standardized tasks, datasets, and evaluation protocols, laying a robust foundation for future computational advances in immunological sequence modeling and therapeutic protein design.2026-05-06T11:09:19Z13 pages, 5 figures, 2 tablesYiming WangWeiyu XiaoJiangbin ZhengStan Z. Lihttp://arxiv.org/abs/2605.03632v1Robust chemotaxis beyond sensing limits: signal, noise, and strategy2026-05-05T11:01:06ZBacterial chemotaxis has long been viewed as operating near the physical limits of sensing, as originally articulated by Berg and Purcell. Recent information-theoretic analyses challenge this view, suggesting that Escherichia coli uses only a small fraction of the information available in ligand arrival statistics to bias its motion. How should such low information efficiency be interpreted at the level of behavior? Here, I argue that chemotactic performance is shaped not only by information transmission and noise, but by the strategy of movement itself. Using simple scaling arguments and minimal models, I show how run-and-tumble chemotaxis can remain robust to noise through symmetry and temporal averaging, even when internal information processing is inefficient. Comparing bacterial and eukaryotic chemotaxis highlights how different sensing strategies convert physical limits into observable behavior. These considerations suggest that low information efficiency need not imply poor performance, but may instead reflect an evolved balance between robustness, simplicity, and function.2026-05-05T11:01:06Z9 pages, 3 figures, 2 appendices. This is the accepted manuscript version; the final published version will appear in Physical Biology (IOP Publishing) and will be available under a CC BY licenseRobert G. Endreshttp://arxiv.org/abs/2511.18883v2Enumeration of Autocatalytic Subsystems in Large Chemical Reaction Networks2026-05-05T07:30:24ZAutocatalysis is an important feature of metabolic networks, contributing crucially to the self-maintenance of organisms. Autocatalytic subsystems of chemical reaction networks (CRNs) are characterized in terms of algebraic conditions on submatrices of the stoichiometric matrix. Here, we derive sufficient conditions for subgraphs supporting irreducible autocatalytic systems in the bipartite Kőnig representation of the CRN. On this basis, we develop an efficient algorithm to enumerate autocatalytic subnetworks and, as a special case, autocatalytic cores, i.e., minimal autocatalytic subnetworks, in full-size metabolic networks. The same algorithmic approach can also be used to determine autocatalytic cores only. As a showcase application, we provide a complete analysis of autocatalysis in the core metabolism of E. coli and enumerate irreducible autocatalytic subsystems of limited size in full-fledged metabolic networks of E. coli, human erythrocytes, and Methanosarcina barkeri (Archea). The mathematical and algorithmic results are accompanied by software enabling the routine analysis of autocatalysis in large CRNs.2025-11-24T08:40:05Z54 Pages (35 main + 19 Supplementary Information), 16 figuresJournal of Chemical Theory and Computation, 2026Richard GolnikThomas GatterPeter F. StadlerNicola Vassena10.1021/acs.jctc.5c01979http://arxiv.org/abs/2406.02522v4Weaving Life into Regolith: Engineered Autotrophic-Heterotrophic Consortia for Autonomous Biofabrication from Granular Feedstocks2026-04-30T18:08:50ZLong-duration human missions to Mars will require autonomous systems capable of converting in situ resources into structural materials, tools, and functional components. More broadly, such systems represent a class of resource-limited bioprocesses relevant to extreme-environment manufacturing. Here, we investigate engineered autotrophic-heterotrophic consortia, inspired by lichen biology, as a platform for autonomous biofabrication from granular feedstocks. We experimentally screened filamentous fungi and paired them with diazotrophic cyanobacteria to identify mutually supportive consortia capable of sustained growth and biomineral production in the presence of Martian regolith simulant as the primary inorganic substrate, without external organic carbon or nitrogen inputs. Selected co-cultures exhibited evidence of metabolic coupling, and untargeted metabolomic analysis revealed coordinated reprogramming consistent with integrated carbon and nitrogen metabolism within the consortia. These systems facilitated mineral consolidation of regolith particles, demonstrating the feasibility of near-closed-loop biomineral production under resource-limited conditions. While integration with additive manufacturing remains conceptual, this study establishes a framework for engineering self-sustaining microbial consortia for biomaterials production and highlights opportunities for coupling metabolism with material synthesis in both extraterrestrial and terrestrial environments.2024-06-04T17:41:25ZNisha RokayaErin C. CarrKumar ShresthaRichard A. WilsonYong HuangCongrui Jinhttp://arxiv.org/abs/2306.10407v3FP-IRL: Fokker--Planck Inverse Reinforcement Learning -- A Physics-Constrained Approach to Markov Decision Processes2026-04-30T16:19:41ZInverse reinforcement learning (IRL) is a powerful paradigm for uncovering the incentive structure that drives agent behavior, by inferring an unknown reward function from observed trajectories within a Markov decision process (MDP). However, most existing IRL methods require access to the transition function, either prescribed or estimated \textit{a priori}, which poses significant challenges when the underlying dynamics are unknown, unobservable, or not easily sampled.
We propose Fokker--Planck inverse reinforcement learning (FP-IRL), a novel physics-constrained IRL framework tailored for systems that can be described by Fokker--Planck (FP) dynamics. FP-IRL simultaneously infers both the reward and transition functions directly from trajectory data, without requiring access to sampled transitions. Our method leverages a correspondence between MDPs and the FP equation, linking reward maximization in MDPs with free energy minimization in FP dynamics. This connection enables inference of the FP potential function using our inference approach of variational system identification, from which the full set of MDP components -- reward, transition, and policy -- can be recovered using analytic expressions.
We demonstrate the effectiveness of FP-IRL through experiments on synthetic benchmarks and a modified version of the Mountain Car problem. Our results show that FP-IRL achieves accurate recovery of agent incentives while preserving computational efficiency and physical interpretability.2023-06-17T18:28:03ZComputer Methods in Applied Mechanics and Engineering, 458, 119010 (2026)Chengyang HuangSiddhartha SrivastavaKenneth K. Y. HoKathy E. LukerGary D. LukerXun HuanKrishna Garikipati10.1016/j.cma.2026.119010http://arxiv.org/abs/2604.27646v1Benchmarking virtual cell models for in-the-wild perturbation response2026-04-30T09:40:23ZVirtual cell (VC) models aim to predict cellular responses to any perturbations in silico and have emerged as a promising approach for drug discovery and precision medicine. Yet, a clear gap still remains: while models routinely reported impressive results on standard benchmarks, it is unclear whether their predictions are truly meaningful in practice. This is mainly due to limitations in current evaluation setups, which are often overly simplified or inconsistent, and do not reflect the complexity and variability of real biological systems. Here, we introduce a standardized and modular benchmarking framework for virtual cell prediction. Our framework evaluates diverse models under in-the-wild challenging scenarios, including unseen cell contexts, unseen perturbations, and cross-dataset generalization, which better reflect practical applications. Our analysis shows that model performance is highly context-dependent and shaped by task design and evaluation criteria. In commonly used setups, performance is often overestimated, and naive dataset aggregation can even reduce performance. When evaluated under more strict conditions, model performance drops markedly, indicating limited robustness to shifts across cellular contexts. In unseen perturbation settings, models including simple linear approaches capture global transcriptional trends but fail to recover fine-grained perturbation-specific effects. In addition, different evaluation metrics focus on different biological properties, leading to substantially different model rankings. Together, our framework provides a more reliable and biologically grounded evaluation, offering clearer guidance for applying virtual cell models in real scenarios.2026-04-30T09:40:23ZXinjie MaoSongming ZhangQianhong WenXiangyu WenKedu JinHao WuShuizhou ChenYuqiang LiLei BaiQi LiuNing DingSiqi SunZhangyang Gaohttp://arxiv.org/abs/2604.26517v1MTCurv: Deep learning for direct microtubule curvature mapping in noisy fluorescence microscopy images2026-04-29T10:32:50ZAccurate quantification of the geometry of curvilinear biological structures is essential for understanding cellular mechanics and disease-related morphological alterations. Microtubule curvature is a key descriptor of filament rigidity and mechanical perturbations. However, reliable curvature extraction from fluorescence microscopy images remains challenging due to noise, low contrast, and partial filament visibility. Existing approaches rely on segmentation pipelines with pre or post-processing, which are highly sensitive to segmentation errors and often fail under adverse imaging conditions. In this work, we propose MTCurv, a deep learning framework for direct, segmenta-tion-free regression of microtubule curvature maps from noisy microscopy images. Leveraging a synthetic dataset with pixel-wise curvature annotations, we reformulated curvature estimation as a regression problem and adapted an attention-based residual U-Net. To reduce hallucinations and enforce spatial coherence, we introduced a gradient-aware loss combining Mean Squared Error with a gradient consistency term. Beyond model and loss design, we evaluated commonly used regression and image quality metrics, revealing that many perceptual and blind metrics are poorly suited for curvature estimation. Correlation-based metrics, particularly Spearman correlation, emerged as more reliable indicators of curvature prediction quality. Experiments on two datasets of increasing difficulty demonstrated that MTCurv accurately recovers local microtubule curvatures, even in the presence of background fluorescence. Ablation studies highlighted the contribution of both residual encoding and attention-based decoding. Overall, this work provides a practical tool for filament curvature analysis and methodological insights for geometry-aware regression in biomedical imaging. Datasets and code are made available.2026-04-29T10:32:50ZAccepted for presentation at the International Conference on Pattern Recognition (ICPR) 2026Achraf Ait LaydiSidi Mohamed Sid'El MoctarYousef El MourabitHélène Bouvraishttp://arxiv.org/abs/2506.11272v2Maximum-Entropy Model of Colored Noise in Superdiffusive Axonal Growth2026-04-27T21:17:27ZWe develop a coarse-grained stochastic theory for axonal growth on micropatterned substrates using the Shannon--Jaynes maximum entropy principle. Starting from a Langevin description of growth cone motion, we infer the effective distribution of traction force relaxation rates from experimentally motivated constraints rather than postulating the colored noise directly. The resulting relaxation rate distribution generates a stationary colored acceleration process with power-law temporal correlations and yields analytical predictions for the axonal mean squared displacement and velocity autocorrelation. The long-time behavior is controlled by the slow-relaxation part of the inferred distribution, corresponding physically to broadly distributed clutch or adhesion engagement times. For biologically relevant parameters, the model predicts a negative correlation exponent $α=-1/2$. This prediction is in close quantitative agreement with measurements on cortical neurons cultured on micropatterned poly-D-lysine-coated PDMS substrates, which are well described by $α\simeq -0.6$ and exhibit superdiffusive mean squared displacement scaling with exponent $1.4$. The same framework accounts for the crossover from early diffusive behavior to long-time anomalous growth and for the corresponding power law decay of the velocity autocorrelation. These results show how entropy-constrained active fluctuations can connect microscopic force generation processes to emergent growth laws in neuronal systems and, more broadly, in active matter.2025-06-12T20:25:22Z20 pages, 5 figuresJulian SutariaCristian Staiihttp://arxiv.org/abs/2604.24673v1Quantifying the effect of phenotype on clustering behaviour in melanoma: from monoculture to co-culture2026-04-27T16:33:43ZMelanoma is an aggressive form of skin cancer. Survival rates are excellent if it is detected early but fall markedly if it metastasises. A key step in early tumour progression is the formation of cell clusters, which can promote metastasis. However, the mechanisms driving cell clustering, and the role of phenotypic heterogeneity in the dynamics of these clusters, remain poorly understood. In this work, we propose a system of ordinary differential equations that models cluster formation dynamics within a coagulation-fragmentation-proliferation framework. Using Bayesian inference, we fit this model to in vitro time-lapse microscopy data from two melanoma phenotypes-proliferative and invasive-to uncover the predominant mechanisms driving cluster formation and how these differ between phenotypes. Additionally, we provide preliminary insights into how clustering behaviour in co-cultures contrasts with that observed in monocultures. The model quantifies phenotypic differences in clustering dynamics: invasive cells in monoculture exhibit nearly threefold higher coagulation rates than proliferative cells, whereas proliferative cells display slightly higher proliferation rates. These differences align with known gene expression profiles. When applied to co-culture data, the model predicts hybrid coagulation behaviour of the clusters influenced by both proliferative and invasive cells but dominated by the invasive cells, and an elevated proliferation rate, suggesting a mutually beneficial effect of phenotypic heterogeneity on cell proliferation.2026-04-27T16:33:43ZNathan SchofieldRichard WhiteRuth BakerHelen Byrnehttp://arxiv.org/abs/2509.25346v2SynthPert: Enhancing LLM Biological Reasoning via Synthetic Reasoning Traces for Cellular Perturbation Prediction2026-04-26T19:07:19ZPredicting cellular responses to genetic perturbations represents a fundamental challenge in systems biology, critical for advancing therapeutic discovery and virtual cell modeling. While large language models (LLMs) show promise for biological reasoning, their application to perturbation prediction remains underexplored due to challenges in adapting them to structured experimental data. We present SynthPert, a novel method that enhances LLM performance through supervised fine-tuning on synthetic reasoning traces generated by frontier models. Using the PerturbQA benchmark, we demonstrate that our approach not only achieves state-of-the-art performance but surpasses the capabilities of the frontier model that generated the training data. Our results reveal three key insights: (1) Synthetic reasoning traces effectively distill biological knowledge even when partially inaccurate, (2) This approach enables cross-cell-type generalization with 87% accuracy on unseen RPE1 cells, and (3) Performance gains persist despite using only 2% of quality-filtered training data. This work shows the effectiveness of synthetic reasoning distillation for enhancing domain-specific reasoning in LLMs.2025-09-29T18:02:41ZLawrence PhillipsMarc Boubnovski MartellAditya MisraJosefa Lia StoisserCesar A. Prada-MedinaRory Donovan-MaiyeKaspar Märtenshttp://arxiv.org/abs/2604.23151v1Hydrodynamic interactions mask the true heterogeneity of a microscopic collective2026-04-25T05:36:41ZCoordinated movement and self-organisation of active self-driven agents is common in nature and is seen across different scales, from herds of animals to collective motion in bacteria. Often, these systems are heterogeneous in composition, with different agents having different intrinsic motilities. Inferring these intrinsic characteristics and quantifying the level of heterogeneity in a collective system is crucial to understanding the observed emergent phenomena. However, when interaction effects dominate, i.e. the observed movement of an agent is strongly influenced by its interacting neighbours, inferring the intrinsic characteristics of agents becomes a challenge. We consider a collective system of agents that undergo purely physical interactions like collisions and long-range hydrodynamic interactions, which resembles a system of microswimmers immersed in a fluid medium. We incorporate heterogeneity into the system through variations in agent motility and examine how the perceived heterogeneity, inferred from measured speeds, depends on the strength of hydrodynamic interactions and the true intrinsic variability. The interplay between short-range collisions, long-range hydrodynamic interactions, and intrinsic heterogeneity makes the inference problem non-trivial. When hydrodynamic effects dominate, true heterogeneity is effectively masked, making even a homogeneous collective appear heterogeneous. The competing effects of collisions, which slow agents down, and hydrodynamic interactions, which enhance their motion, further complicate reliable inference. Hydrodynamic interactions also modify collision angles, rendering them more isotropic. Overall, the findings show highlight experimentally measured properties of microscopic collectives may not accurately reflect their true characteristics.2026-04-25T05:36:41Z9 pages, 6 figuresBalagopal NairArshed NabeelDanny Raj M