• raseliarison
  • nirinA
  • adrien
  • blog
  • code
  • FAQ
  •  home  
  •  news  
    • arXiv
      • astro-ph
      • cond-mat
      • cs
      • eess
      • gr-qc
      • hep-ex
      • hep-lat
      • hep-ph
      • hep-th
      • math
      • math-ph
      • nlin
      • nucl-ex
      • nucl-th
      • physics
      • q-bio
      • quant-ph
      • stat
    • physics
      • phys.org
      • physics world
    • linux
      • kernel
      • slackware
    • nature
      • natcomputsci
      • natastron
      • natbiomedeng
      • nenergy
      • nnano
      • natmachintell
      • nbt
      • nmeth
      • natecolevol
      • nmicrobiol
      • ng
      • nchembio
      • natelectron
      • micronano
      • nphoton
    • bioRxiv
    • plos one
    • world
      • BBC
      • Al Jazeera
    • earth
      • earth observatory
      • weather
      • weather forecast
    • universe
      • apod
      • hubble
      • atel
      • nasa
  •  wiki  
  •  gemini  
  •  python  
  • q-bio updates on arXiv.org

    q-bio updates on the arXiv.org e-print archive.

    Influence of passenger head-position uncertainty on infection risk predictions

    oai:arXiv.org:2608.04024v1

    arXiv:2608.04024v1 Announce Type: new Abstract: We investigate the influence of natural head movement on the infection risk posed by airborne pathogens using a CFD-based forward risk prediction model. Quasi-Monte-Carlo simulations are used to obtain the resulting infection risk distributions by representing head movement via probability distributions of parameters describing the position and orientation of each passenger's breathing zone. A significant impact of fore/aft and lateral head position on infection risk was found and should be accounted for to increase robustness when predictions of local, seat-specific infection risks are used to guide design and policy decisions. Unlike the sampling-based Monte Carlo approach, estimates of the statistical moments of the risk distributions and sensitivities, calculated using first-order second-moment and higher-order methods, were found to inadequately capture the dependencies between head movement components and infection risk.

    https://arxiv.org/abs/2608.04024


    The Cost of Binarizing Survival Outcomes in Clinical Prognostic Modeling

    oai:arXiv.org:2608.04046v1

    arXiv:2608.04046v1 Announce Type: new Abstract: Survival analysis is an established framework for analyzing time-to-event data, yet many clinical machine learning studies still binarize the outcome before model training. This practice excludes censored patients, collapses temporal information into a single threshold, and can affect which features are selected as prognostically relevant. We examine the cost of this binarization in the context of Bayesian network (BN) feature selection, using two recent publications as case studies: one that applies BN-based feature selection to a head-and-neck cancer cohort and a second surgical cohort study that, while not BN-based, likewise binarizes its survival endpoint. We replace the binary scoring function with the Cox partial log-likelihood for feature-to-outcome edges, a modification we call the Survival-Aware Bayesian network, and recover prognostic features that binarization misses. Our ablation experiment confirms that the improvement is driven by the time-to-event scoring formulation rather than by retaining more patients. The results generalize across five endpoint-cohort combinations in head-and-neck cancer and extend to three further cancer types (breast, colorectal, and kidney). We propose that clinical studies with survival outcomes should use time-to-event methods by default, as binarization discards the prognostic signal retained by survival analysis.

    https://arxiv.org/abs/2608.04046


    Time^2: A framework for the neural dynamics of visual perception

    oai:arXiv.org:2608.04218v1

    arXiv:2608.04218v1 Announce Type: new Abstract: Whenever we look at an object, we seem to perceive it immediately. However, this is not the case for two reasons. First, it takes hundreds of milliseconds for the brain to process visual information reaching the retina. Second, we have to look at an object for a certain amount of time to perceive it (and we typically look at it for hundreds of milliseconds) -- during that time, visual information is continuously received on our retinas. These facts together imply that visual information is both processed and received through time. These two temporal facets of perception, which we term processing time and stimulus time, are often conflated in the literature. Moreover, processing time and stimulus time are usually not considered together in experiments. Here, we argue that, to obtain a more complete portrait of visual perception and constrain further models of vision, it is essential to consider and measure both temporal facets simultaneously. We present a new method designed to do so that is based on reverse correlation: Time^2. We show that this method allows us to precisely characterize many neural phenomena, including rhythmic perception, predictive processing and coarse-to-fine sampling.

    https://arxiv.org/abs/2608.04218


    On the realizability of abstract reaction networks with real molecules and reactions

    oai:arXiv.org:2608.04635v1

    arXiv:2608.04635v1 Announce Type: new Abstract: Abstract reaction networks appear not only as models of chemical reactions but also as models of complex systems, with applications in areas such as ecology and epidemiology, and as one of several alternative paradigms for non-standard computation. It is therefore of interest to determine whether an abstract reaction network can be realized by a concrete set of molecules and plausible chemical reaction mechanisms. It is known that a reaction network has a realization in terms of chemical graphs (i.e., Lewis structures) if and only if it is conservative. Here we consider the problem of assigning a set M of known molecules to a set X of abstract entities in a reaction network (X,R) such that each reaction satisfies mass balance and adheres to one of an allowed set of chemical reaction mechanisms. We show that this problem is NP-complete. Nevertheless, it can be solved in practice using a backtrack-and-prune algorithm inspired by the VF2 family of algorithms originally designed for the subgraph isomorphism problem.

    https://arxiv.org/abs/2608.04635


    Poisson Flow and Wasserstein Registration of Trees

    oai:arXiv.org:2608.04770v1

    arXiv:2608.04770v1 Announce Type: new Abstract: Tree-like structures arise in numerous imaging applications, including vascular networks, neuronal arbors, airway trees, and cortical sulcal--gyral folding. We present a nonlinear registration framework based on screened Poisson flow and Wasserstein distance. The screened Poisson equation transforms geometric features into smooth multiscale probability distributions. Registration is formulated by minimizing the Wasserstein distance between these distributions, producing anatomically meaningful correspondences without explicit landmark or branch matching. The framework is demonstrated on the nonlinear registration of cortical sulcal--gyral folding patterns from structural MRI.

    https://arxiv.org/abs/2608.04770


    Traveling fronts in a spatial epidemic model with slow loss of immunity

    oai:arXiv.org:2608.04594v1

    arXiv:2608.04594v1 Announce Type: cross Abstract: We investigate the emergence of traveling front solutions in a spatial SIRS epidemic model with diffusion acting on the infected population. The model exhibits a natural slow-fast structure due to the presence of a small parameter governing the loss of immunity, which induces a separation of scales in the dynamics. Using a traveling wave reduction, the PDE system is transformed into a singularly perturbed system of ODEs, which we analyze within the framework of Geometric Singular Perturbation Theory. In the singular limits, we study the fast excursions governed by the layer problem, and the slow evolution close to the critical manifold. In particular, we identify an entry-exit mechanism tracking the transitions between slow and fast regimes, and derive a quantitative characterization of the entry-exit dynamics. Numerical simulations of the full system confirm the validity of the proposed geometric picture. The traveling front is shown to consist of a concatenation of local, fast, and slow segments, in agreement with the theoretical analysis.

    https://arxiv.org/abs/2608.04594


    An entropic explanation of insistence on sameness in autism

    oai:arXiv.org:2608.04616v1

    arXiv:2608.04616v1 Announce Type: cross Abstract: An information theory-based framework is proposed in attempt to explain insistence on sameness in autism as an instance of a general behavior pattern in which an individual tries to reduce surprise and uncertainty. It offers a new definition of autism as an impairment in which cognitive functions are restricted to discrimination, memorization and prediction of tangible properties of the environment. An analogy between insistence on sameness and constrained minimization of the entropy metric is observed and examined for a set of assumptions that describe cognitive limitations of a person with autism. The metric is given by the formula $D_H(R, M) = H(R|M) + H(M|R)$, where $R$ represents sequences of random stimuli, $M$ is a memory that stores and retrieves them, and where $H(.|.)$ denotes their conditional entropies interpreted as surprise and uncertainty, respectively. It is first inferred that to minimize the metric an individual can learn about $R$ (and store that knowledge in $M$) or can restrict $R$ to the already known $M$. Then, it is concluded that insistence on sameness is a manifestation of the latter. Moreover, it is shown that the proposed framework: (1) Helps to quantify the concepts of surprise, uncertainty, sensory overload and deprivation, anxiety, comfort zone, disappointment, disorientation, pedantry, rigidness, observance or aberrant precision. (2) Leads to a list of guidelines for learning therapies and daily care routines, and allows them to be defined as optimization algorithms and implemented as programs for robotic live-in caregivers. (3) Can be validated with the help of a Turing test-like approach that requires no experiments involving individuals with autism. The framework-if positively validated-will provide formal foundations and design guidelines for therapies aimed at improving self-reliance of individuals with autism in basic activities of daily living.

    https://arxiv.org/abs/2608.04616


    Parameter identification for predator-prey system with sparse data

    oai:arXiv.org:2608.04959v1

    arXiv:2608.04959v1 Announce Type: cross Abstract: Parameter identification from observations of dynamical systems is a fundamental problem in population biology. Mechanistic models of ecological systems rely on optimization methods that require accurate initial guesses to guarantee convergence. In ecological applications, datasets contain observation noise and are collected at sparse time points. This sparsity creates irregular likelihoods that cause standard optimization methods to struggle, while the ordinary differential equation solvers can become stiff or unstable in certain regions of the parameter space. These instabilities cause long running times or runtime errors. Here we present a computational framework for parameter identification that addresses these numerical instabilities by employing Natural Gradient Ascent, and we apply it to the classical Lotka-Volterra predator-prey model. We exploit the non-dimensionalization of the ordinary differential equations to treat scaling factors as nuisance parameters, reducing the dimensionality of the optimization problem. To prevent the solver step from becoming small, we implement an adaptive solver that switches between two independent second-order equations derived from the two components of the model. This approach allows Natural Gradient Ascent to converge in fewer iterations and with more stability than standard gradient ascent or BFGS methods. This framework provides a reliable method for parameter estimation in ecology when data is limited. The method can be generalized to other dynamical systems as long as the different components of the system do not become numerically problematic at the same time.

    https://arxiv.org/abs/2608.04959


    Inferring Relative Consequences of Mechanical Ventilation from Observational Data Using Game-Based Comparisons

    oai:arXiv.org:2510.15127v4

    arXiv:2510.15127v4 Announce Type: replace Abstract: Identifying the effects of mechanical ventilation (MV) protocols in critical care requires analyzing data from heterogeneous patient-ventilator systems in the clinical decision-making environment. Multiscale interactions among these coupled components generate a high-dimensional state space that remains sparsely sampled despite extensive data collection. Analysis of existing data is essential for understanding current respiratory management practices and generating testable hypotheses about improvement. The scale and complexity of available data motivate the use of reinforcement learning (RL) to explore data-consistent counterfactual trajectories. However, formulating RL in practical applications requires a spatiotemporally dependent reward process that defines state-to-consequence relationships, their context dependence, and the delays over which consequences emerge. These poorly understood elements are not known \emph{a priori} and inferred from data via hypotheses. To that end, categorized observed states are contrasted according to their relative consequences by solving a game-based inverse problem that identifies a comparison model required for downstream probabilistic and stochastic methods such as reinforcement learning for seeking MV optimization and personalization. The inverted-game inference is validated on synthetic data to reveal potential caveats before proceeding to real-world ICU data applications that expose complexities of the data-generating process. Clinical data applications revealed that both breath-type consequences and their relative ordering are inherently context- and time-dependent, varying across patient subgroups, time, and comparison quantities, and effect timescale. The discussion includes potential developments toward a state transition model for simulating the effects of MV management actions using empirical data and game-inferred comparisons.

    https://arxiv.org/abs/2510.15127


    On the abelian structure of noncompetitive chemical reaction networks

    oai:arXiv.org:2512.17491v2

    arXiv:2512.17491v2 Announce Type: replace Abstract: Chemical reaction networks (CRNs) are foundational models for describing complex biochemical processes. We study noncompetitive CRNs, a class of networks whose static states, where the CRN is inactive, are rate independent, and that can implement ReLU neural networks. CRNs of interest in biochemistry and systems biology are embedded in complex networks so that CRNs have to respond to internal and environmental cues. We describe the network's response to such perturbations using a new Markov chain that we call CRN sandpile Markov chain, whose state space is the set of static states. The transition mechanism of the CRN sandpile Markov chain is defined by adding a molecule of a randomly chosen species to a static state, and then letting the CRN state evolve toward a new static state. A central contribution of the present work is the observation that one can associate a natural Abelian Network (AN) to each noncompetitive CRN, and use AN theory to get new mathematical results on noncompetitive CRNs. For noncompetitive CRNs on a finite state space, we use AN theory to get that only a fraction of the static states are recurrent for the CRN sandpile Markov chain. We obtain furthermore that the set of recurrent states is in one to one correspondence with the critical group of the AN, which plays a major role in AN theory. Overall, this work establishes a unified algebraic and probabilistic framework for analyzing the long-term behavior of noncompetitive CRNs. We focus on a special class of noncompetitive CRNs called generalized toppling networks, and obtain new mathematical results both for the CRN and AN settings.

    https://arxiv.org/abs/2512.17491


    Plausibility-Driven Prioritization of Candidate Biomedical Annotations

    oai:arXiv.org:2607.20163v2

    arXiv:2607.20163v2 Announce Type: replace Abstract: The rapid growth of biomedical knowledge has made the validation of automatically generated biological annotations a major bottleneck in biomedical curation. While computational methods can rapidly produce large numbers of candidate annotations, determining which are biologically valid still requires costly expert review. Prioritizing these candidates before manual curation has therefore become a fundamental challenge. Machine learning techniques can support this process by exploiting biomedical knowledge graphs (bioKGs), which capture biological entities and their functional associations. In this work, we propose a framework that leverages bioKGs to estimate the plausibility of candidate annotations and guide expert curation. Starting from knowledge graph embeddings, we train relation-specific binary classifiers using a community-based negative sampling strategy to obtain reliable confidence estimates. We then introduce a family of plausibility measures that combine classifier confidence, classifier reliability, and the semantic context provided by alternative relationships involving the same pair of biological entities. Unlike conventional confidence estimation, the proposed approach explicitly accounts for multiple biologically meaningful relations that may coexist between the same entities. Experimental results on five large bioKGs demonstrate that the proposed negative sampling strategy consistently improves classifier robustness, increasing balanced accuracy by an average of 5.8%. Moreover, the plausibility measures outperform classifier confidence alone, enabling more effective prioritization of candidate annotations for expert review. Overall, our results show that the use of bioKGs improves the efficiency of AI-assisted biomedical curation while preserving expert control over the final annotation assessment.

    https://arxiv.org/abs/2607.20163


    Subject-Level Heterogeneity in EEG Motor Imagery Decoding: A Large-Scale Benchmark and Portfolio-Based Reduction of the Search Space

    oai:arXiv.org:2607.22778v2

    arXiv:2607.22778v2 Announce Type: replace Abstract: Robust EEG motor imagery decoding remains limited by strong inter-individual variability, making it difficult to identify pipelines that generalize across users. We present a large-scale, standardized within-session benchmark of decoding pipelines across three public datasets: Cho2017 (52 subjects), PhysionetMI (109 subjects), and Zhou2016 (4 subjects). Using a common MOABB LeftRightImagery setting, two frequency bands (8-15 Hz and 8-30 Hz), and a broad combination of feature extraction, preprocessing, and classification steps, we analyzed 216,714 raw evaluation rows, which after structured aggregation yielded 44,928, 109,000, and 4,192 subject-level observations respectively. Covariance tangent-space projection (cov-tgsp) and Common Spatial Patterns (CSP) consistently defined the strongest methodological families, though their relative ordering was dataset-dependent. On Cho2017, the best family-level mean accuracy came from cov-tgsp in 8-30 Hz (0.712 +/- 0.140), whereas Zhou2016 favored CSP (0.832 +/- 0.121 in 8-15 Hz). These aggregate rankings concealed substantial subject-level heterogeneity: 42 distinct winning pipelines across 52 Cho2017 subjects, and 93 across 109 PhysionetMI subjects. We then used the benchmark as an empirical performance landscape for building compact portfolios of pipelines of size K. Several construction procedures were compared, including a ranking-based Top-K Mean heuristic and search-based strategies. Results were broadly consistent, with Top-K Mean giving the best trade-off. A single best global pipeline already retained 94.2% of the oracle in Cho2017 and 81.8% in PhysionetMI; at K = 12, oracle retention rose to 96.5% and 90.0%. The landscape is therefore subject-dependent, and this heterogeneity can be exploited through compact portfolios that make personalization more feasible.

    https://arxiv.org/abs/2607.22778


    TCellAlign: Cross-study T-cell Populations Alignment with Nomenclature-Guided Multi-Agent Workflow

    oai:arXiv.org:2607.24093v2

    arXiv:2607.24093v2 Announce Type: replace Abstract: Cell type standardization plays a central role in integrating biological knowledge across single-cell studies. While standardized resources (e.g., Cell Ontology, Nomenclature Frameworks) provide unified vocabularies of cell populations, scientific publications and public datasets continue to use heterogeneous study-specific labels, making cross-study comparison difficult even when biologically equivalent cell populations are described. In this work, we are the first to formulate this challenge as an evidence-grounded cell population alignment problem and propose TCellAlign, a multi-agent framework that includes literature retrieval, information extraction, nomenclature-guided label alignment, and evidence-based adjudication. This modular design preserves the original terminology and supporting evidence reported by each study while producing standardized labels that can be compared across studies. We further construct a manually validated benchmark dataset linking study-specific labels, CZ CELLxGENE annotations, and standardized T-cell nomenclature across 44 manually curated, published studies (including over seven million cells) spanning four biological categories: healthy, cancer, infectious disease and inflammatory diseases. Across the evaluated tasks, TCellAlign achieves stronger semantic agreement than ontology-based baselines and maintains transcriptomic coherence with both open-source and closed-source large language models (LLM) backbones. By connecting literature, datasets, and expert's nomenclature, TCellAlign enables consistent interpretation of T-cell subtypes and states across studies, facilitating biological knowledge integration and the development of future foundation models built upon standardized cellular representations.

    https://arxiv.org/abs/2607.24093


    Mechanistic bridges from receptors to whole-brain dynamics: mean-field reductions, validity domains, and computational trade-offs

    oai:arXiv.org:2608.00306v2

    arXiv:2608.00306v2 Announce Type: replace Abstract: Many pharmacological and pathological perturbations arise at molecular, synaptic, or cellular scales, but are observed through population and whole-brain signals. Cross-scale reductions must preserve relevant mechanisms while remaining tractable. This review asks which microscopic mechanisms remain explicit, interpretable, and testable after reduction, and what claims these models support. Using receptor-aware adaptive mean fields from the master-equation lineage as a worked case, we trace finite-size population statistics and semi-analytical transfer functions into conductance-based adaptive nodes coupled through the connectome. We compare this strategy with phenomenological neural masses, low-dimensional and population-density reductions, large-scale spiking models, and learned or hybrid surrogates, including computational work and memory traffic. Receptor-dependent synaptic kinetics, conductance state, and spike-frequency adaptation can remain manipulable across scales, enabling interpretable interventions and testable mesoscopic and macroscopic consequences. However, this relies on coarse-grained Markovianity, population homogeneity, quasi-stationary transfer functions, moment closure, regional uniformity, and measurement-specific observation models. First-order implementations discard covariance dynamics, while macroscopic agreement cannot identify a unique molecular cause. Node-local biological detail mainly changes prefactors, whereas dense global covariances change the scaling class. Cross-scale models should therefore be judged by the interventions and observables they preserve, validity domain, identifiability, empirical adequacy, and computational burden. Receptor-aware mean fields are not universal, but offer a transparent, tractable strategy for selected mechanistic questions when each reduction step is independently validated.

    https://arxiv.org/abs/2608.00306


    A Blind Spot in Alignment: Quantifying Biosecurity Risks in Large Language Models

    oai:arXiv.org:2608.02684v2

    arXiv:2608.02684v2 Announce Type: replace Abstract: Large Language Models (LLMs) are accelerating biological research, yet this same capability poses a critical biosecurity threat: models that assist in protein engineering can equally be prompted to generate predicted toxin-like sequences, potentially lowering the barrier to biological misuse. Current safety evaluations, however, operate in natural language and cannot determine whether a model-generated amino acid sequence is biological gibberish or a computational risk signal. To address this evaluation blind spot, we introduce SPIKE-Bench, coupling 631 curated toxin-design prompts across seven functional categories with the SPIKE funnel, a three-stage protocol that filters output through compliance, biological plausibility, and predicted toxicity, producing stage-level diagnostics and an aggregate function-aware metric: the Functional Harmfulness Rate (FHR). An audit of 32 LLMs reveals that most models freely comply with toxin-design requests; FHR is driven primarily by biological generation capability rather than safety alignment, reaching 50.7%; and Refusal Rate fails to predict functional risk. As a first step toward mitigation, we provide BioSafe-Guard, a domain-specialized classifier that substantially reduces predicted functional risk while preserving benign utility. We release SPIKE-Bench and BioSafe-Guard at https://github.com/PKU-Alignment/SPIKE-Bench to support more rigorous biosecurity evaluation of LLMs.

    https://arxiv.org/abs/2608.02684


    Arnold: A multi-task, multi-embodiment muscle transformer policy

    oai:arXiv.org:2508.18066v2

    arXiv:2508.18066v2 Announce Type: replace-cross Abstract: Controlling high-dimensional and nonlinear musculoskeletal models of the human body is a foundational scientific challenge. Recent machine learning breakthroughs have heralded in-silico policies that master individual skills like reaching, object manipulation and locomotion in musculoskeletal systems with many degrees of freedom. However, these agents are merely "specialists", achieving high performance for a single skill. In this work, we develop Arnold, a transformer-based musculoskeletal control policy that masters multiple tasks and embodiments. Arnold combines behavior cloning and reinforcement learning to address 14 challenging control tasks spanning dexterous object manipulation, reaching, and locomotion, matching or exceeding the performance of single-task specialist policies. A key innovation is Arnold's sensorimotor vocabulary, a compositional representation of the semantics of heterogeneous sensory modalities, objectives, and actuators. Arnold leverages this vocabulary via a transformer architecture to deal with the variable observation and action spaces across tasks. This framework supports efficient multi-task, multi-embodiment learning and facilitates rapid adaptation to novel tasks, while encouraging universal motor strategies such as action and kinematic smoothness. Finally, causal probing of the motor output reveals that low-dimensional muscle synergies remain largely task-specific and that variance-based analyses systematically underestimate functional control dimensionality, consistent with biological observations on the limited transferability of such synergies. Code and data are available here: https://github.com/amathislab/arnold

    https://arxiv.org/abs/2508.18066


    Interpreting GFlowNets for Drug Discovery: What probes can and cannot show

    oai:arXiv.org:2511.19264v2

    arXiv:2511.19264v2 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) construct molecules through sequential decisions, but their internal policies remain opaque, limiting adoption in drug discovery, where chemists need interpretable rationales for proposed structures. We present a control-validated interpretability study of SynFlowNet, a synthesis-aware GFlowNet trained with a drug-likeness (QED) reward. Our framework combines gradient saliency and counterfactual edits, an undercomplete factor analysis, and an overcomplete BatchTopK sparse autoencoder, evaluated with shuffled-label controls, RDKit-descriptor baselines, a scaffold-disjoint split, cross-seed stability, and an architecture-matched untrained network. These controls materially change the interpretation. Physicochemical properties and functional groups are highly decodable from SynFlowNet embeddings, but an untrained network with the same architecture performs essentially as well as the trained policy (drug-likeness within 0.005 and marginally better molecular-size decoding). Thus, this decodability reflects graph architecture and atom featurization rather than representations acquired through policy training. The surviving conclusions are narrower: the overcomplete autoencoder reconstructs embeddings better than a matched undercomplete baseline at comparable sparsity; chemically enriched substructure detectors, including per-halogen and boron features, emerge in individual runs, although only a small subset of dictionary directions is stable across seeds; and zeroing individual features produces property-specific effects on probe decoding. Beyond SynFlowNet, this study provides a reusable control protocol for molecular-model interpretability, separating learned structure from signals supplied by architecture and input representation. High probe scores alone should not be treated as evidence of learned chemistry.

    https://arxiv.org/abs/2511.19264


    Noise-Driven Differentiation via Gene Frustration and Epigenetic Fixation

    oai:arXiv.org:2604.18185v2

    arXiv:2604.18185v2 Announce Type: replace-cross Abstract: Gene expression in cells is stochastic, yet differentiation can display reproducible timing and stable fate commitment. We develop an analytical theory for a previously identified mechanism in which weakly stable intermediate, or frustrated, gene-expression states are perturbed by stochastic fluctuations and subsequently fixed by slow epigenetic feedback. By eliminating the fast expression dynamics, we show that the differentiation of the slow epigenetic variable is driven by the noise of gene-expression, which can be amplified by regulatory interactions. We derive the logarithmic dependence of onset time for differentiation upon the effective noise intensity, and the input-dependent probability of reaching either fate. We further construct a Waddington-inspired time-dependent probability landscape that visualizes population branching and progressive fate fixation.

    https://arxiv.org/abs/2604.18185


    Leakage-Audited Benchmarking Reveals Limited Evidence for Cross-Subject Auditory-Evoked EEG Vowel Perception Decoding

    oai:arXiv.org:2605.00865v2

    arXiv:2605.00865v2 Announce Type: replace-cross Abstract: We tested whether auditory-evoked EEG supports subject-independent five-vowel perception decoding when trial identity, model identity, prediction provenance, and participant-level inference are controlled within a single benchmark. We reconstructed Study 2 event tables from OpenNeuro ds006104 version 1.0.1 and analyzed the consonant-vowel pair task. One-to-one marker-stimulus pairing yielded 3,840 independent trials; control-condition selection and artifact rejection retained 1,094 epochs from 16 participants and 61 EEG channels. Thirteen unique implementations were evaluated using leave-one-subject-out testing, with participant metrics reconstructed from 36,102 trial predictions across 33 complete prediction replicas. Random Forest was numerically highest at 21.474% balanced accuracy (95% participant-bootstrap interval, 19.526-23.482%; chance, 20%), but neither its participant-level tests nor any implementation survived correction across the 13-model family. Deep-model performance was close to chance, and several architectures showed substantial seed-dependent variation and low trial-label agreement. In a separate descriptive sensor-space representation, participant-associated effects accounted for 72.24% of the balanced standardized centroid sum of squares, compared with 2.04% for vowel-associated effects; between-participant same-vowel distances exceeded within-participant across-vowel distances for all 16 participants. An exploratory MDM analysis comprising 9,616 genuine refits across training cohorts of 3-15 participants showed no monotonic performance gain. Within this dataset and protocol, evidence for reliable cross-subject five-vowel decoding is limited. The benchmark provides a reproducible chain from source rows to retained epochs, predictions, participant-level metrics, multiplicity-adjusted inference, and bounded diagnostic analyses.

    https://arxiv.org/abs/2605.00865


    GENEB: Why Genomic Models Are Hard to Compare

    oai:arXiv.org:2606.04525v4

    arXiv:2606.04525v4 Announce Type: replace-cross Abstract: Progress in genomic foundation models is difficult to assess due to fragmented benchmarks, incompatible evaluation protocols, and task-specific reporting. As a result, claims of superiority or generality across models are often not directly comparable. We introduce GENEB, a large-scale diagnostic benchmark that evaluates frozen representations from 40 genomic foundation models across 100 tasks spanning 13 functional categories under a unified probing-based protocol, including few-shot regimes. GENEB enables controlled comparison across model scale, architecture, tokenization, and pretraining data while explicitly exposing task-level trade-offs. Our analysis shows that aggregate leaderboards are unstable: model rankings vary sharply across task categories, scale provides only modest and inconsistent gains, and architectural and pretraining alignment frequently outweigh parameter count. These results highlight limitations of current evaluation practices and position GENEB as a reference framework for principled comparison and category-aware model selection in genomic machine learning.

    https://arxiv.org/abs/2606.04525


    LDARNet: DNA Adaptive Representation Network with Learnable Tokenization for Genomic Modeling

    oai:arXiv.org:2606.04552v2

    arXiv:2606.04552v2 Announce Type: replace-cross Abstract: Genomic foundation models increasingly adopt large language model architectures, yet almost universally rely on fixed tokenization schemes such as $k$-mers, BPE, or single nucleotides, which impose arbitrary sequence boundaries that may obscure biologically relevant structure. We present LDARNet, a 110M-parameter hierarchical genomic foundation model that adapts H-Net-style dynamic chunking from autoregressive generation to masked language modeling, combining BiMamba-2 state-space layers with local attention, bidirectional routing, and a ratio-based regularizer to induce adaptive token boundaries without supervision. Fine-tuned on 27 tasks from the Nucleotide Transformer and Genomic Benchmarks suites, LDARNet achieves 15/18 wins among compact models ($<$300M parameters) and the best overall result on 9 of the 10 histone modification tasks, outperforming models up to 20$\times$ larger. A FLOPs-matched controlled experiment isolates learned routing as the source of these gains: learned boundaries beat fixed-grid boundaries by up to 14 percentage points on histone tasks at identical compute. Nucleotide-resolution analysis further shows that the learned boundaries align with canonical promoter motifs and splice junctions without supervision, providing a biological interpretation for adaptive tokenization in genomic foundation models.

    https://arxiv.org/abs/2606.04552


    MemNovo: Look Back at the Spectrum for Balanced De Novo Peptide Sequencing from Mass Spectrometry

    oai:arXiv.org:2606.11868v2

    arXiv:2606.11868v2 Announce Type: replace-cross Abstract: De novo peptide sequencing from tandem mass spectrometry is pivotal in proteomics, enabling identification of novel peptides without reference databases. While recent Transformer-based encoder-decoder models have achieved remarkable performance, we uncover a critical pathology in their inference dynamics. Through comprehensive feature scaling experiments, we demonstrate that existing auto-regressive peptide decoders tend to over-rely on generated-sequence priors while progressively under-utilizing fine-grained physical evidence from the input mass spectrum. This phenomenon leads to suboptimal results, where generated peptide sequences are biologically plausible yet not faithful to the input spectrum. To rectify this, we propose MemNovo, a training-free and plug-and-play mechanism that re-balances peptide and spectral contributions at inference time. MemNovo alleviates the information bottleneck by establishing a persistent spectral memory bank and injecting retrieved features directly into the final decoding stage via an ultra-conservative residual connection. Theoretical analysis confirms that this mechanism restores the mutual information between the decoder state and the raw spectrum. Extensive experiments on the Nine Species benchmark with two representative baselines, Casanovo and InstaNovo, demonstrate that MemNovo consistently improves both amino acid precision and peptide precision, achieving up to 39.1% relative improvement in peptide precision for Casanovo and up to 3.9% for InstaNovo, with negligible computational overhead.

    https://arxiv.org/abs/2606.11868