2026-08-19 | | Total: 12
Chimeric antigen receptor (CAR) T-cell therapy has transformed the treatment of B-cell acute lymphoblastic leukaemia (B-ALL). Despite high initial response rates, a substantial fraction of patients relapse, often due to loss of CAR T-cell persistence, antigen escape, or immune-privileged sites that shield tumour cells. Prolonged CAR T-cell persistence is clinically associated with durable remission, but why it is required remains poorly understood. To address this, we develop and analyse the BEAM (Blast, Effector, Activated, Memory) model of CAR T-cell dynamics in B-ALL. BEAM extends predator--prey models with three CAR T-cell states (memory, activated, effector) coupled to a logistic growth equation for the blasts, calibrated against the FELIX trial of obecabtagene autoleucel in adult B-ALL. We find that both memory and effector persistence prevent relapse, but for distinct reasons: memory persistence sustains surveillance against low-burden or slowly proliferating residual disease, while effector persistence clears isolated blasts emerging from immune-privileged sites. The model further predicts a trade-off between immediate cytotoxicity and durable surveillance, and identifies initial tumour burden as a key modifiable factor for reducing antigen-negative relapse. Together, these results offer a framework for designing more durable, individually tailored CAR T-cell therapies.
Single-particle cryo-electron microscopy images a macromolecule as many noisy tomographic projections of its electrostatic potential. We reconstruct the protein backbone directly from such projections, as an atomic point cloud, without the intermediate step of reconstructing the 3D electrostatic potential map. We formulate this as an indirect shape matching problem: a point-cloud template of the backbone is deformed until its simulated projections agree with the data, with the structure observed only through the imaging operator. The deformation is computed via a gradient flow on a Lie group, and we derive the resulting framework in a general geometric setting before adapting it for single-particle cryo-electron microscopy. On synthetic data, we recover single- and multichain proteins and capture conformational transitions.
Preterm birth remains a major cause of neonatal morbidity and mortality worldwide. Electrohysterography (EHG), a noninvasive measure of uterine myoelectrical activity, has been studied for preterm-birth prediction, but performance estimates may be biased when segments from the same maternal record are split across training and validation data. We formalized the distinction between segment-level and patient-independent record-grouped validation, established a patient-independent benchmark on the Term-Preterm Electrohysterogram Database, and evaluated class-conditional conformal selective prediction. All 300 records (38 preterm) were analyzed across three prespecified regimes. A 92-feature elastic-net logistic model was evaluated with record-grouped nested cross-validation; preprocessing, model fitting, Platt calibration, and conformal estimation were confined to training data, and performance was calculated from strictly out-of-fold record-level predictions with 1,000 bootstrap resamples. Under patient-independent evaluation, AUROC was 0.493 (95% CI 0.467-0.520), AUPRC 0.122 (0.095-0.152), and Brier score 0.115 (0.094-0.137). AUROC was 0.514 at or before 26 weeks and 0.469 thereafter. At miscoverage alpha=0.10, marginal coverage was 0.897, abstention 72.7%, and singleton-prediction accuracy 0.624. These findings establish a record-separated reference benchmark for EHG prediction and a broader validation principle for segmented physiological data: the unit of resampling should correspond to the unit at which predictive performance is intended to generalize. Conformal prediction further quantifies when the available information supports a singleton classification and when uncertainty warrants deferral.
Motivation: Low-dimensional embeddings are widely used to explore cell-state heterogeneity in single-cell and other high-dimensional biological data. Although many methods preserve local neighborhoods, they may distort the apparent sampling density of processed observations, altering the visual contrast between dense and sparse regions and complicating the interpretation of rare, transitional, or continuous cell-state populations. Results: We present DMT-Dens, a parametric manifold-visualization method built on a latent-token Transformer encoder. The model integrates rank-based manifold alignment with hard-pair aggregation. To preserve density, it optimizes a loss based on the Pearson correlation between k-nearest-neighbor log-radius estimates in the processed input and two-dimensional embedding spaces. Benchmark evaluations demonstrate strong density preservation, particularly on biological datasets, while retaining competitive label separability. Availability: Source code, data-processing scripts, and resolved experiment configurations are available at https://github.com/Ruizhe-wang/DMT-Dens.
Biomolecular design underpins applications from molecular recognition to therapeutics and synthetic biology, yet de novo interaction design remains challenging-especially for DNA/RNA, underexplored non-protein modalities with scarce, heterogeneous complex data and sharper geometric and chemical constraints. We introduce MCTH (Monte Carlo Tree Hallucination), an inference-only framework that casts all-atom sequence-structure co-design as uncertainty-aware planning over hallucinated states from pretrained folding and inverse-folding models, with optional biophysical control within the same decision loop. MCTH treats these models as frozen black-box operators and uses Monte Carlo Tree Search to allocate a fixed inference budget across competing design trajectories, incorporating model confidence and uncertainty, as well as cross-expert consensus/disagreement when multiple predictors are available. Across protein-RNA, protein-DNA, protein-protein, and protein-ligand design, matched-budget experiments show that adaptive search improves over simpler sampling and cycling strategies, while held-out AlphaFold3 and Chai-1 evaluations demonstrate transfer beyond the search-time oracle. MCTH provides a shared planning layer across modalities while allowing task-specific folding, inverse-folding, and biophysical modules, requiring no fine-tuning or backpropagation through component models.
Deep clustering models for single-cell RNA sequencing often assign cells through latent or centroid-based mechanisms that are difficult to inspect. We introduce scDNM-VAE (single-cell Dendritic Neuron Model Variational Autoencoder), a deep clustering framework that combines a variational autoencoder with a dendritic neuron-inspired head. Cluster assignments are governed by learnable signed synaptic weights and thresholds: the weight sign determines the direction of a gate's response to a latent coordinate, its magnitude controls steepness, and the weight-threshold pair determines the transition location. The trained clustering function can therefore be inspected directly without fitting a post-hoc explanation model. We benchmark scDNM-VAE on four datasets spanning immune, cortical, cardiac, and hematopoietic cells against scVI followed by KMeans and an MLP-DEC ablation. scDNM-VAE performs better than scVI on PBMC3k, comparably on the Human Heart Cell Atlas and Paul15, and worse on Zeisel, while producing biologically coherent marker-gene signatures. Ablating each cluster's three highest-magnitude synaptic dimensions causes numerically greater reassignment than random-dimension ablation across all datasets, but the margins are modest and negligible on Zeisel. These results show that signed dendritic gating supports competitive clustering with a parameter-inspectable decision function, while indicating that decision-relevant information is distributed across the latent space.
Competitive interactions can maintain diversity, yet coexistence is often fragile in well-mixed populations, where stochastic fluctuations can lead to extinction. This is the case in non-transitive systems, such as rock-paper-scissors dynamics, where no single type dominates globally. While spatial structure can stabilize these systems by providing refuges in space, it remains unclear whether analogous mechanisms can operate in time in well-mixed environments. Here, we develop a population-genetic framework showing that dormancy can act as a temporal refuge, preserving lineages and preventing collapse to fixation under interaction-driven fluctuations. We introduce a discrete-time Wright-Fisher model that combines generalized seed-banks with frequency-dependent interactions, allowing individuals to inherit their type from potential parents sampled across multiple past generations. This construction provides a tractable framework in which dormancy stores and later reintroduces lost types. In the case of either weak or moderate selection, we prove a multidimensional diffusion limit for the resulting type-frequency process and use it to analyze complex selective interactions. In non-transitive systems, dormancy stabilizes trajectories that would otherwise collapse through stochastic extinction, extends fixation times, and sustains coexistence. These effects cannot be explained solely by an increase in effective population size. Our results show that dormancy introduces temporal memory that qualitatively alters competitive dynamics, stabilizing otherwise fragile systems and enabling long-term coexistence.
What happens when you teach an LLM-based agent the scientific method? Motivation: Scientific discovery emerges from cycles of hypothesis, implementation, empirical testing, and feedback. Can this process be automated? We approach automated algorithm design through the lens of the scientific method, where an LLM-based agent goes through each step of the process in an ordered, iterative fashion. Results: We present The Little Scientist, a framework in which a "Scientist agent" works inside an evaluation environment that benchmarks its code and returns structured per-instance diagnostics. When the Scientist plateaus at a local optimum, a "Kuhn agent" injects a paradigm-shifting conjecture paired with a cross-disciplinary inspiration, forcing exploration of a different region of the LLM's latent space. We demonstrate the framework on two problems that require fundamentally different modes of discovery. For protein fitness prediction, the Scientist discovered Delta V, an ensemble calibration strategy that ranks first on the ProteinGym DMS Substitutions Zero-Shot leaderboard across all five official evaluation metrics, exceeding the #2 model (VenusREM) by +0.033 mean Spearman correlation across 217 DMS assays. For DNA motif discovery, the Scientist wrote an algorithm from scratch--DALE (Dual-seed Algorithm for Latent Enumeration)--that outperforms STREME (the default in the MEME Suite) across 132 ENCODE transcription factors (mean AUROC 0.842 vs. 0.803, Wilcoxon p < 10^{-6}) while running 11x faster. This demonstrates that the framework can produce genuinely novel algorithms, not just optimize existing components. Together, these results show that an LLM agent stepping through the scientific method can discover both new algorithms and new ensemble strategies that outperform prior solutions. The entire research program consumed 704M tokens on a single virtual machine with no GPUs
Biological networks feature recurring motifs, but their uneven abundance remains poorly understood. Feed-forward loops (FFLs) are important motifs in the transcriptional networks of \textit{Escherichia coli} and \textit{Saccharomyces cerevisiae}, yet their eight types appear in highly unequal frequencies. This study presents an information-theoretic framework that decomposes input-output mutual information into pathway and interference components. The interference mutual information (IMI) quantifies the joint influence of direct and indirect regulatory pathways on information transmission. Our results show that IMI closely follows the observed abundance patterns in both organisms, whereas the input-output mutual information and the pathway component do not. We also trace the IMI hierarchy to the strength of pathway-interference interactions and local pathway sensitivities. Notably, a stronger IMI is also associated with reduced robustness to parameter variation, suggesting a trade-off between information processing and resilience in gene regulatory circuits. Our results, thus, identify a topology-dependent information-theoretic feature that relates to the abundance patterns.
Preterm birth (PTB) remains a major global health problem, and reliable non-invasive risk assessment remains difficult. Electrohysterography (EHG) records uterine electrical activity from the maternal abdomen and may support PTB assessment, but performance can be inflated when segments from the same recording are split across training and validation folds. We evaluated empirical mode decomposition (EMD) for term-versus-preterm classification using 26 pregnancy recordings (13 preterm, 13 term) from the public TPEHGT dataset. Annotated intervals and non-overlapping fixed 3-minute windows were compared, and the first four intrinsic mode functions (IMFs) were evaluated. Fourteen features from each of three EHG channels were assessed with nine classifiers using repeated five-fold recording-grouped cross-validation and recording-level aggregation. IMF1 gave the strongest mean performance. With fixed 3-minute IMF1 features, Random Forest achieved mean accuracy 0.8308, F1 0.7969, balanced accuracy 0.8308, MCC 0.6998, ROC-AUC 0.8157, and average precision 0.8877. IMF1 also outperformed matched filtered time-domain features across all reported mean metrics. Preterm recordings showed smaller, more regularly spaced peak-like events, lower temporal-energy measures, and higher entropy. These findings support further evaluation of IMF1-based EHG classification in larger independent cohorts.
The emergence of organized spatiotemporal patterns is ubiquitous in oscillatory systems, from neural populations to engineered networks. Identifying these patterns and tracking how they evolve over time remains challenging, particularly when systems exhibit transient dynamics. Here, we introduce a framework based on spatial ordinal patterns to characterize the spatiotemporal dynamics of oscillatory systems. Our approach acts directly on the phase rather than the amplitude, with additional patterns introduced to account for near-equal phases. This symbolic representation encodes local spatial ordering relations, capturing both phase gradients and synchronized clusters within a single framework. From this construction, we define a spatial permutation entropy that quantifies the diversity of spatiotemporal patterns at each point in time, enabling the detection of transient dynamics and regime transitions as they occur. We show that this approach distinguishes phase-locked states with identical levels of global synchronization but distinct spatial organization and also characterizes partially synchronized states. We demonstrate the method on synthetic oscillator networks across multiple spatiotemporal regimes, and on resting-state EEG recordings from human volunteers, where it distinguishes different conditions within individual volunteers.
Large language models (LLMs) are increasingly proposed for healthcare decision support, but their evaluations still reward single-answer accuracy rather than reasoning about interventions, mechanisms, harms, evidence, and uncertainty. We propose a reproducible, graph-centered evaluation framework for intervention-oriented LLM behavior in healthcare and stress-test it in a cardiovascular pilot. The framework has four components: (i) a domain causal knowledge graph in which assertions are first-class, provenance-preserving nodes with stable identifiers; (ii) a scenario-conditioned subgraph extraction step that, given any clinical scenario, retrieves the relevant reified-assertion subgraph; (iii) four controlled grounding conditions that vary how the retrieved subgraph is composed into the model's context (ungrounded C1, knowledge-graph C2, causal-graph C3, integrated C4); and (iv) an automated scoring pipeline, anchored on assertion identifiers, that computes intervention accuracy, and other evaluation measures on a single pass. To test the framework, we built a category-balanced scenario generator across eight reasoning failure modes and instantiated it on a cardiovascular graph. The metric panel discriminates conditions along interpretable, non-redundant axes: C4 obtains the strongest causal edge F1 (0.838), adverse-effect F1 (0.833), evidence accuracy (0.738), and unsupported claim rate (0.114), while C1 obtains the highest raw intervention accuracy (0.948) with no measurable causal or evidential grounding.