Searcharxiv⌕ Search

arXiv · 2610.03160

Multimodal reasoning for broadly neutralizing antibody discovery from label-free human B cell repertoires across virus families

Abstract

Discovering broadly neutralizing antibodies (bnAbs) from human natural immune repertoires remains a fundamental challenge in immunology, hindered by: the extreme rarity of bnAb, incomplete understanding of their cellular origins across pathogens, and the inability of existing computational tools to generalize across emerging viral threats. Here we present ImmuneAgent, a closed-loop AI system that integrates multimodal reasoning with continual meta-learning and wet-lab feedback to overcome these barriers. Applied to screen the natural BCR repertoires from vaccinated or infected cohorts, the system achieves a ~55% neutralization antibody discovery rate (60 of 110 cloned candidates) and a ~11% bnAb yield (12 of 110), substantially outperforming a state-of-the-art sequence-based neutralization predictor or cofolding models evaluated at the same cloning budget. Five ImmuneAgent-discovered antibodies conferred 100% in vivo protection against lethal influenza challenge, comparable to the clinical-stage therapeutic MEDI8852. The system recovered the cellular and structural determinants of bnAb activity and identified FCRL5+CD27+ atypical memory B cells as a conserved bnAb reservoir and hydrophobic interface enrichment as a cross-viral structural signature, which generalized to unseen antigens, discovering human metapneumovirus (hMPV) cross-neutralizing and human papillomavirus (HPV)-neutralizing antibodies without antigen-specific sorting. These results validate that ImmuneAgent is a generalizable framework for rapid therapeutic antibody discovery against emerging viral threats.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Hantao Lou, Jianqing Zheng, Can Yue, Meihan Zhang, Yuanchao Bao, Yu Chen, Mengting Huang, Yupeng Yang, Qianyu Pan, Nana Fu, Yansong Shi, Hongli Li, Yangyang Chai, Ruyi Chen, Wansheng Li, Zhu Liang, Rongmei Yao, Yuanhan Mo, Lei Wang, Chunmei Wang, Yun Quan, Qiong Zhang, Xiangxi Wang, Xuetao Cao. 2026-10-02. Multimodal reasoning for broadly neutralizing antibody discovery from label-free human B cell repertoires across virus families. https://arxiv.org/abs/2610.03160

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Parameter uncertainty in dynamical models: a practical identifiability index

Ordinary differential equation models are widely used to describe how complex systems change over time, but reliable conclusions depend on how well their parameters can be estimated from limited, noisy data. Common measures of uncertainty, such as confidence intervals and coefficients of variation, do not provide a single, easy-to-interpret scale for comparing uncertainty across parameters, models, error structures, and data-collection designs. We introduce the Practical Identifiability Index (PII), a unit-free, scale-invariant measure of uncertainty for individual positive-valued parameters, defined as the base-10 logarithm of the ratio between the upper and lower bounds of a confidence interval; a value of 1 corresponds to a tenfold ratio between the interval bounds. We study PII using parametric bootstrap experiments in growth and compartmental epidemic models. PII decreases as longer data windows are used for model fitting and tends to increase with observation noise (although not uniformly across settings) and when several parameters must be estimated together. PII remains larger for parameters associated with latent or indirectly observed processes, while adding more measured quantities substantially improves estimation of progression and recovery parameters. We then apply PII to daily influenza incidence from the 1918 San Francisco epidemic. As a benchmark, PII and normalized confidence-interval width (NCIW) with equivalent thresholds give the same classification in 134 of 135 model, scenario, parameter, and error-structure combinations (99.3% agreement), while PII expresses uncertainty directly as a multiplicative range, without normalization by the parameter estimate. PII is meant to complement, rather than replace, other identifiability and uncertainty analyses, and should be considered together with empirical coverage and other diagnostic measures.

q-bio.QM↗

A likelihood-based framework for simultaneously learning both noise and growth dynamics using biologically-informed neural networks

In recent years, neural ordinary differential equation frameworks such as Biologically-Informed Neural Networks (BINNs) have shown promise for learning mechanistic laws from sparse data. However, most existing approaches implicitly assume homoscedastic Gaussian noise, and therefore do not account for potentially meaningful structure in biological variability. Here, we present an extension to the existing BINNs framework that includes a learnable noise model, allowing discovery of the noise model directly from data. Using population growth as an example, we demonstrate that the framework accurately recovers the underlying noise structure and improves predictions of the underlying growth laws compared to existing approaches. As such, this work establishes a general likelihood-based framework for jointly learning dynamics and heteroscedastic noise within mechanistic neural network approaches.

q-bio.QM↗

Reliable mechanistic operator recovery with biologically-informed neural networks: principles for architecture and optimisation design

Many biological processes are governed by complex dynamical mechanisms that remain incompletely understood despite increasing volumes of experimental data. Biologically-informed neural networks (BINNs) seek to address this challenge by embedding differential equations into neural network training, enabling constitutive operators to be recovered directly from sparse and noisy observations. However, the extent to which operator recovery depends on architectural design, optimisation strategy and the information within the data is not yet well understood. We present an empirical study of how these factors influence mechanistic inference using BINNs applied to one-dimensional advection-diffusion-reaction partial differential equations. Across a suite of problems, we investigate how network expressivity, learning rate, loss weighting and batch size influence optimisation behaviour, reconstruction accuracy and operator recovery. We show that mechanistic inference is governed by balancing competing objectives rather than maximising any single aspect. Moderately expressive architectures outperform complex networks, intermediate learning rates balance efficient exploration with optimisation stability, accurate operator recovery requires a balance between data-fitting and PDE residual losses and intermediate batch sizes provide the best compromise between efficient parameter space exploration, computational efficiency and reproducibility. We further identify practical diagnostics for recognising common failure modes, including over-fitting, unstable optimisation and poor mechanistic recovery. These findings establish guidelines for deploying BINNs as credible tools for biological model discovery and demonstrate that reliable mechanistic inference is achieved by appropriately balancing model expressivity, optimisation, physical consistency and data informativeness.

q-bio.QM↗