Searcharxiv⌕ Search

arXiv · 2610.01891

RipplePLM: Structural and Property Decoupling for Protein Mutation Effect Generation

Abstract

Protein mutation effect generation asks a model to describe the functional consequence of a point mutation in natural language. Existing protein-to-text systems typically encode mutation information into undifferentiated representations, overlooking the organization of mutation-induced evidence across structural and biochemical factors. We propose RipplePLM, a mutation-aware generation framework centered on Direct-Distal Cross-Attention (DDCA). By constructing a residue-level Mutation Perturbation Field from pre-trained protein language models, DDCA leverages predicted contact maps to organize mutation representations into two pathways: the mutation site's immediate contact neighborhood and its multi-hop distal context. To complement this structural decomposition, we further introduce the Property Latent Chain (PLChain), which injects expert-guided supervision of biochemical property changes (e.g., thermostability and optimal pH) into the LLM hidden-state pathway through latent property tokens. On MutaDescribe, RipplePLM improves over mutation-specific baselines on temporal and structural splits; under a matched-backbone comparison, average structural-split ROUGE-L increases from {22.23} to {35.65}. Expert evaluation further shows a higher proportion of biologically accurate or relevant descriptions than the mutation-specific baseline. Additional ablations, representation diagnostics, and low-$N$ fitness regression experiments further support the effectiveness of the learned mutation-aware representations. Code: https://github.com/Lyu6PosHao/RipplePLM.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Liuzhenghao Lv, Yuyang Liu, Yuyang Gao, Li Yuan, Yonghong Tian. 2026-10-01. RipplePLM: Structural and Property Decoupling for Protein Mutation Effect Generation. https://arxiv.org/abs/2610.01891

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

How Evaluation Choices Change the Measured Benefit of Cooperative Perception: Evidence from Three V2X Benchmarks

Cooperative perception, in which connected vehicles and roadside infrastructure share sensor information, is a candidate enabler of automated mobility, and benchmark accuracy is the evidence cited when roadside deployment is considered. This paper audits that evidence base across one simulated and two real-world vehicle-to-everything (V2X) benchmarks. In simulation, two widely studied robustness axes leave almost no recoverable headroom: an infrastructure-anchored pose correction returns about one accuracy point at every error level, and corrupting a partner costs 0.7 points. On real data, measurement choices govern the conclusion. An apparent seventeen-fold advantage of infrastructure in partner-poor frames falls below four-fold once the split is broadened, and an ad-hoc class definition measures a far smaller benefit than the official protocol reports. Cooperation is worth 7 to 15 accuracy points, yet 27% of frames on one split offer no partner and almost none do on another, so every regime-conditioned claim must name its split.

cs.CE↗

Names without information: Attention allocation and rent transfer in a zero-fundamental token market

Asset names are associated with investor trading and asset prices, but where names and issuer quality are formed jointly, information, preference and attention-coordination explanations are difficult to distinguish. We study a token launchpad on which tokens issued under the default template use the same contract code, have a fixed supply and carry no cash flows, while names can be registered at almost no cost, cannot be verified and may be reused. Using all 458,174 default-template tokens launched from February to June 2026 and a sample split by creator entity fixed in advance, we find that the share of tokens attracting an outside buyer rises from 34.9% to 52.8% across name-appeal deciles, but within creator and creation minute the effect of appeal is small and does not replicate in the validation sample. Conditional on early capital inflow and creator fixed effects, the residual effects of names on migration, peak market capitalization and post-peak drawdown are statistically equivalent to zero, and buyers who follow appealing names earn no economically meaningful premium. The same name attracts less capital with each reuse, names whose first token draws more capital are copied sooner, and competition over names mainly transfers wealth among buyers.

cs.CE↗

From language-model stock rankings to testable economic rules: A computational audit

We test the stability, reproducibility and investment outcomes of language-model stock rankings. Four models and five numerical comparators share a portfolio engine over 72 monthly holding periods in the Shanghai Stock Exchange (SSE) 50, China Securities Index (CSI) 300 and CSI 500. Rankings use nine characteristics, and five repeated SSE 50 runs measure variation under identical inputs. Linear rules fitted to development-period model preferences are frozen before unseen-month, larger-pool and controlled-intervention tests. Their mean Spearman agreement with model rankings is 0.923-0.984 in the SSE 50 and 0.795-0.985 after transfer. Aggregate rank-change error falls relative to a zero-change prediction in twelve archived feature-group comparisons and eight matched single-feature comparisons, with Holm adjustments applied in separate nine-plus-three and six-plus-two families. Prediction of individual entries and exits remains weak (event Jaccard 0.000-0.125). Historical mean model compound annual growth rates range from 6.26% to 11.17%. At 10 basis points per side and six-month blocks, the twelve-comparison model-minus-rule return family and factor-controlled associations yield no adjusted finding. Three higher-cost, twelve-month-block comparisons favor a Terra rule within their twelve-test slices, with no adjusted finding across the full 144-test sensitivity grid. Two input-intervention batches totaling 5,184 responses supply paired intervention-return tests. The three-model batch has no bootstrap-adjusted finding at the primary block length; a Luna row-order effect appears under heteroskedasticity- and autocorrelation-consistent (HAC) adjustment within its three-test family but not in a pooled 24-test adjustment. Compact rules approximate aggregate rankings; return conclusions depend on comparison families, uncertainty methods and tie priorities.

cs.CE↗