SearcharxivSearch

arXiv subjects

Yuanhao Wu

Publications and source records attributed to Yuanhao Wu.

8 recordsLinked to original sources

An in situ self-adaptive hydrogel coating enables seamless neural interfaces via okra mucilage polysaccharide and α-helical peptide amphiphiles co-assembly

Long-term stability of neural interfaces is frequently compromised by mechanical mismatch and chronic neuroinflammation, often leading to electrode detachment and signal failure. While hydrogel coatings offer a solution, conventional designs typically rely on exogenous conductive fillers that can sacrifice mechanical flexibility or induce toxicity. Here, we report on a soft neural interface based on the supramolecular co-assembly of a renewable natural polysaccharide, okra mucilage polysaccharide (OMP), and an α-helical peptide amphiphiles (APA). The resulting OMP-APA hydrogel (OP gel) exhibits environment-responsive enhancements in bioadhesion and charge-transport capability triggered by physiological pH and electrical stimulation. These properties arise from intrinsic, stimulus-responsive alterations in fibre architecture and orientation, eliminating the need for conductive fillers. Leveraging interfacial liquid-liquid phase separation, we demonstrate the in situ coating of ultra-thin OP-gel coating onto carbon fibre electrodes (CFE). The OP-gel-coated electrodes (OP-CFE) significantly mitigate foreign body responses and glial scarring, enabling stable, high-quality neural recordings in a mouse cortical in vivo model. Our findings provide a versatile strategy for constructing seamless, multifunctional bio-interfaces through supramolecular co-assembly, with broad implications for advancing neural prosthetics and neuroscience research.

physics.bio-ph

OpenGenAlign: A Preference Dataset and Benchmark for Trustworthy Reward Modeling in Open-Ended, Long-Context Generation

Reward Modeling is critical in evaluating and improving the generation of Large Language Models (LLMs). While numerous recent works have shown its feasibility in improving safety, helpfulness, reasoning, and instruction-following ability, its capability and generalization to open-ended long-context generation is still rarely explored. In this paper, we introduce OpenGenAlign, a framework and a high-quality dataset designed to develop reward models to evaluate and improve hallucination-free, comprehensive, reliable, and efficient open-ended long-context generation. We define four key metrics to assess generation quality and develop an automated pipeline to evaluate the outputs of multiple LLMs across long-context QA, Data-to-Text, and Summarization scenarios using o3, ending up with 33K high-quality preference data with a human agreement rate of 81\%. Experimental results first demonstrate that existing reward models perform suboptimally on the held-out benchmark. And Our trained reward model achieves superior performance in the benchmark and effectively improves the generation quality of the policy models using Reinforcement Learning (RL). Additionally, OpenGenAlign could be used for effective guided generation in existing datasets. Furthermore, we demonstrate that the OpenGenAlign could be integrated with reward data from other domains to achieve better performance.

cs.CL

A dispersal recolonisation 3D biofilm in vitro model based on co-assembled peptide amphiphiles and clinical wound fluid

Chronic wound infections are sustained by dynamic 3D biofilm cycles involving maturation, dispersal, and recolonisation, yet existing in vitro models fail to reproduce these temporal and structural complexities. Here, we report a strategy that co-assembles a designed protease-inhibitory peptide amphiphile (PA-GF) with patient-derived wound fluid (WF) to reconstruct the complete biofilm life cycle in vitro. The PA-GF sequence incorporates an HWGF motif capable of binding and inhibiting matrix metalloproteinase-9 (MMP-9), thereby preserving the integrity of recolonised biofilms under proteolytic stress. Co-assembling with WF generated a living material that faithfully mimicked the biochemical and mechanical microenvironment of chronic wounds, supporting the formation of stable 3D biofilms capable of dispersal and recolonisation. Furthermore, we established a controllable polymicrobial infection model and validated its translational relevance through antibiotic susceptibility profiling and spatial microbiological analyses. Notably, the antibiotic response patterns of the PA/WF-derived biofilms closely mirrored those observed in a rat wound infection in vivo model. Collectively, our findings demonstrate that co-assembling living materials can recapitulate the nutritional composition, 3D architecture, and recolonisation dynamics of in vivo infectious biofilms, offering a physiologically relevant and customisable platform for investigating chronic wound infections and accelerating anti-biofilm drug discovery.

physics.med-ph

Mastering energy landscapes via liquid liquid phase separation to program active supramolecular coassembly from the nano to macro scale

The energy landscape dictates pathways and outcomes in supramolecular selfassembly, yet harnessing it from the nano to the macro scales remains a major challenge. Here, we demonstrate liquid liquid phase separation (LLPS) as a powerful tool to navigate and engineer the energy landscapes of coassembly systems comprising disordered proteins and peptides. We quantitatively map the energy barriers and transition states governing structural transitions, enabling predictive on off control of assembly and hierarchical order from nano to macro scales. By integrating supramolecular biofabrication, we achieve spatially organized architectures with life like non equilibrium behaviour. Crucially, assembly stability and scalable selfsorting are shown to depend on accessing minimum energy states, regardless of whether the co assembled structures are disordered or ordered. This work establishes energy landscape mediation via LLPS as a general framework for designing lifelike, hierarchically structured materials.

cond-mat.soft

De novo design of alpha-helical peptide amphiphiles repairing fragmented collagen type I via supramolecular co-assembly

The hierarchical triple-helix structure of collagen type I, Col I, is essential for extracellular matrix support and integrity. However, current reconstruction strategies face challenges such as chain mismatch, preventing proper fibril formation. Here, we report a supramolecular co-assembly strategy using a de novo-designed alpha-helical peptide amphiphile (APA) of just seven amino acids. The APA features a hydrophobic palmitic acid tail, which stabilizes the helical structure and promotes co-assembly upon interaction with complementary molecular structures. This minimal design enables selective recognition of fragmented collagen (FC), restoring triple-helix conformation and guiding fibre formation. We applied this mechanism to engineer FC-rich nanofat (NF) into a mechanically reinforced biomaterial. Integration of APA-NF with coaxial 3D printing enabled spatial control of structure and function. In a porcine model, this platform enhanced in situ vascularized adipose tissue regeneration. Our results demonstrate that hierarchical reconstruction of collagen via peptide-guided supramolecular assembly offers a promising strategy for soft tissue repair.

physics.chem-ph

DuaShepherd: Integrating Stepwise Correctness and Potential Rewards for Mathematical Reasoning

In this paper, we propose DuaShepherd, a novel reward modeling framework that integrates two complementary reward signals, correctness and potential, to enhance the mathematical reasoning capabilities of Large Language Models (LLMs). While correctness-based signals emphasize identification of stepwise errors, potential-based signals focus on the likelihood of reaching the correct final answer. We developed an automated pipeline for constructing large-scale reward modeling dataset with both signals. A unified, multi-head architecture was explored to train the two reward models in a multi-task setup, demonstrating benefits from learning both correctness and potential in parallel. By combining these two signals into a compound probability, our model achieves consistent performance improvements across multiple benchmarks. Empirical evaluations on MATH500 and ProcessBench confirm that this combined reward significantly outperforms models trained on either reward type alone, achieving state-of-the-art performance under comparable resource constraints.

cs.CL

VeraCT Scan: Retrieval-Augmented Fake News Detection with Justifiable Reasoning

The proliferation of fake news poses a significant threat not only by disseminating misleading information but also by undermining the very foundations of democracy. The recent advance of generative artificial intelligence has further exacerbated the challenge of distinguishing genuine news from fabricated stories. In response to this challenge, we introduce VeraCT Scan, a novel retrieval-augmented system for fake news detection. This system operates by extracting the core facts from a given piece of news and subsequently conducting an internet-wide search to identify corroborating or conflicting reports. Then sources' credibility is leveraged for information verification. Besides determining the veracity of news, we also provide transparent evidence and reasoning to support its conclusions, resulting in the interpretability and trust in the results. In addition to GPT-4 Turbo, Llama-2 13B is also fine-tuned for news content understanding, information verification, and reasoning. Both implementations have demonstrated state-of-the-art accuracy in the realm of fake news detection.

cs.CL

RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models

Retrieval-augmented generation (RAG) has become a main technique for alleviating hallucinations in large language models (LLMs). Despite the integration of RAG, LLMs may still present unsupported or contradictory claims to the retrieved contents. In order to develop effective hallucination prevention strategies under RAG, it is important to create benchmark datasets that can measure the extent of hallucination. This paper presents RAGTruth, a corpus tailored for analyzing word-level hallucinations in various domains and tasks within the standard RAG frameworks for LLM applications. RAGTruth comprises nearly 18,000 naturally generated responses from diverse LLMs using RAG. These responses have undergone meticulous manual annotations at both the individual cases and word levels, incorporating evaluations of hallucination intensity. We not only benchmark hallucination frequencies across different LLMs, but also critically assess the effectiveness of several existing hallucination detection methodologies. Furthermore, we show that using a high-quality dataset such as RAGTruth, it is possible to finetune a relatively small LLM and achieve a competitive level of performance in hallucination detection when compared to the existing prompt-based approaches using state-of-the-art large language models such as GPT-4.

cs.CL