Searcharxiv⌕ Search

arXiv subjects

Guanghua Wang

Publications and source records attributed to Guanghua Wang.

3 recordsLinked to original sources

Decoupling Generation and Selection for Budget-Constrained Faithful Summarization

Abstractive summarization models remain vulnerable to factual inconsistency, redundancy, and weak length control. We propose a modular generation-and-selection framework for sentence-budget-constrained summarization. A pretrained generator produces multiple candidate summaries, which are decomposed into sentence-level candidates. A combinatorial selector then constructs the final summary by balancing relevance, factuality, and redundancy under an explicit budget. The framework supports MMR, ILP, and a DPP-inspired log-determinant objective without retraining the generator. Experiments on CNN/DailyMail, Multi-News, FaithBench, and TofuEval show consistent improvements in factuality and source-grounding metrics, especially for multi-document summarization, at the cost of lower reference-overlap scores. Human evaluation further indicates higher perceived consistency, relevance, clarity, and conciseness, with a small reduction in coherence. These results show that decoupling generation from selection provides a model-agnostic mechanism for improving factual grounding. Code is available at https://anonymous.4open.science/r/bcfs-D05E/.

cs.CL↗

Surveying the Landscape of Text Summarization with Deep Learning: A Comprehensive Review

In recent years, deep learning has revolutionized natural language processing (NLP) by enabling the development of models that can learn complex representations of language data, leading to significant improvements in performance across a wide range of NLP tasks. Deep learning models for NLP typically use large amounts of data to train deep neural networks, allowing them to learn the patterns and relationships in language data. This is in contrast to traditional NLP approaches, which rely on hand-engineered features and rules to perform NLP tasks. The ability of deep neural networks to learn hierarchical representations of language data, handle variable-length input sequences, and perform well on large datasets makes them well-suited for NLP applications. Driven by the exponential growth of textual data and the increasing demand for condensed, coherent, and informative summaries, text summarization has been a critical research area in the field of NLP. Applying deep learning to text summarization refers to the use of deep neural networks to perform text summarization tasks. In this survey, we begin with a review of fashionable text summarization tasks in recent years, including extractive, abstractive, multi-document, and so on. Next, we discuss most deep learning-based models and their experimental results on these tasks. The paper also covers datasets and data representation for summarization tasks. Finally, we delve into the opportunities and challenges associated with summarization tasks and their corresponding methodologies, aiming to inspire future research efforts to advance the field further. A goal of our survey is to explain how these methods differ in their requirements as understanding them is essential for choosing a technique suited for a specific setting.

cs.CL↗

Gas permeation through graphdiyne-based nanoporous membranes

Nanoporous membranes based on two dimensional materials are predicted to provide highly selective gas transport in combination with extreme permeability. Here we investigate membranes made from multilayer graphdiyne, a graphene-like crystal with a larger unit cell. Despite being nearly a hundred of nanometers thick, the membranes allow fast, Knudsen-type permeation of light gases such as helium and hydrogen whereas heavy noble gases like xenon exhibit strongly suppressed flows. Using isotope and cryogenic temperature measurements, the seemingly conflicting characteristics are explained by a high density of straight-through holes (direct porosity of ~0.1%), in which heavy atoms are adsorbed on the walls, partially blocking Knudsen flows. Our work offers important insights into intricate transport mechanisms playing a role at nanoscale.

cond-mat.mes-hall↗