SearcharxivSearch

arXiv subjects

Zhipeng Song

Publications and source records attributed to Zhipeng Song.

6 recordsLinked to original sources

Curvature-driven revival of charge density waves in non-Euclidean space

Strongly correlated quantum states, such as charge density waves (CDWs), are exquisitely sensitive to Fermi surface topology and lattice symmetry, and are typically quenched by heavy carrier doping. In two-dimensional (2D) systems, however, macroscopic geometric curvature emerges as a novel structural degree of freedom to modulate microscopic quantum coherence. This raises a compelling physical question: can non-Euclidean geometric deformations compete with extreme electronic perturbations to reshape, or even revive, a quenched macroscopic quantum order? Here, by constructing monolayer TiSe$_2$-NbSe$_2$ heterostructure on a BLG/SiC substrate for the first time, we report the curvature-driven revival of a frustrated charge order in a non-Euclidean space. Low-temperature angle-resolved photoemission spectroscopy (ARPES) reveals a massive interfacial charge transfer, which destroys the global Fermi surface nesting and completely suppresses the long-range CDW order in Euclidean flat regions. Strikingly, high-resolution scanning tunneling microscopy (STM) reveals that a novel, non-linear CDW state miraculously survives, remaining strictly localized within morphologically distorted, non-Euclidean nanoscale curved regions. Atomistic simulations unravel the structural origin of this phenomenon, demonstrating that interfacial twist and lattice mismatch spontaneously generate a corrugated superlattice.

cond-mat.mtrl-sci

CAR: Query-Guided Confidence-Aware Reranking for Retrieval-Augmented Generation

Retrieval-augmented generation (RAG) relies on evidence ranking to determine what information is exposed to the generator, yet existing retrieval and reranking methods primarily estimate query--document relevance. Relevance, however, is not equivalent to generator-side usefulness: a relevant passage may introduce ambiguity or distraction, whereas a lower-ranked passage may stabilize the generator's answer. We present CAR (Confidence-Aware Reranking), a training-free rank-correction framework that uses query-only answer stability as a control and measures each candidate by the change it induces in sampled-answer semantic stability. This controlled contrast estimates a document's marginal contribution to generator behavior without treating semantic stability as relevance or calibrated correctness. CAR converts these confidence changes into coarse precedence constraints and returns the feasible ranking with minimum Kendall distance from the baseline, preserving existing pairwise preferences unless generator-side evidence supports reversing them. Experiments on NQ, HotpotQA and FEVER across sparse and dense retrievers, seven ranking methods and three generator families show robust improvements. In the BM25-centered main analysis, CAR achieves a \textbf{+5.53\% mean relative NDCG@5 gain}; on the fixed NQ-answerable downstream evaluation, it improves token-level F1 by \textbf{+0.43 points}, with ranking and generation gains strongly aligned across rankers ($\rho = 0.93$). These results position CAR as a deployment-friendly, generator-aware correction layer that complements relevance while preserving informative prior rankings. CAR requires neither task-specific training nor access to model internals such as logits or hidden states, making it applicable to black-box LLMs through generated outputs alone.

cs.CL

LLM-Confidence Reranker: A Training-Free Approach for Enhancing Retrieval-Augmented Generation Systems

Large language models (LLMs) have revolutionized natural language processing, yet hallucinations in knowledge-intensive tasks remain a critical challenge. Retrieval-augmented generation (RAG) addresses this by integrating external knowledge, but its efficacy depends on accurate document retrieval and ranking. Although existing rerankers demonstrate effectiveness, they frequently necessitate specialized training, impose substantial computational expenses, and fail to fully exploit the semantic capabilities of LLMs, particularly their inherent confidence signals. We propose the LLM-Confidence Reranker (LCR), a training-free, plug-and-play algorithm that enhances reranking in RAG systems by leveraging black-box LLM confidence derived from Maximum Semantic Cluster Proportion (MSCP). LCR employs a two-stage process: confidence assessment via multinomial sampling and clustering, followed by binning and multi-level sorting based on query and document confidence thresholds. This approach prioritizes relevant documents while preserving original rankings for high-confidence queries, ensuring robustness. Evaluated on BEIR and TREC benchmarks with BM25 and Contriever retrievers, LCR--using only 7--9B-parameter pre-trained LLMs--consistently improves NDCG@5 by up to 20.6% across pre-trained LLM and fine-tuned Transformer rerankers, without degradation. Ablation studies validate the hypothesis that LLM confidence positively correlates with document relevance, elucidating LCR's mechanism. LCR offers computational efficiency, parallelism for scalability, and broad compatibility, mitigating hallucinations in applications like medical diagnosis.

cs.CL

Less is More for RAG: Information Gain Pruning for Generator-Aligned Reranking and Evidence Selection

Retrieval-augmented generation (RAG) grounds large language models with external evidence, but under a limited context budget, the key challenge is deciding which retrieved passages should be injected. We show that retrieval relevance metrics (e.g., NDCG) correlate weakly with end-to-end QA quality and can even become negatively correlated under multi-passage injection, where redundancy and mild conflicts destabilize generation. We propose \textbf{Information Gain Pruning (IGP)}, a deployment-friendly reranking-and-pruning module that selects evidence using a generator-aligned utility signal and filters weak or harmful passages before truncation, without changing existing budget interfaces. Across five open-domain QA benchmarks and multiple retrievers and generators, IGP consistently improves the quality--cost trade-off. In a representative multi-evidence setting, IGP delivers about +12--20% relative improvement in average F1 while reducing final-stage input tokens by roughly 76--79% compared to retriever-only baselines.

cs.CL

Shifted wave equation on noncompact symmetric spaces

Let $G$ be a semisimple, connected, and noncompact Lie group with a finite center. We carry out a detailed analysis of oscillating integrals involving the Harish-Chandra $c$-function, in the case of real rank $l\ge 2$. This allows to obtain two main applications. Consider the Laplace-Beltrami operator $\Delta$ on the homogeneous space $G/K=S$ by a maximal compact subgroup $K$. We obtain pointwise estimates for the kernel of an oscillating function $\exp( it\sqrt{|x|}) \psi(\sqrt{|x|}) $ applied to the shifted Laplacian $\Delta+|\rho|^2$. We obtain a polynomial decay in time of the kernel, and of the $L^p-L^q$ norms of the operator, for $1\le p<2<q\le \infty$. For the related distinguished Laplacian, we obtain bounds for the $L^p-L^p$ norms, $1\le p\le\infty$, with a slower growth in time than predicted by earlier results.

math.AP

Pointwise and uniform bounds for functions of the Laplacian on non-compact symmetric spaces

Let $L$ be the distinguished Laplacian on the Iwasawa $AN$ group associated with a semisimple Lie group $G$. Assume $F$ is a Borel function on $\mathbb{R}^+$. We give a condition on $F$ such that the kernels of the functions $F(L)$ are uniformly bounded. This condition involves the decay of $F$ only and not its derivatives. By a known correspondence, this implies pointwise estimates for a wide range of functions of the Laplace-Beltrami operator on symmetric spaces. In particular, when $G$ is of real rank one and $F(x)={\rm e}^{it\sqrt x}\psi(\sqrt x)$, our bounds are sharp.

math.AP