arXiv · 2509.16598
PruneCD: Contrasting Pruned Self Model to Improve Decoding Factuality
Abstract
To mitigate the hallucination problem in large language models, DoLa exploits early exit logits from the same model as a contrastive prior. However, we found that these early exit logits tend to be flat, low in magnitude, and fail to reflect meaningful contrasts. To address this, we propose PruneCD, a novel contrastive decoding method that constructs the amateur model via layer pruning rather than early exit. This design leads to more informative and well-aligned logits, enabling more effective contrastive decoding. Through qualitative and quantitative analyses, we demonstrate that PruneCD consistently improves factuality with minimal inference overhead, offering a robust and practical approach to mitigating hallucinations in LLMs.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Byeongho Yu, Changhun Lee, Jungyu Jin, Eunhyeok Park. 2025-09-20. PruneCD: Contrasting Pruned Self Model to Improve Decoding Factuality. https://arxiv.org/abs/2509.16598
Cite the original work for its findings. Save a collection to share your selection of sources.