arXiv · 2501.12174
BiMarker: Enhancing Text Watermark Detection for Large Language Models with Bipolar Watermarks
Abstract
The rapid growth of Large Language Models (LLMs) raises concerns about distinguishing AI-generated text from human content. Existing watermarking techniques, like \kgw, struggle with low watermark strength and stringent false-positive requirements. Our analysis reveals that current methods rely on coarse estimates of non-watermarked text, limiting watermark detectability. To address this, we propose Bipolar Watermark (\tool), which splits generated text into positive and negative poles, enhancing detection without requiring additional computational resources or knowledge of the prompt. Theoretical analysis and experimental results demonstrate \tool's effectiveness and compatibility with existing optimization techniques, providing a new optimization dimension for watermarking in LLM-generated content.
Explore related subjects
Keep this discovery
Zhuang Li, Qiuping Yi, Zongcheng Ji, Yijian Lu, Yanqi Li, Keyang Xiao, Hongliang Liang. 2025-01-21. BiMarker: Enhancing Text Watermark Detection for Large Language Models with Bipolar Watermarks. https://arxiv.org/abs/2501.12174
Cite the original work for its findings. Save a collection to share your selection of sources.