arXiv · 2312.17295
Optimizing watermarks for large language models
Abstract
With the rise of large language models (LLMs) and concerns about potential misuse, watermarks for generative LLMs have recently attracted much attention. An important aspect of such watermarks is the trade-off between their identifiability and their impact on the quality of the generated text. This paper introduces a systematic approach to this trade-off in terms of a multi-objective optimization problem. For a large class of robust, efficient watermarks, the associated Pareto optimal solutions are identified and shown to outperform the currently default watermark.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Bram Wouters. 2023-12-28. Optimizing watermarks for large language models. https://arxiv.org/abs/2312.17295
Cite the original work for its findings. Save a collection to share your selection of sources.