SimMark: A Robust Sentence-Level Similarity-Based Watermarking Algorithm for Large Language Models

TOP 文献データベース SimMark: A Robust Sentence-Level Similarity-Based Watermarking Algorithm for Large Language Models

arxiv

AIセキュリティポータルbot

文献データベースの情報は、自動的に収集されています。

Source

https://arxiv.org/abs/2502.02787

PDF

https://arxiv.org/pdf/2502.02787

文献情報

作者: Amirhossein Dabiriaghdam,Lele Wang
公開日: 2025-2-5
更新日: 2025-9-11
所属機関: Department of ECE, University of British Columbia
所属の国: Canada
会議名

AIにより推定されたラベル

透かし設計生成AI向け電子透かしロバスト性分析

※ こちらのラベルはAIによって自動的に追加されました。そのため、正確でないことがあります。
詳細は文献データベースについてをご覧ください。

Abstract

The widespread adoption of large language models (LLMs) necessitates reliable methods to detect LLM-generated text. We introduce SimMark, a robust sentence-level watermarking algorithm that makes LLMs' outputs traceable without requiring access to model internals, making it compatible with both open and API-based LLMs. By leveraging the similarity of semantic sentence embeddings combined with rejection sampling to embed detectable statistical patterns imperceptible to humans, and employing a soft counting mechanism, SimMark achieves robustness against paraphrasing attacks. Experimental results demonstrate that SimMark sets a new benchmark for robust watermarking of LLM-generated content, surpassing prior sentence-level watermarking techniques in robustness, sampling efficiency, and applicability across diverse domains, all while maintaining the text quality and fluency.