Pith. sign in

REVIEW 1 cited by

Improved Unbiased Watermark for Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2502.11268 v3 pith:ZTE3W47R submitted 2025-02-16 cs.CL

classification cs.CL
keywords mcmarkunbiasedlanguagewatermarksai-generateddemonstratedetectabilityexisting
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

As artificial intelligence surpasses human capabilities in text generation, the necessity to authenticate the origins of AI-generated content has become paramount. Unbiased watermarks offer a powerful solution by embedding statistical signals into language model-generated text without distorting the quality. In this paper, we introduce MCmark, a family of unbiased, Multi-Channel-based watermarks. MCmark works by partitioning the model's vocabulary into segments and promoting token probabilities within a selected segment based on a watermark key. We demonstrate that MCmark not only preserves the original distribution of the language model but also offers significant improvements in detectability and robustness over existing unbiased watermarks. Our experiments with widely-used language models demonstrate an improvement in detectability of over 10% using MCmark, compared to existing state-of-the-art unbiased watermarks. This advancement underscores MCmark's potential in enhancing the practical application of watermarking in AI-generated texts.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Can You Detect the Difference?

    cs.CL 2025-07 reject novelty 4.0 of 10

    A 2,000-sample comparison finds diffusion-generated LLaDA text can match human perplexity and burstiness when rephrasing, while LLaMA text is more predictable and easier to flag.

Pith tools