Skip to content
arXiv stat.ML · Papers

Sample Complexities of Estimating Gumbel–Max Watermark Proportions with and without Reduction to Pivotal Statistics

arXiv:2607.00224v1 Announce Type: cross Abstract: Watermarking promises a statistical trace of large language model (LLM) use, but real documents, after editing or paraphrasing, rarely arrive as purely human-written or purely machine-generated. This motivates a quantitative question beyond detection: what proportion of