LR-DWM: Efficient Watermarking for Diffusion Language Models

Watermarking (WM) is a critical mechanism for detecting and attributing AI-generated content. Current WM methods for Large Language Models (LLMs) are predominantly tailored for autoregressive (AR) models: They rely on tokens being generated sequentially, and embed stable signals within the generated sequence based on the previously sampled text. Diffusion Language Models (DLMs) generate text via non-sequential iterative denoising, which requires significant modification to use WM methods designed for AR models. Recent work proposed to watermark DLMs by inverting the process when needed, but suffers significant computational or memory overhead. We introduce Left-Right Diffusion Watermarking (LR-DWM), a scheme that biases the generated token based on both left and right neighbors, when they are available. LR-DWM incurs minimal runtime and memory overhead, remaining close to the non-watermarked baseline DLM while enabling reliable statistical detection under standard evaluation settings. Our results demonstrate that DLMs can be watermarked efficiently, achieving high detectability with negligible computational and memory overhead.

Paper

References (10)

072023. Variance-reduced watermarking for large language modelsAdvances in Neural Information Processing Systems
082024. Discrete diffusion modeling by estimating the ratio of the data distribution to the noise distributionProceedings of the 41st International Conference on Machine Learning
092024. Qwen2.5: Technical reportarXiv preprint
102025. Large language diffusion modelsarXiv preprint

Similar papers

© 2026 NYSGPT2525 LLC