Enhancing LLM Watermark Resilience Against Both Scrubbing and Spoofing Attacks
NeurIPSSpotlight2025
TL;DR
Watermarking is widely regarded as a promising defense against the misuse of large language models (LLMs); however, existing methods are fundamentally constrained by their vulnerability to scrubbing a…
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
large language model language model llm
← All NeurIPS 2025 Spotlight papers · Browse the whole archive