Copyright Traps for LLMs
paper Your tags
Your notes
The canonical trap-sequence methodology for detecting use of copyrighted content in LLM training data (ICML 2024, AISP/de Montjoye with CentraleSupélec) — inject unique sequences, then test membership post-hoc. Policy-relevant and the opening move of the lab's memorization-audit line. Dated by venue.