Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

What happens if we train models (GPT or human students) using the outputs of a model with text havingbthose watermarks? Is there something preventing the watermark from being learnable?


no, and in fact they should be extremely easy to detect if they are substituting a token that is a least-probable synonym




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: