Doesn’t OpenAI’s watermarking affect the quality of the models?

Wait 5 sec.

OpenAI just announced that they will start to apply watermarking on their model’s output text and images to comply with the EU regulation that wants to be able to identify whether a text or an image was produced by an AI model. Anthropic announced the same thing a while ago (they applied worldwide, not just in EU). The way they do that as they explained is by enforcing a “statistical signal in text generated”, meaning preferring not always the most appropriate next token but close enough, in order to meet the “statistical signal” requirement. In my understanding, this deteriorates the output quality of their models, as it introduces KLD>0. And we know that any KLD divergence greater than 0 (which the watermarking certainly creates) may be negligible in small outputs, but it definitely becomes noticeable in multi-turn tasks due to the compounding effect. What do you think?   submitted by   /u/ResearchCrafty1804 [link]   [comments]