OpenAI has built a text watermarking method to detect chatgpt written content

@[email protected] · 6 months ago

OpenAI has built a text watermarking method to detect chatgpt written content

@brucethemoose · edit-2 6 months ago

You have full control of your logit outputs with local LLMs, so theoretically you could “unscramble” them. And any finetuning would just blow that bias away anyway.

OpenAI (IIRC) very notably stopped giving the logprobs of their models. They did this for many reasons, and most of them boil down to “profits” and “they are anticompetitive jerks,” but another reason is to enable watermark methods just like this.

Also, thing about this is that basically no one uses self hosted LLMs compared to OpenAI (or really any API) LLM.

OpenAI has built a text watermarking method to detect chatgpt written content

OpenAI has built a text watermarking method to detect chatgpt written content

OpenAI has built a text watermarking method to detect ChatGPT-written content — company has mulled its release over the past year