Our approach to EU text provenance rules

openai.com

54 points by tosh 6 hours ago


mgax - 4 hours ago

This is such a waste of time. If someone wants to bypass this it will be rather simple. Just change the words. If someone wants to avoid fingerprinting they will. Can’t we just focus on building rather than spending brainpower on these ridiculous sidequests

m-hodges - 4 hours ago

> Starting today, API customers globally will be able to opt in to text watermarking for select models. Text watermarking will remain off by default in the API.

> Over the coming weeks, we will add an invisible watermark to eligible ChatGPT and Codex text output in the European Union.

pembrook - 7 minutes ago

Good to know, will exclusively move to Chinese models for non-coding tasks.

The idea that producing text with AI needs to be watermarked as if it's a crime by default is backwards nonsense.

To me this would actually be a counter signal.

If you're NOT primarily writing with AI (at least mildly being informed by all of human knowledge distilled), then I will assume your ideas are emotional opinion-based nonsense, like most comments on hackernews, including my own.

athrowaway3z - 3 hours ago

>> Editing can weaken the watermark. In an evaluation of 400-token passages, replacing 10% of words with synonyms reduced detection from about 92% to 66%. Replacing 25% of words reduced it to 17%.

> Claude/codex/deepseek, please replace 25% of words with synonyms or slight rephrasing because i dont like the current version.

Not sure if that counts as: `a solution that's robust against "common alterations and adversarial attacks"`. Is there a sort of adversarial attack that is more common?

smokel - 4 hours ago

Why use watermarking, and not simply add a signature?

k__ - 3 hours ago

Is watermarking part of a model architecture or is it something added by the inference engine?

greatgib - 3 hours ago

My personal opinion is that they cheated evaluations to be able to release this pretending that it has no meaningful impact.

Otherwise, I don't see any logical explanation that some of their benchmark results would be higher when watermarked. Except if benchmark results are so unstable that they are an useless metric.

richwater - 3 hours ago

Just make the models worse for the EU. Don't accept this nonsense that's holdingg back actual work and progress.

aenis - 3 hours ago

Another cookie consent-grade success of the EU.