Technology

59708 readers

1807 users here now

This is a most excellent place for technology news and articles.

Our Rules

Follow the lemmy.world rules.
Only tech related content.
Be excellent to each another!
Mod approved content bots can post up to 10 articles per day.
Threads asking for personal tech support may be deleted.
Politics threads may be removed.
No memes allowed as posts, OK to post as comments.
Only approved bots from the list below, to ask if your bot can be added please contact us.
Check for duplicates before posting, duplicates may be removed

Approved Bots

founded 1 year ago

MODERATORS

[email protected]

335

OpenAI has a 99.9% accurate ChatGPT AI text detector, but won't release it. (the-decoder.com)

submitted 3 months ago by ModerateImprovement to c/[email protected]

68 comments fedilink hide all child comments

you are viewing a single comment's thread
view the rest of the comments

[–] [email protected] 16 points 3 months ago* (last edited 3 months ago) (1 children)

That's a bad article. What are they reluctant about? Releasing that detector, or applying watermarks to the generated texts? Do they do that already or doesn't it apply to text generated until then? And how would that affect anything else?

Whats with the error rate? Shouldn't that be near 100% for watermarks? And 0 false positives? What's really holding them back? Is pupils not turning in ChatGPT homework anymore cutting into their business model?

I mean all the major AI companies promised to do AI ethically. Now they don't want the one thing that would solve half the issues people are having with that technology. Kind of fits with OpenAI 🤔

[–] [email protected] 3 points 3 months ago (1 children)

They can't release anything as watermarks can be reverse engineered and people would just wise up and tumble the outputs.

Weirdly, not releasing this tool publicly might be the smartest bet here as all of these bot farms and idiots just blindly use chatgpt outputs without any tumbling or safety.

[–] [email protected] 1 points 3 months ago* (last edited 3 months ago) (1 children)

The issue with that is: Releasing nothing is even worse than releasing something that could be circumvented. I don't see this as a valid argument.

I'm not an expert on text watermarking and how that degrades output. But if they want some stealthy solution that isn't known to the public... Maybe they could attach two watermarks. A simple one that is known to everyone, and an additional, secret one only they know about. It'd be similar to what we do with bank notes. There are some characteristics everyone knows and can use to judge if it's fake money. And they have some additional secret markings in banknotes that only the central bank knows about.

I'm pretty sure a similar thing could be done here. Maybe not for a 280 character tweet. But certainly for other use-cases with longer texts. And in case it has a 0% false positive rate, every match helps someone. Even if it's circumventable. I think even a non-perfect solution that helps several thousands of people is better than helping no-one.

[–] Pika 2 points 3 months ago (1 children)

I agree with not releasing it, but I do find that it defeats the purpose talking about it because if you have it but aren't sharing if what's the point of having it

[–] [email protected] 1 points 3 months ago

I think we're missing half the story. Because I also fail so see a point in doing it like they do.