this post was submitted on 25 Feb 2024
319 points (87.2% liked)

Technology

58369 readers
4274 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


founded 1 year ago
MODERATORS
 

A new report from plagiarism detector Copyleaks found that 60% of OpenAI's GPT-3.5 outputs contained some form of plagiarism.

Why it matters: Content creators from authors and songwriters to The New York Times are arguing in court that generative AI trained on copyrighted material ends up spitting out exact copies.

you are viewing a single comment's thread
view the rest of the comments
[–] kromem@lemmy.world 43 points 7 months ago (2 children)

"Plagiarism detection company claims LLM conditions plagiarism according to their detector."

I wonder how many student written essays also contain 'plagiarism' according to their tool.

[–] Anamnesis@lemmy.world 5 points 7 months ago (1 children)

Probably very few. The bias for these companies is in false negatives, not false positives, since false positives create controversy when students appeal a ruling.

[–] General_Effort@lemmy.world 2 points 7 months ago

The bias here was certainly to come up with a lot of false positives for advertising; kinda like anti-virus companies do it.

[–] CheeseNoodle@lemmy.world 4 points 7 months ago

100% iirc, there are only so many ways to write about how the blue curtains indicate the character is feeling depressed or something.