A new research paper laid out ways in which AI developers should try and avoid showing LLMs have been trained on copyrighted material.

OpenAI now tries to hide that ChatGPT was trained on copyrighted books, including J.K. Rowling’s Harry Potter series::A new research paper laid out ways in which AI developers should try and avoid showing LLMs have been trained on copyrighted material.

@fubo@lemmy.world
link
fedilink
English
50
edit-2
1Y

If I memorize the text of Harry Potter, my brain does not thereby become a copyright infringement.

A copyright infringement only occurs if I then reproduce that text, e.g. by writing it down or reciting it in a public performance.

Training an LLM from a corpus that includes a piece of copyrighted material does not necessarily produce a work that is legally a derivative work of that copyrighted material. The copyright status of that LLM’s “brain” has not yet been adjudicated by any court anywhere.

If the developers have taken steps to ensure that the LLM cannot recite copyrighted material, that should count in their favor, not against them. Calling it “hiding” is backwards.

@StrongFox@lemmy.world
link
fedilink
English
31Y

you bought the book to memorize from, anyway.

@Agent641@lemmy.world
link
fedilink
English
41Y

No, I shoplifted it from an Aldi

Another sensationalist title. The article makes it clear that the problem is users reconstructing large portions of a copyrighted work word for word. OpenAI is trying to implement a solution that prevents ChatGPT from regurgitating entire copyrighted works using “maliciously designed” prompts. OpenAI doesn’t hide the fact that these tools were trained using copyrighted works and legally it probably isn’t an issue.

@khalic@lemmy.world
link
fedilink
English
-21Y

An LLM is not a brain, stop anthropomorphising a fkn vector solver… it’s math, there’s nothing alive about it

Jilanico
link
fedilink
English
01Y

What if you are just a vector solver but don’t realize it? We wouldn’t know we have neurons in our heads if scientists didn’t tell us. What even is consciousness?

@khalic@lemmy.world
link
fedilink
English
21Y

All excellent questions, we need the answer to that. Until then, we don’t know, and can’t make up stuff just because we don’t.

Create a post

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


  • 1 user online
  • 196 users / day
  • 589 users / week
  • 1.38K users / month
  • 4.49K users / 6 months
  • 1 subscriber
  • 7.41K Posts
  • 84.7K Comments
  • Modlog