this post was submitted on 28 Jun 2024
1 points (100.0% liked)
Technology
59566 readers
3371 users here now
This is a most excellent place for technology news and articles.
Our Rules
- Follow the lemmy.world rules.
- Only tech related content.
- Be excellent to each another!
- Mod approved content bots can post up to 10 articles per day.
- Threads asking for personal tech support may be deleted.
- Politics threads may be removed.
- No memes allowed as posts, OK to post as comments.
- Only approved bots from the list below, to ask if your bot can be added please contact us.
- Check for duplicates before posting, duplicates may be removed
Approved Bots
founded 1 year ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
If the model isn't overfitted it's also not even copying. By their nature LLMs are transformative which is the whole point of fair use.
So I have a LLM read a book and paraphrase its contents, that's not stealing?
copyright laws are broken. what seems ethical can be illegal and what seems unethical can be legal.
Have I just stolen The Hitchhikers Guide to the Galaxy and given it to you?
You've probably not infringed the copyright, only the court can decide though; if you were to be challenged by the rights holder.
I think there are lots of factors in your defence:
But add in some more quotes, flesh it out, and then try to sell it . . . each step weakens the 'fair use' defence.
This the the problem for the LLM, it can be used for many things, and if it has no filter or limit, then eventually the collective derived works might add up to commercial, substantial reuse, and might include enough to have copied a substantial portion of the original. Very hard to determine I'd think. Each individual use might be fair, but did the LLM itself go too far at some point?
Copyright holder probably struggles to challenge the LLM on the basis of all the things infinite mokeys might use it for in future.
Again, even an exact copy is not stealing. It's copyright infringement. Theft is a different crime.
But paraphrasing is not copyright infringement either. It's no different than Wikipedia having a synopsis for every single episode of a TV series. Telling someone about what a work contains for informational purposes is perfectly fine.