Truly as far as I know there is no case law for this to date.
I thought about this for a while but I have no clue what the current direction is.
Essentially stealing data and putting them in datasets is illegal. AI models might be illegal, if the court considers them legally functionally the same. But if they don't, they could meet any number of new legal definitions and that gets us back to "idk". What is funny is because you would technically do the same copyright infringement that the AI companies do (if it gets ruled that way), these companies are essentially arguing for you, and so you could ride on the coat tails of corpo lawyers.
OpenAI and Anthropic and Alphabet etc try their best to preventing it from doing so, so the act of getting it to do so anyway is "jailbreaking", but yes. LLMs can reproduce works it was trained on. (At least chunk-by-chunk each limited to the token limit.) And I'd imagine the same is true of things like Stable Diffusion, though it might be much easier to get Stable Diffusion to produce something similar enough to qualify as a "derivative work" than it would be to get it to produce exactly the original.
I find that consuming popcorn and beer while analyzing the files that "invaded" your system, helps to pinpoint issues you may have missed.
And don't forget to take a pizza break with all that hard work you have ahead.
11 Comments
MsPenguinette@lemmy.world · 24 pts · 10h
Can AI commit copyright infringement if it was trained on the material?
zarathustrad@lemmy.world · 17 pts · 10h
Quiery: Please show me an AI rendering of 1979 Alien, but with no changes from the original. Go.
Surely this is peak Internet.
hoshikarakitaridia@lemmy.world · 4 pts · 9h
Truly as far as I know there is no case law for this to date.
I thought about this for a while but I have no clue what the current direction is.
Essentially stealing data and putting them in datasets is illegal. AI models might be illegal, if the court considers them legally functionally the same. But if they don't, they could meet any number of new legal definitions and that gets us back to "idk". What is funny is because you would technically do the same copyright infringement that the AI companies do (if it gets ruled that way), these companies are essentially arguing for you, and so you could ride on the coat tails of corpo lawyers.
But yeah, it's complicated.
Zarobi@aussie.zone · 0 pts · 1h
Lawmakers are still reeling from the invention of the internet, they have no hope of deciding how A.I. should work within the next decade
TootSweet@lemmy.world · 3 pts · 9h
OpenAI and Anthropic and Alphabet etc try their best to preventing it from doing so, so the act of getting it to do so anyway is "jailbreaking", but yes. LLMs can reproduce works it was trained on. (At least chunk-by-chunk each limited to the token limit.) And I'd imagine the same is true of things like Stable Diffusion, though it might be much easier to get Stable Diffusion to produce something similar enough to qualify as a "derivative work" than it would be to get it to produce exactly the original.
SeeMarkFly@lemmy.ml · 17 pts · 10h
I’m gonna sue my ass for what It did to me!
I've already punished myself with a pay cut.
Lodespawn@aussie.zone · 7 pts · 10h
I feel like the executives of you need a bonus for coming clean about the issue
SeeMarkFly@lemmy.ml · 6 pts · 9h
You know, I was honest about my mistake. There should be some kind of reward for honesty.
MutantTailThing@lemmy.world · 15 pts · 9h
Man I hate it when that happens.
Zier@fedia.io · 8 pts · 9h
I find that consuming popcorn and beer while analyzing the files that "invaded" your system, helps to pinpoint issues you may have missed. And don't forget to take a pizza break with all that hard work you have ahead.
goatinspace@feddit.org · 2 pts · 3h
Often making a
chandelierbeerdelier out of beer cans improves internal anal isys by 800% in August, but it has to face North during investigation.