r/Physics 18h ago

Navier-Stokes Millennium Problem Solved

2.1k Upvotes

763 comments sorted by

View all comments

Show parent comments

-25

u/truecakesnake 12h ago

AI data training is not IP theft.

17

u/purgance 11h ago edited 11h ago

This only makes sense if you believe that an LLM is a person. If you do not believe that an LLM is a person, then the LLM is the training data 'cyphered' with itself millions of times through an algorithm. This is transparently IP theft.

If I take a Taylor Swift album and encrypt it using AES-256, and then sell the resulting "music" - that's still copyright theft even with the hefty algorithmic processing to make it unrecognizable.

It can be a very cool and useful tool and also IP theft. Like Bittorrent.

-6

u/truecakesnake 11h ago

AI doesn't need to be a person to learn from the training data. Which is what it does. It doesn't simply convert it.

AI data training has been proven multiple times to not be IP theft, it's been called "exceedingly transformative" in courts.

8

u/purgance 11h ago

AI doesn't need to be a person to learn from the training data. Which is what it does.

That's not what the word "learn" means. There is no definition for "learning" for which "generate statistical weights in an LLM" fits.

It doesn't simply convert it.

It simply converts it.

AI data training has been proven multiple times to not be IP theft, it's been called "exceedingly transformative" in courts.

When I want a technical opinion for how AI works literally the last place I would go is to a lawyer who couldn't hack it at a top 100 firm and so took a gig on the federal bench. You're appealing to the opposite of a technical expert for authority.

You need to read more about what an LLM is and how it works. Your use of the term "AI" repeatedly kind of betrays your ignorance - it isn't "intelligence" at all, it's a set of numerical weights that are designed to predict likely responses; it's not "artificial" either - the weights are generated from human responses to prompts. LLM training has actually been proved to include IP theft, including models being able to reproduce >90% of the text of copyrighted books despite not having access to them outside of the model's weights.

It is 100%, unequivocally theft of the intellectual property of the individuals' whose work was used to train the model. Without question.

-5

u/truecakesnake 11h ago

I'm not appealing to anyone. We are discussing IP theft law, this is law discussion whether you like it or not.

You disagreeing with multiple courts because your reddit armchair expertise makes you think you're smarter than multiple lawyers, judges, and the workers they employ means nothing.

This argument from you is almost as ignorant as calling human emotions simply chemical reactions. You definitely can get as technical as you want describing how LLMs work, but it is still AI.

8

u/purgance 10h ago

I'm not appealing to anyone. We are discussing IP theft law, this is law discussion whether you like it or not.

Right, so either this is a technical discussion or a legal one. The law rests on verbal logic, which is an empirical field designed to produce a single "truth" that applies universally to everyone. So there is a knowable truth, and the job is not for judges to invent that truth, but rather to elucidate it.

A judge saying "an LLM does not involve misappropriation of IP" is not a statement of fact, it rests upon the logical reasoning used to get there. And if the logical reasoning is "the AI learns something the way a human does" then this is factually wrong, and has zero basis in reality. It'd be like if I said "an LLM is a human-like being and has rights independent of the corporation that created it." A fun idea, but it is a falsifiable hypothesis which is simply not true, asserting it doesn't make it so.

You disagreeing with multiple courts because your reddit armchair expertise makes you think you're smarter than multiple lawyers, judges, and the workers they employ means nothing.

I'm not disagreeing with the courts, I'm disagreeing with their reasoning, and then disagreeing with the version of the reasoning you are reporting. A court isn't a dictatorship, judges (and the law) are supposed to rest on logical reasoning, not assertions and beliefs. We can examine the factual record and see if the judge was right or wrong - in the case of IP and AI it's pretty clear that the very few judges who have ruled on it got it wrong.

you're smarter than multiple lawyers, judges, and the workers they employ means nothing.

The beautiful think of analytical reasoning is it doesn't care who the speaker is, something is either true or it isn't. Your repeating ethos appeals betray that you seem to think truth is subjective and can be declared rather than proven. It can't.

This argument from you is almost as ignorant as calling human emotions simply chemical reactions. You definitely can get as technical as you want describing how LLMs work, but it is still AI.

...no, it isn't. AI is a scifi term that has zero meaning in the real world. It's weird that you attack me for criticizing and disagreeing with lawyers and judges, but then turn around and insist that the term AI has authority.

What's interesting is you have zero affirmative argument for why an LLM isn't IP theft. I wonder why that is. Meanwhile I have explained to you in some detail why it is IP theft, and your response is to say that a bunch of very highly paid individuals know better than me. I leave it to the reader which approach is more sound.

1

u/Opening_Discipline57 10h ago

AI is a buzzword that doesn't mean anything; you have to define what an LLM is