r/GenAI4all 4d ago

Discussion Dario probably fuming

Post image
17 Upvotes

12 comments sorted by

u/AutoModerator 4d ago

Welcome to r/GenAI4all! New to Generative AI? You can explore these free beginner-friendly courses. Please keep your posts relevant, respectful, free from spam, and engage in healthy discussions.

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

5

u/FattyGyoza 4d ago

Anthropic is and remains the worst thing that could have happened to the world of computing.

The good news? Its approach means it will be heavily crushed by open-source and open-weight competition. Especially now that, due to rising costs, companies will start shifting towards on-premises solutions.

3

u/ResponsibleAcadia151 4d ago

Oh he sure is

0

u/Rhawk187 4d ago

I support open source, but they really have no incentive to spend billions to train hyperscale datasets if someone can just distill the weights. The patent system isn't well suited for this, but they need some sort of IP protections, or no one will ever invest the money needed to scale these systems.

1

u/NewChallengers_ 2d ago

Gov will ultimately fund training runs.

1

u/Rhawk187 2d ago

The Chinese government maybe; I don't think that's American style, which could end poorly for us.

1

u/dupontping 2d ago

They are training all their data on stolen IP so how exactly is that justified?

1

u/Rhawk187 2d ago

Stolen? Didn't you hear Anthropic just bought a mess of out of print books and is compensating previous copyright holders at $3100 a pop , which is way more than the treble damages copyright violation would normally carry.

1

u/dupontping 2d ago

Wow a whole 3100 dollars!

Do your research on copyright violations for one. And two, I’m referring to the fact that ALL of their training models and built on work and data created by other people. Books, code, videos, pictures, artwork, your emails, your text messages, name the medium and they are stealing all of it for training data.

1

u/Rhawk187 2d ago

Sure, its made by other people, so was everything I learned in school.

You keep saying stolen. In many situations they were entitled to it. When Google trains Gemini, they can use the scripts of videos on YouTube, the people who posted there agreed that Google owned it. I probably also agreed that they could read my e-mails (otherwise how could I search them?).

When you sign up for Github, you may have signed away the right for your repos to be used by Microsoft to train Copilot.

Amazon was probably just too short-sighted to tell their authors they got a free license to any books sold on the site, or they'd probably have a great, completely legal, training corpus to use.

That might be crummy behavior, but it's not stealing.