OpenAI now tries to hide that ChatGPT was trained on copyrighted books, including J.K. Rowling's Harry Potter series

https://lemmy.world/post/3631252

I hope OpenAI and JK Rowling take each other down
Sticky this comment
What’s the issue against openAI?

They’re stealing a ridiculous amount of copyrighted works to use to train their model without the consent of the copyright holders.

This includes the single person operations creating art that’s being used to feed the models that will take their jobs.

OpenAI should not be allowed to train on copyrighted material without paying a licensing fee at minimum.

If they purchased the data or the data is free its theirs to do what they want without violating the copyright like reselling the original work as their own. Training off it should not violate any copyright if the work was available for free or purchased by at least one person involved. Capitalism should work both ways

But they don’t purchase the data. That’s the whole problem.

And copyright is absolutely violated by training off it. It’s being used to make money and no longer falls under even the widest interpretation of free use.

You need to expand on how learning from something to make money is somehow using the original material to make money. Considering that’s how art works in general, I’m having a hard time taking the side of “learning from media to make your own is against copyright”. As long as they don’t reproduce the same thing as the original, I don’t see any issues with it. If they learned from Lord of the rings to them make “the Lord of the rings” then yes, that’d be infringement. But if they use that data to make a new IP with original ideas, then how is that bad for the world/ artists.

Creating an AI model is a commercial work. They’re made to make money. Now these models are dependent on other artists data to train on. The models would be useless if they weren’t able to train on anything.

I hold the stance that using copyrighted data as part of a training set is a violation of copyright. That still hasn’t been fully challenged in court, so there’s no specific legal definition yet.

Due to the requirement of copywritten materials to make the model function I feel that they are using copyrighted works in order to build a commercial product.

Also AI doesn’t learn. LLMs build statistical models based on sentence structure of what they’ve seen before. There’s no level of understanding or inherent knowledge, and there’s nothing new being added.

Also Sam Altman is a grifter who gives people in need small amounts of monopoly money to get their biometric data

Bro I haven’t eaten in days. I am about to lose water access in my apartment. I can’t walk properly because I’m disabled and I have no friends or family to ask help of. I am at literal risk of dying because I don’t have my medication.

Do you think I give the slightest about my biometric data? That it is of ANY value to me at all?

People in need are in need for a reason. We also don’t tend to give a shit about biometric data.

Give the dude hell for OpenAI if you want but your complaint here is just stupid and really ignores the only “PEOPLE WHO ARE IN NEED” aspect.

He’s getting biometric data out of it. Who cares. What have you done to help people in need?

He’s not helping them. That’s my point. He’s taking advantage of them for his grift, so fuck him.

But he is helping them. They’re in need and he’s helping them in exchange for something to meaningless to us.

I’m more akin to say fuck you. You haven’t done anything to help someone but you’re passing judgment on this dude because he doesn’t meet your standard. Your standard is irrelevant. You’re not the one in need of help. It’s the same toxic white Knight behavior as idiots who start complaining about tiktokers helping homeless people for views. I was one of those homeless folks. We don’t care if you’re doing it to help yourself. Were just happy to get help at all when people ignore us.

So like I said. Fuck you. You’re not in need of help but you’re judging US and acting like you know what’s best for US while you don’t lift a finger in return. He’s doing something to help people who would never get help otherwise.

He is a better person than you are. Flat out. Whether he has a motive is unnecessary. He’s helping people and you’re shaming him for that without helping yourself.

Hes a better person than you are.

So hypothetical here. If Dreddit did launch a system that made it so users could trade Karma in for real currency or some alternative, does that mean that all fan fictions and all other fan boy account created material would become copyright infringement as they are now making money off the original works?

“Stealing”.

It cannot be theft as the product is publicly available and the original product is still available to other consumers.

You can not like this and you can argue against it but it isn’t theft. Hasn’t and never will be. The same way piracy isn’t theft.

People might respect this bizarre corporate protection stance if you use the correct terminology. And yes. You’re defending larger companies here, not individual artists. Copyright was invented for companies and corporations. They have extended copyright for decades to be able to hold on to stuff they believe to be theirs. They suppress creatives to take their work and put a copyright on it themselves.

The only people you’re protecting with your argument are massive corporations. Have fun with that.

They used to be a non profit, that immediately turned it into a for profit when their product was refined. They took a bunch of people’s effort whether it be training materials or training Monkeys using the product and then slapped a huge price tag on it.
I didn’t know they were a non profit. I’m good as long as they keep the current model. Release older models free to use while charging for extra or latest features