Technology

37850 readers

460 users here now

A nice place to discuss rumors, happenings, innovations, and challenges in the technology sphere. We also welcome discussions on the intersections of technology and society. If it’s technological news or discussion of technology, it probably belongs here.

Remember the overriding ethos on Beehaw: Be(e) Nice. Each user you encounter here is a person, and should be treated with kindness (even if they’re wrong, or use a Linux distro you don’t like). Personal attacks will not be tolerated.

Subcommunities on Beehaw:

This community's icon was made by Aaron Schneider, under the CC-BY-NC-SA 4.0 license.

founded 3 years ago

MODERATORS

alyaza@beehaw.org

TheRtRevKaiser@beehaw.org

gyrfalcon@beehaw.org

rs5th@beehaw.org

coldredlight@beehaw.org

Los@beehaw.org

SemioticStandard@beehaw.org

TheRtRevKaiser@kbin.social

remington@beehaw.org

242

OpenAI says it’s “impossible” to create useful AI models without copyrighted material (arstechnica.com)

submitted 1 year ago by sculd@beehaw.org to c/technology@beehaw.org

244 comments fedilink hide all child comments

Apparently, stealing other people's work to create product for money is now "fair use" as according to OpenAI because they are "innovating" (stealing). Yeah. Move fast and break things, huh?

"Because copyright today covers virtually every sort of human expression—including blogposts, photographs, forum posts, scraps of software code, and government documents—it would be impossible to train today’s leading AI models without using copyrighted materials," wrote OpenAI in the House of Lords submission.

OpenAI claimed that the authors in that lawsuit "misconceive[d] the scope of copyright, failing to take into account the limitations and exceptions (including fair use) that properly leave room for innovations like the large language models now at the forefront of artificial intelligence."

you are viewing a single comment's thread
view the rest of the comments

[–] frog@beehaw.org 0 points 1 year ago (1 children)

OpenAI are not going to make the source code for their model accessible to all to learn from. This is 100% about profiting from it themselves. And using copyrighted data to create open source models would seem to violate the very principles the open source community stands for - namely that everybody contributes what they agree to, and everything is published under a licence. If the basis of an open source model is a vast quantity of training data from a vast quantity of extremely pissed off artists, at least some of the people working on that model are going to have a "are we the baddies?" moment.

The AI models are also never going to produce a solution to climate change that humans will accept. We already know what the solution is, but nobody wants to hear it, and expecting anyone to listen to ChatGPT and suddenly change their minds about using fossil fuels is ludicrous. And an AI that is trained specifically on knowledge about the climate and technologies that can improve it, with the purpose of innovating some hypothetical technology that will fix everything without humans changing any of their behaviour, categorically does not need the entire contents of ArtStation in its training data. AIs that are trained to do specific tasks, like the ones trained to identify new antibiotics, are trained on a very limited set of data, most of which is not protected by copyright and any that is can be easily licenced because the quantity is so small - and you don't see anybody complaining about those models!

[–] teawrecks@sopuli.xyz 1 points 1 year ago

OpenAI are not going to make the source code for their model accessible to all to learn from

OpenAI isn't the only company doing this, nor is their specific model the knowledge that I'm referring to.

The AI models are also never going to produce a solution to climate change that humans will accept.

It is already being used to further fusion research beyond anything we've been able to do with standard algorithms

We already know what the solution is, but nobody wants to hear it

Then it's not a solution. That's like telling your therapist, "I know how to fix my relationship, my partner just won't do it!"

expecting anyone to listen to ChatGPT and suddenly change their minds about using fossil fuels is ludicrous

Lol. Yeah, I agree, that's never going to work.

categorically does not need the entire contents of ArtStation in its training data.

That's a strong claim to make. Regardless of the ethics involved, or the problems the AI can solve today, the fact is we seeing rapid advances in AI research as a direct result of these ethically dubious models.

In general, I'm all for the capitalist method of artists being paid their fair share for the work they do, but on the flip side, I see a very possible mass extinction event on the horizon, which could cause suffering the likes of which humanity has never seen. If we assume that is the case, and we assume AI has a chance of preventing it, then I would prioritize that over people's profits today. And I think it's perfectly reasonable to say I'm wrong.

And then there's the problem of actually enforcing any sort of regulation, which would be so much more difficult than people here are willing to admit. There's basically nothing you can do even if you wanted to. Your Carlin example is exactly the defense a company would use: "I guess our AI just happened to create a movie that sounds just like Paul Blart, but we swear it's never seen the film. Great minds think alike, I guess, and we sell only the greatest of minds".