OpenAI says it’s “impossible” to create useful AI models without copyrighted material

sculd ( @sculd@beehaw.org ) · 10 months ago

OpenAI says it’s “impossible” to create useful AI models without copyrighted material

vexikron ( @vexikron@lemmy.zip ) · 10 months ago

Well, off the top of my head:

Whole Brain Emulation, attempting to model a human brain as physically accurately as possible inside a computer.

Genetic Iteration (not the correct term for it but it escapes me at the moment), where you set up a simulated environment for digital actors, then simulate quasi-neurons, quasi-body parts dictated by quasi-dna, in a way that mimics actual biological natural selection and evolution, and then you run the simulation millions of times until your digital creature develops a stable survival strategy.

Similar approaches to this have been used to do things like teach an AI humanoid how to develop its own winning martial arts style via many many iterations, starting from not even being able to stand up, much less do anything to an opponent.

Both of these approaches obviously have drawbacks and strengths, and could possibly be successful at far more than what they have achieved to date, or maybe not, due to known or existing problems, but neither of them rely on a training set of essentially the entirety of all content on the internet.

HumbleHobo ( @HumbleHobo@beehaw.org ) · 10 months ago

That sounds like a great idea for making an intelligent agent inside a video game, where you control all aspects of it’s environment. But what about an AI that you want to be able to interact with our current shared reality. If I want to know something that involves synthesis of multiple modalities of knowledge how should that information be conveyed? Do humans grow up inside test tubes that only consume content that they themselves have created? Can you imagine the strange society we would have if people were unleashed upon the world without having any shared experiences until they were fully adults?

I think the OpenAI people have a point here, but I think where they go off the rails is that they expect all of this copyrighted information to be granted to them at zero cost and with zero responsibility to the creators of said content.