OpenAI says it’s “impossible” to create useful AI models without copyrighted material

sculd@beehaw.org · 6 months ago

OpenAI says it’s “impossible” to create useful AI models without copyrighted material

BraveSirZaphod@kbin.social · 6 months ago

AI haters are not applying the same standards to humans that they do to generative AI

I don’t think it should go unquestioned that the same standards should apply. No human is able to look at billions of creative works and then create a million new works in an hour. There’s a meaningfully different level of scale here, and so it’s not necessarily obvious that the same standards should apply.

If it’s spitting out sentences that are direct quotes from an article someone wrote before and doesn’t disclose the source then yeah that is an issue.

A fundamental issue is that LLMs simply cannot do this. They can query a webpage, find a relevant chunk, and spit that back at you with a citation, but it is simply impossible for them to actually generate a response to a query, realize that they’ve generated a meaningful amount of copyrighted material, and disclose its source, because it literally does not know its source. This is not a fixable issue unless the fundamental approach to these models changes.