Yes. Just because they’re in a neural network and not ASCII or unicode doesn’t mean they’re not stored. It’s even more apt a concept sonce apaprently those works can be retrieved fairly easily, even if the references to them are hard to isolate. It seems ChatGPT is storing eidetic copies of data, which would imply what other people have said in this thread, that it is overfitting itself to the data and not learning truly generalisable language.
Comment on Google Researchers’ Attack Prompts ChatGPT to Reveal Its Training Data
tabarnaski@sh.itjust.works 1 year agoYou remember some dialogue from your favorite movie. Does this mean your neurons store copyrighted work?
Excrubulent@slrpnk.net 1 year ago
MxM111@kbin.social 1 year ago
The claim is that it contains entire copies of the book. It does not. AI memory is like our memory, we do not remember books word to word.
Excrubulent@slrpnk.net 1 year ago
They are spitting out, as in the quote above, “verbatim text”, as in, word for word. That is copyrightable.
And that’s not what you said. You it has no memory. That’s clearly wrong.
anlumo@lemmy.world 1 year ago
It’s only under copyright if it’s a significant portion of the work. Single sentences are not enough, unless it’s a short poem.
Fermion@mander.xyz 1 year ago
Shhh. Disney’s lawyers might get ideas.