this post was submitted on 01 Feb 2024
586 points (97.3% liked)
Memes
45727 readers
1403 users here now
Rules:
- Be civil and nice.
- Try not to excessively repost, as a rule of thumb, wait at least 2 months to do it if you have to.
founded 5 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
Tech illiterate guy here. All these Ml models require training data, right? So all these AI companies that develop new ML based chat/video/image apps require data. So where exactly do they? It can't be that their entire dataset is licensed, isn't it?
If so, are there any firms that are using these orgs for data theft? How to know if the model has been trained on your data? Sorry if this is not the right place to ask.
You know how you look at a pic on the internet and don't pay? The AI is basically doing the same thing only it's collecting the effect of the data points ( like pixels in a picture) more accurately. The input no matter what it is only moves a set of weights. That's all. It does not copy anything it is trained on.
Yes it can reproduce with some level of accuracy any work just like a painter or musician could replay a piece they see or hear.
Again, this is not theft any more than u hearing a Song or viewing a selfie.
Isn't that the entire point of creativity. though? What separates an artist from a bad painter is the positioning of pixels on a 2-Dimensional plane? If the model collects the positions of pixels together with the pixel RGB (color? Don't know the technical term for it), then the model is effectively stealing the "pixel configuration and makeup" of that artist which can be reproduced by the said model anywhere if similar prompts were passed to it?
Focus. We are talking about copyright. Copyright doesn't cover this at all.