In the study, published in Nature, the MIT researchers Zheng Dai and David Gifford set out to to test whether a generated image can be traced back to a single piece of training data. They were looking specifically at diffusion models, the systems most often used for generating images and video. Their finding? It comes down to how big the training data set is. They found that the more data a model is trained on, the harder it becomes to attribute its output to any particular piece of training data.
Source: Does generative AI actually copy artists? Researchers say it’s up for debate