{
  "id": 1748960,
  "title": "When AI art has no author: Study finds generated images often can’t be traced to training data",
  "url": "https://urgent.news/2026/08/18/when-ai-art-has-no-author-study-finds-generated-images-often-cant-be",
  "topic": "ai",
  "section": "AI",
  "published": "2026-08-18T16:35:00.000Z",
  "source": {
    "name": "MIT News AI",
    "slug": "mit-news-ai",
    "url": "https://news.mit.edu/2026/when-ai-art-has-no-author-generated-images-often-cant-be-traced-to-training-data-0818"
  },
  "original_language": "en",
  "account": "A new study from MIT's Computer Science and Artificial Intelligence Laboratory (CSAIL) has found that artificial intelligence (AI) image generators often cannot trace generated images to their training data. This phenomenon, dubbed \"attribution decay,\" occurs as AI models are trained on larger datasets. As the amount of data increases, individual training examples become less influential on the generated output. Lead author Zheng Dai explains that if removing a piece of data does not change the AI's output, then that data cannot be considered responsible for the result. This phenomenon challenges the process of assigning credit to artists, assigning liability to companies, and establishing regulations surrounding AI-generated content. The study introduces a new method called \"diffusion ensemble\" which allows for precise removal of individual training examples, providing a counterfactual model that can be used to demonstrate the absence of influence. The researchers trained multiple ensembles on various datasets ranging from 256 to over 160,000 images and found that as the training set grows larger, the \"counterfactual radius\" - the extent to which any single piece of training data could impact the output - shrinks. The phenomenon impacts legal questions about derivative works, copyrightability, and attribution, as well as raising concerns about the privacy implications of AI-generated content.",
  "summary": "A new method for surgically removing training examples from a model reveals that as datasets grow, the link between what a model learns and what it produces dissolves.",
  "key_points": [
    "\"Attribution decay\" occurs as AI models are trained on larger datasets",
    "Counterfactual radius shrinks with larger training sets",
    "Study introduces \"diffusion ensemble\" method for precise data removal"
  ],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 2,
    "also_reported_by": [
      {
        "outlet": "MIT News Research",
        "title": "When AI art has no author: Study finds generated images often can’t be traced to training data",
        "url": "https://urgent.news/2026/08/18/when-ai-art-has-no-author-study-finds-generated-images-often-cant-be-1752578",
        "published": "2026-08-18T16:35:00.000Z"
      }
    ]
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}