Account

Sign in to access your account and subscription

Hidden Details of AI Training Data Set Creates Dilemma for Copyright Holders’ Infringement Claims

How are copyright holders to prove their works were used to train AI models if the details about the vast data sets used for such training are kept secret? That’s a dilemma that surfaced in late August when a federal judge dismissed a claim of direct infringement raised by a group of authors.

7 minute read September 01, 2025 at 12:03 AM
By
Michelle Morgante
Hidden Details of AI Training Data Set Creates Dilemma for Copyright Holders’ Infringement Claims

How are copyright holders to prove their works were used to train AI models if the details about the vast data sets used for such training are kept secret?

This premium content is locked for The Intellectual Property Strategist subscribers only

ENJOY UNLIMITED ACCESS TO THE SINGLE SOURCE OF OBJECTIVE LEGAL ANALYSIS, PRACTICAL INSIGHTS, AND NEWS IN The Intellectual Property Strategist

  • Stay current on the latest information, rulings, regulations, and trends
  • Includes practical, must-have information on copyrights, royalties, AI, and more
  • Tap into expert guidance from top entertainment lawyers and experts

Already have an account? Sign In Now

For enterprise-wide or corporate access, please contact Customer Service at [email protected] or call 1-877-256-2473.

NOT FOR REPRINT

© 2026 ALM Global, LLC, All Rights Reserved. Request academic re-use from www.copyright.com. All other uses, submit a request to [email protected]. For more information visit Asset & Logo Licensing.

Continue Reading

The Copyright Royalty Board (CRB), which works under the umbrella of the Librarian of Congress, sets statutory-license royalty terms and rates. The U.S. Courts of Appeals for the D.C. Circuit recently issued two notable decisions about the CRB.

October 01, 2026