Apple Faces Class-Action Lawsuit Over Alleged Copyright Infringement in AI Training Data

The tech giant Apple is facing a class-action lawsuit filed by two authors, Grady Hendrix and Jennifer Roberson, who accuse the company of copyright infringement for allegedly using their books without permission to train its artificial intelligence models. The lawsuit claims that Apple accessed online repositories known as "shadow libraries," which host illegally distributed books, using its web crawler, Applebot. This allegedly led to the use of pirated copies of copyrighted works, including those of the plaintiffs, in the training process for Apple's AI models.

Key Takeaways:

  • The lawsuit alleges that Apple's actions constituted large-scale copyright theft, stripping authors and creators of control over their intellectual property and devaluing their creative works.
  • The plaintiffs claim that Apple reaped enormous commercial benefits through unlawful means, and that the company's conduct enabled it to profit from pirated literary works.
  • The Apple case is part of a growing wave of legal challenges targeting companies involved in generative AI development, including OpenAI, which is facing lawsuits from The New York Times and The Authors Guild.
  • Anthropic, the company behind the Claude chatbot, recently agreed to a $1.5 billion settlement in a similar class-action lawsuit filed by authors, with reports indicating that the 500,000 authors involved will receive up to $3,000 per work.
  • The Apple case marks a significant escalation in the debate over data sourcing in AI training, with authors and publishers arguing that large-scale data scraping amounts to digital looting and violates copyright law.
  • Legal experts believe that these lawsuits could reshape how AI companies collect and use data, leading to stricter licensing requirements, new compensation models for authors, or major changes in the way AI is developed.

Statistics:

  • The lawsuit alleges that Apple used pirated copies of copyrighted works, including those of the plaintiffs, in the training process for its AI models.
  • The case centers on allegations that the company accessed online repositories known as "shadow libraries," which host illegally distributed books.
  • The Anthropic case involved 500,000 authors and a settlement of $1.5 billion.
  • Authors in the Anthropic case may receive up to $3,000 per work.
  • The number of lawsuits targeting companies involved in generative AI development is growing, with multiple cases brought against OpenAI and others.

Sources:

  • Azernews (no date)
  • [1]

Note: The source Azernews has no date mentioned in the original text.