Skip to content

Meta AI Trained on Pirated Data

Authors Allege Zuckerberg Approved Use of Pirated Books for AI Training

  • A group of authors, including Sarah Silverman, filed a copyright infringement lawsuit against Meta in July 2023.
  • The lawsuit claims Meta used pirated books from the LibGen dataset to train its Llama LLM.
  • Mark Zuckerberg allegedly approved using the dataset despite internal warnings about its illegal nature.
  • Meta engineers reportedly expressed discomfort with downloading pirated material but proceeded anyway.
  • The court dismissed most claims in November 2023, but recent allegations may revive the case.

The lawsuit accuses Meta of knowingly using pirated content from LibGen to train AI models like Llama, with approval from CEO Mark Zuckerberg despite internal concerns.

Share