Authors, including Mike Huckabee, Sue Tech Companies Over Use of Their Work in AI Tools

Authors allege their books were pirated and used in AI datasets

Former Arkansas Governor Mike Huckabee and Christian author Lysa TerKeurst are among a group of writers who have filed a lawsuit against Meta, Microsoft, and other companies for reportedly using their work without authorization to advance AI technology. The authors claim that their written material was unlawfully replicated and incorporated into AI algorithms for training. EleutherAI, an AI research group, and Bloomberg are also named as defendants in the lawsuit.

Synthetic Data Generation: Creating privacy-safe datasets for AI training and data innovation for responsible machine learning (English Edition)

Synthetic Data Generation: Creating privacy-safe datasets for AI training and data innovation for responsible machine learning (English Edition)

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

This proposed class action suit is the latest example of authors accusing tech companies of using their work without permission to train generative AI models. In recent months, popular authors such as George R.R. Martin, Jodi Picoult, and Michael Chabon have also sued OpenAI for copyright infringement.

J. J. Keller Hours of Service Training Driver Handbook (5.25" x 8", English, Softbound) - Addresses Hours of Service Rule Changes

J. J. Keller Hours of Service Training Driver Handbook (5.25" x 8", English, Softbound) – Addresses Hours of Service Rule Changes

  • Includes FMCSA HOS Rule Changes: Details on four recent rule updates
  • Summarized Training Take-Away: Quick reference for training material
  • Training Record and Quizzes: Receipt page and assessment tools

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

The case centers on a controversial dataset called “Books3”

The Huckabee case focuses on a dataset called “Books3,” which contains over 180,000 works used to train large language models. The dataset is part of a larger collection of data called the Pile, created by EleutherAI. According to the lawsuit, companies used the Pile to train their products without compensating the authors.

Unlocking the Business of Writing: Top Tools to Transform Your Craft (Writers Edge Book 9)

Unlocking the Business of Writing: Top Tools to Transform Your Craft (Writers Edge Book 9)

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Microsoft, Meta, Bloomberg, and EleutherAI decline to comment

Microsoft, Meta, Bloomberg, and EleutherAI have not responded to requests for comment on the lawsuit. Microsoft declined to provide a statement for this story.

Amazon

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Debate over compensation for data providers in AI industry

The use of public data, including books, photographs, art, and music, to train AI models has sparked heated debate and legal action. As tools like ChatGPT and Stable Diffusion have become more accessible, questions surrounding how data providers should be compensated have arisen. Getty Images, for instance, sued the company behind AI art tool Stable Diffusion in January, alleging the unlawful copying of millions of copyrighted images for training purposes.

You May Also Like

Can AI Help Personalize Homework at Scale?

Yes, AI can help personalize homework at scale by analyzing each student’s…

AI in Higher Education: How Universities Are Adopting AI

Knowledge of AI’s impact on higher education is expanding rapidly, revealing transformative opportunities and challenges that universities are eager to explore further.

AI’s Impact on Personalized Learning: The Ultimate Guide

Welcome to our in-depth guide on the impact of AI on personalized…

How Coding Robots Make AI Concepts Easier to Teach

Coding robots make AI concepts easier to teach by providing hands-on, interactive…