Big AI companies ingest vast amounts of copyrighted content—books, articles, photos, code—to train large language models without compensation or permission. Microsoft executives internally acknowledged this as "unprecedented theft" and a "doom loop" that destroys the creator supply chain, yet OpenAI and Microsoft defend the practice as fair use while their business model deliberately avoids paying content creators.