Skip to content
Tech News
← Back to articles

‘Sketchy AF’: What to Know About How OpenAI Staff Discussed Book-Pirating

read original more articles
Why This Matters

Internal OpenAI communications reveal employees debated the risks and ethics of using pirated books to train ChatGPT rather than paying for licensed content, weighing costs against convenience. This matters because it exposes the tension between AI companies' rapid development goals and copyright law, fueling ongoing litigation that could reshape how AI firms source training data. The revelations could increase legal and reputational risk for OpenAI and set precedent for the broader AI industry's data practices.

Key Takeaways

Documents recently unsealed in a copyright case show employees weighing the cost of buying books to train early ChatGPT models.