Documents unsealed in a copyright case describe internal discussions among OpenAI employees about the costs and practicality of obtaining books to train early ChatGPT models. The materials include references to whether staff should purchase books and how that would affect training and related expenses.
The outlets reporting on the case characterize the conversations in different ways, including describing the exchanges as raising questions about how training data was accessed. The Wall Street Journal focuses on what the unsealed documents show about employee decision-making around book acquisition. Yahoo Finance highlights the same dispute but uses more informal framing while emphasizing the reported details of those discussions.
Overall, the reporting centers on the documents’ portrayal of staff discussions rather than on a final determination of legal responsibility. The unsealed records are presented as evidence in the copyright litigation and are used to shed light on how training data could have been sourced during ChatGPT’s early development.