Technology•The Guardian 2d ago

Oxford lets OpenAI train its AI models on Bodleian Library

University staff voice concerns over reputational risk of partnering with company behind ChatGPTThe University of Oxford has allowed the company behind ChatGPT to train its AI models on historical texts from its Bodleian Library, as tech companies scour academic institutions for fresh data.The Bodleian material digitised by OpenAI has been used to “populate the OpenAI training set”, according to internal documents. Continue reading...

Oxford lets OpenAI train its AI models on Bodleian Library

Reporting preview

University staff voice concerns over reputational risk of partnering with company behind ChatGPT

The University of Oxford has allowed the company behind ChatGPT to train its AI models on historical texts from its Bodleian Library, as tech companies scour academic institutions for fresh data.

The Bodleian material digitised by OpenAI has been used to “populate the OpenAI training set”, according to internal documents.

Oxford announced a partnership with the company in March 2025, using OpenAI software to digitise texts from the university’s world-famous library, which it said would make the content more widely available for students and researchers.

However, the announcement did not state the material would be used for training OpenAI’s models, which are trained to recognise patterns in words – and thus “learn” to write complete sentences and perform other cognitive tasks – by being fed vast amounts of data.

This preview is a source-linked editorial excerpt. Continue to the original publisher for the complete report, updates and full context.

Continue reading at The Guardian