Text and data mining permissions across two hundred journal policies
Author
Abstract
Policy language remains highly heterogeneous, and in a third of cases the stated permissions conflict with the licence attached to the articles themselves. The analysis draws on a corpus assembled from institutional repositories, preprint servers and publisher metadata feeds, and the full dataset together with the code needed to reproduce every figure is deposited under a permissive licence. Limitations, including coverage gaps in non-English sources, are set out in the discussion.
Graphical Abstract
Keywords
Subjects
