Data set mentions and citations: A content analysis of full‐text publications
通过分析PLoS One上600篇论文,发现不同学科在数据集收集、引用和保存上差异大;多数文章提供免费数据,但DOI和数据引用等正式归属方式使用有限,且数据复用率不足30%。
This study provides evidence of data set mentions and citations in multiple disciplines based on a content analysis of 600 publications in PLoS One . We find that data set mentions and citations varied greatly among disciplines in terms of how data sets were collected, referenced, and curated. While a majority of articles provided free access to data, formal ways of data attribution such as DOIs and data citations were used in a limited number of articles. In addition, data reuse took place in less than 30% of the publications that used data, suggesting that researchers are still inclined to create and use their own data sets, rather than reusing previously curated data. This paper provides a comprehensive understanding of how data sets are used in science and helps institutions and publishers make useful data policies.