Statistics

This page gives an overview of the information about works, people and organizations made available via DataCite Commons. Please reach out to DataCite Support for questions or comments.

Data Sources

The following main data sources are used in DataCite Commons for a total of currently 46,714,569 records:

DataCite

24,431,407 Works
100% of identifiers and metadata.

Crossref

9,976,797 Works
7.88% of identifiers and metadata. Import is ongoing.

ORCID

12,205,898 People
100% of identifiers. Personal and employment metadata.

ROR

100,467 Organizations
100% of identifiers and metadata.
Additional information comes from these data sources:
  • Wikidata: inception year, geolocation and Twitter account for organizations
  • Unpaywall: download link for Open Access content via Crossref

Works

DataCite Commons currently includes 34,408,204 works, with identifiers and metadata provided by DataCite and Crossref. For the three major work types publication, dataset and software, the respective numbers by publication year are shown below.

16,787,809 Publications

10,067,780 Datasets

219,218 Software

6,418,427 out of all 34,412,003 (18.65%) works have been cited at least once, including 0.94% of works registered with DataCite, and 62.01% of works registered with Crossref.

6,170,008 (36.75%) Cited Publications

98,260 (0.98%) Cited Datasets

1,802 (0.82%) Cited Software

People

DataCite Commons includes all 12,205,898 ORCID identifiers, and personal and employment metadata. This information is retrieved live from the ORCID REST API, the respective numbers by registration year are shown below.

12,205,898 People

4,850,562 out of all 34,412,003 (14.10%) works have been claimed (connected) to at least one ORCID record, including 6.03% of works registered with DataCite, and 33.86% of works registered with Crossref.

3,736,309 (22.26%) Claimed Publications

716,160 (7.11%) Claimed Datasets

34,964 (15.95%) Claimed Software

Organizations

DataCite Commons includes all 100,467 Research Organization Registry (ROR) identifiers and metadata. This information is retrieved live from the ROR REST API, the respective numbers by registration year are shown below.

100,467 Organizations

24,749,425 out of all 34,412,003 (71.92%) works are connected with at least one organization via ROR ID or Crossref Funder ID, including 63.59% of works registered with DataCite, and 92.33% of works registered with Crossref.

12,517,621 (74.56%) Connected Publications

6,670,264 (66.25%) Connected Datasets

214,491 (97.84%) Connected Software