Data source
Data from Crossref, delivered clean.
1 dataset pulled from Crossref's releases, checked field by field and shipped the way you want them — daily, weekly, or hourly, your call.
- 1 dataset
- 1 industry
- Real rows on request
What Datadory delivers from Crossref
1Crossref REST API - Scholarly Publishing Metadata
Pick a catch, see the rows.
Name any Crossref dataset and we send real rows from it — not a screenshot of rows. 1,744 datasets. Pick your catch.
Get a sampleAPI, files, or your warehouse. Daily, weekly, or hourly.
Straight answers about Crossref data
How many scholarly works does the Crossref data cover?
185,678,115 DOI-registered works were counted at the August 2026 cataloging pass, alongside roughly 169,622 journals and 45,739 funders. The index grows continuously as publishers register new DOIs, so figures quoted at delivery reflect the pull you receive rather than a frozen brochure number.
Which fields support citation analysis?
Three ride on every work: the deposited `reference` list with optional per-reference DOIs, `reference-count` giving its size, and `is-referenced-by-count` tallying other Crossref records citing it. Together they build directed citation graphs and velocity measures without joining an external bibliometrics product.
How far back does coverage go?
To the registry's beginnings: DOIs have been registered continuously since 2000, and records run from historic backfile through current postings. A sampled Elsevier serial shows the split concretely - 8,459 backfile DOIs against 1,089 current ones out of 9,548 total.
Is the schema stable enough for production pipelines?
Twenty-four fields are documented once with definitions and examples verified against live responses during cataloging, and the confidence flag reads `verified`. Pipelines written against the dictionary behave identically whether they process one chapter or the whole 185-million-row corpus.