Datadory notebook
Goodreads Ratings Data Alternative Data: Dataset Structure and Field Coverage
Datadory delivers goodreads ratings data alternative data covering comprehensive field definitions, entity mappings, and historical time series — structured for direct analytics and delivered on demand.
1,744 datasets. Pick your catch.
How do you build a ratings pipeline without touching Goodreads at all?
A production workflow starts from permitted channels and ends with identifiers that survive across catalogs:
Where to go next
Start with the pillar guide to see how all 19 pooled publishing datasets fit together, then go deeper on the two adjacent builds - scholarly DOI coverage and the OpenAlex citation graph:
Pick up where this leaves off
Every one of these ships with sample rows before you commit to anything.
Goodreads Web Catalog - Structured Book Pages
name · author · isbn …+9 more
Open Library Monthly Data Dumps Data
type · key · revision …+2 more
Google Books APIs Data
subtitle · publisher · publishedDate …+11 more
Open Library Search & Works/Editions API Data
numFound · docs · key …+17 more
Open Library General Catalog - Browsable Lending Library Data
key · title · edition_count …+11 more
OpenAlex API - Scholarly Book & Publisher Graph
doi · title · type …+21 more
Want rows instead of a pitch? Name the datasets.
API, files, or your warehouse. Daily, weekly, or hourly.
Get a sampleQuestions worth asking
How do I match rating data across different book databases?
Join on ISBN-13 through Wikidata property P212, which links editions across catalogs inside a 122,983,238-entity graph queryable via SPARQL. Open Library's dumps carry edition records keyed the same way, and the International ISBN Agency's Range Message XML validates prefixes against 287 registration groups and 1,871 registrant rules before you merge.