The data-sharing index
Empowers the data sharing that empowers science
An open platform that measures the network impact of datasets β and empowers the researchers who share it well. Search any DOI to see its DataRank in seconds.
Why theSindex
Data sharing is recommended by Policy.
Citations and metrics reward Quantity.
theSindex empowers the empowering datasets β and the scientists who share them.
A published dataset is worth the science it empowers.
Why DataRank
A number you can trust. A number that's already cited.
DataRank is built entirely from real, verifiable citations β no black box, no proprietary weighting. It reads a dataset's downstream impact straight off the citation graph, so the score reflects science that has already happened.
Credit that flows from citations
A dataset earns signal from the papers that cite it β and the more influential those citing papers are, the more it counts.
Built only from real citations
No proprietary weighting. The score starts from a dataset's own citation count and blends in the strength of the work that builds on it.
base(p) = log1p(citation_count)
One comparable percentile
Every dataset lands on a 1β100 corpus percentile, giving researchers, funders, and institutions a single clear benchmark.
Phase 1 corpus
NIH-funded biomedical research, scored
Our curated Phase 1 corpus covers NIH-funded data papers β papers whose main contribution is a shared dataset, identified by our DrPaper classifier β scored, ranked, and expanding periodically. Any paper with a DOI can be scored on demand.
Papers indexed
Datasets indexed
Authors tracked
Institutions
Funders indexed
Citation edges
Phase 1 covers NIH-funded biomedical research. The corpus expands periodically.
FAIR Agent
Our FAIR Agent coaches you to share better.
Sharing a dataset is a skill β and most of us were never taught it. The FAIR Agent reads your paper's full text, scores how Findable, Accessible, Interoperable, and Reusable your data really is, then hands you concrete, prioritised steps to make it more reusable for the next scientist.
Illustrative coaching. A parallel quality metric β never folded into DataRank.
API
Build on DataRank
Every score, ranking, and citation network on theSindex is served by a free, open REST API β no key required. Pull DataRank into your own analysis, dashboards, or tools.
Open REST API
Query papers, authors, institutions, and citation graphs over JSON. Interactive docs are auto-generated from the OpenAPI spec.
CSV export
Download ranked papers with full DataRank breakdowns β base score, citation-network contribution, corpus percentile, and more.
DataRank percentile
How it works
From DOI to DataRank in seconds
Enter any DOI
Paste a DOI from any published paper. We query CrossRef and OpenAlex in real time, plus enrichment metadata services.
Start from its own citations
The dataset's base score comes from how often it's cited β growing sub-linearly, so it rewards reuse without runaway scaling.
base(p) = log1p(citation_count)
Add the influence of citing work
We look one hop out to the papers that cite it, weight each by its own citation strength, drop self-citations, and blend the two.
network(p) = Ξ£ log1p(C_q) / outdegree(q) Β· damping d = 0.85
Get your DataRank
Receive a DataRank score, corpus percentile, and a full breakdown of what contributed to the result.
See what your data empowers.
Search any DOI. In seconds you'll get a DataRank score, corpus percentile, and the full story of the science a dataset set in motion.