πŸ† Finalist β€” NIH Data Sharing Index (β€œS-Index”) Challenge

The data-sharing index

Empowers the data sharing that empowers science

An open platform that measures the network impact of datasets β€” and empowers the researchers who share it well. Search any DOI to see its DataRank in seconds.

β€”papersβ€”datasetsβ€”authorsβ€”institutionsβ€”citations
Pilot corpus: NIH-funded biomedical datasets, scored and ranked. Any DOI can be scored on demand β€” network effects sharpen as coverage grows.

Why theSindex

Data sharing is recommended by Policy.

Citations and metrics reward Quantity.

theSindex empowers the empowering datasets β€” and the scientists who share them.

A published dataset is worth the science it empowers.

Why DataRank

A number you can trust. A number that's already cited.

DataRank is built entirely from real, verifiable citations β€” no black box, no proprietary weighting. It reads a dataset's downstream impact straight off the citation graph, so the score reflects science that has already happened.

01

Credit that flows from citations

A dataset earns signal from the papers that cite it β€” and the more influential those citing papers are, the more it counts.

02

Built only from real citations

No proprietary weighting. The score starts from a dataset's own citation count and blends in the strength of the work that builds on it.

base(p) = log1p(citation_count)

03

One comparable percentile

Every dataset lands on a 1–100 corpus percentile, giving researchers, funders, and institutions a single clear benchmark.

Phase 1 corpus

NIH-funded biomedical research, scored

Our curated Phase 1 corpus covers NIH-funded data papers β€” papers whose main contribution is a shared dataset, identified by our DrPaper classifier β€” scored, ranked, and expanding periodically. Any paper with a DOI can be scored on demand.

Papers indexed

Datasets indexed

Authors tracked

Institutions

Funders indexed

Citation edges

Phase 1 covers NIH-funded biomedical research. The corpus expands periodically.

FAIR Agent

Our FAIR Agent coaches you to share better.

Sharing a dataset is a skill β€” and most of us were never taught it. The FAIR Agent reads your paper's full text, scores how Findable, Accessible, Interoperable, and Reusable your data really is, then hands you concrete, prioritised steps to make it more reusable for the next scientist.

FAIR readinessGood Β· 78
F
A
I
R
Deposit the raw data in a repository that mints a dataset DOI.
Add an explicit open license (CC0 or CC-BY) to the record.
Link accessions and the dataset DOI from the paper itself.

Illustrative coaching. A parallel quality metric β€” never folded into DataRank.

API

Build on DataRank

Every score, ranking, and citation network on theSindex is served by a free, open REST API β€” no key required. Pull DataRank into your own analysis, dashboards, or tools.

Open REST API

Query papers, authors, institutions, and citation graphs over JSON. Interactive docs are auto-generated from the OpenAPI spec.

CSV export

Download ranked papers with full DataRank breakdowns β€” base score, citation-network contribution, corpus percentile, and more.

98.4Top 2%

DataRank percentile

How it works

From DOI to DataRank in seconds

1

Enter any DOI

Paste a DOI from any published paper. We query CrossRef and OpenAlex in real time, plus enrichment metadata services.

2

Start from its own citations

The dataset's base score comes from how often it's cited β€” growing sub-linearly, so it rewards reuse without runaway scaling.

base(p) = log1p(citation_count)

3

Add the influence of citing work

We look one hop out to the papers that cite it, weight each by its own citation strength, drop self-citations, and blend the two.

network(p) = Ξ£ log1p(C_q) / outdegree(q) Β· damping d = 0.85

4

Get your DataRank

Receive a DataRank score, corpus percentile, and a full breakdown of what contributed to the result.

See what your data empowers.

Search any DOI. In seconds you'll get a DataRank score, corpus percentile, and the full story of the science a dataset set in motion.