🏆 Finalist — NIH Data Sharing Index (“S-Index”) Challenge

Pilot corpus only

This score is computed over theSindex pilot corpus and does not cover the full scientific literature. Scores are relative to papers we have ingested — papers, authors, and institutions outside the pilot are not represented. Methodology.

Top 5%percentile
0.435Author DataRank

Indexed papers

1in pilot corpus
datarank_citation_only_1hop_v6· scope data_onlyMethodology
Why this DataRank?

An author's DataRank is the sum of the DataRanks of all 1 indexed paper attributed to them. A prolific author with many moderate-impact papers can outrank one with a single high-impact paper.

Author scores recompute whenever paper DataRanks are refreshed, so this number lags the underlying paper scores by at most one batch run.

Read the full methodology →

Top data-sharing exemplar

The highest-impact dataset this researcher has shared, ranked by DataRank — the single contribution doing the most to lift their data-sharing standing.

Top 43%6 citations

Alex Morehead, Chaitanya K Joshi, Zuobai Zhang, Kieran Didi, Simon V. Mathis +6 more

Papers

Driven by 5 papers — median percentile 57. Top paper: Evaluating representation learning on the protein structure universe.

Top 43%6 citations

Alex Morehead, Chaitanya K Joshi, Zuobai Zhang, Kieran Didi, Simon V. Mathis +6 more

6 citations

Lili Wang, Fang Li, J. Tang, Lanying Du, Yang Yang +4 more

23 citations

Renee N. Donahue, Madan Katragadda, Jessica Lowry, Wei Huang, Karunya Srinivasan +15 more

12 citations

Zuobai Zhang, Ke Zhang, Francesca Viggiani, Claire Callahan, J. Tang +3 more

13 citations

Qiao Jin, Kunlun Zhu, Tongxin Yuan, Yichi Zhang, Wangchunshu Zhou +8 more