🏆 Finalist — NIH Data Sharing Index (“S-Index”) Challenge

Pilot corpus only

This score is computed over theSindex pilot corpus and does not cover the full scientific literature. Scores are relative to papers we have ingested — papers, authors, and institutions outside the pilot are not represented. Methodology.

Top 4%percentile
0.752Author DataRank

Indexed papers

2in pilot corpus
datarank_citation_only_1hop_v6· scope data_onlyMethodology
Why this DataRank?

An author's DataRank is the sum of the DataRanks of all 2 indexed papers attributed to them. A prolific author with many moderate-impact papers can outrank one with a single high-impact paper.

Author scores recompute whenever paper DataRanks are refreshed, so this number lags the underlying paper scores by at most one batch run.

Read the full methodology →

Top data-sharing exemplar

The highest-impact dataset this researcher has shared, ranked by DataRank — the single contribution doing the most to lift their data-sharing standing.

Top 27%11 citations

Yung‐Hung Luo, Jason A. Wampfler, Samuel M. Rubinstein, Firat Tiryaki, Ashok Kumar +4 more

Papers

Driven by 5 papers — median percentile 73. Top paper: Efficient and Accurate Extracting of Unstructured EHRs on Cancer Therapy Responses for the Development of RECIST Natural Language Processing Tools: Part I, the Corpus.

Top 27%11 citations

Yung‐Hung Luo, Jason A. Wampfler, Samuel M. Rubinstein, Firat Tiryaki, Ashok Kumar +4 more

100 citations

Zhen Zeng, Yang Jin, Lan Yang, Ting Fan, Zhoufeng Wang +14 more

3 citations

Yalun Li, Vinicius Ernani, Jonathan D’Cunha, Marie‐Christine Aubry, Ping Yang +1 more

0 citations

Han Liu, Jason A. Wampfler, Henry D. Tazelaar, Yalun Li, Tobias Peikert +7 more

7 citations

Han Liu, Jason A. Wampfler, Henry D. Tazelaar, Yalun Li, Tobias Peikert +7 more