🏆 Finalist — NIH Data Sharing Index (“S-Index”) Challenge

Pilot corpus only

This score is computed over theSindex pilot corpus and does not cover the full scientific literature. Scores are relative to papers we have ingested — papers, authors, and institutions outside the pilot are not represented. Methodology.

N/A
0Author DataRank · unranked

Indexed papers

1in pilot corpus
datarank_citation_only_1hop_v6· scope data_onlyMethodology
Why this DataRank?

An author's DataRank is the sum of the DataRanks of all 1 indexed paper attributed to them. A prolific author with many moderate-impact papers can outrank one with a single high-impact paper.

Author scores recompute whenever paper DataRanks are refreshed, so this number lags the underlying paper scores by at most one batch run.

Read the full methodology →

Papers

Driven by 1 paper. Top paper: Data-Juicer: A One-Stop Data Processing System for Large Language Models.

Data-Juicer: A One-Stop Data Processing System for Large Language Models

Companion of the 2024 International Conference on Management of Data(2024)10.1145/3626246.3653385
N/A
0.546DataRank · unranked
37 citations

Yilun Huang, Zhijian Ma, Hesen Chen, Xuchen Pan, Ce Ge +8 more