🏆 Finalist — NIH Data Sharing Index (“S-Index”) Challenge

Raymond Lee

California Institute of Technology

ORCID: 0000-0002-8151-7479

Also affiliated with University of Arizona, Lawrence Berkeley National Laboratory, Joint Genome Institute

Biochemistry, Genetics and Molecular BiologyNeuroscience

Pilot corpus only

This score is computed over theSindex pilot corpus and does not cover the full scientific literature. Scores are relative to papers we have ingested — papers, authors, and institutions outside the pilot are not represented. Methodology.

Top 1%percentile
7.4Author DataRank

Indexed papers

3in pilot corpus
datarank_citation_only_1hop_v6· scope data_onlyMethodology
Why this DataRank?

An author's DataRank is the sum of the DataRanks of all 3 indexed papers attributed to them. A prolific author with many moderate-impact papers can outrank one with a single high-impact paper.

Author scores recompute whenever paper DataRanks are refreshed, so this number lags the underlying paper scores by at most one batch run.

Read the full methodology →

Top data-sharing exemplar

The highest-impact dataset this researcher has shared, ranked by DataRank — the single contribution doing the most to lift their data-sharing standing.

Top 4%2744 citations

James P. Balhoff, Seth Carbon, J. Michael Cherry, Harold Drabkin, Dustin Ebert +90 more

Papers

Driven by 3 papers — median percentile 76. Top paper: The Gene Ontology knowledgebase in 2023.

Top 4%2744 citations

James P. Balhoff, Seth Carbon, J. Michael Cherry, Harold Drabkin, Dustin Ebert +90 more

Top 24%35 citations

Huaiyu Mi, Pascale Gaudet, Anushya Muruganujan, Suzanna Lewis, Dustin Ebert +95 more

Top 27%
0.743DataRank
Top 27%95 citations

James P. Balhoff, Seth Carbon, J. Michael Cherry, Dustin Ebert, Marc Feuermann +95 more