Pilot corpus only
This score is computed over theSindex pilot corpus and does not cover the full scientific literature. Scores are relative to papers we have ingested — papers, authors, and institutions outside the pilot are not represented. Methodology.
Indexed papers
Why this DataRank?
An author's DataRank is the sum of the DataRanks of all 1 indexed paper attributed to them. A prolific author with many moderate-impact papers can outrank one with a single high-impact paper.
Author scores recompute whenever paper DataRanks are refreshed, so this number lags the underlying paper scores by at most one batch run.
Read the full methodology →Top data-sharing exemplar
The highest-impact dataset this researcher has shared, ranked by DataRank — the single contribution doing the most to lift their data-sharing standing.
BanglaLM: Data Mining based Bangla Corpus for Language Model Research
Md. Jashim Uddin, Anik Tahabilder, Md Ruhul Amin, Md. Fahim Shahriar, Md. Shohanur Islam Sobuj +1 more
Papers
Driven by 5 papers — median percentile 44. Top paper: “BanglaLM: Data Mining based Bangla Corpus for Language Model Research”.
BanglaLM: Data Mining based Bangla Corpus for Language Model Research
Md. Jashim Uddin, Anik Tahabilder, Md Ruhul Amin, Md. Fahim Shahriar, Md. Shohanur Islam Sobuj +1 more
Christian C. Thompson, Md. Jashim Uddin, Abu Asaduzzaman
Anoop Vemulapalli, Hiroaki Niitsu, Brenda C. Crews, Connor G. Oltman, Philip J. Kingsley +9 more
Shu Xu, Brenda C. Crews, Ansari M. Aleem, Kebreab Ghebreselasie, Surajit Banerjee +2 more
Christian C. Thompson, Fadi N. Sibai, Md. Jashim Uddin, Abu Asaduzzaman