Alcohol intake and pancreatic cancer risk: An analysis from 30 prospective studies across Asia, Australia, Europe, and North America is a dataset published in PLoS Medicine (2025). On theSindex it has a DataRank of 1.4, placing it in the top 15.1% of the data-sharing corpus. It has been cited 26 times, with 22 citing works in its 1-hop citation network. Its calibrated FAIR score is 13/100.
Ranks in the top 15% for downstream scientific impact
DataRank reads this dataset's downstream impact straight off the citation graph — no black box, no proprietary weighting. How is this computed?
FAIR checklist signals are shown for context only and do not affect DataRank scoring.
Full FAIR picture · advisory
The headline score is computed from the scored criteria — the fact-shaped checks (a repository, an accession, a licence) that two independent models agree on. The advisory criteria below are real FAIR guidance but rest on judgment calls that models read differently, so they inform without moving the number.
No persistent identifier (DOI, Handle, ARK, or repository accession) is given for the dataset.
RDA-F1-01D — FAIR Data Maturity Model: 'Data is identified by a persistent identifier' (priorit · RDA-F1-02D — FAIR Data Maturity Model: 'Data is identified by a globally unique identifier' · FsF-F1-02D — F-UJI/FAIRsFAIR: 'Data is assigned a persistent identifier'
No repository is named as the holder of the data; the data are available upon request from the consortium. [majority verdict 'no' (4/5 passes agreed)]
RDA-F4-01M — FAIR Data Maturity Model: metadata is offered so it can be harvested and indexed ( · NIH DMS Policy Element 4 (NOT-OD-21-014) — name the repository where data will be archived · NSTC Desirable Characteristics of Data Repositories (2022) — 'Long-Term Sustainability', 'Reten
No identifier for the dataset appears in the reference list or body text.
FORCE11 Joint Declaration of Data Citation Principles (2014) — data should be cited as a first- · RDA-F3-01M — metadata clearly and explicitly includes the identifier of the data it describes · FsF-F3-01M — F-UJI: 'Metadata includes the identifier of the data it describes'
Advisory · not in the published score
“Consortium data can be made available upon request to the EPIC website ( https://epic.iarc.who.int/contact-us/ ).”
The data availability statement points to a request process, which is Colavizza category 1.
Colavizza, Hrynaszkiewicz, Staden, Whitaker & McGillivray (2020), 'The citation advantage of li · Springer Nature research data policy — Data Availability Statements: standard statement templat · RDA-F3-01M — metadata clearly and explicitly includes the identifier of the data it describes
“The total study sample consisted of 2,494,432 participants (62% women, 70% alcohol drinkers, 47% never smokers, 64% alcohol drinkers among never smokers) in 30 studies, recruited between 1980 and 2013 with a median age of 57 years.”— not found in the paper; verdict downgraded
The dataset is described in running prose without an itemised inventory. [downgraded to 'no' — no verifiable quote from the paper] [majority verdict 'no' (3/5 passes agreed)]
RDA-F2-01M — 'Rich metadata is provided to allow discovery' (priority Essential) · FsF-F2-01M — F-UJI: 'Metadata includes descriptive core elements to support data findability' · FsF-R1-01MD — F-UJI: 'Metadata specifies the content of the data'
“The sharing of data will be governed by executed data use agreements between the home institution of each cohort studies and the Harvard T.H. Chan School of Public Health, as well as approval from lead investigators for each cohort and/or the cohort study’s leadership.”
The access route carries a stated precondition of a data use agreement and approval, which is a specified followable process. [majority verdict 'partial' (4/5 passes agreed)]
RDA-A1.1-01D — 'Data is accessible through a free access protocol' · FsF-A1-01M — F-UJI: 'Metadata contains access level and access conditions of the data' · NSTC Desirable Characteristics of Data Repositories (2022) — 'Free and Easy Access'
Advisory · not in the published score
“Consortium data can be made available upon request to the EPIC website ( https://epic.iarc.who.int/contact-us/ ). The sharing of data will be governed by executed data use agreements between the home institution of each cohort studies and the Harvard T.H. Chan School of Public Health, as well as approval from lead investigators for each cohort and/or the cohort study’s leadership. Statistical programs can be made available upon request to the corresponding author.”
The paper describes the access process (request, DUA, approvals) but does not apply an explicit access-level label. [majority verdict 'partial' (4/5 passes agreed)]
FsF-A1-01M — F-UJI: 'Metadata contains access level and access conditions of the data' · RDA-A1-01M — metadata contains information to enable the user to get access to the data · COAR Controlled Vocabularies — Access Rights v1.0 (open / embargoed / restricted / metadata-onl
“The sharing of data will be governed by executed data use agreements between the home institution of each cohort studies and the Harvard T.H. Chan School of Public Health, as well as approval from lead investigators for each cohort and/or the cohort study’s leadership.”
The paper names an institutional process (data use agreements and lead investigator approval) as the gatekeeper for accessing the data.
NIH Genomic Data Sharing Policy (NOT-OD-14-124) — controlled-access via a Data Access Committee · RDA-A1.2-01D — 'Data is accessible through an access protocol that supports authentication and · NIH DMS Policy Element 5 (NOT-OD-21-014) — Access, Distribution, or Reuse Considerations (conse
The paper does not mention any retention period or persistence commitment for the data.
NIH DMS Plan Element 4 (NOT-OD-21-014) — Data Preservation, Access, and Associated Timelines · NSTC Desirable Characteristics (2022), Organizational Infrastructure: 'Retention Policy' · RDA-A2-01M — 'Metadata is guaranteed to remain available after data is no longer available'
No file format is named for the released data.
FsF-R1.3-02D — F-UJI: 'Data is available in a file format recommended by the target research co · RDA-R1.3-02D — data is expressed in a machine-understandable community standard · RDA-I1-01D — data uses a knowledge representation expressed in a standardised format
Advisory · not in the published score
“This study is reported as per the Strengthening the Reporting of Observational Studies in Epidemiology (STROBE) guideline ( S1 Checklist ).”
The only named standard is a manuscript reporting guideline (STROBE), not a data or metadata standard.
RDA-R1.3-01M — 'Metadata complies with a community standard' (priority Essential) · RDA-R1.3-01D — 'Data complies with a community standard' · RDA-I2-01M — '(Meta)data use vocabularies that follow FAIR principles'
No identifier for any external resource (e.g., source dataset, reference genome) is provided.
RDA-I3-01M — '(meta)data include references to other (meta)data' · RDA-I3-03M — 'metadata includes qualified references to other metadata' · FsF-I3-01M — F-UJI: 'Metadata includes links between the data and its related entities'
No licence for the data is mentioned; the CC0 license applies to the article only. [majority verdict 'no' (4/5 passes agreed)]
RDA-R1.1-01M — 'Metadata includes information about the licence under which the data can be reu · RDA-R1.1-02M — 'Metadata refers to a standard reuse licence' · RDA-R1.1-03M — 'Metadata refers to a machine-understandable reuse licence'
No version token or date is provided to identify a specific snapshot of the data.
DataCite Metadata Schema 4.6 — the 'Version' property · RDA-R1.2-01M — provenance information (which version was used is provenance) · NSTC Desirable Characteristics of Data Repositories (2022) — 'Provenance', 'Retention Policy'
“Statistical programs can be made available upon request to the corresponding author.”
The code is available only upon request to a person, not via a machine-resolvable locator.
NIH DMS Policy Element 2 (NOT-OD-21-014) — 'Related Tools, Software and/or Code' · FAIR4RS Principles v1.0 (Chue Hong et al., 2022; RDA/FORCE11/ReSA) — FAIR Principles for Resear · FORCE11 Software Citation Principles (Smith, Katz & Niemeyer, 2016, PeerJ CS 2:e86)
“The centralization, checking, harmonization, and statistical analyses of the participant level data from the cohorts were supported by NIH AAA R01 grant (grants: R01AA024770, GR-IARC-20”
A grant number (R01AA024770) is provided for the funding.
DataCite Metadata Schema 4.6 — 'FundingReference' property (funderName, funderIdentifier, award · Crossref Funder Registry — canonical funder identifiers for funding metadata · RDA-F2-01M — rich metadata provided to allow discovery (funding is part of the descriptive reco
Advisory · not in the published score
“Individual-level data from 30 cohorts were pooled and harmonized.”— not found in the paper; verdict downgraded
The processing is described in generic terms without naming specific instruments, kits, or software versions. [downgraded to 'no' — no verifiable quote from the paper] [majority verdict 'no' (4/5 passes agreed)]
RDA-R1.2-01M — 'Metadata includes provenance information according to community- specific standa · FsF-R1.2-01M — F-UJI: 'Metadata includes provenance information about data creation or generati · W3C PROV-O (W3C Recommendation, 2013) — the entity/activity/agent model of provenance
No documentation object (README, codebook, etc.) is mentioned as accompanying the data. [majority verdict 'no' (4/5 passes agreed)]
RDA-R1-01M — '(Meta)data are richly described with a plurality of accurate and relevant attribu · FsF-R1-01MD — F-UJI: 'Metadata specifies the content of the data' · NIH DMS Policy Element 3 (NOT-OD-21-014) — Standards (documentation and metadata to accompany t
Calibrated FAIR score — a parallel quality metric, independent of the DataRank citation score. See the full evaluation →
Base Score Contribution
0.494
From this paper's citation signal
Citation Network Contribution
0.872
From 13 citing papers with measurable signal
Ranked by each citer's contribution to N(p) — log1p(Cq) divided by its reference count — out of 22 citers.
National Institutes of Health
Grant: 5R01CA077398-09
CALIFORNIA TEACHERS STUDY
National Institutes of Health
Grant: 5U01CA086308-05
CLUE STUDIES--EVALUATING BIOMARKERS OF CARCINOGENESIS
National Institutes of Health
Grant: 5P30CA033572-27
Developmental Funds
National Health and Medical Research Council (NHMRC)
Grant: 396414
The Health 2020 Cohort Study (Health 2020)
National Institutes of Health
Grant: 5R01CA144034-05
Prospective studies of cancer etiology and prevention in Shanghai and Singapore
National Institutes of Health
Grant: 5U01CA164973-04
Understanding Ethnic Differences in Cancer: The Multiethnic Cohort Study
National Institutes of Health
Grant: 3R01CA039742-11S1
DISTRIBUTION OF BODY FAT AND CANCER RISK IN WOMEN
National Institutes of Health
Grant: 5UM1CA182876-05
Cancer epidemiology cohorts in Shanghai and Singapore
National Institutes of Health
Grant: 5U01CA167462-07
Infrastructure Support and Pilot Tissue Collection for the CARET Biorepository
National Health and Medical Research Council (NHMRC)
Grant: 1074383
Linking lifestyle and molecular biology to inform "Precision Public Health" for major cancers
UK Research and Innovation
Grant: MR/M012190/1
Health of vegetarians
National Institutes of Health
Grant: 1R13AG013039-01
1995 SUMMER INSTITUTE IN GERIATRIC MEDICINE
National Institutes of Health
Grant: 3P30CA023100-25S4
Specialized Cancer Center Support Grant
National Institutes of Health
Grant: 5U01AG018033-03
GENE-ENVIRONMENT INTERACTIONS: THE ODYSSEY COHORT
National Institutes of Health
Grant: 3U01CA199277-07S1
Genome-wide genotyping of existing samples from Asian American and Pacific Islander participants in the California Teachers Study cohort to facilitate broad and open future research
National Institutes of Health
Grant: 2U01CA167552-06
Cancer Epidemiology Cohort in Male Health Professionals
National Institutes of Health
Grant: 5R01AA024770-04
A pooling project on alcohol use and risk of cancers with inconsistent prior evidence, with an emphasis in non-smokers.
National Institutes of Health
Grant: 5U01HL145386-07
Integrating lifecourse approaches, biologic and digital phenotypes in support of heart and lung disease epidemiologic research
National Health and Medical Research Council (NHMRC)
Grant: 209057
Epidemiology of Chronic Disease, Health Interventions and DNA Studies
National Institutes of Health
Grant: 5U01CA176726-09
Life Course Cancer Epidemiology Cohort in Women
National Institutes of Health
Grant: 3P01CA087969-03S1
DIET, HORMONES AND RISK OF COLORECTAL CANCERS
National Institutes of Health
Grant: 6UM1CA173640-04
Shanghai Men's Health Study
National Institutes of Health
Grant: 5UM1CA164917-02
New Biospecimens to Enhance Research in the California Teachers Study Cohort
National Institutes of Health
Grant: 3U01CA063673-10S4
THE CAROTENE AND RETINOL EFFICACY TRIAL (CARET)
National Institutes of Health
Grant: 5UM1CA186107-05
Long Term Multidisciplinary Study of Cancer in Women: The Nurses Health Study
Fields of Study
Keywords