Pilot corpus only
This score is computed over theSindex pilot corpus and does not cover the full scientific literature. Scores are relative to papers we have ingested — papers, authors, and institutions outside the pilot are not represented. Methodology.
Driven by 3 papers. Top paper: “MedCalc-Bench: Evaluating Large Language Models for Medical Calculations”.
Qiao Jin, Guangzhi Xiong, S.M. Dunn, Serina Applebaum, Zain Anwar +12 more
Qiao Jin, Robert Leaman, Xiaoyu Liu, Guangzhi Xiong, Maame Sarfo-Gyamfi +13 more
Nicholas Wan, Robert Leaman, Shubo Tian, Zhizheng Wang, Yifan Yang +18 more