Awesome Auto Research Hub Papers · Datasets · Projects
← Back to papers

CiteME: Can Language Models Accurately Cite Scientific Claims?

arXiv '24 2024 21 benchmark

Read-first score

Read-first score 21, weighted from topical fit, citation, graph, method, reproducibility, and recency signals.

Recency 11%
75.1

Uses a gentle age decay so recent papers surface without erasing older foundations. 2024

Reproducibility 33%
30

Screens links and visible text for paper, code, dataset, artifact, and repository signals. pdf=True; code=False; dataset=False; markers=none

Topical relevance 56%
4.8

Matches configured research keywords against title, abstract, tags, and analysis text. matched=1

Field roles

Candidate

Rank sensitivity

Stability: volatile; rank range: 18.

Tags