RATIO: A Benchmark for Retrieval Across Typed Ideation Operations in Scientific Literature
TLDR
Introduces RATIO, a benchmark for retrieving scientific literature by ideation operations (Address, Broaden, Specify), built via distant supervision and human/LLM vetting.
Reasoning
The paper presents a novel, large-scale benchmark with clear methodology and empirical evaluation, showing fine-tuning improves retrievers. However, its focus is on retrieval operations, not full automated discovery or experimentation, limiting relevance to broader AI-scientist keywords.
Read-first score
Read-first score 32.3, weighted from topical fit, citation, graph, method, reproducibility, and recency signals. Original total remains 28.
Field roles
Frontier
Rank sensitivity
Stability: volatile; rank range: 29.