When AI Co-Scientists Fail
Dataset Analysis
Recent advances in large language models (LLMs) have fueled the vision of automated scientific discovery, often called AI Co-Scientists. To date, prior work casts these systems as generative co-authors responsible for crafting hypotheses, s...
Provenance
Collected from papers.
Derived from paper: When AI Co-Scientists Fail: SPOT-a Benchmark for Automated Verification of Scientific Research