Can LLMs Generate Novel Research Ideas? A Large Scale Human Study with 100+ NLP Researchers
Read-first score
Read-first score 18.3, weighted from topical fit, citation, graph, method, reproducibility, and recency signals.
Field roles
Candidate
Rank sensitivity
Stability: stable; rank range: 1.