Story
arxiv_cs_cl ยท Jul 17, 2026 ยท paper
arxiv.orgJul 17, 2026
original source linked
In brief
Large language models (LLMs) excel at answering pre-specified questions, yet their ability to navigate the open-ended, pre-conclusion stage of discovery remains largely unmeasured. We introduce Prospective Hypothesis...
Feed lens
evaluation