benchmarks
fact
neutral
RATIO is constructed from millions of full-text scientific papers across CS literature using discourse-marker distant supervision extended to corpus-scale retrieval, combined with LLM and human vetting
RATIO is constructed from millions of full-text scientific papers across CS literature via a general recipe that extends discourse-marker distant supervision - previously used only for classification - to corpus-scale retrieval, combined with extensive LLM and human vetting.
Computation and Language30 Aug 2026