Every score comes with the model's reasoning on the paper page.
"KWBench: Measuring Unprompted Problem Recognition in Knowledge Work" scored 6.5/10 predicted impact on @kurateorg — Clarity 8.5 · Novelty 8.0 · Significance 7.5