Every score comes with the model's reasoning on the paper page.
"Realistic honeypot evaluations for scheming propensity" scored 7.0/10 predicted impact on @kurateorg — Clarity 8.0 · Significance 7.5 · Novelty 7.5