Every score comes with the model's reasoning on the paper page.
"Confidence-Gated Transductive Test Generation for Code Reranking" scored 4.0/10 predicted impact on @kurateorg — Clarity 7.5 · Reproducible 7.5 · Rigor 6.0