Every score comes with the model's reasoning on the paper page.
"When Tool Calls Succeed but Workflows Fail: Anomalies at the Agent-Tool Boundary" scored 6.5/10 predicted impact on @kurateorg — Clarity 7.5 · Reproducible 7.5 · Novelty 7.0