Interactive demo
Disruptiveness scoring
Seeded from this week's #1 paper. Drag factors to recompute the composite — same five-axis model used on cards and articles.
Calibrated to
AI4AI-Bench: Benchmarking LLM Agents in Algorithmic Design for Recursive Self-Improvement
arXiv:2608.20318 · editorial 93/100
Heuristic v1.1 · 6 topic-signal hits (2 in title), 0 boost phrases, claim=no, practical=yes. Editorial review recommended before publish. Cohort-calibrated to 93 (rank 1/20).
composite 86/100
Δ -7 vs editorial
Is this a new idea, proof, architecture, or measurement — or a small delta?
If true, how much do roadmaps, products, or theory change?
How active is this subfield right now on arXiv and in labs?
Can someone act on this soon — devices, code, trials — or is it pure theory?
Does it challenge orthodoxy or invite healthy debate? (Not drama for its own sake.)
Scores are triage, not peer review. See methodology · open this paper · curation demo.
