From d2c57dd0570b9f7eac4c4b4f35efb9fec8c5be77 Mon Sep 17 00:00:00 2001 From: YurenHao0426 Date: Wed, 22 Jul 2026 06:44:34 -0500 Subject: docs: log evidence-to-score changes --- REVIEW_SCORECARD.md | 11 +++++++++++ 1 file changed, 11 insertions(+) diff --git a/REVIEW_SCORECARD.md b/REVIEW_SCORECARD.md index 4cae7a1..1978688 100644 --- a/REVIEW_SCORECARD.md +++ b/REVIEW_SCORECARD.md @@ -61,3 +61,14 @@ Every formal result report records: These are conditional reviewer forecasts, not promised scores. A failed stage leaves its negative result in the record and can lower the score if it invalidates a current claim. Oral-B does not get retroactively reopened by success on standard vision benchmarks. + +## Evidence-to-score log + +| Date / revision | Evidence status | Overall change | Reviewer interpretation | +|:--|:--|:--:|:--| +| 2026-07-22 / `2304e83` | Audit of all completed frozen branches | baseline → 5 | Strong preservation/residualization core, but no standard useful-scale result | +| 2026-07-22 / `c753f51`, `1b24c87`, `6d19078` | Existing 60-run innovation panel promoted to a strict main figure; conditional-projection and norm-direction identities made executable | 5 → 5 | Closes a presentation/theory objection and makes the narrow novelty legible, but adds no new held-out evidence and therefore earns no score inflation | + +Future rows are appended only after an audited frozen stage. A score staying flat is informative: +engineering, theory exposition, or visualization may make the paper more defensible without +resolving the empirical objection that determines the recommendation. -- cgit v1.2.3