diff options
| author | YurenHao0426 <Blackhao0426@gmail.com> | 2026-07-22 12:19:44 -0500 |
|---|---|---|
| committer | YurenHao0426 <Blackhao0426@gmail.com> | 2026-07-22 12:19:44 -0500 |
| commit | 9763a8ecf4d614370cfde3ff926c5070ab766433 (patch) | |
| tree | 4ed0b067bfcf4a16f811f37b5723d728ffb5db74 /REVIEW_SCORECARD.md | |
| parent | 04de9326e60bb9e3592c0ddb5474265d228cbce0 (diff) | |
baselines: complete native author-code audit
Diffstat (limited to 'REVIEW_SCORECARD.md')
| -rw-r--r-- | REVIEW_SCORECARD.md | 7 |
1 files changed, 4 insertions, 3 deletions
diff --git a/REVIEW_SCORECARD.md b/REVIEW_SCORECARD.md index 25d1247..861ba0e 100644 --- a/REVIEW_SCORECARD.md +++ b/REVIEW_SCORECARD.md @@ -47,15 +47,15 @@ Every formal result report records: 3. Innovation was not uniformly beneficial for arbitrary endogenous top-down traffic, so the supported mechanism is narrower than the initial claim. 4. The frozen oral-B screen falsified the desired-velocity and Harnett error-derivative claims. -5. The frozen ResNet-20 A3 run became nonfinite and ended at chance; native Dual Prop is the only - remaining incomplete baseline endpoint. The failed gate forbids the A4 test panel. +5. The frozen ResNet-20 A3 run became nonfinite and ended at chance. The failed gate forbids the + A4 test panel; completed native baselines improve fairness but do not supply SDIL scale evidence. ## Score trajectory and prospective gates | Checkpoint | Overall | What changed | Remaining ceiling | |:--|--:|:--|:--| | Current audited package | 5 | Strong depth-preservation and residual-necessity evidence; negative gates retained | Standard useful scale is absent | -| Native baselines complete | pending | Can close fairness/completeness objections, but cannot by itself establish the main claim | Usually no automatic score increase | +| Native baselines complete | 5 | BurstCCN is below its published endpoint; Dual Prop reproduces 92.46% versus 92.41%, with strict provenance and cost semantics | Fairness objection narrows, but SDIL gains no standard-scale evidence | | Oral-A A1/A2 | 5 | BP reached 91.62%; short channel-gated SDIL reached 41.98% versus tuned DFA at 37.16% | Development screening alone cannot raise the score | | Oral-A A3 fails | 5 | Full ResNet-20 SDIL became nonfinite at epoch 90 and ended at 10%; DFA ended finite at 33.06% | Standard-scale and oral-A claims are closed; A4 remains untouched | | Oral-A A4 | not opened | The prerequisite A3 gate failed | No oral-A confirmation claim is available | @@ -71,6 +71,7 @@ retroactively reopened by success on standard vision benchmarks. | 2026-07-22 / `2304e83` | Audit of all completed frozen branches | baseline → 5 | Strong preservation/residualization core, but no standard useful-scale result | | 2026-07-22 / `c753f51`, `1b24c87`, `6d19078` | Existing 60-run innovation panel promoted to a strict main figure; conditional-projection and norm-direction identities made executable | 5 → 5 | Closes a presentation/theory objection and makes the narrow novelty legible, but adds no new held-out evidence and therefore earns no score inflation | | 2026-07-22 / frozen Oral-A A1--A3 | A1 and A2 pass; full A3 SDIL becomes nonfinite and fails four of six checks; A4 untouched | 5 → 5 | Closes the standard-scale question negatively. The narrow mechanism paper survives, while any standard-ResNet or oral claim does not | +| 2026-07-22 / native C4 | BurstCCN and Dual Prop author-code records pass the strict audit; Dual Prop reproduces 92.46% test in 23119.8 s | 5 → 5 | Closes a baseline-fidelity objection and confirms a strong expensive comparator, but does not repair SDIL's failed A3 evidence | Future rows are appended only after an audited frozen stage. A score staying flat is informative: engineering, theory exposition, or visualization may make the paper more defensible without |
