diff options
| author | YurenHao0426 <Blackhao0426@gmail.com> | 2026-07-22 13:43:05 -0500 |
|---|---|---|
| committer | YurenHao0426 <Blackhao0426@gmail.com> | 2026-07-22 13:43:05 -0500 |
| commit | 203e75987d02cdb8c317e3302e1f05c9a4beb0e3 (patch) | |
| tree | 25c237464e4babaca401222905311bcf295d1cd4 /REVIEW_SCORECARD.md | |
| parent | dba3eb29d5797c1b74d7e1e781523afad9faeb8e (diff) | |
results: close normalized mirror short accuracy gate
Diffstat (limited to 'REVIEW_SCORECARD.md')
| -rw-r--r-- | REVIEW_SCORECARD.md | 2 |
1 files changed, 2 insertions, 0 deletions
diff --git a/REVIEW_SCORECARD.md b/REVIEW_SCORECARD.md index bc5b4ca..93273ba 100644 --- a/REVIEW_SCORECARD.md +++ b/REVIEW_SCORECARD.md @@ -63,6 +63,7 @@ Every formal result report records: | Fixed hierarchical FA short gate fails | 5 | HFA reaches 43.52%, above DFA 37.16% and failed-v1 SDIL 41.98%, but below its frozen 50% full-run threshold | Spatial hierarchy helps; fixed random hierarchy remains insufficient and supplies no SDIL scale evidence | | Hierarchical task-scalar V4 fails | 5 | Stable calibration leaves early alignment near zero; higher rates explode to 58.5x norm or become nonfinite after 800 queries | Correct mechanics and the right spatial family are insufficient when one global scalar estimates 267,904 feedback parameters | | Normalized response mirror capture passes | 5 | Twenty local observations reach 0.446 early alignment and 0.915 feedback/forward cosine at zero task-loss queries | Strongly resolves the engineering bottleneck, but all credit belongs to an inherited weight-estimation baseline until innovation is load-bearing | +| Normalized response mirror short gate fails | 5 | WM reaches 64.04% and 0.939 early alignment at 0.9968x BP MACs, but misses its two accuracy gates by under one point | Strong inherited baseline and useful warning that alignment is insufficient; full run and any SDIL scale claim remain closed | | Oral-A A4 | not opened | The prerequisite A3 gate failed | No oral-A confirmation claim is available | These are conditional reviewer forecasts, not promised scores. A failed stage leaves its negative @@ -84,6 +85,7 @@ retroactively reopened by success on standard vision benchmarks. | 2026-07-22 / fixed HFA S1 | Three clean matched ResNet-20 records; selected HFA reaches 43.52% and 0.0404 early alignment but misses the frozen 50% gate | 5 → 5 | Establishes a stronger zero-query local baseline and confirms hierarchy helps, while closing an uncalibrated full run and leaving standard SDIL evidence absent | | 2026-07-22 / hierarchical V4-1 | Four frozen-forward records: etaA 0.1 leaves early alignment near zero, etaA 1 explodes one feedback norm to 58.5x, and etaA 10 is nonfinite | 5 → 5 | Closes global-task-scalar calibration of the full hierarchy at the fixed budget; richer local information is required before another accuracy endpoint | | 2026-07-22 / response-mirror WM-1 | Four clean frozen-forward records; selected etaM 0.1 reaches 0.4461 early and 0.5422 all-layer alignment with zero task-loss queries | 5 → 5 | Opens a strong inherited-baseline accuracy test and isolates information source as the bottleneck, but supplies no Harnett-specific task evidence | +| 2026-07-22 / response-mirror WM-2 | Two clean short ResNet records; selected WM reaches 64.04% with 0.9393 early alignment and 0.9968x BP MACs, missing both accuracy gates narrowly | 5 → 5 | Substantially strengthens the comparator and cost story, but closes its full run and demonstrates that high credit cosine alone is not a scale result | Future rows are appended only after an audited frozen stage. A score staying flat is informative: engineering, theory exposition, or visualization may make the paper more defensible without |
