Replaying every validation decision.
33 recordings, eight filename groups and 121,325 eligible decisions. Compare original offline scores with two ways to process one available input.
Ordinary single-reading inference fails the unchanged numerical tolerance. Eight copies of the same current input reproduce the original scores, with extra computation. This finite validation result promotes no model; all eight reserved groups remain unopened.
All six frozen profiles
| Checkpoint | Processing method | Failed scores | Alarm changes | Median call | P95 call | Maximum call |
|---|
Check each filename group
| Public filename group | Eligible decisions | Scores outside tolerance | Maximum score difference | Alarm changes |
|---|
What these matches establish
Original offline inference processes all 140,708 grid rows in batches of 256, then masks unusable inputs. The final batch has 164 eligible rows. Both current-only methods use an explicit contiguous boundary; the eight-copy method consumes only the same one available feature, with no future or neighboring row, waiting or cached substitute.
Model, scaler and threshold bytes stay unchanged. Full grid identities and first-128 stored feature bytes match. This is a numerical adaptation on observed validation data, with added compute. A complete original feature tensor was not stored; the newly recorded full feature hash is not an original full-tensor comparison.
The original Mac aggregate replay stays FAILED. Internal kernel causes remain unisolated. Neither useful fresh detection nor physical calibration is established; no inference profile is automatically admitted.
Compute and reproducibility boundary
Every raw call time and first call is retained. Profiles run once in fixed order; timings exclude acquisition and rolling raw-feature work. Lifetime memory includes full-stream matrices and runtime, and is not incremental edge-model memory or a target-device deadline.
Separate saved-array review verifies 727,950 eligible scores, 844,248 full-grid alarm values, 486,400 raw call times, 48 group rows and twelve common/native metric reports. It runs no learned inference or raw measurement decoding and has no independent human sign-off.
Inspect and reproduce
Original output SHA256:
Scientific artifact revision: