The shipped single-person scoring math is a faithful, reproducible port of the canonical engine.
EstablishedDimension: Software reliability and reproducibility
- Evidence
- The boost-app TypeScript scoring core matches the canonical Python engine on 78 of 78 cells, exact to the integer, across every canonical scenario and context (the 5 core plus 8 single-person indices). Output is deterministic: 0 variance over 10 repeated runs of the full pipeline. The three surfaces conform to 0 delta on the FACE-ONLY core at a shared frame rate.
- Method
- A parity shim runs the shipped core and the canonical Python engine on shared single-person fixtures and compares all 13 score fields cell by cell; determinism and conformance are measured by repeated and cross-surface runs on fixed input.
- Effect size and uncertainty
- 78/78 exact match on the 5 core plus single-person indices; 0.0 variance; 0.0 FACE-ONLY cross-surface delta at a shared frame rate.
- Threats to validity
- This is software reliability and reproducibility, not validity. Two scope limits carry forward: (a) parity covers the 5 core and single-person indices only. The multi-person AGGREGATE scoring path (deriveMultiPersonScores), which is the shipped Proof of Impact room read, is NOT covered by the 78/78, and its parity is a separate, unrun exercise. (b) Conformance is 0 delta on the FACE-ONLY core; body-dependent scores diverge up to 10 points (Stress most) across surfaces because body extraction runs only on web (D8, open). Two programs agreeing on a number is not the number being right.
Sources: docs/validation/engine-parity.md, docs/engine/02-conformance.md, docs/validation/03-local-validation-results.md