PhiladelphiaMethodologyeasy
SPREE agrees with our DB on 616/616 cells — Philly's data pipeline is structurally clean
Thesis
Philly's (iii-a) SPREE spot-check shows 100% agreement at ±1.5pp tolerance across 616 paired PSSA/Keystone cells. Mean delta is -0.002, max absolute delta is 0.05. This is structurally tighter than NYC's analogous Snapshot agreement (62-73% across metrics) because Philly's data pipeline keeps SPREE and the public source files numerically aligned end-to-end. The reconciliation gap NYC has between published and ingested doesn't exist here.
Supporting findings
- 616 paired (school, year, metric, subgroup) cells; 616 within ±1.5pp.
- Mean Δ = -0.002, max abs Δ = 0.05.
- NYC comparison: 62-73% agreement on its Snapshot spot check, with deltas ranging -13 to +19 pp.
- Implication: Philly users can trust the served numbers match the family-facing SPREE.
- The structural difference is itself the story: NYC has cohort-attribution divergence between its source and its Snapshot; Philly doesn't.
Reporting directions
- Write up the validation methodology as a public methodology page (already in /philly/methodology).
- Reproduce for 2022-23 SPREE — does the 100% agreement hold across years?
- Methodology comparison piece: NYC vs Philly on validation-gap origin.
Methodology & replication recipe — the queries + sources behind the findings above.
More on methodology
- Philly publishes both Acct and Actual PSSA cuts. They're byte-identical.
- Computed-value (iv) validation: OSS + PSES rollups agree to the cell with the raw source
- Philly PSSA: 2023-24 vs 2024-25 year-over-year — which schools moved most
- With 287 peer-grouped schools and k=40, the rankings barely move at k=20