SeyAero — aviation incident investigation platform logo

Evidence for the validation claim

Validation evidence

Section 5 of the Model Disclaimer states that the wreckage engine is validated against a broad corpus of historical hull losses with known recovery positions, that every case is re-run using only the information searchers held at the time with the recovery position withheld from the model, and that the results are reproducible from a published proof pack. This page is the evidence for each of those statements, and it carries the same limits: validation is retrospective and empirical, and does not guarantee performance on your occurrence.

Scored cases
37

185 scored runs across 5 independent seeds.

90% containment
89.2%

Share of runs where the recovery position fell inside the claimed 90% contour.

Median error
4.95 km

Mean 17.64 km · worst case 165.2 km.

Seed spread
0.75 km

Largest movement of the solution point across seeds — the answer is not seed luck.

1. Datasets summary

37 scored occurrences spanning 19852023, of which 23 ended over water (to a maximum charted depth of 4,900 m) and 14 over land. 2 carry a published real-world swept area, which is what the swept-area reduction figure is measured against. Inputs per case are the last known state, the meteorological and oceanographic layers for the occurrence window, and the elapsed drift interval. The confirmed recovery position is loaded for scoring only and is never visible to the engine.

CategoryCasesOver waterMedian depthMax drift intervalSpan
Shallow-water loss of control9955 m0 h19892021
Deep-water loss of control882,200 m0 h19872016
Land impact800 h19972023
In-flight break-up5336 m0 h19852015
Terrain-masked / CFIT400 h20102016
Ditching / surface drift3315 m18 h19962009

Excluded from scoring

  • MH370 Boeing 777-200ER · 2014-03-08No confirmed main wreckage position. Scored as 'no truth': the engine produces a prior, but there is nothing to score it against. Any product claim about MH370-class geometry is unvalidated.
  • Various Southern Ocean SAR Mixed · Long-drift open-ocean losses where only floating debris was recovered months later. Debris beaching points constrain origin only weakly and are excluded from scoring.

Occurrences without a confirmed recovery position cannot be scored, so they are kept out of every hit rate on this site. They stay in the corpus so the engine's behaviour on genuinely unresolved geometry remains visible and reviewable.

Full per-case detail — flight, airframe, last known state, met layers, recovery position and the published source for each — is rendered live on the validation record, which re-runs the engine in your own browser rather than serving stored numbers.

2. Methodology

  • Scoring is blind: the recovery position is withheld from the engine and used solely to measure error and containment after the run completes.
  • Only information available to searchers at the time is supplied — no hindsight tracks, no post-recovery corrections.
  • Each case is run across five fixed seeds so seed sensitivity is reported rather than hidden behind a single favourable run.
  • Containment is scored against the contour the engine itself claimed, which is what makes the calibration curve meaningful.
  • Baseline ablation compares the engine against a naive last-known-position circle, so the published gain is the gain over doing nothing clever.
  • Cases where the engine performed poorly are published alongside the rest; the worst-case error above is a real case, not a trimmed outlier.

3. Reproducibility

The published pack is deterministic. Re-running it from source must produce the same canonical bytes and therefore the same SHA-256 digest; if it does not, either the engine or the corpus changed and the claim on this page is stale.

Pack version
2.0.0
Published
2026-08-27
Canonical bytes
53,122
Seeds
20260601, 7, 1337, 99991, 424242
SHA-256 digest · 1ecd8b547553f63765d034316e66b5afd79a0d84946d4cb35bc73564db192650
  1. 01
    Fetch the proof

    GET /api/public/validation-proof — no account, no key. The endpoint re-runs the whole pack from source on request and reports whether it reproduced the published digest byte-for-byte.

  2. 02
    Compare digests

    Check the returned digest against 1ecd8b547553f637… and the canonical byte length against 53,122. Any divergent case is named individually in the response.

  3. 03
    Re-derive per case

    Append ?pack=1 to receive the full pack: per-case, per-seed solution point, error, claimed radii, containment, sector rank and swept area — enough to recompute every headline number yourself.

  4. 04
    Re-run in your own browser

    The validation record executes the same Monte Carlo engine client-side against the same inputs, so you can watch the numbers form rather than trust a stored table.

4. What this evidence does not establish

  • It does not establish that wreckage will be found inside any contour on your occurrence. Containment is a property of the corpus, measured retrospectively.
  • It does not validate geometry outside the corpus — long-drift open-ocean losses with no confirmed recovery position remain unvalidated, and are excluded from every figure above.
  • It does not validate the inputs you supply. An incorrect last known position or an inverted wind vector moves the answer by tens of kilometres with no warning in the output.
  • It says nothing about hypothesis ranking or AI transcription accuracy. Those are separate features, are not scored here, and remain unverified until a competent human verifies them.
  • It does not transfer any decision to the platform. Search deployment, grounding, airworthiness action, recommendations and settlement remain with the competent human authority.

This page restates the Model Disclaimer and does not vary it. Where the two differ, the Model Disclaimer and the Terms of Service prevail.