OP

Insights

Patterns across all 27 kit missions: where inspections fail, how battery and configuration matter, and how the dispatcher rules would have behaved on real weather.

Before you trust these charts
Data traps in the kit. Each number is recomputed from the kit files on every load.
  • 77% vs 91% raw

    V1 vs V2 is confounded by batterySee the chart

    V2 is better, but part of its raw lead is battery: V2 never ran below 52%. Same battery band: 40-59%: 85% vs 100%; 60-79%: 86% vs 92%.

    Measured: Action success by configuration, then again inside 20-point battery bands where both configurations ran.

  • 61 of 270

    Asset and location disagree

    Do not map an asset by the action's location alone. Check assets.csv before sending a crew.

    Measured: actions.location_id compared with the asset's location_id in assets.csv.

  • 50 of 270

    Actions with no reading

    A missing reading is not a normal reading. Treat the asset as uninspected.

    Measured: Actions with no row in readings.csv, matched on action_id.

  • 23 of 27 missions

    The replay stream is incomplete

    Do not use the live feed as the record. It stops mid-mission (M023). It has 0 MISSION_FAILED events, yet missions.csv lists 1 failed.

    Measured: Distinct mission_id in replay.ndjson; last mission checked for MISSION_COMPLETED; event types counted.

  • 8 of 40

    Anomalies hide in successful actions

    A SUCCESS action can still find a problem. Review anomalies, not only failed actions.

    Measured: anomalies.csv joined to actions.csv on action_id, counting status SUCCESS.

Battery vs. inspection success
V2 never ran below 52% battery, so part of its lead over V1 is battery, not configuration
0-19%
V127% (3/11)
V2no runs
20-39%
V172% (43/60)
V2no runs
40-59%
V185% (56/66)
V2100% (7/7)
60-79%
V186% (37/43)
V292% (49/53)
80-99%
V1no runs
V287% (26/30)
Least reliable inspections
Share of actions that succeeded, all 27 missions
acoustic70% (30/43)
gauge reading74% (28/38)
navigation82% (59/72)
thermal86% (42/49)
visual91% (62/68)
Configuration V1
77% actions OK · 29 min · 56% battery per mission
Configuration V2
91% actions OK · 22 min · 28% battery per mission
Backtest: would it cry wolf?
Today's rules run on the real archived weather at St. Lucie Plant, FL for each kit round time

1 of 27 rounds held for weather

  • M001 Rain up to 3.5 mm/h before Spot can dock; Spot is rated for light rain and splashes only (IP54)

9 of 27 rounds held for battery (would dock under 20%)

17 of 27 routes include the outdoor leg. Weather: Open-Meteo ERA5 archive, Sep 1-10, 2026. The archive has no thunder data, so this count does not test the thunder rule.

Loading decisions…

Mission data is NextEra's synthetic challenge kit: indoor facility, no weather fields. Kit times are shown as recorded.

Assumptions: site is St. Lucie Plant, FL (assumed site); WP01's entrance apron is the outdoor leg; WP15 is the dock; kit x/y units are meters; Spot walks at 1 m/s; thresholds marked "our assumption" await NextEra's procedure limits. Orbit calls are simulated.