Measure the wait. Remove repeated work.

The evidence behind two focused improvements to Phantom’s engineering loop. Observed on 7 October 2026, with the scope and limits kept beside the numbers.

Release status at preparation: 3.24.3 is published and installed. The 3.24.4 candidate includes concurrent MCP discovery and same-request delivery projection reuse; release qualification is pending. The component preview below uses synthetic data.

A historical release-check baseline

All 15 jobs succeeded in a run lasting 39 min. The chart measures fixtures inside two Mac test lanes, rather than claiming that their durations add up to the complete release time. Runner scheduling, other jobs and native checks have their own costs.

General 3 fixtures account for about 90% of 33 minutes of test wall time; isolated fixtures account for about 98% of 30 minutes.
Observed fixture shares in the historical successful run. The categories reported by Vitest may overlap; they are not an additive decomposition of the whole workflow.
Accessible figure data, rounded for display
Mac laneFixture timeVitest wallShare
General 330 min33 minabout 90%
Isolated29 min30 minabout 98%

Inspect the successful GitHub Actions run. The exact interval, source and numeric values are in the downloadable method and data. This historical run does not qualify the new candidate.

Reuse work within the request

Four alternating baseline/candidate pairs used the same real private Git and preparation store, with controlled delivery, run and artifact witnesses. Output equality and unchanged fixture state were checked. One immediate redundant projection was removed; every new request still revalidates its inputs.

Redundant post-recovery projection reads fall from one to zero across four controlled pairs. Removed projection median: 97 milliseconds.
The median removed projection was 97 ms. Whole-helper medians were roughly 1.5 s and 1.4 s. These small local samples are not a matched whole-mission benchmark or a claimed percentage gain.

The separate real setup and mission tests passed all three cases, including restart, a second scope and Stop, at unchanged deadlines. They establish observed behavior; they do not establish a faster complete mission.

Let independent tool discovery start together

The candidate starts independent downstream MCP tool lists concurrently and merges responses in registry order. Deferred-response tests prove independent start without a fragile wall-clock threshold. Reverse completion, failed peers, native-name collisions and repeated fresh discovery are covered. There is no measured end-to-end server speed claim.

A candidate view of decision latency

Synthetic candidate phone component showing example mean decision response times and an unmeasured category. The image explicitly says source-only preview and no provider calls.
Phantom 3.24.4 candidate component fixture. Synthetic measurements; no provider calls. This is not production account usage. View the original preview.

The candidate makes response-time scope visible: request and decision processing count; model decoding speed is a different metric. Cached decisions are excluded, and missing timing stays unknown.

The numbers in this screenshot are illustrative fixture values. They are deliberately excluded from the measured CSV and performance conclusions.

We are qualifying the new Phantom umbrella homepage separately. phm.dev currently serves Phantom Secrets; this appendix does not claim that its upcoming workbench scene is already live.

Method, source and downloadable data

No private prompts, account names, credit balances, customer records or provider credentials are included. Display quantities are rounded; downloadable CSV retains observed precision.

Source revisions and scope

Historical baseline source · 3.24.4 candidate source · Published 3.24.3 release

Local controlled observations use Node 22.22.3. Background load may affect timings. No matched end-to-end speed comparison, customer adoption result or billing outcome is claimed.