Evidence map
Action batching
The August 2 replication compared one ordered four-click request with four sequential click requests. The batch p50 was 10.34 ms. The sequential p50 was 21.05 ms. The measured p50 speedup was 2.04x. The replication retained 30 samples per arm. It had no failed or replacement samples. Placement and cleanup checks passed. The earlier July 29 artifact reported a 2.33x p50 speedup. Its raw sample arrays are unavailable. The later replication confirms the direction, but it does not reconstruct or supersede the July 29 run.Subprocess execution
The July 30 supporting artifact retains caller-observed, daemon, and transport-overhead sample arrays for three subprocess backends:asynciothreadedisolated-asyncio
Native X11 input
Use the runner matrix as the primary source for the runner-effect claim. The matrix varied two inputs:- Input backend: XTest or
xdotool - Subprocess backend: shared asyncio or isolated asyncio
xdotool move-click, the shared asyncio runner was slower than the isolated runner in every block. The pooled shared-runner daemon mean was 122.62 ms. This value fell within the preregistered 25 percent range around the historical 146.33 ms mean.
The launch-scaling control supports an approximately 45 ms shared-runner contribution per xdotool launch. A move-click uses two launches. This matches the approximately 90 ms matrix delta without treating the historical and clean runs as one distribution.
The XTest move-click control stayed within the matrix gate in every block. The matrix therefore supports a subprocess-runner contribution under the controlled modern conditions. It does not assign the full historical difference to one cause.
The historical 146.33 ms and 1.15 ms values came from a three-sample dirty-worktree diagnostic. The historical source manifest preserves the exact source and rounded aggregates. It does not preserve the original sample arrays or their digests.
The clean replication measured move-click daemon means of 32.78 ms for xdotool and 1.56 ms for XTest. These values reproduce the direction under a different topology. They are not current replacements for the historical values.
Sandbox termination evidence
The matrix runner recordedComputerSandbox.terminate(wait=True) followed by client close for each cell. Focused tests verify that order. They also verify cleanup on benchmark and termination failures.
A later read-only Modal lookup found all 12 Sandboxes in a finished state with exit code 137. Modal documents code 137 for a user-terminated Sandbox. The termination reconciliation does not prove an audit-log event, an invoice entry, or historical handle detachment.
Cost estimate
The recorded run wall time was 520.875 seconds. The formula applies that duration to one 1 CPU, 2 GiB Function and one 1 CPU, 2 GiB Sandbox. It then applies the recorded July 29 rates and a 1.75 region multiplier. The result is $0.06408. The rounded claim is about six cents.billing_reconciled is false. The estimate excludes image builds, storage, network transfer, model usage, plan fees, credits, discounts, taxes, and additional non-overlapping resource lifetimes. Treat it as a list-price reconstruction, not an invoice.
