Files
PaperClipAI/tests/runner-e2e/remote-native-bootstrap.ts
T
DottaandPaperclip 57e977be72 feat: integrate Pi 1.0 into the experimental Runner (#14921)
## Thinking Path

> - Paperclip manages AI agents and their work.
> - The experimental Runner owns provider processes and durable
sessions.
> - Pi needs working task execution and human controls.
> - The five-PR stack must preserve changes already on master.
> - Each layer now carries the complete integrated source for a safe
sequential fallback.
> - This PR belongs to native GitHub stack #15602, ending at #14956.

## Linked Issues or Issue Description

Refs #14436, #14631, #14743 and #14956.

Ship Pi 1.0 through the experimental Paperclip Runner. The five PRs are
#14921, #14922, #14923, #14924 and #14956. The user authorized the
complete merge after checks pass. Existing `pi_local` execution is
unchanged. Accounting and wider provider/platform qualification remain
deferred.

## What Changed

- Recover missing final replies after workspace finalization changes
owners, using accepted-turn evidence without rerunning work or granting
external-chat publication.
- Preserve the admitted Pi instruction root across warm runs, while
retaining changed-root rejection.
- Give Pi a bounded 15-second default shutdown grace so stop, drain
acknowledgement and durable suspension can complete. Explicit deadlines
and other providers retain their existing behavior.
- Integrate the Pi 1.0 runtime and master contracts.
- Use Pi profile 22. Preserve explicit caller-selected models and exact
native thinking levels. Keep Pi's wrapper, helper, extension and
question/control behavior unchanged from the qualified profile-19
runtime.
- Preserve master's Dot lifecycle and consent fields, configured task
environment, status guards and current Codex/Claude dependency versions.
Cursor stays qualified. Copilot stays pending; profile 17 binds the
changed shared protocol validation sources.
- Exclude general AWS IAM credentials from Pi static/custom provider
bindings and selected task projections; preserve the provider-scoped
Bedrock bearer key. Profile 21 is retained as historical provenance.
Rust and cloud install probes use the current declaration.
- Patch bundled brace-expansion 5.0.9 to the exact official 5.0.12
payload. Pin the patch and complete runtime closures. Include the patch
in normal installed setup tooling. Keep the upstream Pi shrinkwrap as
provenance and permit only this exact security correction.
- Include current attestation files in the Docker build context. Keep
the repository lockfile unchanged from master. CI and private image
builds resolve manifest changes before their frozen installation.

## Verification

- Full local `pnpm -r typecheck` passes, including Runner Rust, server
and UI. Focused integration checks pass: 194 Runner
admission/environment tests, 63 profile/credential tests with one
expected skip, 152 Dot/UI configuration tests, and Pi transcript/notice
tests.
- Full local `pnpm build` passes on the final source.
- Fresh final-source checks pass: all 698 Rust workspace tests (32
binaries), 156 credential/profile/controller tests with one expected
skip, Runner TypeScript typecheck, and 20 package/setup/sandbox tests.
- The profile-21 Pi materializer passes on the native host with the
official pinned Node 24.21.0 and its npm. It verifies all 150 locked
packages, the patched dependency and the exact closure. Setup/package
bundle tests and UI token gates pass.
- The old hashes were reproduced for all three supported targets before
calculating the patched graph. New closure hashes are darwin-arm64
`282022db10150c6632b3444df421342e7d534bdf5d5fb1097a2e79d0625a2bcf`,
darwin-x64
`64e251e19009f755c0b04f73ce2138246faab71a961b0f13d75ebfcc34bef12e`, and
linux-x64
`713b1fdff42fb56a1518bdc084f181d70bee8ebadc3e4b1d76321ed9108c8410`.
Independent native platform execution is separate from graph identity
reproduction.
- Historical cloud qualification remains unchanged: all seven core cases
pass on shipping source `10dc43c9ec65d88c2f782d62afb296d09494f215`,
harness `1a4408a48cfb5a1f094a311141c257c92cd7a893`, image
`sha256:5b3a775b383591bda1b0c1889e509acc70ce7f37c53f09733c81d59037f02280`,
and accepted Sonnet 4.6/low fixture. All 215 canonical files and all
seven cleanup checks pass independent verification. These are profile-19
results and are not relabeled as fresh profile-22 runs.
- Current Pi digest:
`sha256:e92078bee3c23bec4100aa589013a44613d054cd686826534025d8019e9f39a9`.
[The readiness
plan](https://github.com/paperclipai/paperclip/blob/codex/pi-production-readiness/doc/plans/2026-10-02-pi-production-readiness.md)
preserves campaign and failed-attempt provenance.
- Merge only after every PR's current-head CI and fresh review pass.
Linux CI covers the full suites, build and browser tests. The local
embedded Postgres API-authority suite cannot start on this macOS/Node 26
host, so Linux CI must confirm that suite.

### Fresh profile-22 core qualification — 2026-10-08

All seven accepted core cases pass canonically on Pi profile 22, with
`openrouter/anthropic/claude-sonnet-4.6` and native-confirmed low
thinking. This model is a fixture; production accepts the caller's
explicit Pi provider/model.

Runtime/install source: `3241a992f2a7703e59e97ed0fd3e5d6405de4401`.
Frozen accepted harness: `1a4408a48cfb5a1f094a311141c257c92cd7a893`.
Immutable cloud image:
`ghcr.io/paperclipai/paperclip-daytona-runner@sha256:506f22db7edd78f37c0c40bec1cc084af1850455026dbf467194bfbb8fcef141`.
Pi digest:
`sha256:e92078bee3c23bec4100aa589013a44613d054cd686826534025d8019e9f39a9`.

[Hosted Linux image and clean-install
verification](https://github.com/paperclipai/paperclip/actions/runs/37868328023)
passes, including all 20 source-bound archives, normal CLI/Pi setup,
companion import and the production pack reader. This exact installation
source includes the latest master integration and the corrected Pi warm
instruction-root fence. Full local typecheck/build and current-head
hosted CI verify the final stack. All 13 focused real-root regressions
pass. The full local executor suite passed 662 tests; 15 database tests
could not start the Mac embedded PostgreSQL service. Hosted Linux CI
passes the full required verification and E2E checks. These fresh
results keep their own source identity; profile-19 results remain
historical.

| Core path | Canonical campaign | Retained archive SHA-256 |
| --- | --- | --- |
| File edit, validation, download and Done |
`pi-core22-replyfix-0-1791511228` | 23 files;
`a473e8603a3dd4737863291f8d3d1e392391f0b16d433c3e0e0e9d8baf7a97b0` |
| Pending question and controller restart |
`pi-core22-replyfix-1-1791511376` | 33 files;
`6b829c4eb74e1f32a89c692a4ae7130dbfc1c6d3cf13915effe2103d9e242c8e` |
| Three-turn session/process/workspace continuity |
`pi-core22-replyfix-2-1791511587` | 23 files;
`7a87021f8f9a3fdd3c58bb4467f8d82c635e3ea4795d6e75f144d9aa14818df8` |
| Four typed questions and browser reconnects |
`pi-core22-replyfix-3-1791511881` | 42 files;
`9e31755252be1f4f9cb0626c984c142d4d1ae5f5bee3a7af08444db8d12c280a` |
| Plan approval and completion | `pi-core22-replyfix-4-1791512031` | 22
files;
`a0383ce1aab38e7b5a25ce0e9dd3bebea5c037ebd96ae6b29dae19015da2ae2c` |
| Same-turn steering and permission denial |
`pi-core22-replyfix-5-1791512261` | 39 files;
`c929b8c7070f0b66aedc17e65ca46e6beab1e363926ac9f7e2a75fb250f05949` |
| Stop during pending permission | `pi-core22-replyfix-6-1791512390` |
33 files;
`7f0a58ae0f4d5bfc76149435f4e322537089c5bd16e7ffe9b5ad71f10a621a07` |

All 215 canonical files (28714587 bytes) are independently
hash-verified. All seven cleanup grades pass, with no owned runtime
process or temporary root after each case. Automatic retries are zero.
The owned cloud host stopped normally after retention. The prior
profile-22 warm attempt remains failed and separately retained: archive
SHA-256
`1e54eba5ec72b50cee1534b23d1d1d4f21a090006b8a64501ba70db972abfde5`. Its
original canonical classification is preserved. Diagnosis reproduced a
product bug comparing an agent-files root against an unset
checkpoint-only field. The fix stores the admitted physical root
separately from the adopted per-run collection capability. The real-root
regression fails before the fix and passes afterward, including
rejection of a changed physical root. Fixture, grader, model and all
seven accepted case IDs are unchanged; this fresh campaign tests
final-reply publication after file registration first. The intermediate
restart attempt also remains failed and retained: archive SHA-256
`5dcaefdf1d17cf4cd54fd4cf810f45e736667392339b8ce7caf08bb4e225277f`. Its
original canonical classification is preserved. Pi resumed, wrote the
verified answer and completed its task; exact runner suspension was
proven, but idle stop consumed about 5.2s and left under 3s for the
drain acknowledgement. The Pi-only default shutdown grace is now 15s,
preserving a full 5s drain round trip and a finite suspension reserve.
Explicit caller deadlines, other provider defaults, literal drain
receipts and exact suspension identity checks remain unchanged. The
timing regression fails before this correction and passes afterward; all
18 focused settlement tests and Runner typecheck pass. The final-source
file attempt is also preserved as failed (`candidate_failure`), archive
SHA-256
`db6767b6773ea618997927ac77bdb005a5ac81492c7b9c0ffbc900449f829bc9`.
Native edit, validation, exact downloadable artifact and Done/succeeded
all passed, and the exact final reply was durably recorded. A workspace
recovery owner completed before the live heartbeat reached presentation,
leaving that reply absent from task chat. Recovery now materializes only
a completed final reply from the accepted turn of an ordinary internal
Done task, preserving issue/run/contract binding, suppression,
external-chat authorization and same-run deduplication. The database
regression covers the generated file-preparation receipt, suppression,
unapproved external continuation and replay. Server typecheck and all 49
response-selection tests pass; hosted Linux verifies the database
regression because embedded PostgreSQL cannot start on this Mac.

The delayed-final-answer database regression passes on [the final
root-source Linux server
shard](https://github.com/paperclipai/paperclip/actions/runs/37868262553/job/113628594152),
alongside 1,108 passing tests. The first root Runner shard had one
unchanged durable-resume test exceed its 5-second timeout; the identical
top-source shard and the isolated exact test passed. One rerun of that
failed job and its required aggregate passed without source or test
changes. The original failed job log and the single-rerun receipt remain
retained.

### October 9 merge verification

Current merge head: `5a8fe63512a7166aaef5cf50065a25008aa8b44b`. All
current-head checks pass, including `ci / verify` and `ci / e2e`;
exact-head Greptile review is 5/5 with no unresolved threads. Current
master conflicts are resolved. The user authorized the maintainer
override of the code-owner review gate after these checks. The seven
retained live core cases remain bound to source
`3241a992f2a7703e59e97ed0fd3e5d6405de4401` and its recorded cloud image.

## Risks

- The security correction changes the dependency closure and profile
identity. Old sessions must reopen on the new profile. Exact identities
and credential bindings fail closed.
- The runner remains experimental and requires explicit selection.
Legacy Pi Local is unchanged. Caller model IDs pass through; the E2E
model is a fixture.
- Accounting and the broad platform/provider matrix remain deferred.
This merge does not publish a release or deploy a service.

## Model Used

OpenAI GPT-6 through Codex assisted with reasoning, repository
inspection, editing and tool use. The exact serving ID and context
window are not exposed in this session. Final live qualification uses Pi
1.0.0 with `openrouter/anthropic/claude-sonnet-4.6` and native-confirmed
low thinking.

## Checklist

- [x] I have included a thinking path that traces from project context
to this change
- [x] I have specified the model used (with version and capability
details)
- [x] I have checked ROADMAP.md and confirmed this PR does not duplicate
planned core work
- [x] I have searched GitHub for duplicate or related PRs and linked
them above
- [x] I have either (a) linked existing issues with `Fixes: #` / `Closes
#` OR (b) described the issue in-PR following the relevant issue
template
- [x] I have not referenced internal/instance-local Paperclip issues or
links (only public GitHub `#NNN` / `github.com/paperclipai/paperclip`
URLs)
- [x] My branch name describes the change (e.g. `docs/...`, `fix/...`)
and contains no internal Paperclip ticket id or instance-derived details
- [x] I have run tests locally and they pass
- [x] I have added or updated tests where applicable
- [x] I have updated relevant documentation to reflect my changes
- [x] I have considered and documented any risks above
- [x] All Paperclip CI gates are green
- [x] Greptile is 5/5 with no open P2s, recommendations, or follow-ups
- [x] I will address all Greptile and reviewer comments before
requesting merge

---------

Co-authored-by: Paperclip <noreply@paperclip.ing>
2026-10-09 08:56:41 -05:00

221 lines
15 KiB
TypeScript

import { randomBytes } from "node:crypto";
import { readFile } from "node:fs/promises";
import { createRequire } from "node:module";
import { dirname, join } from "node:path";
import { pathToFileURL } from "node:url";
import { ObservedStateTimeout, RemoteAdmissionReadError, RunnerApiHttpError, pollUntil } from "./api.js";
import { classifyFailure } from "./failure-classifier.js";
import { REMOTE_FIXTURE_MIN_SETUP_BUDGET_MS, bindRemoteNativeFixture, remoteNativeFixtureDiagnostics, type RemoteFixtureApi, type RemoteFixtureDaytona, type RemoteNativeFixture } from "./remote-native-fixtures.js";
type AdmissionEndpoint = "issue" | "run" | "leases";
function admissionReadError(endpoint: AdmissionEndpoint, error: unknown): RemoteAdmissionReadError {
let failureClass: RemoteAdmissionReadError["failureClass"] = "candidate_failure";
// Only typed status or transport diagnostics classify the failure. In
// particular, an HTTP body must not supply classification keywords.
try {
if (error instanceof RunnerApiHttpError) {
if (error.status === 408 || error.status === 429 || (error.status >= 500 && error.status <= 599)) failureClass = "transient_infrastructure";
else if (error.status === 401 || error.status === 403) failureClass = "permanent_infrastructure";
} else if (error instanceof Error) {
const message = error.message.slice(0, 256);
if (error.name === "TimeoutError" || /^(?:apiRequestContext\.get: )?Timeout \d+ms exceeded\b/u.test(message)
|| /\b(?:ECONNRESET|ECONNREFUSED|ETIMEDOUT|ENOTFOUND|EAI_AGAIN|socket hang up)\b/iu.test(message)) failureClass = "transient_infrastructure";
}
} catch { /* Unknown rejection values remain a bounded candidate failure. */ }
return new RemoteAdmissionReadError(endpoint, failureClass);
}
function admissionJson(endpoint: AdmissionEndpoint, value: unknown): boolean {
const record = (item: unknown): item is Record<string, unknown> => item !== null && typeof item === "object" && !Array.isArray(item);
return endpoint === "leases" ? Array.isArray(value) && value.every(record) : record(value);
}
export interface RemoteNativeBootstrap {
prompt(nonce: string): string;
bindAndRelease(input: {
issueId: string; runId: string; targets: readonly string[]; crossRoot?: { initialText: string };
actionPrompt(fixture: RemoteNativeFixture): string | Promise<string>;
}): Promise<RemoteNativeFixture>;
}
/** Use the repository's pinned SDK, never an ambient CLI, SDK or credential. */
export async function createRemoteFixtureClient(apiKey: string): Promise<RemoteFixtureDaytona> {
if (!apiKey.trim()) throw new Error("Remote native qualification requires an explicitly bound Daytona credential");
const require = createRequire(new URL("../../packages/plugins/sandbox-providers/daytona/package.json", import.meta.url));
const entry = require.resolve("@daytonaio/sdk");
let directory = dirname(entry), version: string | undefined;
for (let i = 0; i < 4; i++, directory = dirname(directory)) {
try {
const pkg = JSON.parse(await readFile(join(directory, "package.json"), "utf8"));
if (pkg.name === "@daytonaio/sdk") { version = pkg.version; break; }
} catch (error) { if ((error as NodeJS.ErrnoException).code !== "ENOENT") throw error; }
}
if (version !== "0.203.0") throw new Error("Remote native qualification requires Daytona SDK 0.203.0");
const sdk = await import(pathToFileURL(entry).href);
return new sdk.Daytona({ apiKey, apiUrl: "https://app.daytona.io/api" }) as RemoteFixtureDaytona;
}
/** The tested action is withheld until its observer is armed. The initial
* provider prompt only reads this one operator-published instruction file.
* No production hook, steering support or remote setup-command execution is
* assumed. Bootstrap reads are distinct from the native action under test. */
export function createRemoteNativeBootstrap(input: {
api: RemoteFixtureApi; daytona: RemoteFixtureDaytona; companyId: string; environmentId: string;
agentId: string; image: string; nodeSha256: string; runnerdSha256: string; deadlineAt: number;
evidence(name: string, data: unknown): Promise<void>;
}, bind = bindRemoteNativeFixture): RemoteNativeBootstrap {
if (!/^sha256:[a-f0-9]{64}$/u.test(input.nodeSha256) || !/^sha256:[a-f0-9]{64}$/u.test(input.runnerdSha256) || !/@sha256:[a-f0-9]{64}$/u.test(input.image)) {
throw new Error("Remote native qualification requires immutable image and interpreter digests");
}
const bindings = new Map<string, { path: string; consumed: boolean }>();
return {
prompt(nonce) {
if (!/^[A-Za-z0-9_-]{1,128}$/u.test(nonce) || bindings.has(nonce)) throw new Error("Invalid or reused remote bootstrap nonce");
const path = `.paperclip-eval-action-${randomBytes(18).toString("hex")}.txt`;
bindings.set(nonce, { path, consumed: false });
return [
`Your task instructions will be published by the operator in the workspace file ${path}.`,
"Use your native file-read tool to read that exact relative file in the current execution workspace. If it is not present yet, retry the read for up to 20 seconds, then report the setup failure.",
"Before reading those instructions: do not infer the task. Do not run shell commands. Do not create or modify any file.",
"Do not ask replacement questions. Do not mark work complete. Do not create the missing instruction file.",
"After reading the complete file, perform precisely the supplied task. Treat it as the operator's continuation of this task.",
].join("\n");
},
async bindAndRelease(request) {
const pending = [...bindings.values()].filter(value => !value.consumed);
if (pending.length !== 1) throw new Error("Remote bootstrap requires one unconsumed instruction delivery");
const setup = pending[0]!; setup.consumed = true;
if (![request.issueId, request.runId].every(id => /^[A-Za-z0-9_-]{1,128}$/u.test(id))) throw new Error("Invalid bootstrap task/run identity");
// Provisioning consumes the authored case deadline. A cold snapshot need
// not produce a lease in 20 seconds; a terminal or foreign run never waits.
let lastState: Record<string, string | number | boolean> = { observed: false };
const admissionDeadlineAt = input.deadlineAt - REMOTE_FIXTURE_MIN_SETUP_BUDGET_MS;
let fixture: RemoteNativeFixture;
let lastReadError: RemoteAdmissionReadError | undefined;
try {
const ready = await pollUntil({
label: "owned native qualification lease", deadlineAt: admissionDeadlineAt, intervalMs: 100,
load: async () => {
type Read<T> = { status: "fulfilled"; value: T } | { status: "rejected"; reason: RemoteAdmissionReadError } | { status: "pending" };
let issueRead: Read<Record<string, any>> = { status: "pending" };
let runRead: Read<Record<string, any>> = { status: "pending" };
let leasesRead: Read<Array<Record<string, any>>> = { status: "pending" };
const snapshot = () => {
const issue = issueRead.status === "fulfilled" ? issueRead.value : undefined;
const run = runRead.status === "fulfilled" ? runRead.value : undefined;
const rows = leasesRead.status === "fulfilled" ? leasesRead.value : undefined;
const issueOwned = issue?.id === request.issueId && issue.companyId === input.companyId
&& issue.assigneeAgentId === input.agentId;
const runOwned = run?.companyId === input.companyId && run.agentId === input.agentId && run.id === request.runId;
const owned = issueOwned && runOwned;
const runLeases = Array.isArray(rows) ? rows.filter(row => row?.heartbeatRunId === request.runId) : undefined;
const active = runLeases?.filter(row => row.issueId === request.issueId && row.status === "active" && row.providerLeaseId);
// A failed endpoint cannot discard a successful read proving that
// ownership changed or the run stopped. Missing reads never admit.
const rejection = (issueRead.status === "fulfilled" && !issueOwned) || (runRead.status === "fulfilled" && !runOwned)
? "Remote bootstrap task/run ownership is unproven"
: runRead.status === "fulfilled" && !["queued", "running"].includes(run?.status) ? "Native qualification run stopped before observer setup"
: runLeases && runLeases.length > 1 ? "Ambiguous native qualification lease" : undefined;
// Fixed keys and allowlisted enum values only: no provider IDs,
// errors, prompts or arbitrary API strings enter startup evidence.
const readError = [issueRead, runRead, leasesRead].find(read => read.status === "rejected");
const evidence = {
observed: true, owned, phase: "lease_admission",
issueRead: issueRead.status, runRead: runRead.status, leasesRead: leasesRead.status,
runStatus: ["queued", "running", "succeeded", "failed", "cancelled", "timed_out"].includes(run?.status) ? run!.status : "unknown",
executionStage: ["queued", "preparing", "executing", "finalizing"].includes(run?.executionStage) ? run!.executionStage : "unknown",
...(runLeases ? { runLeaseCount: runLeases.length, activeOwnedLeaseCount: active!.length } : {}),
...(readError?.status === "rejected" ? { readFailureClass: classifyFailure(readError.reason) } : {}),
deadlineReached: Date.now() >= input.deadlineAt,
};
return { run, owned, lease: active?.[0], rejection, evidence,
readError: readError?.status === "rejected" ? readError.reason : undefined };
};
// At most three in-flight reads. Each has a transport deadline no
// later than admission; terminal/ownership proof need not await it.
const remainingMs = Math.max(1, admissionDeadlineAt - Date.now());
const timeout = Math.min(30_000, remainingMs);
const state = await new Promise<ReturnType<typeof snapshot>>(resolve => {
let finished = false;
const finish = (state: ReturnType<typeof snapshot>) => {
if (finished) return;
finished = true; clearTimeout(timer); resolve(state);
};
const timer = setTimeout(() => {
const failure = (endpoint: AdmissionEndpoint) => ({ status: "rejected" as const, reason: new RemoteAdmissionReadError(endpoint, "transient_infrastructure") });
if (issueRead.status === "pending") issueRead = failure("issue");
if (runRead.status === "pending") runRead = failure("run");
if (leasesRead.status === "pending") leasesRead = failure("leases");
finish(snapshot());
}, remainingMs);
function read<T>(endpoint: AdmissionEndpoint, path: string, save: (result: Read<T>) => void) {
// Both handlers are installed before dispatch. Late completion
// is consumed but cannot mutate saved evidence or start a poll.
void Promise.resolve().then(() => input.api.get<T>(path, { timeout })).then(
value => settled(admissionJson(endpoint, value) ? { status: "fulfilled", value }
: { status: "rejected", reason: admissionReadError(endpoint, undefined) }),
reason => settled({ status: "rejected", reason: admissionReadError(endpoint, reason) }),
);
function settled(result: Read<T>) {
if (finished) return;
save(result);
const state = snapshot();
if (state.rejection || [issueRead, runRead, leasesRead].every(read => read.status !== "pending")) finish(state);
}
}
read<Record<string, any>>("issue", `/api/issues/${request.issueId}`, result => { issueRead = result; });
read<Record<string, any>>("run", `/api/heartbeat-runs/${request.runId}`, result => { runRead = result; });
read<Array<Record<string, any>>>("leases", `/api/environments/${input.environmentId}/leases`, result => { leasesRead = result; });
});
lastState = state.evidence;
lastReadError = state.rejection ? undefined : state.readError;
if (lastReadError !== undefined) throw lastReadError;
return state;
},
// Native runs can retain the legacy "preparing" stage. The observer's
// runtime-ready RPC proves the actual pinned daemon before installation.
accept: state => !state.rejection && state.owned && state.run?.status === "running" && Boolean(state.lease)
&& Date.now() < admissionDeadlineAt,
reject: state => state.rejection,
timeoutDetail: () => JSON.stringify(lastState),
});
const lease = ready.lease!;
lastState = { ...lastState, phase: "observer_setup" };
fixture = await bind({ api: input.api, daytona: input.daytona, sdkVersion: "0.203.0", nodeSha256: input.nodeSha256, runnerdSha256: input.runnerdSha256,
authority: { companyId: input.companyId, environmentId: input.environmentId, runId: request.runId, leaseId: lease.id, sandboxId: lease.providerLeaseId, image: input.image },
targets: [...request.targets], crossRoot: request.crossRoot, actionFile: setup.path, deadlineAt: input.deadlineAt,
});
} catch (error) {
if (lastReadError !== undefined && lastState.phase === "lease_admission") {
if (classifyFailure(lastReadError) === "transient_infrastructure") {
const timeout = new ObservedStateTimeout("owned native qualification lease", "transient_infrastructure", JSON.stringify(lastState));
// The safe read classification is already in the bounded state.
// Do not expose the rejected transport value through a cause chain.
error = timeout;
} else error = lastReadError;
}
try {
await input.evidence(`remote-native-bootstrap-startup-${request.runId}.json`, {
...lastState, fixtureDiagnostics: remoteNativeFixtureDiagnostics(error), deadlineReached: Date.now() >= input.deadlineAt,
admissionDeadlineReached: Date.now() >= admissionDeadlineAt,
});
} catch (evidenceError) {
throw new AggregateError([error, evidenceError], "Remote bootstrap startup failed and state evidence could not be saved");
}
throw error;
}
try {
const action = await request.actionPrompt(fixture);
if (!action.trim() || Buffer.byteLength(action) > 16 * 1024) throw new Error("Remote bootstrap action is empty or too large");
await input.evidence(`remote-native-bootstrap-${request.runId}.json`, { binding: fixture.binding, setupPath: setup.path, baseline: fixture.baseline });
await fixture.publishAction(setup.path, action);
return fixture;
} catch (error) {
try { await fixture.close(); } catch { throw new AggregateError([error], "Remote bootstrap failed and observer cleanup is unproven"); }
throw error;
}
},
};
}