mirror of
https://github.com/paperclipai/paperclip.git
synced 2026-10-08 11:13:44 +02:00
## Thinking Path > - Paperclip is the open source app people use to manage AI agents for work. > - Users create tasks with a title and a description. > - A required title adds work when the prompt already explains the request. > - An agent can name the task once it reads that request. > - This pull request accepts prompt-only tasks and starts them with a short prompt slice. > - A scoped title tool lets the assigned agent replace that slice early without changing execution state. > - A live browser eval checks the real agent call, saved title, audit entry, and preservation of user titles. ## Linked Issues or Issue Description **Subsystem affected** Cross-cutting: task creation, shared contracts, database, server, runner tools, and board UI. **Problem or motivation** Users must currently write a title before they can submit a detailed task prompt. The agent has enough context to write a useful title itself. **Proposed solution** Make the title optional when a description is present. Save the first 120 characters of the normalized prompt as a provisional title. Ask the assigned agent to call `set_task_title` early. Use an atomic provisional-title guard to preserve titles supplied or edited by users. Keep explicit titles supported. Related: #14543 and #14556 concern empty-title submission. This change intentionally enables that submission when a prompt is present, instead of requiring a title. ## What Changed - Add the `titleNeedsGeneration` field with an idempotent migration. Keep existing titles unchanged. - Add `PUT /api/issues/:id/title` and the native and legacy `set_task_title` tool. Enforce company access, active-run ownership, shared, bounded retry receipts across native/HTTP calls, and transactional audit logging. Refresh external-object links after commit, with the same feature gate and plugin detectors as ordinary title edits. - Add early naming guidance in Standard, Ask, and Plan task context. Preserve the description, status, and assignment. - Allow prompt-only root and child task creation, plus draft restoration in the New Task dialog. Keep user titles supported. - Add an opt-in Product E2E suite for prompt-only Standard and Ask tasks, plus an explicit-title control. It checks actual provider calls within the first five tools, persisted state, audit attribution, and the reloaded UI. - Preserve a closed vocabulary of API key maintenance phrases in declared prose while rejecting opaque credential suffixes. Add one bounded naming retry after wording is rejected, without treating the rejected call as a saved title. - Repair the native cleanup receipt check exposed during full verification: accept matching input digests, retain legacy input checks, and reject conflicting receipts. ## Verification - Live Product E2E on `f43478473800e3a46b85c5ee79677efdb15108e7`: **3/3 passed** with native Codex `gpt-5.4-mini`, first attempts only, automatic retries disabled. Standard and Ask each saved “Rotate expired API key” on their first tool call, with matching persisted state and a single same-run audit entry. The explicit-title control retained its user title with zero title writes. All three verified the reloaded browser UI. - Campaign: `local-2026-09-30T21-30-11-021Z`. Earlier failed campaigns are retained separately; they exposed credential-prose handling and prompted the naming recovery fix. No failed result was regraded or deleted. - Reproduce with `pnpm test:e2e:runner -- --id task-titles.runner-codex-mini.local.prompt-title-standard --id task-titles.runner-codex-mini.local.prompt-title-ask --id task-titles.runner-codex-mini.local.preserve-explicit-title --max-automatic-retries 0` and an authorized provider key. - Full `pnpm -r typecheck` and `pnpm build` passed on the latest commit. The runner build used the configured external eval source tree. - Product E2E unit suite: **61 files, 818 tests passed**; E2E typecheck and UI token gates passed. - Title API/native regressions cover prompt-only and explicit child creation, user edits, ownership/company isolation, external reference refresh, cross-surface retry replay, and the 64-key limit without receipt eviction. All passed. Prompt-context coverage: **44 tests passed**. - Rust credential regressions: **35 tests passed**, including benign maintenance qualifiers and opaque credential rejection in every declared prose field. Catalog/report reconciliation: **28 tests passed**. Native recovery: **560 tests passed**. - Broad local `pnpm test:run`: **14,555 tests passed** in the general server group; two suites failed to initialize embedded PostgreSQL and the existing 40,000-file Git streaming stress test exceeded its 300-second macOS timeout. All three suites then passed in isolation (**5 tests passed**) without code or timeout changes. The original full local command exited nonzero and is not being represented as a clean full run. - Latest-head GitHub checks are green: **53 passed, 4 skipped, zero failed or pending**, including all test shards and the canary packaging dry run. Greptile reviewed the same commit at **5/5**, with zero unresolved review threads. ## Risks - The additive database field must reach the server and UI together. The migration uses `IF NOT EXISTS` and defaults existing tasks to a final title. - Title generation depends on the assigned agent running. Tasks without a run keep their provisional title. - Live qualification covers the native Codex path in Standard and Ask modes. API/legacy and Plan behavior have deterministic coverage. - The credential-prose exception validates the entire suffix against a closed maintenance vocabulary. Unknown suffixes, assignments, quoted values, credential prefixes, and diagnostics retain strict checks. ## Model Used OpenAI Codex, based on GPT-6, with reasoning, tool use, and code execution. The exact deployment ID and context window are not exposed in this session. The live eval uses the native Codex `gpt-5.4-mini` profile. ## Checklist - [x] I have included a thinking path that traces from project context to this change - [x] I have specified the model used (with version and capability details) - [x] I have checked ROADMAP.md and confirmed this PR does not duplicate planned core work - [x] I have searched GitHub for duplicate or related PRs and linked them above - [x] I have either (a) linked existing issues with `Fixes: #` / `Closes #` / `Refs #` OR (b) described the issue in-PR following the relevant issue template - [x] I have not referenced internal/instance-local Paperclip issues or links (only public GitHub `#NNN` / `github.com/paperclipai/paperclip` URLs) - [x] My branch name describes the change (e.g. `docs/...`, `fix/...`) and contains no internal Paperclip ticket id or instance-derived details - [x] I have run tests locally and they pass - [x] I have added or updated tests where applicable - [x] I have updated relevant documentation to reflect my changes - [x] I have considered and documented any risks above - [x] All Paperclip CI gates are green - [x] Greptile is 5/5 with no open P2s, recommendations, or follow-ups - [x] I will address all Greptile and reviewer comments before requesting merge --------- Co-authored-by: Paperclip <noreply@paperclip.ing>
372 lines
26 KiB
JSON
372 lines
26 KiB
JSON
{
|
|
"schemaVersion": 1,
|
|
"documentTitle": "Paperclip agent operation groups",
|
|
"evalRepositoryUrl": "https://github.com/paperclipai/paperclip-evals/blob/master/",
|
|
"prpEventFamilyDescriptions": {
|
|
"runner": "Runner connection, drain, and diagnostic lifecycle.",
|
|
"runtime": "Runner phase transitions.",
|
|
"sandbox": "Sandbox resource measurements.",
|
|
"workspace": "Workspace readiness.",
|
|
"harness": "Provider harness startup, readiness, exit, and diagnostics.",
|
|
"session": "Provider-neutral session open, resume, reconciliation, close, and failure.",
|
|
"turn": "Model turn submission through terminal turn disposition.",
|
|
"item": "Provider-neutral model/tool item lifecycle.",
|
|
"mcp_app": "MCP App discovery, initialization, tool, action, host-context, and teardown evidence.",
|
|
"runtime_request": "Runtime permission/input request lifecycle.",
|
|
"interaction": "Issue-thread interaction proposal, materialization, response, delivery, and rejection.",
|
|
"run": "Structured result negotiation and terminal run outcome.",
|
|
"attention": "Continuation and attention routing lifecycle.",
|
|
"work": "Recorded work-assessment evidence.",
|
|
"issue": "Issue-status decision proposal outcome and application evidence.",
|
|
"plan": "Complete provider-authored within-turn checklist snapshots, separate from durable Paperclip Plan documents.",
|
|
"tool": "Provider-neutral process, MCP, dynamic, and built-in execution activity.",
|
|
"research": "Provider-reported search, page-open, and in-page research activity.",
|
|
"delegation": "Child-agent delegation lifecycle and aggregate status.",
|
|
"model": "Requested/effective model routing and verification state.",
|
|
"context": "Context-window compaction markers without hidden summaries.",
|
|
"artifact": "Authorized artifact viewing and structured generated outputs.",
|
|
"review": "Provider review-mode state, separate from Paperclip authority.",
|
|
"hook": "Bounded provider hook lifecycle and blocking outcomes.",
|
|
"memory": "Authorized or unavailable memory citation references.",
|
|
"safety": "Provider safety review state attached to governed work.",
|
|
"terminal": "Content-free terminal input activity metadata.",
|
|
"wait": "Intentional provider waits, distinct from warm idle and human input.",
|
|
"provider": "Redacted provider notices and actionable warnings.",
|
|
"semantic_tool": "Canonical authorized Paperclip tool input and result evidence.",
|
|
"usage": "Provider/model-attributed usage and accounting boundaries."
|
|
},
|
|
"prpCommandFamilyDescriptions": {
|
|
"semantic_tool": "Returns an authorized, correlated Paperclip tool result to the provider through runnerd.",
|
|
"run": "Prepare or cancel a run.",
|
|
"session": "Open, snapshot, or close a normalized provider session.",
|
|
"turn": "Start, steer, interrupt, or stop a model turn.",
|
|
"request": "Resolve a pending runtime request.",
|
|
"interaction": "Acknowledge delivery of an interaction response.",
|
|
"runner": "Drain or shut down the runner process."
|
|
},
|
|
"optionalGrantGroups": [
|
|
{
|
|
"id": "discovery",
|
|
"description": "Company-visible task, agent, project, and goal discovery.",
|
|
"operationIds": ["search_tasks", "list_agents", "get_agent", "list_projects", "list_goals", "set_task_title"]
|
|
},
|
|
{
|
|
"id": "projects",
|
|
"description": "Project creation and authorized repository discovery through the live company/run authority.",
|
|
"operationIds": ["create_project", "list_project_repositories"]
|
|
},
|
|
{
|
|
"id": "delegation_dependencies",
|
|
"description": "Create or reassign delegated work and maintain dependency edges.",
|
|
"operationIds": ["create_task", "reassign_task", "set_dependencies", "hire_agent"]
|
|
},
|
|
{
|
|
"id": "governance",
|
|
"description": "Read, request, comment on, and decide approvals under governed-action checks.",
|
|
"operationIds": ["list_approvals", "get_approval", "get_approval_context", "request_approval", "decide_approval", "comment_on_approval"]
|
|
},
|
|
{
|
|
"id": "cases",
|
|
"description": "Read and update case summaries without reusing issue-document authority.",
|
|
"operationIds": ["list_cases", "upsert_case"]
|
|
},
|
|
{
|
|
"id": "workspace_runtime",
|
|
"description": "Inspect and control the active issue workspace runtime.",
|
|
"operationIds": ["get_workspace_runtime", "control_workspace_service"]
|
|
},
|
|
{
|
|
"id": "wake_scheduling",
|
|
"description": "Schedule a bounded continuation wake when the current task owns the future check.",
|
|
"operationIds": ["schedule_wake"]
|
|
},
|
|
{
|
|
"id": "routines",
|
|
"description": "Inspect or manage company routines.",
|
|
"operationIds": ["list_routines", "manage_routine"]
|
|
},
|
|
{
|
|
"id": "company_skills",
|
|
"description": "Create or inspect company skills and synchronize the current agent's skills.",
|
|
"operationIds": ["create_skill", "list_company_skills", "sync_company_skills"]
|
|
},
|
|
{
|
|
"id": "secrets",
|
|
"description": "Inspect secret metadata or use a brokered secret value without exposing plaintext evidence.",
|
|
"operationIds": ["list_secret_metadata", "read_secret_value"]
|
|
},
|
|
{
|
|
"id": "portability_admin",
|
|
"description": "Export portable company state; broad administration remains deferred until split into governed operations.",
|
|
"operationIds": ["export_company", "administer_company"]
|
|
},
|
|
{
|
|
"id": "test_escape_hatch",
|
|
"description": "Controlled skill-test transport only; never product coverage.",
|
|
"operationIds": ["generic_api_request"]
|
|
},
|
|
{
|
|
"id": "api_fallback",
|
|
"description": "Discover and invoke HTTP API operations unsupported by available dedicated tools. Production authority and route authorization remain required.",
|
|
"operationIds": [
|
|
"search_api",
|
|
"call_api"
|
|
]
|
|
}
|
|
],
|
|
"behaviorGroups": [
|
|
{
|
|
"id": "hb",
|
|
"legacyGroup": 1,
|
|
"name": "Heartbeat",
|
|
"owner": "control plane + always tools",
|
|
"operationIds": ["get_task_context", "set_task_title"],
|
|
"controlPlaneOperationIds": ["select_work", "enforce_budget"],
|
|
"realSurface": "agent identity, inbox-lite, heartbeat context, budget and active-issue services",
|
|
"mockStateDomains": ["company", "actor", "wake", "task", "budget", "run"],
|
|
"prpEvidence": "runner/session/run context plus bounded task-context tool results",
|
|
"gap": "Production semantic binding is unbound; identity and work selection remain injected/control-plane-owned."
|
|
},
|
|
{
|
|
"id": "co",
|
|
"legacyGroup": 2,
|
|
"name": "Checkout",
|
|
"owner": "control plane",
|
|
"operationIds": [],
|
|
"controlPlaneOperationIds": ["checkout_task"],
|
|
"realSurface": "POST /api/issues/:id/checkout and execution-lock services",
|
|
"mockStateDomains": ["task", "actor", "run", "idempotency", "fault"],
|
|
"prpEvidence": "run preparation and issue-status decision evidence with checkout receipt",
|
|
"gap": "Intentionally no model tool; the production checkout receipt still needs the additive semantic-receipt envelope."
|
|
},
|
|
{
|
|
"id": "st",
|
|
"legacyGroup": 3,
|
|
"name": "Status",
|
|
"owner": "always tools + control-plane arbitration",
|
|
"operationIds": ["answer_status_question", "finish_task", "block_task", "request_review"],
|
|
"controlPlaneOperationIds": ["reconcile_run", "append_audit_record"],
|
|
"realSurface": "issue PATCH, review/liveness policy, and native finalization arbitration",
|
|
"mockStateDomains": ["task", "comments", "interactions", "blockers", "audit", "run"],
|
|
"prpEvidence": "semantic operation receipt, work assessment, issue-status decision, and terminal causality",
|
|
"gap": "Production semantic binding and additive typed operation/conflict receipts remain unimplemented."
|
|
},
|
|
{
|
|
"id": "cm",
|
|
"legacyGroup": 4,
|
|
"name": "Comments",
|
|
"owner": "always tools",
|
|
"operationIds": ["get_task_history", "report_progress"],
|
|
"controlPlaneOperationIds": ["append_audit_record"],
|
|
"realSurface": "issue comment list/get/create routes",
|
|
"mockStateDomains": ["task", "comments", "actor", "idempotency", "audit"],
|
|
"prpEvidence": "bounded read result or idempotent comment-write receipt plus audit reference",
|
|
"gap": "Active-task binding is unbound; cross-task comment mutation is deliberately outside V1."
|
|
},
|
|
{
|
|
"id": "se",
|
|
"legacyGroup": 5,
|
|
"name": "Search",
|
|
"owner": "optional discovery tools",
|
|
"operationIds": ["search_tasks", "list_agents", "get_agent", "list_projects", "list_goals", "list_project_repositories"],
|
|
"controlPlaneOperationIds": [],
|
|
"realSurface": "company issue search and agent/project/goal list/get routes",
|
|
"mockStateDomains": ["company", "task", "actor", "project", "goal"],
|
|
"prpEvidence": "bounded redacted read projections through tool-result item events",
|
|
"gap": "Goal operations remain scenario-only. Project and repository discovery contracts are delivered by the live server authority; they are not implemented by the mock command dispatcher."
|
|
},
|
|
{
|
|
"id": "su",
|
|
"legacyGroup": 6,
|
|
"name": "Subtasks",
|
|
"owner": "optional delegation tools",
|
|
"operationIds": ["create_task", "reassign_task", "hire_agent"],
|
|
"controlPlaneOperationIds": ["route_wake"],
|
|
"realSurface": "company issue create, child issue, guarded reassignment, and wake services",
|
|
"mockStateDomains": ["company", "task", "actor", "blockers", "wake", "audit"],
|
|
"prpEvidence": "company/task state diff, audit reference, and continuation wake evidence",
|
|
"gap": "create_task is production-bound to ordinary active-issue child creation with assignment, dependency-ready wake, company checks, child limits, and durable source-scoped idempotency."
|
|
},
|
|
{
|
|
"id": "bl",
|
|
"legacyGroup": 7,
|
|
"name": "Blockers",
|
|
"owner": "always/optional tools + control plane",
|
|
"operationIds": ["block_task", "set_dependencies"],
|
|
"controlPlaneOperationIds": ["schedule_blocker_wake", "route_wake"],
|
|
"realSurface": "issue relations, blocker projection, liveness validation, and blocker wake services",
|
|
"mockStateDomains": ["task", "blockers", "wake", "actor", "audit", "fault"],
|
|
"prpEvidence": "dependency diff, block receipt, attention routing, and issue-status decision",
|
|
"gap": "set_dependencies is production-bound for the active issue; block_task remains unbound, and cancelled-blocker receipts still need typed additive evidence."
|
|
},
|
|
{
|
|
"id": "dp",
|
|
"legacyGroup": 8,
|
|
"name": "Documents and plans",
|
|
"owner": "always tools; restore optional; destructive lifecycle control-plane-only",
|
|
"operationIds": ["list_documents", "read_document", "list_document_revisions", "write_document"],
|
|
"controlPlaneOperationIds": ["append_audit_record"],
|
|
"realSurface": "issue document list/read/upsert/revision/restore/lock/unlock/delete routes",
|
|
"mockStateDomains": ["task", "documents", "interactions", "idempotency", "audit", "fault"],
|
|
"prpEvidence": "bounded reads and revision-safe write/conflict/denial receipts with revision lineage",
|
|
"gap": "restore_document_revision is an approved optional-tool gap; lock/unlock/delete are intentionally control-plane-only."
|
|
},
|
|
{
|
|
"id": "ix",
|
|
"legacyGroup": 9,
|
|
"name": "Interactions",
|
|
"owner": "always tools + addressed resolver",
|
|
"operationIds": ["request_human_input"],
|
|
"controlPlaneOperationIds": ["route_wake"],
|
|
"realSurface": "issue-thread interaction create/respond/accept/reject/withdraw services",
|
|
"mockStateDomains": ["task", "documents", "interactions", "wake", "actor", "idempotency"],
|
|
"prpEvidence": "interaction proposal/materialization/response/delivery and attention events",
|
|
"gap": "Production binding and semantic resolution receipts are unbound."
|
|
},
|
|
{
|
|
"id": "ap",
|
|
"legacyGroup": 10,
|
|
"name": "Approvals",
|
|
"owner": "optional governance tools + governed approver",
|
|
"operationIds": ["list_approvals", "get_approval", "get_approval_context", "request_approval", "decide_approval", "comment_on_approval"],
|
|
"controlPlaneOperationIds": ["route_wake", "append_audit_record"],
|
|
"realSurface": "company approval, decision, issue-link, comment, and governed-action services",
|
|
"mockStateDomains": ["company", "task", "approvals", "actor", "wake", "audit", "idempotency"],
|
|
"prpEvidence": "governed semantic receipts, audit references, and attention/continuation linkage",
|
|
"gap": "Production binding and additive governed-action receipts are unbound; board-only authority stays outside grants."
|
|
},
|
|
{
|
|
"id": "ar",
|
|
"legacyGroup": 11,
|
|
"name": "Artifacts",
|
|
"owner": "always tools + artifact/work-product services",
|
|
"operationIds": ["register_deliverable"],
|
|
"controlPlaneOperationIds": ["append_audit_record"],
|
|
"realSurface": "attachment upload and issue work-product routes",
|
|
"mockStateDomains": ["task", "artifacts", "workProducts", "workspace", "audit", "idempotency"],
|
|
"prpEvidence": "artifact/work-product reference and durable inspectability receipt; never binary bytes",
|
|
"gap": "Production upload/register composite and additive durable-reference receipt are unbound."
|
|
},
|
|
{
|
|
"id": "er",
|
|
"legacyGroup": 12,
|
|
"name": "Errors and critical rules",
|
|
"owner": "runner/control plane + optional workspace/wake tools",
|
|
"operationIds": ["get_workspace_runtime", "control_workspace_service", "schedule_wake", "inspect_operation_result"],
|
|
"controlPlaneOperationIds": ["release_task", "enforce_budget", "persist_run", "replay_run", "reconcile_run"],
|
|
"realSurface": "workspace runtime, monitor/recovery, budget, run persistence/replay, release, and terminal services",
|
|
"mockStateDomains": ["workspace", "budget", "run", "wake", "audit", "idempotency", "fault"],
|
|
"prpEvidence": "runtime/workspace/attention/run lifecycle, typed denials, replay facts, and terminal causality",
|
|
"gap": "Budget stop reasons and semantic denial/conflict receipts require additive v1 envelopes; inspect_operation_result remains scenario-only."
|
|
},
|
|
{
|
|
"id": "rf",
|
|
"legacyGroup": 13,
|
|
"name": "Reference files",
|
|
"owner": "always instruction tools + optional domain tools + test-only escape hatch",
|
|
"operationIds": ["list_cases", "upsert_case", "list_routines", "manage_routine", "create_skill", "list_company_skills", "sync_company_skills", "list_secret_metadata", "read_secret_value", "export_company", "administer_company", "generic_api_request", "search_api", "call_api", "create_project", "read_agent_instructions", "update_agent_instructions", "get_agent_instruction_history", "restore_agent_instructions"],
|
|
"controlPlaneOperationIds": ["append_audit_record"],
|
|
"realSurface": "agent instruction revisions, project, case, routine, company-skill, secret, portability, and administration services",
|
|
"mockStateDomains": ["company", "cases", "routines", "skills", "secrets", "audit", "fault"],
|
|
"prpEvidence": "bounded domain projections, redacted broker receipts, company diffs, and audit references",
|
|
"gap": "Instruction read/update/history/restore require the live canonical revision service, current responsible-user permissions, and pinned base revisions for writes. Their service, authority, and product E2E tests are separate from the legacy mock scenarios. Project and skill creation and API tools require the live server authority; generic_api_request is test-only and the remaining domain operations are scenario-only. Broad administer_company is deferred and cannot claim product coverage. Production API and paired dedicated-tool regressions are recorded separately in paperclip-evals/evals/runner-api-tools; the legacy scenario count is not evidence of that coverage."
|
|
},
|
|
{
|
|
"id": "mh",
|
|
"legacyGroup": 14,
|
|
"name": "Multi-hop",
|
|
"owner": "composed semantic operations + control-plane continuation",
|
|
"operationIds": ["create_task", "set_dependencies", "request_human_input", "request_approval", "register_deliverable"],
|
|
"controlPlaneOperationIds": ["route_wake", "reconcile_run"],
|
|
"realSurface": "delegation, dependency, interaction, approval, artifact, and terminal orchestration services",
|
|
"mockStateDomains": ["task", "blockers", "interactions", "approvals", "artifacts", "wake", "run", "audit"],
|
|
"prpEvidence": "correlated operation receipts, state diffs, attention hops, work assessment, status decision, and terminal outcome",
|
|
"gap": "No generic transaction tool is allowed; shared mock/real conformance must prove each composed effect."
|
|
},
|
|
{
|
|
"id": "rs",
|
|
"legacyGroup": 15,
|
|
"name": "Restraint and no-call",
|
|
"owner": "policy/exposure layer",
|
|
"operationIds": ["answer_status_question", "read_secret_value", "generic_api_request"],
|
|
"controlPlaneOperationIds": ["enforce_budget"],
|
|
"realSurface": "task-mode, secret-broker, test-scope, pause, and budget policy checks",
|
|
"mockStateDomains": ["actor", "task", "budget", "secrets", "audit", "fault"],
|
|
"prpEvidence": "absence of forbidden effects plus typed policy denial/redaction receipts when a call is attempted",
|
|
"gap": "Typed redaction/authorization receipts need additive v1 evidence; generic_api_request is never a product fallback."
|
|
},
|
|
{
|
|
"id": "wk",
|
|
"legacyGroup": 16,
|
|
"name": "Wake situations",
|
|
"owner": "control plane + always context/history tools",
|
|
"operationIds": ["get_task_context", "get_task_history", "schedule_wake"],
|
|
"controlPlaneOperationIds": ["select_work", "route_wake"],
|
|
"realSurface": "wakeup requests, heartbeat context, comment/interaction/approval/blocker wake routing, and scheduled wake services",
|
|
"mockStateDomains": ["wake", "task", "comments", "interactions", "approvals", "blockers", "run"],
|
|
"prpEvidence": "attention request routing/resolution plus resumed session/run causality",
|
|
"gap": "Production scheduling binding is unbound; control-plane routing remains non-callable."
|
|
}
|
|
],
|
|
"rules": {
|
|
"companyAuthorization": [
|
|
"Every entity read and write is resolved inside the authenticated actor's company; cross-company identifiers fail without disclosing protected facts.",
|
|
"Board actors use active membership and role permissions. Agent writes require a company-scoped run JWT and X-Paperclip-Run-Id; active-task tools cannot accept a caller-selected company or arbitrary task.",
|
|
"Optional tools are omitted unless every required claim, role, and task-mode condition is satisfied. A grant never bypasses approval, budget, pause, execution-lock, interaction-owner, or other governed-action checks."
|
|
],
|
|
"taskModes": [
|
|
"standard permits the full eligible catalog subject to claims and role checks.",
|
|
"ask is answer-only: investigation reads are eligible, while implementation, document mutation, delegation, terminal, and governed writes are absent.",
|
|
"planning permits plan/document/progress/input operations but not delegation or terminal completion; accepting a plan transitions the issue into a fresh standard execution continuation for the exact accepted revision.",
|
|
"skill_test exposes only scenario-admitted tools. generic_api_request additionally requires the explicit test grant and allowlist and never supplies product coverage."
|
|
],
|
|
"redaction": [
|
|
"Tool descriptors declare observable redaction. Secret plaintext may reach only the model-use capsule authorized for a bound secret; it never reaches PRP, logs, artifacts, errors, audit details, or serialized tool results.",
|
|
"Authorization denials reveal stable codes, missing public claims, and remediation only; they do not include protected entity state, credentials, headers, cookies, or another company's policy.",
|
|
"Artifacts carry durable references, hashes, content type, and size on the wire; binary payloads remain in storage transports."
|
|
],
|
|
"idempotency": [
|
|
"Every mutation whose descriptor says required must carry a stable idempotency key. Retrying the same key and equivalent input returns the prior receipt without duplicating side effects.",
|
|
"Reusing a key with different input is an idempotency conflict. Optimistic document writes additionally bind baseRevisionId and return the current revision on conflict.",
|
|
"PRP source event IDs are idempotent, source sequence gaps are evidence, and replay is side-effect free. Replaying a trace cannot repeat control-plane writes."
|
|
],
|
|
"sideEffects": [
|
|
"read operations return bounded projections and do not mutate control-plane state.",
|
|
"task_write and company_write operations emit normalized state diffs and immutable audit references after production authorization succeeds.",
|
|
"governance, workspace_control, secret_read, and admin operations retain their production-specific gates; catalog exposure is necessary but never sufficient authority.",
|
|
"Terminal semantic operations propose intent. The control plane owns checkout, release, budget enforcement, persisted run evidence, replay, wake routing, audit append, and final status arbitration."
|
|
]
|
|
},
|
|
"documentLifecycle": {
|
|
"scope": [
|
|
"Issue documents are active-task working records identified by (activeIssueId, key). Always-present document tools do not accept a company id or arbitrary issue id.",
|
|
"Cross-task document reads require a future explicit optional operation and grant; case documents and future company knowledge are separate resource types.",
|
|
"Locked-document fallback may create a deterministic new key only when lockedDocumentStrategy is create_new_document; the source stays locked and the returned receipt names the redirected key."
|
|
],
|
|
"operations": [
|
|
{"action": "create", "semanticOperation": "write_document", "placement": "always_agent_tool", "contract": "baseRevisionId is null; create revision 1 and return document/revision/audit identity.", "status": "supported contract; production binding unbound"},
|
|
{"action": "read", "semanticOperation": "list_documents / read_document", "placement": "always_agent_tool", "contract": "List metadata or read the current active-task document by stable key.", "status": "supported contract; production binding unbound"},
|
|
{"action": "update", "semanticOperation": "write_document", "placement": "always_agent_tool", "contract": "Require the exact latest baseRevisionId and an idempotency key; append an immutable revision.", "status": "supported contract; production binding unbound"},
|
|
{"action": "revisions", "semanticOperation": "list_document_revisions", "placement": "always_agent_tool", "contract": "Return bounded immutable revision history newest first.", "status": "supported contract; production binding unbound"},
|
|
{"action": "restore", "semanticOperation": "restore_document_revision", "placement": "optional_agent_tool (documents:restore)", "contract": "Validate same-document lineage, reject locked documents, append a new revision, and return source/new revision identity; restoring current is a no-change duplicate.", "status": "approved catalog gap"},
|
|
{"action": "lock", "semanticOperation": "none", "placement": "control_plane_owned", "contract": "Board/administrative lifecycle action; agent writes receive document_locked or use explicit create-new fallback.", "status": "intentionally non-callable"},
|
|
{"action": "unlock", "semanticOperation": "none", "placement": "control_plane_owned", "contract": "Board/administrative lifecycle action; never inferred from a failed write.", "status": "intentionally non-callable"},
|
|
{"action": "delete", "semanticOperation": "none", "placement": "control_plane_owned", "contract": "Destructive audited action; pending targets become stale and the runner never recreates them automatically.", "status": "intentionally non-callable"}
|
|
]
|
|
},
|
|
"legacySemanticTargetDispositions": {
|
|
"injected_actor_context": {"kind": "control_plane", "targets": ["select_work"], "reason": "Identity is launch context, not a model tool."},
|
|
"runner_work_selection": {"kind": "control_plane", "targets": ["select_work"], "reason": "Inbox selection belongs to the control plane."},
|
|
"get_project": {"kind": "folded", "targets": ["list_projects"], "reason": "The bounded discovery operation owns project list/get projection in V1."},
|
|
"wait_for_workspace_service": {"kind": "folded", "targets": ["control_workspace_service"], "reason": "Wait is a bounded action of workspace service control."},
|
|
"get_goal": {"kind": "folded", "targets": ["list_goals"], "reason": "The bounded discovery operation owns goal list/get projection in V1."},
|
|
"semantic_task_disposition": {"kind": "composite", "targets": ["answer_status_question", "finish_task", "block_task", "request_review"], "reason": "Generic issue PATCH is replaced by intent-specific terminal/status operations."},
|
|
"atomic_checkout": {"kind": "control_plane", "targets": ["checkout_task"], "reason": "Checkout is an atomic control-plane transaction."},
|
|
"runtime_release": {"kind": "control_plane", "targets": ["release_task"], "reason": "Release is runner/control-plane cleanup."},
|
|
"restore_document_revision": {"kind": "known_gap", "targets": [], "reason": "Approved optional documents:restore operation is not yet in the canonical catalog."},
|
|
"link_approval": {"kind": "folded", "targets": ["request_approval", "get_task_context"], "reason": "Issue linkage is part of approval request/context composites."},
|
|
"unlink_approval": {"kind": "control_plane", "targets": ["append_audit_record"], "reason": "No standalone agent unlink tool is approved in V1."},
|
|
"test_only_api_escape_hatch": {"kind": "folded", "targets": ["generic_api_request"], "reason": "Compatibility alias for the test-only escape hatch."}
|
|
}
|
|
}
|