Skip to content

chore(fork): restore fork and Harness overlay on latest upstream - #87

Merged
mindfn merged 19 commits into
develop_basefrom
merge/main-into-develop-base-20260806
Aug 7, 2026
Merged

chore(fork): restore fork and Harness overlay on latest upstream#87
mindfn merged 19 commits into
develop_basefrom
merge/main-into-develop-base-20260806

Conversation

@mindfn

@mindfn mindfn commented Aug 6, 2026

Copy link
Copy Markdown
Owner

What

Rebuild develop_base as latest upstream/main@06263d99922da6928a21621f617a61b902fb1b5b plus the fork-specific product overlay and F257 Harness/Ledger work.

Exact candidate: 849694b89b9ab03bcc05380bd3392aa59702f06d.

The remote develop_base baseline was reset to that exact upstream SHA before this PR. Its previous tip is recoverable at backup/develop-base-pre-main-reset-20260806@27ea2aec5. Therefore this PR now shows only the overlay: 857 files, +55,695 / -22,340, not the earlier 366k-line upstream synchronization noise.

Overlay contents

  • Fork branding, local workflow/governance overlays, and reserved runtime boundaries.
  • F257 Harness contracts, lifecycle wiring, replay/evaluation surfaces, Ledger storage/projections, and MCP registration.
  • Fork production code plus its repository tests and evidence/docs.
  • Upstream F280 wait lifecycle, F286 canonical MCP governance, and latest LAN cancel provenance fixes are present through the base/history.
  • No runtime config files or persistent data were modified by this rebuild.

PR #83 (fix(F167): mode-aware hold quota with atomic reservation) is not included in this PR and remains a separate open change.

Validation

  • Focused latest-upstream Web tests: 22/22 PASS.
  • Focused shared/MCP/API builds PASS.
  • Focused MCP tests: 65/65 PASS.
  • Full MCP suite: 684/684 PASS.
  • pnpm check PASS.
  • pnpm gate --no-rebase: PASS on exact SHA 849694b89:
    • build PASS
    • tsc including tests PASS
    • Public repository suite: 20,271 tests, 0 failures
    • Web lint PASS
    • repository check PASS
    • total 980s

Risk

  • Behavior: high — broad fork overlay restored over a new upstream baseline.
  • Data: medium — lifecycle/store code moves together; this PR performs no destructive migration or runtime data operation.
  • Security/contract: medium-high — shared, API, MCP, approval governance, and Harness contracts move together.
  • Irreversibility: low — old develop_base tip has a remote backup branch; runtime activation is separate.

Review

Independent exact-SHA review is pending. This is a fork-internal PR; online Codex review is intentionally excluded per operator direction.

Runtime activation note

Upstream L0 compilation now fails closed when a cat lacks an explicit relationshipKey. Runtime config is deliberately outside this PR. Before rebuilding the live develop_base runtime, verify the existing runtime mapping for cat-eqdvbcxw; do not infer breed/model identity in code.

mindfn and others added 18 commits August 3, 2026 11:28
R6 split: A — shared types and segment lifecycle contracts used by runtime base and wiring.
… and verdict publisher

R6 split: B — runtime base. Includes internal services, message stores, routing, prompt-hooks, guard rejection event log, harness eval, local artifact publisher, telemetry, ball-custody, and existing route adaptations. Provenance is optional in MessageStore at this layer so upstream callers still compile. Old Git publisher removed in this commit alongside type sunset.
…d provenance enforcement

R6 split: C — runtime wiring. Adds segment-lifeline routes, prompt-injection override routes, makes MessageStore.provenance required, wires new routes in index.ts, and lands MCP server tools/tests.
R6 split: D — Console/web UI for segment lifeline, replay, eval window provenance, and actionable stage.
R6 split: E — prompt-hook asset updates for variable presentation and governance metadata.
…domains

R6 split: F — F257 feature documentation, eval domain registry, objectives, and bug reports.
R6 split: G — L0 compilation script updates, hook variable population script, and gitignore entries.
Why: the rebuilt develop_base must retain Cat Cafe runtime identity authority and the fork-specific planning/review safeguards without carrying the old mixed commit history.

[砚砚/gpt-5.6-sol🐾]
Why: develop_base is the live runtime baseline, so only the five explicitly governed shared-state files may be committed directly; code continues through feature PRs.

Also restores the fork ROADMAP identity/status overlay without reviving obsolete TeamAct or execution-artifact history.

[砚砚/gpt-5.6-sol🐾]
Why: three consecutive zero-signal 72h windows showed the 3-day schedule generated noise; weekly remains within the 168h SLA and is independently reversible.

[砚砚/gpt-5.6-sol🐾]
* docs(f257): plan annotation-driven objective evaluation

Why: tracing, attribution, and metric evaluation currently share a time-window judgment model that produces misleading per-segment rates. This plan establishes exact episode joins, unified annotations, count thresholds, async semantic analysis, and append-only results before implementation.

* feat(f257): bind harness markers to exact trace episodes

Why: Harness signals must annotate the authenticated invocation's immutable terminal episode instead of writing uncorrelated evaluation facts directly.

* feat(f257): evaluate objective metrics from trace annotations

Why: Objectives need explicit models and count/rate/semantic/replay metrics, while threshold counters must trigger from distinct episodes without inventing denominators.

* feat(f257): auto-evaluate indexed trace annotations

Why: tracing must remain a neutral fact ledger while objective-owned rules validate coordinates, freeze immutable snapshots, and append count or semantic evaluation results asynchronously.

* feat(f257): separate candidates counters and rates

Why: explicit incidents should trigger count evaluation without invented denominators, while uncertain MCP markers remain queued for semantic classification and rate metrics freeze only positive or counterexample samples.

* feat(f257): expose objective metric evaluation read model

Why: the Console must display each segment's registered Objective, Evaluation Model, concrete metrics, collection progress, and immutable result window instead of deriving a generic violation rate from legacy segment judgments.

* feat(f257): show objective metrics and trace replay theater

Why: replace the synthetic blocking lifeline UI with the actual Objective metric read model and complete TraceEpisode replay scenes.

* feat(f257): make template segments directly editable

Why: local owner-authored source overlays are separate from runtime safety-tier controls, and the editor should expose only useful variable metadata plus the exact editable source.

* feat(f257): evaluate unclassified trace episodes asynchronously

Why: tracing must remain a complete invocation fact stream while Objective-owned code/LLM/replay rules classify and evaluate evidence off the response path. Counterexample metrics now trigger from distinct incident counts without fabricated denominators, and semantic writeback is bound to immutable owner-scoped jobs.

* test(f257): isolate task outcome artifact fixtures

Why: fixed temp directory names survived failed runs and admitted unrelated metric YAML into the domain registry reader, making the public suite order-dependent. Keep each fixture under the test-owned root and clean it on every exit path.

* fix(f257): reject legacy signal payloads before storage

Why: report_harness_signal now marks the authenticated invocation trace. Validate that contract before checking Redis so sunset direct-observation requests return 400 and cannot mutate legacy deviation or guard-ledger projections.

* docs(f257): restore active roadmap truth

Why: the F257 feature spec remains in progress, so the repository feature-truth gate requires a matching active roadmap entry. Keep the entry product-scoped and free of fork-only operational state.

* test(f257): add objective evaluation showcase

Why: the existing showcase exercised the retired SegmentJudgment view, so browser acceptance could not verify the new counter/rate/semantic metric model or TraceEpisode theater.

* fix(f257): restore lifecycle-first evaluation navigation

Why: tracing replay and objective metrics belong to a selected version epoch; exposing them as global sibling tabs erased the version lifecycle coordinate and allowed shared-objective data to leak across segments.

* fix(f257): simplify lifecycle stage details

Why: lifecycle navigation should reveal operator-facing version, trace, and metric details without internal implementation noise or ambiguous selection state.

* fix(f257): align lifecycle selection styling

Why: lifecycle stage selection should reuse the existing accent pill language instead of introducing an unrelated dark outline.

* fix(f257): keep evaluation sidecar fail-open

Why: Objective evaluation observes the main runtime and must not abort API startup when its catalog is unavailable. Placeholder segments also cannot produce verdicts before their runtime data is wired.
* fix(f257): focus replay on source context

Why: operator replay should prioritize the originating thread/message and surrounding conversation, while template-render and window-correlated guard internals remain in durable audit storage without dominating the primary UI.

[砚砚/gpt-5.6-sol🐾]

* docs(f257): archive replay context review

Why: preserve the operator requirement, exact-SHA cross-cat verdict, repository gate evidence, and known UI-evidence limitations alongside the implementation before PR publication.

[砚砚/gpt-5.6-sol🐾]
* docs(F257): record replay refinement merge

Why: preserve F257 Feature Truth after fork-internal PR #85 merged under the LI-004 feature-branch guard.

[砚砚/gpt-5.6-sol🐾]

* docs(F257): archive PR 85 truth-sync review request

Why: make the fork-internal cross-individual review scope and evidence reproducible before the Feature Truth sync PR.

[砚砚/gpt-5.6-sol🐾]
Why: preserve fork-only F257 runtime truth and inbound brand boundaries while incorporating the upstream main line through an ordinary two-parent merge.\n\n[砚砚/gpt-5.6-sol🐾]
Why: preserve the fork's ordinary-merge history while reconciling latest main cursor, custody, and plugin inventory changes with the F257 harness ledger contracts.

[砚砚/gpt-5.6-sol🐾]
@mindfn mindfn changed the title merge: sync latest upstream main into develop_base chore(fork): restore fork and Harness overlay on latest upstream Aug 6, 2026
Why: preserve LI-005 routing safety and recover upstream documentation while rebuilding develop_base from the latest upstream main.

[砚砚/gpt-5.6-sol🐾]
@mindfn
mindfn marked this pull request as ready for review August 7, 2026 06:20
@mindfn
mindfn merged commit 8447852 into develop_base Aug 7, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant