Skip to content

fix(ce-code-review): isolate testing reviewer mutation work - #1584

Merged
tmchow merged 5 commits into
EveryInc:mainfrom
khsaurabh:fix/1566-testing-reviewer-worktree-isolation
Aug 31, 2026
Merged

fix(ce-code-review): isolate testing reviewer mutation work#1584
tmchow merged 5 commits into
EveryInc:mainfrom
khsaurabh:fix/1566-testing-reviewer-worktree-isolation

Conversation

@khsaurabh

Copy link
Copy Markdown
Contributor

Dispatch the testing reviewer with isolation: "worktree" so mutation testing cannot write transient lines into the shared checkout that concurrent read-only reviewers observe.

Fixes #1566

Security Disclosure

No security-relevant changes.

Agent Disclosure

  • Model: Hermes Agent · grok-composer-2.5-fast

Dispatch testing with isolation: worktree so concurrent read-only
reviewers cannot observe transient mutations. Fixes EveryInc#1566.

@chatgpt-codex-connector chatgpt-codex-connector Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

💡 Codex Review

Here are some automated review suggestions for this pull request.

Reviewed commit: 3cf1350ad6

ℹ️ About Codex in GitHub

Codex has been enabled to automatically review pull requests in this repo. Reviews are triggered when you

  • Open a pull request for review
  • Mark a draft as ready
  • Comment "@codex review".

If Codex has suggestions, it will comment; otherwise it will react with 👍.

When you sign up for Codex through ChatGPT, Codex can also answer questions or update the PR, like "@codex address that feedback".

Comment thread skills/ce-code-review/references/dispatch-reviewers.md Outdated
khsaurabh and others added 2 commits August 30, 2026 16:07
Native worktrees start from committed HEAD, so local-aligned staged or
unstaged changes were invisible to the testing persona. Require a
faithful snapshot (scratch copy when needed) before isolation.
The exception block led with isolation: "worktree" and then carved out
local-aligned scope and non-isolating hosts. Lead with the invariant (a
tree-mutating persona works only on a faithful snapshot) and name each
mechanism under the condition that makes it valid.

Claude-Session: https://claude.ai/code/session_01EoF39fuXHnmWDVzkBADnzo
@tmchow

tmchow commented Aug 31, 2026

Copy link
Copy Markdown
Collaborator

One thing I fixed directly (pushed 25c178b): the exception block led with isolation: "worktree" and then had to walk it back mid-paragraph for local-aligned scope, where a native worktree of committed HEAD is stale and the scratch snapshot is doing all the work anyway. That read as mechanism-first with carve-outs. I restated it condition-first: the invariant is that a tree-mutating persona operates only on a faithful snapshot, and each mechanism (worktree isolation for committed HEAD, scratch copy for local-aligned or non-isolating hosts) sits under the condition that makes it valid.

Observed live: when the session itself runs inside a managed worktree
(Codex desktop, Cursor, Orca), isolation: "worktree" can cut the
subagent's copy from the primary checkout or default branch, not the
reviewed tree — even at committed HEAD. The persona now verifies its
copy's HEAD equals the reviewed commit before mutating and falls back
to a scratch copy on mismatch.

Claude-Session: https://claude.ai/code/session_01EoF39fuXHnmWDVzkBADnzo
@tmchow

tmchow commented Aug 31, 2026

Copy link
Copy Markdown
Collaborator

One more hardening pass, prompted by a scenario worth testing empirically: most harness apps (Codex desktop, Cursor, Orca) run the whole session inside a managed worktree, so the review itself often runs from one. I probed this live — spawned a worktree-isolated subagent from inside a linked worktree on this repo. It worked mechanically, but the harness cut the isolated copy from the primary checkout's default branch, not from the worktree's checked-out PR branch: its HEAD matched neither the session's HEAD nor the reviewed commit, with 53 files differing.

So "committed HEAD → the harness supplies a faithful snapshot" is false in exactly the most common setup, and a mutation-testing reviewer would have quietly reviewed the wrong tree. Pushed 420714e: the faithful-copy invariant is now verified, never assumed — the persona gets the reviewed commit SHA, checks its copy's HEAD equals it before mutating, and falls back to a scratch copy on any mismatch. Added the smallest test pin for the verify condition. All 85 contract tests pass.

@chatgpt-codex-connector

chatgpt-codex-connector Bot commented Aug 31, 2026

Copy link
Copy Markdown

Codex Review Summary

This comment shows the latest Codex review activity on this pull request.

Review Status Commit Review trigger
📝 Code Review Completed 2026-08-31T18:12:42.038547Z 438db5c New commits
ℹ️ About Codex in GitHub

Your team has set up Codex to review pull requests in this repo. Reviews are triggered when you

  • Open a pull request for review
  • Mark a draft as ready
  • Comment "@codex review" or "@codex security review".

Codex reacts with 👀 while any review is running, comments if it has suggestions, and reacts with 👍 once all reviews finish with no findings.

A harness-created isolated worktree is not automatically a faithful
snapshot of the session's tree; pass the intended commit SHA, verify
HEAD equals it, and fall back to a scratch copy on mismatch.

Claude-Session: https://claude.ai/code/session_01EoF39fuXHnmWDVzkBADnzo
@tmchow

tmchow commented Aug 31, 2026

Copy link
Copy Markdown
Collaborator

Also pushed 438db5c adding docs/solutions/skill-design/verify-harness-worktree-snapshot-fidelity.md so the snapshot-fidelity learning ships with the fix that produced it.

@chatgpt-codex-connector chatgpt-codex-connector Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

💡 Codex Review

Here are some automated review suggestions for this pull request.

Reviewed commit: 438db5c4d7

ℹ️ About Codex in GitHub

Your team has set up Codex to review pull requests in this repo. Reviews are triggered when you

  • Open a pull request for review
  • Mark a draft as ready
  • Comment "@codex review".

If Codex has suggestions, it will comment; otherwise it will react with 👍.

Codex can also answer questions or update the PR. Try commenting "@codex address that feedback".

- **Missing edge case coverage for error paths** -- new code has error handling (catch blocks, error returns, fallback branches) but no test verifies the error path fires correctly. The happy path is tested; the sad path is not.
- **Behavioral changes with no test additions** -- the diff modifies behavior (new logic branches, state mutations, changed API contracts, altered control flow, or error behavior) but adds or modifies zero test files. This is distinct from untested branches above, which checks coverage *within* code that has tests. This check flags when the diff contains behavioral changes with no corresponding test work at all. Non-behavioral changes (formatting, comments, type-only annotations, or dependency/config metadata that does not alter runtime behavior) are excluded.

If you use mutation testing (edit a production file, run the suite, revert), do it only in an isolated worktree or a scratch copy that is a faithful snapshot of the reviewed tree — verify before mutating: your copy's HEAD must equal the reviewed commit (a harness-created worktree may be cut from the primary checkout or default branch instead), and `local-aligned` scope needs the staged/unstaged changes a committed-`HEAD` worktree lacks. On any mismatch, fall back to a scratch copy of the reviewed tree. Never mutate the shared checkout the rest of the reviewer batch is reading.

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P2 Badge Align the leaf prompt with the mutation exception

The assembled testing-reviewer prompt still includes subagent-template.md, whose lines 25 and 148 say the artifact is the only permitted write and explicitly prohibit editing project files. That directly contradicts this new instruction to edit production files in an isolated worktree or create and mutate a scratch copy, so a compliant leaf reviewer cannot perform the mutation test or its mismatch fallback. Move the exception into the owning subagent template, conditioned on a tree-mutating persona and a verified faithful copy, while retaining the unconditional ban on mutating the shared checkout. The added contract test misses this conflict because it only checks for phrases in the dispatch and persona files.

AGENTS.md reference: AGENTS.md:L135-L137

Useful? React with 👍 / 👎.

@tmchow
tmchow merged commit fe02f79 into EveryInc:main Aug 31, 2026
3 checks passed
@github-actions github-actions Bot mentioned this pull request Aug 31, 2026
ethras added a commit to ethras/compound-engineering-orca that referenced this pull request Sep 1, 2026
* fix(ce-commit-push-pr): root PR stacks on the parent PR the user named (EveryInc#1365)

* fix(ce-babysit-pr): decode gh output as UTF-8 on Windows (EveryInc#1368)

* fix(ce-prototype): cover decisions settled by seeing, not just driving (EveryInc#1369)

* perf(tests): cut suite wall time by splitting the largest test file (EveryInc#1370)

* fix(tests): stop the cross-model routes test reading the working tree (EveryInc#1371)

* fix(ce-doc-review): ask only where a real choice exists, batch the rest (EveryInc#1373)

* feat(ce-prototype): add a seeing-mode craft floor and durable storage (EveryInc#1374)

* fix(ce-pov): stop the panel guessing the cross-model host argument (EveryInc#1375)

* chore(cross-model): pin the Grok peer to 4.6 (EveryInc#1376)

* docs(skills): rewrite user skill pages for accuracy and clearer use (EveryInc#1377)

* fix(commit): append known plan unit ids to commit subjects (EveryInc#1379)

* fix(ce-work): stop sandboxed workers committing in linked worktrees (EveryInc#1382)

* fix(ce-doc-review): edit HTML plans in native format (EveryInc#1381)

* fix(ce-code-review): cover adversarial after quota or auth no-review (EveryInc#1380)

* fix(skills): correct a rejected dispatch instead of spending the fallback (EveryInc#1383)

* fix(ce-compound): find Claude sessions started outside the repo root (EveryInc#1378)

* ci(windows-native): retry peer-job-runner smoke on ctypes flake (EveryInc#1384)

* fix(ce-debug): stop asking at the handoff, stop shipping unoffered work (EveryInc#1385)

* docs(solutions): record why skill gates state conditions, not git commands (EveryInc#1386)

* fix(skills): drop the residual-findings record file for real sinks (EveryInc#1387)

* fix(ce-doc-review): run the cross-model pass when CROSS_MODEL_PEERS is unset (EveryInc#1389)

* fix(ce-proof): sync with current Proof v3 contract (EveryInc#1390)

* fix(skill-authoring): make goal-first the default when authoring and reviewing skills (EveryInc#1391)

* fix(cross-model): let reviews run on Fable and pin model/effort from CE config (EveryInc#1392)

* docs(cross-model): point superseded peer benchmarks at the luna/xhigh decision (EveryInc#1393)

* fix(cross-model): discover the Codex.app-bundled codex CLI and name the peer-CLI requirement (EveryInc#1395)

* feat(cross-model): add cross_model_review_mode checkout egress gate (EveryInc#1396)

* fix(ce-compound-refresh): compare knowledge-track learnings against guidance they name (EveryInc#1399)

* docs(solutions): capture the named-guidance contradiction-check learning (EveryInc#1400)

* fix(ce-compound): prefer the repo's own frontmatter vocabulary over the Rails-era enums (EveryInc#1394)

* fix(ce-work): stop asking about branches before starting work (EveryInc#1397)

* fix(review): answer covered cases on skill prose with the condition, not a patch (EveryInc#1401)

* fix(scratch): fall back to $TMPDIR when /tmp cannot host the scratch root (EveryInc#1398)

* feat(ce-skill-work): repo-local skill for authoring, editing, reviewing, and responding to review on skills (EveryInc#1402)

* fix(ce-pov): reject non-final peer positions instead of folding them in (EveryInc#1403)

* feat(manifest): add Agent Plugins v1.0.0 manifest support (EveryInc#1345)

* chore: release main (EveryInc#1354)

* fix(ce-work): run cross-model verification on warm checkouts (EveryInc#1404)

* fix(ce-skill-work): author for Sol/Fable, not Opus-era procedure (EveryInc#1408)

* docs(skill-design): retarget stale learning citations (EveryInc#1409)

* fix(ce-setup): support read-only sandboxes (EveryInc#1407)

* fix(ce-skill-work): require pointer descriptions (EveryInc#1410)

* chore: release main (EveryInc#1405)

* fix(ce-work): let a project-defined shipping process override ce-commit-push-pr (EveryInc#1416)

* fix(ce-doc-review): state the CROSS_MODEL_PEERS gate as a condition at the gate (EveryInc#1421)

* fix(ce-strategy): ground the interview in the repo and share the file safely (EveryInc#1419)

* fix(ce-babysit-pr): fall back for private ref 404 (EveryInc#1418)

Co-authored-by: Trevin Chow <trevin@trevinchow.com>

* fix(ce-commit-push-pr): make medium and large PR descriptions scannable (EveryInc#1422)

* chore: release main (EveryInc#1420)

Co-authored-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>

* fix(ce-plan): make Goal Capsule Objective outcome-shaped with a Means slot (EveryInc#1424)

* fix(manifest): drop Agent Plugins $schema so Codex stops truncating skills at 8KB (EveryInc#1426)

* fix(manifest): also block Agent Plugins $schema while skill frontmatter is non-conformant (EveryInc#1427)

* chore: release main (EveryInc#1425)

Co-authored-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>

* docs(solutions): capture why the Agent Plugins $schema is a host routing switch (EveryInc#1428)

* fix(skills): Make CE portable on no-checkout / shared-workspace hosts (EveryInc#1429)

Co-authored-by: Cursor Agent <cursoragent@cursor.com>
Co-authored-by: Trevin Chow <tmchow@users.noreply.github.com>

* chore: release main (EveryInc#1430)

* docs(readme): fix Grok Bot install instructions (EveryInc#1431)

* docs(readme): fix Grok Bot install path (EveryInc#1432)

* fix(ce-babysit-pr): no unasked base merges, 91% smaller, pipelined stacks (EveryInc#1433)

* docs(skill-design): 8KB is a ceiling, not the target, in the size-restructure playbook (EveryInc#1437)

* fix(ce-skill-work): savings come from structure, not squeezed sentences (EveryInc#1460)

* fix(ce-strategy): 53% smaller SKILL.md, references carry the interview and update run (EveryInc#1436)

* fix(ce-strategy): STRATEGY.md is the agreed shared project doc (EveryInc#1459)

* docs(skill-design): delegating skills need a live eval, not a fake-boundary one (EveryInc#1462)

* fix(ce-commit-push-pr): write the scope map first and audit the opening against it (EveryInc#1457)

* fix(ce-proof): 62% smaller SKILL.md, references carry the API and workflows (EveryInc#1443)

* docs(skill-design): record the three relocation failures from the 8KB sweep (EveryInc#1453)

* fix(ce-handoff): 60% smaller SKILL.md, references carry create and resume (EveryInc#1446)

* fix(ce-handoff): separate the writing agent's voice from the user's in handoffs (EveryInc#1464)

* fix(ce-explain): 28% smaller SKILL.md, orchestration moves to a reference (EveryInc#1451)

* fix(ce-debug): 50% smaller SKILL.md, references carry the phase procedures (EveryInc#1449)

* fix(ce-prototype): 43% smaller SKILL.md, references carry scoping and build (EveryInc#1444)

* fix(ce-sweep): 57% smaller SKILL.md, references/run.md carries the phases (EveryInc#1439)

* fix(ce-debug): name the fixed-but-unpushed pipeline outcome at both ends (EveryInc#1463)

* fix(ce-retune): 39% smaller SKILL.md, references carry the rest (EveryInc#1438)

* fix(ce-commit-push-pr): 69% smaller SKILL.md, references carry the rest (EveryInc#1454)

* fix(ce-compound-refresh): 70% smaller SKILL.md, references carry the rest (EveryInc#1442)

* fix(ce-doc-review): 65% smaller SKILL.md, references carry the rest (EveryInc#1450)

* fix(ce-optimize): 82% smaller SKILL.md, references carry the phase procedures (EveryInc#1456)

* fix(ce-setup): 44% smaller SKILL.md, references carry the rest (EveryInc#1445)

* fix(ce-dogfood): 67% smaller SKILL.md, references/phases.md carries the procedure (EveryInc#1447)

* fix(ce-product-pulse): 51% smaller SKILL.md, references carry config and run (EveryInc#1448)

* fix(ce-ideate): 84% smaller SKILL.md, references carry the phases (EveryInc#1455)

* fix(ce-test-browser): 43% smaller SKILL.md, references carry the rest (EveryInc#1441)

* fix(ce-debug): body enum names fixed-not-pushed like the producer and consumer (EveryInc#1466)

* fix(ce-test-browser): resolve the dev-server port with a bundled script instead of a prose-emitted line (EveryInc#1468)

* fix(ce-resolve-pr-feedback): 31% smaller SKILL.md, references carry the rest (EveryInc#1435)

* fix(ce-resolve-pr-feedback): judge top-level PR comments, not only review threads (EveryInc#1467)

* fix(ce-pov): 66% smaller SKILL.md, references carry the rest (EveryInc#1440)

* fix(ce-explain): 32% smaller SKILL.md, the Phase 6 close moves to a required-read reference (EveryInc#1469)

* fix(ce-debug): audit body pins by provenance, relocate what only wording held in the window (EveryInc#1472)

* fix(ce-code-review): 86% smaller SKILL.md, four step references carry the rest (EveryInc#1471)

* fix(ce-compound): 90% smaller SKILL.md, references carry the phases (EveryInc#1477)

* fix(ce-work): 48% smaller SKILL.md, references carry the execution manual (EveryInc#1478)

* fix(ce-plan): 71% smaller SKILL.md, five phase references carry the rest (EveryInc#1470)

* fix(ce-plan): audit body pins by provenance, relocate what only wording held in the window (EveryInc#1475)

* fix(ce-brainstorm): 86% smaller SKILL.md, six references carry the phases (EveryInc#1476)

* fix(lfg): 72% smaller SKILL.md, six required-read references carry the rest (EveryInc#1479)

* docs(skill-design): third-sweep learnings — place blocks by executing step, audit pins, size evals to reach (EveryInc#1483)

* feat(skill-eval): host-CLI catalog for skill-behavior A/B (EveryInc#1484)

* fix(skills): make model-invoked descriptions context pointers (EveryInc#1486)

* fix(ce-brainstorm): ask only decisions the environment cannot settle (EveryInc#1487)

* fix(skills): complete below-cap refactor sweep (EveryInc#1490)

* fix(ce-resolve-pr-feedback): keep replies out of pending reviews (EveryInc#1491)

* fix(skills): unify needs-human decision handoff (EveryInc#1492)

* fix(cross-model): elevate peer start when host sandbox blocks provider network (EveryInc#1496)

* fix: honor active model-config keys in skill runs (surface + salience) (EveryInc#1489)

Co-authored-by: Trevin Chow <trevin@trevinchow.com>

* fix(skills): remove obsolete dispatch context hook (EveryInc#1499)

* fix(ce-optimize): keep multi-objective wins and screen expensive runs (EveryInc#1506)

* fix(skills): load ce-plan and ce-work phases from references under the 8KB cap (EveryInc#1508)

* test(cross-model): give the subprocess-heavy route suites a 30s per-test ceiling (EveryInc#1510)

* fix(cross-model): accept provider-qualified codex model ids in cross_model_model (EveryInc#1501)

* docs(agent-plugins): attribute the 8000-byte skill cap to its real owner (EveryInc#1511)

* chore: release main (EveryInc#1434)

* fix(ce-code-review): stop the reviewer template from embedding the diff twice (EveryInc#1512)

* chore: release main (EveryInc#1513)

* fix(skills): right-size ceremony for small work in ce-plan, ce-brainstorm, and ce-work (EveryInc#1514)

* fix(cross-model): attest Grok Build as a host and bind its native CLI (EveryInc#1516)

* fix(ce-doc-review): activate product-lens only on a product position (EveryInc#1517)

* docs(ce-skill-work): size a behavioral eval to the decision, not the workflow (EveryInc#1520)

* fix(cross-model): bound peer retries and raise review headroom (EveryInc#1519)

* chore: release main (EveryInc#1515)

* docs: refresh cross-model peer learnings (EveryInc#1521)

* docs(solutions): refresh stale learnings against current code (EveryInc#1524)

* docs(readme): make the README a front door, split detail into linked docs (EveryInc#1525)

* fix(skills): match blocking questions from the current tool list (EveryInc#1526)

* fix(lfg): invoke ce-code-review from the host catalog path (EveryInc#1527)

* fix(orca): reconcile upstream 3.23.2 overlay

* fix(ce-babysit-pr): defer to GitHub for trunk drift on managed stacks (EveryInc#1529)

* fix(ce-code-review): collect async reviewer returns (EveryInc#1530)

* chore: release main (EveryInc#1528)

* fix(cross-model): drop empty Codex usage files and unblock heartbeat teardown (EveryInc#1533)

* fix(skills): anchor the plan Objective above the component being changed (EveryInc#1535)

* docs(solutions): record what a condition costs a literal host (EveryInc#1536)

* fix(ce-skill-work): treat a removed concrete shape as a behavior change (EveryInc#1537)

* fix(ce-code-review): keep missing peer config from skipping the pass (EveryInc#1538)

* docs(strategy): add STRATEGY.md as the project's direction anchor (EveryInc#1539)

* fix(concepts): give CONCEPTS.md a retention lifecycle (EveryInc#1540)

* fix(peer-job-runner): keep Windows reap from killing a recycled pid (EveryInc#1541)

* chore: release main (EveryInc#1534)

* fix(ce-compound-refresh): preserve supported guidance across regressions (EveryInc#1542)

* docs(guides): move skill catalog from docs/skills to skills/guides (EveryInc#1551)

Co-authored-by: Cursor Agent <cursoragent@cursor.com>

* perf(tests): reuse seed git fixtures to cut suite wall time (EveryInc#1548)

Co-authored-by: Cursor Agent <cursoragent@cursor.com>

* fix: preserve release automation ownership after upstream sync

* fix(ce-work): parallelize independent work by default; fix(lfg): narrate progress and gate DONE (EveryInc#1558)

Co-authored-by: Claude <noreply@anthropic.com>

* fix(ce-compound): check in-repo absolute path citations (EveryInc#1561)

Co-authored-by: Cursor Agent <cursoragent@cursor.com>

* fix(guides): keep catalog out of the plugin skills tree (EveryInc#1571)

Co-authored-by: Cursor Agent <cursoragent@cursor.com>

* fix(ce-code-review): take review criteria from a repo-owned standards file (EveryInc#1572)

* docs(guides): document the CODING_STANDARDS.md criteria file (EveryInc#1573)

* fix(ce-work): scrub GIT_INTERNAL_SUPER_PREFIX from inherited git env (EveryInc#1564)

* fix(ce-setup): detect the retired Codex tool map (EveryInc#1577)

* fix(lfg): record unapplied review findings in the PR body, not tickets (EveryInc#1580)

* fix(ce-plan): require a holdable Goal Capsule Objective (EveryInc#1592)

* fix(ce-commit-push-pr): let a PR opening carry what motivates it (EveryInc#1576)

* fix(ce-code-review): anchor reviewer checks to named canonical frameworks (EveryInc#1594)

* fix(ce-commit-push-pr): catch a PR opening that names the mechanism (EveryInc#1595)

* fix(ce-babysit-pr): keep stale-computation degrade off the settle clock (EveryInc#1568)

Co-authored-by: Trevin Chow <trevin@trevinchow.com>

* docs(ce-compound): capture at the completion checkpoint, not at merge (EveryInc#1596)

* fix(ce-work): allow parallel subagent waves in a shared workspace (EveryInc#1598)

* fix(codex): reconcile config-only plugin installs (EveryInc#1599)

* fix(ce-debug): gate hypotheses on a runnable reproduction check (EveryInc#1600)

* docs(lfg): align leftover-findings FAQ with PR-body checklist (EveryInc#1582)

Co-authored-by: chouti <chouti@upai.com>

* fix(ce-commit): make the stage-and-commit example PowerShell-safe (EveryInc#1587)

Co-authored-by: Trevin Chow <trevin@trevinchow.com>

* fix(ce-work): require evidence before unavailable review fallback (EveryInc#1578)

Co-authored-by: Trevin Chow <trevin@trevinchow.com>

* fix(ce-babysit-pr): persist review invariant rounds across heads (EveryInc#1585)

Co-authored-by: Trevin Chow <trevin@trevinchow.com>

* fix(ce-code-review): isolate testing reviewer mutation work (EveryInc#1584)

Co-authored-by: Trevin Chow <trevin@trevinchow.com>

* fix(ce-work): out-of-repo units have no git-derived completion (EveryInc#1583)

Co-authored-by: Trevin Chow <trevin@trevinchow.com>

* fix(ce-work): verify worktree snapshot fidelity before isolated dispatch (EveryInc#1602)

* feat(harness): add opencode as a named peer and work engine (EveryInc#1604)

Co-authored-by: Hally Maschine <hally@rocketable.com>

* fix(ci): absorb the windows _ctypes flake without hiding real failures (EveryInc#1605)

* chore: release main (EveryInc#1543)

Co-authored-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>

* fix(orca): re-pin upstream baseline to 3.24.0

Record compound-engineering-v3.24.0 as the fork baseline, retarget the
pending release identity to 3.24.0-orca.7, and regenerate skill-local
role registries.

---------

Co-authored-by: Trevin Chow <trevin@trevinchow.com>
Co-authored-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
Co-authored-by: Ruslan Kurkebayev <kurkebayev.ruslan@gmail.com>
Co-authored-by: cmbish <carter.m.bish@gmail.com>
Co-authored-by: Cursor Agent <cursoragent@cursor.com>
Co-authored-by: Trevin Chow <tmchow@users.noreply.github.com>
Co-authored-by: Kevin Old <kevin@kevinold.com>
Co-authored-by: Dennis Traub <dennis.traub@gmail.com>
Co-authored-by: Kieran Klaassen <kieranklaassen@users.noreply.github.com>
Co-authored-by: Claude <noreply@anthropic.com>
Co-authored-by: chouti <chouti@gmail.com>
Co-authored-by: Saurabh <saurabhkagent@gmail.com>
Co-authored-by: chouti <chouti@upai.com>
Co-authored-by: Morgan Touverey Quilling <morgan@mtq.io>
Co-authored-by: Thomas Steibl <309042+Rowdy@users.noreply.github.com>
Co-authored-by: Hally Maschine <hally@rocketable.com>
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

ce-code-review: dispatch mutation-testing reviewers in an isolated worktree

2 participants