## batch-core continued — the source-issue cluster Follow-on to #2780 (merged). Scope is still `packages/core` + `packages/dashboard/src`. ### The defect Five places asked the same question — *has this task landed?* — and all five compared against the literal `done`: | surface | consequence on a renamed board | |---|---| | GitHub source-issue commenter | never comments on or closes the source issue | | GitLab source-issue commenter | same | | GitLab `closedAt` backfill reconciler | finds nothing, reports a clean scan | | session-diff boundary | finished tasks diff against an already-merged branch | | tracking-comment transition | (already converted; left alone) | The commenters are the sharpest case: they returned **before reading a single setting**, so on a renamed board the feature looked *disabled* rather than broken — an operator checking `githubCommentOnDone` would see it enabled and still get nothing. The backfill is the quietest: `scanned: N, filled: 0` reads as "nothing to do", so the failure was indistinguishable from success. ### The fix One home: `packages/dashboard/src/task-lifecycle-lanes.ts`. Callers now only ask. Five copies of one question is exactly how the halves drift apart — the motivating incident is FN-6115 → FN-6118 → FN-6123, where the same affordance was fixed three times because it lived in two components. This also folds in the duplicate landed-lane helper I had left in `register-session-diff-routes.ts` in the previous PR, which was the sixth copy waiting to happen. Two helpers, and the difference is deliberate: - **`landedColumnsForTask`** — `complete ∪ archived`. Membership, since a board may declare more than one column carrying either role, and `columnsWithFlag(...)[0]` would silently ignore the second. - **`completeColumnsForTask`** — complete only. The GitLab backfill's own FNXC note records that archived tasks live in `archiveDb` and are *intentionally* excluded, so it must not widen to the archived role just because the shared helper offers it. Today it lists with `includeArchived: false` and would see no archived rows either way — but that is an incidental property of the query, not the contract. The test pins the difference so the two are not later "simplified" into one, which would change that caller's behaviour without touching it. Both treat an **empty** resolved set as *unexpressed*, not absent — the v1 hazard: `synthesizeDefaultColumns` upgrades a v1 graph with `traits: []` on every column, so reading empty as "no complete lane" would stop these surfaces firing on every pre-v2 project. The reconciler is two-stage on purpose: the cheap provider and `closedAt` tests run first and reject almost everything, so a workflow read only happens for real candidates, and it shares one IR cache across the scan — one read per distinct workflow rather than per task. ### Census `batch-core` scope **75 → 72**; repo total **338**. ### Verification - `pnpm --filter @fusion/dashboard exec tsc --noEmit -p tsconfig.json` → 0 errors - `pnpm lint` → 0 errors - commenter + reconciler suites → **63 passed**; helper suite → **5 passed** - **Mutation-verified:** making the helper ignore its resolved set fails 1 of 5. --- ## Round 2 — server.ts, chat.ts, and a correction **Census: 75 → 67** across this PR. ### The correction (see the review thread above) My first pass gated the source-issue commenters on `landedColumnsForTask` (`complete ∪ archived`), which **widened** the trigger — `to === "done"` never fired on archival, and the landed set does. Both commenters now use `completeColumnsForTask`, and the unused `hasTaskLanded` wrapper is gone. The ratchet for it is pinned on the **default** board, deliberately: a widening is visible exactly where the legacy names still apply, so no renamed-board fixture would catch it. ### `chat.ts` — three sites, and a pair that had to move together - **Chat verification** required `column === "in-progress"`, so on a renamed board every chat-driven verification was refused with a message naming a column the board does not have. - **The planner refinement pair.** Two separate guards decide this feature: `createSession` *registers* the tool only for a finished task, and the tool's own `execute()` *refuses* a non-finished source. Both compared `done`. Converting only one half would have offered the tool and then had it refuse itself — the half-converted-pair shape. The new test asserts **both** halves in one case (tool present *and* refinement created), and each half reverted independently fails it. Existing `chat-manager` coverage caught neither revert, which is why the case exists rather than relying on the suite that was already there. Complete-only again, not the landed set: an archived task is off the board and is not a refinement source. ### `server.ts` - **Planner-chat retention** — the archival cutoff was a literal, so on a renamed board task-planner chat sessions were retained forever; the rule this listener exists to enforce never fired. Resolved, and awaited inside the existing fire-and-forget chain rather than by making the listener `async` — `task:moved` has synchronous subscribers whose ordering is load-bearing elsewhere, and a chat-row delete is not the right place to introduce a microtask boundary into that emit. - **`isBadgeEligibleTask` — deliberately NOT converted, and marked as backlog.** On a renamed board it is genuinely wrong: an archived card stays badge-eligible, its snapshot is never evicted, and the cache grows for the daemon's lifetime — the exact memory leak the predicate was added to fix, back under a different column name. What blocks it is measured, not assumed: both callers are synchronous `task:updated` / `task:created` listeners whose next statement is documented as *"Update local cache immediately"*, so awaiting lets a second event for the same task interleave between the eligibility check and the cache write. I did **not** add an optional `archivedColumns` parameter, because nothing could fill it — the callers are the sync listeners. That is the inert-injection shape this PR's own review caught twice on #2780: the predicate would read as converted, its test would pass by injecting the value, and production would keep the literal. The unblocking change (a resolved-archived-lane cache on the badge-snapshot scope, keeping the predicate synchronous) is recorded at the site. ### Verification - `tsc --noEmit` → 0 errors; `pnpm lint` → 0 errors - `chat-manager` → 101 passed; commenter/reconciler/helper/badge suites → 55 passed - Mutation-verified per fix, including each half of the refinement pair separately <!-- This is an auto-generated comment: release notes by coderabbit.ai --> ## Summary by CodeRabbit * **Bug Fixes** * Improved task lifecycle handling for renamed workflow lanes, including completed, archived, landed, and in-progress states. * Task lists now exclude completed tasks regardless of the completion lane’s name. * Chat verification and refinement actions now recognize configured workflow lanes. * GitHub and GitLab completion comments trigger only for genuinely completed tasks, not archived tasks. * Knowledge index refreshes and GitLab metadata updates now support custom completion lanes. * **Tests** * Added regression coverage for renamed completion lanes and archived-task behavior. <!-- end of auto-generated comment: release notes by coderabbit.ai --> --------- Co-authored-by: Claude Opus 5 (1M context) <noreply@anthropic.com>
118 lines
5.3 KiB
TypeScript
118 lines
5.3 KiB
TypeScript
/**
|
|
* Canonical @fusion/core and @fusion/engine mock helpers for dashboard server tests.
|
|
*
|
|
* If a route test starts failing with "No \"X\" export is defined", update this
|
|
* helper first instead of adding another full inline export map in the test file.
|
|
*/
|
|
import { vi, type Mock } from "vitest";
|
|
|
|
type AnyModule = Record<string, unknown>;
|
|
type AnyMock = Mock;
|
|
|
|
const fallbackFns = new Map<string, AnyMock>();
|
|
|
|
function getFallback(name: string): AnyMock {
|
|
if (!fallbackFns.has(name)) fallbackFns.set(name, vi.fn());
|
|
return fallbackFns.get(name)!;
|
|
}
|
|
|
|
function withFallbackFunctions(actual: AnyModule, moduleValue: AnyModule): AnyModule {
|
|
return new Proxy(moduleValue, {
|
|
get(target, prop, receiver) {
|
|
if (typeof prop !== "string") return Reflect.get(target, prop, receiver);
|
|
if (Reflect.has(target, prop)) return Reflect.get(target, prop, receiver);
|
|
if (["then", "catch", "finally"].includes(prop)) return undefined;
|
|
|
|
const actualValue = actual[prop];
|
|
if (typeof actualValue === "function" || actualValue === undefined) {
|
|
const fn = getFallback(prop);
|
|
target[prop] = fn;
|
|
return fn;
|
|
}
|
|
return actualValue;
|
|
},
|
|
});
|
|
}
|
|
|
|
export async function createCoreMock(
|
|
importActual: () => Promise<AnyModule>,
|
|
overrides: AnyModule = {},
|
|
): Promise<AnyModule> {
|
|
const actual = await importActual();
|
|
return withFallbackFunctions(actual, { ...actual, ...overrides });
|
|
}
|
|
|
|
export function createEngineMock(overrides: AnyModule = {}): AnyModule {
|
|
const actual: AnyModule = {};
|
|
return withFallbackFunctions(actual, {
|
|
createFnAgent: vi.fn(),
|
|
promptWithFallback: vi.fn(),
|
|
/*
|
|
FNXC:TestSkills 2026-06-17-19:33:
|
|
Dashboard route tests mock @fusion/engine wholesale, so skill-aware planning lanes need a shaped session-skill helper result instead of the fallback vi.fn() returning undefined.
|
|
*/
|
|
buildSessionSkillContextSync: vi.fn(() => ({
|
|
skillSelectionContext: undefined,
|
|
resolvedSkillNames: [],
|
|
skillSource: "none" as const,
|
|
})),
|
|
// Returns an iterable tool list; dashboard code spreads its result
|
|
// (`...createWorkflowAuthoringTools(...)`), so it must not be undefined.
|
|
createWorkflowAuthoringTools: vi.fn(() => []),
|
|
/*
|
|
FNXC:DashboardRouteTests 2026-06-18-09:07:
|
|
Planning and chat route files can share worker-level @fusion/engine mocks during broad dashboard API quality runs.
|
|
Keep chat task document tools iterable by default so rescuing chat-routes from quarantine does not poison planning route imports with a fallback vi.fn() result.
|
|
*/
|
|
createChatTaskDocumentTools: vi.fn(() => []),
|
|
createChatArtifactTools: vi.fn(() => []),
|
|
/*
|
|
FNXC:MissingWorktreeRetry 2026-07-10-18:45:
|
|
Dashboard route tests mock @fusion/engine wholesale; the retry route must still exercise the upstream #1992 classifier so merge-active unusable-worktree failures are admitted while unrelated merging rows remain rejected.
|
|
*/
|
|
/*
|
|
FNXC:WorkflowResolvedColumns 2026-07-30-07:00 DELIBERATE-LITERAL: a test double mirroring production's own fallback.
|
|
|
|
Production's `isInReviewMissingWorktreeSessionStartFailure` is `(isReviewColumn ?? task.column ===
|
|
"in-review") && ...` — the literal IS its documented degraded path for a caller that has not
|
|
resolved the lane. A double must reproduce that, not improve on it.
|
|
|
|
FIDELITY FIX while marking it: this ignored the second parameter entirely, so a route test that
|
|
passed a resolved `isReviewColumn` got the literal answer anyway and would have reported a pass
|
|
for a renamed board the real classifier handles. Now threaded exactly as production does.
|
|
*/
|
|
isInReviewMissingWorktreeSessionStartFailure: vi.fn((task: { column?: string; error?: unknown }, isReviewColumn?: boolean) => (
|
|
(isReviewColumn ?? task.column === "in-review")
|
|
&& typeof task.error === "string"
|
|
&& (task.error.includes("Refusing to start coding agent in missing worktree:")
|
|
|| task.error.includes("Refusing to start coding agent in incomplete worktree:")
|
|
|| task.error.includes("Refusing to start coding agent in unregistered git worktree:"))
|
|
)),
|
|
// FNXC:McpConfig 2026-07-02-13:45: Planning/mission route tests share this engine mock; MCP resolution must return the full shaped empty result so readonly session creation can proceed without importing real engine stores.
|
|
resolveMcpServersForStore: vi.fn(async () => ({ servers: [], errors: [] })),
|
|
/*
|
|
FNXC:TaskCreateDedup 2026-07-18-15:55:
|
|
FN-8277 routes planning/subtask creation through createAgentTask. The wholesale
|
|
@fusion/engine mock previously fell back to vi.fn() → undefined, so
|
|
`const { task } = await createAgentTask(...)` threw. Default to a thin wrapper
|
|
that uses the real store.createTask and reports wasDuplicate:false.
|
|
*/
|
|
createAgentTask: vi.fn(async (
|
|
store: { createTask?: (input: unknown, options?: unknown) => Promise<unknown> },
|
|
input: unknown,
|
|
_options?: unknown,
|
|
) => {
|
|
if (typeof store?.createTask !== "function") {
|
|
throw new Error("createAgentTask mock requires store.createTask");
|
|
}
|
|
const task = await store.createTask(input, { settings: {} });
|
|
return { task, wasDuplicate: false };
|
|
}),
|
|
...overrides,
|
|
});
|
|
}
|
|
|
|
export function resetDashboardServerMockState(): void {
|
|
for (const fn of fallbackFns.values()) fn.mockReset();
|
|
}
|