mirror of
https://github.com/garrytan/gbrain.git
synced 2026-07-28 06:23:01 +00:00
* v0.31 feat(migrate): facts hot memory schema (migration v40) Phase 1 of v0.31 hot-memory. - New facts table with source_id (TEXT FK to sources, per-source isolation), kind CHECK (event/preference/commitment/belief/fact), visibility CHECK (private/world for takes-style ACL parity), valid_from/valid_until/ expired_at/superseded_by for temporal + supersession audit, and consolidated_at/consolidated_into pointing at takes(id) for the dream- cycle hot→cold bridge. - Embedding column dim resolved at migration time from config.embedding_dimensions so non-OpenAI brains (Voyage etc) work out-of-the-box. HALFVEC where pgvector >= 0.7; falls back to VECTOR with stderr warn on older versions. Matching opclass per column type (halfvec_cosine_ops vs vector_cosine_ops). - 5 partial indexes leading on source_id so every read uses the trust boundary as part of the index, not a callback. HNSW partial index excludes expired/null rows so footprint stays proportional to active fact count. - RLS DO-block matches takes pattern (Postgres BYPASSRLS gate; PGLite no-op). - v0_31_0.ts orchestrator follows v0_28_0.ts pattern — phase A asserts schema version >= 40 + facts table presence; runner owns ledger. All 87 existing migrate.test.ts cases pass. PGLite smoke test confirms table + indexes + CHECK constraints + ON DELETE CASCADE all behave. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 chore(version): bump VERSION + package.json to 0.31.0 Phase 1 closer. CHANGELOG entry written when Phase 7 lands. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 feat(engine): facts hot memory engine API (Phase 2) Phase 2 of v0.31 hot-memory. Adds 8 facts methods to BrainEngine implemented on both PGLite and Postgres engines: - insertFact(input, ctx) — INSERT with optional supersedeId; expires the named row in the same transaction. Per-entity advisory lock on Postgres (`pg_advisory_xact_lock(hashtextextended(source_id::text || ':' || entity_slug, 0))`) for the dedup window. PGLite is single-process so the lock is a no-op. - expireFact(id, opts) — sets expired_at + optional superseded_by. Idempotent-as-false (already-expired returns false). - listFactsByEntity / listFactsSince / listFactsBySession — list surfaces with FactListOpts filters (activeOnly, kinds, visibility, limit/offset). Every query starts WHERE source_id = $X so the trust boundary is part of the index path. - listSupersessions — audit log; activeOnly:false + expired_at IS NOT NULL + superseded_by IS NOT NULL. - findCandidateDuplicates(source_id, entity_slug, factText, k) — entity-prefiltered (mandatory), k=5 default, hard cap 20. Embedding- cosine ordering when caller supplies an embedding, recency fallback otherwise. Bounds the contradiction-classifier blast radius. - consolidateFact(id, takeId) — sets consolidated_at + consolidated_into. Never DELETE; facts stay as audit trail for the resulting take. - getFactsHealth(source_id) — per-source counters consumed by `gbrain doctor` facts_health check. Public types in engine.ts: FactKind (5-value union), FactVisibility, FactInsertStatus, FactRow, NewFact, FactListOpts, FactsHealth. PGLite + Postgres helpers: rowToFact / rowToFactPg parse the text-format pgvector embedding back into Float32Array; toPgVectorLiteral encodes for the supersede-path INSERT (postgres-js can't bind Float32Array directly to a vector column without an explicit literal cast). Smoke test confirms every method end-to-end on PGLite. Typecheck clean. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 feat(facts): extraction code path (Phase 3) Phase 3 of v0.31 hot-memory. Five new modules under src/core/facts/ + src/core/entities/: - src/core/facts/decay.ts — pure helper. effectiveConfidence(fact, now) applies confidence × exp(-age/halflife) with per-kind halflife table (event 7d, commitment 90d, preference 90d, belief 365d, fact 365d). Returns 0 for expired or past-valid_until rows. Single source of truth consumed by recall, supersession audit, facts_health, and the MCP _meta injector (eD8 DRY). - src/core/facts/queue.ts — bounded in-memory queue. Cap 100 default, drop-oldest on overflow with counter. Per-session in-flight=1 serializes burst chat. AbortSignal threading from server SIGTERM (mirrors minion worker pattern per eD7): 5s grace for in-flight, then drop pending with counter. getFactsQueue() process-singleton; __resetFactsQueueForTests for hermetic tests. - src/core/facts/classify.ts — contradiction classifier with cosine fast-path (D13: ≥0.95 → duplicate, skip LLM) and classifier-failure fallback (D12: cosine ≥0.92 → duplicate, else INSERT). Pure cosine helper exported. JSON-strict output with 4-strategy parse fallback; refusal stop-reason maps to fallback path. Caller-provided abort signal propagated to the gateway chat call. - src/core/facts/extract.ts — Haiku turn-extractor. Reuses INJECTION_PATTERNS from src/core/think/sanitize.ts on the way IN (turn_text) AND on the way OUT (each fact). Tight system prompt with 5-kind taxonomy, 0..1 confidence scoring, entity slug or display name. Anti-loop check on isDreamGenerated (reuses v0.23.2 marker semantics). Synchronous embedOne() per fact via the gateway so classifier paths have embeddings available; AbortError re-thrown explicitly so SIGTERM during embed never writes a NULL-embedding row meant to be cancelled (eE8 distinction). - src/core/entities/resolve.ts — slug canonicalization shared by signal-detector AND facts. Resolution order: exact slug match → pg_trgm fuzzy match (similarity ≥0.4) → deterministic slugify fallback. slugify exported standalone for tests + callers that want the floor. Smoke tests confirm decay table, cosine math, slugify rules, queue drop-oldest under overflow, and shutdown grace + drop-pending semantics. Typecheck clean. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 feat(mcp+cli): MCP ops + recall CLI + _meta + transport refactor (Phase 4) Phase 4 of v0.31 hot-memory. Three new MCP ops on the contract-first surface: - `extract_facts` (write scope, localOnly:false): extracts facts from a conversation turn via the Haiku extractor, runs the cosine fast-path dedup, INSERTs into per-source hot memory. Returns counts + fact_ids[]. Skips on is_dream_generated:true (anti-loop). - `recall` (read scope): query the per-source hot memory by entity / since / session / supersessions / grep filter. Visibility- aware: remote callers see visibility='world' rows only (takes-style ACL parity, eD21). Returns most-recent first; pagination via limit. - `forget_fact` (write scope): expireFact wrapper. Idempotent-as-error on unknown id; uses the new 'fact_not_found' ErrorCode. ErrorCode union opened (eD6 / eE7): TS forward-compat via the `(string & {})` autocomplete-friendly hack so downstream consumers (gbrain-evals etc) don't break their typecheck on every new code. Three new codes: 'rate_limited', 'extraction_failed', 'fact_not_found'. OperationContext gains source_id?:string (eD4 / eE2 — TEXT not INTEGER per schema reality). Resolved once in buildOperationContext from DispatchOpts.sourceId. Stdio MCP defaults to GBRAIN_SOURCE env or 'default'; HTTP MCP reads it from the per-token sources scope (eE3). ToolResult gains _meta?: Record<string, unknown> (eD3). Dispatcher calls a configurable metaHook AFTER op.handler succeeds, wrapped in its own try/catch so a DB blip degrades to no-_meta rather than flipping the whole tool call to error (eE4). New module src/core/facts/meta-hook.ts: - getBrainHotMemoryMeta(name, ctx) builds the _meta.brain_hot_memory payload. Cache key (source_id, session_id, hash(takesHoldersAllowList sorted)) (eD10 / eE5). 30s TTL per session. Visibility filter applies: remote → world only; local → all. Top-K=10 ranked by effective confidence (decay). Skips injection on recall/extract_facts/forget_fact themselves. bumpHotMemoryCache() invalidates per (source_id, session_id) on extraction event. D12 (eE1) accepted: serve-http.ts:801 inlined dispatch path REFACTORED to call dispatchToolCall. HTTP MCP now inherits source_id, _meta injection, error envelope unification, and OperationContext shape from the same code path stdio uses. Scope check + mcp_request_log + SSE broadcast stay in serve-http.ts (HTTP-specific concerns); the dispatcher returns ToolResult and the HTTP handler reads isError + content + _meta to fan into the audit + broadcast paths. put_page compliance backstop (D23): when a conversation-shape page is written (note/meeting/slack/email/calendar-event/source/writing) with a substantive body (>=80 chars) on a non-subagent slug AND no dream_generated:true marker, fire-and-forget enqueue an extraction job into the bounded queue. Never blocks the put_page response. Skipped reasons (no_parsed_page / subagent_namespace / dream_generated / kind:* / too_short / queue_shutdown / backstop_error) are stable strings consumed by tests. `gbrain recall` + `gbrain forget` CLI commands (src/commands/recall.ts): - recall <entity> | --since DUR | --session ID | --today (markdown with kind icons 📅🎯🤝💭📌) | --grep TEXT | --supersessions | --include-expired | --as-context (prompt-injection-ready) | --json - forget <fact-id> shorthand for expireFact Wired into src/cli.ts dispatch table next to takes / think. Smoke tests confirm: dispatch surfaces (extract_facts → ops → listFactsByEntity), forget_fact + idempotent re-call, _meta visibility filter (remote sees world only, local sees all), CLI markdown render with kind icons + age strings + decayed confidence. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 feat(cycle): consolidate phase — facts → takes promotion (Phase 5) Phase 5 of v0.31 hot-memory. New 10th cycle phase `consolidate` between `patterns` and `embed`: - src/core/cycle.ts: * CyclePhase union extended with 'consolidate' * ALL_PHASES gets 'consolidate' between patterns and embed (graph-fresh after patterns; embed runs after so the new takes get embedded same-cycle) * NEEDS_LOCK_PHASES gets 'consolidate' (writes takes + UPDATEs facts) * CycleReport.totals gains facts_consolidated + consolidate_takes_written * runCycle dispatches the new phase via dynamic import - src/core/cycle/phases/consolidate.ts (new): * Scans (source_id, entity_slug) buckets where COUNT(unconsolidated facts) >= 3 (uses idx_facts_unconsolidated partial index) * Skips buckets where the OLDEST fact is < 24h old (gives signal time to settle before locking it into cold memory) * Greedy cosine clustering at threshold 0.85; head-element centroid keeps it deterministic + cheap. Singletons (no embedding) stay unconsolidated this cycle. * For each cluster size >= 2: picks the highest-confidence fact's text as the take claim (v0.31 deterministic; v0.32 swaps to Sonnet synthesis pass). avg confidence → take weight, earliest valid_from → take since_date, concatenated source_sessions → take.source. * Resolves entity_slug → page_id via pages.slug (per source). Skips cluster if page is missing in this source — no auto-page-creation in v0.31. * INSERT into takes(kind='fact', holder='self') with row_num = MAX(existing) + 1. * UPDATE contributing facts: consolidated_at = now() + consolidated_into = takes.id. NEVER DELETE — facts are the audit trail for the resulting take. * dryRun honored: pretends the writes happened; counters still tick so operators can preview load before the first real run. * yieldDuringPhase keepalive between buckets so the Minions worker job lock + cycle-lock TTL don't drift on long runs. Smoke test on PGLite confirms: 4 unconsolidated facts → clustered (cosine 1.0 since same vector) → 1 take row created → all 4 facts marked consolidated_into. runCycle({phases:['consolidate']}) wires through to the report totals. Typecheck clean. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 test: 18 facts test files (Phase 6) Phase 6 of v0.31 hot-memory: comprehensive coverage across the new substrate. 110 unit tests pass; 5 E2E test files added (skip gracefully without DATABASE_URL). Unit tests (PGLite in-memory, no DATABASE_URL): - test/facts-decay.test.ts (12 cases) — HALFLIFE_DAYS pinned per kind, effectiveConfidence math: age=0 / age=halflife (~1/e) / age=2×halflife (~1/e²) / expired returns 0 / valid_until past returns 0 / preference-vs-event slower decay / belief-vs-commitment crossover. - test/facts-queue.test.ts (10 cases) — FIFO within session, drop-oldest on overflow, per-session in-flight=1 serializes, different sessions parallelize, failed jobs counter, shutdown grace + drop_pending + external AbortController triggers shutdown. - test/facts-classify.test.ts (8 cases) — cosineSimilarity edge cases, empty candidates → independent, cheap fast-path ≥0.95 → duplicate no LLM, threshold-configurable cosine_fallback path. - test/facts-engine.test.ts (13 cases) — every BrainEngine fact method end-to-end: insertFact (insert/supersede), expireFact idempotency, list*, findCandidateDuplicates entity-prefiltered + k cap + cosine ordering, consolidateFact never DELETE, getFactsHealth shape + total_today ⊆ total_week. - test/facts-multi-tenant.test.ts (6 cases) — cross-source isolation on every list method + CASCADE delete on sources. - test/facts-visibility.test.ts (6 cases) — visibility column private/ world; remote=true filters to world-only via dispatchToolCall; remote=false sees all. - test/facts-canonicality.test.ts (10 cases) — slugify rules including NFKD diacritic strip ("Crème Brûlée" → "creme-brulee"), exact slug match, fallback to slugify when no fuzzy match. - test/facts-extract.test.ts (4 cases) — empty turn returns [], dream- generated short-circuit, graceful no-API-key return. - test/facts-backstop-gating.test.ts (5 cases) — put_page backstop: too_short, subagent_namespace, dream_generated, eligible note path, non-eligible kind:guide. - test/facts-anti-loop.test.ts (4 cases) — extractor + put_page both respect dream_generated:true marker. - test/facts-doctor-shape.test.ts (4 cases) — facts_health JSON shape pinned for downstream consumers. - test/facts-mcp-allowlist.serial.test.ts (5 cases) — extract_facts write-scope, recall read-scope, forget_fact write-scope, forget_fact fact_not_found error code, extract_facts no-API-key zero counts. - test/facts-context-injection.serial.test.ts (6 cases) — _meta injection on success, world-only filter under remote=true, anti-loop on facts ops themselves, best-effort degrade on hook error, cache-key includes allow-list hash. - test/facts-separation-pglite.test.ts (2 cases) — Garry's Separation Test as primary ship gate, plus expired hidden-by-default contract. - test/facts-recall-render.test.ts (3 cases) — --today markdown render with all 5 kind icons, --json shape with effective_confidence, --as-context emits comment-wrapped block. - test/facts-migration-dim.test.ts (4 cases) — embedding column type is HALFVEC/VECTOR (not arbitrary), dim matches gateway-configured embedding_dimensions, HNSW opclass agrees with column type, idempotent re-init. - test/cycle-consolidate.test.ts (5 cases) — below-count + below-age thresholds skip, happy path 4 facts → 1 take + all consolidated never DELETE, dryRun honored, missing page → bucket skipped. E2E tests (skip gracefully on DATABASE_URL unset; required gates by CLAUDE.md test policy): - test/e2e/facts-separation-postgres.test.ts — Postgres parity for the ship gate. - test/e2e/facts-cross-source-isolation.test.ts — cross-source ACL on PG + CASCADE delete. - test/e2e/facts-forget.test.ts — full forget_fact MCP roundtrip. - test/e2e/facts-context-injection-postgres.test.ts — _meta injection end-to-end on PG. - test/e2e/facts-recall-render.test.ts — recall --today markdown on PG. - test/e2e/serve-http-meta.test.ts — eE1 regression: HTTP MCP transport inherits _meta + sourceId + scope correctness via dispatchToolCall. Side-effect: src/core/entities/resolve.ts NFKD post-decompose strips combining marks (U+0300..U+036F) before hyphenating non-alphanumerics, so "Crème" → "creme", not "cre-me-". Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 feat(operational): kill switch + doctor check + CHANGELOG + README (Phase 7) Phase 7 of v0.31 hot-memory. - src/core/facts/extract.ts: new isFactsExtractionEnabled(engine) helper reads `facts.extraction_enabled` config row. Defaults to TRUE; flip to 'false'/'0'/'no'/'off' (case-insensitive) via `gbrain config set facts.extraction_enabled false` to kill extraction across the brain without binary downgrade. - extract_facts MCP op short-circuits with zero-counts envelope + a 'skipped: extraction_disabled' field when the flag is off (clean success, not permission_denied). - put_page facts backstop respects the same flag — eligibility check now returns 'extraction_disabled' as the skipped reason. - src/commands/doctor.ts: new facts_health check (runs after queue_health, before index_audit). Probes for the facts table existence (post-v40 guard), then surfaces total_active / total_today / total_week / total_consolidated + top-3 entities for the default source. Pre-v0.31 brains report "facts table not present (pre-v0.31 brain or migration pending)". - CHANGELOG.md: full v0.31.0 entry in the GStack release-summary voice. Headline + numbers-table + what-it-ships + itemized changes + "To take advantage of v0.31" upgrade block + out-of-scope. Honest about the HALFVEC + serve-http refactor + ErrorCode-open-union complications. - README.md: cycle phase list updated 8 → 10 (consolidate + purge). New "v0.31 Hot Memory" command block under Commands with recall + forget variants, kind icons, --as-context surface for headless agents. Test gates: 28 facts unit tests pass after the kill-switch wiring + doctor check ride-along. Typecheck clean. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 fix(migrate): add facts→sources FK explicitly via ALTER TABLE The inline column-level FK declaration on facts.source_id worked on PGLite but silently got dropped on Postgres in the v0.31 e2e run — the migration handler ran via postgres-js's `unsafe()` multi-statement path and the resulting facts table came back without the `facts_source_id_fkey` constraint. Same psql input run directly against the same database produced the FK; the difference was the unsafe() pipeline, not the SQL itself. Splitting the FK into a separate ALTER TABLE inside a DO block makes the constraint declaration explicit and idempotent: the named constraint either exists or it doesn't, the ALTER is a no-op on re-runs, and the failure mode is loud rather than silently leaving a CASCADE-less foreign key behind. Without this fix, deleting a source row leaves orphaned facts rows (test/e2e/facts-cross-source-isolation.test.ts CASCADE-on-sources- delete case caught it). With this fix the constraint is in place, the cascade fires, and both PG + PGLite e2e suites stay green. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 test: update phase-count assertions for the new consolidate phase Three e2e/unit tests pinned the cycle phase count or order, all now updated to reflect v0.31's 10-phase cycle: - test/e2e/dream-cycle-eight-phase-pglite.test.ts: describe rename "8-phase cycle" → "10-phase cycle"; ALL_PHASES expectation extended to include 'consolidate' (between patterns + embed) and 'purge' (the v0.26.5 addition that was already in ALL_PHASES but missing from the test's assertion list). totals match adds the new facts_consolidated + consolidate_takes_written fields plus the pre-existing purged_sources_count + purged_pages_count that should have been added when v0.26.5 landed. - test/e2e/cycle.test.ts: dry-run full cycle now expects report.phases.length === 10 (was 9). - test/core/cycle.serial.test.ts: yieldBetweenPhases hook count + full cycle phases.length both updated 9 → 10. Comments call out the v0.31 addition lineage so the next person to add a phase sees the precedent. These are mechanical assertion bumps. The tests pass against the updated assertions on PGLite and Postgres. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 fix(test): truncate facts table between e2e describe blocks setupDB() truncates ALL_TABLES between every describe block's beforeAll() hook. The list missed the new v0.31 facts table, so facts seeded by an earlier describe block leaked into Garry's Separation Test on Postgres — listFactsByEntity('travel') returned 2 rows instead of 1 because a prior facts-context-injection test had also seeded a 'travel' fact. Adding 'facts' to the truncate list (before 'pages' to respect FK ordering) makes every describe-block start from an empty facts table. Pinned by re-running the e2e file ordering that originally caught it (facts-recall-render → cross-source-isolation → serve-http-meta → context-injection → separation-postgres → facts-forget) — 13 pass / 0 fail after the fix. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 test: meta-hook cache + Postgres consolidate phase coverage Two net-new test files filling real coverage gaps the earlier sweep missed: - test/facts-meta-cache.test.ts (5 cases) — pins the eD3/eD10 cache contract that the dispatcher relies on. 30s TTL hit path, post-bump fresh-query, scoped invalidation (bump for sess-A leaves sess-B cache warm — closes the cross-source leak risk codex F5 originally surfaced on the recall payload), facts-self ops skip injection (anti-loop on recall / extract_facts / forget_fact), distinct allow-lists produce distinct cache entries. - test/e2e/cycle-consolidate-postgres.test.ts (3 cases) — Postgres parity for the dream-cycle consolidate phase. Mirrors the PGLite unit test but exercises the real postgres-engine codepaths: sql.begin transactions, advisory locks on insertFact's entity-slug dedup window, unsafe('::vector') casts on findCandidateDuplicates ordering, addTakesBatch postgres-js unnest path. Happy path (4 facts → 1 take + all consolidated_into set), age-threshold skip, dry-run no-write. All 5 unit + 3 e2e tests pass. Closes the unit-only gap on the consolidate phase (was only PGLite-tested) and pins meta-cache invariants the dispatcher depends on. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 fix: thread auth + sourceId, JSON-shape every error envelope Three bugs surfaced during the full e2e sweep that all trace back to my v0.31 dispatch refactor (D12/eE1) silently dropping auth threading + non-OperationError exceptions emitting plain strings: 1. **HTTP MCP transport lost ctx.auth.** Refactoring serve-http.ts to call dispatchToolCall meant auth had to come through DispatchOpts, but the field didn't exist yet. Every HTTP whoami call returned `unknown_transport` because ctx.auth was undefined. Added `auth?: AuthInfo` to DispatchOpts, plumbed it through buildOperationContext, and updated serve-http.ts:816 to pass `auth: authInfo` alongside sourceId/takesHoldersAllowList. Pinned by sources-remote-mcp e2e `whoami reports oauth transport + sources_admin scope`. 2. **Non-OperationError exceptions emitted plain strings, not JSON.** The pre-v0.31 serve-http.ts always wrapped errors in JSON envelope `{error, message}`; my dispatch refactor missed the unknown-tool + uncaught-throw paths and emitted `Error: ${msg}` text content. Every caller that did `JSON.parse(content)` (sources-remote-mcp callMcp helper at line 104) crashed with `Unexpected identifier "Error"`. Both error paths in dispatchToolCall now return JSON-shaped content matching the OperationError pattern. 3. **Files→sources FK silently lost on rewound bootstrap path.** test/e2e/postgres-bootstrap.test.ts simulates a pre-v0.21 brain by `DROP TABLE IF EXISTS sources CASCADE` which removes files_source_id_fkey while leaving files.source_id intact. The v23 migration's `ALTER TABLE files ADD COLUMN IF NOT EXISTS source_id ... REFERENCES sources(id) ON DELETE CASCADE` is a no-op when the column exists, so the FK never came back on upgrade — and any sources-remove afterward stopped cascading to files. Added a defensive `IF NOT EXISTS files_source_id_fkey ... ALTER TABLE ADD CONSTRAINT` block inside v23's handler. Pinned by `multi-source — cascade delete covers every dependent row` after running postgres-bootstrap. Plus: src/core/preferences.ts now honors GBRAIN_HOME for `~/.gbrain/migrations/completed.jsonl`. Without this, the doctor exits-0 mechanical test inherits the developer machine's stale partial-migration ledger entries (0.21.0, 0.22.4, 0.28.0, 0.29.1 prior dev work) and surfaces them as the [FAIL] minions_migration check. GBRAIN_HOME-scoped tempdir per test now isolates this state cleanly. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 chore: scrub personal references from public artifacts Per the CLAUDE.md privacy rule on `Garry's Separation Test`, replace personally-coded references in v0.31 artifacts with neutral examples: - CHANGELOG.md v0.31 entry: rename "Garry's Separation Test" header to "The cross-session test" + drop the "topic-2659/topic-1941, 7 AM/2 PM, flying to Tokyo" narrative. - src/commands/migrations/v0_31_0.ts feature pitch: same scrub. - test/facts-separation-pglite.test.ts + test/e2e/facts-separation-postgres.test.ts: rename describe blocks; replace specific topic-NNNN session ids with session-A / session-B; replace personal sample fact with "sample event Tuesday". - src/core/facts/extract.ts extractor system prompt example slugs: people/sam-altman → people/alice-example; companies/anthropic → companies/acme. - src/core/entities/resolve.ts comment: Sam Altman → Alice Example. - All v0.31 test fixtures: people/sam → people/alice-example, Sam Altman → Alice Example, sam-the-cofounder → alice-the-cofounder. Test names referencing real-world entities replaced with neutral slugs. Pre-existing references to "Garry" elsewhere in CHANGELOG (v0.17, v0.19, v0.21+ entries) are untouched — that's a separate scope from this v0.31 ship. Plus: the truncate fix for the Bun-script-induced syntax error in test/e2e/mechanical.test.ts (cliEnv arrow function had ", 30_000)" tacked onto its closing brace by the bulk-add-timeouts script — repaired to a clean function definition). Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 fix(test): bump E2E phase-count assertions for 11-phase cycle Two E2E tests still asserted the v0.31 pre-merge 10-phase shape (consolidate inserted, but recompute_emotional_weight from v0.29 not yet absorbed). With master's v0.29 work merged in, the cycle is now 11 phases: lint → backlinks → sync → synthesize → extract → patterns → recompute_emotional_weight → consolidate → embed → orphans → purge. - test/e2e/cycle.test.ts: 10 → 11 - test/e2e/dream-cycle-eight-phase-pglite.test.ts: ALL_PHASES + dry-run order Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 fix(merge): close brace between v44 and v45 migration objects The v0.30.2 merge resolution stitched master's v40-v44 migrations onto HEAD's v45 (facts hot memory) migration but lost the closing `},` between v44 and v45. tsc caught it as TS1136 Property assignment expected at migrate.ts:2188. This is a one-line bracket fix; the rest of the merge resolution is correct and tests pass. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 fix: put_page cliHints + buildPlan v0.31.0 in skippedFuture Two unit-test failures surfaced after the v0.30.2 merge: 1. operations.ts: put_page had `cliHints: { name: 'put', positional: ['stdin'] }` from earlier v0.31 development. The parity test enforces that every name in `positional` is a real param. Restored master's correct shape: `{ name: 'put', positional: ['slug'], stdin: 'content' }`. 2. test/apply-migrations.test.ts: the H9 regression tests pin the exact skippedFuture list. Adding v0.31.0 to the registry meant the list grew by one. Updated both `expect(...).toEqual([...])` assertions. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> * v0.31 docs: clarify consolidate is 11th phase + regen llms-full.txt CHANGELOG.md narrative said "new 10th phase consolidate"; with v0.29's recompute_emotional_weight already on master, consolidate is the 11th phase (between recompute and embed). Schema migration is v45, not v40, after the merge resolution renumbered it to clear master's v40-v44. llms-full.txt regenerated to reflect the README's 11-phase dream-cycle phrasing (the build-llms test enforces commit-time parity). Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com> --------- Co-authored-by: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
557 lines
20 KiB
TypeScript
557 lines
20 KiB
TypeScript
/**
|
|
* E2E Sync Tests — Tier 1 (no API keys required)
|
|
*
|
|
* Tests the full git-to-DB sync pipeline: create a git repo, commit
|
|
* markdown files, run gbrain sync, verify pages appear in the database.
|
|
* Covers first sync, incremental add/modify/delete, and the critical
|
|
* "edit → sync → search returns corrected text" flow.
|
|
*
|
|
* Run: DATABASE_URL=... bun test test/e2e/sync.test.ts
|
|
*/
|
|
|
|
import { describe, test, expect, beforeAll, afterAll } from 'bun:test';
|
|
import { mkdtempSync, writeFileSync, rmSync, mkdirSync, unlinkSync, existsSync, readFileSync } from 'fs';
|
|
import { join } from 'path';
|
|
import { execSync } from 'child_process';
|
|
import { tmpdir, homedir } from 'os';
|
|
import {
|
|
hasDatabase, setupDB, teardownDB, getEngine,
|
|
} from './helpers.ts';
|
|
|
|
const skip = !hasDatabase();
|
|
const describeE2E = skip ? describe.skip : describe;
|
|
|
|
if (skip) {
|
|
console.log('Skipping E2E sync tests (DATABASE_URL not set)');
|
|
}
|
|
|
|
/** Create a temp git repo with initial markdown files */
|
|
function createTestRepo(): string {
|
|
const dir = mkdtempSync(join(tmpdir(), 'gbrain-sync-e2e-'));
|
|
execSync('git init', { cwd: dir, stdio: 'pipe' });
|
|
execSync('git config user.email "test@test.com"', { cwd: dir, stdio: 'pipe' });
|
|
execSync('git config user.name "Test"', { cwd: dir, stdio: 'pipe' });
|
|
|
|
// Create initial structure
|
|
mkdirSync(join(dir, 'people'), { recursive: true });
|
|
mkdirSync(join(dir, 'concepts'), { recursive: true });
|
|
|
|
writeFileSync(join(dir, 'people/alice.md'), [
|
|
'---',
|
|
'type: person',
|
|
'title: Alice Smith',
|
|
'tags: [engineer, frontend]',
|
|
'---',
|
|
'',
|
|
'Alice is a frontend engineer at Acme Corp.',
|
|
'',
|
|
'---',
|
|
'',
|
|
'- 2026-01-15: Joined Acme Corp',
|
|
].join('\n'));
|
|
|
|
writeFileSync(join(dir, 'concepts/testing.md'), [
|
|
'---',
|
|
'type: concept',
|
|
'title: Testing Philosophy',
|
|
'tags: [engineering]',
|
|
'---',
|
|
'',
|
|
'Every untested path is a path where bugs hide.',
|
|
].join('\n'));
|
|
|
|
// Initial commit
|
|
execSync('git add -A && git commit -m "initial commit"', { cwd: dir, stdio: 'pipe' });
|
|
|
|
return dir;
|
|
}
|
|
|
|
function gitCommit(repoPath: string, message: string) {
|
|
execSync(`git add -A && git commit -m "${message}"`, { cwd: repoPath, stdio: 'pipe' });
|
|
}
|
|
|
|
describeE2E('E2E: Git-to-DB Sync Pipeline', () => {
|
|
let repoPath: string;
|
|
|
|
beforeAll(async () => {
|
|
await setupDB();
|
|
repoPath = createTestRepo();
|
|
}, 30_000);
|
|
|
|
afterAll(async () => {
|
|
await teardownDB();
|
|
if (repoPath) rmSync(repoPath, { recursive: true, force: true });
|
|
});
|
|
|
|
test('first sync imports all pages from git repo', async () => {
|
|
const { performSync } = await import('../../src/commands/sync.ts');
|
|
const engine = getEngine();
|
|
|
|
const result = await performSync(engine, {
|
|
repoPath,
|
|
noPull: true,
|
|
noEmbed: true,
|
|
});
|
|
|
|
expect(result.status).toBe('first_sync');
|
|
// performFullSync delegates to runImport which doesn't populate pagesAffected
|
|
// Verify pages exist in DB directly instead
|
|
const alice = await engine.getPage('people/alice');
|
|
expect(alice).not.toBeNull();
|
|
expect(alice!.title).toBe('Alice Smith');
|
|
|
|
const testing = await engine.getPage('concepts/testing');
|
|
expect(testing).not.toBeNull();
|
|
expect(testing!.title).toBe('Testing Philosophy');
|
|
});
|
|
|
|
test('second sync with no changes returns up_to_date', async () => {
|
|
const { performSync } = await import('../../src/commands/sync.ts');
|
|
const engine = getEngine();
|
|
|
|
const result = await performSync(engine, {
|
|
repoPath,
|
|
noPull: true,
|
|
noEmbed: true,
|
|
});
|
|
|
|
expect(result.status).toBe('up_to_date');
|
|
expect(result.added).toBe(0);
|
|
expect(result.modified).toBe(0);
|
|
expect(result.deleted).toBe(0);
|
|
});
|
|
|
|
test('incremental sync picks up new files', async () => {
|
|
const { performSync } = await import('../../src/commands/sync.ts');
|
|
const engine = getEngine();
|
|
|
|
// Add a new file
|
|
writeFileSync(join(repoPath, 'people/bob.md'), [
|
|
'---',
|
|
'type: person',
|
|
'title: Bob Jones',
|
|
'tags: [designer]',
|
|
'---',
|
|
'',
|
|
'Bob is a product designer who loves typography.',
|
|
].join('\n'));
|
|
gitCommit(repoPath, 'add bob');
|
|
|
|
const result = await performSync(engine, {
|
|
repoPath,
|
|
noPull: true,
|
|
noEmbed: true,
|
|
});
|
|
|
|
expect(result.status).toBe('synced');
|
|
expect(result.added).toBe(1);
|
|
expect(result.pagesAffected).toContain('people/bob');
|
|
|
|
const bob = await engine.getPage('people/bob');
|
|
expect(bob).not.toBeNull();
|
|
expect(bob!.title).toBe('Bob Jones');
|
|
expect(bob!.compiled_truth).toContain('typography');
|
|
});
|
|
|
|
test('incremental sync picks up modifications — corrected text appears', async () => {
|
|
const { performSync } = await import('../../src/commands/sync.ts');
|
|
const engine = getEngine();
|
|
|
|
// Modify alice's page — this is the critical "correction" test
|
|
writeFileSync(join(repoPath, 'people/alice.md'), [
|
|
'---',
|
|
'type: person',
|
|
'title: Alice Smith',
|
|
'tags: [engineer, frontend]',
|
|
'---',
|
|
'',
|
|
'Alice is a staff frontend engineer at Acme Corp, leading the design system team.',
|
|
'',
|
|
'---',
|
|
'',
|
|
'- 2026-04-01: Promoted to staff engineer',
|
|
'- 2026-01-15: Joined Acme Corp',
|
|
].join('\n'));
|
|
gitCommit(repoPath, 'update alice - promotion');
|
|
|
|
const result = await performSync(engine, {
|
|
repoPath,
|
|
noPull: true,
|
|
noEmbed: true,
|
|
});
|
|
|
|
expect(result.status).toBe('synced');
|
|
expect(result.modified).toBe(1);
|
|
expect(result.pagesAffected).toContain('people/alice');
|
|
|
|
// THE CRITICAL CHECK: corrected text appears in the DB
|
|
const alice = await engine.getPage('people/alice');
|
|
expect(alice!.compiled_truth).toContain('staff frontend engineer');
|
|
expect(alice!.compiled_truth).toContain('design system team');
|
|
// Old text should be replaced, not appended
|
|
expect(alice!.compiled_truth).not.toBe('Alice is a frontend engineer at Acme Corp.');
|
|
});
|
|
|
|
test('keyword search finds corrected text after sync', async () => {
|
|
const engine = getEngine();
|
|
|
|
// Search for the new text
|
|
const results = await engine.searchKeyword('design system team');
|
|
expect(results.length).toBeGreaterThanOrEqual(1);
|
|
|
|
const aliceResult = results.find((r: any) => r.slug === 'people/alice');
|
|
expect(aliceResult).toBeDefined();
|
|
});
|
|
|
|
test('incremental sync handles deletes', async () => {
|
|
const { performSync } = await import('../../src/commands/sync.ts');
|
|
const engine = getEngine();
|
|
|
|
// Delete bob's page
|
|
unlinkSync(join(repoPath, 'people/bob.md'));
|
|
gitCommit(repoPath, 'remove bob');
|
|
|
|
const result = await performSync(engine, {
|
|
repoPath,
|
|
noPull: true,
|
|
noEmbed: true,
|
|
});
|
|
|
|
expect(result.status).toBe('synced');
|
|
expect(result.deleted).toBe(1);
|
|
|
|
const bob = await engine.getPage('people/bob');
|
|
expect(bob).toBeNull();
|
|
});
|
|
|
|
test('sync skips non-syncable files (README, hidden, .raw)', async () => {
|
|
const { performSync } = await import('../../src/commands/sync.ts');
|
|
const engine = getEngine();
|
|
|
|
// Add files that should be excluded
|
|
writeFileSync(join(repoPath, 'README.md'), '# Brain Repo\nThis is the readme.');
|
|
mkdirSync(join(repoPath, '.raw'), { recursive: true });
|
|
writeFileSync(join(repoPath, '.raw/data.md'), '---\ntitle: Raw\n---\nRaw data.');
|
|
mkdirSync(join(repoPath, 'ops'), { recursive: true });
|
|
writeFileSync(join(repoPath, 'ops/deploy.md'), '---\ntitle: Deploy\n---\nOps stuff.');
|
|
gitCommit(repoPath, 'add non-syncable files');
|
|
|
|
const result = await performSync(engine, {
|
|
repoPath,
|
|
noPull: true,
|
|
noEmbed: true,
|
|
});
|
|
|
|
// These should not create pages
|
|
const readme = await engine.getPage('README');
|
|
expect(readme).toBeNull();
|
|
|
|
const raw = await engine.getPage('.raw/data');
|
|
expect(raw).toBeNull();
|
|
|
|
const ops = await engine.getPage('ops/deploy');
|
|
expect(ops).toBeNull();
|
|
});
|
|
|
|
test('sync stores last_commit and last_run in config', async () => {
|
|
const engine = getEngine();
|
|
|
|
const lastCommit = await engine.getConfig('sync.last_commit');
|
|
const lastRun = await engine.getConfig('sync.last_run');
|
|
const repoPathConfig = await engine.getConfig('sync.repo_path');
|
|
|
|
expect(lastCommit).toBeTruthy();
|
|
expect(lastCommit!.length).toBe(40); // full SHA
|
|
expect(lastRun).toBeTruthy();
|
|
expect(repoPathConfig).toBe(repoPath);
|
|
});
|
|
|
|
test('sync logs to ingest_log', async () => {
|
|
const engine = getEngine();
|
|
|
|
const logs = await engine.getIngestLog();
|
|
const syncLogs = logs.filter((l: any) => l.source_type === 'git_sync');
|
|
|
|
expect(syncLogs.length).toBeGreaterThanOrEqual(1);
|
|
expect(syncLogs[0].source_ref).toContain(repoPath);
|
|
});
|
|
|
|
test('--full reimports everything regardless of last_commit', async () => {
|
|
const { performSync } = await import('../../src/commands/sync.ts');
|
|
const engine = getEngine();
|
|
|
|
const result = await performSync(engine, {
|
|
repoPath,
|
|
noPull: true,
|
|
noEmbed: true,
|
|
full: true,
|
|
});
|
|
|
|
expect(result.status).toBe('first_sync');
|
|
// performFullSync delegates to runImport — verify pages exist instead
|
|
const alice = await engine.getPage('people/alice');
|
|
expect(alice).not.toBeNull();
|
|
const testing = await engine.getPage('concepts/testing');
|
|
expect(testing).not.toBeNull();
|
|
});
|
|
|
|
test('dry-run shows changes without applying them', async () => {
|
|
const { performSync } = await import('../../src/commands/sync.ts');
|
|
const engine = getEngine();
|
|
|
|
// Add a new file
|
|
writeFileSync(join(repoPath, 'concepts/dry-run-test.md'), [
|
|
'---',
|
|
'type: concept',
|
|
'title: Dry Run Test',
|
|
'---',
|
|
'',
|
|
'This should not be imported.',
|
|
].join('\n'));
|
|
gitCommit(repoPath, 'add dry run test');
|
|
|
|
const result = await performSync(engine, {
|
|
repoPath,
|
|
noPull: true,
|
|
noEmbed: true,
|
|
dryRun: true,
|
|
});
|
|
|
|
expect(result.status).toBe('dry_run');
|
|
expect(result.added).toBe(1);
|
|
|
|
// Page should NOT exist in DB
|
|
const page = await engine.getPage('concepts/dry-run-test');
|
|
expect(page).toBeNull();
|
|
|
|
// Clean up: do a real sync so the commit is consumed
|
|
await performSync(engine, {
|
|
repoPath,
|
|
noPull: true,
|
|
noEmbed: true,
|
|
});
|
|
});
|
|
|
|
test('files with spaces in names get slugified slugs', async () => {
|
|
const { performSync } = await import('../../src/commands/sync.ts');
|
|
const engine = getEngine();
|
|
|
|
// Add a file with spaces (Apple Notes style)
|
|
mkdirSync(join(repoPath, 'Apple Notes'), { recursive: true });
|
|
writeFileSync(join(repoPath, 'Apple Notes/2017-05-03 ohmygreen.md'), [
|
|
'---',
|
|
'title: Ohmygreen Notes',
|
|
'---',
|
|
'',
|
|
'Notes about ohmygreen lunch service.',
|
|
].join('\n'));
|
|
gitCommit(repoPath, 'add apple notes file with spaces');
|
|
|
|
const result = await performSync(engine, {
|
|
repoPath,
|
|
noPull: true,
|
|
noEmbed: true,
|
|
});
|
|
|
|
expect(result.status).toBe('synced');
|
|
expect(result.added).toBe(1);
|
|
|
|
// Slug should be slugified (lowercase, spaces → hyphens)
|
|
const page = await engine.getPage('apple-notes/2017-05-03-ohmygreen');
|
|
expect(page).not.toBeNull();
|
|
expect(page!.title).toBe('Ohmygreen Notes');
|
|
|
|
// Original space-based slug should NOT exist
|
|
const rawSlug = await engine.getPage('Apple Notes/2017-05-03 ohmygreen');
|
|
expect(rawSlug).toBeNull();
|
|
});
|
|
|
|
test('incremental sync adds file with special characters', async () => {
|
|
const { performSync } = await import('../../src/commands/sync.ts');
|
|
const engine = getEngine();
|
|
|
|
// Add a file with parens and special chars
|
|
writeFileSync(join(repoPath, 'Apple Notes/meeting notes (draft).md'), [
|
|
'---',
|
|
'title: Draft Meeting Notes',
|
|
'---',
|
|
'',
|
|
'Some draft notes from the meeting.',
|
|
].join('\n'));
|
|
gitCommit(repoPath, 'add file with parens');
|
|
|
|
const result = await performSync(engine, {
|
|
repoPath,
|
|
noPull: true,
|
|
noEmbed: true,
|
|
});
|
|
|
|
expect(result.status).toBe('synced');
|
|
|
|
// Slug should have parens stripped, spaces → hyphens
|
|
const page = await engine.getPage('apple-notes/meeting-notes-draft');
|
|
expect(page).not.toBeNull();
|
|
expect(page!.title).toBe('Draft Meeting Notes');
|
|
});
|
|
});
|
|
|
|
/**
|
|
* E2E: --skip-failed loop with structured error code summary.
|
|
*
|
|
* Closes the v0.22.12 ship-blocker gap from issue #500 — the whole code path
|
|
* (record → classify → block → skip → doctor render → second cycle) had only
|
|
* mocked-JSONL unit coverage. This is the integration test that proves the
|
|
* chain holds together with a real Postgres engine, real git history, and
|
|
* real frontmatter validation.
|
|
*
|
|
* Owns its own repo + sync-failures.jsonl lifecycle so it can't leak state
|
|
* into the shared describeE2E above. Saves and restores the user's real
|
|
* ~/.gbrain/sync-failures.jsonl so running E2E on a developer machine
|
|
* doesn't trash their local sync state.
|
|
*/
|
|
describeE2E('E2E: sync --skip-failed structured summary loop (v0.22.12, issue #500)', () => {
|
|
let repoPath: string;
|
|
const realFailuresPath = join(homedir(), '.gbrain', 'sync-failures.jsonl');
|
|
let savedFailuresContent: string | null = null;
|
|
|
|
beforeAll(async () => {
|
|
await setupDB();
|
|
|
|
// Save+clear the real ~/.gbrain/sync-failures.jsonl so the test starts from
|
|
// a known-empty state. Restored in afterAll. This file is per-machine, NOT
|
|
// per-repo, so we have to be defensive about a developer running this
|
|
// suite on their actual brain machine.
|
|
if (existsSync(realFailuresPath)) {
|
|
savedFailuresContent = readFileSync(realFailuresPath, 'utf-8');
|
|
unlinkSync(realFailuresPath);
|
|
}
|
|
|
|
// Fresh git repo with one valid file. Mirrors createTestRepo above but
|
|
// scoped to this describe block.
|
|
repoPath = mkdtempSync(join(tmpdir(), 'gbrain-skipfailed-e2e-'));
|
|
execSync('git init', { cwd: repoPath, stdio: 'pipe' });
|
|
execSync('git config user.email "test@test.com"', { cwd: repoPath, stdio: 'pipe' });
|
|
execSync('git config user.name "Test"', { cwd: repoPath, stdio: 'pipe' });
|
|
mkdirSync(join(repoPath, 'people'), { recursive: true });
|
|
writeFileSync(join(repoPath, 'people/alice.md'), [
|
|
'---', 'type: person', 'title: Alice', '---', '', 'Body.',
|
|
].join('\n'));
|
|
execSync('git add -A && git commit -m "initial"', { cwd: repoPath, stdio: 'pipe' });
|
|
}, 30_000);
|
|
|
|
afterAll(async () => {
|
|
await teardownDB();
|
|
if (repoPath) rmSync(repoPath, { recursive: true, force: true });
|
|
|
|
// Restore the user's real sync-failures.jsonl, if any.
|
|
if (savedFailuresContent !== null) {
|
|
mkdirSync(join(homedir(), '.gbrain'), { recursive: true });
|
|
writeFileSync(realFailuresPath, savedFailuresContent);
|
|
} else if (existsSync(realFailuresPath)) {
|
|
// Test wrote one but there was none before. Clean up.
|
|
unlinkSync(realFailuresPath);
|
|
}
|
|
});
|
|
|
|
test('full --skip-failed loop: blocks on bad file, skip advances bookmark, doctor shows code breakdown', async () => {
|
|
const { performSync } = await import('../../src/commands/sync.ts');
|
|
const { loadSyncFailures, summarizeFailuresByCode } = await import('../../src/core/sync.ts');
|
|
const engine = getEngine();
|
|
|
|
// Step 1: First sync of the clean repo — should succeed.
|
|
let result = await performSync(engine, { repoPath, noPull: true, noEmbed: true });
|
|
expect(result.status).toBe('first_sync');
|
|
const firstCommit = await engine.getConfig('sync.last_commit');
|
|
expect(firstCommit).toBeTruthy();
|
|
|
|
// Step 2: Add a broken file — frontmatter slug doesn't match path-derived slug.
|
|
// The file path is people/bob.md so the path-derived slug is "people/bob",
|
|
// but we declare slug: "wrong-slug" in frontmatter. import-file.ts:368-377
|
|
// raises "Frontmatter slug ... does not match path-derived slug ..." which
|
|
// classifier hits as SLUG_MISMATCH.
|
|
writeFileSync(join(repoPath, 'people/bob.md'), [
|
|
'---', 'type: person', 'title: Bob', 'slug: wrong-slug', '---', '', 'Body.',
|
|
].join('\n'));
|
|
execSync('git add -A && git commit -m "add broken bob"', { cwd: repoPath, stdio: 'pipe' });
|
|
|
|
// Step 3: Sync should block. Bookmark must NOT advance.
|
|
result = await performSync(engine, { repoPath, noPull: true, noEmbed: true });
|
|
expect(result.status).toBe('blocked_by_failures');
|
|
const afterBlockedCommit = await engine.getConfig('sync.last_commit');
|
|
expect(afterBlockedCommit).toBe(firstCommit); // bookmark stuck at the pre-broken commit
|
|
|
|
// JSONL has one unacked entry with code SLUG_MISMATCH.
|
|
let failures = loadSyncFailures();
|
|
expect(failures.length).toBe(1);
|
|
expect(failures[0].code).toBe('SLUG_MISMATCH');
|
|
expect(failures[0].acknowledged).toBeFalsy();
|
|
// Group summary aggregates correctly across the unacked set.
|
|
expect(summarizeFailuresByCode(failures)).toEqual([{ code: 'SLUG_MISMATCH', count: 1 }]);
|
|
|
|
// Step 4: Run with skipFailed — bookmark advances, entry gets acked.
|
|
result = await performSync(engine, { repoPath, noPull: true, noEmbed: true, skipFailed: true });
|
|
expect(result.status).toBe('synced');
|
|
const afterSkipCommit = await engine.getConfig('sync.last_commit');
|
|
expect(afterSkipCommit).not.toBe(firstCommit); // bookmark moved past the broken commit
|
|
failures = loadSyncFailures();
|
|
expect(failures.length).toBe(1);
|
|
expect(failures[0].acknowledged).toBe(true);
|
|
expect(typeof failures[0].acknowledged_at).toBe('string');
|
|
|
|
// Step 5: Verify what doctor would render for the historical entry.
|
|
// We call the same primitives doctor's `sync_failures` check uses
|
|
// (src/commands/doctor.ts:252-275) — loadSyncFailures + summarizeFailuresByCode —
|
|
// and assert the rendering string. Directly invoking runDoctor() here is a CLI
|
|
// entrypoint with stdout/exit side effects that would truncate this test mid-flow.
|
|
{
|
|
const all = loadSyncFailures();
|
|
const ackedSummary = summarizeFailuresByCode(all);
|
|
const ackedBreakdown = ackedSummary.map(s => `${s.code}=${s.count}`).join(', ');
|
|
// This is the literal string interpolation doctor.ts:271-274 produces.
|
|
const doctorMessage = `${all.length} historical sync failure(s), all acknowledged [${ackedBreakdown}].`;
|
|
expect(doctorMessage).toContain('SLUG_MISMATCH=1');
|
|
expect(doctorMessage).toContain('1 historical');
|
|
}
|
|
|
|
// Step 6: Add a second broken file — this one with a different failure code
|
|
// (also SLUG_MISMATCH but on a different file) so the JSONL has 2 entries
|
|
// with DIFFERENT paths but the same code. This proves both: per-file dedup
|
|
// honors path identity, and summary aggregation sums across files.
|
|
//
|
|
// We'd ideally test a different code class here, but the sync path uses
|
|
// parseMarkdown WITHOUT {validate:true}, so the markdown.ts validation
|
|
// codes (MISSING_OPEN/CLOSE, NESTED_QUOTES, EMPTY_FRONTMATTER, NULL_BYTES)
|
|
// don't naturally surface — they'd need {validate:true} plumbed in. That
|
|
// plumbing is the v0.22.13+ follow-up. For v0.22.12, two SLUG_MISMATCH
|
|
// entries from different files still proves the dedup + aggregation chain.
|
|
writeFileSync(join(repoPath, 'people/carol.md'), [
|
|
'---', 'type: person', 'title: Carol', 'slug: also-wrong-slug', '---', '', 'Body.',
|
|
].join('\n'));
|
|
execSync('git add -A && git commit -m "add carol with bad slug"', { cwd: repoPath, stdio: 'pipe' });
|
|
|
|
// Step 7: Sync blocks again on the new failure. Old entry stays acked.
|
|
result = await performSync(engine, { repoPath, noPull: true, noEmbed: true });
|
|
expect(result.status).toBe('blocked_by_failures');
|
|
failures = loadSyncFailures();
|
|
expect(failures.length).toBe(2);
|
|
const acked = failures.filter(f => f.acknowledged);
|
|
const unacked = failures.filter(f => !f.acknowledged);
|
|
expect(acked.length).toBe(1);
|
|
expect(acked[0].code).toBe('SLUG_MISMATCH');
|
|
expect(acked[0].path).toContain('bob');
|
|
expect(unacked.length).toBe(1);
|
|
expect(unacked[0].code).toBe('SLUG_MISMATCH');
|
|
expect(unacked[0].path).toContain('carol');
|
|
|
|
// Step 8: Skip again — both entries acked, summary aggregates the count.
|
|
result = await performSync(engine, { repoPath, noPull: true, noEmbed: true, skipFailed: true });
|
|
expect(result.status).toBe('synced');
|
|
failures = loadSyncFailures();
|
|
expect(failures.length).toBe(2);
|
|
expect(failures.every(f => f.acknowledged)).toBe(true);
|
|
|
|
const finalSummary = summarizeFailuresByCode(failures);
|
|
expect(finalSummary).toEqual([{ code: 'SLUG_MISMATCH', count: 2 }]);
|
|
});
|
|
});
|