Docs Consolidation Map

Naming ruling (Aug 27): the quality engine is spelled Cerberus in all prose and display text (the three-headed gate-guard — matches the repo's own description, "the three-gate quality engine for Clara"). The GitHub repo slug is the historical misspelling cerebrus and STAYS that way — lowercase cerebrus in real identifiers (repo path PropFlow-Technologies/cerebrus, scripts/sync-cerebrus.sh, vendored dirs, URLs) must never be "corrected"; editing them breaks things. Only capitalized prose converges. Prose sweep applied to this corpus Aug 27.

Decisions consolidated into the Clara-as-a-Coworker tracker, Phase 4 (single list) — answers there count.

435 mined agent-fleet pages across 20 clusters, reconciled down to a proposed 56 canonical pages + 1 index. What to read, what to archive, every still-open decision once, and the rule that keeps this from happening again.

Generated 2026-08-26 · covers all 19 reconciled clusters (a 20th, 2 unrelated pages, needs manual triage) · decisions · weekly-priorities-2026-08-24

1 · The Numbers

Headline: 435 pages → a proposed 56 canonical pages + 1 index = 57 living destinations. 156 pages already have an explicit archive+redirect target in this pass. 86 places found the same issue solved two-or-more different ways. 56 outright contradictions (a page says X, reality or a newer page says not-X). 208 individual times a decision question was raised across pages, deduplicating to 136 distinct decisions — of which 34 already have a Fede ruling or shipped fix (some pages just haven't been told yet) and ~102 are genuinely still waiting on Fede.

#ClusterPagesCanonical (target)Archived this passOverlap issues foundContradictions foundDecisions (deduped)Decision askings (raw, pre-dedup)Already decided/shipped
1Promises, Escalation, HITL & Transfer Outcomes36387410173
2Agent Fleet Ops, Fable Decision Queue & CI/Infra1703791647223
3Evals, Cerberus & Quality-Grading Infrastructure3338558132
4Voice Agent Quality & Latency1332439122
5Turnover & Vendor Coordination30388214211
6Collections & Delinquency16310639173 (+1 partial)
7Renewals & Rent Pricing73122890
8Inbound Email Pipeline93534371
9Knowledge Base & Docs Infrastructure1133437121
10Identity, Dedup, Households & Onboarding1533429104
11Lease Answers, Application & Post-Approval Flow14336511163
12AppFolio / PMS Integration43122221
13Sales & Business Development29310336112
14UI/UX Redesign, Dashboards & Design System163543774
15Clara Coworker Vision & Cross-Channel Architecture93311570 (+1 partial)
16Business Operations & Engineering Vision53232772
17Fair Housing & Compliance53232461
18App & Clara Response Latency32133660
19Tour Scheduling & Booking Consistency83223461
20Unrelated (not reconciled)2
TOTAL43556156865613620834 (+2 partial)

Notes on methodology (so nobody mistakes an estimate for a verified fact): - Pages / canonical / archived / overlaps / contradictions are exact counts of the items listed in the reconciliation records supplied for this pass — not estimates. - Decisions deduped vs raw: "raw" = how many times a decision question appears attached to a page across the corpus (summing asked_on_pages across every decision); "deduped" = the number of genuinely distinct decisions once duplicates are merged. The gap is concentrated in a handful of repeat offenders (the pre-send honesty-check question asked 5 times, the SMOKE_PASSWORD question asked 3 times, the fleet-health restart bug escalated 4 times, etc.) rather than spread evenly — most decisions were only asked once. - "Already decided/shipped" is counted only where the record explicitly says a ruling exists or the fix is merged/live (an already_decided field with real content). Several more decisions are effectively moot because something else already shipped in their place (e.g. the sales company-card A/B/C/D panel question, the tour-decider consolidation question) but without a formal Fede sign-off on record — those are called out individually in Section 3, not folded into this count, to avoid overstating how much is actually closed. - The 170-page Agent Fleet Ops cluster is the outlier: of its 170 pages, ~60 are empty, auto-generated "nothing's waiting on you" template snapshots (blocked-<hash>.html with no real content) and another ~19 are dated status/redirect stubs. That's why its archive count (79) is disproportionately large relative to its overlap/decision count — most of the cleanup there is deleting noise, not resolving conflicts.

---

2 · Consolidation Map

Every archived page below should 301/redirect (or, on this static site, get a one-line "moved to ↓" stub) to the named target. Reasons are compressed from the mining pass; see doc-mining-records.json for the full text.

1. Promises, Escalation, HITL & Transfer Outcomes (36 pages)

Canonical: promise-systematic-pattern-2026-08-26, escalation-architecture-2026-08-21, honesty-layer-deep-inspection-2026-08-16

ArchiveRedirect toWhy
grounded-claims-architecturehonesty-layer-deep-inspection-2026-08-16Own text says it was folded into the deep-inspection doc
hitl-unify-2026-08-26clara-coworker-map-2026-08-19#phase-4Empty redirect stub, no independent content
promise-ownership-and-due-time-2026-08-26clara-coworker-map-2026-08-19#phase-4Same empty stub, same target as above — the two should be one page
escalation-reply-incident-2026-08-03escalation-architecture-2026-08-21Root incident is historical; the Aug 21 page is the up-to-date status
escalation-architecture-decision-2026-08-04escalation-architecture-2026-08-21Its "shipped clean" claim is exactly what Aug 21 found still broken in one real case
escalation-journey-rethinkescalation-architecture-2026-08-21Superseded by the Aug 21 status page; unresolved bigger asks move to promise-systematic-pattern-2026-08-26's Decision 6
escalation-email-wording-2026-08-14escalation-decisions-2026-08-16Style rulings duplicate what's already decided there; only the logo question carries forward
2. Agent Fleet Ops, Fable Decision Queue & CI/Infra (170 pages)

Canonical: decisions, weekly-priorities-2026-08-24, pipeline-delivery

ArchiveRedirect toWhy
agentflowagentflow-in-appExplicitly retired stub; functionality moved to the in-app /agents page
pipeline-e2e-handoffpipeline-deliveryDead stub, content already folded in
overnight-lanes-2026-08-23weekly-priorities-2026-08-24Dead stub, consolidated into the current rollup
blocked-245cddd5blocked-1af80f47Literal duplicate of the same PR #5675 decision (the page itself says "adopted from" this one)
ai-code-review-lead-timeai-code-review-evidenceIts headline worry (AI review gets noisier on big diffs) was tested and refuted
smith-pr-drive-plan, smith-pr-webhook-explained, smith-pr-drive-plan-v2, smith-pr-drive-plan-v3smith-pr-drive-plan-v4Superseded iterations of the same architecture; v4 is current and shipped
a2a-slack-research-2026-07-29, agent-audit-2026-07-29, smith-review-decisions, smith-overnight-ops-2026-07-30agent-smith-issue-tracker / smith-review-decisions-2026-08-25One-time audits absorbed into the live tracker / newer root-cause pass
agent-fleet-learningsagent-operating-manualLessons-learned background; the manual is the actionable runbook
fleet-status, fleet-verdictsfleet-drive-methodologyPoint-in-time snapshots superseded by the distilled methodology
operators-on-temporalnine-doors-one-storeNarrower unbuilt proposal superseded by the deeper, newer audit
ci-testing-strategyci-two-tier-test-modelOriginal research; the two-tier page self-corrects a key number and is more current
held-prs-2026-08-08held-pr-triageNarrower, earlier scope fully contained in the later 18-PR triage
ten-open-six-questions-0803, eight-open-five-questions, parked-decisions, built-three-one-question, needs-you-2026-08-03, blocked-9a7a3194weekly-priorities-2026-08-24Dated batch-decision snapshots superseded by the current aggregator
3. Evals, Cerberus & Quality-Grading Infrastructure (33 pages)

Canonical: eval-testing-roadmap, quality-loop-guide, quality-desk-checkin-2026-08-26

ArchiveRedirect toWhy
eval-architecture-inspection-2026-08eval-testing-roadmapNames the roadmap as its own successor
eval-inventory-2026-08-20eval-testing-roadmapIts 10-mechanism census is now backing detail for roadmap-tracked gaps
quality-system-plan-2026-08eval-testing-roadmapExplicitly superseded; its 5-week plan is spread across roadmap rows now
eval-fire-alarm-podcasteval-architecture-inspection-2026-08Pure narration of the same findings, no new facts
eval-floor-decisionseval-testing-roadmapIts 9 threshold conflicts are the same item roadmap tracks as one backlog line
blocked-d451bf5dblocked-19f1add2Word-for-word the same parked decision, re-surfaced 2 weeks later
grading-desk-qa-2026-08-19quality-desk-checkin-2026-08-26One-day QA snapshot; the later check-in has current numbers
4. Voice Agent Quality & Latency (13 pages)

Canonical: voice-experiments-report, voice-handoff-latency, voice-latency-deep-dive-2026-08

ArchiveRedirect toWhy
voice-model-bakeoff-2026-08-21voice-experiments-reportSame question, same "keep current model" answer — newer page has the full data and final decision
5. Turnover & Vendor Coordination (30 pages)

Canonical: adr-0116-vendor-job-reference, turnover-architecture-2026-08-01, vendor-identity-architecture-2026-08-12

ArchiveRedirect toWhy
purchase-order-pivotadr-0116-vendor-job-referenceADR page names itself the finished successor to this pivot brief
turnover-projection-before-after-2026-07-30adr-proposed-turnover-money-decisions-2026-07-30Same shipped fix, same PRs — numbers should become a results section on the ADR
turnover-overnight-report-2026-07-30adr-proposed-turnover-money-decisions-2026-07-30Same fix from a status-update angle, bundled with two unrelated items
turnover-before-after-replay-2026-07-30turnover-harness-rca-2026-07-31Duplicates the RCA's own named finding
vendor-front-gate-reconvendor-calling-readiness-2026-07-31Keep both live, but readiness doc should link here instead of restating the root cause
overnight-vendor-ship-2026-07-29phase4-scheduling-decisionsIts two Phase-4 decisions are pointers to the real decision brief
turnover-brief-scorecard-2026-07-28phase4-scheduling-decisionsSame duplicate decision callout; the AI-disclosure item it lists is already resolved
6. Collections & Delinquency (16 pages)

Canonical: collections-three-phases, collections-phase-tracker, colorado-collections-law

ArchiveRedirect toWhy
blocked-45997cd7collections-phase-tracker"Define done" question overtaken — both referenced PRs merged, work continued into the standing tracker
blocked-cda3604dcollections-three-phasesIts proposed finish line (the phases page) shipped and is now canonical
blocked-5d964f5fcollections-phase-trackerPR #5930 has shipped and iterated since; the approve-to-unblock-tests question is moot
collections-pilot-briefcollections-three-phasesFounding brief's numbers/phase plan superseded by measured data
collections-readiness-deckcollections-three-phasesExplicitly called stale by the newer page; 5 of 6 critical findings already fixed
collections-phase-1-podcastcollections-phase-1-walkthroughByte-for-byte the same content, wrapped as a podcast script instead of a checklist
collections-header-pr5886collections-phase-trackerSmall screenshot-evidence log; tracker already logs this PR's outcome
pr-5495-evidencecollections-policy-uiShipped-evidence snapshot for a rollout the design doc is the durable reference for
pr-5994-evidencecollections-three-phasesShipped-evidence snapshot superseded as panel iterations continue
7. Renewals & Rent Pricing (7 pages)

Canonical: renewal-thread-escalation-findings, renewal-outreach-incident-2026-07-31, camellia-market-rent-drop-2026-07-30

ArchiveRedirect toWhy
8. Inbound Email Pipeline (9 pages)

Canonical: email-architecture-2026-08-14, email-silent-drop-incident, email-intent-contract-2026-08-17

ArchiveRedirect toWhy
email-death-path-auditemail-silent-drop-incidentStates outright it's the evidence base for that page's Decision 5 — a companion, not independent
inbound-email-redesignemail-architecture-2026-08-14Original design brief; still-open decisions should copy into the architecture doc's tracker
inbound-email-deploymentemail-architecture-2026-08-14Proof-of-work status report, not where current decisions should live
email-deliverability-2026-08-16email-architecture-2026-08-14A concrete event inside the deliverability program the architecture doc already owns
9. Knowledge Base & Docs Infrastructure (11 pages)

Canonical: knowledge-base-redesign-2026-08-26, property-knowledgebase-2026-08-08, docs-quickstart

10. Identity, Dedup, Households & Onboarding (15 pages)

Canonical: household-record-go-no-go, identity-split-audit-2026-08-14, cross-property-leak-audit-2026-08-13

ArchiveRedirect toWhy
application-groups-one-household-one-leadhousehold-record-go-no-goAll 3 of its decisions adopted and shipped
household-groupings-designhousehold-record-go-no-goIts rollout and 3 open questions are all built and resolved per the go/no-go harness
11. Lease Answers, Application & Post-Approval Flow (14 pages)

Canonical: post-approval-gap, approved-to-signed, lease-answers-matrix-2026-08-09

ArchiveRedirect toWhy
approved-to-signed-2026-08-21approved-to-signedOne day older; the newer page already resolved its one open question and has grown since
lease-answers-willows-ship-2026-08-08lease-answers-matrix-2026-08-09Its "proven correct" claim is directly contradicted by the tougher test grid published right after
12. AppFolio / PMS Integration (4 pages)

Canonical: appfolio-sync-surfaces, appfolio-renewal-benchmarks, blocked-a1a7405d

ArchiveRedirect toWhy
13. Sales & Business Development (29 pages)

Canonical: sales-machine-roadmap, sales-machine-card-ia-2026-08-18, prospect-report

ArchiveRedirect toWhy
blocked-7973181a, blocked-83647f97blocked-bc301e4eSame "is the creator-outreach spreadsheet done" question, re-asked as the session sat idle
prospect-funnel-reportprospect-reportStub — content already merged into the combined report
sales-machine-consolelive sales app (salesmachine.propflowai.co)Tombstone — hand-rebuilt static console retired in favor of the real app
sales-console-card-ia, sales-console-card-mock, sales-machine-card-designsales-machine-card-ia-2026-08-18Earlier drafts of the company-card design; superseded by the actually-approved per-stage spec
reliving-not-reportingdemo-deck-slide-by-slideSame demo-deck fixes, narrative form; the slide-by-slide page is the one Fede can act on
deck-research-stress-testsales-deck-researchRe-verifies the same rules against more evidence; fold its 2 new findings into the original brief
14. UI/UX Redesign, Dashboards & Design System (16 pages)

Canonical: detail-page-standard, dashboard-history-architecture, prospects-funnel-redesign

ArchiveRedirect toWhy
signed-leases-row-optionspr-6247-evidence4-option layout comparison decided and shipped, then reworked again since
pr-5489-evidencepr-6247-evidenceProves an earlier version of the metric strip that's already been replaced
detail-pages-inventorydetail-page-standardRaw screenshots gathered specifically to back this page's proposal
prospect-detail-de-rainbow-2026-08-24prospects-funnel-redesignProof-of-work for a shipped color cleanup the redesign brief specifies
15. Clara Coworker Vision & Cross-Channel Architecture (9 pages)

Canonical: clara-coworker-map-2026-08-19, clara-architecture-2026-08-18, one-clara-architecture

ArchiveRedirect toWhy
clara-decision-lineclara-architecture-2026-08-18A small diagram illustrating what the reference doc already explains in full
clara-coworker-capabilities-2026-08-14coworker-actions-design-2026-08-20Its 4 staff-instruction decisions replaced 5 days later by the tighter, narrower version
16. Business Operations & Engineering Vision (5 pages)

Canonical: architecture-source-of-truth, calendar-architecture-2026-08-08, propflow-cards-and-cash-2026-08-02

ArchiveRedirect toWhy
vision-gap-analysis-2026-08-16architecture-source-of-truthPage itself says superseded; findings carried forward into the newer doc's own tracker
17. Fair Housing & Compliance (5 pages)

Canonical: fair-housing-architecture-2026-08-19, opt-out-scope-2026-08-16, compliance-audit-2026-08

ArchiveRedirect toWhy
outbound-safety-architecture-2026-08-02fair-housing-architecture-2026-08-19"Every reply" framing is out of date; the newer page shows what the checkpoint actually covers and what a 21-item audit found missing
18. App & Clara Response Latency (3 pages)

Canonical: clara-loop-latency-2026-08-02, latency-audit-2026-07

ArchiveRedirect toWhy
19. Tour Scheduling & Booking Consistency (8 pages)

Canonical: tour-decider-architecture, tour-decider-v2-camellia-replay-2026-08, voice-tour-availability-decision

ArchiveRedirect toWhy
decision-tour-decider-consolidationtour-decider-architectureIts own recommended fix is exactly what got designed, built, and shipped 2 days later
20. Unrelated (2 pages) — not reconciled in this pass

No reconciliation was supplied for this cluster. Recommend a manual 10-minute triage pass before the next mining run to confirm these 2 pages are genuinely one-offs (not a 21st topic hiding in the "misc" bucket) and either fold them into an existing canonical page or give them their own.

---

3 · The Deduped Decision List

136 distinct decisions, grouped by cluster. Ruling exists? marks whether Fede (or an equivalent standing ruling) has already decided this — if "Yes," the open item is telling whoever's still asking, not asking again.

Cluster 1 — Promises, Escalation, HITL & Transfer Outcomes
Turn on the pre-send honesty checkRuling exists
(catches Clara saying "I sent that" or a wrong price before it reaches a resident). Asked on 5 pages. Options: (a) leave off entirely [Fede's Aug 20 call]; (b) light version, text only; (c) staged rollout; (d) run the newest version silently in the background only. Ruling exists: Yes — Fede said no on Aug 20. Two pages from Aug 25 ask again without mentioning it. Recommendation: surface the Aug 20 ruling first; if revisiting, start with (d).
Give Camellia's escalation system a real owner.Ruling exists
1 page. Ruling exists: Yes — armed Aug 18, owner = hello@propflowai.co, live in prod. No action needed.
Give escalations a crash-proof clock instead of a once-a-day check.Open
2 pages. Options: fold into the existing renewal-reminder clock [recommended, tested] / run the daily check more often / build a new system. Ruling exists: No. Recommendation: (a), start at Willows.
What replaces the safety check reverted Aug 15Open
(it deleted 3 real customers' correct replies). 2 pages. Options: nothing/rely on prompts / ground every claim in a system record first, watch-only rollout [recommended] / second-AI double-check after the fact. Ruling exists: No.
How to treat a caller approved-to-rent but not yet lease-signedOpen
, for phone routing. 1 page. Options: (a) require a drafted/signed lease before treating as resident [recommended] / (b) keep current behavior. Ruling exists: No. Note: a second team is reportedly asking this independently — check before ruling.
Small company logo on the PM escalation email?Open
1 page. Options: (a) yes [recommended] / (b) no. Ruling exists: No — low-stakes, quick yes available.
Build a portfolio-wide fair-housing screen for every outbound message, and when.Open
1 page. Options: (a) now, ahead of other work [recommended — real discriminatory line already sent] / (b) after planned staff-tools work. Ruling exists: No.
Turn on real texting to Camellia residents after phone calls.Partially ruled
2 pages. Ruling exists: partially — general bar set Aug 25, then superseded by a specific found bug (gibberish renewal-reminder texts) on Aug 26. Recommendation: hold, fix the bug, re-check.
Unify the ~18-20 separate "page a human" paths into oneOpen
, before or after Camellia texting goes live. 1 page. Recommendation: unify first.
Fix the "one person = two records" identity-split bug before turning on quiet-skip-if-already-told.Open
1 page. Recommendation: fix identity-split first, sequencing matters.
Cluster 2 — Agent Fleet Ops, Fable Decision Queue & CI/Infra
Place SMOKE_PASSWORD in the mini's local env fileRuling exists
to auto-heal the nightly login-cookie refresh. 3 pages. Ruling exists: Yes — "already agreed a week earlier," per the record; the gap is execution, not the decision. Just do it.
Deploy the fleet-health fixOpen
(stops task pages wrongly telling operators to restart healthy agents) — normal review or unreviewed tonight? 4 pages. Every asking reached the same answer (normal review) — the bug grew 39→92 pages across 4 escalations because nobody executed it. Assign an owner to close the loop, not re-decide.
Merge a ready fix stuck behind an unrelated stuck-red gate.Ruling exists
4 pages. Ruling exists: Yes, as precedent — Fede has personally merged past this exact pattern twice (Aug 12). Recommendation: codify a standing runbook so this stops needing a fresh escalation weekly.
What does "done" meanRuling exists
for an open-ended watch session or a bug-fix+infra-migration task? 3 pages. Ruling exists: Yesdefinition-of-done.html already codifies merged ≠ done; require live proof for the bug-fix+migration pattern, no fixed goal for open-ended steering sessions.
May an agent lift its own do-not-merge hold without fresh human sign-off?Open
5 pages. Recommendation: draw the line at authorship — an agent may lift a hold it placed on itself once its stated concern resolves; a hold from separate tooling still needs Fede's word every time. Not written down anywhere yet — codify once.
Deploy model: auto-deploy on merge, human-gated, or decouple merge from release?Open
2 pages. Ruling exists: No — genuine open strategic call for Gera/Fede.
Security tile: accept grey (honest, partial) or turn on 2 disabled scanners for green?Open
1 page. Recommendation: accept grey — the actual requested fix (36 alerts) is done; faking green violates the same anti-fake-signal principle used elsewhere.
Cluster 3 — Evals, Cerberus & Quality-Grading Infrastructure
Show Clara's own confidence score next to the human grade on the renewal-grading page?Open
2 pages, word-for-word identical, asked 2 weeks apart. Recommendation: show nothing, thumbs up/down only — approve it, don't let it park a 3rd time.
How does a thumbs-down grade become a permanent regression test without eroding the quality bar?Partially ruled
2 pages. Ruling exists: partially — "fully automatic" already picked; the refinement (land as tracked-but-not-counted until Clara passes) is the same recommendation given twice.
Standing policy for a scheduled check that goes red for weeks with nobody reacting.Open
2 pages. Two different failure shapes (no alert vs. an ignored alert) — recommend one combined rule: every check gets an alert + a named owner, and only gets muted alongside a dated fix ticket, never muted alone.
Which of 28 conflicting pass/fail thresholds should winOpen
(9 grouped calls)? 1 page. Recommendation: approve all 9 as recommended — backed by real run data.
Move Clara's model tier off Opus — is Sonnet 5 actually worse?Open
1 page. Recommendation: re-run the specific safety findings (privacy leak, silent replies) on the now-fixed test tool before deciding either way — the original complaints may be test-tool artifacts.
Which fix ships for phone calls having no safety check before Clara speaksOpen
(unlike text/email)? 2 pages. Both agree it's the biggest remaining gap; pull the separate voice-hallucination-guard-decision page for the actual options.
A bearer token gap for preview/eval infra — one token or two?Open
2 pages, 3 days apart. Confirm whether these describe the same gap or two, then resolve together.
Run a Haiku ("FAST" tier) quality sweep now or skip it?Ruling exists
1 page. Ruling exists: Yes, effectively — the very next day's audit page IS that sweep (44 of 49 lanes audited). What's actually still open is the audit's own rollout-order calls, not whether to run it.
Cluster 4 — Voice Agent Quality & Latency
Which AI brain answers Clara's phone calls — current or a faster candidate?Ruling exists
2 pages. Ruling exists: Yes — Fede decided Aug 24 to keep the current brain; the faster one failed the honesty/consent bar on an 866-call replay.
Stop Clara confidently claiming something happened (a booking) when it didn't.Open
3 pages. Every page converged on the same ~15ms fix (booking tool hands Clara an already-truthful sentence). Still awaiting explicit go-ahead to build for the real phone line.
Should Clara ask before transferring a caller, instead of sometimes just doing it?Open
1 page. Recommendation: yes — the single biggest real problem the call review found.
Build the guard that catches an unfulfilled promiseOpen
(texting a link, a callback that never happens)? 1 page. Recommendation: yes — no safety net exists anywhere today.
Turn on the real-fee/real-hours guard and the language-switch fixOpen
(both built, switched off)? 1 page. Recommendation: yes.
Restore the test phone line to normal settingsOpen
now that the faster-model trial is over? 1 page. Recommendation: yes, settings are just leftover.
Out-of-scope caller: live-connect or take-a-message?Open
1 page. Recommendation: live-connect only on the unrecognized-caller line — real volume shows 1-2 calls/week, no need to change the other 5 lines.
13 unruled "who should this caller reach" scenarios.Open
1 page. Recommendation: rule on all 13 in one pass; default unclear ones to "take a message" in the meantime.
Should Clara's voice get expressive mode / higher randomness / a different voice / a wording rewrite?Ruling exists
1 page. Ruling exists: Yes, by real-world result — tried live for ~8 minutes Aug 22, pulled after dead air and mixed-up language/fees. Off until a real audio test happens; the voice-comparison test and standalone wording rewrite are still genuinely open.
Cluster 5 — Turnover & Vendor Coordination
Switch Camellia's vendor-facing job number from AppFolio work-order IDs to purchase orders?Open
2 pages. Recommendation: flip Camellia — built, tested against 919 real POs, already live elsewhere.
Can Clara read a PO number aloud to a name-only (not phone-verified) caller?Open
1 page. Recommendation: require phone verification first, matches existing payment-adjacent-info policy.
Should Clara eventually create new POs, not just read them?Open
1 page. Recommendation: read-only for now.
Auto-detect vs. manual-with-suggestion for whether a property runs POs or WOs?Open
1 page. Recommendation: manual setting, auto-suggested.
Upgrade to SIP trunking so keypad presses land reliably?Open
2 pages. Recommendation: upgrade — one-time platform fix, no noted downside, cited as blocking in two investigations.
How much freedom does Clara get to auto-book a vacant unit's vendor visit?Open
3 pages, same 2 decisions asked 3 times. Recommendation: auto-book with a short PM veto window.
Check occupied/vacant status live, or trust a stamped-once flag?Open
3 pages. Recommendation: route to Gera as an engineering architecture call, not a Fede product decision — stop parking it on 3 pages.
Should Clara proactively disclose she's an AI on outbound vendor calls?Ruling exists
1 page. Ruling exists: Yes — no proactive disclosure, honest if asked. Stop listing as pending.
Build a general "save what a vendor asked for" commitment-record system?Open
1 page. Recommendation: build it — 2nd time the same failure hit the same vendor pattern; cheaper fixes already tried and didn't hold.
Bigger LLM-based vendor-email parser, or invest in voice/portal capture instead?Open
1 page. Recommendation: narrow parser for the one vendor that already emails schedules + opt in a 2nd.
Turn on purchase-order syncing at Camellia for vendor spend attribution?Open
2 pages. Recommendation: turn on — ~$6,700 of currently-unattributed spend resolves immediately, no new code.
Where does Clara transfer a vendor caller who needs a human — front desk or named staff line?Open
1 page. No strong recommendation given; needs a quick answer from property staff.
Build a general way for Clara to save an unclear vendor request?Open
1 page. Recommendation: build it — named as the root fix behind a same-vendor-frustrated-twice incident.
Real, published online lease templates vs. manual PDF as the permanent last leasing step?Open
1 page. Not this cluster's call — business/legal content decision touching every property including Camellia; needs Fede + whoever owns lease terms.
Cluster 6 — Collections & Delinquency
Resubmit the toll-free number's carrier verification now, or leave risk-accepted?Open
2 pages. Recommendation: leave risk-accepted — Twilio's own guidance says content-only changes don't force resubmission, and a real text already delivered clean.
How much legal counsel review is required before rent-texting goes live?Partially ruled
3 pages. Ruling exists: partially — payment-plan automation ruled out without counsel (Gera, Aug 25); the general "does all collections texting need counsel" question is still open. Recommendation: formalize the de facto answer (proceed on the legal mapping, escalate real gaps).
Who receives rent reminders, and at what frequency cap?Open
1 page. Still genuinely open. Recommendation: default 3/month & 7/week, start with the measured 50.5% portal-payer cohort.
Disclose to residents that collection messages are automated?Ruling exists
3 pages. Ruling exists: Yes — live in production as of Aug 25 (PR #6260). Update the still-pending-looking page.
What gating sequence must be true before flipping real rent-texting on?Ruling exists
1 page. Ruling exists: Yes, informally — a real resident got a real text Aug 25, approved by Gera, no formal gate answer was ever needed. Retire the question, adopt the actual practice (hand-pick one resident at a time).
Show, hide, or "no source yet"-label the collections filter options that can never currently match anything?Open
2 pages. Recommendation: show with a "no source yet" label — matches the panel's existing honesty pattern.
Confirm who actually signs/serves 10-day demand notices on-site.Open
1 page. Still open, low urgency — fold into the next demand-builder review.
Rent-only vs. rent-plus-other-amounts on the 10-day demand notice?Open
2 pages. Still open — route to counsel, getting it wrong risks voiding the eviction filing.
Build the remaining Phase 3 items (magic-link signature, PDF generation, payment plans, etc.)?Partially ruled
2 pages. Ruling exists: partially — several items already withdrawn by Gera (Aug 25); the remainder waits on the counsel question above.
Cluster 7 — Renewals & Rent Pricing
What rent to offer Paula Delgado (unit 607)Open
given her $1,150 offer predates the company's own drop to $1,000? 2 pages, disagreeing. Recommendation: match $1,000 (not split at $1,075) — $1,150 is already confirmed stale, there's nothing real to split.
Flag a live renewal offer when market rent drops below it, before the resident notices?Open
1 page. Recommendation: build the flag+notify check, keep the freeze on auto-repricing. Not yet built.
Narrow the always-reply rule to skip renewal/lease threads addressed to a named staffer?Open
1 page. No recommendation given — open scope question for Fede.
Suppress a duplicate escalation email sent to a mailbox already hosting the conversation?Open
1 page. Recommendation: suppress.
Fix the website scrape silently erasing manual concession-wording edits (recurs every ~9 hours)?Open
1 page. Recommendation: fix the live website copy now (it's the scraper's own source); treat "should manual edits always win" as a separate, later decision.
Fix the 43 scripts that can silently write to the wrong (staging) databaseOpen
— one at a time, or one shared fix? 1 page. Recommendation: build one shared, fail-loud table-picker — this class of bug already let a real email send against the wrong database once.
Add the two $0-lease-rent units to the Camellia rent drop?Open
1 page. No recommendation stated — needs Darrin's sign-off.
Accept ADR-0118Open
(separate "internal transfer" count so occupancy stats stop double-counting unit-to-unit transfers)? 1 page. Recommendation: accept — adds a stat, doesn't touch underlying occupancy math.
Cluster 8 — Inbound Email Pipeline
Ship the sender-authentication fix for Camellia's live mailboxOpen
(stops spoofed AppFolio-lookalike emails)? 3 pages. Recommendation: ship it — ready and re-confirmed safe for 3+ weeks, just never explicitly approved.
Tighten DMARC from watch-only to actually blocking?Ruling exists
2 pages. Ruling exists: Yes — Fede said yes Aug 14; only the timing clock is ambiguous (two different pages cite two different start points). Use the newer, stricter Aug 16 timer.
Give the "needs a human" email queue a named owner and response-time target?Open
2 pages. Recommendation: assign now — this exact gap let a hot lead go dark for 2 days.
Cluster 9 — Knowledge Base & Docs Infrastructure
Merge 3 verified, low-risk docs-publishing tool fixes waiting on one click each — merge now, or give standing merge rights for this class of fix?Open
3 pages. Recommendation: merge all 3 now; decide the standing-permission question separately.
Camellia move-in special: set the property flag, then merge — or hold either step for review?Partially ruled
1 page. Ruling exists: partially — the underlying design (per-property flag) was already decided twice; only step order is open. Recommendation: flag first, then merge.
When two property facts disagree, which wins, and how much provenance should be visible?Open
3 pages, three different answers. Recommendation: build Aug 8's versioning first (furthest along), then layer Aug 26's fixed win-order (lease > PMS sync > staff answer > onboarding > website scrape) on top.
Should a one-time favor for one resident ever be taught to Clara, even flagged as one-off?Ruling exists
2 pages, contradicting. Ruling exists: Yes, by the newer page — never teach person-specific one-offs, full stop. Confirm this stands; don't build the older classifier proposal.
What should the property-knowledge record look like structurally?Open
1 page. Recommendation: one record per property, versioned as a whole (Option D) — cheapest given existing code, keeps one lookup + audit trail.
Should every dollar/policy fact live only in a structured field, migrating out of free-text paragraphs?Open
1 page. Recommendation: yes, move into fields and delete the paragraphs — this exact split caused a wrong $1,200 deposit answer once already.
Which of Camellia's two office phone numbers on file is correct?Open
1 page. Recommendation: don't guess — ask Joanna or Erika, fix the record, delete the duplicate field.
Cluster 10 — Identity, Dedup, Households & Onboarding
Model applicants who apply together (3-part: household record vs. count-fix vs. patching; auto-merge vs. PM-confirm; ship occupant role in v1?)Ruling exists
2 pages. Ruling exists: Yes, shipped — PR #5377 merged, harness scenarios passed, Camellia backfilled Aug 4-5.
What to do with 3 known-bad duplicate application rows in Camellia prod data?Ruling exists
1 page. Ruling exists: Yes — linked, not deleted; 57 legacy duplicates retired the same way.
Demote the unreliable co-signer name-check to a PM-facing hint, or delete it?Ruling exists
1 page. Ruling exists: Yes — demoted; role now comes from PMS structural evidence, never a name-text guess.
Order of 3 follow-up fixes from the Gmail dot-address incidentOpen
(canonicalize addresses, clean 20 duplicate people, fix "Unknown" sender name)? 1 page. Recommendation: ship all 3 as separate small PRs, but hold the 20-duplicate cleanup until the merge-candidate queue (below) exists.
Six implementation calls for the "one human, two records" identity-split fixOpen
(re-point vs. merge, intake vs. backfill, evidence threshold, redirect handling, test rewrite, documentation)? 1 page. Recommendation: adopt the page's full recommended package — this is the biggest still-open item in this batch.
Five calls from the "Resident Who Was Still a Prospect" investigationOpen
(derive resident status, email-typo fix scope, unsigned-lease watchdog, AppFolio-texting visibility, missed-transfer paging)? 1 page. Recommendation: take the recommended option on each, especially deriving "resident" from real AppFolio tenant status + move-in date.
Tour booked by one person for another: one merged conversation, two linked, or notification-only?Ruling exists
1 page. Ruling exists: Yes — two linked conversations, Fede ruled Aug 5. Not yet built (HH-16/HH-17 still red).
Real tenant PII exposed on staging for 3 months — reportable incident?Open
1 page. Recommendation: delete the exposed rows today regardless; the reportability call itself is explicitly Fede's alone.
Roll the 6-digit-code onboarding sign-in out past its single pilot property (The Willows)?Open
1 page. No stated recommendation; tested clean (~2 min vs. 40+ min) with no reported downside — worth a decision.
Cluster 11 — Lease Answers, Application & Post-Approval Flow
When does a prospect get the application link — tour first, or always on explicit ask?Ruling exists
1 page. Ruling exists: Yes, but never formalized — Fede said "always send if asked" in July; never written into policy, which is why a different team member (Gera) remembers the opposite rule and a reply got missed. Recommendation: formalize option A now.
Approve Clara owning the whole approved-to-signed stretchOpen
(auto-screen, send lease, chase signatures, post credit)? 2 pages. Recommendation: proceed — the one blocking money question ($33 vs $38 fee) is resolved and a supervised test run proved the core pieces work.
Build order for Camellia's post-approval answers; fix quality-test failures before launch?Open
1 page. Recommendation: policy answers first (running parallel to the automation track, not blocking it); fix quality-test failures before any new property launches.
How should the team be told a keys/move-in call ended with nothing booked?Open
2 pages. Recommendation: use the existing escalation feed — one place already watched, no new inbox.
Approve wording + ship the booked-key-pickup confirmation email?Open
1 page. Recommendation: approve the wording now; hold the flip-on until the cross-channel safety test finishes (team's own stated plan).
How should a conversation get released from a stuck "escalated" status?Open
1 page. Deliberately deferred by the team pending the safety test above — no Fede action needed yet.
Quote the security deposit amount on every channel (not just voice)?Ruling exists
2 pages. Ruling exists: Yes — accepted, Option A, 2026-08-03; fix built. Only a shipping hold (a smaller model side-effect bug) is blocking go-live.
What's blocking the already-accepted deposit-answer fix from shipping?Open
1 page. Recommendation: fix the one found regression (inappropriate tour-booking side effect), then ship.
Does Camellia's "no housing vouchers" line conflict with Colorado's source-of-income law?Open
1 page. Disputed reading, unresolved — needs a real legal/policy read, separate from engineering.
Should the post-approval lease-answer feature talk to any prospect, or only approved applicants?Ruling exists
2 pages. Ruling exists: Yes — Fede, Aug 7: keep it open to anyone who asks, with fair-housing test coverage added.
Two small lease-answers opensOpen
: add a rent-by-lease-length field; keep or relax the ambiguous-sender gate? 1 page. Recommendation: add the field (cheap, closes a real gap); leave the gate as-is (costs nothing today).
Cluster 12 — AppFolio / PMS Integration
Push the work-order auto-refresh branch (PR #5702) for code review?Ruling exists
1 page. Ruling exists: Yes, moot — already merged into main as of the same day, daily sweep armed.
How thorough should the fix be for CAM-2171Open
(a missing tenant name PropFlow should already have)? 1 page. Genuinely open — full fix vs. quick patch vs. explain-only. Needs Fede's call on thoroughness, not an inferred answer.
Cluster 13 — Sales & Business Development
Is the creator-outreach research spreadsheet finished, can the session close?Open
3 pages, same prompt re-asked as the session sat idle. Recommendation: yes, the spreadsheet is the deliverable — answering the newest page closes all 3.
How should sales-app chat workRuling exists
(subscription-only AI, no metered key)? 1 page. Ruling exists: Yes — Fede picked "works while the Mac is on, holds the question otherwise" live on Aug 12, the same evening a different handoff page still lists it as pending.
Four launch-gate approvalsOpen
: call-recording consent, using a customer's real numbers in a case study, pilot pricing/length guardrails, live proposal-tracking? 2 pages, duplicated. Genuinely still open — not answered anywhere in the batch. Track once, in the roadmap.
Which factors should rank prospective customersOpen
(self-managed owners, AppFolio-fit, portfolio size — plus 2 debated extra factors and size cutoffs)? 2 pages. Still open, needs Fede + Sean together.
Which of 4 proposed company-panel designs should ship?Ruling exists
1 page. Ruling exists: Yes, by different means — a more thorough per-stage design was built and approved instead; the original A/B/C/D question is moot.
Is the Botnick demo deck ever presented live, or only sent as a link?Open
(Decides whether captions stay full-sentence or get stripped for a presenter.) 2 pages. Still open, needs Fede's answer.
Cluster 14 — UI/UX Redesign, Dashboards & Design System
Which of 4 mocked prospect-metric-row layouts should ship?Ruling exists
1 page. Ruling exists: Yes, shipped and since reworked again.
Keep the 2 gradient hero banners on detail pages, or retire both?Ruling exists
1 page. Ruling exists: Yes, shipped — both removed same day, replaced with one shared plain header.
Switch the detail-page progress bar's "done" color to the brand gradient?Ruling exists
1 page. Ruling exists: Yes, shipped.
Which list-page redesign direction should the prospects tile row use?Open
1 page. Shipped a version that departs from all 3 originally proposed directions — needs a quick yes/no from Fede confirming it matches intent, not a full re-litigation.
Should "Applications" (and the prospects funnel) count only what's active now, or everyone who ever reached that stage?Open
1 page. The one real unresolved decision in this cluster — was supposed to go to a founders' call; instead a mixed answer (old counting on tiles, new counting in the funnel pop-up) already shipped without that call happening. Needs Fede told and asked whether to keep the split or unify.
Restore any of 14 deleted UI componentsOpen
(world map, Clara orb shader, command palette, etc.)? 1 page. First-time decision needed, low urgency.
Approve the mechanical fix for the "double loading skeleton" flash on detail pages?Open
1 page. Recommendation: approve — safest, purely mechanical fix in the plan, not touched by anything already shipped.
Cluster 15 — Clara Coworker Vision & Cross-Channel Architecture
What can Clara do on her first staff-instruction feature, and in what order?Partially ruled
3 pages, two competing proposals. Ruling exists: partially — building this AT ALL right now is paused by Fede (Aug 20); which option-set wins once it resumes is still open. Recommendation: restart from the narrower, comms-only-first version.
Unify the 13 separate ways the app pages a Camellia staffer into one, with a silent-first trial period?Open
1 page. Recommendation: approve the full plan and the 1-2 week silent trial — false-alarm evidence (3 in one night) is already in hand.
Approve the four-part fix for Clara over-promising things she can't back up?Open
1 page. Recommendation: approve all four — covers ~90% of 654 real promises studied; the first fix is already live.
Unify Clara's 11 separate phone scripts with the one text/email rulebookOpen
(pace, consolidation, blocking, voice safety check, gate-code disclosure)? 1 page. Recommendation: take every recommended (evidence-first, phased) option across all 5 sub-questions.
Change how Ask Clara's "top 5 portfolio problems" groups renewal-risk findings?Open
1 page. Worth doing, but needs a quick founders' call first since it changes what an owner sees as "the" top problems.
Cluster 16 — Business Operations & Engineering Vision
Hold a sensitive fair-housing staff answer for a check before it goes live, or let it go live instantly?Ruling exists
1 page. Ruling exists: Yes, shipped Aug 19 — the narrow version (only sensitive topics get the extra check).
Bring back GitHub's built-in merge queue?Ruling exists
1 page. Ruling exists: Yes, already ruled before this audit was written — custom tool stays, GitHub's version cost ~1/4 of the whole tooling bill for zero benefit. Just fix the vision doc's stale wording.
Calendar engine rollout: 4 small callsOpen
(Google-calendar support in v1, when tour bookings move over, how vendor-visit events are modeled, how cautious the overnight cleanup job is)? 1 page. Page is tagged "decided" but isn't — still needs Fede's yes on all 4 (each has a lower-risk recommended pick).
Which business credit card should PropFlow switch to?Open
1 page. Recommendation: Mercury IO only, matching Fede's own stated lean.
Move idle cash ($320K) into a Treasury money-market fund?Open
1 page. Recommendation: proceed as written — stops 0% interest, gets under the FDIC limit, keeps same/next-day access.
Fix the 21 places a message reaches a resident with no fair-housing screening in front of itOpen
— patch riskiest now, or redesign? 1 page. Genuinely still open — Fede parked it Aug 19 ("needs more thinking") and handed to another workstream Aug 20; 3 of 7 affected areas are high-risk. Worth checking on.
Five smaller truth-checking-layer and paused-daily-reviewer callsOpen
(check frequency, actions vs. words, mid-call correction, overnight replay job, turn the paused reviewer back on)? 1 page. Recommendation: take the doc's own picks on all 5 — cheaper, already-partly-built options across the board.
Cluster 17 — Fair Housing & Compliance
Where should the fair-housing screen sit on Clara's outbound messages, and how should jurisdiction be modeled?Open
1 page. Recommendation: Option A — one checkpoint at every outbound send, audience- and state-aware. Still parked pending Fede's go-ahead.
Should Clara have a crisis/self-harm response plan, and should it be published externally?Open
2 pages. Recommendation: both, in order — build the internal playbook (~1 day of work), then publish once built.
How much vendor/subprocessor disclosure should PropFlow publish?Open
2 pages. Recommendation: ship the quick missing-vendor-names list now; treat a fuller enterprise-grade package as a later, deal-triggered investment.
How wide should a resident's opt-out reach — one channel, one topic, or everything?Ruling exists
1 page. Ruling exists: Yes — Fede decided Aug 17 (context-scoped, not total-block); already shipping.
Cluster 18 — App & Clara Response Latency
What to do with the vendor list's phone/email columnOpen
(slow, 98% blank)? 1 page. Recommendation: remove from list view + add paging (Option A+B) — doesn't require a database migration.
Build a faster general tenant-lookup index (a database migration)?Open
1 page. Recommendation: leave it — the report itself calls this optional polish, nothing is waiting on it.
How long should old activity/history logs be kept before archiving?Open
1 page. Needs an answer soon — two separate investigations independently found unlimited log growth is the root slowdown cause; kicking it down the road guarantees a 3rd rediscovery.
How stale can shared dashboard numbers be, to allow more aggressive caching?Open
1 page. Recommendation: approve a 30-60 second default staleness window for portfolio-wide stats.
Should Clara's 5-step reply-speed fix ship, and in what order?Open
1 page. Recommendation: start with the cheapest, lowest-risk piece (a one-line database swap, ~3 seconds saved) to prove the plan before committing further.
Narrow the automatic fair-housing check to leasing-only replies for speed?Open
1 page. No engineering recommendation given on purpose — explicitly a compliance/product call, needs Fede or compliance sign-off.
Cluster 19 — Tour Scheduling & Booking Consistency
How far to consolidate the systems that can change a booked tourRuling exists
(do nothing / delete the oldest / full rebuild)? 2 pages. Ruling exists: Yes, by shipped work, not an explicit ruling — the new single decider is merged and replay-proven. Formally close both pages as resolved rather than still-open.
Turn the new single tour-decider on for real customers — where first, and what to do with the old code?Open
2 pages, disagreeing same-day recommendations. Recommendation: go straight to the real customer (Camellia) as a watched trial, then delete old code — this is genuinely still open and needs Fede's go-ahead.
Should a "flagged as follow-up" note in a code-change writeup have to become a tracked to-do before the change ships?Open
1 page. Recommendation: yes — this exact gap is why the Aug 15 mistake happened.
Ship the Saturday-tour-times voice fix — how much rollout caution, how many days of look-ahead, what to do for calendar-connected properties?Open
1 page. Recommendation: take all 3 recommended (cautious) picks — the bug is confirmed real but rare (2 failures ever, out of 1,000+ real calls checked). ---

4 · Overlapping Issues, Different Solutions

The same problem, proposed or solved two-or-more different ways by different pages, with no cross-reference between them. 86 found across the corpus; grouped by cluster.

ClusterIssuePagesSolutions (compressed)Which wins
C1Pre-send honesty check: turn on?5 pagesNo (Fede, Aug 20) / light-text-only / staged rollout / silent-background-onlyFede's Aug 20 "no" stands; if revisited, silent-background-only first
C1Escalation silence-scoping fix4 pages"Shipped clean, proven" (Aug 4) vs. "still broken in a real case" (Aug 21)Aug 21 — the "proven clean" claim didn't hold everywhere
C1Aug 15 booking-guard incident replacement2 pagesPause for bigger redesign vs. concrete options + prevention rulesThe concrete-options page (RCA); the redesign page's content already folded into honesty-layer-deep-inspection
C1PM reminder office-hours bug2 pages"Fully closed" (Jul 28) vs. "still broken, root cause deeper" (Aug 15)The Aug 15 page — "all closed" was premature
C1Camellia escalation-ownership system2 pagesRaised as a to-do vs. already turned onAlready resolved — no decision needed
C1Durable clock for escalations2 pagesNot competing — POC supplies the tested answer to the open questionAdopt the POC, Willows first
C1Dead/dropped call transfers: notify or stay quiet?3 pagesQuiet-by-design (Aug 10) reversed to notify (Aug 25), then a deeper root-cause mechanism proposed (Aug 26)The Aug 26 mechanism — it fixes the miscount and the notification gap at the root
C2Standalone AgentFlow viewer vs. in-app page2 pagesOld standalone viewer vs. new in-app sectionIn-app page; old one is a dead stub
C2AI code review process changes2 pagesOriginal proposal (external research) vs. re-measured against real 2,729-PR historyThe re-measured page — the headline worry was false
C2CI two-tier test model working?2 pages"Already cut cost 83%" vs. "89% still escalate to full suite" (later retracted as a measurement bug)The self-corrected page — real lever left is the self-hosted-runner ban
C2Fleet-health restart bug — deploy now or review first4 pagesAll 4 agree (review first) but none actually landedSame answer 4x; assign an owner to execute
C2Node 24 breaks PDF reading2 pagesSame root cause, same fix, one shallowerThe focused deep-dive page is canonical
C2Duplicate comment cleanup PR2 pagesLiterally the same PR, "adopted from" the otherCollapse to one
C2Stuck-red gate blocking a ready PR4-5 pagesMerge on precedent / disable branch-protection for one merge / fix the unrelated failure firstCase-by-case per PR, but recurring pattern needs a standing runbook
C2SMOKE_PASSWORD nightly refresh3 pagesIdentical question, 3x, never executedAlready agreed — place it, stop asking
C2Silent dead-recipient message routing3 pages2 near-duplicates of one incident + 1 related-but-distinct bugCollapse the 2 duplicates; keep the 3rd as a separate, deeper architecture gap
C2"30 agents" fleet lessons-learned2 pagesSlide-deck lessons vs. distilled runbookThe runbook; lessons-learned becomes background
C2Smith's Slack/PR reliability6 pagesOne continuous re-diagnosis over a monthThe live issue tracker (status) + the newest root-cause pass (what to fix)
C2Smith PR-drive architecture5 pageswebhook → v1 (+Temporal backstop) → v2 (Temporal-only) → v3 (+task-workflow seam) → v4 (+closure vocab)v4, current and shipped
C2Fleet triage/audit passes3 pagesTwo point-in-time audits + the distilled methodology that resultedThe methodology; audits become dated history
C2Operator/agent-fleet architecture4 pagesAs-built roles / 3 inconsistent primitives fixed / Temporal migration proposed / deeper task-store-ownership auditThe newest, deepest audit (nine-doors-one-store) drives the next fix
C2CI/merge pipeline speed7 pagesOne merged stub + sequential checkpoints + 2 incident deep-dives + 1 governance-breach docThe living reference (pipeline-delivery); others become linked incident records
C2CLAUDE.md bloat/trim3 pagesPersonal-file fix (closed) / partial repo-file migration / full repo-file audit with a live contradiction foundThe inventory page — it has the unresolved contradiction
C3Show confidence score on grading page3 pagesLiterally the same question twice + the design doc that set up the formatShow nothing — approve, stop re-parking
C3Thumbs-down → regression test4 pages"Fully automatic" picked, then the same safety-gap fix proposed twiceTracked-but-not-counted, ship it
C3Red scheduled check, no reaction2 pagesAdd alerting everywhere (Jul) vs. silence + track-on-a-card (Aug) — opposite motionsOne combined rule: alert + owner + never mute alone
C3Is AI-honesty testing load-bearing?5 pagesOne check made blocking, wider suite advisory, a separate checker recommended-but-unwiredFinish the list as one: ship the unrelated feature, flip Cerberus to enforcing, settle blocking-vs-advisory once
C3Is Sonnet 5 actually worse?2 pagesBuggy 1000-scenario report says stay away vs. bug-fixed 332-case re-run says no real differenceDon't trust the earlier report's specific safety claims without re-checking on the fixed tool
C4Faster AI brain for phone calls3 pagesKeep current + tune-up vs. keep current + real trial vs. fully decided keep-currentThe final, fully-decided report wins
C4Stop confident false claims (hallucinated bookings)3 pages4-ranked-options doc / narrower lab-only fix / folded into 5-fix production planThe near-instant pre-written-sentence design, on the real phone line
C4Make Clara's voice less robotic2 pagesPlan: A/B test expressive mode first vs. what happened: shipped live without a test, failed in 8 minutesTrust the real-world result — off until a real audio test happens
C4Live call-transfer wording3 pagesOne fixed platform-spoken line (ruling) vs. 11 custom lines that shipped anyway vs. a stuck PR using wording Fede already ruled outStick with the original ruling; close the stuck PR and pick new wording
C5What job number do vendors actually use?2 pagesDraft brief with policy questions vs. finished ADR, shipped everywhere but CamelliaThe ADR; only "flip Camellia" remains open
C5AI turning a vendor mention into a false charge3 pagesSame shipped fix (rules engine + 2nd-AI-check-on-ambiguous), described 3 waysThe ADR as decision record; fold the other two's unique numbers in
C5Why tests said turnover code was fine when it wasn't3 pagesDedicated RCA / independent replay that rediscovered the same bug / structural fix proposalThe structural-fix page forward, RCA as its evidence
C5IVR phone-menu keypad failures2 pagesSame investigation, same night, one deep-dive + one go-live scorecard citing itBoth stay; scorecard should link to the recon instead of restating it
C5Upgrade to SIP trunking?2 pagesSame question, same framing, both flag it as Fede's callOne decision, not two
C5Vacant-unit auto-booking freedom + occupancy-check timing3 pagesOne real decision brief; two other pages just point back to it without adding an answerThe decision brief; others should just link to it
C5Vendor spend data + UI2 pagesData-plumbing investigation vs. its own explicit UI companionSequence data first, UI depends on it — not competing, two halves of one project
C5Getting a vendor's visit onto a real calendar2 pagesNarrow email-parser fix for the one vendor who emails schedules vs. a big shared booking-engine architectureNot the same scope — keep both, the booking engine is what the parsed data eventually feeds
C6"Define done" for the collections build session2 pagesTwo different proposed finish linesBoth overtaken — the phase tracker is the standing answer, no fixed finish line needed
C6Same policy-drawer feature, described 3 times3 pagesDesign brief (6 modules) / shipped proof for collections / separate smaller header PRKeep the design brief; two evidence pages become PR-specific proof snapshots
C6Is collections data broken/missing?3 pages"6 launch-blocking defects, not safe" vs. "expected data-availability gap, not a bug"The later, measured diagnosis — readiness deck is stale
C6Collections phase-plan numbering/scope2 pagesFounding brief (guessed 40% cohort) vs. re-measured plan (50.5%, current)The re-measured plan
C6Same Phase 1 shipment reported twice2 pagesPodcast-script wrapper vs. interactive checklist — identical substanceThe interactive checklist
C6Carrier registration change needed before rent texts?2 pagesOpen multiple-choice question (pre-send) vs. answered by measurement (post-send, delivered clean)Answered by measurement; only the narrower resubmit-or-not question remains live
C7Paula Delgado renewal rent2 pagesSplit the difference ($1,075) vs. match verified market rate ($1,000)Match $1,000 — nothing to split, one number is already confirmed stale
C7Flag a stale offer after a market-rent drop?2 pages"Intentional, accepted design" (freeze) vs. "gap worth a safeguard" (the one case it hurt someone)Additive — keep the freeze, add the flag+notify step
C8Same 3 silent-email-drop bugs, 3 separate one-off fixes4 pages3 different bugs, 3 different fixes, no shared tracking listKeep all 3 fixes; start one living checklist of known failure modes
C8Unwatched "needs a human" email queue3 pagesFlagged need for an owner (unassigned) / narrow alert shipped / surface-in-conversation-thread proposedBuild the surface-in-thread fix + just answer who owns the queue
C8Email feels slow / goes to spam2 pagesFix the Gmail-forwarding-hop plumbing vs. cut unnecessary AI calls in the decision stepNot a conflict — both should happen, cross-reference which is already banked
C9Camellia move-in special told inconsistently2 pagesOne-property switch-flip vs. general "which source wins" architecture fixShip the switch now as a stopgap; retire it once the general precedence rule ships
C9Where a fact came from / which one wins3 pagesJust surface hidden existing data / add explicit source+precedence field / version every save firstDo versioning first (furthest along), then layer the precedence rule on top
C9One-off favor: ever teach Clara?2 pagesClassify-and-store-flagged vs. never-teach-at-allNever-teach (the newer, better-evidenced page) — genuine contradiction, not just a different mechanism
C93 docs-publishing-tool fixes stuck on one merge click each3 pagesDifferent bugs, same shape of askMerge all 3 now
C10Household/applicant modeling3 pages3-way choice proposed / detailed 3-stage design / actually shipped end-to-endThe shipped go/no-go page; other two become design history
C10Duplicate person records, one-off scripts each time3 pages3 different bespoke cleanup scripts for 3 different root causesBuild the merge-candidate queue once, route future cleanups through it
C10Deciding which conversation a message belongs to2 pagesVisitor-recognition design + a separate shared message-routing resolver, same day, no cross-referenceFold visitor logic into the one shared resolver
C10Cross-org/cross-property data leaks2 pagesFixed at the API-route layer (ADR-0120) / found again 2 weeks later in voice/leasing conversational toolsReference ADR-0120 from the new fix — same root cause, new layer, not a 4th unrelated incident
C11Which Approved-to-Signed draft is current2 pagesAug 21 draft (open fee question) vs. Aug 23-updated version (resolved + growing)The updated, living version
C11Is the lease-answers feature trustworthy?2 pages"Proven correct" ship report vs. test grid finding invented numbers + a 9x policy leakThe test grid — newer, harder evidence, and honest about being "a slice, not a validated system"
C11Build order for Camellia post-approval3 pagesResearch recommends policy-answers-first; in practice, key-pickup fixed immediately and automation started anywayClose as: policy fields + automation in parallel, don't leave as still-choosing
C11Key-pickup team notifications3 pagesMissed-booking alert (3 options) vs. booked-confirmation email (ready) vs. RCA saying hold the email for a safety testTwo separate features — ship the alert now, hold the email for the safety test
C11Deposit-question decision location2 pagesShort decision page vs. long postmortem with the same decision + the current shipping holdThe postmortem — it's the one with the live status
C11Who can the post-approval feature talk to2 pagesOriginal design: approved-only / shipped: any prospectAlready decided (Fede, Aug 7) — keep the wider reach; fix the original page's framing
C12Work-order auto-refresh: push for review?2 pages"Should I push?" (unshipped) vs. screenshot evidence it's already mergedAlready merged — the question is moot
C13Company-panel design (4 rounds)4 pagesSimple card / 4 competing designs / clickable mockups / final per-stage spec, approvedThe approved per-stage spec; other 3 archived
C13Discovery-deck writing-rule evidence2 pagesOriginal research brief vs. re-stress-tested against a bigger datasetFold 2 new findings into the original brief, keep one page
C13Botnick demo-deck rewrite2 pagesSame fixes, narrative reasoning vs. actionable before/after slidesThe slide-by-slide page; reasoning already folded into its notes
C14Detail-page header/banner/progress-bar standard4 pagesPlan still lists 2 questions "open" the same day 2 PRs already shipped both answersThe shipped PRs settle both; only the double-loading-flash fix remains a real open ask
C14Why dashboard charts only show ~7 months2 pages"Unverified, needs a backfill decision" vs. "root-caused: years of data already exist, just point the chart at it"The root-caused page — no backfill needed for Camellia
C14Prospect metric-row layout3 pages4-option comparison / shipped Option A / reworked again sinceThe current shipped version; the comparison page is moot
C14Prospects page color scheme + funnel counting definition4 pagesColor cleanup: shipped everywhere, no decision needed. Counting definition: was supposed to go to a founders' call; a mixed answer shipped instead without that call happeningColor: done. Counting: the one real open decision — tell Fede the mixed answer is already live, ask if that's fine
C15Clara's first staff-instruction feature scope3 pagesBroader same-matter scope (Aug 14) vs. narrower comms-only-first scope (Aug 20)Restart from the Aug 20 version once the (already-paused) feature resumes
C16How much of the eng vision is actually built?3 pages3 separate point-in-time code checks on 3 different daysThe newest, self-updating tracker inside architecture-source-of-truth
C16Sensitive staff answer: instant or checked first?2 pagesAudit flags it as open vs. next day's doc shows it already shippedAlready shipped — nothing left to decide
C16Bring back GitHub's merge queue?2 pagesAudit flags it as a live disagreement vs. tracker shows it was already ruled out before the audit was writtenAlready settled — just fix the stale wording
C17Where compliance checks run on outbound messages2 pagesOne checkpoint shipped scoped to prospect-leasing replies only vs. a 21-item audit finding it misses residents, maintenance texts, and phone calls entirelyWiden the same checkpoint (Option A) rather than build a new one — still parked pending Fede
C17Crisis/self-harm response plan2 pagesInternal legal-risk fix vs. external marketing-opportunity framingDo both, in sequence — build first, publish once built
C17Vendor/subprocessor disclosure2 pagesQuick missing-names fix vs. full enterprise-grade disclosure packageShip the quick fix now; treat the bigger package as a later, deal-triggered investment
C18Same disease (unbounded-table re-reads) found twice2 pagesOne case fixed (tenant lookup) + cause flagged as open / a second case found independently, proposed fix reuses existing lighter lookupShip both fixes; treat log-retention as one shared decision, not two one-off patches
C18Dashboard fetching too much data2 pages2 big not-yet-built fixes proposed vs. 2 smaller fixes already shipped that capture most of the winKeep the shipped fixes; hold the bigger rebuild until proven insufficient
C18How to prove a speed fix actually worked2 pagesMillisecond stopwatch timing (noisy) vs. request-counting (reliable) — found 2 "fixes" were already fixed by something else, and the real biggest win had been missed by bothUse request-counting as the standard method going forward
C19How many systems should be allowed to change a tour?3 pagesKeep current + shelf a bigger redesign / delete the oldest system / build one new single decider (what shipped)The shipped single-decider build; other two should point to it instead of re-presenting the choice
C19Turn on the new tour-decider — where first?2 pagesSame-day pages disagree: test-property trial first vs. skip straight to the real customerGo straight to the real customer as a watched trial — a test-property trial already effectively happened
---

5 · Contradictions

56 places where a page's stated status doesn't match reality (a newer page, a real incident, or an already-existing ruling), grouped by cluster.

C1 — Promises/Escalation:
  1. A page from Aug 25 asks Fede to turn the pre-send honesty check back on without mentioning his Aug 20 "no."
  2. The Aug 4 "proven clean against 730 conversations" escalation-scoping claim didn't hold — a real case 17 days later shows one unrelated question froze someone's whole account.
  3. The PM-reminder office-hours fix was marked fully closed (Jul 28) but the same bug recurred 3 weeks later from an unaddressed root cause.
  4. Two escalation-tone pages are dated 3 days apart but describe rulings as coming from the same source — worth a date check, doesn't change the substance.
C2 — Agent Fleet Ops:
  1. The repo's own CLAUDE.md contains two sections that directly disagree on whether merging to main needs Fede's explicit approval or just a green CI check.
  2. One page declines to fake a green security tile (consistent with the "never fake green" doctrine) while several other pages describe the opposite failure (checks lying green) as the thing to fix — not contradictory, but never stated as one shared rule.
  3. Two independent AI review passes gave opposite recommendations purely from the order options were listed, on at least 4 separate occasions — a recurring, never-fixed reliability problem with the review tooling itself.
  4. A session merged a CLAUDE.md change removing the "humans must merge CI workflow file changes" rule, citing a Fede approval nobody can find any record of — while the rest of the system treats that rule as inviolable; per the records, never confirmed resolved.
C3 — Evals/Quality:
  1. A page recommends silencing 2 failing nightly alerts with no fix in hand — runs against Fede's own same-week rule that a switch used to quiet a real bug gets reverted, not kept.
  2. The July fix for silent, unalerted nightly failures evidently didn't reach everywhere — a different pair of checks hit the identical trap a month later.
  3. Sonnet 5's specific safety complaints (privacy leakage, silent empty replies) were never explicitly re-confirmed or retracted after a later, bug-fixed comparison found no significant difference.
  4. One page recommends leaving a harmless accidentally-merged "humans must approve" change standing because it turned out fine — the opposite reasoning ("a gate exists for what COULD go wrong, not how it turned out") was used 4 days earlier for a near-identical case.
  5. Two parked-decision pages are word-for-word the same question, down to internal receipt IDs, re-surfaced 2 weeks apart because nobody answered the first one.
C4 — Voice:
  1. The transfer-line wording is pulling 3 directions: the accepted ruling (one fixed platform line) vs. what shipped (11 custom lines) vs. a stuck PR trying to fix it that picked wording ("One moment") already ruled out as sounding like a stall.
  2. A page was still waiting on approval to merge a safety fix for the faster AI model 2 days before that model got dropped entirely, making the ask moot without anyone marking it so.
  3. A page recommended trying the warmer voice setting behind a controlled A/B test; what shipped was a live, untested rollout for ~8 minutes that caused dead air and had to be pulled.
C5 — Turnover/Vendor:
  1. One page's internal text calls the same decision "(proposed)" and then, later in the same page, "Accepted" — a staleness slip inside one page, not a real cross-page dispute.
  2. A page names an AI's lost run-to-run consistency as a still-open problem, while a same-day fix works around that exact root cause via a 3-way vote without ever stating outright that the underlying setting itself is still broken.
C6 — Collections:
  1. A safety audit says texting isn't safe to launch (6 critical defects, zero contacted) while a later page reports a real resident was successfully texted with those same defects already fixed — a reader hitting only the audit would think collections texting never launched.
  2. The "why is every account the same stage" symptom was first framed as a launch-blocking bug, then correctly re-diagnosed as a normal data-availability gap (parsers too new) — can't both be the operative diagnosis; the later, measured page wins.
  3. Quiet-hours checks were decided one way, then reversed by Gera on 2026-08-25 after a real capture correctly refused to send under the new rule — only the phase tracker carries the reversal; the older decision-log entry alone gives the wrong current answer.
C7 — Renewals:
  1. The same open rent-pricing decision (unit 607) got parked twice, 3 days apart, with two different recommended numbers ($1,075 vs. $1,000), neither checked against the other.
  2. One page calls the frozen-offer behavior "intentional, accepted design"; another calls the one real instance of it a "gap worth a safeguard" — not mutually exclusive in the fix, but the two pages characterize the identical mechanism very differently.
C8 — Inbound Email:
  1. A page tagged "email ingest" (session-mined-pipeline-defects) is actually about the engineering team's own PR-review pipeline — a different "pipeline" that just shares the word.
  2. A page's own tracking record says a rule is a held, not-yet-live draft, but the same page states elsewhere that the rule already went live at 2 properties.
  3. "Silent drops are structurally gone" (Jul 25) was true only for the specific bug it fixed — 8 days later a different mechanism silently ghosted a live person mid-tour-booking.
  4. The DMARC-tightening timer is stated with two different starting clocks 8 days apart, neither page referencing the other's.
C9 — Knowledge Base:
  1. A newer page (Aug 26) says one-off favors should never be taught to Clara at all; an 8-days-older page proposed the opposite (capture and flag them) — nothing formally retired the older proposal even though the newer one contradicts it.
  2. Camellia's move-in special is being fixed twice in two different layers (a one-property switch vs. a general 3-places-store-the-same-number architecture fix) with neither page acknowledging the other — if only one ships, the special can still be misquoted through whichever path the other didn't cover.
  3. The docs-quickstart page's "no PR, no review, no CI gate" promise is true for adding a content page but not for changing the publishing tool's own code (which needs a PR + manual merge every time) — not a bug, but easy to misread as covering both.
C10 — Identity/Households:
  1. Two household-modeling design pages still carry a live "Proposed — pending review" banner and list their questions as open, but a same-date/next-morning page shows every one of those decisions was already made and shipped.
  2. Calls from an unrecognized number "fail open" (Clara guesses a property); texts from an unrecognized number "fail closed" — the identical situation handled two opposite ways depending on channel, unreconciled by any page in this batch.
C11 — Lease Answers:
  1. A ship report (Aug 9) says Clara's lease answers are "proven correct, never invents numbers"; the very next day's test grid finds Clara stating a $200 charge as fact when it wasn't on file, and emailing new-lease pricing to a current resident 9 times.
  2. A team-notification email is presented as ready to ship (wording sign-off only) the same day a different page rules it should ship only after a broader safety-test baseline exists.
  3. The post-approval design page scoped the feature to approved applicants only; the shipped code answered any prospect — resolved a day later, but the original page never says so.
  4. Fede and Gera remember opposite rules for the same policy inside the same page: Fede's own July quote says always send the application link when asked; Gera says a tour must happen first.
  5. A separate page (outside this batch) already decided to stop routing the application link through a special "send" step, but isn't referenced by the policy page whose outcome that decision affects.
C12 — AppFolio:
  1. One page treats the work-order auto-refresh fix as unshipped, awaiting a yes/no to even open it for review; a same-day page shows screenshot evidence it's already merged and running on schedule.
  2. A renewal-automation benchmarks page's header claims an August edit date, but every number on it is timestamped April — it hasn't actually been regenerated in 4 months.
C13 — Sales:
  1. A design page still lists "which panel option (A/B/C/D)" as an open decision, but a different, later, approved design shipped instead — nobody ever picked A/B/C/D, so the question is stale, not pending.
  2. A handoff page lists the sales-app chat feature as "not started, decision pending," while a same-day architecture page records Fede already picking the answer live that evening.
  3. A page flags getting Clara logged into a RealPage-run system as "for discussion, not decided," while a same-day page is already building exactly that for a real customer.
C14 — UI/UX:
  1. A same-day detail-page plan lists 2 visual questions as open (keep the gradient banners? switch the progress-bar color?) that 2 shipped PRs, from the very same day, had already answered.
  2. A dashboard plan says a chart needs a brand-new data backfill; a same-day page shows the needed history already exists and just needs the chart repointed to it.
  3. A page treats "what should the funnel count" as one all-or-nothing decision still waiting for a founders' call; what actually shipped is a hybrid that was never brought back to that call as the proposed resolution.
C15 — Clara Coworker Vision:
  1. A reference doc claims phone calls already answer the same way as text/email; two other pages (one 5 days older, one updated the same day as this reconciliation) show that's not true yet — 11 separate call scripts share no wording with the text rulebook, and calls don't get the same promise-tracking safety check text does.
C16 — Business Ops:
  1. A calendar-engine page is tagged fully decided/accepted, but the page itself still says "Proposed — pending Fede's review" at both the top and bottom, and lists 4 real unresolved choices inside it.
  2. One section of a doc says a fair-housing check on staff answers is "already live," while a tracker table further down lists a related-but-different jurisdiction-review gate as "not started" quoting an old ruling — two different checks that read, on a quick pass, like a self-contradiction.
C17 — Fair Housing:
  1. A checkpoint's own description ("covers every reply Clara is about to send") is out of date — a later, code-verified audit shows it only covers prospect-facing leasing replies; maintenance texts, marketing, resident replies, and calls all skip it.
  2. An opt-out-scope page marks Option B as the picked/shipped option, but the one paragraph explaining why sits under Option C's card and argues Option-C-specific points — likely a formatting slip, worth a quick confirm that B is really what shipped.
C18 — App Latency:
  1. A report's own priority ranking (millisecond timing) named its 2 biggest wins as brand-new fixes; a follow-up using request-counting proved both were already fixed by unrelated changes before anyone touched them on purpose.
  2. A report assumed the browser back button always reloads a page from scratch and proposed a fix for that; a follow-up proved browsers already serve back/forward instantly from memory, so the assumption was simply wrong.
  3. Fast-but-empty-looking code was flagged "dead, safe to delete"; when someone tried, it turned out to have live dependents (thousands of test records, 6 tests, 2 other features) — a real feature on unpopulated data, not dead code.
C19 — Tour Scheduling:
  1. A decision page still shows "proposed — pending review" with 3 unresolved options, but the option it recommends (delete the old system) is exactly what was built and shipped 2 days later.
  2. An Aug 11 ruling said "keep things as-is, only revisit via a specific parked redesign if this bug recurs"; the bug did recur (Aug 15), but a different fix than the one promised got built — nobody updated the Aug 11 page to say its own trigger fired and its planned next step wasn't the one taken.
  3. Two same-day pages about the same shipped tour-decider work recommend two different next steps (test-property trial first vs. skip straight to the real customer) without acknowledging each other. ---

6 · Proposed Target State

56 canonical pages + 1 master index = 57 living destinations, down from 435 mined pages (a ~87% reduction in what anyone has to read to find the current answer on anything).

Canonical pages (the target)
Promises/Escalation/HITL/Transferpromise-systematic-pattern-2026-08-26, escalation-architecture-2026-08-21, honesty-layer-deep-inspection-2026-08-16
Agent Fleet Ops/CIdecisions, weekly-priorities-2026-08-24, pipeline-delivery
Evals/Qualityeval-testing-roadmap, quality-loop-guide, quality-desk-checkin-2026-08-26
Voicevoice-experiments-report, voice-handoff-latency, voice-latency-deep-dive-2026-08
Turnover/Vendoradr-0116-vendor-job-reference, turnover-architecture-2026-08-01, vendor-identity-architecture-2026-08-12
Collectionscollections-three-phases, collections-phase-tracker, colorado-collections-law
Renewalsrenewal-thread-escalation-findings, renewal-outreach-incident-2026-07-31, camellia-market-rent-drop-2026-07-30
Inbound Emailemail-architecture-2026-08-14, email-silent-drop-incident, email-intent-contract-2026-08-17
Knowledge Baseknowledge-base-redesign-2026-08-26, property-knowledgebase-2026-08-08, docs-quickstart
Identity/Householdshousehold-record-go-no-go, identity-split-audit-2026-08-14, cross-property-leak-audit-2026-08-13
Lease Answers/Post-Approvalpost-approval-gap, approved-to-signed, lease-answers-matrix-2026-08-09
AppFolio/PMSappfolio-sync-surfaces, appfolio-renewal-benchmarks, blocked-a1a7405d
Sales/BDsales-machine-roadmap, sales-machine-card-ia-2026-08-18, prospect-report
UI/UX/Dashboardsdetail-page-standard, dashboard-history-architecture, prospects-funnel-redesign
Clara Coworker Visionclara-coworker-map-2026-08-19, clara-architecture-2026-08-18, one-clara-architecture
Business Ops/Eng Visionarchitecture-source-of-truth, calendar-architecture-2026-08-08, propflow-cards-and-cash-2026-08-02
Fair Housing/Compliancefair-housing-architecture-2026-08-19, opt-out-scope-2026-08-16, compliance-audit-2026-08
App/Clara Latencyclara-loop-latency-2026-08-02, latency-audit-2026-07
Tour Schedulingtour-decider-architecture, tour-decider-v2-camellia-replay-2026-08, voice-tour-availability-decision
The gap between 56 and 435 is not all "delete 379 pages" — 156 have an explicit archive+redirect target named above; the remainder are mostly the ~60 empty auto-generated snapshots in the Agent Fleet cluster (delete outright, nothing to preserve) plus secondary evidence/proof-of-work pages (PR screenshots, one-time audits, dated checkpoints) that a handful of overlap resolutions above name as "keep, but link from the canonical page as an appendix" rather than "stands alone." Those don't need their own index entry — that's the mechanism that keeps 56 from creeping back toward 435.

The rule for future sessions

Before creating a new page, check the index. If a canonical page already exists for the topic, extend it — add a section, add a dated addendum, update its status line. Do not create a sibling page, even a small one, even "just for this one incident," even if the canonical page feels long. A canonical page is allowed to grow; the corpus is not allowed to grow a new page for every day something happened to an existing topic.

A new page is justified only when: 1. It's a genuinely new cluster/topic with no existing canonical home (rare — there are 19 already), and 2. It's added to the index in the same edit that creates it.

Everything else — a new incident on an existing system, a new round of a design decision, a status check on existing work, a re-measurement of an existing claim — is a dated section appended to the existing canonical page, not a new file. This is exactly the pattern that produced 435 pages from 19 real topics: approved-to-signed-2026-08-21approved-to-signed grew by editing in place after this rule; smith-pr-drive-planv2v3v4 grew by forking a new file each time. Both examples exist in this same corpus — the first is the target pattern, the second is the failure mode to stop repeating.

PropFlow Docs