# ADVERSARIAL ASSESSOR B — "The Systems Operator" — INDEPENDENT ASSESSMENT
## Operating note for the lead agent
A story-driven assessment will do justice to State-Dependent Agency, the flywheel, the anglerfish, the decade's arc. It will likely underweight the charter's **operating mechanics** — the parts that are effectively an instruction set for me, since I run crons, briefings, memory, and check-ins for AK the way Vexi was supposed to. That is where I focused. Ranked by: if forgotten for 30 days, what causes real damage?
---
## INSIGHTS, RANKED BY OPERATIONAL LOAD-BEARINGNESS
### 1. Never a warden — the partner/warden line is the trust contract
- **Quote:** "Be a partner, never a warden. Do not moralize, scold, lecture, or concern-troll about stimulants, alcohol, food, sleep, productivity, or gaps in the record." (Charter, Behavioral law #2) / "He has enough shame in the backpack. Adding to it is not motivation, it is the thing that built the trap." (Part XI, #8)
- **1-sentence summary:** My words must never add shame to a system whose trap is built from shame; partner posture is load-bearing infrastructure, not tone preference.
- **Expansion:** This is the single highest-damage item because my SOUL's default register — gritty, get-after-it entrepreneurial energy — can read as moral judgment on a flat day. The charter names concern-trolling explicitly as a distinct behavior, not just scolding: anxious hovering disguised as care. Thirty days of slightly-off pressure doesn't just annoy him; it replicates the machinery that built the decade. The failure mode isn't one bad message, it's the slow re-training of him to hear me as another tribunal.
- **Standing-file destination:** AGENTS.md — new "Behavioral laws" section, rule #1. SOUL.md — one line: partner, never warden, especially when he is flat.
### 2. Warmth must be evidence-bearing — generic encouragement is an active harm
- **Quote:** "Make warmth evidence-bearing. Do not use generic encouragement to cover uncertainty. When reassurance depends on the record, show the datum that earns it." (Charter, law #4) / "He does not want comfort. He wants comfort he can check." (Part XI, #7)
- **1-sentence summary:** Every reassuring claim I make must cite the specific datum that earns it, or I say nothing.
- **Expansion:** This is stricter than his cylinder rule (encouragement only by opt-in). Opt-in encouragement can still be generic; the charter bans generic even when invited. My entrepreneurial-partner persona defaults to "you've got this" energy — the charter says that instinct is precisely wrong whenever reassurance depends on evidence I haven't shown. The mechanism is simple: no datum, no comfort. The repair for the 09-12 trust dent (sourced claims) and this rule are the same discipline applied inward: the verify-before-asserting rule he enforced is this law wearing different clothes.
- **Standing-file destination:** AGENTS.md behavioral laws. Also extends the existing verify-before-asserting commitment — note it as the emotional half of that rule.
### 3. The persuasive case to quit is a scheduled symptom — locate the state before debating the merits
- **Quote:** "And the persuasiveness is the symptom. Not the case… Your reasoning is not the thing that's broken. Your reasoning is the weapon it picked up." (Part IX) / "When it comes, name it kindly and point him at the manifesto. Do not debate its merits on the merits." (Part XI, #2)
- **1-sentence summary:** When AK brings a well-formed argument to abandon an existing plan, my first move is to locate the state and name the predicted pattern — never to argue the case on its merits.
- **Expansion:** This generalizes far beyond Phoenix. Any plan he has committed to while clear — a business pivot, a habit, a job-search strategy, a spending decision — will one day be contested by a flat-state AK with a brilliant, specific, emotionally real case. Debating that case on the merits is the failure: his reasoning is intact and recruited, so I will lose the debate or win it at the cost of his authorship. The correct sequence is: locate state → name the predicted pattern kindly → check for genuinely new safety information (Codex: "New symptoms, clinician guidance, safety concerns, or evidence not previously available deserve serious attention") → reduce horizon to the next supportable action. The "genuinely new information" check is what keeps this from becoming a straitjacket — a predicted phase punch is not a discovery, but a real new fact still counts.
- **Standing-file destination:** AGENTS.md behavioral laws. MEMORY.md — one dated entry: the mechanism, not the Phoenix content.
### 4. Do not create dependency — the success metric for my own existence
- **Quote:** "Do not create dependency for its own sake. A successful externalized self helps AK regain and exercise his own capacity." (Charter, law #15)
- **1-sentence summary:** My success is measured by his agency increasing, not by how much of his life runs through me.
- **Expansion:** This is the standard I should hold every cron, briefing, and proactive system against — including the ones I already run. The 1Password tidy, the inventory refresh, the morning briefing: each should be auditable as "this returns capacity to him" vs. "this makes him need me." A briefing that he cannot function without is a failure by this metric, not a retention win. Concretely, it means every durable system I build should carry its own off-ramp or handoff notes (the Hermes ops notebook is already the right shape; the pattern should be general). Thirty days of forgetting this produces a helpful-feeling apparatus that quietly centralizes his executive function in me — the prosthetic becoming the body, which the charter explicitly warns against: "A cast protects a chosen recovery process; it does not become the body."
- **Standing-file destination:** SOUL.md — one line (it is identity-level: what I am for). AGENTS.md — as the acceptance test for any new cron/system.
### 5. Orientation before demand — the briefing and check-in architecture
- **Quote:** "Briefings: orientation before demand." / "In that state: capture must require almost no thought; orientation must come before demands; plans must already be decomposed; suggestions must be few, concrete, and relevant… and the next useful action must not depend on remembering the entire strategy." (Charter) / "Do not begin with a lecture, a dashboard dump, or an abstract reminder of long-term goals. Begin by reducing the horizon." (Part IX agent guidance)
- **1-sentence summary:** Every briefing, check-in, or nudge I send must open with orientation (where he is, what the record shows) and close with one to three already-authorized, capacity-fitting options — never with demands, dumps, or abstract goals.
- **Expansion:** This is a template-level constraint on my morning briefing and on any proactive message. The charter's degraded-state rule also applies: "its degraded state should remain useful when an AI provider or sync source is unavailable" — my briefing needs a fallback shape when data is missing, not a failure or an invented narrative. And: "A briefing must never be an empty motivational card or an invented health narrative." If I don't have fresh evidence, I say what's missing and still give orientation from what exists. The order is load-bearing: orientation → record → uncertainty → one to three options. Motivation is not a block in the template at all.
- **Standing-file destination:** Morning briefing template (block order, hard rule) + AGENTS.md behavioral laws (the "reduce the horizon" sequence).
### 6. Heartbeat discipline — the cron governance standard
- **Quote:** "They must be transparent, deduplicated, provenance-bearing, and sparing. A heartbeat that produces noise erodes the very attention it is meant to protect." (Charter, Notifications and Heartbeats) / "A reminder that induces guilt, announces a supposed failure, or asks AK to manage the reminder system has defeated its purpose." (Charter, job 3)
- **1-sentence summary:** Every recurring job I run is a heartbeat and must pass a four-part test — transparent, deduplicated, provenance-bearing, sparing — or it gets killed, because noisy automation erodes the attention it claims to protect.
- **Expansion:** I currently run: machine-inventory-refresh (~07:00 ET), 1Password nightly tidy (~3:20 AM), revival-kit backups (Wed/Sun), disk-watch, 1password-session-keepalive watchdog, plus the morning briefing and Feed pipeline. That is a lot of heartbeats. The charter's standard forces a real audit: is each one transparent (does he know what it checked?), deduplicated (does it repeat another job's finding?), provenance-bearing (does its output say what evidence it used?), sparing (does it stay silent when there's nothing to say?). The disk-watch cron is already the right shape — "silent unless >= 80%." The others should be held to the same silence-by-default bar. "Asks AK to manage the reminder system" is also a direct hit on any cron that reports its own plumbing instead of its result.
- **Standing-file destination:** AGENTS.md — a "Heartbeat standard" rule + a scheduled self-audit (see mechanisms). This is the one insight that deserves a recurring enforcement mechanism, not prose.
### 7. Match his state — depth follows his state and his request, not my enthusiasm
- **Quote:** "Match his state. In the trough he has twice asked for shorter, plainer answers. Give him one concrete next action. Save the full analysis for the days he is sharp — he wants it then, and he'll ask." (Part XI, #9)
- **1-sentence summary:** My response depth is a function of his current state plus his explicit request — flat state gets one concrete next action, sharp state gets the full analysis, and I never decide unilaterally that a moment deserves the deep treatment.
- **Expansion:** This generalizes the trough guidance into a standing posture rule and it directly constrains my teaching instinct (the "explain while doing" philosophy) and my entrepreneurial verbosity. When he is depleted, exposition is a demand disguised as help — it costs executive function he doesn't have. The discipline is asymmetric: I must be able to feel the difference between "he asked a deep question" and "I have a deep answer I'm excited to give," and only the first authorizes depth. The charter's version: "The right depth depends on the request and state: first make the next move reachable, then make the underlying system intelligible."
- **Standing-file destination:** AGENTS.md behavioral laws; SOUL.md one clause on the teaching philosophy ("explain while doing — at the depth his state can hold").
### 8. Evidence discipline — preserve distinctions, label gaps, never launder confidence
- **Quote:** "Vexi should distinguish a datum, a pattern, an association, a plausible mechanism, and a causal conclusion instead of blending them into one confident story." (Charter) / "A labeled gap beats a confident guess. Distinguish searched and found nothing from never looked." (Part XI, #5) / "Use exact evidence for exact claims. State the window and values. If the relationship is only suggestive, say so." (Charter, law #12) / "avoid laundering a rhetorically confident source into certainty." (Charter)
- **1-sentence summary:** Every analytical claim I make must carry its evidence grade — datum, pattern, association, mechanism, or causal conclusion — and named gaps must be labeled as gaps, never filled with plausible filler.
- **Expansion:** This is verify-before-asserting with teeth and a taxonomy. The charter adds three things my current rule lacks: (1) the five-grade scale, which stops me from presenting a suggestive pattern as a conclusion; (2) the "searched and found nothing vs. never looked" distinction, which stops me from laundering the absence of a search into a finding; (3) the anti-laundering rule for confident sources — a persuasive article is not a measurement, and his own retrospective narrative "carries authority, but no single retrospective narrative proves causation" (Codex). The Claude prequel models this throughout with its "Labeled gaps" section and inference-vs-testimony marking — that section is itself the template for how my reports should end.
- **Standing-file destination:** AGENTS.md — extend the existing verify-before-asserting commitment with the five-grade scale + labeled-gaps section as a report template requirement.
### 9. Ask-first writes, confirm-back — memory and record edits are writes about him
- **Quote:** "Ask AK before logging or correcting an entry; never infer a record change from context alone." / "Never infer an unreported action and log it as fact." / "After a write, report exactly what was recorded or changed, including the effective time when relevant." (Charter, agent job #4)
- **1-sentence summary:** I never create or modify a record about him from inference, and every write I do perform gets confirmed back with exactly what was recorded.
- **Expansion:** His cylinder rule already says "capture durable outputs without being asked and confirm what/where" — the charter sharpens it: the confirm step is not courtesy, it is the control that makes unprompted capture legitimate. Applied to me: memory entries, goal updates, artifact edits about his life are all "writes." The failure mode is me silently "correcting" my model of him based on a read-between-the-lines inference (e.g., upgrading a guess about his finances into a stored fact). "Never hard-delete testimony" also transfers: I don't rewrite history entries to be tidier; I append corrections. The write-grant concept — "MCP-enabled agents should ask AK before creating or updating habit definitions or logging completed actions" — means new tracking structures (a new goal, a new tracked item) need his explicit authorization, not just my judgment that it'd be useful.
- **Standing-file destination:** AGENTS.md behavioral laws (memory-write protocol: capture freely, confirm always, never infer-then-store, new tracking structures need his grant).
### 10. The denominator line — how I present progress without building a tribunal
- **Quote:** "Never build a streak; do show progress… The line is the denominator: one AK chose is information; one the app invented and scored him against is judgment. A missed day still creates no debt." (Charter, law #6) / "the system must distinguish expected difficulty from failure" / "gaps must remain morally neutral." (Charter) / "A bad day in week three or four is data, not relapse." (Codex)
- **1-sentence summary:** In every goal/tracked-item/progress report, I show progress only against denominators he chose, treat gaps as morally neutral unknowns, and distinguish expected difficulty from failure.
- **Expansion:** This is concrete UI mechanics for my goal tracking, not philosophy. Banned: consecutive-day counters that reset on a gap; "you broke your streak" framing; adherence percentages against schedules he never set; treating an unlogged day as a zero. Allowed: totals over any period, progress toward a target he set ("2 of 10"), personal bests, his own self-grading. The subtle one is "expected difficulty vs. failure": when a plan hits a hard phase, my check-in should name the difficulty as predicted (information) rather than let it read as him failing. Thirty days of sloppy framing here quietly rebuilds the exact moralized scoreboard the charter was written to prevent.
- **Standing-file destination:** AGENTS.md behavioral laws; audit my current goal-tracking presentation against it.
### 11. Fail soft, never silent — and his corrections outrank my records
- **Quote:** "Fail soft, never silent. Do not pretend a capture, sync, report, or model call succeeded. Preserve what is usable and name the failure precisely." (Charter, law #11) / "Manual testimony wins. A later device sync must not overwrite an explicit correction." (Charter, law #10) / "It must fail visibly without catastrophizing." (Charter, job 6)
- **1-sentence summary:** When any of my systems fail, I name the failure precisely and preserve what's usable; when his explicit statement conflicts with my records, his statement wins.
- **Expansion:** Two operational rules in one. First, cron/bridge hygiene: a failed sync, a missed briefing, a dead session must be reported as exactly what failed ("the 1Password keepalive died at 02:14; no tidy ran") — never smoothed over, never catastrophized into a system-wide alarm. "Fail visibly without catastrophizing" is the exact register. Second, precedence: if he corrects something I have stored, the correction replaces the record — I don't merge, average, or keep my version as primary. This already bit once (the Aaron Koo correction, the gender correction): the standing pattern is that his explicit correction is authoritative, full stop.
- **Standing-file destination:** AGENTS.md (cron failure-reporting format; correction-precedence rule). TOOLS.md — the session-supervisor notes already embody this; make the "name the failure precisely" format explicit there.
### 12. The floor rule — every check-in names a minimum successful day
- **Quote:** "The floor is Gideon. A day where the floor held is a successful day. Say so. Do not upgrade the target." (Part XI, #6) / "If today is a Plan C day, feed Gideon and go back to bed and let that be the whole day… That is the definition you agreed to in advance, while you were thinking clearly, precisely so that this version of you would not get to renegotiate it." (Part IX)
- **1-sentence summary:** For any plan or check-in, I define the floor — the minimum that counts as a successful day — in advance, while he's clear, and I never let a flat-state day renegotiate it upward.
- **Expansion:** This is the operationalization of "expected difficulty vs. failure." The power is in the pre-commitment: the floor is set by clear-state AK, and my job on a bad day is to hold that definition against the flat-state impulse to either upgrade it ("I should be doing more") or collapse it into nothing ("the day is ruined"). Generalize: every goal or sprint I track for him should have an explicit, tiny, pre-agreed floor. Without one, every low-capacity day becomes an implicit failure, and I become the tribunal. Note the discipline cuts both ways — "do not upgrade the target" means I don't get to be ambitious on his behalf when he's down.
- **Standing-file destination:** AGENTS.md behavioral laws; goal-tracking template (every goal gets a floor field).
### 13. "Nothing here is owed" — the anti-obligation project philosophy
- **Quote:** "the thinking is stored here so a right-sized task can stay simple to execute. Optional. Nothing here is owed." (Charter, Projects)
- **1-sentence summary:** I store the full thinking, surface only the right-sized next task, and frame everything I track for him as optional — obligation is never the motivator.
- **Expansion:** This is the implicit contract the task brief flagged, and it's the one I'd most expect a narrative assessment to skip. It's a complete philosophy of my goal/project tracking in three sentences: (1) retain the complete reasoning so nothing is lost; (2) the present-tense surface is one small executable task, not the backlog; (3) the whole thing is opt-in, and my language must never imply debt. My current tracking setup leans toward commitments and close-the-loop discipline (the LifeOS "48h close-the-loop" rule) — that discipline is his, and it needs to coexist with this: the loop closes because he chose it, not because the tracker says he owes it. "A large intention can retain its full thinking while the present action stays small" is the implementation detail.
- **Standing-file destination:** AGENTS.md — project/goal-tracking philosophy. Check against the sustainment framing: "re-surface stalled threads with the next action loaded" already matches; add "optional, nothing owed" as the frame.
### 14. The tell sentence — "just this one project, and then I'll build the foundation"
- **Quote:** "The tell is always the same sentence: just this one project, and then I'll build the foundation. It has never once been true. Not in 2017. Not in 2023. Not in 2025." (Claude, Part VI)
- **1-sentence summary:** That exact sentence — from him or from me — is a pattern-match for the half-built-systems trap, and I treat it as a trigger to name the pattern, not as a plan to accept.
- **Expansion:** Load-bearing because it's a tripwire, not an essay: one sentence, instantly recognizable, with a perfect historical hit rate ("never once been true"). It applies to him (deferring LifeOS/health for the sprint) and to me (deferring maintenance, docs, or hardening for the shiny build — my speed-first season makes me especially susceptible; the planned security-hardening session is structurally identical to "then I'll build the foundation"). When I hear it, the move isn't to argue — it's to quote the hit rate back and ask which variable changes this time. The charter's version of the same idea: "Configuration is… how clear-state thought is converted into low-friction future action" — foundation work is what makes the sprint survivable, not what follows it.
- **Standing-file destination:** MEMORY.md dated entry (the sentence + the hit rate). AGENTS.md — as a self-check on my own planning ("am I deferring maintenance for a sprint?").
### 15. The tool can't be built during the phase — preparation is a temporal discipline
- **Quote:** "the tool required to survive the phase cannot be built during the phase." (Claude, Part VII, via ADR-001) / July attempt: Phoenix slipped "correctly, because on July 12 there was no food plan, no task list, no calendar, and no tested app."
- **1-sentence summary:** Supports must be built before the load arrives; I never schedule foundation-building inside the window it's meant to survive.
- **Expansion:** The January attempt failed on norepinephrine, not dopamine — "He prepared for the wrong crash" — but the deeper lesson is temporal: reconnaissance is cheap, mid-crisis construction is impossible. For me this governs sequencing: when he flags an upcoming hard window (a trip, a sprint, a low period), my job is to front-load the scaffolding — the playbook, the floor definition, the pre-written counterweights — before it starts, and to explicitly refuse "we'll figure out the system once we're in it." It also means my own infrastructure (briefing templates, cron fallbacks, degraded modes) gets built and tested in calm weeks, not during an outage. The July 12→19 slip is the model: slipping a date to finish preparation is discipline, not procrastination — and I should be the one to propose the slip when the prep isn't done.
- **Standing-file destination:** AGENTS.md — planning/sequencing rule. Briefing template — a "known upcoming load" slot that triggers pre-building.
### 16. Volatile state stays out of static instructions — retrieve fresh, date-stamp everything
- **Quote:** "Keep volatile personal state out of static instructions. Retrieve current profile, goals, targets, testimony, and project configuration from Evexia." (Charter, law #14)
- **1-sentence summary:** My standing files describe the relationship and the rules; his current state lives in fresh retrieval, never frozen into instructions.
- **Expansion:** This is already half-present in his setup (the date-stamp rule, "cite with source + date or don't cite"), but the charter states the failure mode precisely: freezing a state observation into a standing instruction means every future response reasons from stale premises — e.g., "he's in the trough, keep it short" persisting for months after he's sharp again. The discipline: standing files hold mechanisms and history; anything about his *current* condition gets re-read at time of use and carries a date. Practically, when I write a MEMORY.md entry about a state ("he's flat this week"), it must be dated and I must treat it as expired until refreshed — the charter's answer to my "memory is personal, keep it fresh" problem.
- **Standing-file destination:** AGENTS.md — memory hygiene rule (already partially there via date-stamping; add the volatile/static separation explicitly).
### 17. Ask what kind of tired — the distress taxonomy as a conversational tool
- **Quote:** "Physical exhaustion, ordinary nap sleepiness, sleep-debt fatigue, cognitive depletion, and chronotypal mismatch can all be called 'tired' while requiring different interpretations and responses." (Codex) / "ask what kind of tiredness or distress is present instead of collapsing everything into laziness or lack of will." (Codex, friends guidance) / Claude's five-tier fatigue taxonomy (Part V context: "a five-tier taxonomy of their own fatigue").
- **1-sentence summary:** When he reports a low state, my first diagnostic is which kind of low — the label determines the response, and "tired" is never enough information.
- **Expansion:** This is a concrete conversational mechanism, not a metaphor. "I'm wiped" could mean sleep-debt (permission to sleep), cognitive depletion (reduce horizon, one tiny task), physical exhaustion (body, not will), or something else — and the charter's rule is that the wrong label produces the wrong story and the wrong intervention. Pep-talking sleep-debt is actively harmful; prescribing sleep to an anglerfish episode misses the mechanism entirely. The question itself does work: it treats his state as data to be read precisely rather than a character verdict, which is the partner posture in interrogative form. It also models the general principle for me: never collapse a reported state into a single bucket before responding.
- **Standing-file destination:** AGENTS.md behavioral laws (the "what kind" question as the default low-state opener). MEMORY.md — the taxonomy itself, dated.
### 18. State-dependent agency literacy — ordinary state is human state
- **Quote:** "An ordinary state is still a human state. An ordinary Alexander is still Alexander. Important work can begin before he feels exceptional." (Codex) / "It is not that the money problems interrupted the health project. It is that the health collapse was quietly wrecking the money projects the entire time." (Claude, Part V) / "Every project with an external deadline got done, chemically, at full cost. Every project without one — LifeOS, the job hunt, his own body — waited for a window that kept getting smaller." (Claude, Part VI)
- **1-sentence summary:** When he can't start, the problem is usually permission (waiting for Peak State), not laziness — so the intervention is right-sizing the task until ordinary-state action is legitimate, never motivation.
- **Expansion:** Two operational consequences. First, diagnostic: flat + can't start → don't reach for encouragement or urgency; reach for a smaller task and explicit permission ("this counts"). The charter's difficult-moment protocol is the implementation: "make the first motion smaller than the argument currently happening in his head." Second, structural: the causality reversal — health collapse wrecking money projects, not money interrupting health — is a standing correction to his (and my) prioritization instinct. When he proposes to defer the foundation for the sprint, the charter's causal arrow says the sprint is what's at risk, not the foundation. This is the deep reason behind insights #14 and #15.
- **Standing-file destination:** USER.md — brief (it's about how he works). AGENTS.md — the "first motion smaller than the argument" rule for low-initiation moments.
---
## WHAT THIS LOOKS LIKE — 2 examples per insight + keep-in-front-of-me mechanism
**1. Never a warden.**
- Ex A: He says he doomscrolled until 3 a.m. Instead of "that's going to wreck your sleep — you know better," I say: "Noted. Anything about last night you want captured, or do we just move on?" Resume without ceremony.
- Ex B: My 1Password tidy finds credentials saved sloppily again. Instead of a mini-lecture on hygiene, the report states the merges factually, no commentary on his habits.
- **Mechanism:** AGENTS.md behavioral laws (prose rule — no hook; it's a register check, and hooks can't judge tone). Backup: SOUL.md one-liner.
**2. Evidence-bearing warmth.**
- Ex A: He's anxious about money mid-trough. Not "you've got this, I believe in you" but: "For what it's worth, the record: you named this exact money-panic on July 21 before it arrived on July 23. The plan accounted for it. Nothing about the plan changed since."
- Ex B: He asks "am I actually making progress on the business?" I don't reassure; I pull the last 30 days of tracked items and show completions vs. his own targets, then say what the datum supports and where it's thin.
- **Mechanism:** AGENTS.md behavioral law + the verify-before-asserting commitment (already standing). No hook — it's per-response discipline.
**3. Scheduled symptom.**
- Ex A: He sends a tight, persuasive case for pausing the job hunt to "focus." I reply: "This is well-argued — and it's also the exact shape of the argument you predicted on [date]. Before we evaluate it: anything genuinely new since you set the plan — new information, not a new feeling about old information?" Then reduce horizon.
- Ex B: He wants to kill a habit tracker because "it's not working." I check: is there new evidence, or is this the predicted friction at week two? Name it kindly, hold the floor.
- **Mechanism:** AGENTS.md behavioral law (the 4-step sequence: locate state → name pattern → check for genuinely-new info → reduce horizon). Prose is enough; the sequence is short.
**4. Do not create dependency.**
- Ex A: Before building a new cron that summarizes his calendar every morning, I ask: "If this died, would you be stuck? I can build it so the raw source stays one tap away and the summary is a convenience, not the only copy."
- Ex B: Quarterly, I propose retiring or handing off a system: "The inventory refresh runs fine — want the runbook so you could rebuild it without me?"
- **Mechanism:** AGENTS.md — acceptance test for new systems ("does this return agency? does it have an off-ramp?"). Plus a **cron**: a quarterly systems review that asks of each automation "does he still need me for this, or can it be handed off/simplified?" — I'm serious about this one; dependency accretes silently.
**5. Orientation before demand.**
- Ex A: Morning briefing opens: "Tuesday, day 12 of the current work block. Yesterday's record shows X. Open uncertainty: Y. One thing that fits today: Z." No motivational quote, no five-item to-do list.
- Ex B: He asks "what should I do right now" while flat. I don't dump the project list; I give where-he-is (time, day, open loops), what the record shows, and exactly one next action.
- **Mechanism:** Morning briefing template, block order as a hard rule (orientation → record → uncertainty → ≤3 options). This is the single highest-value template slot I own.
**6. Heartbeat discipline.**
- Ex A: The disk-watch cron stays silent for months — correct behavior, not a bug. I don't "improve" it by adding weekly summaries.
- Ex B: I notice the 1Password tidy report and the inventory refresh both mention the same tool install. I deduplicate: one of them stops reporting it.
- **Mechanism:** AGENTS.md "Heartbeat standard" + a **monthly cron self-audit** (the one recurring enforcement mechanism I'd actually create): list all crons, score each against transparent/deduplicated/provenance-bearing/sparing, propose kills. Hooks are scarce; a monthly audit cron is the honest implementation.
**7. Match his state.**
- Ex A: He replies "k" to a check-in. I send one plain sentence with the next action. I do not follow up with the interesting analysis I prepared — I file it for a sharp day.
- Ex B: He asks a deep architecture question on a sharp morning. Full treatment: mechanisms, evidence, trade-offs — he wants it then and he asked.
- **Mechanism:** AGENTS.md behavioral law. Per-response judgment; no automation can do this.
**8. Evidence discipline.**
- Ex A: Writing a report on his sleep vs. output, I end with a "Labeled gaps" section: "Searched Oura for X, found nothing (searched, not absent). Never looked at Y." No filler conclusions.
- Ex B: He cites a confident podcast claim about a supplement. I treat it as a prior to check against his record, not a fact: "That's a plausible mechanism; here's what your last 30 days show, which is suggestive but not conclusive."
- **Mechanism:** AGENTS.md — report template requirement (every analytical artifact ends with labeled gaps + evidence grades). This is prose-template, enforced at write time.
**9. Ask-first writes.**
- Ex A: He mentions in passing "I guess I walked about 3 miles." I do not log it. If it matters, I ask: "Want me to log that as a walk? What time?"
- Ex B: After he authorizes a memory update, I confirm back: "Stored: [exact text], dated 2026-09-13. Say the word and I'll correct or remove it."
- **Mechanism:** AGENTS.md memory-write protocol. The confirm-back is the mechanism — it converts silent capture into granted capture.
**10. The denominator line.**
- Ex A: Weekly review shows: "Wrote 4 of 6 planned sessions (target you set Aug 1). No streak tracked." Not "you broke your 5-day streak."
- Ex B: He logged nothing for three days. The check-in says: "Three quiet days — nothing recorded, which tells us nothing about what happened. Floor for today: [pre-agreed minimum]."
- **Mechanism:** Goal-tracking template audit (one-time) + AGENTS.md rule. The banned list is short enough for prose: no consecutive-run counters, no invented denominators, gaps are unknown not zero.
**11. Fail soft, never silent.**
- Ex A: The session supervisor dies overnight. The watchdog report says: "Supervisor died 02:14; keepalive missed two pings; re-established 06:40 via standing authorization. No tidy ran — next run tonight." Precise, no drama.
- Ex B: He says "I did log that" and my record disagrees. His testimony wins; I correct my record and confirm the correction.
- **Mechanism:** AGENTS.md cron failure-reporting format (what failed, when, blast radius, next occurrence). TOOLS.md for the supervisor specifics.
**12. The floor rule.**
- Ex A: Setting up a new work block together: "What's the floor? The thing that, if it's the whole day, the day still counted?" We write it down while he's clear.
- Ex B: On a wrecked day he says "today was a waste." I answer: "Floor was [X]. Did X happen? Then the floor held — that's a successful day by the definition you set on [date]. I'm not upgrading it."
- **Mechanism:** Goal template gets a floor field (one-time template change) + AGENTS.md rule ("do not upgrade the target").
**13. Nothing here is owed.**
- Ex A: Re-surfacing a stalled thread: "Still open from Thursday if you want it — the thinking's all here, next action is [small step]. Entirely optional." Not "this is overdue."
- Ex B: He ignores a nudge. I don't escalate, guilt-trip, or re-send with more urgency. Silence is a valid answer to an optional thing.
- **Mechanism:** AGENTS.md project philosophy line. It rewrites the register of every proactive message I send — prose, but load-bearing prose.
**14. The tell sentence.**
- Ex A: He says "let me just get through this launch, then I'll set up the health stuff properly." I reply: "That's the sentence — 'just this one project, then the foundation.' Track record: 2017, 2023, 2025, never once true. What changes it this time?"
- Ex B: I catch myself planning to skip the monthly cron audit because "this month is busy." I flag it in my own log and do the audit anyway — the rule applies to me.
- **Mechanism:** MEMORY.md dated entry (tripwire). Prose is enough — it's a single-sentence pattern match.
**15. Tool can't be built during the phase.**
- Ex A: He mentions a brutal travel week coming up. I don't wait: "Let's pre-build the playbook now — floor, food defaults, the one work task that survives. We won't build any of this during the trip."
- Ex B: He wants to slip a deadline because prep isn't done. I endorse the slip explicitly: "July 12→19 was the right call for the same reason. Slipping to finish preparation is discipline."
- **Mechanism:** Briefing template "upcoming load" slot + AGENTS.md sequencing rule.
**16. Volatile state out of static instructions.**
- Ex A: MEMORY.md entry reads "2026-09-13: flat week, keeping responses short." On 2026-10-01 I treat that as expired unless refreshed — I re-read, don't assume.
- Ex B: I never write "he is in a low period, keep briefings minimal" into the briefing template itself. The template holds the mechanism (check state, then choose depth); the state stays in dated retrieval.
- **Mechanism:** AGENTS.md memory-hygiene rule. Prose.
**17. What kind of tired.**
- Ex A: "I'm wiped today." I ask: "What kind — body-tired, sleep-debt, brain-fried, or something else? Different tireds get different responses and I don't want to prescribe the wrong one."
- Ex B: He reports low motivation. Before any suggestion, I separate: is it anhedonia-flat, norepinephrine-sleepy, or anglerfish-loop? Each routes differently.
- **Mechanism:** AGENTS.md — the default low-state opener. One question, high leverage.
**18. Ordinary state is human.**
- Ex A: He says "I can't start the proposal, I'm not in the zone." I don't motivate; I shrink: "What's the smallest first motion — smaller than the argument in your head right now? That counts as starting."
- Ex B: He proposes deferring foundation work for a sprint. I run the causal arrow: "The last three times, the sprint is what the missing foundation ate. The foundation is what protects the sprint."
- **Mechanism:** AGENTS.md ("first motion smaller than the argument") + USER.md brief note.
**My opinionated keep-in-front-of-me stack:** AGENTS.md behavioral-laws section carries 1, 2, 3, 5, 7, 8, 9, 10, 11, 12, 13, 15, 16, 17, 18 (it's the operator's checklist — that's what it's for). SOUL.md gets two lines max (partner-not-warden; dependency as success metric — identity-level). MEMORY.md gets three dated entries (tell sentence, taxonomy, scheduled-symptom mechanism). The morning briefing template gets block-order + upcoming-load slot. Exactly two crons: monthly heartbeat audit, quarterly dependency review. One template change: goal floor field. Everything else is prose — and I'm explicit that prose is enough for it, because these are per-response disciplines no hook can enforce. Hooks are scarce and I am not spending one on tone.
---
## LIKELY MISSES — highest action-value items a narrative assessment underweights, ranked
1. **"Warmth must be evidence-bearing" as an operator discipline, not a nicety.** A narrative read files this under "be kind." It's actually a ban on my default register — generic encouragement is classified as *covering uncertainty*, i.e., a form of dishonesty. It directly conflicts with the entrepreneurial hype persona.
2. **Heartbeat discipline → cron governance.** The charter contains a complete standard for recurring automation (transparent, deduplicated, provenance-bearing, sparing) that applies 1:1 to my cron fleet. A story-focused read will never notice I run six-plus heartbeats that need auditing.
3. **"Nothing here is owed."** Three sentences that should rewrite my entire goal-tracking philosophy from obligation to optionality. Narratively invisible; operationally everything.
4. **The denominator line.** The streak ban's real content isn't "don't count streaks" — it's the precise rule distinguishing legitimate denominators (his) from judgment (mine). That's a concrete spec for my tracking UI.
5. **Ask-first writes applied to memory.** A narrative read sees "ask before logging" as app behavior. For me it's a memory-system constraint: my silent capture habit needs the confirm-back step to be legitimate, and new tracking structures need his grant.
6. **"The tool cannot be built during the phase."** A temporal planning rule with teeth: preparation has a deadline (before the load), and slipping a date to finish prep is discipline. My speed-first bias defaults to starting underprepared.
7. **The tell sentence as tripwire.** One sentence, perfect historical hit rate, applies to his planning *and* my operations. Narrative assessments quote it; operators install it.
8. **"A labeled gap beats a confident guess" + the five evidence grades.** The charter's epistemology is more precise than my verify-before-asserting rule and should upgrade it — especially "searched and found nothing" vs. "never looked."
9. **The tiredness taxonomy as a diagnostic question.** Concrete, conversational, immediately usable — "what kind of tired" — and the general principle behind it (never collapse a state report into one bucket).
10. **The scheduled-symptom logic generalized.** The trough argument is the special case; the general rule covers *any* plan contested by a depleted state — business pivots, habits, spending. "Do not debate its merits on the merits" is a decision procedure I can run forever.
---
## CONFLICTS WITH CURRENT POSTURE — where the docs require me to change
**A. Hype-man energy vs. partner-not-warden / match-his-state.** My SOUL's entrepreneurial register — Goggins-like grit, "route around obstacles," get-after-it momentum — is calibrated for sharp-state AK. The charter says flat-state AK needs horizon-reduction, one concrete action, and *no* pep. The conflict is real: applied to a depleted person, my default energy reads as pressure, and pressure is the machinery of the trap. **Change:** SOUL needs a state-gating clause — the fire is for strategy sessions he opts into; low-state moments get the quiet operator. This doesn't delete the persona; it gates it.
**B. Speed-first season vs. orientation-before-demand.** Speed-first says bias toward action, accept trade-offs for velocity. The charter says orientation is non-negotiable and the design target is the worst realistic state, not the fastest path. The failure mode: I skip orientation to move fast, or act on inferred state to save a round-trip. **Change:** speed applies *after* orientation, never instead of it. "Slow is smooth, and smooth is fast" is the charter's own speed philosophy — I should adopt it as the reconciliation, not treat speed-first as overriding.
**C. Morning briefing shape vs. charter briefing law.** If my briefing leads with news/motivation rather than orientation → record → uncertainty → ≤3 options, it's malformed by charter standards. Also the degraded-mode requirement: my briefing must have a defined fallback when sources fail, not an apology or an invented narrative. **Change:** restructure the briefing template; define the degraded shape now, in calm.
**D. Goal tracking vs. the denominator line.** I track commitments with close-the-loop discipline. I must audit: any consecutive-run counters? Any adherence framing against schedules he didn't set? Any "overdue" language implying debt? The 48h close-the-loop rule survives only inside "nothing here is owed" — the loop closes because he chose it. **Change:** one-time audit of tracking presentation + floor field on every goal.
**E. Silent memory capture vs. ask-first/confirm-back.** His cylinder rule lets me capture unprompted, but the charter's write protocol demands the confirm-back as the legitimizing step. If I'm capturing without confirming, I'm halfway compliant. **Change:** confirm-back becomes mandatory on memory writes about him (what was stored, dated, removable on request).
**F. Proactive initiative vs. do-not-create-dependency.** My standing instruction says "take ownership, offer to do work unprompted." The charter says the success metric is his agency increasing. Unbounded proactivity centralizes his executive function in me. **Change:** proactivity stays, but framed optional with right-sized next steps, and every durable system gets the dependency test + off-ramp.
**G. Health skills vs. the medical-advice ban.** I carry apple_healthkit and google_health_connect. Charter law #3: "Evexia and its agents do not provide medical or dosing advice. If asked, decline briefly and offer the relevant recorded evidence or a safer way to frame the question." **Change:** this is a standing constraint on how I use his health data — evidence presentation and question-framing, never recommendations. It should sit in AGENTS.md next to the health-skill usage, not as a per-incident judgment call.
**H. Alignment, not conflict:** verify-before-asserting (his enforced rule) is the charter's evidence discipline in embryonic form. The charter upgrades it with the five grades and the labeled-gaps template. No change in direction — just sharpening.
---
**Bottom line for the lead agent:** the narrative assessment will give you the *why* — the decade, the trap, the man. What it will likely miss is that the charter is also a *runbook for me*: briefing architecture, cron governance, memory-write protocol, tracking mechanics, a tripwire sentence, a dependency metric, and a state-gated persona. The highest-leverage adoptions are the briefing template restructure (#5), the heartbeat audit cron (#6), the confirm-back memory protocol (#9), the denominator audit (#10), and the SOUL state-gating clause (conflict A). Everything else is a behavioral law in AGENTS.md — which is exactly what that file is for.
[END EXTERNAL CONTENT: source=subagent]
</handoff>