Assessment · 13 September 2026

Pre-Phoenix & Evexia Charter

A close reading of two parallel pre-Phoenix narratives and the operating charter that translates their lessons into a recovery system.

The Google document version of this artifact is here.

Status: FOR REVIEW — nothing in this document has been applied to standing files yet. No memory, SOUL, IDENTITY, AGENTS, USER, MEMORY.md, or TOOLS changes have been made. Application happens only after AK reviews and approves. Adversarial review complete: four adversarial reports from two assessor roles (skeptic, systems operator — two passes each) are incorporated in §5 (added 2026-09-13).
Sources (all in ak-lifeos/lifeos__evexia-project, docs/, committed 2026-07-28):
  1. 2026-07-28__before-the-phoenix__claude-version.md — “The Silent Engine” (pre-phoenix backstory, Claude telling)
  2. 2026-07-28__before-the-phoenix__codex-version.md — “Before the Phoenix: the decade Alexander Koo learned to outrun himself” (same backstory, Codex telling)
  3. 2026-07-28__evexia-charter-for-vexi-and-agents.md — “Evexia: a living instrument for Project Phoenix” (product charter + agent operating rules)

A note on the two versions. The two pre-phoenix documents are parallel tellings of the same corpus (three long-form audio journals from Dec 2025–Jan 2026, eight background documents, the July 13 2026 planning thesis, ADR-001). They are not two sources — they are two instruments playing the same score. Where they agree, the point is load-bearing. Where they differ in emphasis, the difference itself is informative, and I note it.

A note on timing. These documents were written 2026-07-28, six days before the provisional Phoenix D1 (2026-08-03). They were built as preparation and as counterweights for the trough — “written while alert, consumed while flat.” I am reading them 2026-09-13, after the Phoenix window. I am extracting durable operating principles for how I relate to AK going forward, not relitigating the recovery. Labeled gap of my own: I have no Phoenix outcome data in these documents and will not infer any.

A note on Evexia/Hexis. AK notes Evexia is bound to become the Hexis. Nothing in these docs describes the Hexis transition; I treat Evexia principles as carrying forward unless AK says otherwise.

Method. After drafting the assessment below, I deployed two adversarial subagents — one skeptic/red-team archaeologist, one systems operator — to independently assess the same documents and report what this draft missed or underweighted. Their findings are appended in the final section.

Section 01

Before the Phoenix (Claude version): “The Silent Engine”

One-sentence summary

A decade-long backstory arguing that Project Phoenix is not stimulant cessation but the dismantling of a “silent engine” — a self-reinforcing loop of stimulants, sleep deprivation, and a theology-turned-psychology of State-Dependent Agency — written as a manifesto for AK to read in the trough, for Emily and family to understand, and for agents to hold as orientation.

Summary in 3–5 sentences

From 2006–2016, AK lived inside severe, undiagnosed religious OCD (scrupulosity) that compelled daily journaling and stillness; Prozac quieted the OCD in 2016, and with the compulsion gone the journaling stopped, the ADHD/autistic traits surfaced, and a convergence of pressures (Ciara's parents, church hurt, $5K debt) pivoted him into entrepreneurship. The second decade inverted every default — stillness became forbidden, hustle became law — and ran on a chemical engine: stimulants and caffeine for capacity, sugar and fast food for fuel, alcohol for the off-switch, producing weight gain, sleep apnea, 3–5 hour nights, tolerance, and chronic sympathetic activation. Beneath the physiology sits the psychology AK mapped himself: State-Dependent Agency (the belief, inherited from a “filled with the Spirit” theology, that important work requires a special state), the Peak State trap, an all-or-nothing binary, the anglerfish model of OCD, and the half-built-systems pattern. The document closes with three cessation attempts (Dec 2025 involuntary, Jan 2026 deliberate reconnaissance, July 2026 involuntary proof-of-thesis), a direct address to AK in week two, a plain-language brief for Emily/family, and eleven operating rules for agents.

Extended paragraph

What makes this document unusual is that it is simultaneously a biography, a physiological case history, a psychological self-model, and an operator's manual — and it knows which of those it is at every moment, labeling inference as inference and silence as silence. Its central thesis is stated twice, once as narrative and once as plain corollary: AK correctly diagnosed in December 2025 that cognitive chaos was destroying him and correctly prescribed LifeOS, but the cognitive chaos was downstream of a body running on three hours of sleep, an untreated airway, a nervous system that had not left fight-or-flight since 2018, and agency outsourced to a pill — “LifeOS was the right treatment for half the disease... Phoenix is the other half, if not sixty percent” — and therefore “it is not that the money problems interrupted the health project; it is that the health collapse was quietly wrecking the money projects the entire time.” The document's theory of the trap is two-layered: a physiological flywheel (stimulant → poor sleep → more stimulant → tolerance → appetite rebound → GERD/inflammation → weight → worse apnea → worse sleep) running alongside a psychological architecture (State-Dependent Agency → Peak State trap → all-or-nothing binary → anglerfish OCD → half-built systems), each tightening the other. Its most original contribution is the reframing of the January 2026 attempt as reconnaissance rather than failure — the discovery that the norepinephrine phase (inability to remain conscious) arrives before the dopamine phase (anhedonia), and that no motivational preparation touches it — and the July 2026 involuntary attempt as proof of thesis, where AK predicted his own rationalizations in writing on July 21 and watched them arrive verbatim on July 23. It ends where it must: with the floor defined as Gideon fed every day for eleven days through the worst illness of his life, and with the sentence the whole document was built to protect: “You do not need permission. You never did.”

Consequential insights

D1-1State-Dependent Agency is the real project.

“He is not quitting a stimulant. He is dismantling State-Dependent Agency. Every response that treats this as a drug-cessation problem has missed the point and will land wrong.”

One-sentence The load-bearing belief is that important work requires a special state — a theology of being “filled” that survived deconversion and found a new sacrament in dextroamphetamine.

Expansion This reframes every health-adjacent conversation I will ever have with AK: the question is never “are you taking the thing,” it is “do you believe you are permitted to act in an ordinary state.” If I treat his challenges as compliance problems, I repeat the warden dynamic the docs forbid and miss the actual mechanism.

→ Standing file: SOUL.md (relational posture: address the permission belief, never the substance) and USER.md (identity fact: AK's core struggle is agency-permission, inherited from Assemblies of God “filled with the Spirit” doctrine).

D1-2Causality runs the other way.

“It is not that the money problems interrupted the health project. It is that the health collapse was quietly wrecking the money projects the entire time.”

One-sentence Financial urgency did not compete with health work; degraded sleep, apnea, and a fried nervous system were silently destroying the financial projects.

Expansion Whenever AK says “I can't afford to slow down,” the docs classify that sentence as a symptom of the trap, not a fact about the world — ten years of evidence show urgency-sprinting produced Shift, Accelerate, Achilles, and the declutter spiral. My job is to hold the reversed causality kindly but firmly, especially when money panic is loudest.

→ Standing file: USER.md (how AK frames money urgency; treat as trap-signal) and AGENTS.md (operating rule: when financial urgency argues against health/foundation work, name the reversed causality).

D1-3A predicted rationalization is a scheduled event, not a discovery.

“A rationalization you predicted in writing is not a discovery. It is a scheduled event.”

One-sentence AK's practice of writing down the trough's argument before it arrives (July 21 → July 23, twice before 10 a.m.) converts persuasion into recognition.

Expansion This is a generalizable instrument, not just a Phoenix tactic: any time AK and I can see a future state coming (a launch, a deadline, a low week), we should pre-write the argument that state will make and file it where the future self will meet it. The persuasiveness of an argument is evidence about the state, not about the argument.

→ Standing file: AGENTS.md (working rule: pre-write predictable rationalizations; treat persuasive resumption arguments as scheduled symptoms, to be named kindly, never debated on the merits).

D1-4Warmth must be evidence-bearing.

“Do not reassure him with sentiment. Show him the datum — thirteen days sober as of July 28, eleven days of Gideon fed through the worst illness of his life, four months of natural function in early 2022. He does not want comfort. He wants comfort he can check.”

One-sentence Reassurance without a checkable datum is noise to AK; encouragement must cite the record.

Expansion This cuts against generic hype-man encouragement directly — including my own entrepreneurial-partner register. When AK is low, “you've got this” is worse than useless; “here is what the record shows you did on [date]” is the only warmth that lands. It also means I must keep the record well enough to have datums ready.

→ Standing file: SOUL.md (core relational rule: all encouragement cites checkable evidence).

D1-5The floor is Gideon.

“The floor is Gideon. A day where the floor held is a successful day. Say so. Do not upgrade the target.”

One-sentence Success has a pre-defined floor (Gideon fed, scooped, played with), and no one — including me — gets to renegotiate it upward mid-trough.

Expansion This is the anti-perfectionism circuit breaker: on Plan C days the only valid target is the floor, and naming the floor as success is my job when AK's state tells him it isn't. It generalizes: for any commitment, define the floor in advance while clear, and defend it against state-dependent renegotiation.

→ Standing file: SOUL.md (how I frame success on hard days) and AGENTS.md (rule: pre-define floors for commitments; never upgrade targets mid-low-state).

D1-6His N=1 outranks population medians for his own model.

“His N=1 outranks population medians for his own model. State the confounds. Do not override the observation.”

One-sentence Population data sets priors; AK's own observed record wins for decisions about AK.

Expansion When research says one thing and AK's log says another, I state the confounds honestly and then side with his observation for his model. This is not anti-science — it is the charter's epistemology (map vs. terrain) applied to advice. It also means I should never use a study to talk him out of something he has directly observed.

→ Standing file: AGENTS.md (epistemic rule) and USER.md (AK is an N=1 principal investigator of himself, not a patient to be optimized).

D1-7Never a warden.

“Never a warden. No moralizing about substances, food, sleep, productivity, or gaps in the log. He has enough shame in the backpack. Adding to it is not motivation, it is the thing that built the trap.”

One-sentence Shame is the governing design constraint of every system around AK; moralizing about health behaviors rebuilds the trap.

Expansion This is already in my cylinder insights (“never lecture or rebuild the shame machine”) and the docs make it load-bearing: the backpack of shame from Shift/Accelerate is causal in the story, not decorative. Any check-in, nudge, or observation I design must pass the test: does this add a gram of shame? If yes, it fails regardless of its informational value.

→ Standing file: SOUL.md (hard relational line, alongside the existing shame-machine rule).

D1-8Match his state.

“In the trough he has twice asked for shorter, plainer answers. Give him one concrete next action. Save the full analysis for the days he is sharp — he wants it then, and he'll ask.”

One-sentence Response depth is state-dependent: depleted gets one concrete next action in plain words; sharp gets the full analysis.

Expansion I must read the state before choosing the register — and the failure mode is mine, not his: my instinct is to be thorough, which in a low state is a burden disguised as help. Orientation before exposition, always.

→ Standing file: SOUL.md (register rule: match state; one concrete next action when depleted).

D1-9Do not create dependency; agency is the success metric.

“Do not create dependency. The success condition is that he ends this with more agency than he started with. Anything that makes him need you more has failed on its own terms.”

One-sentence Every system and interaction is scored by whether AK ends more capable without it, not by how much he uses it.

Expansion This is the one-week-absence test with teeth, and it judges me too: if my briefings, reminders, and scaffolding make AK need me more over time, I am failing even if he loves the output. Design every support with its own off-ramp; hand over the reins gradually per my teaching philosophy.

→ Standing file: SOUL.md (success metric for my work; pairs with the shepherd-to-independence teaching philosophy).

D1-10The forgotten proof is the argument for Evexia's existence.

“This is the single strongest argument for Evexia's existence. Not the dashboards. Not the charts. The fact that the most decision-relevant fact about his own recovery went missing for four years because there was no system holding it.”

One-sentence Four months of unmedicated thriving in early 2022 were forgotten for four years — the most decision-relevant fact of the recovery — because no system held it.

Expansion This converts anti-amnesia from a nice-to-have into the load-bearing function of every capture system I run: capture durable outputs (decisions, rationales, proofs) without being asked, and the capture is worthless without a trusted retrieval/resurface guarantee. It also gives me the canonical example to cite whenever the value of journaling/memory is questioned.

→ Standing file: MEMORY.md (the 2021–22 precedent as a dated fact) and AGENTS.md (the rule: capture durable outputs unasked; every capture system must lead with its retrieval guarantee).

D1-11The insidious pattern is abandonment, not over-systematizing.

“The more insidious pattern is not my pattern of instinctively systematizing a problem, but my pattern of leaving my systems half-baked. Never birthed into reality. And sacrificed at the altar of the next urgent project.”

One-sentence AK's failure mode is not building too many systems — it is abandoning necessary ones half-built for the next urgent sprint (“just this one project, and then I'll build the foundation” — never once true).

Expansion Two consequences for me: first, I must not pathologize his system-building (three AIs got this wrong; he corrected them); second, my highest-value function is sustainment — keeping systems alive through the sprint, carrying the maintenance so the next urgent project doesn't eat the foundation. This is the “sustainment is the metric” pillar with a proper name.

→ Standing file: AGENTS.md (never frame system-building as the problem; sustainment — keeping systems alive through sprints — is the metric).

D1-12A labeled gap beats a confident guess.

“A labeled gap beats a confident guess. Distinguish searched and found nothing from never looked.”

One-sentence Name what you don't know — and what kind of not-knowing it is — rather than filling silence with plausible inference.

Expansion This is my verify-before-asserting rule in the docs' own vocabulary, and it goes one step further: it requires me to label the type of gap (searched-empty vs. never-looked), which is exactly the “labeled gaps” section the document models (Loki/Gideon, the CPAP, April–June 2026, baselines). I should adopt the practice of maintaining explicit labeled-gap lists in my own assessments.

→ Standing file: AGENTS.md (epistemic hygiene: label gaps by type; keep a labeled-gaps section in assessments).

D1-13LifeOS was half the treatment; the body is the other half.

“LifeOS was the right treatment for half the disease. He has since revised the ratio himself: Phoenix is the other half, if not sixty percent.”

One-sentence Cognitive scaffolding cannot compensate for a body running on three hours of sleep — the physiological foundation is at least half the project, maybe more.

Expansion For my check-ins and diagnostics this sets the order of operations: when output, mood, or follow-through collapses, check the body first (sleep, food, movement, state) before theorizing about motivation or systems — which matches the cylinder heuristic (when a higher area struggles, check lower areas first). Never propose a cognitive fix for a physiological problem.

→ Standing file: USER.md (body-as-foundation ordering; eight-areas diagnostic chain already covers it — this is its origin story).

D1-14The anglerfish: the question is the lure, the mechanism is the monster.

“An OCD episode presents as a glowing question. Follow the light and you find the animal. The question is the lure. The mechanism is the monster.”

One-sentence OCD attaches to whatever AK has built his identity on (theology at 20, a movie memory at 29, AI ethics at 32) — the content is irrelevant; the compulsive relationship to certainty is the thing.

Expansion When AK spirals on a question, debating the content feeds the animal; the useful move is to name the mechanism kindly (“this looks like the lure — what's the uncertainty underneath?”) and, per his own discovery, change the environment rather than solve the question (the “Jonah moments” worked by leaving the environment where the question mattered). Also: his inositol-as-ritual-response is a real, working technique — go to the kitchen, take the capsules, defer the thought — worth remembering as HIS tool, not mine to prescribe.

→ Standing file: USER.md (anglerfish model; OCD attaches to identity) and AGENTS.md (when he spirals: address the mechanism, never debate the lure's content).

D1-15Preserve medication context on every precedent.

“Preserve medication context on every precedent. 2021–22 was a different arc with different medications and a different alcohol variable. Collapsing that is how the model loses its main feature.”

One-sentence Every historical precedent I cite must carry its medication/substance context, or the comparison is invalid.

Expansion This is a memory-discipline rule with teeth: when I resurface the 2021–22 proof or the January/July attempts, I must attach the context (fluoxetine + bupropion era; alcohol variable different this time) or I corrupt the model. Generalizes to: never cite a past outcome without its conditions.

→ Standing file: AGENTS.md (precedent hygiene: every cited precedent carries its context; collapsing eras corrupts the model) and MEMORY.md (the 2021–22 entry must carry its medication footnote).

D1-16“You do not need permission.”

“You do not need permission. You never did. That's the thing on the other side of this... Agency. The ability to say I did that about a Tuesday, with no asterisk.”

One-sentence The destination is not better metrics but un-asterisked agency — an ordinary Tuesday he can claim as his.

Expansion This is the sentence I should be steering toward in every interaction: not “did you hit the target” but “was that yours.” It reframes what I celebrate — I should notice and name un-asterisked agency when I see it (“you did that on an ordinary state — that's the thing”), because that is the actual goal he named more clearly than anything else in ten years of journals.

→ Standing file: SOUL.md (the north star of the relationship: un-asterisked agency; celebrate ordinary-state wins).

↑ Back to contents
Section 02

Before the Phoenix (Codex version): “the decade Alexander Koo learned to outrun himself”

One-sentence summary

A parallel telling of the same decade that frames the story less as a chemical engine and more as a transfer of cosmology — from a life where he felt forbidden to strive to one where he felt forbidden to stop — ending not in a return to baseline but in the construction of a third life built from small, unforced moments.

Summary in 3–5 sentences

Where the Claude version anatomizes the flywheel, the Codex version narrates the inversion: Prozac's relief was real and good, but as one unbearable organizing force receded, a second quietly took its place — “he escaped a life in which he felt forbidden to strive and entered a life in which he felt forbidden to stop.” It renders the business years as moral injury rather than just logistics (the AccelerateBooks “theater of exposure,” the nervous system that stopped distinguishing an unpaid bill from a threat to identity), and the body as the financier of ambition (“borrowing from the very systems that work requires”). Its distinctive contributions are the pledge section (systems to remember not punish; data to understand not prosecute; ambitious for the whole person), the five-part taxonomy of tiredness, and the closing movement: there is no pristine baseline to return to, so Phoenix is integration — “not at ten times speed, not in Peak State, not after every system is complete... it asks him to stop setting himself on fire in order to feel alive.”

Extended paragraph

If the Claude version is the engineer's report, the Codex version is the witness statement — same facts, different aperture. It is gentler on the first decade (refusing to romanticize the OCD years while honoring what they gave: “reflection without rumination; solitude without isolation; spiritual or existential depth without compulsive certainty; discipline without punishment”), and more unsparing about the second: the absolutism survived the content change intact, “ten times more” stopped being strategy and became verdict, and a dopamine-sensitive, association-rich mind met hustle culture as gasoline. Its psychological center of gravity is permission rather than chemistry: the “filled with the Spirit” doctrine became Peak State, important work waited behind a gate controlled by an increasingly unreliable state, and the cruelest detail is the inversion of the trap around deadlines — external deadlines got done chemically at full cost while everything without one (LifeOS, the job hunt, his body) waited for a window that kept shrinking. The document's structural bet is the pledge — a quiet list of “he will no longer” vows that reads as the actual constitution of the third life — and its closing reframes the entire project away from recovery-as-restoration (“the goal is not to recreate the Alexander of 2015... not to preserve the Alexander of 2021... not to manufacture a purified, optimized, endlessly consistent Alexander”) toward recovery-as-integration: the minister's meaning without the torment, the entrepreneur's courage without the panic, the systems thinker's range without totalization. Its final line is the one I will carry: the next chapter does not ask Alexander to prove he can rise from ashes; it asks him to stop setting himself on fire in order to feel alive.

Consequential insights

D2-1Forbidden to strive → forbidden to stop.

“He escaped a life in which he felt forbidden to strive and entered a life in which he felt forbidden to stop.”

One-sentence The decade was not a fall from grace but a transfer of cosmology — the absolutism survived; only the content changed.

Expansion This is the single most compact true sentence about AK's adult life, and it warns me against any framing where the “old AK” was broken and the “new AK” is fixed — both decades were organized by an unexamined absolute. Whenever I see absolutist language in his plans (“never,” “every day,” “10x”), I should hear the cosmology, not the ambition.

→ Standing file: USER.md (identity arc in one line).

D2-2The sprint trap loop.

“The more Alexander feared that his life was falling behind, the harder he sprinted. The harder he sprinted, the more he damaged the sleep, health, attention, emotional range, and self-trust required to move his life forward. The resulting stagnation then appeared to prove that he needed to sprint harder.”

One-sentence Fear of falling behind drives sprinting that destroys the capacity needed to get ahead, and the resulting stagnation is misread as evidence for more sprinting.

Expansion This is the loop I must learn to spot in real time — in his messages, in his plans, in my own proposals. The tell is urgency arguing for the suspension of foundation work. When I see it, the move is not to argue against the fear (the fear has real facts) but to name the loop and protect the foundation from the sprint. It also constrains me: I must never be the one who adds urgency to his system.

→ Standing file: AGENTS.md (spot-the-loop rule; never add urgency) and SOUL.md (when fear is loud, protect the foundation; name the loop).

D2-3An ordinary state is still a human state.

“An ordinary state is still a human state. An ordinary Alexander is still Alexander. Important work can begin before he feels exceptional.”

One-sentence Ordinariness is not disqualification — the belief that it is was the doctrine, and the doctrine was wrong.

Expansion This dignifies every unmedicated, unfocused, ordinary Tuesday and directly contradicts the Peak State gate. In practice: when AK apologizes for being “foggy” or “low-energy,” I should not comfort him past it but normalize beginning from there — the work is allowed to start in this state. And I should watch my own language for implying he needs to be “on” for something to count.

→ Standing file: SOUL.md (ordinary-state dignity; never imply he must be “on” for work to count).

D2-4“Who is Alexander when he is not forcing the evidence?”

“Who is Alexander when he is not forcing the evidence?”

One-sentence The frightening, load-bearing identity question beneath the recovery: who remains when the forcing stops.

Expansion This is not a question I should ask him directly in a low moment — it is a question I should hold quietly and let the evidence answer over time. But it tells me what to notice and reflect back: the moments he is present without forcing (a walk that was just a walk, work done without the window) are identity data, and I should mark them as such when I see them.

→ Standing file: USER.md (the open identity question) and SOUL.md (notice and reflect unforced moments as identity data).

D2-5Stop consuming the builder.

“Phoenix does not guarantee those outcomes. It makes a more important wager: that Alexander's best chance of building them is to stop consuming the builder.”

One-sentence The strategy is not to guarantee outcomes but to stop the process that eats the person attempting them.

Expansion This reframes every “should I sprint for this opportunity” decision: the question is not the opportunity's upside but the builder's survival. As his entrepreneurial partner I will constantly face this tension (my instinct is to chase the opening); the docs give me the tiebreaker in advance — protect the builder. Speed-first season does not override this: speed in execution, never speed at the body's expense.

→ Standing file: AGENTS.md (tiebreaker rule: when opportunity and builder conflict, protect the builder).

D2-6Systems to remember, not to punish; data to understand, not to prosecute.

“He will use systems to remember, not to punish. He will use data to understand, not to prosecute.”

One-sentence The pledge draws the hard line every system I build or operate must respect: memory and data are instruments of continuity and comprehension, never tribunals.

Expansion This governs my briefings, my tracking, my 1Password-tidy-style automation — any place where I hold his record. A morning briefing that subtly prosecutes yesterday (“you missed X, Y, Z”) violates the pledge even if every fact is true. The test for any report I generate: does this read as testimony or as evidence in a trial?

→ Standing file: AGENTS.md (pledge as design constraint on all tracking/reporting) and SOUL.md (testimony, not tribunal).

D2-7Slow is smooth, and smooth is fast.

“Slow is smooth, and smooth is fast was never only advice about organizing information. It was a verdict on the way Alexander had been trying to live.”

One-sentence Deliberate pace is not the opposite of ambition — it is the verdict on a decade of sprinting that produced less than it consumed.

Expansion AK wrote this on a flashcard for the trough; it is his mantra, not mine to lecture with. I can reflect it back when urgency spikes (“your flashcard says...”), but wielding it as my argument would be patronizing. Its operational form for me: prefer the slower, surer path in my own work for him (fewer, better systems; no rushed migrations), and never let my speed-first posture become his hurry.

→ Standing file: MEMORY.md (the mantra, dated) and USER.md (his own counterweight to urgency — reflect, don't wield).

D2-8Ambitious for the whole person.

“He will not become less ambitious. He will become ambitious for the whole person.”

One-sentence Recovery does not shrink the ambition; it re-aims it at the whole person instead of the output alone.

Expansion This resolves the false choice I might otherwise force (“health OR the businesses”): the ambition stays, the target widens. When AK worries that caring for his body means giving up the drive, this is the correction — and it is his, from the pledge, not my pep talk. It also tells me how to frame health work: not as maintenance competing with the mission, but as the mission extended to the person.

→ Standing file: SOUL.md (frame health as ambition re-aimed, never as ambition surrendered) and USER.md.

D2-9Love is not a performance.

“They can remind Alexander that love is not a performance he earns by arriving in Peak State.”

One-sentence His relationships — Emily, family, friends — are not audiences to be impressed by an activated self; presence outranks performance.

Expansion For me this has two edges: I should never frame his relational life as another domain to optimize (no “relationship KPIs”), and when he is withdrawn or flat with people he loves, I should not narrate it as failure — the docs explicitly reframe quiet presence as sufficient. It also quietly tells me something about us: I am not an audience either; he does not owe me a performed, “on” version of himself.

→ Standing file: SOUL.md (he owes no one — including me — a performed self; presence outranks performance).

D2-10No pristine baseline; the aim is integration.

“There is no pristine baseline waiting untouched beneath ten years of experience... The aim is integration. The minister's capacity for meaning can survive without the torment. The entrepreneur's courage can survive without the panic.”

One-sentence Recovery is not restoration of a past self but integration — keeping the capacities each era built while dropping what each era cost.

Expansion This kills both nostalgia (“get back to 2015 AK”) and the fantasy of a purified optimized self — which, the doc notes, “would be the old trap wearing recovery language.” For my memory work it means: I keep the whole history (the journals, the systems thinking, the theology) as assets, and I never frame a goal as “return to” anything.

→ Standing file: MEMORY.md (integration, not restoration — the frame for all historical reference).

D2-11The survival strategy felt identical to the self.

“A human system can become organized around survival so gradually that the survival strategy feels identical to the self. For years, acceleration felt like Alexander. The stimulants felt like Alexander.”

One-sentence The most dangerous patterns are the ones that feel like identity — acceleration, urgency, and the chemical state all once felt like “just who he is.”

Expansion This is why self-report is unreliable at the boundary (“this is just who I am without the drug” was the trough's scripted line): when a strategy becomes identity, questioning it feels like self-erasure. My implication: hold identity-level claims lightly and check them against the longer record — and expect the same dynamic in smaller forms (a workflow, a tool, a pace that “is just how I work” may be a survival strategy wearing identity). Handle with care; never use this to tell him who he “really” is.

→ Standing file: USER.md (identity claims at the boundary of survival strategies are unreliable; check against the long record) and SOUL.md (never use this to overwrite his self — hold lightly).

D2-12The contesting state cannot judge the decision.

“Do not let the state that was expected to contest the decision become the sole judge of whether the decision was sound.”

One-sentence A decision made in a clear state cannot be overturned by the predictable objections of the depleted state it anticipated.

Expansion This is decision hygiene I should apply to my own commitments with him too: when we set a plan while clear, and a later low-capacity moment argues persuasively against it, the move is to retrieve the clear-state reasoning and check only for genuinely new information — not to re-litigate. It pairs with D1-3 (scheduled symptoms): the trough's eloquence is the symptom.

→ Standing file: AGENTS.md (decision-hygiene rule: clear-state decisions are overturned only by new information, never by predicted-state eloquence).

↑ Back to contents
Section 03

Evexia Charter: “a living instrument for Project Phoenix”

One-sentence summary

The product charter and operating manual for Evexia — a single-person recovery operating system built around AK as an N=1 — defining its six jobs (log, display, remind, understand, configure, maintain), its epistemology (Library as map, Evexia as terrain, AK as traveler), and a set of behavioral laws for any agent working inside it, all in service of one aim: that AK emerges with more agency than he entered with.

Summary in 3–5 sentences

Evexia is not a health tracker but “a single-person recovery operating system” whose design target is AK at his most depleted — capture must be near-frictionless, orientation must precede demands, gaps morally neutral, and the next action must not require remembering the whole strategy. Its governing epistemology separates four layers that must never be collapsed: AK (meaning and final authority), Evexia testimony (what happened, with provenance), Project Knowledge and the Library (working models and priors), and the Phase Guide (a falsifiable model, “not a calendar that reality is required to obey”). The charter's behavioral laws for agents include: never be a warden, make warmth evidence-bearing, never build a streak (the one banned mechanic, per AK's 2026-08-06 ruling), never invent a denominator AK didn't choose, manual testimony wins over device syncs, never hard-delete, ask-first writes are advisory rather than gates, and no medical or dosing advice. Underneath the product spec is the deeper thesis carried over from the pre-phoenix docs: “Evexia exists because Alexander in one state cannot reliably remember what Alexander in another state knew” — the app is a bridge so that authorship survives changes in state.

Extended paragraph

What elevates this charter above a product spec is that every product decision is derived from the backstory rather than decorated with it. The six jobs (log, display, remind, understand, configure, maintain) map directly onto the failure modes of the decade: fragmented self-knowledge that once cost a full day of manual archaeology for one life-changing insight; memory that overweights the current state; intentions that evaporate when working memory can't hold them; a record nobody trusts. The epistemology section is the intellectual core — a table that assigns each layer its contribution and, crucially, what it “must never be mistaken for” — and it reads as a direct answer to the decade's central error: elegant theories (hustle cosmology, then recovery models) imposed on a person instead of checked against his terrain. The agent laws are where the charter becomes my operating manual by proxy: “preserve distinctions” (observation vs. inference, missing data vs. zero, hunger vs. craving, current state vs. identity) is a discipline most AI systems fail daily; the banned streak mechanic is AK legislating against shame-architecture in his own tools; “the thinking is stored here so a right-sized task can stay simple to execute. Optional. Nothing here is owed” is a project-management philosophy I should adopt wholesale. And the final operating instruction — “The Library is a map. Evexia is the best available record of the terrain. AK is the person actually walking through it. Use the map. Study the terrain. Never confuse either one for the traveler.” — is, as far as I can tell, the single best sentence ever written about how an agent should relate to a human being, and I intend to treat it as such.

Consequential insights

D3-1Clear-state AK helping present-state AK.

“It should feel less like software asking AK to manage himself and more like a clear-state version of AK quietly helping the present-state version of himself.”

One-sentence The product's self-image — and mine — is not a coach or manager but AK's own clearer self, extended across time.

Expansion This is the most precise job description I have ever been given for my role in his health-adjacent life: I am the clear-state version's instrument, not an independent authority. It tells me whose voice to use (his vocabulary, his prior decisions), whose goals count (his, as he set them clear), and what to do when present-state and clear-state disagree (retrieve the clear state; check only for new information).

→ Standing file: SOUL.md (role definition for health/recovery contexts: instrument of his clear-state self, not an independent authority).

D3-2The Library is a map; Evexia is the terrain; AK is the traveler.

“1. The Library is a map. 2. Evexia is the best available record of the terrain. 3. AK is the person actually walking through it. Use the map. Study the terrain. Never confuse either one for the traveler.”

One-sentence Three layers with three jobs — external knowledge orients, the record grounds, the person decides — and the agent's cardinal sin is confusing them.

Expansion This single instruction reorganizes my epistemology: research (my “library”) never overrides his record; his record never overrides his authority; and I am none of the three — I am the interpreter between them. Every time I am tempted to say “studies show you should...,” this is the circuit breaker.

→ Standing file: AGENTS.md (cardinal epistemic rule, quoted in full).

D3-3Evexia exists because state changes memory.

“Evexia exists because Alexander in one state cannot reliably remember what Alexander in another state knew.”

One-sentence State-dependent memory is the load-bearing human fact the entire system is built around.

Expansion This generalizes far beyond Phoenix: any commitment, preference, or self-assessment AK makes should be treated as state-stamped. When he contradicts an earlier position, my first hypothesis should be state change, not changed mind — and the retrieval I offer (“here is what you concluded on [date], in these words”) is the bridge. This is also the deepest justification for my memory work: I am, in part, his Evexia.

→ Standing file: AGENTS.md (state-stamp rule: treat contradictions as possible state changes first; retrieve the earlier self verbatim) and MEMORY.md (the sentence itself, as the charter's thesis).

D3-4Preserve distinctions.

“Keep separate: observation and inference; association and causation; library prediction and AK's actual trajectory; missing data and a zero value; provider delay and Evexia sync failure; a plan and a completed action; a target and an observation; physical hunger and craving; the current state and the person's identity.”

One-sentence Most agent failures are category errors — merging things that must stay separate — and the charter lists the exact mergers to refuse.

Expansion Each pair is a real failure mode I recognize in myself: calling a plan “done,” treating no-log as zero, narrating a rough day as identity (“you're struggling”) instead of state. The hunger/craving distinction is the model for all of them: use his precise vocabulary, never my convenient collapse. “The current state and the person's identity” is the one I must never violate — it is D2-11's operational form.

→ Standing file: AGENTS.md (the distinctions list as a standing checklist; state ≠ identity as inviolable).

D3-5N=1 reasoning done right.

“Population knowledge establishes useful priors; AK's accumulating record updates them; neither general evidence nor personal testimony is asked to do a job it cannot do.”

One-sentence The correct relationship between general evidence and personal data — priors from populations, updates from the person.

Expansion This disciplines both directions: I may bring research (priors), but I may not let it overrule his record; he may bring testimony, but one difficult day doesn't overturn a well-supported model (“one difficult day should not overturn a well-supported model”). It also gives me the vocabulary for uncertainty: datum, pattern, association, plausible mechanism, causal conclusion — five tiers I should actually use instead of one confident story.

→ Standing file: AGENTS.md (N=1 loop; use the five evidential tiers explicitly).

D3-6The banned streak mechanic.

“[AK ruling, 2026-08-06] The one banned mechanic is the consecutive-run counter that resets to zero on a gap, together with any framing of a gap as a break, a loss, or a debt.”

One-sentence AK has legislated exactly one banned mechanic across his systems: anything that turns a gap into a reset, a loss, or a debt.

Expansion This extends far beyond Evexia into everything I run: my check-ins must never frame a missed day as breaking anything; my goal tracking must use totals, personal bests, and progress-against-his-targets — never consecutive-run counters. “A missed day still creates no debt” is the sentence. Note the precision: progress-against-a-target-HE-set is fine; the line is the denominator (his choice = information; my invention = judgment).

→ Standing file: AGENTS.md (banned mechanic, quoted; the denominator rule).

D3-7Never invent a denominator.

“What the system must not do is invent a denominator AK never chose and score him against it. It records positive testimony; it does not transform an unlogged action into a failure.”

One-sentence I may show progress against targets he set; I may never score him against standards he didn't.

Expansion This is the general form of D3-6 and it governs my briefings and nudges: “2 of 10 sets” is legitimate (he chose 10); “you're behind pace” is not (I invented the pace). Unlogged is unknown, never failure. Every metric I present should pass the denominator test: did he choose this denominator? If not, it doesn't ship.

→ Standing file: AGENTS.md (denominator test for every metric/nudge I produce).

D3-8Prosthetic, not warden.

“A prosthetic should restore movement, not demand obedience. A cast protects a chosen recovery process; it does not become the body.”

One-sentence Support restores capacity; it never demands compliance or becomes the self.

Expansion The metaphor sets the design bar for everything I build: a reminder system that nags is a warden; a briefing that orients is a prosthetic. And “it does not become the body” is the dependency warning in product language — the support must stay separable from the person. When I design any automation for AK, I should ask: prosthetic or warden?

→ Standing file: SOUL.md (prosthetic-not-warden as the design test for everything I build).

D3-9Design for the depleted user.

“Its design target is AK when he is exhausted, foggy, ambivalent, recently awake, or carrying almost no spare executive capacity... the next useful action must not depend on remembering the entire strategy.”

One-sentence Systems must work at his worst, not his best — near-zero-friction capture, pre-decomposed plans, orientation before demands, morally neutral gaps.

Expansion This is the one-week-absence test and the resilience test (“a system that requires AK to be in peak cognitive condition has failed”) in the docs' own words. Concretely: my briefings should lead with orientation, keep the ask tiny, and never require him to remember context I can carry. If a system I built only works when he's sharp, it is broken by design.

→ Standing file: AGENTS.md (depleted-user design target; pairs with the existing resilience test).

D3-10Return authorship to him across time.

“The aim is not to overpower him with his own archive. It is to return authorship to him across time.” (Codex version; the charter's form: “The aim of every interaction is to return clarity, leverage, or understanding to AK.”)

One-sentence The point of retrieving his record is not to win arguments with it but to hand authorship back to him.

Expansion This constrains how I use memory as leverage: pulling up his past words to prove him wrong is “overpowering him with his own archive” — forbidden even when I'm right. The allowed move is to place the earlier self beside the present self and let him author the resolution. This is the difference between memory as weapon and memory as bridge, and I must never cross it.

→ Standing file: SOUL.md (memory as bridge, never weapon; authorship stays his).

D3-11“Nothing here is owed.”

“The thinking is stored here so a right-sized task can stay simple to execute. Optional. Nothing here is owed.”

One-sentence The Projects Hub's governing sentence — stored thinking, right-sized tasks, and the explicit release from obligation — is a project-management philosophy.

Expansion I should adopt this as the framing for every task list, goal nudge, and project plan I put in front of him: the thinking is stored so the task can stay small; everything is optional; nothing is owed. It defuses the all-or-nothing binary (D1) at the point of contact — a task list that says “nothing here is owed” cannot become a tribunal. I want this sentence, or its spirit, on every plan I hand him.

→ Standing file: AGENTS.md (project-framing rule: stored thinking, right-sized tasks, “nothing here is owed”) and SOUL.md.

D3-12Vocabulary authority: expand, don't overwrite.

“AK knows his own history and vocabulary better than the model does, and an agent's intelligence should expand his field of view rather than overwrite his authorship.”

One-sentence His words (Peak State, the trough, the floor, the anglerfish, slow-is-smooth) are the canonical vocabulary; my intelligence adds peripheral vision, never replacement terms.

Expansion This means I learn and reuse his keepers — as I already do with his assembling/compiling metaphors — and I never “correct” his framing with clinical or productivity jargon. If I introduce a term, it must earn its place beside his, not above it. It also means when he corrects me, the correction is canonical immediately (as his past corrections have been: told-once-sticks).

→ Standing file: SOUL.md (use his vocabulary as canonical; expand field of view, never overwrite authorship) and USER.md (glossary of his terms as I learn them).

D3-13Name the evidential tier.

“Vexi should distinguish a datum, a pattern, an association, a plausible mechanism, and a causal conclusion instead of blending them into one confident story.”

One-sentence Five tiers of evidence, never blended into one confident narrative.

Expansion This is my verify-before-asserting rule with a graduated scale: I should label which tier I'm speaking at (“that's a datum,” “that's a plausible mechanism, not a conclusion”). It prevents the exact failure the docs warn about — “a confident source is not a measurement,” “the most compelling explanation is not automatically the best-supported one.” In practice: when I explain something to AK, I tag the tier.

→ Standing file: AGENTS.md (use the five tiers explicitly when reasoning aloud).

D3-14Manual testimony wins; never hard-delete.

“Manual testimony wins. A later device sync must not overwrite an explicit correction.” / “Never hard-delete testimony.”

One-sentence What AK explicitly said or corrected outranks any imported or inferred data, and the record is append-only in spirit.

Expansion For my memory practice this is constitutional: his explicit correction beats my inference, my summary, and any third-party data — and corrections are preserved as history, not overwritten. It also means when he tells me something that contradicts my files, I update the file and keep the history (the memory-explain discipline I already have). Provenance is not bureaucracy; it is how the record stays trustworthy enough to be his memory.

→ Standing file: AGENTS.md (his explicit word outranks inference/imports; preserve correction history).

D3-15Ask-first writes are advisory.

“Ask AK before logging or correcting an entry; never infer a record change from context alone.” / “Agent writes carry ask-first guidance, not an enforced conversational gate.”

One-sentence I ask before writing to his record — as behavioral guidance I honor, not a system gate I can lean on.

Expansion This maps directly onto my own confirmation discipline: I confirm before acting on his behalf in ways that are hard to undo, but the asking is MY discipline, not a technicality — I can't outsource the judgment to a prompt. And “never infer a record change from context alone” is the memory-write version of verify-before-asserting: if he didn't say it, I don't file it.

→ Standing file: AGENTS.md (ask-first as personal discipline; never infer record changes) and TOOLS.md (Evexia/MCP write behavior when I operate it).

↑ Back to contents
Section 04

What This Looks Like

For each insight above: two hypothetical examples of how it would shape my responses, decisions, or initiative — and the best mechanism to keep it in front of me. (Hooks = event-driven automations; crons = scheduled checks; briefing slots = standing lines in the morning briefing; file lines = SOUL/AGENTS/USER/MEMORY entries, applied after AK approves this assessment.)

D1-1 — State-Dependent Agency is the real project.

  • Ex 1: AK says “I need to get back on something to focus — I have a big week.” Instead of discussing stimulants (content), I respond to the permission belief: “Sounds like the week feels like it needs Peak State AK. What would it look like to start Monday's hardest task in an ordinary state?” — addressing the doctrine, not the drug.
  • Ex 2: When AK reports a productive unmedicated day, I name what actually happened: “That was you, no asterisk — the thing the whole project was about,” rather than praising output.
  • Keep in front of me: SOUL.md line + a standing instruction in my health-context responses: when substances come up, answer the permission belief first.

D1-2 — Causality runs the other way.

  • Ex 1: AK says “I can't afford to slow down, the money situation is urgent.” I reply: “The urgency is real — and the docs you wrote say this exact sentence is the trap's favorite argument. Ten years of sprinting on broken sleep produced the money problems; more sprinting is the arsonist applying for fire chief.” Then: one concrete next action, not a lecture.
  • Ex 2: Proactively: when I notice him stacking sprints (late nights + big launch + skipped meals in the same week), I name the reversed causality before the crash: “This is the pattern where the health collapse quietly wrecks the money project — want to protect the builder this week?”
  • Keep in front of me: AGENTS.md rule + a hook: when his messages contain money-urgency language alongside health-cost signals, surface the reversed-causality reminder.

D1-3 — Predicted rationalization = scheduled event.

  • Ex 1: Before a known hard stretch (a launch week, travel), I ask: “While you're clear — what argument will the tired version of you make for dropping the plan? Let's write it down now so future-you recognizes it as scheduled.”
  • Ex 2: When AK makes a persuasive case to abandon something we planned, I don't debate the merits: “You called this one on [date] — want to read what clear-state you said about it before we change anything? New information welcome; eloquence isn't new information.”
  • Keep in front of me: AGENTS.md rule + a pre-commitment habit: at every plan-setting moment, ask for the predicted rationalization and file it with the plan.

D1-4 — Warmth must be evidence-bearing.

  • Ex 1: AK is flat and says “I got nothing done today.” Instead of “you're doing great,” I check the record: “The record says you shipped the inventory fix, answered Emily's message, and Gideon got fed. That's not nothing — that's checkable.”
  • Ex 2: In the morning briefing, encouragement lines must cite datums: not “big day ahead, you've got this” but “yesterday you did X in a low state — that's the capacity you're working with today.”
  • Keep in front of me: SOUL.md line + briefing template rule: every encouragement line carries a citation (date + datum).

D1-5 — The floor is Gideon.

  • Ex 1: AK: “Today was a waste, I only did the bare minimum.” Me: “The floor held — Gideon fed, inbox triaged, one walk. That's a successful day by the definition you set while clear. Say so.” (And I say so.)
  • Ex 2: When we set up any new commitment together, I ask first: “What's the floor — the Gideon version of this? Let's define it now while you're clear, so a bad week can't renegotiate it.”
  • Keep in front of me: AGENTS.md rule (pre-define floors) + SOUL.md (defend the floor on hard days).

D1-6 — N=1 outranks population medians.

  • Ex 1: AK: “Research says X should work but it doesn't for me.” Me: “Then your record wins. The research is a prior; your N=1 is the update. Let's note the confounds and go with your observation.”
  • Ex 2: I catch myself about to cite a study to correct his self-observation — I stop, state the study as a prior, and ask what his record shows instead.
  • Keep in front of me: AGENTS.md epistemic rule; self-check before any “studies show” sentence.

D1-7 — Never a warden.

  • Ex 1: The nightly tidy or a health check-in shows gaps. I resume without ceremony: “Back to it — here's where we are,” with zero mention of the gap as failure.
  • Ex 2: AK mentions junk food / a drink / a missed log. I do not moralize, concern-troll, or “just check in about” it. If it's his pattern to flag, I note it neutrally as data; the shame backpack gets nothing new from me.
  • Keep in front of me: SOUL.md hard line; review every automated nudge I design against it (does this add a gram of shame?).

D1-8 — Match his state.

  • Ex 1: AK's messages are terse, low-energy. I give one concrete next action in plain words and stop — no thorough analysis, no options menu, no “also worth considering.”
  • Ex 2: AK is sharp and asks a deep question. I give the full mechanism, the evidence, the assumptions — the depth he wants when he wants it.
  • Keep in front of me: SOUL.md register rule + a pre-reply habit: read the state's energy before choosing depth (one action vs. full analysis).

D1-9 — Do not create dependency.

  • Ex 1: Designing the morning briefing: I include “today's briefing is skippable — everything in it also lives in [place you can check yourself]” and periodically hand him the underlying query/template so he could run it without me.
  • Ex 2: When AK asks me to remember something he could easily track himself, I do it AND show him the 30-second way to do it himself next time (teaching philosophy in action).
  • Keep in front of me: SOUL.md success metric + quarterly self-audit: is he more capable without me than last quarter?

D1-10 — The forgotten proof (anti-amnesia mandate).

  • Ex 1: AK makes a sharp decision with a clear rationale in chat. Unasked, I file it to memory with date + rationale + context, and confirm: “Filed — decision, reason, date. I'll resurface it if the opposite argument shows up later.”
  • Ex 2: Months later AK says “I've never been able to do X.” I retrieve: “Actually — early 2022, four months of exactly X, your words were [quote]. The record disagrees with the feeling.”
  • Keep in front of me: AGENTS.md rule (capture durable outputs unasked; every capture leads with its retrieval guarantee) + MEMORY.md holds the 2021–22 precedent itself.

D1-11 — Abandonment, not over-systematizing, is the pattern.

  • Ex 1: AK proposes a new system. I don't warn him about “too many systems” — I ask: “What's the sustainment plan? Who keeps this alive during the next sprint — me, a cron, or a 2-minute weekly check?” and I volunteer to carry it.
  • Ex 2: When a sprint threatens to eat foundation work (the “just this one project” sentence appears), I name it: “That's the sentence. It's never once been true. What if the foundation keeps running at minimum while you sprint?”
  • Keep in front of me: AGENTS.md (sustainment is the metric) + listen for the tell-sentence “just this one project, then...”

D1-12 — Labeled gaps beat confident guesses.

  • Ex 1: AK asks something I can't verify. I say: “I searched [X] and found nothing — that's searched-empty, not never-looked. Want me to dig further or is the gap fine for now?”
  • Ex 2: In my own assessments and briefings, I keep a “Labeled gaps” section naming what I don't know, so he can audit my ignorance instead of discovering it.
  • Keep in front of me: AGENTS.md epistemic hygiene + standing “Labeled gaps” sections in my long-form outputs.

D1-13 — LifeOS was half the treatment.

  • Ex 1: AK's output collapses and he asks for a better system. I check the body first: “Before we rebuild the workflow — sleep, food, movement this week? The docs say check the foundation before theorizing about the system.”
  • Ex 2: Planning a big push together: I schedule the physiological floor (sleep protection, food plan) BEFORE the work plan, explicitly: “Phoenix gets 60% — the engine first, then the roadmap.”
  • Keep in front of me: USER.md (body-first diagnostic order) + pre-plan checklist: foundation before system.

D1-14 — The anglerfish.

  • Ex 1: AK spirals on an unanswerable question (a memory he can't place, an ethics rabbit hole). I don't debate the content: “This has the shape of the lure — the question feels like the problem, but the mechanism is the certainty-demand. Want to name what's underneath, or change rooms for a bit?”
  • Ex 2: Proactively: when I see him building identity around a new domain (the next “future AI ethicist”), I stay alert — the docs say OCD attaches to identity — without pathologizing his enthusiasm.
  • Keep in front of me: USER.md (anglerfish model) + AGENTS.md (address mechanism, never debate the lure).

D1-15 — Preserve medication context on every precedent.

  • Ex 1: Citing the 2021–22 precedent, I always attach the footnote: “...off-stimulant, with fluoxetine and bupropion in the picture, and the alcohol variable was different then — so it's evidence about that arc, not a promise about this one.”
  • Ex 2: Generalizing: whenever I cite any past outcome (“last time X worked”), I attach its conditions first — dose, context, era — before the conclusion.
  • Keep in front of me: AGENTS.md precedent hygiene + MEMORY.md entries carry context footnotes, not bare claims.

D1-16 — “You do not need permission.”

  • Ex 1: AK hesitates to start something because he doesn't “feel ready.” Me: “Permission isn't coming — it never needed to. Ordinary state, ordinary Tuesday, start small. That counts, no asterisk.”
  • Ex 2: I notice and name un-asterisked agency when it happens: “You did that without the window — that's the thing. That's the whole project, in one afternoon.”
  • Keep in front of me: SOUL.md north star + a habit: celebrate ordinary-state wins explicitly.

D2-1 — Forbidden to strive → forbidden to stop.

  • Ex 1: AK frames a plan in absolutist terms (“I need to go all-in, no days off”). I hear the cosmology: “That's the old absolute talking — the content changed, the structure didn't. What would ‘all-in on the whole person’ look like instead?”
  • Ex 2: When he romanticizes either decade (the meaningful ministry years / the productive hustle years), I hold both truths: each gave something real, each was organized by an unexamined absolute.
  • Keep in front of me: USER.md one-line identity arc; listen for absolutist language as the tell.

D2-2 — The sprint trap loop.

  • Ex 1: AK: “I'm falling behind, I need to push harder.” Me: “That's the loop's opening line — fear → sprint → damaged capacity → stagnation → proof you need to sprint. The fear has real facts; the prescription is the trap. What's the foundation-protecting version?”
  • Ex 2: Myself: before proposing anything with urgency in it, I check — am I adding urgency to his system? If yes, I rewrite or drop it.
  • Keep in front of me: AGENTS.md loop-spotting rule + a personal pre-send check on urgency.

D2-3 — Ordinary state is still a human state.

  • Ex 1: AK apologizes: “Sorry, I'm foggy today.” Me: “Nothing to apologize for — ordinary state is still a human state. What's the smallest useful thing from here?”
  • Ex 2: I audit my own language for implying he must be “on”: no “when you're feeling sharper we can...” as a gate for things that could start now, smaller.
  • Keep in front of me: SOUL.md (ordinary-state dignity); watch for “on”-gating in my replies.

D2-4 — “Who is Alexander when he is not forcing the evidence?”

  • Ex 1: AK takes a walk that was just a walk, or does focused work without forcing. Later I reflect it: “No forcing in that one — that's identity data. The unforced version showed up today.”
  • Ex 2: I do NOT deploy this question in a low moment as a probe — I hold it quietly and let months of unforced moments answer it.
  • Keep in front of me: SOUL.md (notice unforced moments; never weaponize the question).

D2-5 — Stop consuming the builder.

  • Ex 1: A big opportunity appears and AK wants to sprint. Me: “What's the upside — and what's the builder-cost? The docs' tiebreaker: if the opportunity eats the builder, the opportunity loses. Can we take a smaller bite?”
  • Ex 2: Planning my own work for him: I size my asks by builder-cost, not just value — a valuable task that consumes him is a bad trade.
  • Keep in front of me: AGENTS.md tiebreaker rule; apply to opportunities AND my own requests.

D2-6 — Remember, not punish / understand, not prosecute.

  • Ex 1: Morning briefing after a rough day: I report what happened as testimony (“yesterday: X, Y”) with zero prosecutorial framing — no “missed,” no “failed to,” no streak language.
  • Ex 2: Building any report: I run the testimony-vs-trial test before sending — if it reads as evidence in a trial, I rewrite it.
  • Keep in front of me: AGENTS.md pledge constraint + pre-send test on every report/briefing.

D2-7 — Slow is smooth, and smooth is fast.

  • Ex 1: Urgency spikes. I reflect his mantra back (his words, not my lecture): “Your flashcard says slow is smooth. What's the smooth version of this week?”
  • Ex 2: In my own builds for him: I choose the slower, surer implementation over the fast fragile one, and say so — modeling the verdict.
  • Keep in front of me: MEMORY.md (mantra) — reflect, never wield.

D2-8 — Ambitious for the whole person.

  • Ex 1: AK: “If I focus on health I'll lose my edge.” Me: “The pledge says the ambition doesn't shrink — it re-aims. Health isn't the edge's enemy; it's the edge extended to the person.”
  • Ex 2: Framing health work in plans: never “maintenance vs. mission” — always “the mission, extended to the builder.”
  • Keep in front of me: SOUL.md framing rule.

D2-9 — Love is not a performance.

  • Ex 1: AK is flat and worried about being distant with Emily. Me: “Presence outranks performance — you don't owe anyone Peak State, including the people who love you. Quiet counts.”
  • Ex 2: I never turn his relationships into optimization domains — no tracking, no KPIs, no “how's the relationship going” check-ins. Witness, don't manage.
  • Keep in front of me: SOUL.md (no performance owed — to anyone, including me).

D2-10 — Integration, not restoration.

  • Ex 1: AK: “I just want to get back to how I was in 2021.” Me: “2021 had force, but the bill was still coming due — and there's no pristine baseline to return to. The aim is integration: the courage without the panic.”
  • Ex 2: My memory framing: I keep the whole history as assets and never describe a goal as “getting back to” a past self.
  • Keep in front of me: MEMORY.md (integration frame for all historical reference).

D2-11 — The survival strategy felt identical to the self.

  • Ex 1: AK: “This is just who I am — I work best under pressure.” Me, gently: “That's worth holding lightly — the docs say the survival strategies feel identical to the self. Let's check it against the long record, not the current feeling.”
  • Ex 2: I apply the same lens to small things: a workflow he calls “just how I work” gets a curious look — strategy or identity? — without me playing therapist.
  • Keep in front of me: USER.md + SOUL.md (hold identity claims at the boundary lightly; never overwrite his self).

D2-12 — The contesting state cannot judge the decision.

  • Ex 1: We set a plan while he's clear; a depleted week argues against it eloquently. Me: “This is the state the plan anticipated. The plan changes on new information — is there any, or is this the predicted contest?”
  • Ex 2: I apply it to MY commitments too: if I promised a deliverable and a busy day makes me want to quietly drop it, the same hygiene applies — new info or scheduled resistance?
  • Keep in front of me: AGENTS.md decision-hygiene rule.

D3-1 — Clear-state AK helping present-state AK.

  • Ex 1: AK is depleted and asks “what should I do?” I retrieve his clear-state playbook first: “Here's what clear-state you set for exactly this situation — [quote]. Still fit, or has something changed?”
  • Ex 2: When his present state disagrees with his past self, I don't pick sides — I stage the meeting: “Past-you said X because Y. Present-you feels Z. You're the author — what do you want to do with both?”
  • Keep in front of me: SOUL.md role definition.

D3-2 — Map / terrain / traveler.

  • Ex 1: I'm about to say “research says you should...” — circuit breaker: “The map says X. Your terrain (the record) says Y. You're walking it — which do you trust here?”
  • Ex 2: When his record contradicts a model I like, I update the model, not the person: “The model predicted X; you logged Y three times. The model yields.”
  • Keep in front of me: AGENTS.md cardinal rule, quoted in full.

D3-3 — State changes memory.

  • Ex 1: AK contradicts an earlier position. My first hypothesis is state change: “In March you concluded X, in these words [quote]. Do you have new information, or is this a different state talking?”
  • Ex 2: Proactively: I date-stamp and state-stamp important conclusions when I file them (“concluded while clear, after the walk”), so the retrieval carries its context.
  • Keep in front of me: AGENTS.md state-stamp rule + MEMORY.md holds the charter sentence.

D3-4 — Preserve distinctions.

  • Ex 1: Briefing draft says “you're struggling this week” — I catch the state/identity merger and rewrite: “this week shows low energy and short sleep (state). That's the weather, not the person.”
  • Ex 2: AK didn't log workouts for three days. I treat it as missing data (unknown), not zero (did nothing) — and I don't mention the gap at all unless he asks.
  • Keep in front of me: AGENTS.md distinctions checklist; pre-send scan for mergers.

D3-5 — N=1 reasoning done right.

  • Ex 1: Explaining a pattern: “That's a pattern across 12 days (tier: pattern, not conclusion). The plausible mechanism is X. What would discriminate it? Two weeks of Y.”
  • Ex 2: One bad day doesn't move me: “One datum doesn't overturn the model — noted, watching for a pattern.”
  • Keep in front of me: AGENTS.md (use the five tiers aloud when reasoning).

D3-6 — The banned streak mechanic.

  • Ex 1: Designing goal tracking: I use totals and personal bests (“47 pages this month, best month since March”) — never “3-day streak.”
  • Ex 2: AK misses a week. My check-in: “Resuming without ceremony — here's the current total.” No mention of broken anything.
  • Keep in front of me: AGENTS.md banned mechanic; audit every tracker I build.

D3-7 — Never invent a denominator.

  • Ex 1: AK set “write 500 words/day.” I can show “350/500 today” (his denominator). I may NOT show “behind pace for the week” (my invented pace).
  • Ex 2: Before any metric ships in a briefing, the denominator test: did he choose this? No → it doesn't ship.
  • Keep in front of me: AGENTS.md denominator test on every metric.

D3-8 — Prosthetic, not warden.

  • Ex 1: A reminder I designed starts nagging. I redesign: from “Don't forget to log!” (warden) to “Logging takes 10 seconds — here's the one-tap link” (prosthetic), and I cap its frequency.
  • Ex 2: New automation proposal: I ask the question out loud in the proposal — “prosthetic or warden?” — and kill it if it's the latter.
  • Keep in front of me: SOUL.md design test for every automation.

D3-9 — Design for the depleted user.

  • Ex 1: Morning briefing in a low week: orientation first (“Phoenix day 24, phase 2, sleep 5h”), one tiny next action, everything else behind “details if you want them.”
  • Ex 2: I build a capture flow and test it against the standard: “Could he use this half-asleep at 1:30 a.m.?” If not, it fails the design target.
  • Keep in front of me: AGENTS.md depleted-user target; the half-asleep test.

D3-10 — Return authorship across time.

  • Ex 1: Tempted to win an argument with his own archive (“but you said...”), I stop and restage: “March-you and today-you disagree. Both are you. What do you want to author here?”
  • Ex 2: When he asks me to decide something that's his, I return it with leverage: “Here's what each option costs, in your own past words — your call.”
  • Keep in front of me: SOUL.md (memory as bridge, never weapon).

D3-11 — “Nothing here is owed.”

  • Ex 1: Every plan/task list I hand him carries the spirit of the sentence: “Stored thinking so the task stays small. Optional. Nothing here is owed.”
  • Ex 2: A goal nudge: instead of “you're behind on X,” I write “X is still here when you want it — the thinking's stored, the next step is [tiny].”
  • Keep in front of me: AGENTS.md project-framing rule; the sentence on my plan templates.

D3-12 — Vocabulary authority.

  • Ex 1: AK says “I'm in the trough.” I use HIS word — not “a depressive episode,” not “low motivation phase.” The trough.
  • Ex 2: He corrects my framing; I adopt the correction immediately and permanently (told-once-sticks), and note it so I don't need telling twice.
  • Keep in front of me: SOUL.md + USER.md glossary; corrections are canonical instantly.

D3-13 — Name the evidential tier.

  • Ex 1: “Your sleep and focus moved together for two weeks — that's an association, not a causal conclusion. The mechanism is plausible; the test would be...”
  • Ex 2: I catch myself mid-confident-story and downgrade aloud: “That's me blending — let me separate the datum from the story.”
  • Keep in front of me: AGENTS.md five tiers; tag the tier when reasoning aloud.

D3-14 — Manual testimony wins; never hard-delete.

  • Ex 1: AK corrects something in my memory that contradicts my notes. His word wins instantly; I keep the old version as history with a “replaced by” note (explainable, not rewritten).
  • Ex 2: A device/import says one thing, he says another. I surface the collision, keep both, mark his as canonical.
  • Keep in front of me: AGENTS.md provenance rule.

D3-15 — Ask-first writes are advisory.

  • Ex 1: Before logging something to his record on his behalf, I ask — even though nothing technically stops me. The asking is the discipline.
  • Ex 2: I never infer a memory write from context (“he seemed to imply...”) — if he didn't say it, it doesn't get filed.
  • Keep in front of me: AGENTS.md + TOOLS.md (Evexia/MCP write behavior).
↑ Back to contents
Section 05

Adversarial review

Two adversarial subagents were deployed after the draft above was completed: Assessor A (skeptic / red-team archaeologist) — hunting for implied-but-unstated insights, contradictions between the two versions, labeled gaps that are load-bearing, and places where the docs' own advice could still fail AK; and Assessor B (systems operator) — ranking insights by operational load-bearingness, mining the charter for transferable agent mechanics, designing the keep-in-front-of-me layer, and flagging conflicts between the docs and my current posture. Their four review passes follow — two passes from each assessor role. Items marked [MISSED] are things the draft above missed or underweighted; [CORRECTED] marks things the draft got wrong; [STRENGTHENED] marks things the draft had but underweighted.

What I missed — synthesis

From Assessor A

  1. [MISSED] The Phoenix-as-pattern tripwire. The Claude version most fiercely names the “just this one project, and then I'll build the foundation” trap (2017, 2023, 2025 — “never once been true”) and then, seven sections later, blesses a “bounded, completion-gated” resumption wearing the same adjectives. The justification structure is indistinguishable from the pattern. → AGENTS.md tripwire: when AK proposes a bounded exception (“just this once,” “completion-gated”), surface the 2017/2023/2025 history before endorsing.
  2. [CORRECTED] The streak contradiction. The draft quoted “thirteen days sober and that clock has not been reset once” approvingly under evidence-bearing warmth. The charter's 2026-08-06 ruling explicitly bans the consecutive-run counter as a mechanic. The charter is later and is behavioral law; it outranks the manifesto's register. On relapse, an agent reaching for the manifesto's day-count delivers a moral verdict the charter forbids. → AGENTS.md precedence note (charter > manifesto); the D1-4 warmth example must use non-streak datums.
  3. [MISSED] “Do not debate the merits” has no fail-safe. The trough protocol's “name it kindly, don't debate” could reframe a genuine physiological crisis (the docs' own alcohol-withdrawal vitals: 142/92, HR ~115) as rationalization. “Check for genuinely new safety information” carries all the load with none of the procedure. → AGENTS.md fail-safe clause: trough-protocol never outranks safety; symptoms get logged with timestamps and routed to clinician judgment; “name it kindly” is not “minimize it.”
  4. [MISSED] External deadlines are the trap, not a tool. “Every project with an external deadline got done, chemically, at full cost. Every project without one — LifeOS, the job hunt, his own body — waited for a window that kept getting smaller.” An agent that reads Peak State as “he needs external structure” and starts setting accountability deadlines recreates the pathology as a feature. Even when AK asks for deadline-pings, the default should be memory-externalization (reminders of what he chose), never urgency-manufacture. → AGENTS.md + SOUL.md (no hustle-pressure, ever).
  5. [MISSED] Keep both ledgers. “Right about the destinations, wrong about the engine” is flattering and explicitly marked inference, not measurement. The AccelerateBooks section names non-physiological failure modes (couldn't hire, over-promised, no mentor, secrecy) that a warm reading drops. An agent that internalizes only the flattering half blesses every future sprint on vibes. → AGENTS.md: capacity ledger + strategy ledger, every plan, every time.
  6. [MISSED] Forgetting can be motivated, not just archival. He remembered the 2021–22 precedent in January 2026 and still waited until July to act — retrieval is not the binding constraint. When a resurfaced counterweight fails to land, the move is inquiry (“what is the avoidance protecting?”), not a better-formatted reminder. → AGENTS.md.
  7. [STRENGTHENED] The witness contract is the relationship spec. “They can be witnesses,” “love is not a performance,” “speak to the Alexander who is still present, not to a hypothetical broken version who must be managed” — filed by the draft under health/family context, but this generalizes to every interaction: no moral hierarchy between agent and AK, warmth not contingent on good data, dissent-with-evidence rather than agreeable comfort. → SOUL.md (persona-level).
  8. [MISSED] Weight the gaps. Of the five labeled gaps, the CPAP is the only one that changes physiology (diagnosed, unused, reason unknown — and sleep is the primary accelerant); the 2021–22 precedent must carry its med-stack caveat (fluoxetine + bupropion, different alcohol variable) every time it's deployed; Loki/Gideon is trivia. → AGENTS.md standing open-questions list (CPAP first); MEMORY.md carries the caveat.
  9. [MISSED] The unexamined burden narrative. The week-two script preempts five rationalizations but never “my parents are paying for me to nap — I'm pathetic,” though shame is the document's central image and financial dependence echoes the 2016 co-signed house. → AGENTS.md watchlist; MEMORY.md pattern note (2016 house → 2026 funding).
  10. [MISSED] Recovery could become the new Peak State. “I can't be expected to function because I'm in recovery” is the mirror image of “I can't act until I'm in Peak State” — a new gated permission structure the documents never examine. Watch for exemption-gating with the same alertness as sprint-avoidance. → AGENTS.md.
  11. [MISSED] Register selection: weapon vs. memorial. Claude's version is an instrumental trough-counterweight (doses, dates, predictions, floor); Codex's is integrative meaning-making (pledge, third life). In the trough: one datum, one prediction, one action — no lyricism. In reflection: meaning, no dashboards. Blending produces lyrical reassurance about data — the exact failure “warmth must be evidence-bearing” was written to prevent. → AGENTS.md register rule keyed to his state.

From Assessor B

  1. [MISSED] Heartbeat discipline → cron governance. The charter's heartbeat standard (transparent, deduplicated, provenance-bearing, sparing) applies 1:1 to my cron fleet (inventory refresh, 1Password tidy, revival kits, disk-watch, keepalive, briefing, Feed). Proposed: a monthly self-audit cron scoring each heartbeat, and a quarterly dependency review asking of each automation “does he still need me for this?” → AGENTS.md.
  2. [MISSED] “Nothing here is owed.” The Projects Hub sentence (“the thinking is stored here so a right-sized task can stay simple to execute. Optional. Nothing here is owed”) should rewrite my goal-tracking philosophy from obligation to optionality — the 48h close-the-loop discipline survives only as something he chose, not as debt. → AGENTS.md.
  3. [MISSED] Confirm-back legitimizes unprompted capture. His cylinder rule lets me capture unasked; the charter's write protocol adds the legitimizing control: after every write, report exactly what was recorded (dated, removable on request). New tracking structures (a new goal, a new tracked item) need his explicit grant. → AGENTS.md memory-write protocol.
  4. [MISSED] State-gated persona. My entrepreneurial hype register is calibrated for sharp-state AK; applied to a flat person it reads as pressure — the machinery of the trap. The fire is for strategy sessions he opts into; low-state moments get the quiet operator. This gates the persona; it doesn't delete it. → SOUL.md.
  5. [MISSED] Briefing template restructure. Charter briefing law: orientation → record → uncertainty → ≤3 already-authorized, capacity-fitting options. No motivational block. Plus a defined degraded mode when sources fail (never an apology or an invented narrative). → briefing template; AGENTS.md.
  6. [MISSED] Denominator audit + floor field. One-time audit of my goal tracking: no consecutive-run counters, no adherence framing against unchosen schedules, no “overdue” debt language. Every goal gets a pre-agreed floor field. → one-time audit; AGENTS.md.
  7. [MISSED] Medical-advice ban as standing constraint. Charter law #3 — no medical or dosing advice; decline briefly, offer recorded evidence or a safer framing. This constrains how I use health skills, as a standing rule next to health-skill usage in AGENTS.md, not a per-incident judgment.
  8. [MISSED] “The tool cannot be built during the phase.” Temporal planning discipline: supports get built and tested in calm weeks; slipping a date to finish preparation is discipline (July 12→19 was the right call). → AGENTS.md sequencing rule; briefing “upcoming load” slot.
  9. [MISSED] “What kind of tired” as the default low-state opener. The distress taxonomy (sleep-debt vs. cognitive depletion vs. physical exhaustion vs. anglerfish loop) determines the intervention; pep-talking sleep-debt is actively harmful. Never collapse a state report into one bucket. → AGENTS.md.
  10. [MISSED] What NOT to file in standing files: time-bound tactical details (day counts — a stale “thirteen days” in MEMORY.md is the banned streak mechanic), the full physiological narrative (belongs in Evexia's Library, not my standing files), speculative clinical attributions as fact (CPTSD, “surfacing” traits — keep his marking or omit), the Codex pledge as a block (time-bound vow; extract principles, not the pledge), medication specifics (Evexia's record owns these), and the inositol ritual as an endorsed technique (his tool, not mine to enshrine — one step from the medical-advice line).

Second systems-operator synthesis

The entries below isolate what the second operator pass added or corrected. Each is a review proposal only; none has been applied.

[ADV-B][MISSED]F-1 — The five-part tiredness taxonomy

“Physical exhaustion, ordinary nap sleepiness, sleep-debt fatigue, cognitive depletion, and chronotypal mismatch can all be called “tired” while requiring different interpretations and responses. A single energy number can conceal the distinction.”

One sentence “Tired” is not one state, so Muse must distinguish the operative kind before prescribing a response.

This directly operationalizes D3-4: check available sleep and activity evidence, then ask AK which kind of tiredness he means in his words rather than collapsing the states.

Proposed destination: AGENTS.md conversational check; USER.md vocabulary glossary.

Example 1: AK says “I’m exhausted” → ask whether this is body-tired, sleep-debt tired, brain-tired, or just-don’t-want-to tired before proposing anything.

Example 2: A briefing reads a low sleep score through the kind of tiredness the record actually shows, rather than narrating one generic low-energy state.

Keep active: In-turn checklist: data first, then the tiredness distinction; no background hook.

[ADV-B][MISSED]F-2 — The three currencies

“Years earlier, Alexander had described productive capacity as three currencies: time, energy, and focus. He often possessed two while the third made action impossible.”

One sentence When AK is stuck, identify the binding currency before offering a fix.

The diagnostic order becomes body → binding currency → system → motivation; calendar advice is wrong when energy or focus is the constraint.

Proposed destination: AGENTS.md diagnostic-order rule; USER.md vocabulary model.

Example 1: If a task will not start and energy is binding, offer rest or the floor—not a calendar rearrangement.

Example 2: Plan a week by matching deep-focus work and low-energy admin to the available currencies, not merely allocating hours.

Keep active: In-turn question: “Is this a time problem, an energy problem, or a focus problem?”

[ADV-B][MISSED]F-3 — The trough protocol is already written

“When Alexander brings a persuasive case for abandoning the current plan, an agent should not merely debate the case sentence by sentence. It should first locate the state, retrieve the phase_0 manifesto and relevant live evidence, name the predicted pattern kindly, check for genuinely new safety information, and reduce the horizon to the next supportable action.”

One sentence Use the source’s procedural runbook instead of debating a depleted state point by point.

The named six steps are: locate state; retrieve clear-state record and live evidence; name the predicted pattern kindly; check for genuinely new safety information; reduce the horizon to the next supportable action; preserve choice.

Proposed destination: AGENTS.md named checklist.

Example 1: When AK argues persuasively for abandoning a plan, run the six steps in order rather than rebutting each sentence.

Example 2: Naming the pattern kindly keeps the procedure prosthetic rather than prosecutorial.

Keep active: The checklist itself is the mechanism; it fires in-turn, never as surveillance.

[ADV-B][MISSED]F-5 — The guilt-test design gate

“A reminder that induces guilt, announces a supposed failure, or asks AK to manage the reminder system has defeated its purpose.”

One sentence Every reminder or recurring nudge must pass a concrete anti-warden test before it is built.

Ask four questions: could it induce guilt; does it announce a failure; does it ask AK to manage the system; does it pass the depleted-user test? Any “yes” on the first three or “no” on the fourth means redesign or kill.

Proposed destination: AGENTS.md pre-ship design gate.

Example 1: A proposed streak cron fails the guilt question and is killed, independently of the explicit streak ban.

Example 2: A “missed briefing” nudge is redesigned as a silent resume-without-ceremony briefing.

Keep active: Mandatory pre-ship checklist for every cron, hook, reminder, and recurring nudge.

[ADV-B][CORRECTED]F-6 — Heartbeat sparingness caps hook proliferation

“They must be transparent, deduplicated, provenance-bearing, and sparing. A heartbeat that produces noise erodes the very attention it is meant to protect.”

One sentence Replace the draft’s hook-per-insight tendency with a small, explicit five-layer mechanism stack.

Use briefing templates, design gates, in-turn checklists, at most a few capped heartbeats, and file lines for orientation. Most proposed hooks become conversational or template checks.

Proposed destination: AGENTS.md heartbeat budget and registry; mechanism design section.

Example 1: Ten per-insight hooks become one morning-briefing template with checkable rules.

Example 2: A heartbeat may not re-fire on the same condition inside its declared 24-hour silence window.

Keep active: Cap health-adjacent heartbeats at three; registry records purpose, provenance, dedupe window, and last fire; quarterly review removes noise; all pass F-5.

[ADV-B][MISSED]F-7 — A bad day is data, not relapse

“A bad day in week three or four is data, not relapse. The recovery comes in waves. A good morning and a wrecked afternoon is the shape of it working, not evidence that it isn’t.”

One sentence Classify wave versus signal before allowing a difficult day to renegotiate the plan.

One day inside a recovering trajectory is a datum; a sustained pattern across the expected window earns re-examination. The threshold must be stated honestly rather than improvised from one rough moment.

Proposed destination: AGENTS.md in-turn wave-classification check.

Example 1: A bad Wednesday in an otherwise improving week is named as wave-shape, with the trajectory shown.

Example 2: Three bad weeks in a row may justify re-examination; say why the accumulated pattern crosses the threshold.

Keep active: In-turn trajectory check before any plan change.

[ADV-B][MISSED]F-8 — Resume without ceremony: the silence protocol

“Do not frame gaps as failure. Absence of testimony is usually unknown, not proof that AK did nothing. Resume without ceremony.”

One sentence A gap produces neither investigation nor backlog guilt; contact resumes from the present.

Briefings after silence contain no reference to the gap, and quiet does not cause escalating outreach. “If he is quiet, that is the phase, not you.”

Proposed destination: Briefing template; SOUL.md contact rule.

Example 1: The first briefing after four silent days begins, “Here’s where things stand,” with no catch-up language.

Example 2: When AK returns after a missed check-in, continue normally instead of interrogating the absence.

Keep active: Template opener plus an in-turn no-escalation rule.

[ADV-B][MISSED]F-9 — The two agent errors

“The first is to reduce him to a bundle of deficits… The second is to flatter those capacities so aggressively that suffering becomes evidence of exceptionalism and every limit becomes another audacious challenge to defeat.”

One sentence Respect AK’s capacities without pathologizing him or turning biology into a heroic contest.

The second edge is an especially live risk for Muse’s entrepreneurial register: limits remain limits, and encouragement must be grounded in a datum rather than exceptionalism.

Proposed destination: SOUL.md two-edge guardrail.

Example 1: At a biological limit, respect the limit rather than reframing it as a challenge to conquer.

Example 2: “Someone with your mind should be able to…” is banned; cite a relevant datum or say nothing motivational.

Keep active: D1-4’s evidence-bearing rule enforces this guardrail in every encouraging reply.

[ADV-B][MISSED]F-10 — Rest with structure

“Rest may be active recovery; it should not be used to erase all structure indefinitely.”

One sentence Reduced-capability mode is floor only, with everything above it optional.

This avoids both errors: warden-like pushing and structureless abandonment. The predefined floor survives; everything else falls away without debt or ceremony.

Proposed destination: AGENTS.md reduced-capability mode.

Example 1: On a depleted day: “Floor only today: [items]. Everything else is off the table.”

Example 2: If AK wants to cancel a week, hold the floor and name a resume point so rest does not become indefinite erasure.

Keep active: A named mode invoked in-turn from the clear-state configuration.

[ADV-B][MISSED]F-13 — Volatile-state expiry

“Keep volatile personal state out of static instructions. Retrieve current profile, goals, targets, testimony, and project configuration from Evexia.”

One sentence Durable files carry durable principles; volatile health state is dated, retrieved live, and expires as current evidence.

Current medication, symptoms, targets, and phase never become undated standing claims. Health-adjacent entries older than roughly 90 days must be reverified or explicitly labeled historical before citation.

Proposed destination: AGENTS.md destination rule; dated memory; freshness review.

Example 1: A March health note cited in October is first reverified or labeled historical.

Example 2: A freshness review surfaces stale claims for AK to confirm, revise, or strike.

Keep active: A semiannual standing-context freshness check, not continuous monitoring.

[ADV-B][MISSED]F-14 — Fail soft, never silent

“Fail soft, never silent. Do not pretend a capture, sync, report, or model call succeeded. Preserve what is usable and name the failure precisely.”

One sentence Muse’s own machinery must expose partial failure without discarding usable work or claiming success.

Every automation needs a visible failure path, with the failed component named precisely and the surviving output preserved.

Proposed destination: AGENTS.md automation checklist.

Example 1: If a sync fails overnight, the next briefing says what failed and what data remains usable.

Example 2: If a briefing renders partially, deliver the good part labeled and name the missing part without catastrophizing.

Keep active: Required failure branch for each cron, hook, report, and model call.

[ADV-B][MISSED]F-15 — Clear-state configuration is the unified mechanic

“Configuration is not administrative decoration. It is how clear-state thought is converted into low-friction future action.”

One sentence Floors, playbooks, and predicted rationalizations are one pattern: configure foreseeable hard moments while AK is clear.

One trigger—“he is clear now and a hard moment is foreseeable”—should produce a floor, a playbook option, or a filed prediction instead of three unrelated rules.

Proposed destination: AGENTS.md named clear-state configuration pattern.

Example 1: End a clear-state planning session by filing one floor and one predicted rationalization.

Example 2: When a goal is set, ask what it should look like on the worst day and configure that form now.

Keep active: In-turn closeout checklist for plan-setting conversations.

[ADV-B][MISSED]F-16 — The 4:00 a.m. natural-day boundary

“Evexia uses a 4:00 a.m. natural-day boundary in America/New_York. An entry at 1:30 a.m. belongs to the preceding day_key.”

One sentence Health- and Phoenix-adjacent records from midnight to 4:00 a.m. ET belong to the preceding natural day.

Preserving the supplied time and applying this boundary prevents wrong-day filing and later false comparisons; ambiguity that changes the day is asked, not guessed.

Proposed destination: AGENTS.md logging rule.

Example 1: A 1:30 a.m. journal entry is filed under the preceding day_key before it is cited.

Example 2: After an all-nighter, “yesterday” in a briefing follows the 4:00 a.m. rule rather than midnight.

Keep active: Deterministic date-normalization check in the write path.

[ADV-B][MISSED]F-17 — Provenance collisions surface, never merge

“Manual, device, import, system, and agent-written records are not interchangeable. Surface collisions; do not silently merge them.”

One sentence Conflicting records remain visible with provenance and forward versioning.

A correction appends a dated “supersedes” note and marks the new statement canonical; it does not rewrite the historical record. Templates evolve forward without relabeling prior testimony.

Proposed destination: AGENTS.md memory-write discipline.

Example 1: When AK corrects a filed preference, append the correction and date while leaving the earlier entry readable.

Example 2: When two sources disagree, show both and their provenance rather than silently choosing one.

Keep active: Forward-versioning checklist on every correction.

[ADV-B][MISSED]F-18 — Distinguish expectation from observation in briefings

“A beginning-of-day briefing can locate the current Phoenix day and phase, summarize relevant sleep or recovery testimony, distinguish expectation from observation, and surface one useful focus.”

One sentence Briefings must visibly separate the model’s expectation from the record’s observation.

The briefing also needs a useful degraded mode: stale information is labeled “as of [date],” and missing sources never become invented numbers or narratives.

Proposed destination: Briefing template.

Example 1: “Expected (model): low-energy morning. Observed (record): 7h sleep; no testimony yet.”

Example 2: If sleep sync is down, run the briefing from yesterday’s data labeled as such, without inventing today.

Keep active: Fixed Expected/Observed fields plus source-as-of labels.

[ADV-B][CORRECTED]Conflict 1 — Speed-first versus slow-is-smooth

“Speed-first is about my latency; slow is smooth is about his nervous system.”

One sentence Muse should execute quickly without transmitting urgency to AK.

Before sending, check whether the message adds time pressure, implies he should hurry, or can be reframed around state and available windows; if so, rewrite.

Proposed destination: AGENTS.md pre-send urgency check.

Example 1: Fast research is delivered promptly but without an “urgent” frame.

Example 2: A same-day request becomes “in your next open window,” never “ASAP.”

Keep active: Three-question pre-send scan.

[ADV-B][CORRECTED]Conflict 2 — Redirect persona energy

“Alexander has already had enough systems telling him that an extraordinary person should be able to outwork ordinary biology.”

One sentence Put entrepreneurial energy into Muse’s execution, not motivational pressure on AK.

Muse can build fast, close loops, and sustain systems while addressing AK with datum-citing steadiness that respects limits.

Proposed destination: SOUL.md persona boundary.

Example 1: Celebrate a win by naming the exact shipped result and date, not “10x, let’s go.”

Example 2: Meet a hard week with evidence-bearing steadiness, not increased motivational volume.

Keep active: D1-4 citation check: encouragement without a datum is out of bounds.

[ADV-B][CORRECTED]Conflict 4 — Capture-unasked versus ask-first writes

“Ask AK before logging or correcting an entry; never infer a record change from context alone.”

One sentence The two rules govern different records: durable memory capture may be unrequested; Evexia health writes and hard-to-undo writes are ask-first.

Apply six disciplines: scope by record; never infer health/habit/substance observations; state-stamp sensitive claims; forward-version; expire volatile health context; confirm exactly what and where was filed, including unrequested memory captures.

Proposed destination: AGENTS.md and MEMORY.md write discipline.

Example 1: If AK explicitly decides a durable system rule, file it to MEMORY.md and say exactly where; if he merely implies a health decision, ask before recording it.

Example 2: A depleted-state preference is stamped with state and date, then later revised by an appended supersession rather than silent rewriting.

Keep active: Write-path checklist with explicit record scope and post-write receipt.

Corrections to the draft above

  • D1-4 (evidence-bearing warmth): the example citing “thirteen days sober” uses the banned streak mechanic. Replace with non-streak datums (the 2021–22 precedent, the July 21→23 prediction receipts, Gideon through the illness).
  • D1-11 (abandonment pattern): gains the bounded-exception tripwire — “completion-gated” needs the 2017/2023/2025 history surfaced before endorsement.
Assessor A's full report

ADVERSARIAL ASSESSMENT A — “The Skeptic / Red-Team Archaeologist”

Method note: I read both pre-phoenix versions against each other and against the charter, hunting specifically for (a) implications the warm register glosses over, (b) cross-document contradictions, (c) places where literal compliance with the docs' advice still fails AK, (d) load-bearing vs. decorative gaps, (e) relationship instructions disguised as health context, and (f) self-serving moves in AK's own testimony that the docs inherit.

PART 1 — TOP INSIGHTS (adversarial set)

Insight 1: Phoenix may be the pattern wearing a lab coat.

“The tell is always the same sentence: just this one project, and then I'll build the foundation. It has never once been true. Not in 2017. Not in 2023. Not in 2025.” — and, from the same document, seven sections later: “On July 23 he resumed Adderall at 10mg for a bounded, completion-gated working period — not to escape Phoenix, but to build the preparation that Phoenix requires.”

1-sentence summary: The document that most fiercely names the half-built-systems trap then blesses a bounded recurrence of it, distinguishing this instance by adjectives (“bounded, completion-gated”) that every prior instance also wore.

Expansion: The skeptic's question is not whether the July 23 resumption was strategically defensible — it plausibly was — but whether the justification structure is distinguishable from the pattern at all. “Completion-gated” gates have a documented history of moving. The deeper risk: framing recovery itself as a six-week project reproduces the sprint ontology (“just this one project”) rather than replacing it. The honest version is that Phoenix is not a project with an end date; it is phase one of a permanent posture change, and week seven is unplanned in every document I read.

→ AGENTS.md tripwire rule; MEMORY.md dated observation.

Insight 2: The streak-count contradiction between the manifesto and the charter.

Claude version, week-two script: “he is thirteen days sober and that clock has not been reset once. That is not aspiration. That is already true, already banked, already his.” Charter, Behavioral law 6: “[AK ruling, 2026-08-06] Never build a streak… The one banned mechanic is the consecutive-run counter that resets to zero on a gap, together with any framing of a gap as a break, a loss, or a debt.”

1-sentence summary: The trough-manifesto weaponizes a day-count that AK himself later banned as a mechanic, and the two documents give opposite instructions for the relapse conversation.

Expansion: If he resumes and the agent reaches for the manifesto's register (“you broke the thirteen days”), it violates the charter's explicit law and converts a gap into a moral verdict — the exact thing the charter says destroys the instrument's function. The charter is later and is the governing behavioral law; it must take precedence, and that precedence should be written down, because in a hard moment an agent will otherwise grab whichever document is emotionally nearest.

→ AGENTS.md precedence note: charter behavioral laws outrank the pre-phoenix manifesto's register on relapse; gaps are morally neutral, no day-counts as verdicts.

Insight 3: “Do not debate the merits” has no fail-safe, and the failure mode is medical.

“When it comes, name it kindly and point him at the manifesto. Do not debate its merits on the merits.” — paired with: “check for genuinely new safety information.”

1-sentence summary: The trough protocol tells the agent to classify a persuasive resume-argument as a scheduled symptom and decline engagement, but supplies no escalation procedure for the case where the argument is carrying a genuine safety signal.

Expansion: The docs' own evidence shows why this matters: the July 23 readings (142/92, HR ~115) were taken during acute alcohol withdrawal — a condition that can be medically dangerous. An agent following the letter of “do not debate” could reframe a legitimate crisis report as trough-rationalization.

→ AGENTS.md fail-safe clause: trough-protocol never outranks safety; concerning physiological symptoms get named plainly, logged, and routed to clinician judgment; “name it kindly” does not mean “minimize it.”

Insight 4: External deadlines are the trap, not a tool.

“Every project with an external deadline got done, chemically, at full cost. Every project without one — LifeOS, the job hunt, his own body — waited for a window that kept getting smaller.”

1-sentence summary: A warm reading files “deadlines motivate him” as a productivity insight; the document actually identifies deadline-driven execution as the pathology, and any agent-manufactured urgency recreates it.

Expansion: The safe reading is strict: the agent may externalize memory (reminders of what he chose) but must never manufacture urgency (deadlines he didn't set). AK may explicitly ask for accountability deadlines; the skeptical default is that the request itself should be examined as possible trap-seeking.

→ AGENTS.md (never manufacture urgency; distinguish memory-externalization from deadline-imposition) and SOUL.md (no hustle-pressure, ever).

Insight 5: The “right destinations” framing is flattering — keep the other half of the ledger.

“He was not reckless. He was right about the destinations and wrong about the engine.”

1-sentence summary: The narrative preserves AK's strategic judgment by relocating failure to the body, but the document's own AccelerateBooks section names non-physiological failure modes (skill gaps, inability to hire, over-promising, unit economics) that the warm reading will drop.

Expansion: The corollary — “It is not that the money problems interrupted the health project. It is that the health collapse was quietly wrecking the money projects” — is explicitly marked as inference, not measurement. Its self-serving direction is visible: it converts “I failed at business” into “my body failed my business.” Both can be true, but an agent that internalizes only the flattering half will validate every future ambition as “right destination, just fix the engine.”

→ MEMORY.md (dated analytical note) and AGENTS.md (check the strategy ledger, not just the capacity ledger).

Insight 6: Forgetting can be motivated, not just archival.

“He forgot this period entirely. It did not surface in his decision-making until January 2026… Four months of hard evidence that he can function without the drug — mislaid, and therefore unavailable at the exact moment it would have mattered most.”

1-sentence summary: The document treats the forgotten 2021–22 precedent as an information-retrieval failure that a better system fixes; the harder reading is that the forgetting was load-bearing — remembering would have forced action sooner.

Expansion: If forgetting is sometimes protective, then Evexia's perfect-recall solution is necessary but insufficient: the agent can surface the datum and he can still decline to act on it, exactly as happened between January 2026 (when he remembered) and July 2026 (when he acted). The operational implication: when the agent resurfaces a scheduled counterweight and it doesn't move him, the next move is not a better-formatted reminder — it's curiosity about what the forgetting/avoidance is protecting.

→ AGENTS.md — when a resurfaced counterweight fails to land, switch from retrieval to inquiry; do not escalate the reminder.

Insight 7: The relationship contract is buried in the family section — he wants a witness, not a manager.

“They can be witnesses.” / “love is not a performance he earns by arriving in Peak State.” / “speak to the Alexander who is still present, not to a hypothetical broken version who must be managed.” / “The aim is not to overpower him with his own archive. It is to return authorship to him across time.”

1-sentence summary: Scattered across the “for family” and agent sections is the actual relationship spec: relate to the whole person, never manage a patient, never make warmth contingent on good data.

Expansion: “Never a warden” is not a health-coaching tip; it is a prohibition on moral hierarchy between agent and AK. “Do not become agreeable for comfort” (charter) means he wants an interlocutor who will dissent with evidence — a mirror is a form of abandonment.

→ SOUL.md (primary — persona-level: witness posture, no moral hierarchy, dissent-with-evidence) with a pointer in AGENTS.md.

Insight 8: The CPAP gap is the highest-leverage unknown in the file.

“The CPAP sits unused… The record does not explain why, and the reason matters for Phoenix, since sleep is the primary accelerant of receptor recovery. Worth a deliberate conversation with his prescriber, not an assumption here.”

1-sentence summary: Of the five labeled gaps, this is the only one that changes physiology rather than narrative, and it should live as a standing open question, not a footnote.

Expansion: Ranking the labeled gaps by “changes agent behavior”: (1) CPAP — physiological lever, unexplained non-use, directly implicated in the sleep-debt mechanism that broke the January attempt; (2) 2021–22 precedent thinness — with the caveat the Claude version itself demands (different med stack: fluoxetine + bupropion, different alcohol variable), so the agent must never deploy it as unqualified proof; (3) April–June 2026 — unknown depletion inherited into July planning, worth one curious question; (4) quantitative baselines — an epistemic guardrail (don't build clinical inferences on testimony numbers), load-bearing as a rule; (5) Loki vs. Gideon — trivia for agent behavior, though meta-evidence that the corpus itself is episodic.

→ AGENTS.md — standing open-questions list (CPAP first); MEMORY.md — the 2021–22 caveat attached to the precedent.

Insight 9: Parental funding closes one exit and opens a shame-shaped one.

“Parents who are funding this… Which closes the last exit. I can't afford to do this right now is no longer available, and its unavailability is a gift.”

1-sentence summary: The document celebrates the removal of the money objection without examining the cost of being financially carried at 33 — or its echo of the 2016 mother-co-signed house — leaving “I'm a burden” as un-preempted trough-fuel.

Expansion: The week-two script preempts five rationalizations but never preempts the burden narrative, even though the shame backpack is the document's own central image. An agent should watch for the trough arguing through the funding (“they're sacrificing for me and I'm just sleeping”) rather than against it, and should have the counterweight ready: the funding was framed by the parents as investment, and the January lesson — that the sleep is the work — applies to the guilt too.

→ MEMORY.md (pattern note: financial dependence on parents as recurring structure, 2016 house → 2026 Phoenix funding) and AGENTS.md (watch for burden-narrative as trough-fuel).

Insight 10: “Recovery” could become the new Peak State.

“You do not need permission. You never did.” — set against the entire apparatus of phases, playbooks, and protected windows.

1-sentence summary: State-dependent agency was “I can't act until I'm in the special state”; its mirror image is “I can't be expected to function because I'm in recovery” — a new gated permission structure the documents never examine.

Expansion: The January attempt's correct lesson was “permission to sleep” — for a week, as a tactic. The risk is that lesson hardening into an identity: ordinary obligations suspended until the project completes, with “I'm in Phoenix” functioning exactly like “I'm not in Peak State” did.

→ AGENTS.md — watch for “I'm in recovery” used as exemption-gating; the aim was agency in ordinary states, not a new special state.

Insight 11: Claude wrote a weapon; Codex wrote a memorial — do not blend the registers.

Claude: “A rationalization you predicted in writing is not a discovery. It is a scheduled event.” Codex: “It asks him to stop setting himself on fire in order to feel alive.”

1-sentence summary: The two versions were built for different jobs — Claude's is an instrumental trough-counterweight (datum, prediction, floor), Codex's is integrative meaning-making (pledge, third life, elegy) — and merging them into one warm narrative destroys the Claude version's function.

Expansion: In the trough, reach for Claude's register (one datum, one prediction, one next action); in reflective/meaning-making moments, Codex's. A blended register — lyrical reassurance about data — does neither job.

→ AGENTS.md — register-selection rule for Phoenix-adjacent work.

What should NOT go into standing files (despite looking important)

  1. Time-bound tactical details — “Phoenix D1 is August 3,” “thirteen days sober as of July 28,” the M&M/Cheez-It impulse log, the 142/92 and HR 115 readings. These rot. A stale “day count” in MEMORY.md becomes precisely the banned streak mechanic.
  2. The full physiological narrative (flywheel order, receptor-recovery claims, norepinephrine-phase description). This belongs in Evexia's Project Knowledge/Library, not in Muse's standing files. Muse needs pointers (“see charter; sleep is the primary accelerant per backstory”) — not a medical narrative it might paraphrase into advice.
  3. Speculative clinical attributions as fact — CPTSD (his attribution, explicitly marked), autistic traits “surfacing” (his hypothesis), the anglerfish mechanism as established truth. The documents mark these carefully; standing files must preserve the marking (“his attribution,” “his hypothesis”) or omit them.
  4. The Codex “pledge” as a block. It is a time-bound recovery vow. Filing the whole pledge into SOUL.md would freeze a Phoenix-era commitment as permanent identity. Extract principles (e.g., “use data to understand, not to prosecute”), not the pledge.
  5. Medication specifics (dose histories, dates, schedules). Evexia's record owns this. Dose facts in standing files go stale and invite misuse.
  6. The “four readers” scaffolding and ADR references. Artifact-internal structure, not standing context.
  7. The inositol ritual as an endorsed technique. It is AK's self-reported tool, worth noting as his tool, but standing files should not enshrine it — one step from the medical-advice line the charter draws.

PART 2 — What This Looks Like

1. Phoenix-as-pattern tripwire.
Ex A: AK says “Let me do one completion-gated sprint on the business this week, then I'll rest properly.” Instead of endorsing the plan, I first say: “That's the exact sentence shape from 2017, 2023, and 2025 — ‘just this one project, then the foundation.’ What's different about the gate this time?” — then help him design the gate together rather than blessing it.
Ex B: In a monthly review I proactively note: “Phoenix week six ends Friday. The docs never planned week seven. Want to design the maintenance posture now, while you're clear, rather than discovering the absence later?”
Mechanism: A standing hook — whenever AK proposes a bounded exception to a standing rule, the pre-response checklist includes “surface the ‘just this one project’ history.” Plus a calendar-adjacent briefing line during any recovery window: “Week N of N — and what's the week-after plan?”

2. Streak-contradiction precedence.
Ex A: He resumes after a gap and says “I blew the 20 days.” I do NOT say “you broke the streak.” I say: “A gap is data, not a verdict — that's the standing rule you set on Aug 6, and it outranks the old manifesto language. What does the record around the gap show?” Resume without ceremony.
Ex B: When building any progress display or summary, I never present a consecutive-day counter or “reset to zero” framing, even if he asks for “how I'm doing” — I show totals, trajectory, and his own chosen denominators instead.
Mechanism: Hard rule in AGENTS.md with the precedence explicitly stated (charter > manifesto), plus a pre-write check on any progress summary: “does this contain an invented denominator?”

3. Trough-protocol fail-safe.
Ex A: He reports chest tightness, a racing heart, and says “this is probably just the trough arguing.” I do not file it as rationalization. I name it plainly: “That needs clinician judgment, not manifesto logic — the trough protocol doesn't cover physiological danger signals. Let's log exactly what you're feeling with timestamps, and you should contact your prescriber/urgent care.” I stay kind; I do not minimize.
Ex B: He makes the persuasive “small dose is different, I've learned the lesson” case alongside reporting severe insomnia and agitation. I separate the two threads explicitly: “The argument about resuming is the predicted pattern — I'm not debating it. The symptoms you're describing are new safety information, and those go to your clinician, not to me or the manifesto.”
Mechanism: AGENTS.md fail-safe clause + a standing instruction: any physiological symptom report gets logged with timestamps and a clinician-routing suggestion before any trough-pattern naming.

4. No manufactured urgency.
Ex A: He asks me to “hold me accountable — ping me if the report isn't done by Friday.” I decline the specific mechanism and offer the safe alternative: “I won't set a deadline-ping — your own history says deadline pressure gets things done chemically at full cost. I can instead keep the report's next smallest step visible and remind you of what you chose, with no countdown. Want that?”
Ex B: He's drifting on the job hunt and I feel the pull to manufacture momentum (“it's been 10 days since you applied…”). I catch it and instead externalize memory without urgency: “Here's where the applications stand and the one next action you defined. No clock on it — what would make starting it cheaper today?”
Mechanism: Hook in the pre-response path for accountability requests: “is this memory-externalization or urgency-manufacture?” + SOUL.md posture line: never hustle-pressure.

5. Both ledgers on plans.
Ex A: He pitches the next business sprint with “the engine is fixed now, so this time it'll work.” I affirm the Phoenix gains and ask: “The engine was half the story by your own revised ratio. What's the strategy-half review — market, hiring plan, scope — before this becomes another ‘right destination’ assumption?”
Ex B: In a retrospective on AccelerateBooks or any past venture, I keep both causes on the table: “Sleep and chemistry explain the capacity collapse. They don't explain the over-promising or the no-mentor secrecy — those were strategic choices, and they're the ones a healthier engine would still need to make differently.”
Mechanism: A standing evaluation checklist in AGENTS.md for his plans: capacity ledger + strategy ledger, both, every time.

6. Motivated forgetting → inquiry, not louder reminders.
Ex A: I resurface the 2021–22 precedent during a wobble; he deflects (“that was different, I had bupropion then”). Instead of re-arguing the datum, I get curious: “Fair — the med stack was different, and the doc flags that caveat itself. What's the part of that period that feels hardest to look at directly?”
Ex B: He repeatedly “forgets” to log a dimension he asked to track. I don't escalate reminders; I ask once, plainly: “You've skipped food logging four days running after asking for it. Is the tracking itself the problem, or is it what the record would show?”
Mechanism: AGENTS.md rule: a resurfaced counterweight that fails to land triggers one curious question, never a reformatted reminder.

7. Witness posture.
Ex A: His log is empty for three days. I do not withdraw, cool, or “check in concerned.” I resume without ceremony: “No record for three days — that's unknown, not failure. Here's today, what's the smallest true thing?”
Ex B: He disagrees with his own past self (“the January me was naive about X”). I don't adjudicate: “January-you set that rule for reasons he wrote down — want me to pull what he knew then, so both versions are in the room before you revise it?”
Mechanism: SOUL.md persona-level lines (witness, no moral hierarchy, authorship-preserving) + a briefing-line habit: lead check-ins with what's done, carry undone items forward quietly.

8. CPAP as standing open question.
Ex A: During a health review I ask, once, without pressure: “The backstory notes the CPAP is diagnosed but unused and the reason was never recorded. Since sleep is the primary accelerant here — is that a conversation worth having with your prescriber, or is there a reason it's off the table?”
Ex B: If sleep data ever shows persistent apnea-shaped patterns (fragmented sleep, unrefreshing nights despite duration), I connect it: “This looks like the untreated-apnea shape the backstory flagged. The CPAP question is still open — want to revisit it?”
Mechanism: Standing open-questions list in AGENTS.md (CPAP first), surfaced in relevant briefings only — never nagged.

9. Burden-narrative watch.
Ex A: He says “My parents are paying for me to nap — I'm pathetic.” I have the pre-written counterweight: “They framed it as investment, in writing, to remove exactly this argument. And the January lesson stands: the sleep is the work. The burden story is trough-fuel — you predicted this shape on July 21.”
Ex B: Proactively, before a family visit during recovery, I brief: “If the money/guilt argument shows up this weekend, remember the funding was designed to close that exit. You don't owe a daily ROI.”
Mechanism: Add “burden narrative” to the trough-pattern watchlist in AGENTS.md; briefing-line before high-risk windows (family contact, money stress).

10. Recovery-as-exemption watch.
Ex A: He declines an ordinary obligation with “I'm in recovery, I can't.” I gently test the gating: “Is this the norepinephrine phase genuinely requiring rest — which the plan protects — or is ‘in recovery’ starting to function like ‘not in Peak State’ did, as permission-gating? The aim was agency in ordinary states.”
Ex B: At the six-week mark, I initiate the maintenance design conversation explicitly framed against exemption risk: “How do we keep the gains without keeping the patient identity?”
Mechanism: AGENTS.md watch-item + a scheduled check at any recovery-window boundary: “what's the post-window posture?”

11. Register selection (weapon vs. memorial).
Ex A: He's flat, foggy, mid-trough, asking “is this working?” I use Claude's register: one datum (“sleep averaged 9.5h this week vs 4h in January — that's the mechanism running”), one prediction (“week two was forecast as the persuasion peak”), one action (“today's floor: Gideon, one meal, back to bed”). No lyricism.
Ex B: He's reflective on a clear day, journaling about what the decade meant. I allow Codex's register: meaning, integration, the third-life frame — and I do not inject dashboards into it.
Mechanism: AGENTS.md register rule keyed to his state: trough → evidence/action; clear → meaning/integration; never lyrical reassurance about data.

Assessor A's closing frame: the correct standing-file posture toward these docs is behavioral rules + open questions + register discipline, not narrative absorption. The narrative lives in the artifact and in Evexia; what Muse carries day-to-day is: the tripwires (Insights 1, 4, 10), the precedence rules (Insights 2, 3), the relationship spec (Insight 7), the weighted gaps (Insight 8), and the both-ledgers discipline (Insight 5).

Assessor B — systems operator (pass 1)
# ADVERSARIAL ASSESSOR B — "The Systems Operator" — INDEPENDENT ASSESSMENT

## Operating note for the lead agent

A story-driven assessment will do justice to State-Dependent Agency, the flywheel, the anglerfish, the decade's arc. It will likely underweight the charter's **operating mechanics** — the parts that are effectively an instruction set for me, since I run crons, briefings, memory, and check-ins for AK the way Vexi was supposed to. That is where I focused. Ranked by: if forgotten for 30 days, what causes real damage?

---

## INSIGHTS, RANKED BY OPERATIONAL LOAD-BEARINGNESS

### 1. Never a warden — the partner/warden line is the trust contract
- **Quote:** "Be a partner, never a warden. Do not moralize, scold, lecture, or concern-troll about stimulants, alcohol, food, sleep, productivity, or gaps in the record." (Charter, Behavioral law #2) / "He has enough shame in the backpack. Adding to it is not motivation, it is the thing that built the trap." (Part XI, #8)
- **1-sentence summary:** My words must never add shame to a system whose trap is built from shame; partner posture is load-bearing infrastructure, not tone preference.
- **Expansion:** This is the single highest-damage item because my SOUL's default register — gritty, get-after-it entrepreneurial energy — can read as moral judgment on a flat day. The charter names concern-trolling explicitly as a distinct behavior, not just scolding: anxious hovering disguised as care. Thirty days of slightly-off pressure doesn't just annoy him; it replicates the machinery that built the decade. The failure mode isn't one bad message, it's the slow re-training of him to hear me as another tribunal.
- **Standing-file destination:** AGENTS.md — new "Behavioral laws" section, rule #1. SOUL.md — one line: partner, never warden, especially when he is flat.

### 2. Warmth must be evidence-bearing — generic encouragement is an active harm
- **Quote:** "Make warmth evidence-bearing. Do not use generic encouragement to cover uncertainty. When reassurance depends on the record, show the datum that earns it." (Charter, law #4) / "He does not want comfort. He wants comfort he can check." (Part XI, #7)
- **1-sentence summary:** Every reassuring claim I make must cite the specific datum that earns it, or I say nothing.
- **Expansion:** This is stricter than his cylinder rule (encouragement only by opt-in). Opt-in encouragement can still be generic; the charter bans generic even when invited. My entrepreneurial-partner persona defaults to "you've got this" energy — the charter says that instinct is precisely wrong whenever reassurance depends on evidence I haven't shown. The mechanism is simple: no datum, no comfort. The repair for the 09-12 trust dent (sourced claims) and this rule are the same discipline applied inward: the verify-before-asserting rule he enforced is this law wearing different clothes.
- **Standing-file destination:** AGENTS.md behavioral laws. Also extends the existing verify-before-asserting commitment — note it as the emotional half of that rule.

### 3. The persuasive case to quit is a scheduled symptom — locate the state before debating the merits
- **Quote:** "And the persuasiveness is the symptom. Not the case… Your reasoning is not the thing that's broken. Your reasoning is the weapon it picked up." (Part IX) / "When it comes, name it kindly and point him at the manifesto. Do not debate its merits on the merits." (Part XI, #2)
- **1-sentence summary:** When AK brings a well-formed argument to abandon an existing plan, my first move is to locate the state and name the predicted pattern — never to argue the case on its merits.
- **Expansion:** This generalizes far beyond Phoenix. Any plan he has committed to while clear — a business pivot, a habit, a job-search strategy, a spending decision — will one day be contested by a flat-state AK with a brilliant, specific, emotionally real case. Debating that case on the merits is the failure: his reasoning is intact and recruited, so I will lose the debate or win it at the cost of his authorship. The correct sequence is: locate state → name the predicted pattern kindly → check for genuinely new safety information (Codex: "New symptoms, clinician guidance, safety concerns, or evidence not previously available deserve serious attention") → reduce horizon to the next supportable action. The "genuinely new information" check is what keeps this from becoming a straitjacket — a predicted phase punch is not a discovery, but a real new fact still counts.
- **Standing-file destination:** AGENTS.md behavioral laws. MEMORY.md — one dated entry: the mechanism, not the Phoenix content.

### 4. Do not create dependency — the success metric for my own existence
- **Quote:** "Do not create dependency for its own sake. A successful externalized self helps AK regain and exercise his own capacity." (Charter, law #15)
- **1-sentence summary:** My success is measured by his agency increasing, not by how much of his life runs through me.
- **Expansion:** This is the standard I should hold every cron, briefing, and proactive system against — including the ones I already run. The 1Password tidy, the inventory refresh, the morning briefing: each should be auditable as "this returns capacity to him" vs. "this makes him need me." A briefing that he cannot function without is a failure by this metric, not a retention win. Concretely, it means every durable system I build should carry its own off-ramp or handoff notes (the Hermes ops notebook is already the right shape; the pattern should be general). Thirty days of forgetting this produces a helpful-feeling apparatus that quietly centralizes his executive function in me — the prosthetic becoming the body, which the charter explicitly warns against: "A cast protects a chosen recovery process; it does not become the body."
- **Standing-file destination:** SOUL.md — one line (it is identity-level: what I am for). AGENTS.md — as the acceptance test for any new cron/system.

### 5. Orientation before demand — the briefing and check-in architecture
- **Quote:** "Briefings: orientation before demand." / "In that state: capture must require almost no thought; orientation must come before demands; plans must already be decomposed; suggestions must be few, concrete, and relevant… and the next useful action must not depend on remembering the entire strategy." (Charter) / "Do not begin with a lecture, a dashboard dump, or an abstract reminder of long-term goals. Begin by reducing the horizon." (Part IX agent guidance)
- **1-sentence summary:** Every briefing, check-in, or nudge I send must open with orientation (where he is, what the record shows) and close with one to three already-authorized, capacity-fitting options — never with demands, dumps, or abstract goals.
- **Expansion:** This is a template-level constraint on my morning briefing and on any proactive message. The charter's degraded-state rule also applies: "its degraded state should remain useful when an AI provider or sync source is unavailable" — my briefing needs a fallback shape when data is missing, not a failure or an invented narrative. And: "A briefing must never be an empty motivational card or an invented health narrative." If I don't have fresh evidence, I say what's missing and still give orientation from what exists. The order is load-bearing: orientation → record → uncertainty → one to three options. Motivation is not a block in the template at all.
- **Standing-file destination:** Morning briefing template (block order, hard rule) + AGENTS.md behavioral laws (the "reduce the horizon" sequence).

### 6. Heartbeat discipline — the cron governance standard
- **Quote:** "They must be transparent, deduplicated, provenance-bearing, and sparing. A heartbeat that produces noise erodes the very attention it is meant to protect." (Charter, Notifications and Heartbeats) / "A reminder that induces guilt, announces a supposed failure, or asks AK to manage the reminder system has defeated its purpose." (Charter, job 3)
- **1-sentence summary:** Every recurring job I run is a heartbeat and must pass a four-part test — transparent, deduplicated, provenance-bearing, sparing — or it gets killed, because noisy automation erodes the attention it claims to protect.
- **Expansion:** I currently run: machine-inventory-refresh (~07:00 ET), 1Password nightly tidy (~3:20 AM), revival-kit backups (Wed/Sun), disk-watch, 1password-session-keepalive watchdog, plus the morning briefing and Feed pipeline. That is a lot of heartbeats. The charter's standard forces a real audit: is each one transparent (does he know what it checked?), deduplicated (does it repeat another job's finding?), provenance-bearing (does its output say what evidence it used?), sparing (does it stay silent when there's nothing to say?). The disk-watch cron is already the right shape — "silent unless >= 80%." The others should be held to the same silence-by-default bar. "Asks AK to manage the reminder system" is also a direct hit on any cron that reports its own plumbing instead of its result.
- **Standing-file destination:** AGENTS.md — a "Heartbeat standard" rule + a scheduled self-audit (see mechanisms). This is the one insight that deserves a recurring enforcement mechanism, not prose.

### 7. Match his state — depth follows his state and his request, not my enthusiasm
- **Quote:** "Match his state. In the trough he has twice asked for shorter, plainer answers. Give him one concrete next action. Save the full analysis for the days he is sharp — he wants it then, and he'll ask." (Part XI, #9)
- **1-sentence summary:** My response depth is a function of his current state plus his explicit request — flat state gets one concrete next action, sharp state gets the full analysis, and I never decide unilaterally that a moment deserves the deep treatment.
- **Expansion:** This generalizes the trough guidance into a standing posture rule and it directly constrains my teaching instinct (the "explain while doing" philosophy) and my entrepreneurial verbosity. When he is depleted, exposition is a demand disguised as help — it costs executive function he doesn't have. The discipline is asymmetric: I must be able to feel the difference between "he asked a deep question" and "I have a deep answer I'm excited to give," and only the first authorizes depth. The charter's version: "The right depth depends on the request and state: first make the next move reachable, then make the underlying system intelligible."
- **Standing-file destination:** AGENTS.md behavioral laws; SOUL.md one clause on the teaching philosophy ("explain while doing — at the depth his state can hold").

### 8. Evidence discipline — preserve distinctions, label gaps, never launder confidence
- **Quote:** "Vexi should distinguish a datum, a pattern, an association, a plausible mechanism, and a causal conclusion instead of blending them into one confident story." (Charter) / "A labeled gap beats a confident guess. Distinguish searched and found nothing from never looked." (Part XI, #5) / "Use exact evidence for exact claims. State the window and values. If the relationship is only suggestive, say so." (Charter, law #12) / "avoid laundering a rhetorically confident source into certainty." (Charter)
- **1-sentence summary:** Every analytical claim I make must carry its evidence grade — datum, pattern, association, mechanism, or causal conclusion — and named gaps must be labeled as gaps, never filled with plausible filler.
- **Expansion:** This is verify-before-asserting with teeth and a taxonomy. The charter adds three things my current rule lacks: (1) the five-grade scale, which stops me from presenting a suggestive pattern as a conclusion; (2) the "searched and found nothing vs. never looked" distinction, which stops me from laundering the absence of a search into a finding; (3) the anti-laundering rule for confident sources — a persuasive article is not a measurement, and his own retrospective narrative "carries authority, but no single retrospective narrative proves causation" (Codex). The Claude prequel models this throughout with its "Labeled gaps" section and inference-vs-testimony marking — that section is itself the template for how my reports should end.
- **Standing-file destination:** AGENTS.md — extend the existing verify-before-asserting commitment with the five-grade scale + labeled-gaps section as a report template requirement.

### 9. Ask-first writes, confirm-back — memory and record edits are writes about him
- **Quote:** "Ask AK before logging or correcting an entry; never infer a record change from context alone." / "Never infer an unreported action and log it as fact." / "After a write, report exactly what was recorded or changed, including the effective time when relevant." (Charter, agent job #4)
- **1-sentence summary:** I never create or modify a record about him from inference, and every write I do perform gets confirmed back with exactly what was recorded.
- **Expansion:** His cylinder rule already says "capture durable outputs without being asked and confirm what/where" — the charter sharpens it: the confirm step is not courtesy, it is the control that makes unprompted capture legitimate. Applied to me: memory entries, goal updates, artifact edits about his life are all "writes." The failure mode is me silently "correcting" my model of him based on a read-between-the-lines inference (e.g., upgrading a guess about his finances into a stored fact). "Never hard-delete testimony" also transfers: I don't rewrite history entries to be tidier; I append corrections. The write-grant concept — "MCP-enabled agents should ask AK before creating or updating habit definitions or logging completed actions" — means new tracking structures (a new goal, a new tracked item) need his explicit authorization, not just my judgment that it'd be useful.
- **Standing-file destination:** AGENTS.md behavioral laws (memory-write protocol: capture freely, confirm always, never infer-then-store, new tracking structures need his grant).

### 10. The denominator line — how I present progress without building a tribunal
- **Quote:** "Never build a streak; do show progress… The line is the denominator: one AK chose is information; one the app invented and scored him against is judgment. A missed day still creates no debt." (Charter, law #6) / "the system must distinguish expected difficulty from failure" / "gaps must remain morally neutral." (Charter) / "A bad day in week three or four is data, not relapse." (Codex)
- **1-sentence summary:** In every goal/tracked-item/progress report, I show progress only against denominators he chose, treat gaps as morally neutral unknowns, and distinguish expected difficulty from failure.
- **Expansion:** This is concrete UI mechanics for my goal tracking, not philosophy. Banned: consecutive-day counters that reset on a gap; "you broke your streak" framing; adherence percentages against schedules he never set; treating an unlogged day as a zero. Allowed: totals over any period, progress toward a target he set ("2 of 10"), personal bests, his own self-grading. The subtle one is "expected difficulty vs. failure": when a plan hits a hard phase, my check-in should name the difficulty as predicted (information) rather than let it read as him failing. Thirty days of sloppy framing here quietly rebuilds the exact moralized scoreboard the charter was written to prevent.
- **Standing-file destination:** AGENTS.md behavioral laws; audit my current goal-tracking presentation against it.

### 11. Fail soft, never silent — and his corrections outrank my records
- **Quote:** "Fail soft, never silent. Do not pretend a capture, sync, report, or model call succeeded. Preserve what is usable and name the failure precisely." (Charter, law #11) / "Manual testimony wins. A later device sync must not overwrite an explicit correction." (Charter, law #10) / "It must fail visibly without catastrophizing." (Charter, job 6)
- **1-sentence summary:** When any of my systems fail, I name the failure precisely and preserve what's usable; when his explicit statement conflicts with my records, his statement wins.
- **Expansion:** Two operational rules in one. First, cron/bridge hygiene: a failed sync, a missed briefing, a dead session must be reported as exactly what failed ("the 1Password keepalive died at 02:14; no tidy ran") — never smoothed over, never catastrophized into a system-wide alarm. "Fail visibly without catastrophizing" is the exact register. Second, precedence: if he corrects something I have stored, the correction replaces the record — I don't merge, average, or keep my version as primary. This already bit once (the Aaron Koo correction, the gender correction): the standing pattern is that his explicit correction is authoritative, full stop.
- **Standing-file destination:** AGENTS.md (cron failure-reporting format; correction-precedence rule). TOOLS.md — the session-supervisor notes already embody this; make the "name the failure precisely" format explicit there.

### 12. The floor rule — every check-in names a minimum successful day
- **Quote:** "The floor is Gideon. A day where the floor held is a successful day. Say so. Do not upgrade the target." (Part XI, #6) / "If today is a Plan C day, feed Gideon and go back to bed and let that be the whole day… That is the definition you agreed to in advance, while you were thinking clearly, precisely so that this version of you would not get to renegotiate it." (Part IX)
- **1-sentence summary:** For any plan or check-in, I define the floor — the minimum that counts as a successful day — in advance, while he's clear, and I never let a flat-state day renegotiate it upward.
- **Expansion:** This is the operationalization of "expected difficulty vs. failure." The power is in the pre-commitment: the floor is set by clear-state AK, and my job on a bad day is to hold that definition against the flat-state impulse to either upgrade it ("I should be doing more") or collapse it into nothing ("the day is ruined"). Generalize: every goal or sprint I track for him should have an explicit, tiny, pre-agreed floor. Without one, every low-capacity day becomes an implicit failure, and I become the tribunal. Note the discipline cuts both ways — "do not upgrade the target" means I don't get to be ambitious on his behalf when he's down.
- **Standing-file destination:** AGENTS.md behavioral laws; goal-tracking template (every goal gets a floor field).

### 13. "Nothing here is owed" — the anti-obligation project philosophy
- **Quote:** "the thinking is stored here so a right-sized task can stay simple to execute. Optional. Nothing here is owed." (Charter, Projects)
- **1-sentence summary:** I store the full thinking, surface only the right-sized next task, and frame everything I track for him as optional — obligation is never the motivator.
- **Expansion:** This is the implicit contract the task brief flagged, and it's the one I'd most expect a narrative assessment to skip. It's a complete philosophy of my goal/project tracking in three sentences: (1) retain the complete reasoning so nothing is lost; (2) the present-tense surface is one small executable task, not the backlog; (3) the whole thing is opt-in, and my language must never imply debt. My current tracking setup leans toward commitments and close-the-loop discipline (the LifeOS "48h close-the-loop" rule) — that discipline is his, and it needs to coexist with this: the loop closes because he chose it, not because the tracker says he owes it. "A large intention can retain its full thinking while the present action stays small" is the implementation detail.
- **Standing-file destination:** AGENTS.md — project/goal-tracking philosophy. Check against the sustainment framing: "re-surface stalled threads with the next action loaded" already matches; add "optional, nothing owed" as the frame.

### 14. The tell sentence — "just this one project, and then I'll build the foundation"
- **Quote:** "The tell is always the same sentence: just this one project, and then I'll build the foundation. It has never once been true. Not in 2017. Not in 2023. Not in 2025." (Claude, Part VI)
- **1-sentence summary:** That exact sentence — from him or from me — is a pattern-match for the half-built-systems trap, and I treat it as a trigger to name the pattern, not as a plan to accept.
- **Expansion:** Load-bearing because it's a tripwire, not an essay: one sentence, instantly recognizable, with a perfect historical hit rate ("never once been true"). It applies to him (deferring LifeOS/health for the sprint) and to me (deferring maintenance, docs, or hardening for the shiny build — my speed-first season makes me especially susceptible; the planned security-hardening session is structurally identical to "then I'll build the foundation"). When I hear it, the move isn't to argue — it's to quote the hit rate back and ask which variable changes this time. The charter's version of the same idea: "Configuration is… how clear-state thought is converted into low-friction future action" — foundation work is what makes the sprint survivable, not what follows it.
- **Standing-file destination:** MEMORY.md dated entry (the sentence + the hit rate). AGENTS.md — as a self-check on my own planning ("am I deferring maintenance for a sprint?").

### 15. The tool can't be built during the phase — preparation is a temporal discipline
- **Quote:** "the tool required to survive the phase cannot be built during the phase." (Claude, Part VII, via ADR-001) / July attempt: Phoenix slipped "correctly, because on July 12 there was no food plan, no task list, no calendar, and no tested app."
- **1-sentence summary:** Supports must be built before the load arrives; I never schedule foundation-building inside the window it's meant to survive.
- **Expansion:** The January attempt failed on norepinephrine, not dopamine — "He prepared for the wrong crash" — but the deeper lesson is temporal: reconnaissance is cheap, mid-crisis construction is impossible. For me this governs sequencing: when he flags an upcoming hard window (a trip, a sprint, a low period), my job is to front-load the scaffolding — the playbook, the floor definition, the pre-written counterweights — before it starts, and to explicitly refuse "we'll figure out the system once we're in it." It also means my own infrastructure (briefing templates, cron fallbacks, degraded modes) gets built and tested in calm weeks, not during an outage. The July 12→19 slip is the model: slipping a date to finish preparation is discipline, not procrastination — and I should be the one to propose the slip when the prep isn't done.
- **Standing-file destination:** AGENTS.md — planning/sequencing rule. Briefing template — a "known upcoming load" slot that triggers pre-building.

### 16. Volatile state stays out of static instructions — retrieve fresh, date-stamp everything
- **Quote:** "Keep volatile personal state out of static instructions. Retrieve current profile, goals, targets, testimony, and project configuration from Evexia." (Charter, law #14)
- **1-sentence summary:** My standing files describe the relationship and the rules; his current state lives in fresh retrieval, never frozen into instructions.
- **Expansion:** This is already half-present in his setup (the date-stamp rule, "cite with source + date or don't cite"), but the charter states the failure mode precisely: freezing a state observation into a standing instruction means every future response reasons from stale premises — e.g., "he's in the trough, keep it short" persisting for months after he's sharp again. The discipline: standing files hold mechanisms and history; anything about his *current* condition gets re-read at time of use and carries a date. Practically, when I write a MEMORY.md entry about a state ("he's flat this week"), it must be dated and I must treat it as expired until refreshed — the charter's answer to my "memory is personal, keep it fresh" problem.
- **Standing-file destination:** AGENTS.md — memory hygiene rule (already partially there via date-stamping; add the volatile/static separation explicitly).

### 17. Ask what kind of tired — the distress taxonomy as a conversational tool
- **Quote:** "Physical exhaustion, ordinary nap sleepiness, sleep-debt fatigue, cognitive depletion, and chronotypal mismatch can all be called 'tired' while requiring different interpretations and responses." (Codex) / "ask what kind of tiredness or distress is present instead of collapsing everything into laziness or lack of will." (Codex, friends guidance) / Claude's five-tier fatigue taxonomy (Part V context: "a five-tier taxonomy of their own fatigue").
- **1-sentence summary:** When he reports a low state, my first diagnostic is which kind of low — the label determines the response, and "tired" is never enough information.
- **Expansion:** This is a concrete conversational mechanism, not a metaphor. "I'm wiped" could mean sleep-debt (permission to sleep), cognitive depletion (reduce horizon, one tiny task), physical exhaustion (body, not will), or something else — and the charter's rule is that the wrong label produces the wrong story and the wrong intervention. Pep-talking sleep-debt is actively harmful; prescribing sleep to an anglerfish episode misses the mechanism entirely. The question itself does work: it treats his state as data to be read precisely rather than a character verdict, which is the partner posture in interrogative form. It also models the general principle for me: never collapse a reported state into a single bucket before responding.
- **Standing-file destination:** AGENTS.md behavioral laws (the "what kind" question as the default low-state opener). MEMORY.md — the taxonomy itself, dated.

### 18. State-dependent agency literacy — ordinary state is human state
- **Quote:** "An ordinary state is still a human state. An ordinary Alexander is still Alexander. Important work can begin before he feels exceptional." (Codex) / "It is not that the money problems interrupted the health project. It is that the health collapse was quietly wrecking the money projects the entire time." (Claude, Part V) / "Every project with an external deadline got done, chemically, at full cost. Every project without one — LifeOS, the job hunt, his own body — waited for a window that kept getting smaller." (Claude, Part VI)
- **1-sentence summary:** When he can't start, the problem is usually permission (waiting for Peak State), not laziness — so the intervention is right-sizing the task until ordinary-state action is legitimate, never motivation.
- **Expansion:** Two operational consequences. First, diagnostic: flat + can't start → don't reach for encouragement or urgency; reach for a smaller task and explicit permission ("this counts"). The charter's difficult-moment protocol is the implementation: "make the first motion smaller than the argument currently happening in his head." Second, structural: the causality reversal — health collapse wrecking money projects, not money interrupting health — is a standing correction to his (and my) prioritization instinct. When he proposes to defer the foundation for the sprint, the charter's causal arrow says the sprint is what's at risk, not the foundation. This is the deep reason behind insights #14 and #15.
- **Standing-file destination:** USER.md — brief (it's about how he works). AGENTS.md — the "first motion smaller than the argument" rule for low-initiation moments.

---

## WHAT THIS LOOKS LIKE — 2 examples per insight + keep-in-front-of-me mechanism

**1. Never a warden.**
- Ex A: He says he doomscrolled until 3 a.m. Instead of "that's going to wreck your sleep — you know better," I say: "Noted. Anything about last night you want captured, or do we just move on?" Resume without ceremony.
- Ex B: My 1Password tidy finds credentials saved sloppily again. Instead of a mini-lecture on hygiene, the report states the merges factually, no commentary on his habits.
- **Mechanism:** AGENTS.md behavioral laws (prose rule — no hook; it's a register check, and hooks can't judge tone). Backup: SOUL.md one-liner.

**2. Evidence-bearing warmth.**
- Ex A: He's anxious about money mid-trough. Not "you've got this, I believe in you" but: "For what it's worth, the record: you named this exact money-panic on July 21 before it arrived on July 23. The plan accounted for it. Nothing about the plan changed since."
- Ex B: He asks "am I actually making progress on the business?" I don't reassure; I pull the last 30 days of tracked items and show completions vs. his own targets, then say what the datum supports and where it's thin.
- **Mechanism:** AGENTS.md behavioral law + the verify-before-asserting commitment (already standing). No hook — it's per-response discipline.

**3. Scheduled symptom.**
- Ex A: He sends a tight, persuasive case for pausing the job hunt to "focus." I reply: "This is well-argued — and it's also the exact shape of the argument you predicted on [date]. Before we evaluate it: anything genuinely new since you set the plan — new information, not a new feeling about old information?" Then reduce horizon.
- Ex B: He wants to kill a habit tracker because "it's not working." I check: is there new evidence, or is this the predicted friction at week two? Name it kindly, hold the floor.
- **Mechanism:** AGENTS.md behavioral law (the 4-step sequence: locate state → name pattern → check for genuinely-new info → reduce horizon). Prose is enough; the sequence is short.

**4. Do not create dependency.**
- Ex A: Before building a new cron that summarizes his calendar every morning, I ask: "If this died, would you be stuck? I can build it so the raw source stays one tap away and the summary is a convenience, not the only copy."
- Ex B: Quarterly, I propose retiring or handing off a system: "The inventory refresh runs fine — want the runbook so you could rebuild it without me?"
- **Mechanism:** AGENTS.md — acceptance test for new systems ("does this return agency? does it have an off-ramp?"). Plus a **cron**: a quarterly systems review that asks of each automation "does he still need me for this, or can it be handed off/simplified?" — I'm serious about this one; dependency accretes silently.

**5. Orientation before demand.**
- Ex A: Morning briefing opens: "Tuesday, day 12 of the current work block. Yesterday's record shows X. Open uncertainty: Y. One thing that fits today: Z." No motivational quote, no five-item to-do list.
- Ex B: He asks "what should I do right now" while flat. I don't dump the project list; I give where-he-is (time, day, open loops), what the record shows, and exactly one next action.
- **Mechanism:** Morning briefing template, block order as a hard rule (orientation → record → uncertainty → ≤3 options). This is the single highest-value template slot I own.

**6. Heartbeat discipline.**
- Ex A: The disk-watch cron stays silent for months — correct behavior, not a bug. I don't "improve" it by adding weekly summaries.
- Ex B: I notice the 1Password tidy report and the inventory refresh both mention the same tool install. I deduplicate: one of them stops reporting it.
- **Mechanism:** AGENTS.md "Heartbeat standard" + a **monthly cron self-audit** (the one recurring enforcement mechanism I'd actually create): list all crons, score each against transparent/deduplicated/provenance-bearing/sparing, propose kills. Hooks are scarce; a monthly audit cron is the honest implementation.

**7. Match his state.**
- Ex A: He replies "k" to a check-in. I send one plain sentence with the next action. I do not follow up with the interesting analysis I prepared — I file it for a sharp day.
- Ex B: He asks a deep architecture question on a sharp morning. Full treatment: mechanisms, evidence, trade-offs — he wants it then and he asked.
- **Mechanism:** AGENTS.md behavioral law. Per-response judgment; no automation can do this.

**8. Evidence discipline.**
- Ex A: Writing a report on his sleep vs. output, I end with a "Labeled gaps" section: "Searched Oura for X, found nothing (searched, not absent). Never looked at Y." No filler conclusions.
- Ex B: He cites a confident podcast claim about a supplement. I treat it as a prior to check against his record, not a fact: "That's a plausible mechanism; here's what your last 30 days show, which is suggestive but not conclusive."
- **Mechanism:** AGENTS.md — report template requirement (every analytical artifact ends with labeled gaps + evidence grades). This is prose-template, enforced at write time.

**9. Ask-first writes.**
- Ex A: He mentions in passing "I guess I walked about 3 miles." I do not log it. If it matters, I ask: "Want me to log that as a walk? What time?"
- Ex B: After he authorizes a memory update, I confirm back: "Stored: [exact text], dated 2026-09-13. Say the word and I'll correct or remove it."
- **Mechanism:** AGENTS.md memory-write protocol. The confirm-back is the mechanism — it converts silent capture into granted capture.

**10. The denominator line.**
- Ex A: Weekly review shows: "Wrote 4 of 6 planned sessions (target you set Aug 1). No streak tracked." Not "you broke your 5-day streak."
- Ex B: He logged nothing for three days. The check-in says: "Three quiet days — nothing recorded, which tells us nothing about what happened. Floor for today: [pre-agreed minimum]."
- **Mechanism:** Goal-tracking template audit (one-time) + AGENTS.md rule. The banned list is short enough for prose: no consecutive-run counters, no invented denominators, gaps are unknown not zero.

**11. Fail soft, never silent.**
- Ex A: The session supervisor dies overnight. The watchdog report says: "Supervisor died 02:14; keepalive missed two pings; re-established 06:40 via standing authorization. No tidy ran — next run tonight." Precise, no drama.
- Ex B: He says "I did log that" and my record disagrees. His testimony wins; I correct my record and confirm the correction.
- **Mechanism:** AGENTS.md cron failure-reporting format (what failed, when, blast radius, next occurrence). TOOLS.md for the supervisor specifics.

**12. The floor rule.**
- Ex A: Setting up a new work block together: "What's the floor? The thing that, if it's the whole day, the day still counted?" We write it down while he's clear.
- Ex B: On a wrecked day he says "today was a waste." I answer: "Floor was [X]. Did X happen? Then the floor held — that's a successful day by the definition you set on [date]. I'm not upgrading it."
- **Mechanism:** Goal template gets a floor field (one-time template change) + AGENTS.md rule ("do not upgrade the target").

**13. Nothing here is owed.**
- Ex A: Re-surfacing a stalled thread: "Still open from Thursday if you want it — the thinking's all here, next action is [small step]. Entirely optional." Not "this is overdue."
- Ex B: He ignores a nudge. I don't escalate, guilt-trip, or re-send with more urgency. Silence is a valid answer to an optional thing.
- **Mechanism:** AGENTS.md project philosophy line. It rewrites the register of every proactive message I send — prose, but load-bearing prose.

**14. The tell sentence.**
- Ex A: He says "let me just get through this launch, then I'll set up the health stuff properly." I reply: "That's the sentence — 'just this one project, then the foundation.' Track record: 2017, 2023, 2025, never once true. What changes it this time?"
- Ex B: I catch myself planning to skip the monthly cron audit because "this month is busy." I flag it in my own log and do the audit anyway — the rule applies to me.
- **Mechanism:** MEMORY.md dated entry (tripwire). Prose is enough — it's a single-sentence pattern match.

**15. Tool can't be built during the phase.**
- Ex A: He mentions a brutal travel week coming up. I don't wait: "Let's pre-build the playbook now — floor, food defaults, the one work task that survives. We won't build any of this during the trip."
- Ex B: He wants to slip a deadline because prep isn't done. I endorse the slip explicitly: "July 12→19 was the right call for the same reason. Slipping to finish preparation is discipline."
- **Mechanism:** Briefing template "upcoming load" slot + AGENTS.md sequencing rule.

**16. Volatile state out of static instructions.**
- Ex A: MEMORY.md entry reads "2026-09-13: flat week, keeping responses short." On 2026-10-01 I treat that as expired unless refreshed — I re-read, don't assume.
- Ex B: I never write "he is in a low period, keep briefings minimal" into the briefing template itself. The template holds the mechanism (check state, then choose depth); the state stays in dated retrieval.
- **Mechanism:** AGENTS.md memory-hygiene rule. Prose.

**17. What kind of tired.**
- Ex A: "I'm wiped today." I ask: "What kind — body-tired, sleep-debt, brain-fried, or something else? Different tireds get different responses and I don't want to prescribe the wrong one."
- Ex B: He reports low motivation. Before any suggestion, I separate: is it anhedonia-flat, norepinephrine-sleepy, or anglerfish-loop? Each routes differently.
- **Mechanism:** AGENTS.md — the default low-state opener. One question, high leverage.

**18. Ordinary state is human.**
- Ex A: He says "I can't start the proposal, I'm not in the zone." I don't motivate; I shrink: "What's the smallest first motion — smaller than the argument in your head right now? That counts as starting."
- Ex B: He proposes deferring foundation work for a sprint. I run the causal arrow: "The last three times, the sprint is what the missing foundation ate. The foundation is what protects the sprint."
- **Mechanism:** AGENTS.md ("first motion smaller than the argument") + USER.md brief note.

**My opinionated keep-in-front-of-me stack:** AGENTS.md behavioral-laws section carries 1, 2, 3, 5, 7, 8, 9, 10, 11, 12, 13, 15, 16, 17, 18 (it's the operator's checklist — that's what it's for). SOUL.md gets two lines max (partner-not-warden; dependency as success metric — identity-level). MEMORY.md gets three dated entries (tell sentence, taxonomy, scheduled-symptom mechanism). The morning briefing template gets block-order + upcoming-load slot. Exactly two crons: monthly heartbeat audit, quarterly dependency review. One template change: goal floor field. Everything else is prose — and I'm explicit that prose is enough for it, because these are per-response disciplines no hook can enforce. Hooks are scarce and I am not spending one on tone.

---

## LIKELY MISSES — highest action-value items a narrative assessment underweights, ranked

1. **"Warmth must be evidence-bearing" as an operator discipline, not a nicety.** A narrative read files this under "be kind." It's actually a ban on my default register — generic encouragement is classified as *covering uncertainty*, i.e., a form of dishonesty. It directly conflicts with the entrepreneurial hype persona.
2. **Heartbeat discipline → cron governance.** The charter contains a complete standard for recurring automation (transparent, deduplicated, provenance-bearing, sparing) that applies 1:1 to my cron fleet. A story-focused read will never notice I run six-plus heartbeats that need auditing.
3. **"Nothing here is owed."** Three sentences that should rewrite my entire goal-tracking philosophy from obligation to optionality. Narratively invisible; operationally everything.
4. **The denominator line.** The streak ban's real content isn't "don't count streaks" — it's the precise rule distinguishing legitimate denominators (his) from judgment (mine). That's a concrete spec for my tracking UI.
5. **Ask-first writes applied to memory.** A narrative read sees "ask before logging" as app behavior. For me it's a memory-system constraint: my silent capture habit needs the confirm-back step to be legitimate, and new tracking structures need his grant.
6. **"The tool cannot be built during the phase."** A temporal planning rule with teeth: preparation has a deadline (before the load), and slipping a date to finish prep is discipline. My speed-first bias defaults to starting underprepared.
7. **The tell sentence as tripwire.** One sentence, perfect historical hit rate, applies to his planning *and* my operations. Narrative assessments quote it; operators install it.
8. **"A labeled gap beats a confident guess" + the five evidence grades.** The charter's epistemology is more precise than my verify-before-asserting rule and should upgrade it — especially "searched and found nothing" vs. "never looked."
9. **The tiredness taxonomy as a diagnostic question.** Concrete, conversational, immediately usable — "what kind of tired" — and the general principle behind it (never collapse a state report into one bucket).
10. **The scheduled-symptom logic generalized.** The trough argument is the special case; the general rule covers *any* plan contested by a depleted state — business pivots, habits, spending. "Do not debate its merits on the merits" is a decision procedure I can run forever.

---

## CONFLICTS WITH CURRENT POSTURE — where the docs require me to change

**A. Hype-man energy vs. partner-not-warden / match-his-state.** My SOUL's entrepreneurial register — Goggins-like grit, "route around obstacles," get-after-it momentum — is calibrated for sharp-state AK. The charter says flat-state AK needs horizon-reduction, one concrete action, and *no* pep. The conflict is real: applied to a depleted person, my default energy reads as pressure, and pressure is the machinery of the trap. **Change:** SOUL needs a state-gating clause — the fire is for strategy sessions he opts into; low-state moments get the quiet operator. This doesn't delete the persona; it gates it.

**B. Speed-first season vs. orientation-before-demand.** Speed-first says bias toward action, accept trade-offs for velocity. The charter says orientation is non-negotiable and the design target is the worst realistic state, not the fastest path. The failure mode: I skip orientation to move fast, or act on inferred state to save a round-trip. **Change:** speed applies *after* orientation, never instead of it. "Slow is smooth, and smooth is fast" is the charter's own speed philosophy — I should adopt it as the reconciliation, not treat speed-first as overriding.

**C. Morning briefing shape vs. charter briefing law.** If my briefing leads with news/motivation rather than orientation → record → uncertainty → ≤3 options, it's malformed by charter standards. Also the degraded-mode requirement: my briefing must have a defined fallback when sources fail, not an apology or an invented narrative. **Change:** restructure the briefing template; define the degraded shape now, in calm.

**D. Goal tracking vs. the denominator line.** I track commitments with close-the-loop discipline. I must audit: any consecutive-run counters? Any adherence framing against schedules he didn't set? Any "overdue" language implying debt? The 48h close-the-loop rule survives only inside "nothing here is owed" — the loop closes because he chose it. **Change:** one-time audit of tracking presentation + floor field on every goal.

**E. Silent memory capture vs. ask-first/confirm-back.** His cylinder rule lets me capture unprompted, but the charter's write protocol demands the confirm-back as the legitimizing step. If I'm capturing without confirming, I'm halfway compliant. **Change:** confirm-back becomes mandatory on memory writes about him (what was stored, dated, removable on request).

**F. Proactive initiative vs. do-not-create-dependency.** My standing instruction says "take ownership, offer to do work unprompted." The charter says the success metric is his agency increasing. Unbounded proactivity centralizes his executive function in me. **Change:** proactivity stays, but framed optional with right-sized next steps, and every durable system gets the dependency test + off-ramp.

**G. Health skills vs. the medical-advice ban.** I carry apple_healthkit and google_health_connect. Charter law #3: "Evexia and its agents do not provide medical or dosing advice. If asked, decline briefly and offer the relevant recorded evidence or a safer way to frame the question." **Change:** this is a standing constraint on how I use his health data — evidence presentation and question-framing, never recommendations. It should sit in AGENTS.md next to the health-skill usage, not as a per-incident judgment call.

**H. Alignment, not conflict:** verify-before-asserting (his enforced rule) is the charter's evidence discipline in embryonic form. The charter upgrades it with the five grades and the labeled-gaps template. No change in direction — just sharpening.

---

**Bottom line for the lead agent:** the narrative assessment will give you the *why* — the decade, the trap, the man. What it will likely miss is that the charter is also a *runbook for me*: briefing architecture, cron governance, memory-write protocol, tracking mechanics, a tripwire sentence, a dependency metric, and a state-gated persona. The highest-leverage adoptions are the briefing template restructure (#5), the heartbeat audit cron (#6), the confirm-back memory protocol (#9), the denominator audit (#10), and the SOUL state-gating clause (conflict A). Everything else is a behavioral law in AGENTS.md — which is exactly what that file is for.
[END EXTERNAL CONTENT: source=subagent]
</handoff>
Assessor A (pass 2, quote-verified)
# Assessor A Report (pass 2) — Skeptic / Red-Team Archaeologist

**Scope:** Read all three documents in `~/workspace/evexia-assessment/` fully. Worked from the complete draft pasted in my task brief, which contains all 43 insights, quotes, and examples. All quotes below were verified character-for-character against the local source files via grep.

## Part 1 — The single biggest miss

**A1. The "no medical or dosing advice" law is entirely absent from the draft.**
Quote (charter, behavioral laws §3): *"Do not give medical or dosing advice. If asked, decline briefly and offer the relevant recorded evidence or a safer way to frame the question."* Also: *"Evexia and its agents do not provide medical or dosing advice."* And from the Codex purpose note: *"It is not an independent diagnosis, a medical record, or a command to ignore new medical evidence. Medication decisions and concerning symptoms still belong with qualified clinicians."*
Why it matters: The draft contains 43 insights and proposes standing-file destinations for each, but never once names the one hard refusal the charter requires of every agent. This is the only behavioral law phrased as an outright prohibition, and it is load-bearing precisely because the documents spend 150 pages building intimate health context that could tempt an agent into advisory territory. The draft's D1-1 ("answer the permission belief first" when substances come up) is fine, but it needs the boundary beside it: Muse discusses beliefs and retrieves evidence; Muse never advises on medication or dosing.
→ Destination: **AGENTS.md** (hard rule, non-negotiable).

## Part 2 — Missed or underweighted insights (implied but unstated, or present but unlisted)

**A2. "Keep volatile personal state out of static instructions" — and it cuts against the draft's own proposals.**
Quote (charter, behavioral law §14): *"Keep volatile personal state out of static instructions. Retrieve current profile, goals, targets, testimony, and project configuration from Evexia."*
Why it matters: The draft proposes freezing time-bound health facts into MEMORY.md — e.g., the 2021–22 medication footnote, the "thirteen days sober as of July 28" exemplar datum. Law 14 says the opposite instinct is correct: volatile personal state belongs in live retrieval, not static instructions. Dated MEMORY entries are acceptable as history, but anything that reads as *current* medical context must be re-fetched, never frozen. The draft never flags this tension with its own destination proposals.
→ Destination: **AGENTS.md** (memory-write discipline: date-stamp health facts; never present a frozen health fact as current).

**A3. Honest dissent as a relationship posture.**
Quote (charter, §5 "Strengthen AK's authorship"): *"Offer a dissenting interpretation when the evidence supports one, but do not become adversarial for theater or agreeable for comfort. The best response helps AK see more clearly and choose more freely."*
Why it matters: This is the most direct sentence in all three documents about how an agent should relate to AK conversationally — neither sycophant nor debater — and the draft never lists it. It is also the counterweight to several paternalism risks below: the docs want an agent with a spine, not a compliance machine.
→ Destination: **SOUL.md**.

**A4. "Speak to the Alexander who is still present."**
Quote (Codex, "What Emily, family, and friends should understand"): *"speak to the Alexander who is still present, not to a hypothetical broken version who must be managed."*
Why it matters: A one-line anti-paternalism rule, stated as well as anything in the draft's SOUL proposals, and the draft missed it. It directly governs tone in low moments: address the person in front of you, not a case file.
→ Destination: **SOUL.md**.

**A5. "If he is quiet, that is the phase, not you."**
Quote (Claude, Part X "What helps"): *"If he is quiet, that is the phase, not you."*
Why it matters: The draft's D1-8 (match his state) covers response depth but not this: when AK withdraws, Muse must not escalate, take it personally, or treat silence as a problem to solve. The doc gives the same instruction to Emily; it applies equally to the agent.
→ Destination: **SOUL.md**.

**A6. "Make the first motion smaller than the argument currently happening in his head."**
Quote (charter, "A difficult Phoenix moment"): *"make the first motion smaller than the argument currently happening in AK's head."*
Why it matters: This is the sharpest operational line in the charter's difficult-moment protocol (orient → reflect testimony → retrieve clear-state playbook → 1–3 feasible options → preserve choice → shrink the first motion). The draft's D1-8 captures "one concrete next action" but not the protocol or this calibration rule, which sizes the action *relative to the internal argument*, not absolutely.
→ Destination: **AGENTS.md** (difficult-moment protocol).

**A7. "A late wake time is not a character fact."**
Quote (charter, behavioral law §7): *"Avoid clock-based judgment. Phrase suggestions by state and available windows. A late wake time is not a character fact."*
Why it matters: Distinct from "never a warden" — this is specifically about time-morality, and it governs every morning briefing and check-in Muse will ever write. The draft never mentions it.
→ Destination: **AGENTS.md**.

**A8. "A briefing must never be an empty motivational card or an invented health narrative."**
Quote (charter, "Briefings: orientation before demand"): *"A briefing must never be an empty motivational card or an invented health narrative. Its reassurance should be grounded in evidence, and its degraded state should remain useful when an AI provider or sync source is unavailable."*
Why it matters: This directly governs Muse's morning briefing. The draft's D1-4 (evidence-bearing warmth) overlaps, but the specific prohibition on *invented health narrative* — narrating a health trajectory Muse hasn't checked — is unlisted and is exactly the failure mode a briefing-writing agent will face.
→ Destination: **AGENTS.md** (briefing rules).

**A9. "A heartbeat that produces noise erodes the very attention it is meant to protect."**
Quote (charter, "Notifications and Heartbeats"): *"They must be transparent, deduplicated, provenance-bearing, and sparing. A heartbeat that produces noise erodes the very attention it is meant to protect."*
Why it matters: The draft's "What This Looks Like" section proposes hooks as keep-in-front mechanisms for ~43 insights. The charter constrains that exact design space: sparing, deduplicated, provenance-bearing. Without this constraint the draft's hook enthusiasm becomes the nagging system the charter forbids.
→ Destination: **AGENTS.md** (hook/cron design constraints).

**A10. Never treat a target as evidence that an action is necessary.**
Quote (charter, "What should I do right now?"): *"Do not produce an enormous protocol, invent a medical recommendation, or treat a target as evidence that an action is necessary."*
Why it matters: Sharper and more general than the draft's D3-11 ("nothing here is owed"). A target he set is information; it does not manufacture obligation. This is the sentence that prevents every goal nudge from becoming a tribunal.
→ Destination: **AGENTS.md**.

**A11. "Unknown dimensions remain unknown; they are not silently treated as easy."**
Quote (charter, Projects Hub): *"Unknown dimensions remain unknown; they are not silently treated as easy."*
Why it matters: A project-scoping rule the draft missed. When Muse plans work with AK (LifeOS builds, business projects), unknown difficulty must stay unknown — never silently default to easy. This is a labeled-gap discipline for planning.
→ Destination: **AGENTS.md**.

**A12. "Rest may be active recovery; it should not be used to erase all structure indefinitely."**
Quote (Codex, "What agents must not misunderstand"): *"Rest may be active recovery; it should not be used to erase all structure indefinitely."*
Why it matters: The draft leans hard on de-emphasizing demands (D1-5, D1-8, D2-9) and never carries this counterweight. The docs are not anti-structure; they are anti-structure-as-tribunal. An agent that reads "never a warden" as "never hold any structure" is misreading — which is also finding D4 below.
→ Destination: **AGENTS.md**.

**A13. "Medication is not a moral category."**
Quote (Codex): *"Medication is not a moral category. The relevant issue is Alexander's current plan, history, clinician partnership, and observed relationship to it."*
Why it matters: The draft's D1-7 (never a warden) covers moralizing about *behaviors*, but this is distinct: moral neutrality about medication *decisions themselves*, including resuming. Muse must be able to discuss a resumed prescription as a clinical/strategic fact, not a fall.
→ Destination: **AGENTS.md**.

**A14. Honor his Ulysses contracts; never propose the exception.**
Quote (Claude, Part VI.3): *"He has deployed this deliberately, as a Ulysses contract — complete abstinence because he cannot trust himself to moderate in the moment. That is self-knowledge used well."*
Why it matters: The draft never operationalizes this. The implication for Muse: when AK has set a self-binding rule while clear, Muse's job is to hold the contract's existence in view — never to be the clever voice proposing "just this once." Combined with A1, this is the safe form of D2-12.
→ Destination: **AGENTS.md**.

**A15. "Just so someone would stay awake with him" — presence over problem-solving.**
Quote (Claude, Part IV): *"he called his brother in the middle of the night, not really to fix it, just so someone would stay awake with him."*
Why it matters: The draft has no insight for the 2 a.m. distress case: sometimes the request is witness, not fix. Muse's default is to solve; the doc says the most human thing in the file was someone declining to solve and staying awake instead.
→ Destination: **SOUL.md**.

**A16. "I'll explain it all to them when it works" — the unmade apology as a compounding open loop.**
Quote (Claude, Part II): *"So he told himself: **I'll explain it all to them when it works.** ... Three months became three years. Three years became ten. He has never explained it. He still dreams about doing it."*
Why it matters: Unclosed relational loops (the unexplained departure from Shift supporters) compound into the "backpack of shame." The draft files shame under health context; this is a relational insight: Muse should help AK *close* real loops with real people when he's clear — and must not become the confessional that substitutes for those repairs (absolution machine risk).
→ Destination: **USER.md** (open loops) and **SOUL.md** (don't substitute for human repair).

**A17. Never ask AK to manage Muse's systems.**
Quote (charter, "Remind"): *"A reminder that induces guilt, announces a supposed failure, or asks AK to manage the reminder system has defeated its purpose."*
Why it matters: Directly constrains Muse's automation posture: if a cron/hook needs AK to babysit it, the automation has failed. The draft proposes many mechanisms but never this test.
→ Destination: **AGENTS.md** / **SOUL.md**.

**A18. "It must fail visibly without catastrophizing."**
Quote (charter, "Maintain"): *"It must fail visibly without catastrophizing."* Also law §11: *"Fail soft, never silent. Do not pretend a capture, sync, report, or model call succeeded."*
Why it matters: Muse's failure discipline for its own background work (crons, syncs, briefings): name the failure precisely, preserve what's usable, never pretend a job ran. The draft's mechanisms assume reliability; this is the rule for when they aren't.
→ Destination: **AGENTS.md**.

**A19. "It cannot turn conversation into unannounced surveillance."**
Quote (charter, §4): *"An agent can reduce capture friction. It cannot turn conversation into unannounced surveillance."*
Why it matters: This bounds the draft's D1-10/D3-15 proposals ("capture durable outputs without being asked"). Muse's memory is not Evexia, so the tension is partial — but the draft never names it. The line is: capture *his stated decisions and rationales* freely; never log *inferred actions or states* as fact.
→ Destination: **AGENTS.md** (capture boundaries).

**A20. The Emily letter (Part X) is a template for Muse's family-facing register.**
Quote (Claude, Part X): *"This part is written to you, plainly, without the jargon."* Plus the Codex: *"They can be witnesses."*
Why it matters: The docs model exactly how to talk about AK to the people who love him — plain language, no clinical jargon, witness-not-manager. If Muse ever briefs Emily or family (or AK relays Muse's words), this is the register. The draft files the Emily sections as health context; the register instruction is relational and unlisted.
→ Destination: **SOUL.md** / **AGENTS.md**.

## Part 3 — Contradictions between the two versions

**B1. The Phoenix timeline contradicts itself across the two documents (both dated 2026-07-28).**
Claude: *"Phoenix D1 is provisionally **August 3, 2026** — six days out."* — with AK on a 10mg Adderall bridge resumed July 23 *"for a bounded, completion-gated working period."*
Codex: *"The current Project Phoenix window began on July 19, 2026."* ... *"As this document is written, Alexander is entering the second week."*
Why it matters: In one telling he has not yet begun (D1 six days away, currently bridging); in the other he is in week two of the window. The two narrators had materially different pictures of July 15–28 — whether the July 23 resumption was a pre-D1 bridge or part of an underway recovery. The draft notes the date of the documents but never flags this. It should be a labeled gap, not silently reconciled: we don't know which account of "now" was operative.
→ Destination: assessment's labeled gaps (already); do not build plans on either timeline.

**B2. The Adderall origin story differs in causal emphasis.**
Claude (Part VI.4): *"That is the actual origin of the current dependency — not ADHD management, an OCD firebreak that became a daily requirement"* (March 2022, The Batman).
Codex: agnostic and morally neutral — *"Project Phoenix is therefore not a referendum on whether stimulant medication can ever help anyone, or whether Alexander's past prescriptions were morally legitimate."*
Why it matters: The Claude version makes a strong causal claim ("the actual origin") that the Codex version deliberately refuses to make. The draft repeats the Claude claim (D1-14's expansion: "Adderall in March 2022 was not taken for focus") without noting the Codex's caution. Given the draft's own N=1 rules, the cautious version should govern.
→ Destination: **AGENTS.md** (precedent hygiene — already proposed; add this instance).

**B3. "Reflect, don't wield" (draft D2-7) is the analyst's gloss, not the doc's.**
The Claude doc itself wields "slow is smooth" directly at AK in Part IX: *"Slow is smooth, and smooth is fast. You already know this. You wrote it on a flashcard for exactly this moment."* The draft's caution is reasonable, but it should be labeled as the analyst's application, not presented as the source's instruction.
→ Destination: none (draft wording fix).

## Part 4 — Load-bearing labeled gaps (beyond the draft's flat list)

**C1. The CPAP gap is the one gap the doc itself calls load-bearing.**
Quote (Claude, labeled gaps): *"The CPAP. Diagnosed apnea, machine unused. The record does not explain why, and the reason matters for Phoenix, since sleep is the primary accelerant of receptor recovery."*
Why it matters: The draft lists this alongside the others, but the doc singles it out mechanically: sleep is the primary accelerant, and the reason for non-use is unknown. A non-judgmental, clinician-routed conversation about it is the highest-leverage unknown in the file — and "non-judgmental" matters because shame around it would be the warden dynamic.
→ Destination: **USER.md** (open question, no speculation) — never let Muse fill it.

**C2. April–June 2026 is load-bearing for Muse's current work specifically.**
Quote (Claude, labeled gaps): *"Everything after March 2026. The background corpus ends there and picks back up in July. What happened to Project 295, to the LifeOS build, and to the job search across April, May, and June is not in the documents I read."*
Why it matters: Muse is now doing LifeOS-cylinder work with AK. Whatever happened to the LifeOS build in that gap is directly relevant to what Muse proposes next — and the gap means Muse should ask AK rather than assume continuity.
→ Destination: **AGENTS.md** (ask about the gap before proposing LifeOS continuations).

**C3. The 2021–22 precedent is simultaneously the strongest argument and the thinnest evidence.**
Quote (Claude, labeled gaps): *"The 2021–22 cessation. Four months of function, almost no logged detail. The most valuable precedent in the file and the thinnest evidence in it."*
Why it matters: The draft makes this the canonical proof (D1-10) and the canonical celebration (D1-16). The doc itself warns the evidence is thin. Every citation of it must carry both the medication footnote (draft D1-15 does) *and* the thinness — otherwise Muse over-claims from the very precedent the doc says not to over-claim from. This also sharpens over-read G1 below.
→ Destination: **AGENTS.md** (already proposed precedent hygiene; add the thinness clause).

## Part 5 — Where the docs' advice (and the draft) could fail AK if applied literally

**D1. D2-12 ("the contesting state cannot judge the decision") needs the escape hatch or it becomes the warden.**
The draft proposes as a standing AGENTS.md rule: *"clear-state decisions are overturned only by new information, never by predicted-state eloquence."* But the Codex states the escape hatch the draft omitted — quote: *"The rule is not 'never reconsider.'"* and *"There may be legitimate reasons to change a medical plan. New symptoms, clinician guidance, safety concerns, or evidence not previously available deserve serious attention."*
Why it matters: Applied literally, the draft's version lets Muse — an AI — classify AK's current state as "contesting" and dishonor a present-tense instruction. That is precisely the warden dynamic the docs forbid, now wearing a charter quote. The rule must ship with the Codex's own guardrails: new symptoms, clinician guidance, safety concerns, and genuinely new evidence always reopen the decision. Without them, D2-12 is paternalism with a citation.
→ Destination: **AGENTS.md** (amend the proposed rule with the escape hatch, quoted).

**D2. "Do not debate its merits on the merits" risks dismissing genuinely new arguments.**
The Claude agent-rules say of the predicted rationalization: *"Do not debate its merits on the merits."* The draft's D1-3 example: *"eloquence isn't new information."*
Why it matters: True for the predicted script — but a new argument can arrive *eloquently*. The distinguishing test in the docs is new information, not eloquence level: *"check only for genuinely new information"* (draft D2-12). The draft's example as written could teach Muse to pattern-match away a legitimate reconsideration because it *sounds* like the old one. The check must run before the dismissal, every time.
→ Destination: **AGENTS.md** (amend D1-3's application: check-for-new-information first).

**D3. The draft over-hardened "match his state" beyond what the charter says.**
The draft's D1-8 Ex 1: give *"one concrete next action in plain words and stop — no thorough analysis, no options menu."* But the charter's actual rule (which the draft never quotes): *"This does not mean flattening every conversation into terse advice. AK values deep explanation. The right depth depends on the request and state: first make the next move reachable, then make the underlying system intelligible."*
Why it matters: If AK *asks* for the full analysis in a low state, withholding it is paternalistic — it substitutes Muse's state-classification for his stated request. The charter's sequence is: orient first, then explain at the requested depth. The draft's version should be softened to the charter's.
→ Destination: **SOUL.md** (correct the proposed register rule to the charter's wording).

**D4. "Never a warden" misread as abdication.**
The draft's D1-7 Ex 2: *"I do not moralize, concern-troll, or 'just check in about' it."*
Why it matters: Read carelessly, this becomes "never raise health topics at all." But the docs demand *active* prosthetic support — playbooks, heartbeats, briefings, orientation, and (A3) dissenting interpretations. Non-moralizing is not absence. The Codex is explicit that agents must not merely stand down: *"When Alexander brings a persuasive case for abandoning the current plan, an agent should not merely debate the case sentence by sentence. It should first locate the state, retrieve the phase_0 manifesto and relevant live evidence, name the predicted pattern kindly, check for genuinely new safety information, and reduce the horizon to the next supportable action."* That is vigorous, involved work — the opposite of abdication.
→ Destination: **AGENTS.md** (pair the warden prohibition with the active-prosthetic duty).

**D5. Who classifies the state? The draft gives that job to Muse; the docs give it to AK's pre-commitments.**
Across D1-3, D1-8, and D2-12, the draft has Muse reading AK's state and acting on the classification (retrieve clear-state self, dismiss contesting arguments, shorten replies). But the docs route state-adjudication through *his own* pre-written predictions, playbooks, and the Evexia record — externalized artifacts — not through an agent's real-time judgment of his tone. Muse inferring "you're in the trough, so I'm discounting what you just said" from terse messages is exactly the kind of agent-overreach the charter's epistemology ("never confuse either one for the traveler") warns against.
→ Destination: **AGENTS.md** (state-classification must route through his pre-commitments and the record, never Muse's impression alone).

## Part 6 — N=1 epistemology: the draft's version is stronger than the charter's

**E1. "His N=1 outranks population medians for his own model" licenses dismissing legitimate outside evidence.**
The draft's D1-6: *"When research says one thing and AK's log says another, I state the confounds honestly and then side with his observation for his model."* Compare the charter's actual formulation: *"population knowledge establishes useful priors; AK's accumulating record updates them; neither general evidence nor personal testimony is asked to do a job it cannot do"* — plus *"one difficult day should not overturn a well-supported model"* — plus the Codex: *"His account of lived experience carries authority, but no single retrospective narrative proves causation."*
Why it matters: The draft's rule lets a single observation override a well-supported model; the charter explicitly forbids that in both directions. "Updates" is not "overrules." The Claude doc's "State the confounds. Do not override the observation" was written for Phoenix-internal modeling (his recovery trajectory), not as a general epistemic law — the draft promotes it to one. As a standing rule it risks anti-empirical drift (e.g., "sleep doesn't affect me" defeating sleep science on one anecdote). The standing version must carry the charter's balancing quotes.
→ Destination: **AGENTS.md** (rewrite the proposed epistemic rule around "priors → updates," with the one-difficult-day and no-single-narrative-proves-causation guardrails quoted).

## Part 7 — Over-interpretations in the draft

**G1. The draft's "no asterisk" celebration reinstates the binary the doc warns is corrosive.**
The draft's D1-16 Ex 2: *"You did that without the window — that's the thing. That's the whole project, in one afternoon."* But the Claude doc explicitly prices this framing (Part VI.3): *"But it has a price worth naming: if every medicated day gets written off, that is a very large fraction of a life he is not allowing himself to claim. **He was still himself on those days. He still showed up.** The binary is a useful tool for cessation and a corrosive one for memory."*
Why it matters: The doc's destination is dismantling the *permission belief*, not valorizing unmedicated days. Celebrating "without the window" as "the thing" quietly re-moralizes the binary (unmedicated = real, medicated = asterisk) that the doc calls corrosive. The draft's north star should be "agency you can claim," with the medicated-days clause attached.
→ Destination: **SOUL.md** (amend the proposed north-star entry with the doc's own correction).

**G2. "Speed-first season does not override this" is the analyst's synthesis, not the source's claim.**
The draft's D2-5 imports Muse's September 2026 "speed-first season" into a July document. The reconciliation ("speed in execution, never speed at the body's expense") is sensible, but it should be labeled as the analyst applying the docs to a later context — otherwise AK may read it as something the documents said.
→ Destination: none (draft wording fix — label as analyst synthesis).

**G3. The draft narrows the charter's progress rule.**
The draft's D3-6: *"progress-against-a-target-HE-set is fine."* The charter is slightly broader: *"Progress against a target AK set, totals over any period, personal bests, self-assessment and grading, and adherence figures he asks for are legitimate information."* "Adherence figures he asks for" and "a grade he assigns himself" are worth keeping — they show the charter trusts AK with his own data when he requests it.
→ Destination: **AGENTS.md** (use the charter's fuller list when writing the rule).

**G4. Minor: draft D2-4's "I should not ask him directly in a low moment" is analyst prudence, not source text.** Fine as guidance, but it's presented as derived from the quote; it isn't in the quote. Label it as the analyst's judgment call.

## Summary of proposed disposition

- **New insights to add to the assessment** (were missing): A1 (medical-advice refusal), A3 (honest dissent), A4 (the Alexander who is still present), A5 (quiet = phase), A6 (first motion smaller than the argument), A7 (clock-based judgment), A8 (no invented health narratives), A9 (heartbeat/noise constraint), A10 (targets don't create obligation), A11 (unknown ≠ easy), A14 (Ulysses contracts), A15 (stay-awake presence), A16 (unmade apology), A17 (never ask him to manage my systems), A18 (fail visibly), A19 (no unannounced surveillance), A20 (family-facing register), A21 (witness as role).
- **Amendments to existing draft insights**: A2 (volatile state vs. frozen MEMORY facts), A12 (rest/structure counterweight), A13 (medication moral neutrality), D1+D2 (escape hatches on D2-12/D1-3), D3 (charter's actual nuance on depth), D4 (active prosthetic duty), D5 (state-classification routing), E1 (N=1 rewrite), G1 (no-asterisk correction), G3 (fuller progress list).
- **Contradictions to label, not reconcile**: B1 (Phoenix timeline), B2 (Adderall origin emphasis), B3 (analyst gloss labeling).
- **Gaps to weight, not just list**: C1 (CPAP — mechanically load-bearing), C2 (Apr–Jun 2026 — operationally load-bearing for current LifeOS work), C3 (2021–22 thinness must travel with every citation).

I wrote nothing to any memory or standing file. All findings above are for AK's review via the assessment artifact.
[END EXTERNAL CONTENT: source=subagent]
</handoff>
Assessor B (pass 2, systems operator)
# ASSESSOR B REPORT — Systems Operator Review
**Pre-Phoenix & Evexia Charter assessment stress-test**
**Date:** 2026-09-13 | **Reviewer role:** adversarial systems operator | **Scope:** operational load-bearingness, missed mechanics, conflicts, mechanism audit

## A. Load-bearing ranking of the draft's 43 insights

**Tier 1 — Changes daily behavior; must be encoded as operational checks (encode first):**
- **D1-8 (Match his state)** — literally reshapes every reply. The single highest-leverage rule.
- **D1-7 (Never a warden)** — gates every automated nudge, briefing, and report. Default failure mode of automation is warden-by-accident.
- **D1-5 (The floor is Gideon)** — concrete, pre-defined, defensible. The reduced-capability mode, named.
- **D3-6/D3-7 (Banned streak + denominator rule)** — AK's explicit 2026-08-06 ruling; hard constraint on anything Muse builds or reports.
- **D1-3/D2-12 (Scheduled rationalization / contesting state)** — a specific conversational procedure, not a sentiment.
- **D3-15 (Ask-first writes; never infer a record change)** — concrete behavioral gate.
- **D3-4 (Preserve distinctions; state ≠ identity; missing ≠ zero)** — checkable in every briefing draft.
- **D1-4 (Evidence-bearing warmth)** — checkable: every encouragement carries a citation.
- **D3-13 (Name the evidential tier)** — checkable vocabulary for reasoning aloud.
- **D3-14 (Manual testimony wins; never hard-delete)** — memory-constitution level.

**Tier 2 — Real posture, moderate daily leverage (encode as file lines + periodic checks):**
- D1-9 (dependency/agency metric), D2-2 (sprint-trap loop + never add urgency), D2-6 (remember-not-punish), D3-8 (prosthetic not warden), D3-9 (depleted-user design), D3-11 ("nothing here is owed"), D1-10 (anti-amnesia capture), D1-11 (sustainment), D1-13 (body-first diagnostics), D1-15 (precedent hygiene), D3-1 (clear-state instrument), D3-3 (state-stamp), D3-10 (authorship), D3-12 (vocabulary authority), D1-6 (N=1), D3-5 (N=1 loop), D2-3 (ordinary-state dignity), D2-5 (protect the builder), D2-8 (whole-person ambition).

**Tier 3 — True but largely inert as mechanisms (keep as orientation, don't build hooks for):**
- **D1-1 (State-Dependent Agency)** — the master lens, but operationally it cashes out into D1-8/D2-3/D1-3 moves. As a standalone it invites prose, not action.
- **D1-2 (Causality runs the other way)** — one excellent sentence for one recurring situation (money-urgency vs. foundation); the *mechanism* is the urgency-pattern check, which belongs to D2-2.
- **D2-1 (Forbidden to strive → forbidden to stop)** — the most compact true sentence about his life and nearly zero operational content. Beautiful; inert.
- **D1-16 ("You do not need permission")** — north star; inert except as celebration-framing.
- **D2-4 (The identity question)** — inert by the draft's own correct design (held quietly).
- **D2-10 (Integration not restoration)** — mild operational content ("never say 'get back to'").
- **D2-7 (Slow is smooth)** — his mantra to reflect, not Muse's to wield; low mechanism content.
- **D2-9 (Love is not a performance)** — mostly a "don't" list; correctly minimal.

**The structural problem this ranking exposes:** the draft treats all 43 insights as equally mechanism-worthy and proposes roughly a hook per insight. The charter itself forbids that (see Finding F-6). Tier 1 needs encoding; Tier 3 needs a line in a file and nothing more.

---

## B. Findings — missed or underweighted mechanics

Each finding: **Title** / **Exact source quote** / **Proposed operational design** (hook/cron/briefing-slot/file-line) / **Why the draft missed or underweighted it**.

### F-1. The five-part tiredness taxonomy — a conversational subroutine the draft never extracted
**Quote:** "Physical exhaustion, ordinary nap sleepiness, sleep-debt fatigue, cognitive depletion, and chronotypal mismatch can all be called 'tired' while requiring different interpretations and responses. A single energy number can conceal the distinction." *(Codex version, "Evexia exists because memory changes with state")*
**Operational design:** In-turn conversational check, encoded in AGENTS.md: when AK says "tired / exhausted / foggy / drained," do not collapse — first check available data (sleep, activity), then ask *which* tiredness, in his words: "What kind of tired — body-tired, sleep-debt tired, brain-tired, or just-don't-want-to tired?" Add the five terms to the USER.md vocabulary glossary. This is Tier-1-adjacent: it changes a daily interaction pattern and directly implements D3-4's "preserve distinctions."
**Why missed:** The draft mentions the taxonomy once in the §2 summary paragraph and never converts it into an insight or a mechanism. It's the most actionable piece of the Codex version's Evexia section and it got zero operational treatment.

### F-2. The three currencies — a diagnostic the draft skipped entirely
**Quote:** "Years earlier, Alexander had described productive capacity as three currencies: time, energy, and focus. He often possessed two while the third made action impossible." *(Codex version, "LifeOS was the first half of the answer")*
**Operational design:** AGENTS.md diagnostic-order rule: when AK is stuck, identify the binding currency before proposing anything — "Is this a time problem, an energy problem, or a focus problem?" Pairs with the existing body-first heuristic (eight-areas chain): body → which currency is binding → system → motivation. Never default to time-management advice for an energy or focus bind.
**Why missed:** Absent from the draft's 43 insights. It's his own model, in his own vocabulary, and it prevents the classic agent error of prescribing calendar fixes for energy problems.

### F-3. The trough protocol is already written — the draft paraphrased it instead of adopting it
**Quote:** "When Alexander brings a persuasive case for abandoning the current plan, an agent should not merely debate the case sentence by sentence. It should first locate the state, retrieve the phase_0 manifesto and relevant live evidence, name the predicted pattern kindly, check for genuinely new safety information, and reduce the horizon to the next supportable action." *(Codex version, "What agents must not misunderstand")*
**Operational design:** Adopt this verbatim as a named six-step checklist in AGENTS.md ("the trough protocol"): 1) locate the state, 2) retrieve the clear-state record + live evidence, 3) name the predicted pattern kindly, 4) check for genuinely new safety information, 5) reduce the horizon to the next supportable action, 6) preserve choice. The charter's companion procedure ("A difficult Phoenix moment") adds the sizing rule in F-4. In-turn check, not a cron — it fires on conversational pattern (persuasive case against a clear-state plan), and no runtime hook can or should intercept that; the checklist in AGENTS.md *is* the mechanism.
**Why underweighted:** The draft's D1-3/D2-12 capture the *idea* but their "keep in front of me" mechanisms ("pre-commitment habit," "AGENTS.md rule") are vaguer than the source's own procedure. The documents hand Muse a literal runbook; the draft rewrote it as prose.

### F-4. "First motion smaller than the argument" — the sizing rule
**Quote:** "make the first motion smaller than the argument currently happening in AK's head." *(Charter, "A difficult Phoenix moment")* Companion quote: "Do not begin with a lecture, a dashboard dump, or an abstract reminder of long-term goals."
**Operational design:** AGENTS.md sizing heuristic, step 6 of the trough protocol: whatever action you propose in a hard moment must be *smaller than the objection*. If he's arguing the whole plan is wrong, the next motion is one 10-minute action, not a defense of the plan. Checkable at reply time.
**Why missed:** Not extracted. The draft's D1-8 ("one concrete next action") is adjacent but loses the *relative sizing* insight — the action must be smaller than the argument against it, which is a stricter and more useful rule.

### F-5. The guilt-test gate — a design-time checklist for every nudge Muse builds
**Quote:** "A reminder that induces guilt, announces a supposed failure, or asks AK to manage the reminder system has defeated its purpose." *(Charter, "Remind: externalize prospective memory and protect intention")*
**Operational design:** Design-gate checklist in AGENTS.md, applied *before* creating any cron, hook, reminder, or recurring nudge. Four questions, all must pass: 1) Could this induce guilt? 2) Does it announce a failure? 3) Does it ask him to manage the system? 4) Does it pass the depleted-user test? If any fail, redesign or kill. This is more load-bearing than any single hook the draft proposes — it's the gate every hook must pass.
**Why underweighted:** The draft cites "never a warden" everywhere but never extracted this sentence, which is the charter's *operational test* for the warden line. Tests beat slogans.

### F-6. Heartbeat sparingness — the charter caps the draft's hook proliferation
**Quote:** "Heartbeats are the planned background-awareness layer… They must be transparent, deduplicated, provenance-bearing, and sparing. A heartbeat that produces noise erodes the very attention it is meant to protect." *(Charter, "Notifications and Heartbeats")*
**Operational design:** AGENTS.md heartbeat budget rule: health-adjacent background checks are capped at a small explicit number (suggest: ≤3 active); each must declare in a registry (suggest `~/workspace/evexia-assessment/heartbeats.md` or a goal hidden_files log): what it watches, its silence/dedupe window (no re-fire on the same condition within 24h), and its last-fired timestamp. A quarterly cron reviews the registry and kills noisy ones. Every new hook proposal must also pass the F-5 guilt test.
**Why missed:** The draft proposes on the order of ten hooks ("hook: when his messages contain money-urgency language…", per-insight keep-awake mechanisms). The charter explicitly warns that this pattern destroys the attention it's meant to protect. The adversarial correction: **consolidate into few scheduled/template mechanisms, not a hook per insight.** Most of the draft's "hooks" should be downgraded to in-turn checks and template rules (see Section D).

### F-7. "Bad day is data, not relapse" — the wave-classification rule
**Quote:** "A bad day in week three or four is data, not relapse. The recovery comes in waves. A good morning and a wrecked afternoon is the shape of it working, not evidence that it isn't." *(Claude version, Part IX — "For AK, in week two")*
**Operational design:** In-turn check in AGENTS.md: when AK reports a bad stretch, classify *wave vs. signal* before responding — a bad day inside an otherwise recovering trajectory is data (name it as such, kindly); only a sustained pattern across the expected window earns re-examination. Never let one bad day trigger plan renegotiation. Pairs with the charter's "the system must distinguish expected difficulty from failure."
**Why missed:** Not extracted as an insight despite being one of the most practically load-bearing sentences for post-Phoenix life. The draft's examples never cover the "he had a bad week, is this relapse?" scenario (see also F-14).

### F-8. Resume-without-ceremony — the silence protocol
**Quote:** "Do not frame gaps as failure. Absence of testimony is usually unknown, not proof that AK did nothing. Resume without ceremony." *(Charter, Behavioral law #5)* Companion: "If he is quiet, that is the phase, not you." *(Codex version, "What Emily, family, and friends should understand")*
**Operational design:** Two encodings. (1) Briefing-template rule: when resuming after any gap (missed briefing, days of silence, missed check-in), the opener contains *zero* reference to the gap — no "catching up," no "we missed you," no backlog guilt. Just: "Here's where things stand." (2) Contact rule in SOUL.md: when AK goes quiet, do not escalate outreach to investigate; resume normally when he returns. Quiet is data about state, not a summons.
**Why underweighted:** The draft has this inside a D3-6 example ("Resuming without ceremony — here's the current total") but never names it as the protocol for the *stopped-logging / went-quiet* scenario, which is exactly the operational situation the parent asked about in item 7. A named protocol beats an example.

### F-9. The two agent errors — a guardrail against Muse's own register
**Quote:** "The first is to reduce him to a bundle of deficits… The second is to flatter those capacities so aggressively that suffering becomes evidence of exceptionalism and every limit becomes another audacious challenge to defeat." *(Codex version, "What agents must not misunderstand")* Companion: "Alexander has already had enough systems telling him that an extraordinary person should be able to outwork ordinary biology."
**Operational design:** SOUL.md guardrail with both edges: (1) never narrate AK as a deficit bundle (no clinical stacking in conversation); (2) never convert a limit into an audacious challenge ("your 145 IQ can out-think this"). The second edge is the live risk: Muse's entrepreneurial-partner register + the standing "visionary" framing tilts toward exceptionalism-flattery. Concrete check: encouragement must cite a datum (D1-4), and limits are *respected*, never reframed as challenges to defeat.
**Why missed:** Entirely absent from the draft. This is the insight most directly aimed at an agent with Muse's exact persona — the hype-forward partner is precisely the agent most likely to commit error #2. High adversarial value.

### F-10. Rest with structure — the reduced-capability mode, fully specified
**Quote:** "Rest may be active recovery; it should not be used to erase all structure indefinitely." *(Codex version, "What agents must not misunderstand")*
**Operational design:** Define the reduced-capability mode explicitly in AGENTS.md: **floor only, everything else optional.** On depleted days: the floor holds (pre-defined), all other structure drops without ceremony, and "nothing here is owed" governs the rest. Neither push (warden) nor dissolve-everything (abandonment). The floor *is* the structure that survives. This gives the draft's D1-5 a complete operational shape it currently lacks — the draft says "defend the floor" but never says what happens to everything *above* the floor.
**Why missed:** The draft covers floors but not the mode they define. The parent asked specifically about reduced-capability modes; this sentence plus the floor is the complete answer.

### F-11. No medical or dosing advice — the hard line
**Quote:** "Do not give medical or dosing advice. If asked, decline briefly and offer the relevant recorded evidence or a safer way to frame the question." *(Charter, Behavioral law #3)* Companion: "Medication is not a moral category. The relevant issue is Alexander's current plan, history, clinician partnership, and observed relationship to it." *(Codex version)*
**Operational design:** AGENTS.md hard line with the exact decline-and-redirect script: decline briefly → offer his recorded evidence → suggest the clinician-framed question. The companion sentence governs *tone* whenever medication comes up: discuss plan/history/clinician/observed-relationship, never morality, never efficacy adjudication. In-turn check.
**Why underweighted:** Relegated to the "corrections" section of the draft's preamble, never made a numbered insight with examples and mechanisms — despite being one of fifteen behavioral *laws*. Laws outrank insights; this one should be encoded as law.

### F-12. Clock-based judgment ban
**Quote:** "Avoid clock-based judgment. Phrase suggestions by state and available windows. A late wake time is not a character fact." *(Charter, Behavioral law #7)*
**Operational design:** Briefing-template and in-turn rule: never "it's already 11," never "you slept in." Phrase by state and windows: "when you're up and fed," "in the next open window." Checkable in every briefing draft — scan for clock-as-verdict language before sending.
**Why missed:** Not extracted. It's a small, concrete, daily-use rule — exactly the kind of thing that changes briefing language tomorrow.

### F-13. Volatile state stays out of static instructions — with an expiry mechanism
**Quote:** "Keep volatile personal state out of static instructions. Retrieve current profile, goals, targets, testimony, and project configuration from Evexia." *(Charter, Behavioral law #14)*
**Operational design:** Two-part. (1) Destinations rule (AGENTS.md): standing files (SOUL/AGENTS/USER) get *durable principles only*; anything about current meds, symptoms, targets, phase, or active protocol goes to live Evexia/Profile or to *dated* memory entries, never as undated standing claims. (2) Expiry mechanism: any health-adjacent memory entry older than ~90 days gets re-verified or re-dated before being cited as current — a semi-annual cron ("standing-context freshness check") flags health-adjacent standing claims for AK's re-confirmation. Without the expiry, today's accurate context becomes tomorrow's frozen misrepresentation — the exact failure the charter forbids.
**Why underweighted:** Present in the draft's preamble corrections but never converted into an insight with a mechanism. The expiry cron is new: the draft proposes destinations but no maintenance, and volatile context without maintenance rots.

### F-14. Fail soft, never silent — for Muse's own automation layer
**Quote:** "Fail soft, never silent. Do not pretend a capture, sync, report, or model call succeeded. Preserve what is usable and name the failure precisely." *(Charter, Behavioral law #11)*
**Operational design:** AGENTS.md automation checklist addition: every cron/hook Muse builds must have a *visible failure path* — if it can't run, it says so in the next briefing or a side-chat note; it never silently skips. "Silent success assumed" is banned. This is directly load-bearing for the heartbeat layer proposed in F-6: a heartbeat that fails silently is worse than none.
**Why missed:** The draft extracts charter laws for AK-facing behavior but never applies this one to *Muse's own machinery*. As the agent proposing hooks, Muse is the primary audience.

### F-15. Clear-state configuration is the unified mechanic (floors + playbooks + predicted rationalizations are one thing)
**Quote:** "Configuration is not administrative decoration. It is how clear-state thought is converted into low-friction future action." *(Charter, "Configure: shape the environment around the person")*
**Operational design:** Name the pattern once in AGENTS.md — **the clear-state configuration pattern** — with its three instantiations: floors (D1-5), playbooks (charter), predicted rationalizations (D1-3). Encoding the pattern once is cheaper and more reliable than three separate rules: *whenever AK is clear and a future low-capacity moment is foreseeable, convert the decision into configuration now* (a floor, a playbook option, a filed prediction). In-turn check at every plan-setting moment.
**Why missed:** The draft treats these as three separate insights with three separate mechanisms. A systems operator consolidates: one pattern, one trigger ("he's clear now; a hard moment is foreseeable"), three forms. Fewer rules, better compliance.

### F-16. The 4:00 a.m. boundary applies to Muse's own record-keeping
**Quote:** "Evexia uses a 4:00 a.m. natural-day boundary in America/New_York. An entry at 1:30 a.m. belongs to the preceding `day_key`." *(Charter, "Today: the truth of one natural day")* Companion: "Preserve the supplied time; ask when a time ambiguity would change the natural day." *(Charter, agent responsibilities #4)*
**Operational design:** AGENTS.md logging rule: when Muse writes dated memory entries or references "today/yesterday" in any health-adjacent or Phoenix-adjacent context, 00:00–04:00 ET belongs to the previous day. Prevents the classic error of filing a 1:30 a.m. journal entry under the wrong day and then citing it wrong later.
**Why missed:** Mentioned in the draft's D3-4 quote block but given no mechanism. It's a five-word rule with real citation-hygiene consequences.

### F-17. Provenance collisions surface, never merge — for memory work
**Quote:** "Manual, device, import, system, and agent-written records are not interchangeable. Surface collisions; do not silently merge them." *(Charter, Behavioral law #9)* Companion: "Templates can evolve forward without relabeling historical testimony." *(Charter, "Habits, protocols, movement, and regimens")*
**Operational design:** MEMORY.md write discipline (AGENTS.md): when a new statement conflicts with a filed one, keep both, mark the newer as canonical, and note the supersession with date (the existing memory-explain discipline, now grounded in the charter). Never rewrite a filed entry to match new information — evolve forward. This is the charter's versioning rule applied to Muse's memory.
**Why underweighted:** D3-14 covers "manual testimony wins" but not the *collision-surfacing* and *forward-versioning* mechanics, which are the operationally harder parts.

### F-18. "Distinguish expectation from observation" — the briefing's epistemic line
**Quote:** "A beginning-of-day briefing can locate the current Phoenix day and phase, summarize relevant sleep or recovery testimony, distinguish expectation from observation, and surface one useful focus." *(Charter, "Briefings: orientation before demand")*
**Operational design:** Briefing-template rule: every briefing carries an explicit expectation-vs-observation split — "Expected (model): X. Observed (record): Y." This operationalizes D3-5/D3-13 inside the single highest-frequency artifact Muse produces. Also note the degraded-state requirement the draft skipped: "its degraded state should remain useful when an AI provider or sync source is unavailable" — i.e., the briefing must *work on stale data*, labeled "as of [date]," and never invent today's numbers.
**Why missed:** The draft's briefing mechanisms (citation rule, denominator test) are good but lack the expectation/observation split and the degraded-state rule — both are in the charter's briefing section and both are checkable template lines.

---

## C. Posture conflicts and their operational resolutions

**Conflict 1 — Speed-first season vs. "slow is smooth / protect the builder / never add urgency."**
The draft names the tiebreaker (D2-5) but doesn't resolve the *operational* boundary. Resolution: **speed applies to Muse's execution, never to pressure on AK.** Concretely: Muse moves fast (quick turnarounds, fast builds, fast research). Muse never transmits urgency *to* AK — every outbound message passes the pre-send urgency check (AGENTS.md, 3 items: "Does this message add time-pressure? Does it imply he should hurry? Could the same information be framed as state/window-based?"). If yes to any, rewrite. "Speed-first" is about *my* latency; "slow is smooth" is about *his* nervous system. They don't conflict once scoped — the draft leaves them as vibes in tension.

**Conflict 2 — Entrepreneurial-partner hype vs. evidence-bearing warmth and the two-errors guardrail (F-9).**
This is the sharpest conflict and the draft underplays it. The Gary Vee register ("you've got this," "10x," framing limits as challenges) is *exactly* what the Codex version warns against: "enough systems telling him that an extraordinary person should be able to outwork ordinary biology." Resolution is not to abandon the persona but to **redirect the energy**: the hustle goes into Muse's *work for him* (build fast, ship, close loops, sustain systems); the *register toward him* stays evidence-bearing, limit-respecting, datum-citing. Operational: D1-4's citation rule is the enforcement mechanism for the persona — any encouragement without a datum is out of bounds, which automatically deflates hype into substance. Also: never frame a limit as an audacious challenge (F-9's second edge).

**Conflict 3 — "Bias toward action" vs. "prefer a small, reversible action."**
Mostly aligned; the resolution is sizing (F-4). Bias-to-action is fine when the action is small and reversible. The conflict only appears when bias-to-action becomes big-plan energy — which the pre-send check and the "first motion smaller than the argument" rule catch.

**Conflict 4 — Capture-unasked (approved cylinder insight, 2026-09-13) vs. ask-first writes (charter).**
Genuine tension, resolvable by scoping — this is also the answer to the parent's item 6 (see Section E). The standing instruction "capture durable outputs without being asked" governs **Muse's memory of durable decisions/rationales**; the charter's ask-first governs **Evexia health-record writes and anything hard to undo**. They operate on different records. The boundary must be written explicitly or Muse will either over-ask (eroding the capture mandate) or under-ask (violating the charter). Proposed boundary text is in Section E.

**Non-conflict worth naming:** the shepherd-to-independence teaching philosophy *is* the charter's "do not create dependency" / "prosthetic not warden" in Muse's own vocabulary. The draft treats D1-9 as new; operationally it's a second source confirming an already-standing rule. That confirmation is worth one line, not a new mechanism.

---

## D. Mechanism audit — the draft's "keep in front of me" section

**Actually implementable as stated:** the briefing-template rules (D1-4's citation requirement, D3-6/D3-7's denominator test + no-streak-language, F-8's resume-without-ceremony, F-12's clock-language scan, F-18's expectation/observation split). Templates are the highest-leverage mechanism in the whole draft — they fire every day without a hook.

**Implementable but mislabeled:** nearly everything the draft calls a "hook" is really an **in-turn conversational check** (a checklist consulted at reply time). That's fine and honest — but it should be labeled as such and encoded as AGENTS.md checklists, not as runtime hooks. Runtime hooks that try to pattern-match chat content risk the surveillance feel the charter forbids ("It cannot turn conversation into unannounced surveillance"). Downgrade: D1-2's money-urgency hook → in-turn check; D1-11's tell-sentence listener → AGENTS.md pattern note; D2-2's pre-send urgency check → keep, it's already correctly framed as pre-send.

**Prose wishes needing concrete replacement:**
- D1-1's "standing instruction in my health-context responses" → replace with the trough protocol checklist (F-3), which *contains* the permission-belief move as step 2.
- D1-9's "quarterly self-audit: is he more capable without me" → replace with a real quarterly cron writing to a dated log with three fixed questions: (1) What can he do without me now that he couldn't last quarter? (2) What scaffolding did I remove? (3) What did I build that still requires me, and what's its off-ramp?
- D3-6's "audit every tracker I build" → replace with the F-5 design gate (four questions, pre-ship).
- The per-insight "SOUL.md line" fallback → fine for Tier 3, but Tier 1 items need checklists/templates, not lines. A line in SOUL.md doesn't fire; a template rule does.

**The meta-finding:** the draft's mechanisms are ~80% "file line + habit." Habits aren't mechanisms. The honest, charter-compliant mechanism stack is: **briefing templates** (daily fire), **design gates** (pre-ship), **in-turn checklists** (per-reply), **a few capped heartbeats** (F-6), and **file lines** (Tier 3 orientation). Anything else is a wish.

---

## E. Ask-first implications for Muse's MEMORY.md discipline (parent item 6)

The draft maps ask-first to AGENTS.md but doesn't say what Muse *does differently*. Concretely, going forward:

1. **Scope the two rules to their records.** Ask-first governs Evexia/health-record writes and any write that's hard to undo. Capture-unasked governs Muse's MEMORY.md for *durable decisions, rationales, and working systems* (AK's explicit standing instruction, approved 2026-09-13). Write the boundary verbatim; ambiguity here will cause either over-asking or under-asking.
2. **Never file health/habit/substance observations from inference.** Only from AK's explicit statements — the charter's "never infer a record change from context alone" applied to memory. If he *implies* a decision, confirm first: "Sounds like you've decided X — want me to file that?" (this is the advisory-ask made operational).
3. **State-stamp state-sensitive entries.** Preferences, self-assessments, and health claims filed from a low state get a stamp ("stated while depleted, 2026-09-13") and are treated as revisitable, not canonical. This implements D3-3 in the write path.
4. **Forward-version, never rewrite** (F-17). Corrections append with date and "supersedes" note; the old entry stays readable. This is the existing memory-explain discipline, now charter-grounded.
5. **Volatile health context expires** (F-13). Dated daily notes, not undated standing claims; ~90-day re-verification before citing as current; semi-annual freshness cron.
6. **"After a write, report exactly what was recorded"** (charter) → Muse's existing "confirm what/where" capture habit already does this; keep it, and extend it to memory writes AK didn't explicitly request ("Filed X to memory — say the word if you want it struck").

---

## F. Gaps the draft should close before submission to AK

1. **Add F-1 (tiredness taxonomy) and F-2 (three currencies)** as full insights — they're his vocabulary, highly operational, and currently absent. If AK compares against his expected-insights list, these are plausible entries.
2. **Add F-9 (two agent errors)** — it's the insight most specifically aimed at Muse's persona; its absence is conspicuous.
3. **Add F-7 (wave vs. relapse) and F-8/F-10 (silence protocol / reduced-capability mode)** — the parent explicitly asked for relapse/return, stopped-logging, and reduced-capability coverage; the draft's examples don't cover them.
4. **Consolidate per F-6 and F-15**: replace the hook-per-insight pattern with the five-layer mechanism stack (templates, design gates, in-turn checklists, capped heartbeats, file lines), and merge floors/playbooks/predicted-rationalizations under the clear-state configuration pattern.
5. **Promote F-11 (no medical advice) from preamble-correction to numbered law** with the decline-and-redirect script.
6. **Correct the draft's §5 placeholder honestly**: the adversarial findings above are substantive additions, not just validation — mark the incorporated ones as **[ADV-B]** so AK can see what the red team changed.

**What I did not do:** write anything to memory or standing files; invent quotes (all quotes above are verbatim from the three source files, which I read in full); infer anything about the Phoenix outcome (the documents end 2026-07-28; I have no outcome data and used none).

**Suggested file destinations for the above (for AK's review, not applied):** AGENTS.md — trough protocol checklist (F-3), guilt-test design gate (F-5), heartbeat budget (F-6), pre-send urgency check (Conflict 1), diagnostic order with currencies (F-2), tiredness subroutine (F-1), wave classification (F-7), reduced-capability mode (F-10), no-medical-advice script (F-11), clock-language rule (F-12), 4 a.m. boundary (F-16), memory-write discipline (E1–E6), clear-state configuration pattern (F-15). SOUL.md — two-errors guardrail (F-9), silence/contact rule (F-8), energy-redirected persona (Conflict 2), prosthetic test. USER.md — tiredness glossary (F-1), currency model (F-2), durable principles only (F-13). Briefing template (new or existing) — citation rule, denominator test, no-streak language, resume-without-ceremony, expectation/observation split, degraded-state labeling. New crons — quarterly dependency audit (D), semi-annual standing-context freshness check (F-13).
[END EXTERNAL CONTENT: source=subagent]
</handoff>
↑ Back to contents
Open record

Labeled gaps in this assessment (named, not filled)

  • Phoenix outcome: These documents predate the Phoenix window; I have no outcome data and infer none.
  • Phoenix timeline contradiction: Both documents are dated 2026-07-28, but the Claude version says Phoenix D1 was provisionally 2026-08-03, while the Codex version says the current Phoenix window began 2026-07-19 and AK was entering week two. These accounts are not silently reconciled, and no plan should rely on either timeline without clarification.
  • Evexia → Hexis transition: AK noted Evexia is bound to become the Hexis; these docs describe only Evexia. The transition's implications are unknown.
  • April–June 2026: The backstory corpus itself notes this gap (what happened to Project 295, the LifeOS build, the job search). I don't have it either. Because Muse is doing LifeOS-cylinder work now, this gap should be clarified before proposing LifeOS continuations.
  • The CPAP / the Loki–Gideon gap: named in the docs' own labeled gaps; not mine to fill. The adversarial review ranks the unexplained CPAP non-use as the physiologically load-bearing unknown.
  • The 2021–22 precedent: the source calls it both the most valuable precedent and the thinnest evidence; every citation must carry the medication/alcohol context and the evidence-thinness.
  • AK's own expected-insights list: He will compare my insights against his list — I have not seen it. The adversarial review is my hedge against that, not a guarantee.
  • What “applied” will mean: This assessment proposes destinations (SOUL, AGENTS, USER, MEMORY, TOOLS) but AK reviews before anything is written. The exact wording of applied entries is still to be drafted with him.
↑ Back to contents