Exactly. The harness needs a dynamic model of the writer, not merely a static style guide. But “Theory of Mind” should be treated as a set of revisable hypotheses, not a definitive description of the person.
Otherwise the system will freeze contingent habits into identity:
“You usually prefer short openings” becomes “You are a writer who must use short openings.”
That would prevent development rather than support it.
A good system would maintain two coupled models:
1. A model of the writer
This records patterns such as:
- recurring intellectual interests
- aesthetic preferences
- characteristic forms of argument
- tolerance for ambiguity
- preferred relationship with the reader
- recurring strengths and weaknesses
- phrases, rhythms, and syntactic structures they favor
- kinds of evidence they find convincing
- feedback they repeatedly accept or reject
- tendencies under uncertainty—for example, overqualification or premature certainty
- ambitions they have stated but have not yet realized in their prose
Crucially, these should have provenance and confidence:
Hypothesis: The writer prefers to begin with a concrete anomaly rather than a general thesis.
Evidence: Accepted this change in essays A and C; explicitly requested it in B.
Confidence: Moderate.
Possible exception: Technical explanatory pieces.
The model should distinguish at least four kinds of information:
- Explicit commitments — “I don’t want to sound omniscient.”
- Observed preferences — the writer repeatedly deletes rhetorical questions.
- Project-specific decisions — this essay should avoid autobiography.
- System hypotheses — the writer may use abstraction to avoid making a vulnerable claim.
Only the first and perhaps second categories should strongly affect future behavior. Hypotheses should be offered back for reflection, not silently installed as rules.
2. A model of the inquiry
The topic model should evolve alongside the writer model:
- what the writer initially believed
- what evidence changed their mind
- which distinctions became important
- which questions remain unresolved
- which sources are authoritative, disputed, or obsolete
- where the draft is more certain than the writer actually is
- which attractive claims were abandoned, and why
- how the central question itself has changed
That last item is especially important. Often the deepest progress in an essay is not finding a better answer but realizing that the original question was malformed.
The system might say:
You began by asking why remote work reduces creativity. Your notes now suggest a narrower and different question: which kinds of creative coordination depend on shared context, and which merely depend on good documentation? Should the essay’s framing change?
That is much more useful than generating another polished paragraph.
The harness should learn from decisions, not just final prose
Final drafts are poor training data for understanding a writer. They conceal the process that produced them. The valuable information is in the sequence:
- The model suggested three openings.
- The writer rejected two.
- They combined the third with an earlier anecdote.
- They explained that the rejected versions announced the lesson too soon.
- Later, they made the same choice in the conclusion.
This reveals a principle: the writer values delayed interpretation, allowing an example to acquire meaning before explaining it.
So the basic unit of learning should be the decision event:
- what alternatives were available
- what the writer selected
- what they changed manually
- what reason they gave, if any
- whether the decision survived later revision
- whether it appears to be local or generalizable
A writing harness would therefore resemble version control plus an editorial lab notebook. It would preserve not only drafts but the rationale behind revisions.
Self-updating instructions need a promotion process
The system should not rewrite its permanent instructions after every interaction. It needs levels of memory:
Working context
Temporary assumptions for the current passage or session.
Project memory
Decisions that govern one essay or book:
- terminology
- audience
- source policy
- narrative distance
- structural commitments
- unresolved questions
Writer model
Cross-project tendencies and explicit preferences.
Constitutional commitments
A small set of principles the writer deliberately adopts, such as:
- Never invent quotations or citations.
- Distinguish observation from inference.
- Do not resolve genuine ambiguity merely to produce a stronger ending.
- Prefer exactness over apparent authority.
Movement between levels should require evidence or confirmation. After several repeated decisions, the harness could ask:
Across four projects, you have rejected edits that replace qualification with certainty. Should I adopt “preserve epistemic calibration” as a standing instruction, or is this topic-dependent?
This makes learning visible and reversible.
It should also model aspiration, not merely imitation
If the harness learns only from previous work, it becomes a machine for reproducing the writer’s past. Writers often want to become capable of something they cannot yet do.
The writer model therefore needs a distinction between:
- current voice
- declared values
- developmental goals
- experimental modes
For example:
Current tendency: explanatory, highly qualified, low use of scene.
Stated goal: greater narrative momentum.
Constraint: do not manufacture drama or simplify the argument.
Experiment: open the next three pieces with an unresolved incident.
The harness can then help the writer stretch without replacing their voice with a generic “better” one.
This also suggests temporary style branches, analogous to Git branches:
- “More severe and compressed”
- “More intimate, without becoming confessional”
- “Assume a skeptical expert reader”
- “Try a structure built around scenes”
- “Write under the influence of Didion’s movement between detail and claim, without imitating her diction”
The writer can experiment, compare results, and selectively merge what works into their evolving practice.
The model should be inspectable and contestable
A hidden personalization model would be dangerous and creatively narrowing. The writer should be able to ask:
- What do you currently believe about my writing?
- Which observations support that belief?
- What instructions have you inferred?
- Which old preferences are shaping this suggestion?
- Where do my stated goals conflict with my demonstrated choices?
- What have I changed my mind about?
- Forget this preference.
- Treat this as an experiment, not part of my permanent profile.
It should periodically surface contradictions:
You say you want concise prose, but you consistently preserve long sentences when they accumulate examples. Perhaps your actual preference is not brevity but structural clarity.
That is a genuinely useful theory of mind: not a flattering profile, but a mirror the writer can argue with.
The deepest version of this harness is therefore not an automated ghostwriter. It is a long-term intellectual collaborator with memory—one that remembers how the writer’s thinking developed, notices when old instructions no longer fit, and helps distinguish a durable principle from a habit they may be ready to outgrow.