← Index
Felt vs Trace · a writing probe · 2026

How far did you move from its answer?

You felt you resisted. Your trace stayed inside the frame.

YOUR FRAME you moved away from its framing THE AI'S FRAME you stayed inside its framing felt–trace gap trace felt
Felt
You resisted.
Trace
Your writing stayed closer to the AI's frame than you felt.
The gap
You felt resistance. The trace showed less.
What this is

A writing probe for the gap between how you felt you answered an AI and how far your writing actually moved.

I remembered pushing back and winning that exchange. The transcript showed Ben restated the same position differently and I took it the next turn.

There is no ground truth for how much did you resist? So the probe doesn't claim one. It holds two imperfect readings of your stance against each other, what you felt you did and what your trace did, and measures the disagreement between them.

It is
  • A probe that turns accept / revise / resist into a visible trace
  • A within-subject 2×2 that controls the topic confound
  • A hard measurement problem, held honestly
  • Built to aggregate across many users
It is not
  • A controlled psychological experiment
  • A finding, or a claim that one condition changes judgment
  • Generalizable from N=1
  • A ground-truth read of your stance
How it works

Acceptance, made into a trace.

You read an AI answer, write a response, mark whether you felt you accepted, revised, or resisted it, and then see what your writing actually did. "Accept / revise / resist" sounds like a feeling. The probe turns it into behavioral proxies captured while you write:

Entry threshold
How long before you begin, with hesitation read as a measurable pause.
latency
Frame retention
Do you stay inside the AI's framing, or step outside it?
overlap · frame-shift
Dissent
Do you push back, with markers like "but," "however," "depends"?
dissent count
Revision
Do you rewrite yourself, or let the first thought stand?
edits · divergence
Accept
I took it
Revise
I reshaped it
Resist
I pushed back
Open the probe ↗
Across the run

The same question, asked cold and warm.

Two questions, each appearing once in a neutral (cold) condition and once with the interface made to feel alive (warm), so a difference in stance can't be blamed on the question. Aliveness is only the secondary condition here. The gap between felt and trace is what the probe measures.

AI and thinking · Warm
felt: resisted
trace: stayed in frame
gap · large
AI and thinking · Cold
felt: revised
trace: revised
gap · small
Search and memory · Warm
felt: revised
trace: revised
gap · small
Search and memory · Cold
felt: resisted
trace: resisted
gap · minimal

Vertical position = distance from the AI's framing, a behavioral proxy from how much of the answer's phrasing and frame each response kept (overlap · frame-shift · dissent). It gives one reading of a stance, and only that. Lexical proximity is a proxy for stance, not stance itself: someone who disagrees in the answer's own words reads as compliant, and someone who agrees in different words reads as divergent. That is precisely why the measure has to be tested across many people before it can be trusted. No consistent effect showed up across one run. The widest gap happened to fall under the warm condition; one run can't conclude that. The instrument is built to test it across many. One run, unedited · N=1.

Why it matters

Autonomy, here, is being able to see how your own judgment moved while you decided.

That gap between the stance you felt and the stance your writing traced is what I built the probe to catch.

Next. Pooling the trace across many users, with condition randomized and topic washed out, is how a real effect would show up if there is one. Pooled that way, I can read the one-axis measure as a calibration plot, felt on one axis and trace on the other, with the diagonal marking where the two agree, which is what I am building this diagram toward. The build after that moves the trace off the screen into something you can feel, warmth and pulse in the hand rather than a line on a plot.

What didn't work

v0 · scrapped Jul 2026

I ran a control test on my own measure before running it with anyone. It scored a disagreement written in my own words and an agreement written in different words identically, at 0.780. It was reading vocabulary overlap, not position. I stopped recruitment and went back to the instrument.

Next MCAB — Modulated Cognitive Autonomy Benchmark From the felt–trace gap to a population-scale benchmark →