Return, at the end, to where you began. The mark on the wall has not moved. You are the only thing here that might have.
At Station I you were asked to take a set of readings and told explicitly that they were not a score to improve. They were a mark on the wall.
Take them again now.
This is not a graduation and it is not a test. It is the last application of the discipline the whole ring has run on: guess first, then check. You have spent months with these practices and you have a belief about whether they worked. That belief is a prediction. This station is where you find out how good it was — which is, in the end, the same skill Station I was training, applied to the training itself.
| Arena | What this looks like there |
|---|---|
| Any sport | The end-of-season review that is usually done on feeling alone. |
| Endurance | Retesting a benchmark piece, which no athlete would skip, and its interior equivalent, which almost everyone does. |
| Team sports | A squad that assumes its culture has improved, and the measure nobody took at the start. |
| Coaching | Whether the psychological work you introduced actually changed anything. |
| Masters sport | Change measured against last year rather than against twenty-five years ago. |
| Rehabilitation | Return-to-play criteria for the body, and their complete absence for the state. |
| Beyond the arena | A year of effort, and no baseline against which to know. |
Two failures, and both are about what the number is taken to mean.
The first is the verdict. The retest read as a grade — improvement as success, stasis as failure, a decline as evidence the whole thing was pointless. This misreads what the instruments do. A ZSR trend is a description of a period, and periods contain injuries, bereavements, examinations, selection decisions and winters. An athlete whose scores held flat through a genuinely brutal year has a substantial result.
The second is the dismissal. The retest waved away — I know how I'm doing — which returns the athlete precisely to the uncalibrated instrument of Station III, now with eleven stations of technique layered on top of a perception nobody has checked in months.
The correct posture is the one from Station I, applied at longer range: register, note, put it down. It is the same glance. It is simply taken across a season rather than across a morning.
Predict the delta, then measure it, and keep the difference.
Before you retake anything, write down what you expect to have changed and by roughly how much. Be specific. Then take the instruments. The gap between your prediction and the result is the last and best piece of data this ring produces — because it tells you not how regulated you are, but how accurately you can now see yourself, which is the capacity everything else was built on.
A worked example. Before the retest, in writing: I expect the reappraisal score up meaningfully, suppression down slightly, ZSR roughly flat, resting HRV up a little. I expect no change in how I handle selection news. Sealed. Then the battery.
The last sentence is the valuable one. Predictions about what has not changed are consistently more accurate than predictions about what has, and an athlete who can specify what remains unmoved after nine months has demonstrated the exact capacity the ring set out to build.
There is a habit in the Sufi orders called muḥāsaba — the accounting — which is performed daily, and a longer version performed at intervals across a life.
Al-Muhāsibī, whose name literally means the one who takes account of himself, built an entire method on the assumption that the self is an unreliable narrator requiring periodic audit. Not because it lies deliberately, but because it is too close to its own material to be trusted with the summary.
What is notable is the tone the tradition insists on. The accounting is not a tribunal. It is performed with the same attitude a competent steward brings to a ledger: interested, unsentimental, neither defensive nor self-punishing. A merchant who cannot bear to look at the books does not thereby improve the accounts.
The Stoics kept the same practice at daily scale — Seneca's evening review, in which he questions his own conduct with, in his phrase, nothing hidden and nothing passed over, and then goes to sleep. The Confucian version is Zengzi's daily threefold self-examination in the Analects. Three traditions, the same instrument: a periodic, dispassionate, honest reading.
They differ from this ring in one respect only. They had no way to check their reading against anything external, and you do.
The evidence on retesting is mostly evidence about how badly people estimate their own change.
Retrospective self-report is unreliable. People systematically misremember their prior state in ways that conform to their beliefs about whether they have changed — recalling themselves as worse before an intervention they believe worked, and as similar before one they believe did not. Pre-measurement is the only defence, and it has to be taken before rather than reconstructed after.
Response shift. A well-documented complication: as people become more perceptive about a construct, their internal standard for rating it changes. An athlete who has spent months learning to detect states may rate their regulation lower than at baseline while having genuinely improved, because they can now see what they previously could not. This is the single most important caveat at this station and it should be expected rather than explained away.
Trait change is slow and real. Longitudinal work on contemplative and biofeedback interventions shows resting-parameter shifts over months, with effect sizes that are modest and durable — the opposite profile to the large, temporary effects of acute techniques.
Regulatory flexibility predicts better than proficiency. Bonanno and Burton (2013) suggest the outcome that matters is not skill with any single strategy but the ability to select appropriately by context and to abandon a strategy that is not working. That is what a twelve-station ring is actually building, and it is poorly captured by any single instrument.
Evidence status. Recall bias in retrospective self-report: strong. Response shift: well documented in health outcomes research, rarely accounted for in sport. Durability of trait-level change: reasonable evidence, small effects. Measurement of regulatory flexibility: an open problem, and this ring does not pretend otherwise.
A composite case. An athlete finishes the ring convinced it has not worked. Her retest confirms it: her self-reported regulation score has fallen slightly against baseline. Her resting HRV has risen, her recovery slopes have shortened, and her coach reports fewer disrupted sessions across the same period. She has not deteriorated. She has become considerably better at noticing what she previously could not see, and is now rating herself against a standard she did not possess in March. This is the most common result of a well-run regulation block and the one most likely to be misread as failure.
What the ring cannot measure. The outcome that matters most — regulatory flexibility, the ability to select the right strategy for the situation and abandon one that is not working — has no good instrument. Every measure in the battery captures capacity with a particular strategy; none captures the judgement that decides between them. That gap is worth naming rather than papering over, and it is the reason the final read includes a coach's observation and not only a set of numbers.
Why retrospective judgement fails. The recall problem is not memory decay; it is reconstruction. People estimate their past state by taking their present state and adjusting it in the direction their theory of change predicts. Someone who believes an intervention worked recalls themselves as having been worse beforehand; someone who believes it failed recalls themselves as having been much the same. The adjustment is invisible from inside and it is large enough to manufacture an entire result.
This is the whole justification for having taken a baseline in March that nobody wanted to take. A measurement made before the belief existed is the only version of the past that has not been edited by the present.
| Ring | What it says at this station |
|---|---|
| Sufi | The home tradition. Muḥāsaba, the accounting — conducted as a steward reads a ledger, not as a court passes sentence. |
| Stoic | Seneca's evening review: nothing hidden, nothing passed over, and then sleep. The audit explicitly ends. |
| Confucian | Zengzi's threefold daily self-examination, and the tradition's assumption that cultivation without review is not cultivation. |
| Buddhist | The recognition of one's own progress as itself a hazard — the tradition's warnings about attainment are warnings about this exact station. |
| Zen | Dōkus, the private interview: progress checked against someone else's eyes, precisely because your own are unreliable here. |
| Gītā | Sthitaprajña, the person of steady wisdom — described by behavioural signs rather than by inner report, which is a measurement decision. |
| Ubuntu | The check performed by the community. Others can see a change in you that you cannot, and their reading is data. |
Absent, and why.
Where they disagree. There is a real objection to this station and it should be given its due. The Taoist position is that measuring one's own progress interferes with the process being measured — that the person checking whether they have become less self-preoccupied has just been self-preoccupied. The Buddhist material makes a related and sharper point: attainment is itself a hazard, and the practitioner pleased with their progress has acquired a new problem.
Against that stands the plain fact that people are poor judges of their own change and systematically rewrite their own baselines. The compromise in this station is the annual cadence and the single reading: look once, honestly, and put it down — which is the same instruction Station I gave about a morning, applied to a year.
Five moves, and the first has to happen before you look at anything.
The measured interior. The comparison worth making is not baseline against today. It is variance then against variance now.
The clearest signature of trait-level regulatory change is not a higher mean but a narrower spread: fewer extreme days, faster recovery slopes, less catastrophic response to ordinary provocations. Look at the distribution across a month, not the level on a morning.
And then, finally, the instruction that has governed every station: the reading belongs to you. It is not a report to a coach, a qualification, or a number to defend. It is an account of a period in your own life, returned to the person who lived it. Read it once, honestly, and put it down.
How you know it worked. Not from any single instrument. The convergent picture is what counts: a narrower distribution across a month, shorter recovery slopes after difficult events, a prediction gap smaller than it was in March, and a coach's independent observation of fewer disrupted sessions. Any one of those alone is weak. Together they are about as good as this kind of evidence gets.
And one final marker, which is behavioural and slightly perverse: whether you are willing to do this again in a year. An athlete who found the retest useful rather than threatening has completed the thing the ring was actually building — a relationship with their own interior that can survive being looked at.
A week, and then a decision about the year.
Day 1 · Predict. Write the full prediction before touching an instrument. Include what you expect to have not changed, which is usually the more accurate half.
Days 2–3 · Retest. Full battery, matched conditions. Do not look at baseline values while completing them — the contamination is real and well documented.
Day 4 · Write the period. Before comparing anything, write what the months contained. Injuries, load, selection, life. This is the denominator and the results are meaningless without it.
Day 5 · Compare. Baseline, retest, prediction. Three columns. Note the prediction gap specifically — it is the most informative figure on the page.
Days 6–7 · Decide the year. One practice, one anchor, one maintenance dose. Write it down. Then set a date, twelve months out, to do this again.
What usually goes wrong. The comparison is made without the denominator. An athlete who compares month one to month nine without first writing down what the nine months contained will attribute a season of injury, examinations and a house move to their own regulatory failure.
The second failure is keeping everything. The ring closes with a single practice at a maintenance dose for a reason, and the athlete who leaves with a twelve-item programme has not finished the ring. They have collected it.
The last coaching note in the ring, and it is the same one as the first.
The retest is the athlete's, not yours. If these results feed selection, contracts or comparison, you will get managed data from here on and you will have destroyed the instrument's only function.
Ask for their prediction before you show them anything. Every time, to the end. This has been the method for twelve stations and it does not stop being the method because the ring is finishing.
Expect and explain the response shift. An athlete whose self-report has dropped while their physiology has improved will read that as failure unless someone tells them otherwise. Tell them beforehand.
Interpret against the year. You know what the season contained. The athlete is often the worst-placed person to weigh their own results against their own circumstances, and this is where a coach genuinely adds something.
Help them cut to one. The strongest thing you can do at the close is prevent an enthusiastic athlete from attempting to maintain everything. Choose one, protect it in the programme, and revisit in a year.
And do it yourself. A coach who has run the ring alongside their squad has a baseline too, and the squad's regulatory state has a great deal to do with yours.
Beyond the arena. Almost nobody outside sport takes a baseline before attempting to change themselves, which is why almost nobody knows whether anything they have tried has worked. The transferable habit is not the battery. It is the mark on the wall, taken before, by someone who intends to look again.
Twelve stations, and the claim at the end is deliberately small.
You can detect a state earlier than you could. You can open a gap, and buy one when it will not open, without paying for it in silence. You can change a reading when a truer one exists, recruit a charge when it fits the task, and stop pushing when nothing will move. You have laid some of it down by repetition, and you know who you borrow from.
What you are not is regulated, finished, or calm by disposition. The weather still arrives. It always did — and this ring has never once suggested that the point was to stop the weather.
The point was to be someone who can read it, stand in it, and act anyway. That capacity is not achieved. It is maintained, in unremarkable increments, on ordinary days, mostly when nothing is wrong. Which is why the last instruction in the ring is the least dramatic one available: choose one practice, keep it, and come back in a year.
Sources & further reading.
Traditional. Al-Muḥāsibī, on muḥāsaba. Al-Qushayri, Risāla. Seneca, De Ira III.36, the evening review. Analects I.4, Zengzi’s threefold examination. Bhagavad Gītā II, on sthitaprajña. Homer, Odyssey XXIII, recognition established by test.
Research. Ross, M. (1989). Relation of implicit theories to the construction of personal histories. Howard, G. S., et al. (1979). Internal invalidity in pretest–posttest self-report evaluations. Schwartz, C. E., & Sprangers, M. A. G. (1999). Methodological approaches for assessing response shift. Bonanno, G. A., & Burton, C. L. (2013), as above. Lehrer, P. M., et al. (2003), as above. Goldberg, S. B., et al. (2018). Mindfulness-based interventions for psychiatric disorders: a systematic review and meta-analysis.
Instruments referenced are SportsFlow EPAB measures; see the Research Library for scoring and administration.