Sports Flow · Research Library
Learning to Regulate · No. 03

The Instrumentand the Hand

ZHĒNG MÍNG
You have been reading yourself with an instrument of unknown calibration. This station is what happens when you check it against something that cannot be argued with.
By Noah WickliffeThe Still Water · Learning to Regulate
station 03 · movement one · detect
zhēng míng · calling things by their right names
guess first, then check
§I  

The Call

A scale that has never been checked still gives a number, and gives it with perfect confidence. This station is about the quiet act of holding a known weight against it.

Two stations of looking inward have produced a great deal of data and one unanswered question: is any of it accurate?

Confidence in your own readings and the correctness of those readings are separate things, and in the interoception literature they are only loosely related. An athlete can be certain, consistent, articulate — and wrong. Certainty is a feeling. It is not evidence.

The instrument you have been reading yourself with was calibrated by your childhood, and nobody has checked it since.

This is not a small problem in this population. Athletes are selected, over years, for the ability to override internal signals — to train through fatigue, to discount discomfort, to report readiness. The same trait that makes a good competitor makes an unreliable narrator of their own interior.

And for some people the calibration problem is not physiological but historical. Anyone whose account of their own experience was routinely contradicted by the adults around them learns to distrust the reading itself. For them this station is not a technical exercise. It is the return of something that was taken.

ArenaWhat this looks like there
Rowing“I felt strong” against a split that says otherwise, and the honest question of which to believe.
WeightliftingRPE 7 that was actually a 9, three sessions running.
CyclingPerceived effort drifting away from power output across a block.
SwimmingFeeling smooth and swimming slow — the most common disagreement in the sport.
Team sportsA player certain they were fine in a match the data says they faded in.
Endurance runningPacing by feel, and discovering feel has a systematic bias.
Beyond the arenaBeing sure the meeting went badly, and finding out it did not.
SensibilityHow certain you areFeels like knowledgeHigh in most athletesTrainable by confidence aloneNot the same as accuracyAccuracyWhether you are rightMeasurable against a signalModest in most athletesOnly trainable with feedbackThe variable that mattersGarfinkel's third construct — awareness — is whether your certainty tracks your accuracy at all.
Fig. 1 The interoception dissociation. Certain and wrong is a common, consequential profile, and it is invisible from the inside — which is the entire argument for an external check.
§II  

Two Failures

The two failures here are the two ways people relate to being measured.

The first is refusal. The athlete who will not look, or who looks and dismisses. Usually framed as trusting the body over the machine, which sounds like wisdom and is frequently avoidance — because a number can disagree with you, and disagreement is the point.

The second is surrender. The athlete who stops consulting themselves at all. The morning begins with the app, and the app's verdict becomes the state: told they are recovered, they feel recovered; told they are strained, they discover strain. This is the more modern failure and it destroys the exact capacity Stations I and II were building.

One athlete will not be told anything. The other will be told everything. Both have stopped learning.

There is no midpoint between them. There is an order: your reading first, written down, committed — and the instrument second, as a check on a guess you have already made. Anything else and you are not calibrating. You are being trained by a wristband.

Refusal · will not looklearns nothingSurrender · looks firstperception atrophiesSequence · predicts, then lookscalibratesThe middle row is the modern failure and the one most likely to be mistaken for diligence.
Fig. 4 Three relationships with being measured. Only the third produces calibration, and the difference between it and the second is nothing but the order of two actions.
§III  

The Line

Guess, write it down, then look.

The order is the entire method. A prediction made before the data is an act of perception. A prediction made after it is a memory. Only the first one teaches you anything, and only the first one generates the number that actually matters at this station: the size of the gap.

You are not trying to be right. You are trying to be less wrong over time, in a direction you can see.

A worked example. Before opening anything: readiness 7, sleep 6, mood 6. Written down, committed. Then the device: readiness 4, sleep 5, mood not measured. The useful figure is not the 4. It is the three-point gap on readiness, and the fact that it is the third such gap this week and always in the same direction.

The sentence that closes it is one line and no more: I run optimistic on readiness after poor sleep. That is a calibration finding. It will be true next month, it can be corrected for, and it was not available to you before you started writing predictions down.

§IV  

The Story

Confucius was asked what he would do first if given charge of a state. His answer disappointed everybody.

He said he would rectify the names.

His questioner thought this was pedantry — a country to run, and the philosopher wants to tidy the vocabulary. The reply is one of the sharpest things in the Analects: if names are not correct, language does not fit the facts; when language does not fit the facts, affairs cannot be carried to completion; and in the end the people do not know where to put hand or foot.

If the names are wrong, nothing downstream can be right — and the people do not know where to put hand or foot.

Applied to an interior rather than a state, this is exactly the calibration problem. Zhēng míng is not about honesty in the moral sense; it is about the correspondence between the word and the thing. An athlete who calls exhaustion laziness, or dread excitement, or a training-load problem a character problem, has misnamed the fact — and every decision built on that name will be subtly wrong in the same direction, for years.

The Confucian tradition pairs this with shèn dú, watchfulness in solitude, and the pairing is the point. Private accuracy is not a private virtue in this system. It is the foundation of everything public that follows.

§V  

The Science

Three findings make the case for measurement, and one makes the case for keeping it in its place.

Interoception is three things, not one. Garfinkel et al. (2015) separate accuracy (are your readings correct), sensibility (how confident you are), and awareness (whether your confidence tracks your accuracy). They dissociate. High sensibility with low accuracy — certain and wrong — is a common and consequential profile, and it is invisible without an external check.

Self-report and physiology diverge under load. In monitoring research, subjective wellness questionnaires and objective markers agree in ordinary conditions and come apart precisely when it matters: heavy blocks, competition weeks, accumulated fatigue. The divergence is not a flaw in either measure. It is data.

Prediction improves calibration. The forecasting literature is consistent: people who commit to an explicit prediction and then receive feedback improve; people who receive the same feedback without predicting first do not improve nearly as much. Guessing is not a formality. It is the mechanism.

And the limit. There is no instrument that measures an emotion. HRV, sleep, resting heart rate and coherence measure autonomic and recovery states that correlate with affect and are not identical to it. A low reading is a question — poor sleep, a virus, a hard block, a bad week, or all four — and treating it as an emotional verdict is precisely the misnaming this station exists to prevent.

Evidence status. The three-way dissociation of interoception: well replicated. Subjective–objective divergence under load: consistent across sports and monitoring systems. Calibration through prediction plus feedback: strong in judgement research, largely untested in athlete self-monitoring specifically — the transfer is reasoned rather than demonstrated. Stated as such.

ordinary trainingheavy block →Self-report wellnessObjective markersIllustrative of the divergence pattern reported across athlete-monitoring studies.
Fig. 2 Subjective and objective measures agree in ordinary conditions and come apart precisely when it matters. The divergence is not a flaw in either measure. It is the data.

A composite case. A cyclist runs the prediction column for three weeks and discovers that his estimates of readiness are almost perfectly accurate on good days and systematically optimistic on bad ones — a bias that is invisible on average and obvious in the gaps. He is not a poor perceiver. He is a good perceiver with a directional error, which is a different and far more tractable problem: he begins subtracting a fixed amount from his own estimate before hard sessions, and the divergence closes without his perception changing at all.

What calibration is not. It is worth separating this from the idea that athletes should defer to data. A well-calibrated athlete disagrees with their device regularly and knows when they are entitled to. The purpose of the check is to earn that entitlement — to establish, over weeks, the conditions under which your reading is better than the instrument's and the conditions under which it is not. Neither party wins in advance.

Why prediction is the active ingredient. Feedback alone produces surprisingly little learning, which is one of the more counter-intuitive findings in the judgement literature. A person shown their readiness score every morning for a year does not become better at estimating it. A person who commits to an estimate first, and is then shown the score, does. The mechanism appears to be error signalling: an unpredicted outcome carries no information about a model, because no model was declared. The commitment is what makes the feedback informative.

This has a direct implication for athlete monitoring as it is usually practised. Dashboards that present data without eliciting a prior estimate are not building perception; they are supplying it, and the athlete's own instrument quietly degrades from disuse across a season.

§VI  

The Traditions

RingWhat it says at this station
ConfucianThe home tradition. Zhēng míng: fix the names or nothing downstream can be right.
StoicThe dichotomy of control begins with an accurate description of the situation; a Stoic assessment that misnames the facts is worthless however calm its tone.
BuddhistYathābhūta — seeing things as they actually are, which the tradition treats as prior to, and more difficult than, any change of state.
SufiMuḥāsaba, the daily accounting. The self is explicitly held to be an unreliable narrator that requires audit rather than trust.
ZenThe koan tradition's function is to break a false certainty. Not to supply the right answer — to dismantle the confident wrong one.
TaoistChuang Tzu's caution against confusing the name with the thing: the finger pointing at the moon is a caution about instruments as much as words.
GītāViveka, discrimination — the trained ability to distinguish what is actually the case from what one prefers to be the case.

Absent, and why.

Where they disagree. There is a genuine split here between traditions that treat the self as auditable and those that treat the auditor as the problem. Confucian and Sufi practice assume a self that can be inspected by a sufficiently honest observer, and build daily accounting on that assumption. The Zen and Taoist positions are less confident that the inspector is any more reliable than the inspected, and prefer devices — koans, the teacher's eye — that break certainty from outside rather than examining it from within.

The instrument in this station belongs to the second camp, oddly enough. A number is an outside device. Its only virtue is that it does not share your assumptions, which is precisely what the koan tradition was after with a considerably less convenient technology.

§VII  

The Practice

Five moves. The first three are the method; the last two keep it from becoming a cage.

  1. Predict, in writing, before you look. Every morning: predicted readiness, predicted sleep quality, predicted mood, on whatever scale you like. Then open the app. Two columns, every day.
  2. Log the gap, not the score. The only figure worth tracking here is the difference between prediction and measure. It should shrink. If it does not, your detection is not yet stable — return to Station I for a fortnight.
  3. Find your bias. After three weeks, ask which direction you err in. Most athletes are systematically optimistic about readiness and systematically pessimistic about sleep. A known bias is worth more than an unknown accuracy.
  4. Name it correctly, once. When prediction and measure disagree sharply, write one sentence naming what you now think was actually happening. Not a story — a name. That was load, not mood.
  5. Then close the laptop. One look, one entry, done. The check happens once a day and then the instrument is out of the room. Any athlete consulting it a second time before the next morning has crossed into the second failure.

The measured interior. This is the station where the full battery earns its keep. The EPAB instruments — EIS‑32 for emotional intelligence, CPS‑32 for coping, ARI‑32 for arousal regulation — give a standing profile that a wristband cannot, and the ZSR‑48 from Station I gives you the mark on the wall. Take them now, properly, and do not take them again until Station XII.

Daily, the check is narrow: predicted versus actual on three variables. That is all. The instrument's job at this station is not to inform your training. It is to inform your perception.

A number cannot be gaslit. That is its whole value — and the reason it must be handed to the athlete rather than held over them.

For anyone whose interior was routinely overruled by someone else, a measurement that does not negotiate is not a technology. It is a form of restitution. It is also, for exactly that reason, the thing most capable of doing harm if a coach or a platform takes possession of it. The reading belongs to the person it came from. That is not a courtesy; it is the condition on which the whole method is allowed to operate.

Predictin writingCommitbefore lookingMeasureone lookName the gapone sentenceClose itdone for the dayReverse any two of these and the instrument trains you to report what the dashboard already said.
Fig. 3 The order is the method. A prediction made before the data is an act of perception; a prediction made after it is a memory, and only the first one teaches anything.

How you know it is working. Not by the gap closing to zero — it will not, and an athlete whose predictions match perfectly is either exceptional or looking first. The signs are subtler. The gaps become consistent in size, which means the noise has gone out of them. The direction becomes stable, which means you have a bias rather than a fog. And the exceptions become interesting: a day where you are wrong in an unusual direction starts to feel like information rather than failure.

At that point the instrument has done its job. You are no longer being told about yourself; you are checking a reading you already made, which is a different relationship entirely and the one this station exists to establish.

§VIII  

The Deeper Block

Three weeks, and a decision at the end of it.

Days 1–5 · Predict blind. Two columns daily. Do not analyse, do not adjust, do not try to be right. You are establishing what your untrained estimate looks like.

Days 6–10 · Find the direction. Look only at the sign of the gap, not the size. Are you consistently high or consistently low? Almost everyone is one or the other, and knowing which is most of the value.

Days 11–15 · Correct for it. Predict, then apply your known bias, then look. If you run optimistic on readiness, subtract before checking. Calibration is not perception improving; it is perception plus a correction factor.

Days 16–19 · Add the hard case. Predict on the days you least want to — after poor sleep, mid-block, the morning after a bad session. Divergence is largest exactly there, which is where the learning is.

Days 20–21 · Decide the standing rate. Choose how often you will run this from now on. Daily is for a calibration block. Weekly is enough for a season. Continuous is the second failure.

What usually goes wrong. Athletes stop predicting once they discover their bias, on the reasonable-sounding grounds that they now know it. Bias is not stable: it shifts with training phase, season, sleep and age, and a correction factor established in November is a guess by March.

The other failure is looking first, just once, because the phone was already open. It contaminates that day's entry completely and, more importantly, it establishes that the order is negotiable. Keep the device out of reach until the prediction is written; this is the one place in the ring where a physical arrangement does more work than a resolution.

§IX  

For the Coach

This station is where coaching most often goes wrong, and the error is almost always made with good intentions.

The data belongs to the athlete. If squad readiness scores are visible to staff and used in selection, you have not built a monitoring system. You have built an incentive to lie, and you will receive exactly what you have incentivised within about three weeks.

Ask for the prediction before you show the number. Every time. An athlete who learns that their guess will be heard before the device speaks is an athlete who keeps guessing. One who learns otherwise stops.

Do not diagnose from a wristband. A low HRV morning means look, not conclude. Say “something's showing up, what do you notice?” rather than “you're stressed.” The first is a check. The second is the misnaming this station is about.

Be careful whose reading you overrule. Some athletes arrive already trained by their history to defer to any external authority about their own experience. If you consistently know their interior better than they do, you are not coaching them. You are confirming the lesson that made them this way.

And watch the numbers for what they are for. Monitoring exists to protect athletes. The moment it becomes a means of comparing them, it stops doing that job and starts doing damage.

Beyond the arena. The same discipline applies to any domain where people estimate their own state and act on it: fatigue at work, readiness for a decision, how a conversation actually went. Prediction plus feedback is the only known route to calibration, and almost nobody applies it to themselves outside of a sport that keeps score.

§X  

The Turn

Detection is complete. You can read the spike, the tide, and the reliability of your own instrument.

Nothing so far has asked you to change a single state. That was the design: three stations of perception with the deliberate instruction, each time, to look and stop. If that has been uncomfortable, the discomfort is informative — it is the reflex this ring is about to work on.

Movement Two begins where the trigger lands. Not with what you feel, but with the fraction of a second between the feeling and the deed — the space that decides whether a state becomes an action you have to explain afterwards.

Hold the reading lightly. It is a description of a morning, and mornings are not verdicts.

Sources & further reading.

Traditional. Analects XIII.3, on the rectification of names. Doctrine of the Mean, on watchfulness in solitude. Al-Muḥāsibī, on muḥāsaba. Zhuangzi, on names and things.

Research. Garfinkel, S. N., et al. (2015), as above. Saw, A. E., Main, L. C., & Gastin, P. B. (2016). Monitoring the athlete training response: subjective self-reported measures. Halson, S. L. (2014). Monitoring training load to understand fatigue in athletes. Lichtenstein, S., Fischhoff, B., & Phillips, L. D. (1982). Calibration of probabilities. Buchheit, M. (2014). Monitoring training status with HR measures: do all roads lead to Rome?

Instruments referenced are SportsFlow EPAB measures; see the Research Library for scoring and administration.