HOME / THE GUIDE / READINESS
HEALTH DATA

Readiness score: what it actually measures

PUBLISHED 2026-09-11 · OPERATÖR 45
Short answer: The score is not a measurement. It is a weighted blend of a few numbers from the night that has just ended, usually heart rate variability, resting heart rate, sleep and yesterday's activity. A review of 14 such scores across 10 manufacturers found that not one of them discloses its formula, and that few offer peer-reviewed evidence for the number. So read it as a summary of yesterday, not a verdict on today. And always factor in the measure with the strongest research behind it for tracking training response: your own answers to the same three questions every morning.

It is five to six in the morning. You reach for the phone before your feet hit the floor, and there it is. 41 out of 100. Red ring. "Take it easy today."

So now you are sitting on the edge of the bed negotiating with a number. Should you cancel the session? Your body feels fine, honestly. But surely the number knows something you do not?

It does not. It knows less than you think, and it knows it in a way worth understanding, because that understanding is the difference between a tool and a boss.

What the score is made of

In April 2025, the journal Translational Exercise Biomedicine published a review with an unusually plain purpose. The researchers wanted to know what these scores are actually assembled from. They worked through everything the manufacturers had made public: technical white papers, user manuals, the explanations inside the apps themselves, and whatever research existed.

They found 14 composite health scores across 10 of the largest manufacturers.

ManufacturerName of the score
OuraReadiness, Resilience
WhoopStrain, Recovery, Stress Monitor
GarminBody Battery, Training Readiness
FitbitDaily Readiness
PolarNightly Recharge
SamsungEnergy Score
SuuntoBody Resources
UltrahumanDynamic Recovery
CorosDaily Stress
WithingsHealth Improvement Score

Then they counted the ingredients. Four measurements turned up more often than anything else.

IngredientShare of the scores
Heart rate variability86 %
Resting heart rate79 %
Physical activity71 %
Sleep duration71 %

Read that table twice, because it gives something away. There is no secret sensor. There is no measurement of "recovery" as an organ or a state. The score is built from numbers you can already see individually on the same screen.

What separates the manufacturers is not what they measure. It is how they weight and combine those measurements. And that is where the review lands its uncomfortable sentence: none of the ten disclosed their exact algorithmic formula, and few provided empirical validation or peer-reviewed evidence supporting the accuracy or clinical relevance of their scores.

What that claim is and is not: it is a verdict on the evidence base, not proof that the number is wrong. The ingredients are well studied individually. What is missing is independent testing of the blending itself. The difference between "not shown to work" and "shown not to work" is the whole difference, and it is worth holding onto in both directions.

Why two devices disagree about one night

The most common frustration with these scores is not that they exist. It is that they do not agree. The ring says 78, the watch says 44, and you only slept one night.

Two things explain that, and they act at the same time.

The first is the raw material. In August 2025, Physiological Reports published a validation study in which thirteen healthy adults, six of them women, wore an ECG reference and several consumer devices at the same time while they slept. That produced 536 nights. Each device's nocturnal resting heart rate and heart rate variability were then compared against the reference.

DeviceError, resting HRError, variability
Oura Gen 41.94 %5.96 %
Oura Gen 31.67 %7.15 %
Whoop 4.03.00 %8.17 %
Garmin Fenix 6Not analysed10.52 %
Polar Grit X Pro2.71 %16.32 %

The figures are mean absolute percentage error against ECG. Lower is better. Garmin was excluded from the resting heart rate analysis because of methodological inconsistencies in how the device reports the value, which is itself a reminder that two devices do not always measure the same thing under the same name.

Look at the right-hand column. The best sat at roughly 6 percent error, the worst at just over 16. The error at the bottom of the list is almost three times the error at the top, measured against the same ECG on the same nights. Which device you happen to wear is therefore part of the answer to what your number turns out to be.

The second explanation sits on top of the first. Even if two devices had measured identically, they would still produce different scores, because they weight the ingredients differently and draw them from different time windows. The blending is proprietary at all ten companies. Two numbers built from different raw material and different recipes cannot possibly land in the same place.

Read the study with the right expectations: thirteen participants is a small sample, all of them were healthy adults, and each result applies to the device generation that was tested. The ranking is not a verdict for all time. What holds is the general lesson: the margin of error varies a lot between devices, and it is large enough to show up in your score.

If you want to understand why the wrist is a harder place to measure than the chest, that whole question is covered in the guide to chest strap versus wrist heart rate.

The score looks backwards

This is where the real misconception sits, and it is built into the word. "Readiness" sounds like a forecast. A verdict on what you can handle today.

But go back to the ingredients. Heart rate variability from the night that just ended. Resting heart rate from the night that just ended. Sleep duration from the night that just ended. Activity from the day before. Every single ingredient is history.

So the score is a summary, not a prediction. It says roughly this: this night looked different from your recent weeks, in this direction, by about this much. That is genuinely useful. It is just not the same thing as knowing how your session will go.

The gap is most obvious when the low number has an obvious cause. You ate late. You had a glass of wine. You ran intervals at nine in the evening. The score is then telling you exactly what you already knew, with two decimal places of confidence. What a late session does to your night is covered in the guide to exercise and sleep.

There is one exception worth taking seriously. When the number is low with no cause you can point to, and it stays there for several mornings in a row, it has told you something you did not know. That is the only pattern in this kind of data that earns the right to change your week.

The cheapest instrument has the best evidence

Now for the finding that turns the whole question around, and that no manufacturer has any interest in advertising.

In 2016 the British Journal of Sports Medicine published a systematic review of how athlete well-being is best tracked over time. The researchers looked for studies that measured objective markers, meaning heart rate, blood markers and performance tests, alongside subjective ones, meaning what the person reported about mood, fatigue and perceived stress. Fifty-six original studies met the criteria.

Two results came out of it.

The first was that the objective and the subjective measures generally did not correlate with each other. They simply were not measuring the same thing.

The second was that the subjective measures reflected both acute and chronic training load with superior sensitivity and consistency compared with the objective ones. Subjective well-being typically got worse with an acute increase in training load and with chronic training, and improved when the load was reduced.

In other words: the question "how does it feel today?" tracked training load better than the instruments did.

Two caveats that belong here: the evidence concerns athletes, not forty-somethings with three children and a commute. And the subjective measures were not gut feeling but structured questionnaires, meaning the same questions asked the same way at the same time of day. That is what makes the answer comparable. A vague sense of things at six in the morning is not the same instrument.

Build your own readiness in three lines

The conclusion from the two sections above is not to bin the watch. It is to put the number where it belongs: as one witness among several, not as the judge.

Ask three questions every morning, before you look at the phone. Rate each answer from 1 to 5, where 5 is best.

  1. How did I sleep? Your own experience, not the app's sleep score.
  2. How is my energy right now? Before the coffee, not after.
  3. How heavy does my body feel? Your legs on the stairs say more than you would think.

The total lands between 3 and 15. After a couple of weeks you have a normal range, and only then does the number mean anything. Exactly as with individual metrics, the information is carried by the deviation from your own baseline, never by the absolute figure. How to establish a baseline in practice is covered in the guide to HRV.

Then let the watch score be the fourth voice. Here is how to read the outcome.

SituationWhat it probably isWhat you do
Low score, you feel fineMeasurement noise or a known causeTrain as planned
High score, you feel wreckedYour body knows more than the sensorLower the ambition, keep the session
Both low, one dayOne bad nightEasier session, earlier bedtime
Both low, three daysAccumulated load or a brewing infectionEasy week, nothing hard

Notice what is missing from the right-hand column. The word "rest" does not appear once on the single-day rows. A low morning is almost never a reason to skip a session. It is a reason to make the session easier, which is a different thing and a far smaller price. How often the body actually needs a full day off is covered in the guide to rest days between workouts.

Row two is the most important row in the table. When the number says 88 and you can barely tie your shoes, trust yourself. The 2016 review points in exactly that direction.

After 40: what actually changes

Two things make this kind of data trickier in midlife, and neither of them means your body is broken.

The first is that your week contains more disturbances that are not training. A child who wakes up. A deadline. An evening that ran late for reasons you did not choose. All of it lands in the same score as yesterday's run, and the score cannot tell them apart. A low number after a night with a sick child is not a training verdict.

The second follows from the first. The more disturbances your week contains, the more noise there is in the series, so you need more days before a pattern can be told apart from chance. That argues for patience with the numbers, and against reading every morning as a fresh verdict.

The practical consequence is simple. Look at the week, not the day. Four of seven mornings below your normal range means something. A single Tuesday rarely does. Your watch's VO2 max figure needs even longer than that: months rather than weeks, as the guide to what VO2 max is explains.

When it is not about training

There is a point where this page stops being a training page, and it is worth saying out loud.

Tiredness that does not lift is not a question for your watch. No readiness score can make that assessment, and none of them are designed to.

When to seek care. The NHS advises that you see a GP if you have been feeling tired for a few weeks and you are not sure why, if your tiredness affects your daily life, or if you feel tired and have other symptoms such as weight loss or mood changes, or you have been told you make gasping, snorting or choking noises while asleep. In Sweden, contact a vårdcentral or call 1177 for health advice if you are unsure where to turn. This page is general health information, not medical advice.

One more thing, since it sits close by: persistent daytime sleepiness despite enough time in bed has a common cause that no score will pick up. That one is covered in the guide to snoring and sleep apnoea.

What this means on a Monday morning

The summary fits in four lines.

The score is not a measurement but a blend of four numbers you already have, and the recipe is proprietary at every manufacturer. It looks backwards, at the night that has ended, not forwards at the day you still have. The margin of error in the raw material is enough to make two devices disagree about the same night. And the instrument with the best research behind it is free: the same three questions, every morning, written down.

Use the number as one witness among several. Never let it be the one that decides. A missed session is data, not a failure, and a red ring at five to six in the morning is not a verdict on you either.

Frequently asked questions

What does a readiness score actually mean?

It is a composite number your watch or ring calculates from a handful of measurements taken during the night that has just ended. A review published in Translational Exercise Biomedicine in 2025 catalogued 14 such scores across 10 manufacturers. The most common ingredients were heart rate variability in 86 percent of the scores, resting heart rate in 79 percent, physical activity in 71 percent and sleep duration in 71 percent. So the score is a summary of what has already happened, not a measurement of some new organ in your body. In practice it says: this night looked different from your own recent weeks, in this direction, by roughly this much.

Why do my watch and my ring show different readiness on the same morning?

For two reasons that act at the same time. The raw ingredients differ, and the recipes differ. In a validation study published in Physiological Reports in 2025, thirteen adults wore an ECG reference alongside several consumer devices while they slept, across 536 nights. For heart rate variability, the mean absolute percentage error ranged from 5.96 percent for the most accurate device to 16.32 percent for the least. On top of that, each manufacturer applies its own weighting, and none of them publish the formula. Two numbers built from different raw material and different weights cannot possibly land in the same place.

Should I skip training when my readiness is low?

Rarely entirely, often a little. A single low morning after a late session, a glass of wine or a broken night is expected, and it is not a stop sign. What is worth acting on is a number that stays low for several days with no cause you can point to, or a low number that matches how you actually feel. In practice that usually means keeping the session but lowering the ambition: swap the intervals for an easy run, drop a set, shorten the session. Removing training altogether is rarely the right answer to a number on its own.

Are readiness scores scientifically validated?

Not in the way the word suggests. The 2025 review found that none of the ten manufacturers disclosed their exact algorithmic formula, and that few provided empirical validation or peer-reviewed evidence supporting the accuracy or clinical relevance of their scores. That is a verdict on the evidence base, not proof that the number is wrong. The individual ingredients, especially resting heart rate and heart rate variability, are well studied on their own. What is missing is independent testing of the way they are combined.

What works better than looking at the score?

Asking yourself the same questions every morning and writing the answers down. A systematic review in the British Journal of Sports Medicine in 2016 examined 56 studies that measured subjective and objective markers of athlete well-being at the same time. The two kinds of measure generally did not correlate, and the subjective ones reflected both acute and chronic training load with superior sensitivity and consistency. The evidence comes from athletes rather than from forty-somethings with three children, and the subjective measures were structured questionnaires asked the same way every time, not a vague feeling. The point still stands: the cheapest instrument you own has strong support behind it.

Want someone to read the numbers for you?

Operatör 45 builds a personal coach out of your answers: one that weighs up your sleep, your week and your training and turns them into a verdict in plain language, and knows the difference between a bad night and a pattern. 7 days free, then 99 kr/month or 799 kr/year.

Start free →