← Blog
MindSeptember 29, 2026By Johnson

Screen Time Is the Wrong Number

The weekly figure counts minutes because minutes are the only thing an operating system can see. Four hours of calls with family is not four hours of a feed with no bottom, and the counter cannot tell them apart. Two better questions, and how far the evidence on screens and mental health actually goes.

The notification arrives on a Sunday morning. Four hours eleven minutes a day, down six per cent. Something about the format implies a verdict has been handed down, and nothing in it says what the four hours were.

Four hours of video calls with a parent in another country, photographs of a niece, a chapter of a book, a map, a language app and a work thread is not the same four hours as a feed with no bottom in it. The counter cannot tell those apart. It was not built to, and almost every article written about this uses the figure anyway, because a number that goes up and down is easy to put in a headline and the thing that actually varies is not.

r = 0.38

how closely self-reported screen time tracks the logged figure, across 106 comparisons

5%

share of those studies in which the self-report was an accurate reflection of the log

16%

of US adults whose only route onto the internet is a phone

Duration became the measure because duration is what a phone can see

An operating system can observe when the screen woke, which app was in front, and when it went dark. It cannot observe whether the app was opened on purpose. It cannot observe whether somebody was in the room. It cannot observe whether the reaching hand arrived before the thought did, which is the single most diagnostic thing about the whole behaviour and also the one thing no sensor registers.

So the dashboard reports minutes. Not because minutes are the important variable, but because minutes are the available one.

This is the ordinary failure mode of measurement and it is worth naming plainly, because once it is named the rest of the subject reorganises itself. What is cheap to count becomes what gets counted, what gets counted becomes what gets reported, and what gets reported becomes what everyone treats as the problem. The counter did not decide that duration was the issue. It decided that duration was loggable, and the culture downstream of it supplied the rest.

Lukoff and colleagues went looking for the thing the counter misses, using interviews, repeated in-the-moment sampling and logs of 86,402 app sessions. What predicted whether a stretch of phone use felt meaningful was not how long it lasted. Habitual use scored lower on meaningfulness than use with a specific purpose, and the uses people rated highest were productivity, looking something up, and actually talking to somebody. The same twenty minutes landed in different places depending on whether it had been decided on.

How far the evidence actually goes

This is the part where it would be convenient to overstate, so here is where it stops.

Most of the research that frightens people about screens does not use a log at all. It asks people how much they use their phone. Parry and colleagues collected every study they could find that had measured both ways and pooled 106 comparisons from 47 studies. Self-reported use and logged use correlated at r = 0.38, with a confidence interval from 0.33 to 0.42, and the self-report was an accurate reflection of the log in about five per cent of the studies. The errors did not even run in one direction. People overstated and understated in roughly equal measure, which rules out the obvious fix of adjusting everything downward.

Studies that measured both ways: did the self-report reflect the log
  • An accurate reflection5% · 5%
  • Not an accurate reflection95% · 95%
  • An accurate reflection: 5%
  • Not an accurate reflection: 95%

Note carefully what that does and does not undermine. It does not mean the figure in the settings app is wrong; that figure is a log and it is accurate about minutes. It means that a large share of the literature reporting associations between screen time and mental health has been measuring how people describe their phone use, which turns out to be a different variable with its own relationship to mood. Somebody feeling low in the week they fill in the questionnaire is not a neutral reporter of their own fortnight.

Where duration is measured properly, the associations are small. Godard and Holtzman pooled 897 effect sizes from 141 studies covering roughly 145,000 people, looking at the active and passive use distinction that was supposed to rescue the field, and found most associations negligible at absolute values below 0.10. Their own summary of fifteen years of this work is that the return on investment has been poor. Meier and Reinecke's meta-review, 34 reviews and 594 publications, lands in the same place: a small negative association overall, strongly dependent on which indicator anybody happened to pick.

Most of it is cross-sectional too, which means a single snapshot of two self-reports, from which no direction of causation can be recovered.

Now the part that does not support the argument and belongs here anyway.

Something in this literature is stable, and it is not duration. Elhai and colleagues reviewed studies using standardised problematic-use scales, the questionnaires asking whether use feels out of hand rather than how many hours it ran, and depression severity related to those scores consistently and at at least medium effect sizes. Anxiety related consistently too, at smaller ones. That is a stronger and more replicable finding than anything the hour count has produced.

There is a loop worth flagging, since problematic-use scales contain items about distress and will therefore correlate with distress by construction. But set that aside and a shape remains, and Parry's meta-analysis sharpens it from the other side: problematic-use measures correlate with the logs even less well than ordinary self-reports do. The questionnaires that predict how somebody is doing are the ones least connected to how long the phone was on.

Which is the finding, stated as plainly as it can be. What tracks distress is the experience of use being out of hand. Duration is not that experience and does not stand in for it.

The two questions

Nothing above argues for a better number. A better number is not available, and the replacement is not a metric.

Was this time chosen. And after putting the phone down, is the feeling slightly better or slightly worse.

Two sentences, answerable from memory without any instrumentation, and between them they capture most of what the hour count was reaching for and missing.

The questionWhat the counter can answer
How long was the screen onExactly, to the minute
Which app held itYes, by name
Was the time decided on, or did it arriveNo
Was somebody else in the roomNo
Did putting it down feel like relief or like lossNo
Did it displace something that had been plannedNo
Would the same minutes have been fine on a different eveningNo

Six of those seven are what people mean when they say they are worried about their phone. The counter answers the other one.

Four hours that were chosen
Four hours that happened
Opened for a reason that can still be named afterwards
Opened without a reason, and the reason was never missed
Ends when the reason ends
Ends when something outside interrupts it
The rest of the evening survives it
The evening was the thing that went
Remembered the next day
Mostly not remembered
Reported to other people without rounding down
Reported with a number that has been quietly reduced

That last row is Parry's finding arriving in domestic form. The use that gets understated is the use that already feels wrong, which is why the understatement is itself the signal.

The phone check is built around this rather than around a total. Fifteen questions read three things separately: how automatic the reach has become, how often the pull wins once the phone is open, and what it has been taking. A high reach with a low cost is an extremely common shape and is not a finding about anybody's character. It is what it looks like when the map, the bank, the camera, the radio and the people you love are one object in a pocket.

What the phone is carrying for people

Advice that opens with a detox has usually not thought about who is reading it.

Sixteen per cent of US adults own a smartphone and have no home broadband, so the phone is not one of their routes online, it is the only one. Among adults in households under thirty thousand dollars that figure is thirty four per cent, against four per cent in households above a hundred thousand. Globally the asymmetry is sharper still: the ITU's position is that in developing countries mobile broadband is the principal and frequently the only access to the digital world.

Underneath those percentages are specific people. Somebody living alone whose entire social contact arrives through a screen. Somebody managing a chronic condition whose appointments, prescriptions and results are in an app. Somebody holding a relationship together across six time zones. Somebody for whom the phone is an accessibility device and the day does not work without it. A week-long fast is written for a life that can absorb one, and recommending it universally is a failure of attention rather than a hard truth.

So the goal is not fewer hours. Nobody gets anything from a smaller total as such. The goal is that the hours were chosen, and a long total of chosen time is not a problem waiting to be solved.

The cost column is real

None of which makes the worry imaginary, and an article that ended here would have argued something it does not believe.

The clearest evidence is about other people. Dwyer and colleagues ran a field experiment in a restaurant with 304 participants, groups of friends and family randomly assigned to keep phones on the table or put them away, and the meal was less enjoyable in the phone condition. That is an experiment rather than a correlation, and what it measures is not a total. It measures where the phone was during an hour that was supposed to belong to somebody else.

Evening use and sleep is a real chain with a large literature, and it belongs to the entry on screens and sleep rather than to this piece.

Attention is the claim most often overstated, so it is worth pricing honestly. The widely circulated version is that a phone merely sitting on the desk consumes cognitive capacity. Böttger and colleagues pooled 43 effects from 22 studies and the mere-presence effect came out at g = -0.14, statistically significant overall, strongest for memory, and not significant in European or North American samples taken on their own. Real, then, and much smaller than its circulation suggests. The larger cost is the switching, and the reason switching is expensive is not the seconds lost in the switch but what stays behind on the thing just left, which the entry on attention residue covers properly.

And then the one that never shows up in any study because it is not a variable: the thing somebody keeps meaning to do, which keeps not happening, in a week with four hours a day available.

Deliberate design, and a charger that is still yours

The patterns are not accidents, and pretending otherwise makes the subject impossible to think about. Monge Roffarello, Lukoff and De Russis reviewed 43 papers and built a typology of eleven attention-capture deceptive designs, from infinite scroll outward. What the eleven share is instructive: they exploit known psychological vulnerabilities, automate the experience so decisions stop being made, cause the goal to be lost, cause the sense of time and control to be lost, and end in regret.

Read that list again next to the hour count. Those patterns are specifically engineered to break the link between time spent and time chosen. Which means duration was always going to be a poor proxy, and gets worse as the design gets better.

The conclusion people reach from there is often that none of it is their responsibility, and that is where the argument stops being useful. The design is not going to be redesigned on anybody's behalf. Where the charger goes overnight is still a decision. Both things are true at once, and the writing on this subject almost always picks one.

Weak will and helplessness are the same mistake pointed in opposite directions. Each one hands the whole problem to a single party and then has nothing left to do.

Where the number is not the question at all

One boundary, so the right readers go to the right place. If the answer is mostly about nights, if the difficulty is not the phone's place in the day but getting to bed at all, that is a different question with four quite different causes and the phone is only one of them. The bedtime check sorts between those four without scoring severity. The phone check scores how the phone sits across the whole day, including the queue, the red light and the middle of a sentence somebody was saying. Taking both is reasonable and they do not repeat each other.

Children and teenagers are a separate subject and this piece does not cover them. The evidence there is genuinely contested rather than settled, and anybody sounding completely certain in either direction is worth treating with caution.

Four hours eleven minutes, down six per cent. The only honest reading of that sentence is that it is a count of minutes, accurate, and silent on everything that was being asked.

Sources

Frequently asked questions

Is there a recommended daily screen time limit for adults?

No health authority publishes one, and the absence is not an oversight. A threshold would have to apply to a quantity that mixes a work inbox, a video call with a parent, a map, a banking app and a feed into one total, and no cut point can mean the same thing across that mixture. Guidance for young children exists because the question there is about displacing sleep, movement and play in a developing body, which is a different question with a different evidence base. For adults the honest answer is that the number has no threshold because it has no units anyone can interpret.

So the Screen Time figure in my settings is useless?

Not useless, just badly used. It is accurate about duration, and it is genuinely informative in one narrow way: compared against the same phone a month earlier, a large move in it is a real signal that something in the week has changed. What it cannot support is the comparison people actually make with it, which is against other people or against a figure they read somewhere. Treat it as a thermometer for your own trend rather than a score.

Do app blockers and time limits work?

Reasonably well at the thing they do, which is adding friction, and poorly at the thing they are sold as, which is resolve delivered by software. The limits that survive are the ones that remove a step rather than add a dialogue box, because a dialogue box asking whether to continue is a decision handed back at the worst possible moment and most people tap through it. Deleting an app and using it in a browser instead tends to outperform a timer on the same app, for no reason more sophisticated than the extra steps.

My partner says I am always on my phone and the number says four hours. Who is right?

Probably both, because the two claims are not about the same thing. Four hours spread across a day in two hundred short reaches reads as constant to somebody sitting opposite, and the same four hours in one block while alone does not. Complaints about phone use are almost never complaints about a total. They are complaints about distribution, and specifically about the reaches that happen mid conversation. That is a measurable thing in principle and the counter does not measure it.

Does work count toward it?

The counter counts it and the question of whether it should is where the whole measure falls apart. An hour of work messages answered on a sofa at nine in the evening and an hour of the same messages answered at a desk at eleven in the morning appear identically in the total, and they are not the same hour in any sense that matters. If a figure has to be used at all, the useful version is a count of time on the phone that was neither work nor chosen, which no operating system currently reports and which is the only slice most people were worried about in the first place.

Why does the number still bother me when I know it measures the wrong thing?

Because a quantity that arrives weekly and goes up and down behaves like a score whether or not it is one, and scores are hard to disregard. There is a second reason that is less comfortable. For a lot of people the figure is not frightening because of what it measures but because of what they already suspected before it arrived, and the number simply gave the suspicion a digit. That reaction is worth attending to. It is information, just not the information the counter was reporting.

See this pattern in your own numbers

✦

This article, about your situation

Luna has read the same research — and, if you let her, your results.

✦ Talk it through

More notes on people