Everybody has a number in their head. Put yours down before you read on, then we will show you the one famous experiment about this, the caveat that usually gets left out, and what it means for your own list.
You still scroll it. You just stop choosing from it, and the switch happens without any announcement.
Straight answer first: there is no fixed number, nobody has measured one for watchlists, and anyone who gives you a confident figure is guessing. The useful limit is not the size of your whole archive at all. It is how many things you compare at the moment you are choosing, and that is a number you control without deleting anything. The evidence around this is genuinely mixed. One famous supermarket experiment about jam found people bought far more often from six options than from twenty-four, and a much less famous follow-up found that pooling fifty similar experiments gave an average effect near zero. Long watchlists do have several features that can make choosing harder, which we go through below. Put your number in first, then we will get to it.
How many saved items is the point where a list stops being useful to you?
No looking it up, no thinking too hard. The number that came into your head when you read the headline.
Share of people who bought jam, by how many jars were on the tasting table.
Iyengar and Lepper's supermarket study. Six options, thirty per cent bought. Twenty-four options, three per cent bought.
One more question, and this is the one that actually decides whether you clean or cut. How many different kinds of evening can you name for the things on your list? Short and funny, needs a clear head, watching with somebody, background while cooking, that sort of thing.
A supermarket, a tasting table, and jam. Some days the table held six jars, other days twenty-four. The big table drew more people over, and then about 30% of the people who stopped at the small table bought a jar, against about 3% at the big one. That gap is why this one experiment has been quoted in roughly every article about choice for twenty years, including this one.
The part that usually gets left out: when researchers later pooled around fifty experiments on choice overload, the average effect came out close to zero. Too many options cause problems in some situations, but not in every situation. More options hurt when the options are similar, when you have no expertise to lean on, and when you do not know in advance what you want. More options help when you know exactly what you are after, because a bigger catalogue then means a better chance of having it.
The jam study is Iyengar and Lepper, summarised at digitalwellbeing.org. The meta-analysis picture is summarised at atticusli.com. We are flagging both because it would be easy to write this page as though a shorter list were scientifically proven to be better, and that would turn a real finding into a law it is not.
Nobody has tested a watchlist the way they tested the jam. What we can do is run a list against those three conditions, and a watchlist ticks all three, which at least tells you what kind of situation you are in.
Everything on your list passed the same test: it looked worth watching. So there is no obvious loser to eliminate, which is the shape that makes choosing slow.
You are not looking for a specific thing. You are looking for something that suits an evening, and the list holds no information about evenings at all.
This is the one the research on jam tables did not have to deal with. Every hard condition lands at the exact moment you have the least capacity for it.
So the useful question is not how many items you have. It is how many you compare at the moment you decide. Those are different numbers, and only the second one hurts.
If you only change one thing, note down what kind of evening each thing is for as you save it. That single word is what turns two hundred comparisons into about twenty, without deleting anything.
You can shorten a list by deleting items, or make it easier to use by dividing it into smaller groups. People reach for deleting first because it feels virtuous, and it is almost always the weaker move.
Think about what actually happens on a Tuesday. You open the list and start comparing everything against everything, because there is nothing telling you which part of it applies to tonight. Deleting twenty items leaves you comparing 180 things instead of 200, which changes nothing you can feel. Splitting the same list into four labelled parts leaves you comparing about fifty, and you feel that immediately.
| Size of the group you are choosing from | What to do |
|---|---|
| Under about 20 | Nothing. It still reads in one go, and organising it is effort you will not get back |
| About 20 to 50 | Deleting still helps. This is the band where removing the obvious never-going-to-happen items shortens the scroll |
| Over about 50 | Stop deleting and start dividing. Split by the kind of evening, not by genre |
Those three numbers are a rule of thumb, not a finding. They come from what fits on a screen without scrolling twice, nothing more. No study supports 20 or 50, and we would rather label them than dress them up.
Why genre is the weaker split. Say you have forty things filed under drama. On a Tuesday you are not looking for a drama, you are looking for something you can follow while half asleep, and that forty contains both a four-hour Hungarian film and a comfortable crime series. The genre label did not remove a single option for you. Split the same list by evening and the tired group has maybe eight things in it, all of which fit the night you are actually having.
What that looks like in practice, on a real-sized list. You are not deleting anything here, only labelling.
| Group | What goes in it | Roughly how many |
|---|---|---|
| Tired | Familiar, short, nothing to keep track of. Half-seen series count double here | 35 |
| Focused | The demanding ones. Subtitles, long runtimes, the masterpieces you keep deferring | 45 |
| With family | Anything everybody can sit through, which is a much smaller set than you think | 25 |
| Under 30 minutes | Comedy specials, single episodes, the filler for when there is half an evening | 30 |
| Not sure yet | Everything you could not label in two seconds. It is allowed to be the biggest pile | 65 |
What changed on Tuesday. You are tired, you have about ninety minutes, and you are on your own. That is one group, 35 items, and if you also apply the ninety minutes it lands nearer 12. You are choosing between roughly a dozen things you already know suit tonight, instead of comparing two hundred. Nothing was deleted, and the archive is exactly as big as it was this morning.
The “not sure yet” pile is the honest part. A third of the list will resist labelling, and that is fine. Those are usually things saved on somebody else’s enthusiasm. Leave them. They get labelled naturally the next time you go looking, or they quietly reveal themselves as things to remove.
Put it in both, if your list allows it. This is labelling, not filing: nothing is being moved out of anywhere, so there is no cost to a film appearing in tired and in under 30 minutes. If your tool only allows one label, pick the evening you are most likely to be having.
Yes, and it is the one bit of deleting worth doing on sight. A watched item is pure noise: it passes every filter and it can never be the answer. Clear those whenever you notice them.
Almost never. Groups describe your evenings, and your evenings change with your life rather than with the seasons. Label new saves as they arrive and the groups look after themselves. If a group has grown past about fifty, split it rather than rebuilding everything.
No, and this is the conclusion we would push back on hardest. Saving is cheap, useful, and it is how you catch things you would otherwise forget. The problem was never the saving. It was that everything gets dumped into one undifferentiated pile with no way back out.
Then you are fine. Genuinely. The people this matters to are the ones who scroll a long list and put on something they have already seen. If that is not happening to you, your system is working and you should ignore all of this.
It solves a different problem. Recommendations answer what is popular and similar to what you watched. Your saved list is the record of what you decided you wanted, which is a much rarer and more personal thing, and no recommender is trying to serve it back to you.
The conclusion here is deliberately not save less. It is save as much as you like, and make sure something can hand the right part of it back at the right moment.
That is the whole point of dEssence. You save the thing with the reason attached, then ask Tracy the way you would say it out loud, along the lines of what did I save that is short and funny, and you get a handful back rather than the whole pile. It lets you search a large collection with a specific question instead of scrolling it. It does not connect to your streaming services, it will not tidy anything on its own, and it has no opinion about how long your list should be.
Two hundred items you can ask a question of beats forty you have to scroll through.
Nobody has measured this for watchlists. There is no study that finds the item count where a saved list stops working, and we did not find anyone credible claiming one. The tool above compares your number with a jam experiment because that is the nearest real evidence, not because a jam table is a watchlist.
The jam figures are real. Six options, about 30% bought. Twenty-four options, about 3%. Iyengar and Lepper's supermarket study, summarised at digitalwellbeing.org. We have drawn only those two measured points, and we have deliberately not drawn a curve between them, because the shape in between is not something anybody measured.
The caveat is real too. A meta-analysis pooling around fifty choice-overload experiments found an average effect near zero, meaning the effect depends heavily on the situation. Summary at atticusli.com. Any article quoting the jam study without this is telling you half of it.
The bands and the verdict are ours. Twenty, fifty per group, twelve per kind of evening: all editorial rules of thumb about what fits on a screen, not research findings. Treat them as a starting point to argue with.
Nothing leaves the page. Your numbers stay in your browser, there is no account and no upload, and closing the tab throws it away. We do not collect answers, which is also why there is no other-readers average anywhere on this page.