Most productivity apps fail their users for a reason that has nothing to do with features and everything to do with a category-wide mistake: they help you capture and organise work, and then leave the hardest part, deciding what to actually do, entirely to you. They are excellent at making lists and useless at making choices, and since the choice is the thing you were struggling with, the app quietly becomes one more surface to maintain rather than a tool that moves you. You end up with a beautifully organised backlog and the same paralysis you started with.

This is not a knock on any single tool. Todoist, Things, Notion, the Obsidian task plugins, all of them are well built and many people get real value from them. The failure is structural and shared: the industry solved capture and presentation years ago and then kept polishing those, because they are the parts that demo well, while the deciding problem sat untouched because it is genuinely hard and does not screenshot nicely.

The list is not the problem you have

Walk through what a productivity app actually does. It gives you a fast way to write tasks down, some fields to tag and date them, and flexible views to slice and sort them. Every one of those is capture and presentation. Not one of them decides. Sorting by due date or priority is still not a decision; it is a rearrangement you then have to read and judge, item by item, to figure out where to start.

For a short list this is fine, because reading five items and picking one is trivial. The failure appears at scale, and it appears reliably. Once the backlog holds more items than you can hold in your head, the daily act of scanning it to choose becomes the exact friction the app was supposed to remove. You did not want a place to store two hundred tasks. You wanted to know which three matter today, and a store cannot tell you that.

The engagement trap makes it worse

There is a second failure layered on the first, and it comes from how many of these products are built to be measured. An app judged by daily active users and session length has a quiet incentive to keep you in the app: more notifications, more streaks, more badges, more reasons to open it and poke around. None of that helps you finish work. It helps the app’s metrics, which is not the same thing and is sometimes the opposite.

The cost lands on your attention. Field research on knowledge work found people already switch tasks on average every three minutes, and the American Psychological Association notes that the mental cost of switching between tasks can reach 40% of productive time. An app that pings you to preserve its engagement numbers is manufacturing exactly the interruptions that research says are most expensive. A tool that respected your attention would try to get you out of itself and into your work as fast as possible, which is close to the opposite of what an engagement-optimised product is rewarded for doing.

Why the deciding problem stays unsolved

If deciding is the real gap, why has the category not closed it? Because deciding well is hard in a way capture is not. To rank your tasks properly, a tool needs to weigh signals most apps never collect: not just due dates and priorities, but which tasks sit in work you are actively pushing on, which have gone stale, which are blocked, which keep resurfacing in your notes. Most productivity apps are islands that hold only the task text and a couple of fields, so they simply lack the raw material to decide, and they retreat to letting you sort what little they have.

The apps that could decide are the ones that already sit on top of your real work. This is the quiet argument for keeping tasks where their context lives, and it is a large part of why an Obsidian task management setup can go further than a standalone app: the tasks are surrounded by the notes, projects and links that reveal what actually matters. The raw material for a real decision is already there; it just needs a layer that reads it.

What a tool that actually decides looks like

A productivity tool that does not fail at scale has to cross the line from presenting to deciding, and doing that credibly has requirements. It has to read more than the task text, pulling in the surrounding signals, recency, link centrality, salience, blockers, that separate what matters from what is merely written down; the case for using those structural signals is laid out in what makes a note important. It has to produce a genuine ranking, not another sort, so the output is a short answer rather than a long list to re-judge. And it has to explain that ranking, because a decision you cannot inspect is one you will abandon the first time it errs, a reaction documented in research on algorithm aversion and the argument in why your task app should explain itself.

What most apps optimiseWhat a deciding tool optimises
capture speed and pretty viewsthe quality of the daily decision
time spent in the apptime spent on the actual work
how much you can storehow little you have to re-judge
sorting the list you gave itranking on signals you did not have to enter

That is the shape of a tool that respects both your attention and your judgement: it does the deciding you were stuck on, shows its reasoning so you stay in control, and then gets out of your way. ZPXE is one attempt at exactly this for Obsidian vaults, a deciding layer on top of the tasks you already keep, but the point is bigger than any product. The category’s failure is the missing decision, and any tool that genuinely closes that gap, wherever it comes from, is doing the thing the others only promised.

The graveyard of abandoned setups

If you want evidence the failure is real, look at your own history with these tools. The pattern is almost universal among people who take productivity seriously: a burst of enthusiasm for a new app, a weekend spent building the perfect system, a few good weeks, and then a slow drift back to a scratch list or nothing at all. The app did not break. Nothing crashed. It simply stopped being where the deciding happened, because it never did the deciding in the first place, and maintaining an elaborate store you no longer consult is pure cost.

This cycle is expensive in a way that is easy to miss, because each individual migration feels productive. Setting up the new tool is satisfying, the fresh start feels like progress, and the honeymoon is real. But the underlying problem, that you still have to rank your work by eye every morning, travels with you from app to app untouched, so the churn produces motion without resolution. The tools that break this cycle are not the ones with more capture features or prettier views; they are the ones that change what happens after capture, by making the decision for you and showing their reasoning. Recognising that the gap is the decision, not the storage, is what finally lets you stop app-hopping, because it changes what you are shopping for.

SymptomRoot causeWhat actually fixes it
A tidy backlog and no idea where to startthe app sorts, it never decidesa ranking that weighs many signals
Notifications that pull you back inengagement optimised over attentiona tool that gets you out and into the work
Endless app-hoppingyou keep buying storage, not decidingchoosing for the decision, not the features
The system rots after a few weeksit never did the hard part for youthe deciding done for you, transparently

When the standard apps are not failing you

It would be unfair to imply these tools fail everyone. If your task life is small, a plain list app is not failing you at all; it is doing exactly the right amount, and adding a deciding layer would be overhead. If your work is mostly externally scheduled, meetings, shifts, deadlines someone else sets, then a calendar and a simple list genuinely cover it, because the deciding was done for you. The structural failure bites specifically when you have more self-directed work than you can rank by eye, which is the situation of anyone with a large, self-managed backlog and no boss handing them the next task.

Key takeaways: why productivity apps fail and what to want instead

Productivity apps fail because they perfect capture and presentation and leave deciding, the part you were actually stuck on, entirely to you, then compound it by optimising for engagement over your attention. The failure is structural, not a flaw in any one product, and it appears reliably once your backlog outgrows what you can rank in your head. Want a tool that crosses from presenting to deciding: one that reads the signals around a task, produces a genuine ranking rather than another sortable list, explains that ranking so you stay in control, and then gets out of your way. Stop grading task apps on how much they let you store and start grading them on whether they make tomorrow’s first move obvious.

Quick answers

Why do most productivity apps fail to make you more productive?

Because they solve capture and presentation and leave deciding to you. They give you fast ways to write, tag and sort tasks, but sorting is not deciding, and once your backlog is larger than you can hold in your head, scanning it to choose becomes the exact friction the app was meant to remove. Many also optimise for engagement, pinging you in ways that fragment the attention you needed to finish work. The missing piece is the decision.

Are apps like Todoist and Notion bad?

No, they are well built and genuinely useful, especially for smaller or externally scheduled workloads. The failure is structural and shared across the category, not a flaw in one product: these tools are excellent stores and presenters and were never designed to rank your work by the signals that actually determine what matters. They fail specifically when you have more self-directed tasks than you can prioritise by eye, because that is the job they do not do.

What should a productivity tool do differently?

Cross from presenting to deciding. It should read more than the task text, pulling in recency, link context, salience and blockers, produce a real ranking rather than another sortable list, explain that ranking so you can inspect and correct it, and then get out of your way instead of competing for your attention. The test is simple: does it make tomorrow’s first move obvious, or hand you a tidy list and the same decision you were already stuck on?

When is a simple task app enough?

When your workload is small or mostly set by others. If you keep a short list, reading it and picking a task is trivial and a deciding layer is pure overhead. If your day is driven by meetings, shifts and externally imposed deadlines, a calendar and a plain list already cover you, because the prioritising was done for you. The structural failure only bites when you have a large, self-directed backlog and no one handing you the next thing to do.

Does adding AI fix the deciding problem?

Not on its own, and sometimes it makes trust worse. A language model can produce a confident ranking, but if you cannot inspect or durably correct it, you will abandon it the first time it errs, which research on automated judgement says happens fast. Fixing the deciding problem is less about adding intelligence and more about reading the right signals and showing your reasoning. A transparent, deterministic ranking you can audit beats a fluent one you have to trust blindly.