Personal project · AI agents · Product design · Concept

Predicting what an AI task will cost, so people can make tradeoffs and still finish the work that matters

AI tools are starting to show how much usage you have left. They still don’t predict what a task will take, or help you decide what to trade off when you’re running low. I designed a task usage predictor that estimates the cost up front and helps people choose how to finish the critical work, right in the conversation.

Jump to the design
A pile-up of usage limit messages from several AI tools: rate limits, model limits, daily limits and upgrade prompts

TL;DR

THE PROBLEM

Agents spend your usage for you, across several tools. You can see your balance, but not what a task will cost, so you find out it was too much only after the task stops.

WHAT I DID

I designed a task usage predictor that helps people make tradeoffs: an estimate before you run a task, a verdict while it runs, limits for every connected tool, and options to finish the critical work instead of failing, like batching, deferring or saving a checkpoint.

WHAT I FOUND

Tools show your balance and warn you at the limit. None estimate what a task will cost, and when you’re running low, the only choices they offer are upgrade or buy more.

01 · DEFINING THE SPACE

A price you can’t see

Every request to an agent spends usage, but you don’t see the price until afterward. Tools are getting better at showing your balance. They still don’t show what a task will cost, or give you any say in how it’s spent.

I couldn’t tell what a request would cost before I sent it. Checking meant leaving the conversation for settings, and even there I only saw a balance, not a price.

When I did worry about a usage limit, I asked the agent to do less, which made its work worse, without knowing whether I needed to.

A settings modal showing session and weekly usage bars

02 · TOOL AUDIT

What exists already?

I audited six AI tools and in doing so found examples of trackers people have built for themselves.

Every tool shows a balance or warns you at the limit. None estimate what a task will cost before you run it, and none offer choices beyond upgrading.

The best in-context patterns are in developer tools. None estimate what a task will cost before you run it. This felt like a design opportunity.

How six AI tools show usage limits. Researched September 2026.
ToolWhere usage livesWarns before limit?In-context indicator?Estimates task cost?Notes
Claude (app)Settings → Usage: session and weekly bars, plus a pace forecastPace forecast in Settings; notice near the limitYes — ring in the composer near the limit (Sep 2026)NoNow shows your balance, a forecast and a ring. Suggested actions: Upgrade or Buy more.
Claude Code/usage command; limits can be piped into a customizable status lineOnly if you build itOpt-in, via scriptsNoOpen GitHub feature request asks for in-CLI usage and threshold alerts to avoid mid-task cutoffs.
ChatGPTNo dedicated usage page in the chat appNo — banner when you hit a capNoNoText chats uncapped since Aug 2026; images, Deep Research, agent and Codex still capped. You learn by hitting them.
CursorWeb dashboardIndicator appears after ~50% usedYes, once past halfwayNoLimit banner leads with an upgrade CTA. Popular tracker extension was abandoned because pricing changed too often.
GitHub CopilotCopilot icon in VS Code status bar → usage %Budget alerts at 75 / 90 / 100%One click from status barNoMoved to usage-based AI credits in June 2026. Users have asked for remaining balance directly in the editor.
GeminiProfile menu → usage viewYes — in-chat message near limitOnly near the limitNoCompute-based limits: 5-hour refresh inside a weekly cap. Warning names the reset time.

03 · HOW I FRAMED IT

From a meter to decisions inside the task

I rebuilt the Claude chat interface in Figma (with Claude) to match the real product and ideate on potential solutions with the real thing.

Mapped from my own experience: confident, then uneasy, interrupted, rationing, and blocked. The after journey ties each moment to the feature that changes it.

CURRENT STATE

Journey today: running an agentic task in Claude

Based on my own experience running a multi-step design task with Claude and the Figma connector. Evidence cards point to the screenshots I captured.

Six stages of a task today, from starting it to hitting the wall
Start the taskWork in flowStart to worryGo checkRation blindlyHit the wall
DoingAsk Claude to build a design in Figma.Agent runs many Figma calls; I watch it work.Notice how many calls it is making. No usage shown anywhere in the chat.Open Settings → Usage in a modal, several clicks away. Figma’s limit isn’t shown at all.Tell the agent to “do less pulls” so I don’t get blocked.In Cursor: “You’ve hit your usage limit” mid-task, with an upgrade button.
Thinking“This should be quick.”“Great, it’s moving.”“Am I close to a limit? Which one?”“38%… so was I worried for nothing?”“I’ll just ask for less and hope.”“I was in the middle of something.”
Pain point—Cost is invisible while work happens.No way to tell which tool will run out first.Checking breaks flow; data is session-wide, not task-specific.Rationing without information makes the agent do worse work.Work stops mid-task. I’ll have to remember what I was thinking later.
Evidence——My chat: “I don’t want to get blocked by the limit.”Settings → Usage modal screenshot.My chat: “Try to optimize so we are doing less pulls.”Cursor limit banner.
Feeling ConfidentFocusedUneasyInterruptedRationingBlocked

04 · THE DESIGN

Four moments

Each moment answers a question people have mid-task: what will this cost, am I on track, which tool will run out, and how do I still finish?

See the cost before you ask

Before a heavy request, the agent estimates what it will use as a range, across itself and every tool it will call. Run it, or ask for a lighter version. It never blocks you.

Replaces: guessing, and rationing requests just in case.

Pre-flight card estimating a task will use 8 to 12 percent of the session and about 20 Figma calls

05 · REFLECTION

Reflections and what’s next

Since I started, Claude added a usage ring and a pace forecast in settings, which validated the direction. What’s still missing is visibility inside the task and choices beyond buying more. Next, I’d test whether task-level estimates earn trust, then build this into an extension to test usability.

Show the answer, not the math

A verdict like "Enough to finish" answers the real question, with the numbers one step away.


Validate: How accurate do estimates need to be before people trust them?

Quiet beats constant

Color and prompts only when they're needed kept the meter from becoming a source of stress.


Validate: A/B test an always-on ring against one that appears past a threshold.

Design across the chain

With agents, the limit that stops you often isn't the one you're watching.


Validate: Do people take the cheaper paths or push on anyway with a higher-cost command?

Visibility isn’t agency

Showing a balance helps, but people need decisions, not just numbers. The strongest moments give a choice, not a warning.


Validate: Do people finish more tasks when offered choices than when they’re only warned?

Next project

Benefit Management Dashboard Redesign

View next case study