Pattern 22 · Cost & limits

Rate limits & quotas

Provider limits and plan quotas hit every AI product — most show a generic error and a dead feature.

By Aleksey StepikinUpdated October 20263 min readLive demo
Live demo · try it

An interactive mock built in plain HTML, CSS and JavaScript. Data is fictional; no model is called.

FitWhen to use it — and when not

Use it when

  • Plan-based message or credit limits
  • Provider rate limits (429s) at peak load
  • Premium models with tighter caps

Skip it when

  • Hiding limits entirely until they hit — that is the anti-pattern this fixes

AnatomyThe parts of the pattern

  1. MeterRemaining usage, visible before it runs out.
  2. Early warningA heads-up at ~20% left.
  3. Limit stateWhat happened, when it resets (live timer).
  4. OptionsFallback model, queue for later, upgrade — with trade-offs.

GuidelinesDo & don’t

Do

  • Offer a fallback that keeps the user working.
  • Show the exact reset time.
  • Distinguish "you hit your plan limit" from "we are overloaded".

Don’t

  • Show a raw 429 or "Too many requests".
  • Lead with upgrade as the only option.
  • Reset counters at a time you do not show.

In productionHow it looks in a shipped product

AI quota and rate-limit screen: provider throttling, agents routing to a fallback model, three remediation options
In production — Atlas quota state: a provider is throttling, agents auto-route to a fallback model, with three explicit options. See the Atlas case

In the wildReal-world examples

ChatGPTClaudeCursorPerplexity

Products named for reference only — no affiliation, and the demo above is an original illustration, not a copy of their UI.

For engineersImplementation notes

  • Route through a gateway that knows quotas per user and provider limits; return a typed limit state with reset_at.
  • Implement model fallback chains server-side and tag responses with the model actually used.
  • Queue deferred requests with idempotency keys so "run later" is safe.

Cost & limitsRelated patterns

All 26 LLM UX patterns

Building an AI product?

I design and ship AI products end to end — LLM interfaces, agents, RAG, billing — from concept to a live product in weeks, not quarters. Tell me what you are building and get a fixed estimate.

Get an estimateBook a call

Create bold.
Deliver better.

See our workGet in touch