AI UX PlaygroundNewsletterJoin 2K+ AI designers and PMs on Substack. New teardowns, patterns, and prompts as they drop.

Performance

Rate Limit Warnings

Warn before someone hits quota, RPM, or plan ceilings, and say what happens next: wait, upgrade, or switch model. A hard stop mid-task with no warning is a product failure.

Interactive demo

Approaching limit

5 requests left

API requests45 / 50
Pace your requests

Slow down now to avoid hitting the hard limit.

5:00

Until limit resets

Overview

How might we design rate limit warnings so people can trust and act on AI output?

When to use

  • Essential for applications using AI APIs with rate limits, developer tools, and platforms where proactive limit management prevents service interruptions.

When to skip

  • Truly unlimited internal deployments where limits never apply.
  • Background batch jobs better served by queues than interactive warnings.
  • When limits are so high that constant nags train users to ignore them.

Rules

  • Failing only after submit with a cryptic HTTP 429 and no retry guidance.

  • Warnings that push upgrade without showing remaining allowance.

  • Inconsistent units between the warning and the billing page.

  • Blocking the composer with no estimate of when capacity returns.

Evidence

ProductImplementation
ChatGPT / Claude ProUsage and limit messaging when approaching plan caps.
CursorPremium model and agent usage warnings against plan quotas.
API consoles (OpenAI, Anthropic)Rate limit headers and dashboard alerts for RPM/TPM.
PerplexityPro/feature gates when free limits are exhausted.

FAQ

What are rate limit warnings in AI UX?

Rate limit warnings tell users they are nearing API or plan quotas before a hard failure, and explain wait, upgrade, or model-switch options.

When should a rate limit warning appear?

Show a soft warning as the user approaches the ceiling (for example under 20% remaining) and a hard, actionable state at the limit, not only a raw error after send.

How do rate limits relate to cost transparency?

Cost transparency shows price per action. Rate limits show throughput or plan caps. Users need both when agents can burn quota quickly.