Rate limits, compared honestly
Almost no pricing comparison site publishes this, even though it's often what actually decides whether a production launch can handle traffic. Researched directly from each provider's own current docs — including the providers that don't publish real numbers at all.
As of Aug 25, 2026. This is hand-researched, not part of the automated daily pipeline — re-verify against the source link before relying on a specific number.
How to read this
Rate limits are not a level playing field across providers, and this page tries to show that honestly rather than force every provider into the same table shape. Five providers (OpenAI, Anthropic, Google, xAI, Moonshot) publish real, spend-gated tier ladders — start on a low tier and it rises automatically as you spend or build usage history. Qwen is different: its limits are fixed per model and identical for every account regardless of spend, with increases handled as one-off support requests. DeepSeek publishes no RPM or TPM ceiling at all — only a per-model concurrency cap. And Mistral and Z.AI both fall short of a fully public numeric table: Mistral's tier thresholds are documented but the exact per-tier numbers live behind a login, and Z.AI publishes no numeric limits anywhere in its official docs.
Self-checking your actual limit
Only Anthropic exposes a dedicated API to read your account's configured rate limits directly. OpenAI is the next-best option — every API response carries your current remaining-requests and remaining-tokens in its headers, so one cheap authenticated call tells you where you stand. Every other provider on this list requires checking a web console.
Why we don't auto-update this page
Unlike pricing, rate-limit tiers and thresholds aren't published in a form that can be safely scraped and re-published automatically — several providers gate the real numbers behind a login, and formats vary too much across the nine to normalize reliably. This page is refreshed by hand periodically; see the methodology page for which parts of the site are live-pipelined and which are dated snapshots like this one.
