Safe steady capacity
160.75exact-derivedBinding constraint: daily budget.
[1][2][3]Portfolio 14 · quota envelope
Find the minimum of request, input-token, output-token, concurrency, and daily-budget capacity, then show how retries and explicit quota windows change the result.
Decision this answers: Which documented or account-specific limit binds steady traffic, and will the actual burst fit its rolling or fixed window?
Safe steady capacity
160.75exact-derivedBinding constraint: daily budget.
[1][2][3]Daily request plan
231481exact-derivedSteady capacity multiplied by 1,440 minutes.
[1]Retry quota multiplier
1.08model estimateExpected attempts counted against all applicable quotas.
[1]Example burst headroom
12.5%exact-derivedA 3,500-request window against the 4,000 RPM example.
[1]The tables below are the accessible source of truth; bar lengths never carry identity alone.
| Candidate | Request quota | Token quota | Window | Retry treatment | Guarantee | Evidence |
|---|---|---|---|---|---|---|
| Anthropic | Tier/account-specific | Input/output categories | Documented rolling behavior | Counts by failure point | Quota compliance only | published list[1] |
| OpenAI | Account/project-specific | Model-specific | Response headers/dashboard | Retries consume capacity | Quota compliance only | published list[2] |
| Gemini | Tier/project-specific | Model-specific | Documented tier windows | Retries consume capacity | Quota compliance only | published list[3] |
Use the official surface to confirm native prices, counting rules, limits, and product eligibility.
Provider docs define quota categories and windows. This page reconciles them with measured latency, request shape, retry behavior, and a daily budget without claiming provider availability.
Every result-affecting reference is visible here without JavaScript and is retained in the JSON export.
rate limit tiers and token bucketsrate limit headersGemini API tiers