ModelRail

Docs
Dashboard

Limits and billing

Project limits may differ from the defaults below. These are not SLAs, plan names, or guaranteed pricing.

Common defaults

LimitTypical default
Requests per minuteCommonly 60
Concurrent requestsCommonly 10
Monthly spendWhen enabled → spending_limit_exceeded
Request body size2 MB (413)
Default max output4096 tokens if neither max_tokens nor max_completion_tokens is set
Gateway timeoutsFinite; longer for reasoning workloads

On 429, honor Retry-After and rate-limit headers — see Errors.

Cost notes

Note

Failed upstream attempts may still incur cost. A compatible fallback may create an additional attempt. See Models and aliases.