How ModelRush works
Live
Open in ChatGPT
(opens in new tab)Last verified: 2026-08-11Request lifecycle
- Your server authenticates with a ModelRush API key.
- The gateway validates the payload and model/endpoint compatibility.
- Region routing selects an eligible execution path.
- The provider executes the workload.
- ModelRush normalizes the response and records status, latency, errors, and billable usage.
Stable application boundary
Keep your product code coupled to the ModelRush endpoint contract, not to provider-specific SDKs. Provider differences remain visible where they affect parameters, availability, or data handling.
Traceability
Store response and generation IDs with your own user or job ID. Never place personal data in log labels merely to make requests searchable.
PreviousRegions and pricing