Status
All systems operationalChecked live from this page every minute. Last check 2026-09-17 09:30:54 UTC.
What is checked. The gateway probe requests /health over HTTPS and passes when it returns 200. The catalog probe lists /v1/models and passes when the list is non-empty; the count shown is distinct models, not aliases. The upstream probe confirms the inference backend answers at all; a model can still be temporarily unavailable, in which case requests return 503 model_unavailable and the error is deterministic to retry. Probes run from this server, cached for one minute, so timings include one network hop from the site to the API.
If your requests fail while everything here is green, check the error docs first: 401 and 402 are account-side, and 404 model_not_found means the id is not in the catalog.
Incident history
| Date | Event |
|---|---|
| 2026-09-16 | Gateway v2 deployed: added /v1/audio/speech, /v1/audio/transcriptions, streaming usage, per-unit metering. No downtime. |
| 2026-09-16 | TLS certificate for api.inferenceapis.com renewed after a lapse earlier in the day. Requests during the lapse failed certificate validation. |
Something wrong that is not shown here? Tell us.
