Other
Chat completion (proxied to inference-gateway).
Chat completion (proxied to inference-gateway). Phase 2: models[]/sort ordered fallback with retry + 30s outage exclusion, @preset/<slug>, guardrails, opt-in response cache.
Not yet described in the overlay (TI-20).
Authorizations
Authorization: Bearer ozk_…. ozk_ platform API key (Authorization: Bearer ozk_…)Authorization: Bearer sk-ocean-…. Legacy sk-ocean- inference key: still accepted on the money plane until the sk-ocean- sunset; new integrations use the ozk_ keyBody
application/jsonMoney plane, ozk_ auth. Transparent proxy to inference-gateway — any OpenAI-compatible field is passed through untouched; the fields below are the ones Phase 2 routing adds on top.
Available options: price, latency, throughput
Default: false
Response
Completion (model = the slug that actually served it, not necessarily the one requested). stream:true returns text/event-stream instead — see the developer guide for the streaming caveat.
application/jsontext/event-stream
Response headers
Missing or invalid API key
application/json
byok_required — this provider needs your own connected key on this plan
application/json
model_not_allowed (routing candidates excluded by an allow-list) or content_blocked (prompt matched a guardrail regex)
application/json
preset_not_found — @preset/<slug> does not exist or is inactive
application/json
rate_limit_exceeded (per-key/per-provider) or spend_cap_exceeded (guardrail cap reached)
application/json
money_plane_unavailable — INFERENCE_GATEWAY_SERVICE not bound on this tier
application/json