Skip to content
Inference

Limits and support

Limits

SettingLimit
Context and outputPer-model limits
Request body50 MiB total JSON, including base64 images and history
Image formatsJPEG and PNG; base64 adds roughly one third to file size
Rate and concurrencyNo fixed per-key allowance is published; contact support for capacity commitments
Serving deadlinesUp to 10 minutes for prefill; 15 minutes for a non-streaming response. Client/network timeouts may be shorter.

Set a client timeout explicitly, such as timeout=600.0 on OpenAI(...), and use streaming for long outputs. Incomplete streams cannot be resumed; retries are new requests and may incur additional charges. Do not replay completed tool actions.

Errors

StatusAction
400 / 422Fix the request or unsupported parameter.
401 / 403Check the key and account/model access.
402Check your Console balance.
404Check the model ID and endpoint.
413Reduce request size.
429 / 503Reduce concurrency; back off with jitter and respect Retry-After.
500 / 502 / 504Check service status; retry transient failures with a retry limit.

Support

Service status · support@river.ai

Include the model, endpoint, request time/timezone, status code, and x-request-id when available. Never include your API key.

Data handling

Default Commercial terms: no training on your inputs or outputs; User Content retained up to 2 days after the session ends, with legal, safety/security/compliance, and agreed exceptions. De-identified aggregate statistics may be retained longer. Testing terms and signed addenda may differ; store: false is not a zero-data-retention agreement.