Vora Cloud

Errors

Errors use the OpenAI shape. OpenAI SDKs raise their usual exceptions.

{
  "error": {
    "message": "Rate limit reached: 60 requests per minute. Retry in 12s.",
    "type": "rate_limit_error",
    "param": null,
    "code": "rate_limit_requests"
  }
}

Codes

StatusCodeMeaningRetry
400invalid_jsonThe body is not valid JSON.Don't retry
400invalid_requestA parameter is missing or has the wrong type.Don't retry
400context_length_exceededPrompt plus reply exceed the 16,384-token context window.Don't retry
400unsupported_parameterWe don't support this parameter, e.g. n > 1 or logit_bias.Don't retry
400unsupported_contentA message has non-text content, e.g. an image.Don't retry
401missing_api_keyNo Authorization: Bearer header.Don't retry
401invalid_api_keyThe key does not exist.Don't retry
403key_revokedThe key was revoked. Create a new one.Don't retry
403account_disabledThe account is disabled. Contact us.Don't retry
404model_not_foundUnknown model. See GET /v1/models.Don't retry
404unknown_endpointWe serve chat completions and models only.Don't retry
429rate_limit_requestsOver 60 requests in a minute.Retry after Retry-After
429rate_limit_tokensOver 500,000 tokens today.Retry after Retry-After
429concurrency_limitOver 2 requests at once on this key.Retry after Retry-After
500internal_errorSomething failed on our side.Retry with backoff
503model_unavailableThe model is offline or loading.Retry with backoff
503engine_busyThe model is at capacity.Retry with backoff
503model_timeoutThe model did not answer in time. Try a shorter prompt.Retry with backoff
503stream_interruptedThe model stopped mid-reply.Retry with backoff

When to retry

Rate-limit headers

Every accepted request returns these headers. A 429 returns Retry-After instead.

HeaderMeaning
x-ratelimit-limit-requestsRequests allowed per minute.
x-ratelimit-remaining-requestsRequests left in the current minute.
x-ratelimit-reset-requestsTime until the minute resets, e.g. 12s.
x-ratelimit-limit-tokensTokens allowed per day.
x-ratelimit-remaining-tokensTokens left today.
x-request-idThis request's id. Include it when you contact us.

Errors in a stream

If the model stops unexpectedly during a stream, the last event is an error, not [DONE]:

data: {"error":{"message":"The model stopped unexpectedly. Retry the request.","type":"server_error","param":null,"code":"stream_interrupted"}}

Check each event for an error field before you read choices.