Response format
OpenAI-compatible interfaces (/v1/*) return a standard structure on error:
/api/query/v1/*) use a unified envelope:
Common HTTP status codes
Common error types (type / code)
authentication_error/invalid_api_key: invalid token.invalid_request_error/model_not_found: model name does not exist or is unavailable.invalid_request_error/context_length_exceeded: context length exceeded.rate_limit_error: rate limit triggered.insufficient_quota: insufficient or exhausted quota.permission_error: account not enabled for this model or capability.
Request ID in error responses
Error responses may carry a request ID (see Request ID):Troubleshooting tips
- First confirm the token is valid, unexpired, and has the required permissions.
- Record the
X-Oneapi-Request-Idresponse header for reconciliation and ticket tracking. - For
401/403, check your auth method and permission points first (see Authentication). - For
429, reduce concurrency, retry with backoff, or contact support to raise quota. 500/502/503are usually transient; retry later. Submit a ticket if they persist.