Vora Cloud

Models & limits

What you can call, how much, and which parameters work.

Models

FactQwen3.5 9B
Model idqwen3.5-9b
VendorQwen (Alibaba)
Parameters9B
QuantizationQ4_K_M
Context window16,384 tokens (prompt + reply)
Max output2,048 tokens
InputText
CapabilitiesChat, Streaming, Tool calling, JSON mode
LicenseApache-2.0
PriceFree during beta
HostingSpain (EU)

GET /v1/models returns the same list as JSON.

Limits

Beta limits apply per API key.

LimitValueWhen you hit it
Requests per minute60429 rate_limit_requests
Tokens per day500,000429 rate_limit_tokens. Resets at 00:00 UTC.
Concurrent requests2429 concurrency_limit

Every accepted request reports where you stand in rate-limit headers.

Supported

Not supported

These return 400 with code unsupported_parameter or unsupported_content:

Thinking mode is off. The model answers directly.

Any other parameter is ignored.