Should you use Qwen3.5-9B?
Qwen3.5-9B is listed for chat workloads and supports a 131K context window.
Use it when OVHcloud AI Endpoints's free tier is enough for evaluation, demos, or light production traffic.
https://oai.endpoints.kepler.ai.cloud.ovh.net/v1 qwen3.5-9b Qwen3.5-9B is listed for chat workloads and supports a 131K context window.
Use it when OVHcloud AI Endpoints's free tier is enough for evaluation, demos, or light production traffic.
We only found this qwen3-5-9b listing on OVHcloud AI Endpoints in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | Qwen3.5-9B | Free tier | 131K | OpenAI-style | 2 RPM (anonymous) |
| Model | Provider | Context | Access |
|---|---|---|---|
| Qwen3.5-397B-A17B | OVHcloud AI Endpoints | 131K | Free tier |
| Meta-Llama-3_3-70B-Instruct | OVHcloud AI Endpoints | 131K | Free tier |
| Qwen3.6-27B | OVHcloud AI Endpoints | 131K | Free tier |
| Qwen3-32B | OVHcloud AI Endpoints | 131K | Free tier |
| Qwen3-Coder-30B-A3B-Instruct | OVHcloud AI Endpoints | 262K | Free tier |
Qwen3.5-9B is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
Qwen3.5-9B is listed with free API access on OVHcloud AI Endpoints, subject to the provider's quota and account policy.
The model ID shown in this catalog is qwen3.5-9b.
The listed free tier limit is 2 RPM (anonymous). Limits can change per account tier, so confirm against the provider dashboard.
The listed context window is 131K tokens with up to 8K output tokens.
Multimodal reasoning model for visual analysis, planning, and tool use
For API keys, setup steps, and provider-level limits, see the OVHcloud AI Endpoints provider page.