NVIDIA NIM logo

qwen3.5-397b-a17b API status on NVIDIA NIM

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

Large open Qwen multimodal MoE for visual agents and long technical tasks

Check providerOpenAI compatibleReasoningTool callingJSON modeTextImage
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://integrate.api.nvidia.com/v1
Model ID
qwen/qwen3.5-397b-a17b
API format OpenAI Chat Completions
Technical Details

qwen3.5-397b-a17b specifications

Provider and model catalog
Context window 262K
Max output 66K
Status Check provider
Family qwen
Knowledge cutoff 2026-01
Released Feb 15, 2026
Last updated Jul 26, 2026
Free listing since Feb 16, 2026
Input text, image, video, audio
Output text
Capabilities reasoning, tool calling, structured output, file attachments, temperature control
Open weights Yes

External benchmark references

Benchmark Score Metric Date
SWE-Bench Verified 76.4 resolved Not listed
AI Recommendation

Should you use qwen3.5-397b-a17b?

qwen3.5-397b-a17b is listed for chat workloads and supports a 262K context window.

Check NVIDIA NIM's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Structured JSON output

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
Benchmark Overview

Benchmark signals for qwen3.5-397b-a17b

Measured data
Intelligence General reasoning and instruction following
33.7/100
Coding Programming and code generation
48.2/100
Agentic Tool use and multi-step tasks
19.8/100
Speed Observed generation speed
68 tok/s
Context Maximum listed context window
262K
Pricing

qwen3.5-397b-a17b pricing per 1M tokens

Check provider
Input $0.6 per 1M tokens
Output $3.6 per 1M tokens
Free access Check provider NVIDIA NIM
Rate limit Up to 40 RPM provider policy
Availability

qwen3.5-397b-a17b availability by provider

2 alternatives

We found 3 provider listings for qwen3-5-397b-a17b. Check model ID, quota, pricing, and API format before switching providers.

Provider Model listing Access Context API Limits
NVIDIA NIM qwen/qwen3.5-397b-a17b Check provider 262K OpenAI-style Up to 40 RPM
OVHcloud AI Endpoints Qwen3.5-397B-A17B Free tier 131K OpenAI-style 2 RPM (anonymous)
ModelScope Qwen/Qwen3.5-397B-A17B Free tier 8K Native Varies
View NVIDIA NIM setup guide →
Typical Use Cases

qwen3.5-397b-a17b use cases

Chat

qwen3.5-397b-a17b is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

qwen3.5-397b-a17b free API FAQ

Is qwen3.5-397b-a17b free to use?

qwen3.5-397b-a17b appears in the free model catalog for NVIDIA NIM, but its current endpoint availability should be confirmed with the provider before use.

What is the qwen3.5-397b-a17b model ID?

The model ID shown in this catalog is qwen/qwen3.5-397b-a17b.

What are the qwen3.5-397b-a17b free tier rate limits on NVIDIA NIM?

The listed free tier limit is Up to 40 RPM. Limits can change per account tier, so confirm against the provider dashboard.

What context window does qwen3.5-397b-a17b support?

The listed context window is 262K tokens with up to 66K output tokens.

More about qwen3.5-397b-a17b

Qwen3.5 397B (17B active via MoE) is Alibaba's largest publicly-hosted model, free on NVIDIA NIM with up to 40 RPM and no daily token cap. The enormous expert pool at 397B total parameters brings frontier-level capability while 17B active parameters maintain manageable inference cost. OpenAI-compatible API. Requires NVIDIA Developer Program membership and phone verification.

For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.