Should you use Step-3.5-Flash?
Step-3.5-Flash is listed for chat workloads and supports a 8K context window.
Use it when ModelScope's free tier is enough for evaluation, demos, or light production traffic.
https://api-inference.modelscope.cn/v1 stepfun-ai/Step-3.5-Flash Step-3.5-Flash is listed for chat workloads and supports a 8K context window.
Use it when ModelScope's free tier is enough for evaluation, demos, or light production traffic.
We only found this Step-3.5-Flash listing on ModelScope in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | stepfun-ai/Step-3.5-Flash | Free tier | 8K | Native | Varies |
| Model | Provider | Context | Access |
|---|---|---|---|
| Qwen/Qwen3.5-35B-A3B | ModelScope | 131K | Check provider |
| Qwen/Qwen3.5-27B | ModelScope | 131K | Check provider |
| MiniMax-M2.5-highspeed | ModelScope | 205K | Check provider |
| Kimi K2.5 | ModelScope | 262K | Check provider |
| Qwen/Qwen-Image | ModelScope | 131K | Check provider |
Step-3.5-Flash is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
Step-3.5-Flash is listed with free API access on ModelScope, subject to the provider's quota and account policy.
The model ID shown in this catalog is stepfun-ai/Step-3.5-Flash.
The listed context window is 8K tokens with up to 4K output tokens.
StepFun Step 3.5 Flash is available free on NVIDIA NIM with up to 40 RPM and no daily token cap. StepFun's latest flash-optimized model prioritizes speed and efficiency for high-throughput tasks while retaining strong general performance. OpenAI-compatible API. Requires free NVIDIA Developer Program membership and phone verification.
For API keys, setup steps, and provider-level limits, see the ModelScope provider page.