ModelScope logo

DeepSeek-V4-Flash Free API on ModelScope

Free API Verified
★★★★★★★★★★ 3.5 Benchmark-backed score

Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work

Free APIOpenAI compatibleReasoningTool callingJSON mode
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api-inference.modelscope.cn/v1
Model ID
deepseek-ai/DeepSeek-V4-Flash-0731
API format OpenAI Chat Completions + OpenAI Responses
Technical Details

DeepSeek-V4-Flash specifications

Provider and model catalog
Context window 8K
Max output 4K
Status Online
Family deepseek-flash
Knowledge cutoff 2025-05
Released Apr 24, 2026
Last updated Jul 7, 2026
Free listing since Apr 24, 2026
Input text
Output text
Capabilities reasoning, tool calling, structured output, temperature control
Open weights Yes

External benchmark references

Benchmark Score Metric Date
SWE-Bench Verified 79 resolved Not listed
AI Recommendation

Should you use DeepSeek-V4-Flash?

DeepSeek-V4-Flash is listed for chat workloads and supports a 8K context window.

Use it when ModelScope's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Reasoning mode listed
  • Tool calling support
  • Structured JSON output
  • Open weights available

Watch outs

  • Vision support is not listed
Benchmark Overview

Benchmark signals for DeepSeek-V4-Flash

Measured data
Intelligence General reasoning and instruction following
40.3/100
Coding Programming and code generation
56.2/100
Agentic Tool use and multi-step tasks
31.1/100
Context Maximum listed context window
8K
Typical Use Cases

DeepSeek-V4-Flash use cases

Chat

DeepSeek-V4-Flash is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

DeepSeek-V4-Flash free API FAQ

Is DeepSeek-V4-Flash free to use?

DeepSeek-V4-Flash is listed with free API access on ModelScope, subject to the provider's quota and account policy.

What is the DeepSeek-V4-Flash model ID?

The model ID shown in this catalog is deepseek-ai/DeepSeek-V4-Flash-0731.

What context window does DeepSeek-V4-Flash support?

The listed context window is 8K tokens with up to 4K output tokens.

More about DeepSeek-V4-Flash

DeepSeek V4 Flash is available free on NVIDIA NIM with up to 40 RPM and no daily token cap. As DeepSeek's latest-generation flash variant, it prioritizes speed and efficiency while retaining strong general-purpose performance. NVIDIA's OpenAI-compatible endpoint makes it a drop-in replacement for any tool that accepts a custom base URL. NVIDIA Developer Program membership (free) required; phone verification needed for API key.

For API keys, setup steps, and provider-level limits, see the ModelScope provider page.