Should you use deepseek-v4-flash?
deepseek-v4-flash is listed for chat workloads and supports a 1.0M context window.
Use the comparison and availability sections before choosing this listing, because it is not currently marked as free.
https://api.ollama.com deepseek-v4-flash:preview | Benchmark | Score | Metric | Date |
|---|---|---|---|
| SWE-Bench Verified | 79 | resolved | Not listed |
deepseek-v4-flash is listed for chat workloads and supports a 1.0M context window.
Use the comparison and availability sections before choosing this listing, because it is not currently marked as free.
We found 3 provider listings for deepseek-v4-flash. Check model ID, quota, pricing, and API format before switching providers.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | deepseek-v4-flash | Paid | 1.0M | OpenAI-style | Session/weekly limits (unpublished) |
| | deepseek-ai/deepseek-v4-flash | Free tier | 1.0M | OpenAI-style | Up to 40 RPM |
| | deepseek-ai/DeepSeek-V4-Flash | Free tier | 8K | Native | Varies |
| Model | Provider | Context | Access |
|---|---|---|---|
| deepseek-ai/deepseek-v4-flash | NVIDIA NIM | 1.0M | Free tier |
| deepseek-ai/DeepSeek-V4-Flash | ModelScope | 8K | Free tier |
| minimax-m3 | Ollama Cloud | 1.0M | Free tier |
| gpt-oss:20b | Ollama Cloud | 131K | Free tier |
| nemotron-3-ultra | Ollama Cloud | 262K | Free tier |
deepseek-v4-flash is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
deepseek-v4-flash is not currently marked as free on Ollama Cloud.
The model ID shown in this catalog is deepseek-v4-flash:preview.
The listed free tier limit is Session/weekly limits (unpublished). Limits can change per account tier, so confirm against the provider dashboard.
The listed context window is 1.0M tokens with up to 131K output tokens.
Fast DeepSeek model for efficient chat, coding help, and agent loops
For API keys, setup steps, and provider-level limits, see the Ollama Cloud provider page.