Should you use deepseek-v4-flash?
deepseek-v4-flash is listed for chat workloads and supports a 1.0M context window.
Use it when Cline's free tier is enough for evaluation, demos, or light production traffic.
See provider docs deepseek/deepseek-v4-flash | Benchmark | Score | Metric | Date |
|---|---|---|---|
| SWE-Bench Verified | 79 | resolved | Not listed |
deepseek-v4-flash is listed for chat workloads and supports a 1.0M context window.
Use it when Cline's free tier is enough for evaluation, demos, or light production traffic.
We only found this deepseek-v4-flash listing on Cline in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | deepseek-v4-flash | Free tier | 0 | Native | See provider page |
| Model | Provider | Context | Access |
|---|---|---|---|
| nemotron-3.5-lightning | Cline | 0 | Free tier |
| laguna-s-2.1:free | Cline | 0 | Free tier |
deepseek-v4-flash is tagged for chat in this catalog.
deepseek-v4-flash is listed with free API access on Cline, subject to the provider's quota and account policy.
The model ID shown in this catalog is deepseek/deepseek-v4-flash.
The listed context window is 1.0M tokens with up to 384K output tokens.
DeepSeek V4 Flash is a Mixture-of-Experts model with 284B total parameters (13B active per token), available free on OpenRouter. It features a 1M context window with hybrid attention for efficient long-context processing. Supports configurable reasoning effort (high/xhigh levels). Strong performance on coding, reasoning, and agent workflows. Model weights available on Hugging Face. Text-only. OpenAI-compatible via OpenRouter. Free tier: 200 RPD (or 1,000 with $10 lifetime credit).
For API keys, setup steps, and provider-level limits, see the Cline provider page.