Should you use ling-3.1-flash?
ling-3.1-flash is listed for chat workloads and supports a 262K context window.
Use it when Kilo Code's free tier is enough for evaluation, demos, or light production traffic.
https://api.kilo.ai/api/gateway inclusionai/ling-3.1-flash ling-3.1-flash is listed for chat workloads and supports a 262K context window.
Use it when Kilo Code's free tier is enough for evaluation, demos, or light production traffic.
We found 2 provider listings for ling. Check model ID, quota, pricing, and API format before switching providers.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | inclusionai/ling-3.1-flash | Free tier | 262K | OpenAI-style | 200 req/hr |
| | inclusionAI: Ling 3.1 Flash | Free tier | 262K | OpenAI-style | 200 req/day (free tier) |
| Model | Provider | Context | Access |
|---|---|---|---|
| inclusionAI: Ling 3.1 Flash | OpenRouter | 262K | Free tier |
| nvidia/nemotron-3-ultra-550b-a55b:free | Kilo Code | 1.0M | Free tier |
| stepfun/step-3.7-flash:free | Kilo Code | 262K | Free tier |
| nvidia/nemotron-3-super-120b-a12b:free | Kilo Code | 262K | Free tier |
| nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free | Kilo Code | 256K | Free tier |
ling-3.1-flash is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
ling-3.1-flash is listed with free API access on Kilo Code, subject to the provider's quota and account policy.
The model ID shown in this catalog is inclusionai/ling-3.1-flash.
The listed free tier limit is 200 req/hr. Limits can change per account tier, so confirm against the provider dashboard.
The listed context window is 262K tokens with up to 32K output tokens.
Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.
For API keys, setup steps, and provider-level limits, see the Kilo Code provider page.