Should you use ling-3.0-flash?
ling-3.0-flash is listed for chat workloads and supports a 262K context window.
Use it when Kilo Code's free tier is enough for evaluation, demos, or light production traffic.
https://api.kilo.ai/api/gateway inclusionai/ling-3.0-flash:free ling-3.0-flash is listed for chat workloads and supports a 262K context window.
Use it when Kilo Code's free tier is enough for evaluation, demos, or light production traffic.
We only found this ling-3.0-flash listing on Kilo Code in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | inclusionai/ling-3.0-flash:free | Free tier | 262K | OpenAI-style | ~200 req/hr |
| Model | Provider | Context | Access |
|---|---|---|---|
| nvidia/nemotron-3-ultra-550b-a55b:free | Kilo Code | 1.0M | Free tier |
| stepfun/step-3.7-flash:free | Kilo Code | 262K | Free tier |
| nvidia/nemotron-3-super-120b-a12b:free | Kilo Code | 262K | Free tier |
| nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free | Kilo Code | 256K | Free tier |
| poolside/laguna-s-2.1:free | Kilo Code | 262K | Free tier |
ling-3.0-flash is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
ling-3.0-flash is listed with free API access on Kilo Code, subject to the provider's quota and account policy.
The model ID shown in this catalog is inclusionai/ling-3.0-flash:free.
The listed free tier limit is ~200 req/hr. Limits can change per account tier, so confirm against the provider dashboard.
The listed context window is 262K tokens with up to 32K output tokens.
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
For API keys, setup steps, and provider-level limits, see the Kilo Code provider page.