Kilo Code logo

ling-3.0-flash API status on Kilo Code

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*.

Check providerOpenAI compatibleReasoningTool callingTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api.kilo.ai/api/gateway
Model ID
inclusionai/ling-3.0-flash:free
API format OpenAI-style
Technical Details

ling-3.0-flash specifications

Provider catalog
Context window 262K
Max output 32K
Status Check provider
Released Aug 4, 2026
Last updated Aug 6, 2026
Free listing since Aug 4, 2026
Input text, reasoning
Output text
Capabilities reasoning, tool calling
AI Recommendation

Should you use ling-3.0-flash?

ling-3.0-flash is listed for chat workloads and supports a 262K context window.

Check Kilo Code's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Works with OpenAI-style SDKs

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for ling-3.0-flash

Measured data
Intelligence General reasoning and instruction following
24.9/100
Coding Programming and code generation
50.6/100
Agentic Tool use and multi-step tasks
21/100
Speed Observed generation speed
312 tok/s
Context Maximum listed context window
262K
Pricing

ling-3.0-flash pricing per 1M tokens

Check provider
Input $0.07 per 1M tokens
Output $0.22 per 1M tokens
Access Check provider Kilo Code
Rate limit ~200 req/hr provider policy
Availability

ling-3.0-flash availability by provider

Current provider only

We only found this ling-3.0-flash listing on Kilo Code in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
Kilo Code inclusionai/ling-3.0-flash:free Check provider 262K OpenAI-style ~200 req/hr
View Kilo Code setup guide →
Typical Use Cases

ling-3.0-flash use cases

Chat

ling-3.0-flash is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

ling-3.0-flash free API FAQ

Is ling-3.0-flash free to use?

ling-3.0-flash appears in the free model catalog for Kilo Code, but its current endpoint availability should be confirmed with the provider before use.

What is the ling-3.0-flash model ID?

The model ID shown in this catalog is inclusionai/ling-3.0-flash:free.

What are the ling-3.0-flash free tier rate limits on Kilo Code?

The listed free tier limit is ~200 req/hr. Limits can change per account tier, so confirm against the provider dashboard.

What context window does ling-3.0-flash support?

The listed context window is 262K tokens with up to 32K output tokens.

More about ling-3.0-flash

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

For API keys, setup steps, and provider-level limits, see the Kilo Code provider page.