Kilo Code logo

ling-3.0-flash Free API on Kilo Code

Free API
★★★★★★★★★★ 3.5 Benchmark-backed score

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*.

Free APIOpenAI compatibleReasoningTool callingTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api.kilo.ai/api/gateway
Model ID
inclusionai/ling-3.0-flash:free
API format OpenAI-style
Technical Details

ling-3.0-flash specifications

Provider catalog
Context window 262K
Max output 32K
Status Online
Released Aug 4, 2026
Last updated Aug 6, 2026
Free listing since Aug 4, 2026
Input text, reasoning
Output text
Capabilities reasoning, tool calling
AI Recommendation

Should you use ling-3.0-flash?

ling-3.0-flash is listed for chat workloads and supports a 262K context window.

Use it when Kilo Code's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Works with OpenAI-style SDKs

Watch outs

  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for ling-3.0-flash

Measured data
Intelligence General reasoning and instruction following
37.4/100
Coding Programming and code generation
50.6/100
Agentic Tool use and multi-step tasks
29.6/100
Speed Observed generation speed
280 tok/s
Context Maximum listed context window
262K
Pricing

ling-3.0-flash pricing per 1M tokens

Free tier listed
Input $0.07 per 1M tokens
Output $0.22 per 1M tokens
Free access Available Kilo Code
Rate limit ~200 req/hr provider policy
Availability

ling-3.0-flash availability by provider

Current provider only

We only found this ling-3.0-flash listing on Kilo Code in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
Kilo Code inclusionai/ling-3.0-flash:free Free tier 262K OpenAI-style ~200 req/hr
View Kilo Code setup guide →
Typical Use Cases

ling-3.0-flash use cases

Chat

ling-3.0-flash is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

ling-3.0-flash free API FAQ

Is ling-3.0-flash free to use?

ling-3.0-flash is listed with free API access on Kilo Code, subject to the provider's quota and account policy.

What is the ling-3.0-flash model ID?

The model ID shown in this catalog is inclusionai/ling-3.0-flash:free.

What are the ling-3.0-flash free tier rate limits on Kilo Code?

The listed free tier limit is ~200 req/hr. Limits can change per account tier, so confirm against the provider dashboard.

What context window does ling-3.0-flash support?

The listed context window is 262K tokens with up to 32K output tokens.

More about ling-3.0-flash

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

For API keys, setup steps, and provider-level limits, see the Kilo Code provider page.