Kilo Code logo

ling-3.1-flash Free API on Kilo Code

Free API
★★★★★★★★★★ 3.5 Benchmark-backed score

Efficient model for low-latency assistance, extraction, and routine automation

Free APIOpenAI compatibleReasoningTool callingTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api.kilo.ai/api/gateway
Model ID
inclusionai/ling-3.1-flash
API format OpenAI-style
Technical Details

ling-3.1-flash specifications

Provider and model catalog
Context window 262K
Max output 32K
Status Online
Family ling
Released Sep 29, 2026
Last updated Oct 6, 2026
Free listing since Sep 29, 2026
Input text
Output text
Capabilities reasoning, tool calling
AI Recommendation

Should you use ling-3.1-flash?

ling-3.1-flash is listed for chat workloads and supports a 262K context window.

Use it when Kilo Code's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Works with OpenAI-style SDKs

Watch outs

  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for ling-3.1-flash

Measured data
Intelligence General reasoning and instruction following
41.1/100
Speed Observed generation speed
211 tok/s
Context Maximum listed context window
262K
Pricing

ling-3.1-flash pricing per 1M tokens

Free tier listed
Input $0.3 per 1M tokens
Output $0.9 per 1M tokens
Access Available Kilo Code
Rate limit 200 req/hr provider policy
Typical Use Cases

ling-3.1-flash use cases

Chat

ling-3.1-flash is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

ling-3.1-flash free API FAQ

Is ling-3.1-flash free to use?

ling-3.1-flash is listed with free API access on Kilo Code, subject to the provider's quota and account policy.

What is the ling-3.1-flash model ID?

The model ID shown in this catalog is inclusionai/ling-3.1-flash.

What are the ling-3.1-flash free tier rate limits on Kilo Code?

The listed free tier limit is 200 req/hr. Limits can change per account tier, so confirm against the provider dashboard.

What context window does ling-3.1-flash support?

The listed context window is 262K tokens with up to 32K output tokens.

More about ling-3.1-flash

Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.

For API keys, setup steps, and provider-level limits, see the Kilo Code provider page.