Groq logo

kimi-k2-instruct API status on Groq

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

Kimi K2 Instruct on Groq is Moonshot AI's second-generation model, delivering 262K context and 262K output — one of the highest output ceilings available on any free endpoint.

Check providerOpenAI compatibleTool callingText
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api.groq.com/openai/v1
Model ID
kimi-k2-instruct
API format OpenAI Chat Completions + OpenAI Responses
Technical Details

kimi-k2-instruct specifications

Provider catalog
Context window 262K
Max output 262K
Status Check provider
Family kimi-k2
Released Sep 5, 2025
Last updated Jun 15, 2026
Free listing since Sep 5, 2025
Input text
Output text
Capabilities tool calling
AI Recommendation

Should you use kimi-k2-instruct?

kimi-k2-instruct is listed for chat workloads and supports a 262K context window.

Check Groq's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Tool calling support
  • Works with OpenAI-style SDKs

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for kimi-k2-instruct

Measured data
Intelligence General reasoning and instruction following
26.3/100
Coding Programming and code generation
22.1/100
Agentic Tool use and multi-step tasks
24.3/100
Speed Observed generation speed
26 tok/s
Context Maximum listed context window
262K
Pricing

kimi-k2-instruct pricing per 1M tokens

Check provider
Input $0.6 per 1M tokens
Output $2.5 per 1M tokens
Free access Check provider Groq
Rate limit 30 RPM, 14,400 RPD provider policy
Availability

kimi-k2-instruct availability by provider

2 alternatives

We found 3 provider listings for kimi-k2. Check model ID, quota, pricing, and API format before switching providers.

Provider Model listing Access Context API Limits
Groq kimi-k2-instruct Check provider 262K OpenAI-style 30 RPM, 14,400 RPD
Cloudflare Workers AI @cf/moonshotai/kimi-k2.7-code Free tier 262K Native 10K neurons/day (shared)
NVIDIA NIM moonshotai/kimi-k2.6 Free tier 262K OpenAI-style Up to 40 RPM
View Groq setup guide →
Typical Use Cases

kimi-k2-instruct use cases

Chat

kimi-k2-instruct is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

kimi-k2-instruct free API FAQ

Is kimi-k2-instruct free to use?

kimi-k2-instruct appears in the free model catalog for Groq, but its current endpoint availability should be confirmed with the provider before use.

What is the kimi-k2-instruct model ID?

The model ID shown in this catalog is kimi-k2-instruct.

What are the kimi-k2-instruct free tier rate limits on Groq?

The listed free tier limit is 30 RPM, 14,400 RPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does kimi-k2-instruct support?

The listed context window is 262K tokens with up to 262K output tokens.

More about kimi-k2-instruct

Kimi K2 Instruct on Groq is Moonshot AI's second-generation model, delivering 262K context and 262K output — one of the highest output ceilings available on any free endpoint. This makes it uniquely suited for long-form generation tasks: full document drafts, extended code refactors, or comprehensive analysis reports. Running on Groq's LPU hardware ensures responsive latency even with the high token throughput. The free tier offers 14,400 requests per day at 30 RPM. OpenAI SDK compatible; registration required, no credit card needed.

For API keys, setup steps, and provider-level limits, see the Groq provider page.