OVHcloud AI Endpoints logo

gpt-oss-20b API status on OVHcloud AI Endpoints

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

Open GPT reasoning model for self-hosted agents and controllable deployments

Check providerOpenAI compatibleReasoningTool callingJSON modeTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://oai.endpoints.kepler.ai.cloud.ovh.net/v1
Model ID
gpt-oss-20b
API format OpenAI-style
Technical Details

gpt-oss-20b specifications

Provider and model catalog
Context window 128K
Max output 8K
Status Check provider
Family gpt-oss
Released Aug 5, 2025
Last updated Jul 30, 2026
Free listing since Aug 5, 2025
Input text
Output text
Capabilities reasoning, tool calling, structured output, temperature control
Open weights Yes
AI Recommendation

Should you use gpt-oss-20b?

gpt-oss-20b is listed for chat, coding workloads and supports a 128K context window.

Check OVHcloud AI Endpoints's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
  • Coding
Strengths & Weaknesses

Strengths

  • Useful for coding workflows
  • Long context window
  • Reasoning mode listed
  • Tool calling support

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for gpt-oss-20b

Measured data
Intelligence General reasoning and instruction following
14.9/100
Coding Programming and code generation
20.7/100
Agentic Tool use and multi-step tasks
3.1/100
Speed Observed generation speed
173 tok/s
Context Maximum listed context window
128K
Pricing

gpt-oss-20b pricing per 1M tokens

Check provider
Input $0.06 per 1M tokens
Output $0.2 per 1M tokens
Free access Check provider OVHcloud AI Endpoints
Rate limit 2 RPM (anonymous) provider policy
Availability

gpt-oss-20b availability by provider

2 alternatives

We found 3 provider listings for gpt-oss-20b. Check model ID, quota, pricing, and API format before switching providers.

Provider Model listing Access Context API Limits
OVHcloud AI Endpoints gpt-oss-20b Check provider 128K OpenAI-style 2 RPM (anonymous)
Ollama Cloud gpt-oss:20b Free tier 131K OpenAI-style Session/weekly limits (unpublished)
LLM7.io GPT OSS 20B Free tier 131K Native Varies
View OVHcloud AI Endpoints setup guide →
Typical Use Cases

gpt-oss-20b use cases

Chat

gpt-oss-20b is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

Coding

gpt-oss-20b is tagged for coding in this catalog and works with OpenAI-compatible client libraries.

FAQ

gpt-oss-20b free API FAQ

Is gpt-oss-20b free to use?

gpt-oss-20b appears in the free model catalog for OVHcloud AI Endpoints, but its current endpoint availability should be confirmed with the provider before use.

What is the gpt-oss-20b model ID?

The model ID shown in this catalog is gpt-oss-20b.

What are the gpt-oss-20b free tier rate limits on OVHcloud AI Endpoints?

The listed free tier limit is 2 RPM (anonymous). Limits can change per account tier, so confirm against the provider dashboard.

What context window does gpt-oss-20b support?

The listed context window is 128K tokens with up to 8K output tokens.

More about gpt-oss-20b

Open-weight GPT model for self-hosted reasoning and instruction-following workloads

For API keys, setup steps, and provider-level limits, see the OVHcloud AI Endpoints provider page.