OVHcloud AI Endpoints logo

Llama-3.1-8B-Instruct API status on OVHcloud AI Endpoints

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

Llama-3.1-8B-Instruct — free model from OVHcloud AI Endpoints.

Check providerOpenAI compatibleTool callingJSON modeText
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://oai.endpoints.kepler.ai.cloud.ovh.net/v1
Model ID
llama-3-1-8b-instruct
API format OpenAI-style
Technical Details

Llama-3.1-8B-Instruct specifications

Provider catalog
Context window 131K
Max output 4K
Status Check provider
Family llama-3-1-8b-instruct
Released Jul 23, 2024
Last updated Jul 30, 2026
Free listing since Jul 23, 2024
Input text
Output text
Capabilities tool calling, structured output
AI Recommendation

Should you use Llama-3.1-8B-Instruct?

Llama-3.1-8B-Instruct is listed for chat workloads and supports a 131K context window.

Check OVHcloud AI Endpoints's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Tool calling support
  • Structured JSON output
  • Works with OpenAI-style SDKs

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for Llama-3.1-8B-Instruct

Measured data
Intelligence General reasoning and instruction following
7.6/100
Coding Programming and code generation
5.4/100
Agentic Tool use and multi-step tasks
0.5/100
Speed Observed generation speed
140 tok/s
Context Maximum listed context window
131K
Pricing

Llama-3.1-8B-Instruct pricing per 1M tokens

Check provider
Input $0.08 per 1M tokens
Output $0.09 per 1M tokens
Free access Check provider OVHcloud AI Endpoints
Rate limit 2 RPM (anonymous) provider policy
Availability

Llama-3.1-8B-Instruct availability by provider

3 alternatives

We found 4 provider listings for llama-3-1-8b-instruct. Check model ID, quota, pricing, and API format before switching providers.

Provider Model listing Access Context API Limits
OVHcloud AI Endpoints Llama-3.1-8B-Instruct Check provider 131K OpenAI-style 2 RPM (anonymous)
Hugging Face Meta-Llama-3.1-8B-Instruct Free tier 128K Native Credit-metered
Cloudflare Workers AI @cf/meta/llama-3.1-8b-instruct-fp8 Free tier 8K Native Varies
NVIDIA NIM meta/llama-3.1-8b-instruct Free tier 8K Native Varies
View OVHcloud AI Endpoints setup guide →
Typical Use Cases

Llama-3.1-8B-Instruct use cases

Chat

Llama-3.1-8B-Instruct is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

Llama-3.1-8B-Instruct free API FAQ

Is Llama-3.1-8B-Instruct free to use?

Llama-3.1-8B-Instruct appears in the free model catalog for OVHcloud AI Endpoints, but its current endpoint availability should be confirmed with the provider before use.

What is the Llama-3.1-8B-Instruct model ID?

The model ID shown in this catalog is llama-3-1-8b-instruct.

What are the Llama-3.1-8B-Instruct free tier rate limits on OVHcloud AI Endpoints?

The listed free tier limit is 2 RPM (anonymous). Limits can change per account tier, so confirm against the provider dashboard.

What context window does Llama-3.1-8B-Instruct support?

The listed context window is 131K tokens with up to 4K output tokens.

More about Llama-3.1-8B-Instruct

Llama-3.1-8B-Instruct — free model from OVHcloud AI Endpoints.

For API keys, setup steps, and provider-level limits, see the OVHcloud AI Endpoints provider page.