Nscale logo

DeepSeek-R1-Distill-Llama-70B API status on Nscale

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

DeepSeek R1 Distill Llama 70B on OVHcloud AI Endpoints combines DeepSeek's reasoning distillation with Llama's 70B architecture, delivered from OVHcloud's European infrastructure.

Check providerNscale nativeReasoningTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://inference.api.nscale.com/v1
Model ID
deepseek-r1-distill-llama-70b
API format Nscale native
Technical Details

DeepSeek-R1-Distill-Llama-70B specifications

Provider catalog
Context window 128K
Max output 32K
Status Check provider
Family deepseek-r1-distill-llama-70b
Released Jan 20, 2025
Last updated Jun 15, 2026
Free listing since Jan 20, 2025
Input text, reasoning
Output text
Capabilities reasoning
AI Recommendation

Should you use DeepSeek-R1-Distill-Llama-70B?

DeepSeek-R1-Distill-Llama-70B is listed for chat, reasoning workloads and supports a 128K context window.

Check Nscale's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
  • Reasoning
Strengths & Weaknesses

Strengths

  • Strong reasoning profile
  • Long context window
  • Reasoning mode listed

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Vision support is not listed
  • Tool calling is not confirmed
Benchmark Overview

Benchmark signals for DeepSeek-R1-Distill-Llama-70B

Measured data
Coding Programming and code generation
11.4/100
Speed Observed generation speed
49 tok/s
Context Maximum listed context window
128K
Pricing

DeepSeek-R1-Distill-Llama-70B pricing per 1M tokens

Check provider
Input $0.7 per 1M tokens
Output $1.05 per 1M tokens
Free access Check provider Nscale
Rate limit Fair-use provider policy
Availability

DeepSeek-R1-Distill-Llama-70B availability by provider

Current provider only

We only found this deepseek-r1-distill-llama-70b listing on Nscale in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
Nscale DeepSeek-R1-Distill-Llama-70B Check provider 128K Native Fair-use
View Nscale setup guide →
Typical Use Cases

DeepSeek-R1-Distill-Llama-70B use cases

Chat

DeepSeek-R1-Distill-Llama-70B is tagged for chat in this catalog.

Reasoning

DeepSeek-R1-Distill-Llama-70B is tagged for reasoning in this catalog.

FAQ

DeepSeek-R1-Distill-Llama-70B free API FAQ

Is DeepSeek-R1-Distill-Llama-70B free to use?

DeepSeek-R1-Distill-Llama-70B appears in the free model catalog for Nscale, but its current endpoint availability should be confirmed with the provider before use.

What is the DeepSeek-R1-Distill-Llama-70B model ID?

The model ID shown in this catalog is deepseek-r1-distill-llama-70b.

What are the DeepSeek-R1-Distill-Llama-70B free tier rate limits on Nscale?

The listed free tier limit is Fair-use. Limits can change per account tier, so confirm against the provider dashboard.

What context window does DeepSeek-R1-Distill-Llama-70B support?

The listed context window is 128K tokens with up to 32K output tokens.

More about DeepSeek-R1-Distill-Llama-70B

DeepSeek R1 Distill Llama 70B on OVHcloud AI Endpoints combines DeepSeek's reasoning distillation with Llama's 70B architecture, delivered from OVHcloud's European infrastructure. With chain-of-thought reasoning, 131K context, and 32K output, it is well-suited for complex analytical tasks that benefit from step-by-step deliberation. OpenAI-compatible API; registration required. A good choice for EU developers who need reasoning capabilities with GDPR-compliant hosting.

For API keys, setup steps, and provider-level limits, see the Nscale provider page.