OVHcloud AI Endpoints logo

Qwen2.5-VL-72B-Instruct Free API on OVHcloud AI Endpoints

Free API
Catalog metadata matched Tool use models.dev metadata

Qwen vision-language model for visual reasoning, documents, and agent tasks

Free APIOpenAI compatibleTool callingJSON modeTextImage
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://oai.endpoints.kepler.ai.cloud.ovh.net/v1
Model ID
qwen2.5-vl-72b-instruct
API format OpenAI-style
Technical Details

Qwen2.5-VL-72B-Instruct specifications

Provider and model catalog
Context window 128K
Max output 8K
Status Online
Family qwen
Knowledge cutoff 2024-04
Released Sep 1, 2024
Last updated Aug 6, 2026
Free listing since Sep 1, 2024
Input text, image
Output text
Capabilities tool calling, structured output, temperature control
Open weights Yes
AI Recommendation

Should you use Qwen2.5-VL-72B-Instruct?

Qwen2.5-VL-72B-Instruct is listed for chat workloads and supports a 128K context window.

Use it when OVHcloud AI Endpoints's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Tool calling support
  • Structured JSON output
  • Open weights available

Watch outs

  • Free-tier rate limits apply
Pricing

Qwen2.5-VL-72B-Instruct pricing per 1M tokens

Free tier listed
Input $1.01 per 1M tokens
Output $1.01 per 1M tokens
Free access Available OVHcloud AI Endpoints
Rate limit 2 RPM (anonymous) provider policy
Typical Use Cases

Qwen2.5-VL-72B-Instruct use cases

Chat

Qwen2.5-VL-72B-Instruct is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

Qwen2.5-VL-72B-Instruct free API FAQ

Is Qwen2.5-VL-72B-Instruct free to use?

Qwen2.5-VL-72B-Instruct is listed with free API access on OVHcloud AI Endpoints, subject to the provider's quota and account policy.

What is the Qwen2.5-VL-72B-Instruct model ID?

The model ID shown in this catalog is qwen2.5-vl-72b-instruct.

What are the Qwen2.5-VL-72B-Instruct free tier rate limits on OVHcloud AI Endpoints?

The listed free tier limit is 2 RPM (anonymous). Limits can change per account tier, so confirm against the provider dashboard.

What context window does Qwen2.5-VL-72B-Instruct support?

The listed context window is 128K tokens with up to 8K output tokens.

More about Qwen2.5-VL-72B-Instruct

Qwen2.5 VL 72B is Alibaba's large vision-language model, available free on OVHcloud AI Endpoints. With 128K context, OpenAI-compatible API, and full multimodal support (text + vision), it is a strong choice for image understanding, visual Q&A, and document analysis with diagrams. The 8K output per request is sufficient for most vision tasks. Hosted on OVHcloud's European infrastructure; registration required.

For API keys, setup steps, and provider-level limits, see the OVHcloud AI Endpoints provider page.