NVIDIA NIM logo

bge-m3 Free API on NVIDIA NIM

Free API Verified
Catalog profile API ready provider catalog metadata

Flagship model for demanding analysis, coding, and production agent workflows

Free APIOpenAI compatibleText
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://integrate.api.nvidia.com/v1
Model ID
baai/bge-m3
API format OpenAI Chat Completions
Technical Details

bge-m3 specifications

Provider catalog
Context window 131K
Max output 8K
Status Online
Released Jan 30, 2024
Last updated Aug 6, 2026
Free listing since Jan 30, 2024
Input text
Output text
AI Recommendation

Should you use bge-m3?

bge-m3 is listed for chat workloads and supports a 131K context window.

Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Works with OpenAI-style SDKs
  • Live API verification available

Watch outs

  • Free-tier rate limits apply
  • Vision support is not listed
  • Tool calling is not confirmed
Pricing

bge-m3 pricing per 1M tokens

Free tier listed
Input $0 per 1M tokens
Output $0 per 1M tokens
Free access Available NVIDIA NIM
Rate limit Up to 40 RPM provider policy
Availability

bge-m3 availability by provider

Current provider only

We only found this bge-m3 listing on NVIDIA NIM in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
NVIDIA NIM baai/bge-m3 Free tier 131K OpenAI-style Up to 40 RPM
View NVIDIA NIM setup guide →
Typical Use Cases

bge-m3 use cases

Chat

bge-m3 is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

bge-m3 free API FAQ

Is bge-m3 free to use?

bge-m3 is listed with free API access on NVIDIA NIM, subject to the provider's quota and account policy.

What is the bge-m3 model ID?

The model ID shown in this catalog is baai/bge-m3.

What are the bge-m3 free tier rate limits on NVIDIA NIM?

The listed free tier limit is Up to 40 RPM. Limits can change per account tier, so confirm against the provider dashboard.

What context window does bge-m3 support?

The listed context window is 131K tokens with up to 8K output tokens.

More about bge-m3

Flagship model for demanding analysis, coding, and production agent workflows

For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.