Should you use Mistral-Small-3.1?
Mistral-Small-3.1 is listed for chat workloads and supports a 128K context window.
Use it when GitHub Models's free tier is enough for evaluation, demos, or light production traffic.
Mistral Small 3.1 is Mistral AI's efficient 24B-parameter model, available free on GitHub Models.
https://models.github.ai/inference mistral-small-3-1 Mistral-Small-3.1 is listed for chat workloads and supports a 128K context window.
Use it when GitHub Models's free tier is enough for evaluation, demos, or light production traffic.
We only found this Mistral-Small-3.1 listing on GitHub Models in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | Mistral-Small-3.1 | Free tier | 128K | OpenAI-style | 15 RPM, 150 RPD |
| Model | Provider | Context | Access |
|---|---|---|---|
| Phi-4 | GitHub Models | 131K | Free tier |
| Mistral Large (24.11) | GitHub Models | 131K | Free tier |
| AI21 Jamba 1.5 Large | GitHub Models | 256K | Free tier |
| gpt-5 | GitHub Models | 200K | Free tier |
| gpt-4.1 | GitHub Models | 1.0M | Free tier |
Mistral-Small-3.1 is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
Mistral-Small-3.1 is listed with free API access on GitHub Models, subject to the provider's quota and account policy.
The model ID shown in this catalog is mistral-small-3-1.
The listed free tier limit is 15 RPM, 150 RPD. Limits can change per account tier, so confirm against the provider dashboard.
The listed context window is 128K tokens with up to 4K output tokens.
Mistral Small 3.1 is Mistral AI's efficient 24B-parameter model, available free on GitHub Models. It delivers strong multilingual performance and instruction-following at a fraction of the cost of larger models, making it a practical choice for chat applications, content generation, and text analysis. With 128K context, OpenAI SDK compatibility, and 15 RPM / 150 RPD limits, it is well-suited for prototyping and personal projects. The main limitation on GitHub Models is the 4K per-request output cap — adequate for chat turns and short-form generation, but constraining for long-form writing tasks.
For API keys, setup steps, and provider-level limits, see the GitHub Models provider page.