Directory of Free LLM APIs: Compare 410+ Models
Showing 410 of 410 free LLM models
Discover and filter 410+ free LLM models across 31 providers. Find APIs by capability (vision, reasoning), rate limits, or no-credit-card requirements, and get the perfect free AI model for your project.
| Provider | Model | Score | Context | Modality | Rate Limit | Status |
|---|---|---|---|---|---|---|
| kimi-k3 Paid kimi-k3 | 96 | 128K | Session/weekly limits (unpublished) | Online | ||
| DeepSeek: DeepSeek V4 Flash 0423 | 90 | 1.0M | 200 req/day (free tier) | Online | ||
| Gemini 3.6 Flash Verified Gemini 3.6 Flash | 90 | 1.0M | 15 RPM, 1,500 RPD | Online | ||
| Kimi K3 Verified | 89 | 1.0M | Online | |||
| Tencent: Hy3 Paid | 88 | 262K | 200 req/day (free tier) | Online | ||
| MiniMax: MiniMax M3 Paid MiniMax: MiniMax M3 | 88 | 1.0M | 200 req/day (free tier) | Online | ||
| deepseek-ai/deepseek-v4-flash Verified DeepSeek: DeepSeek V4 Flash 0423 | 87 | 1.0M | Up to 40 RPM | Online | ||
| minimax-m3 Verified minimax-m3 | 87 | 1.0M | Session/weekly limits (unpublished) | Online | ||
| Gemini 3.7 Flash Verified | 86 | 1.0M | 15 RPM, 1,500 RPD | Online | ||
| 86 | 262K | 200 req/day (free tier) | Online | |||
| qwen/qwen3.8-27b Paid | 86 | 8K | Online | |||
| qwen/qwen3.8-27b Paid | 86 | 8K | Online | |||
| qwen/qwen3.8-27b Paid | 86 | 8K | Online | |||
| 85 | 262K | ~200 req/hr | Online | |||
| minimaxai/minimax-m3 Verified MiniMax: MiniMax M3 | 85 | 1.0M | Up to 40 RPM | Online | ||
| Nex AGI: Nex-N2-Pro Paid Nex AGI: Nex-N2-Pro | 85 | 262K | 200 req/day (free tier) | Online | ||
| deepseek-ai/DeepSeek-V4-Pro-0813 Verified | 84 | 8K | Online | |||
| Qwen/Qwen3.8-27B Verified | 84 | 8K | Online | |||
| Gemini 3.5 Flash Verified Gemini 3.5 Flash | 84 | 1.0M | 15 RPM, 1,500 RPD | Online | ||
| NVIDIA: Nemotron 3 Ultra (free) Verified NVIDIA: Nemotron 3 Ultra (free) | 83 | 1.0M | 200 req/day (free tier) | Online | ||
| deepseek-v4-pro Paid deepseek-v4-pro | 82 | 128K | Session/weekly limits (unpublished) | Online | ||
| deepseek-ai/deepseek-v4-pro Verified deepseek-v4-pro | 82 | 1.0M | Up to 40 RPM | Online | ||
| agnes-2.0-flash Verified | 81 | 256K | 30 RPM | Online | ||
| MoonshotAI: Kimi K2.6 | 81 | 262K | 200 req/day (free tier) | Online | ||
| NVIDIA: Nemotron 3 Ultra (free) | 80 | 1.0M | ~200 req/hr | Online | ||
| moonshotai/kimi-k2.6 Verified MoonshotAI: Kimi K2.6 | 80 | 262K | Up to 40 RPM | Online | ||
| Z.ai: GLM 5.1 Paid | 79 | 205K | 200 req/day (free tier) | Online | ||
| Gemini 3.5 Flash-Lite Verified Gemini 3.5 Flash-Lite | 79 | 1.0M | 30 RPM, 1,500 RPD | Online | ||
| qwen/qwen3.6-27b Paid qwen/qwen3.6-27b | 78 | 131K | 30 RPM, 1,000 RPD | Online | ||
| qwen/qwen3.6-27b | 75 | 131K | 2 RPM (anonymous) | Online | ||
| 74 | 262K | ~200 req/hr | Online | |||
| Nemotron 3 Ultra 550B A55B Verified NVIDIA: Nemotron 3 Ultra (free) | 72 | 1.0M | Online | |||
| 71 | 1.0M | 200 req/day (free tier) | Online | |||
| nemotron-3-ultra Verified | 71 | 262K | Session/weekly limits (unpublished) | Online | ||
| MiniMax: MiniMax M3 | 70 | 205K | 200 req/day (free tier) | Online | ||
| Google: Gemma 4 31B (free) Verified Google: Gemma 4 31B (free) | 70 | 262K | 200 req/day (free tier) | Online | ||
| LiquidAI: LFM2.5-2.6B (free) Verified | 69 | 66K | 200 req/day (free tier) | Online | ||
| deepseek-ai/DeepSeek-V4-Pro Verified deepseek-v4-pro | 69 | 8K | Online | |||
| Poolside: Laguna S 2.1 (free) Verified Poolside: Laguna S 2.1 (free) | 69 | 262K | 200 req/day (free tier) | Online | ||
| Cohere: North Mini Code (free) Verified Cohere: North Mini Code (free) | 69 | 256K | 200 req/day (free tier) | Online | ||
| Google: Gemma 4 26B A4B (free) Verified Google: Gemma 4 26B A4B (free) | 68 | 262K | 200 req/day (free tier) | Online | ||
| Poolside: Laguna S 2.1 (free) | 67 | 262K | ~200 req/hr | Online | ||
| Cohere: North Mini Code (free) | 67 | 256K | ~200 req/hr | Online | ||
| Qwen3.5-397B-A17B | 67 | 131K | 2 RPM (anonymous) | Online | ||
| MiniMax-M2.7 | 67 | 128K | 20 RPM, 20 RPD, 200K TPD | Online | ||
| deepseek-ai/DeepSeek-V4-Flash Verified DeepSeek: DeepSeek V4 Flash 0423 | 66 | 8K | Online | |||
| qwen/qwen3.6-27b Paid | 66 | 8K | Online | |||
| qwen/qwen3.6-27b Paid | 66 | 8K | Online | |||
| qwen/qwen3.6-27b Paid | 66 | 8K | Online | |||
| qwen/qwen3.6-27b Paid | 66 | 8K | Online | |||
| qwen/qwen3.6-27b Paid | 66 | 8K | Online | |||
| qwen/qwen3.6-27b Paid | 66 | 8K | Online | |||
| Poolside: Laguna S 2.1 (free) | 66 | 262K | ~200 req/hr | Online | ||
| NVIDIA: Nemotron 3 Nano Omni (free) Verified NVIDIA: Nemotron 3 Nano Omni (free) | 66 | 256K | 200 req/day (free tier) | Online | ||
| mistral-large-3:675b Paid | 65 | 128K | Session/weekly limits (unpublished) | Online | ||
| poolside/laguna-xs-2.1 Verified Poolside: Laguna S 2.1 (free) | 65 | 262K | Up to 40 RPM | Online | ||
| gemma-4-31b Paid gemma-4-31b | 65 | 131K | 15 RPM, 30K TPM, 1M TPD | Online | ||
| NVIDIA: Nemotron 3 Super (free) Verified NVIDIA: Nemotron 3 Super (free) | 65 | 262K | 200 req/day (free tier) | Online | ||
| Poolside: Laguna XS 2.1 (free) Verified | 63 | 262K | 200 req/day (free tier) | Online | ||
| NVIDIA: Nemotron 3 Nano Omni (free) | 63 | 256K | ~200 req/hr | Online | ||
| NVIDIA: Nemotron 3 Super (free) | 63 | 262K | ~200 req/hr | Online | ||
| MiniMax-M2.7 Verified MiniMax-M2.7 | 62 | 205K | Online | |||
| GLM-4.7-Flash Verified GLM-4.7-Flash | 62 | 200K | 1 concurrent request | Online | ||
| 61 | 262K | 200 req/day (free tier) | Online | |||
| @cf/google/gemma-4-26b-a4b-it Verified Google: Gemma 4 26B A4B (free) | 61 | 256K | 10K neurons/day (shared) | Online | ||
| o4-mini | 61 | 200K | 10 RPM, 50 RPD | Online | ||
| 01-ai/yi-large Verified | 60 | 131K | Up to 40 RPM | Online | ||
| meta/codellama-70b Verified | 60 | 131K | Up to 40 RPM | Online | ||
| nvidia/llama3-chatqa-1.5-70b Verified | 60 | 131K | Up to 40 RPM | Online | ||
| writer/palmyra-fin-70b-32k Verified | 60 | 131K | Up to 40 RPM | Online | ||
| writer/palmyra-med-70b Verified | 60 | 131K | Up to 40 RPM | Online | ||
| writer/palmyra-med-70b-32k Verified | 60 | 131K | Up to 40 RPM | Online | ||
| Muse Glimmer 30B Verified | 60 | 131K | Online | |||
| @cf/moonshotai/kimi-k2.7-code Verified MoonshotAI: Kimi K2.6 | 60 | 262K | 10K neurons/day (shared) | Online | ||
| Gemma 4 31B IT Verified Google: Gemma 4 31B (free) | 60 | 262K | Online | |||
| Ling 3.0 Flash Fin (free) Verified | 59 | 262K | 200 req/day (free tier) | Online | ||
| MiniMax: MiniMax M2.7 (free) Verified | 59 | 197K | 200 req/day (free tier) | Online | ||
| 58 | 128K | 20 RPM, 20 RPD, 200K TPD | Online | |||
| 58 | 131K | See provider page | Online | |||
| 58 | 131K | See provider page | Online | |||
| Nemotron 3 Super 120B A12B Verified | 58 | 262K | Online | |||
| GLM-4.6V-Flash Verified GLM-4.6V-Flash | 58 | 128K | 1 concurrent request | Online | ||
| gpt-4.1 | 58 | 1.0M | 10 RPM, 50 RPD | Online | ||
| 57 | 1.0M | 200 req/day (free tier) | Online | |||
| Thinking Machines: Inkling (free) Verified | 57 | 1.0M | 200 req/day (free tier) | Online | ||
| 57 | 512K | 200 req/day (free tier) | Online | |||
| Qwen: Qwen3 Coder 480B A35B | 57 | 262K | 200 req/day (free tier) | Online | ||
| Qwen3.5-9B | 57 | 131K | 2 RPM (anonymous) | Online | ||
| OpenAI: gpt-oss-120b Paid OpenAI: gpt-oss-120b | 57 | 131K | 200 req/day (free tier) | Online | ||
| Gemini 3.1 Flash-Lite Verified Gemini 3.1 Flash-Lite | 57 | 1.0M | 30 RPM, 1,500 RPD | Online | ||
| Z.ai: GLM 5.2 (free) Verified | 56 | 256K | 200 req/day (free tier) | Online | ||
| ibm/granite-34b-code-instruct Verified | 56 | 131K | Up to 40 RPM | Online | ||
| Gemma 4 31B IT Verified Google: Gemma 4 31B (free) | 56 | 262K | Varies by model and account | Online | ||
| Aion 3.0 | 56 | 128K | 15 RPM, 20K TPD | Online | ||
| nvidia/llama-3.1-nemotron-ultra-253b-v1 | 56 | 131K | Up to 40 RPM | Online | ||
| qwen3.5:397b Paid qwen3.5:397b | 56 | 131K | Session/weekly limits (unpublished) | Online | ||
| Qwen: Qwen3 Coder 30B A3B Instruct | 56 | 262K | 200 req/day (free tier) | Online | ||
| agnes-1.5-flash Verified | 55 | 256K | 30 RPM | Online | ||
| Qwen/Qwen3.5-397B-A17B Verified Qwen3.5-397B-A17B | 55 | 8K | Online | |||
| NVIDIA: Nemotron 3 Nano Omni (free) | 55 | 256K | Online | |||
| Qwen: Qwen3 VL 235B A22B Instruct | 55 | 262K | 200 req/day (free tier) | Online | ||
| Aion 3.0 Mini | 55 | 128K | 15 RPM, 20K TPD | Online | ||
| Qwen/Qwen3.5-122B-A10B Verified Qwen/Qwen3.5-122B-A10B | 55 | 8K | Online | |||
| Ling 3.0 Flash Fin Verified | 54 | 262K | Online | |||
| 54 | 256K | See provider page | Online | |||
| 54 | 256K | 10 RPM, 50 RPD | Online | |||
| Nemotron 3 Super 120B A12B Verified NVIDIA: Nemotron 3 Super (free) | 54 | 262K | Online | |||
| Qwen: Qwen3 235B A22B | 54 | 131K | 200 req/day (free tier) | Online | ||
| gpt-5 | 54 | 200K | 10 RPM, 50 RPD | Online | ||
| GPT OSS 120B Paid | 53 | 131K | Online | |||
| google/diffusiongemma-26b-a4b-it Verified | 53 | 8K | Online | |||
| bigcode/starcoder2-15b Verified | 53 | 131K | Up to 40 RPM | Online | ||
| google/deplot Verified | 53 | 131K | Up to 40 RPM | Online | ||
| google/gemma-2b Verified | 53 | 131K | Up to 40 RPM | Online | ||
| google/recurrentgemma-2b Verified | 53 | 131K | Up to 40 RPM | Online | ||
| microsoft/kosmos-2 Verified | 53 | 131K | Up to 40 RPM | Online | ||
| microsoft/phi-3-vision-128k-instruct Verified | 53 | 131K | Up to 40 RPM | Online | ||
| microsoft/phi-3.5-moe-instruct Verified | 53 | 131K | Up to 40 RPM | Online | ||
| mistralai/mixtral-8x22b-v0.1 Verified | 53 | 131K | Up to 40 RPM | Online | ||
| 53 | 131K | Up to 40 RPM | Online | |||
| nvidia/nemotron-4-340b-instruct Verified | 53 | 131K | Up to 40 RPM | Online | ||
| nvidia/nemotron-4-340b-reward Verified | 53 | 131K | Up to 40 RPM | Online | ||
| nvidia/nemotron-parse Verified | 53 | 131K | Up to 40 RPM | Online | ||
| nvidia/neva-22b Verified | 53 | 131K | Up to 40 RPM | Online | ||
| nvidia/nvclip Verified | 53 | 131K | Up to 40 RPM | Online | ||
| nvidia/riva-translate-4b-instruct Verified | 53 | 131K | Up to 40 RPM | Online | ||
| nvidia/vila Verified | 53 | 131K | Up to 40 RPM | Online | ||
| writer/palmyra-creative-122b Verified | 53 | 131K | Up to 40 RPM | Online | ||
| Gemma 4 26B A4B IT Verified Google: Gemma 4 26B A4B (free) | 53 | 262K | Varies by model and account | Online | ||
| Qwen: Qwen3 Coder 30B A3B Instruct | 53 | 262K | 2 RPM (anonymous) | Online | ||
| Qwen: Qwen3 Next 80B A3B Instruct | 53 | 262K | 200 req/day (free tier) | Online | ||
| Z.ai: GLM 4.5 Air Paid Z.ai: GLM 4.5 Air | 53 | 131K | 200 req/day (free tier) | Online | ||
| @cf/openai/gpt-oss-120b Verified | 52 | 128K | 10K neurons/day (shared) | Online | ||
| 52 | 256K | ~1 RPS, 500K TPM | Online | |||
| stepfun-ai/Step-3.7-Flash Verified | 52 | 8K | Online | |||
| GPT OSS 120B Paid | 52 | 131K | Online | |||
| GPT OSS 120B Paid | 52 | 131K | Online | |||
| GPT OSS 120B Paid | 52 | 131K | Online | |||
| GPT OSS 120B Paid | 52 | 131K | Online | |||
| Venice: Uncensored Paid | 52 | 128K | 200 req/day (free tier) | Online | ||
| adept/fuyu-8b Verified | 52 | 131K | Up to 40 RPM | Online | ||
| aisingapore/sea-lion-7b-instruct Verified | 52 | 131K | Up to 40 RPM | Online | ||
| 52 | 131K | Up to 40 RPM | Online | |||
| google/codegemma-1.1-7b Verified | 52 | 131K | Up to 40 RPM | Online | ||
| google/codegemma-7b Verified | 52 | 131K | Up to 40 RPM | Online | ||
| ibm/granite-3.0-8b-instruct Verified | 52 | 131K | Up to 40 RPM | Online | ||
| ibm/granite-8b-code-instruct Verified | 52 | 131K | Up to 40 RPM | Online | ||
| 52 | 131K | Up to 40 RPM | Online | |||
| 52 | 131K | Up to 40 RPM | Online | |||
| zyphra/zamba2-7b-instruct Verified | 52 | 131K | Up to 40 RPM | Online | ||
| agnes-image-2.0-flash Verified | 52 | 4K | 30 RPM (1K) | Online | ||
| agnes-image-2.1-flash Verified | 52 | 4K | 30 RPM (1K) | Online | ||
| NVIDIA: Nemotron 3.5 Content Safety (free) | 52 | 128K | 200 req/day (free tier) | Online | ||
| gpt-oss-120b Paid gpt-oss-120b | 52 | 131K | 5 RPM, 30K TPM, 1M TPD | Online | ||
| Qwen: Qwen3 VL 8B Thinking | 52 | 131K | 200 req/day (free tier) | Online | ||
| Qwen: Qwen3 VL 8B Instruct | 52 | 262K | 200 req/day (free tier) | Online | ||
| NVIDIA: Nemotron 3 Nano 30B A3B (free) | 52 | 262K | 200 req/day (free tier) | Online | ||
| Qwen: Qwen3 235B A22B Thinking 2507 | 52 | 131K | 200 req/day (free tier) | Online | ||
| groq/compound Verified | 51 | 131K | 30 RPM, 250 RPD | Online | ||
| 51 | 256K | ~1 RPS, 500K TPM | Online | |||
| Laguna S 2.1 Verified | 51 | 1.0M | Online | |||
| nvidia/cosmos-reason2-8b Verified | 51 | 131K | Up to 40 RPM | Online | ||
| gemma-4-31b | 51 | 128K | 20 RPM, 20 RPD, 200K TPD | Online | ||
| Gemini 2.5 Flash Verified Gemini 2.5 Flash | 51 | 1.0M | 15 RPM, 1,500 RPD | Online | ||
| Gemini 2.5 Pro Verified Gemini 2.5 Pro | 51 | 1.0M | 5 RPM, 50 RPD | Online | ||
| mistralai/mistral-large-3-675b-instruct-2512 | 51 | 8K | Online | |||
| groq/compound-mini Verified | 50 | 131K | 30 RPM, 250 RPD | Online | ||
| Nemotron 3 Ultra 550B A55B Verified | 50 | 1.0M | Online | |||
| OpenAI: gpt-oss-120b | 50 | 131K | 200 req/day (free tier) | Online | ||
| nvidia/nemotron-3.5-content-safety Verified NVIDIA: Nemotron 3.5 Content Safety (free) | 50 | 128K | Up to 40 RPM | Online | ||
| Llama-4-Scout-17B-16E-Instruct | 50 | 512K | 15 RPM, 150 RPD | Online | ||
| DeepSeek-V3.1 | 50 | 128K | 20 RPM, 20 RPD, 200K TPD | Online | ||
| Qwen: Qwen3 Next 80B A3B Thinking | 50 | 262K | 200 req/day (free tier) | Online | ||
| Google: Lyria 3 Pro Preview Verified Google: Lyria 3 Pro Preview | 50 | 1.0M | 200 req/day (free tier) | Online | ||
| Google: Lyria 3 Clip Preview Verified Google: Lyria 3 Clip Preview | 50 | 1.0M | 200 req/day (free tier) | Online | ||
| Qwen: Qwen3 30B A3B Paid Qwen: Qwen3 30B A3B | 50 | 131K | 200 req/day (free tier) | Online | ||
| gpt-4o | 50 | 128K | 10 RPM, 50 RPD | Online | ||
| GLM-4.5-Flash Verified GLM-4.5-Flash | 50 | 128K | 1 concurrent request | Online | ||
| 49 | 128K | 15 RPM, 150 RPD | Online | |||
| agnes-video-v2.0 Verified | 49 | 4K | 2 RPM | Online | ||
| Gemma 4 31B IT Verified | 49 | 262K | Online | |||
| GPT OSS 120B Verified | 49 | 131K | Online | |||
| Gemini 3.1 Flash Lite Verified Gemini 3.1 Flash-Lite | 49 | 1.0M | Online | |||
| mistralai/codestral-22b-instruct-v0.1 | 49 | 131K | Up to 40 RPM | Online | ||
| DeepSeek-R1 | 49 | 64K | 15 RPM, 150 RPD | Online | ||
| gpt-oss:20b Verified gpt-oss:20b | 49 | 131K | Session/weekly limits (unpublished) | Online | ||
| Qwen2.5-VL-72B-Instruct | 49 | 128K | 2 RPM (anonymous) | Online | ||
| Mistral-Small-3.2-24B-Instruct | 49 | 128K | 2 RPM (anonymous) | Online | ||
| Mistral-Nemo-Instruct-2407 | 49 | 128K | 2 RPM (anonymous) | Online | ||
| gpt-4.1-mini | 49 | 1.0M | 15 RPM, 150 RPD | Online | ||
| 48 | 256K | ~1 RPS, 500K TPM | Online | |||
| MiMo-V2.5 Verified | 48 | 1.0M | Online | |||
| mistralai/mistral-7b-instruct-v0.3 Verified | 48 | 131K | Up to 40 RPM | Online | ||
| Gemini 3.1 Flash Lite Verified Gemini 3.1 Flash-Lite | 48 | 1.0M | 30 RPM, 1,500 RPD | Online | ||
| Qwen: Qwen3 32B Paid Qwen: Qwen3 32B | 48 | 131K | 200 req/day (free tier) | Online | ||
| stepfun-ai/Step-3.5-Flash Verified | 47 | 8K | Online | |||
| GPT OSS 120B Verified | 47 | 131K | Online | |||
| @cf/nvidia/nemotron-3-120b-a12b Verified | 47 | 8K | Online | |||
| @cf/baai/bge-large-en-v1.5 Verified | 47 | 8K | Online | |||
| 47 | 10.0M | 10K neurons/day (shared) | Online | |||
| 47 | 256K | ~1 RPS, 500K TPM | Online | |||
| 47 | 256K | ~1 RPS, 500K TPM | Online | |||
| GPT OSS 20B Paid | 47 | 131K | Online | |||
| 47 | 436K | 20 RPM | Online | |||
| Qwen: Qwen3 30B A3B Thinking 2507 | 47 | 82K | 200 req/day (free tier) | Online | ||
| Qwen: Qwen3 14B Paid Qwen: Qwen3 14B | 47 | 131K | 200 req/day (free tier) | Online | ||
| Gemini 2.5 Flash-Lite Verified Gemini 2.5 Flash-Lite | 47 | 1.0M | 30 RPM, 1,500 RPD | Online | ||
| Tencent-Hunyuan/Hy3 Verified | 46 | 8K | Online | |||
| 46 | 32K | 2 RPM (anonymous) | Online | |||
| 46 | 288K | 20 RPM | Online | |||
| 46 | 288K | 20 RPM | Online | |||
| DeepSeek V4 Flash 0731 Verified | 46 | 1.0M | Online | |||
| Mixtral 8x7B Verified | 46 | 33K | Unlimited for free models | Online | ||
| GLM-4.7-FlashX Verified | 46 | 200K | Online | |||
| 46 | 256K | ~1 RPS, 500K TPM | Online | |||
| ibm/granite-3.0-3b-a800m-instruct Verified | 46 | 131K | Up to 40 RPM | Online | ||
| GPT OSS 120B Paid | 46 | 131K | Online | |||
| GPT OSS 120B Paid | 46 | 131K | Online | |||
| GPT OSS 120B Paid | 46 | 131K | Online | |||
| GPT OSS 120B Paid | 46 | 131K | Online | |||
| GPT OSS 120B Paid | 46 | 131K | Online | |||
| GPT OSS 120B Paid | 46 | 131K | Online | |||
| 46 | 128K | 20 RPM | Online | |||
| 46 | 128K | 20 RPM | Online | |||
| 46 | 128K | 20 RPM | Online | |||
| 46 | 128K | 20 RPM | Online | |||
| Aion 2.0 | 46 | 128K | 15 RPM, 20K TPD | Online | ||
| DeepSeek V4 Flash Verified | 45 | 1.0M | Online | |||
| 45 | 131K | See provider page | Online | |||
| GPT OSS 20B Paid | 45 | 131K | Online | |||
| GPT OSS 20B Paid | 45 | 131K | Online | |||
| GPT OSS 20B Paid | 45 | 131K | Online | |||
| GPT OSS 20B Paid | 45 | 131K | Online | |||
| 45 | 128K | 15 RPM, 20K TPD | Online | |||
| mistralai/mistral-large-2-instruct Verified | 45 | 131K | Up to 40 RPM | Online | ||
| Free Models Router Verified | 45 | 200K | 200 req/day (free tier) | Online | ||
| nvidia/nemotron-nano-3-30b-a3b Verified nvidia/llama-3.1-nemotron-ultra-253b-v1 | 45 | 131K | Up to 40 RPM | Online | ||
| Qwen: Qwen3 32B | 45 | 131K | 2 RPM (anonymous) | Online | ||
| @cf/mistralai/mistral-small-3.1-24b-instruct | 45 | 128K | 10K neurons/day (shared) | Online | ||
| Qwen: Qwen3 8B Paid Qwen: Qwen3 8B | 45 | 131K | 200 req/day (free tier) | Online | ||
| 44 | 128K | 20 RPM | Online | |||
| whisper-large-v3 Verified | 44 | 131K | 20 RPM, 2,000 RPD | Online | ||
| whisper-large-v3-turbo Verified | 44 | 131K | 20 RPM, 2,000 RPD | Online | ||
| Mistral 7B Verified | 44 | 33K | See provider page | Online | ||
| 44 | 33K | See provider page | Online | |||
| GLM-5.1 Verified | 44 | 1.0M | Online | |||
| 44 | 9K | 20 RPM | Online | |||
| 44 | 131K | 30 RPM, 60K TPM | Online | |||
| meta/llama-3.2-90b-vision-instruct Verified | 44 | 8K | Online | |||
| mistralai/mistral-nemotron Verified | 44 | 8K | Online | |||
| 44 | 8K | Online | ||||
| nvidia/ising-calibration-1.5-31b Verified | 44 | 8K | Online | |||
| nvidia/riva-translate-4b-instruct-v2 Verified | 44 | 8K | Online | |||
| 44 | 8K | Online | ||||
| MiniMax-M3 Verified | 44 | 512K | Online | |||
| DeepSeek-R1 | 44 | 131K | Community-powered, no hard cap | Online | ||
| Llama-3.3-70B-Instruct | 44 | 131K | 15 RPM, 150 RPD | Online | ||
| 43 | 8K | Online | ||||
| 43 | 8K | Online | ||||
| 43 | 8K | Online | ||||
| Qwen/Qwen3.8-Flash-Next Verified | 43 | 8K | Online | |||
| 43 | 16K | 20 RPM | Online | |||
| qwen/qwen3.6-27b Paid | 43 | 8K | Online | |||
| qwen/qwen3.6-27b Paid | 43 | 8K | Online | |||
| qwen/qwen3.6-27b Paid | 43 | 8K | Online | |||
| qwen/qwen3.6-27b Paid | 43 | 8K | Online | |||
| 43 | 33K | See provider page | Online | |||
| @cf/zai-org/glm-4.7-flash Verified | 43 | 8K | Online | |||
| nvidia/llama-3.1-nemotron-ultra-253b-v1 | 43 | 131K | Up to 40 RPM | Online | ||
| Meta-Llama-3_3-70B-Instruct Verified Llama-3.3-70B-Instruct | 43 | 131K | 2 RPM (anonymous) | Online | ||
| Phi-4 | 43 | 131K | See provider page | Online | ||
| Gemma 4 31B IT Verified | 42 | 262K | Online | |||
| 42 | 32K | 15 RPM, 20K TPD | Online | |||
| Qwen/Qwen3.5-27B Verified | 42 | 8K | Online | |||
| Qwen/Qwen3.5-35B-A3B Verified | 42 | 8K | Online | |||
| ai21labs/jamba-1.5-large-instruct Verified | 42 | 131K | Up to 40 RPM | Online | ||
| GPT OSS 20B Verified | 42 | 131K | Online | |||
| nvidia/llama-nemotron-embed-vl-1b-v2 Verified nvidia/llama-3.1-nemotron-ultra-253b-v1 | 42 | 131K | Up to 40 RPM | Online | ||
| meta/llama-3.2-11b-vision-instruct Verified meta/llama-3.2-11b-vision-instruct | 42 | 131K | Up to 40 RPM | Online | ||
| Qwen/Qwen3-235B-A22B-Instruct-2507 Verified Qwen/Qwen3-235B-A22B-Instruct-2507 | 42 | 8K | Online | |||
| GLM-4.5-Air Verified GLM-4.5-Air | 42 | 131K | Online | |||
| 41 | 33K | See provider page | Online | |||
| 41 | 131K | 200 req/day (free tier) | Online | |||
| 41 | 128K | ~1 RPS, 500K TPM | Online | |||
| MedAIBase/AntAngelMed Verified | 41 | 8K | Online | |||
| MusePublic/Qwen-Image-Edit Verified | 41 | 8K | Online | |||
| OpenGVLab/InternVL3_5-241B-A28B Verified | 41 | 8K | Online | |||
| PaddlePaddle/ERNIE-4.5-21B-A3B-PT Verified | 41 | 8K | Online | |||
| PaddlePaddle/ERNIE-4.5-300B-A47B-PT Verified | 41 | 8K | Online | |||
| PaddlePaddle/ERNIE-4.5-VL-28B-A3B-PT Verified | 41 | 8K | Online | |||
| Qwen/Qwen-Image-Edit Verified | 41 | 8K | Online | |||
| Qwen/Qwen3-4B Verified | 41 | 8K | Online | |||
| Shanghai_AI_Laboratory/Intern-S1 Verified | 41 | 8K | Online | |||
| 41 | 8K | Online | ||||
| early-access/EA-29B-A4B Verified | 41 | 8K | Online | |||
| @cf/deepseek-ai/deepseek-r1-distill-qwen-32b | 41 | 32K | 10K neurons/day (shared) | Online | ||
| nvidia/embed-qa-4 Verified | 40 | 131K | Up to 40 RPM | Online | ||
| 40 | 131K | Up to 40 RPM | Online | |||
| nvidia/llama-3.2-nv-embedqa-1b-v1 Verified | 40 | 131K | Up to 40 RPM | Online | ||
| nvidia/nemotron-3-embed-1b Verified | 40 | 131K | Up to 40 RPM | Online | ||
| nvidia/nv-embedqa-mistral-7b-v2 Verified | 40 | 131K | Up to 40 RPM | Online | ||
| snowflake/arctic-embed-l Verified | 40 | 131K | Up to 40 RPM | Online | ||
| databricks/dbrx-instruct Verified | 40 | 131K | Up to 40 RPM | Online | ||
| meta/llama2-70b Verified | 40 | 131K | Up to 40 RPM | Online | ||
| @cf/baai/bge-m3 Verified | 40 | 8K | Online | |||
| @cf/google/gemma-2b-it-lora Verified | 40 | 8K | Online | |||
| @cf/moonshotai/kimi-k2.6 Verified | 40 | 8K | Online | |||
| @cf/ibm-granite/granite-4.0-h-micro Verified | 40 | 8K | Online | |||
| @cf/baai/bge-small-en-v1.5 Verified | 40 | 8K | Online | |||
| @cf/zai-org/glm-5.2 Verified | 40 | 8K | Online | |||
| @cf/baai/bge-base-en-v1.5 Verified | 40 | 8K | Online | |||
| 40 | 8K | Online | ||||
| @cf/openai/gpt-oss-20b Verified | 40 | 8K | Online | |||
| @cf/moondream/moondream3.1-9B-A2B Verified | 40 | 8K | Online | |||
| mistral-Nemo-Instruct-2407 Verified | 40 | 8K | Online | |||
| @cf/qwen/qwen3.8-27b Verified | 40 | 8K | Online | |||
| meituan-longcat/LongCat-Flash-Lite Verified | 40 | 8K | Online | |||
| Qwen/Qwen3-Coder-30B-A3B-Instruct Verified Qwen: Qwen3 Coder 30B A3B Instruct | 40 | 8K | Online | |||
| Qwen/Qwen3-VL-235B-A22B-Instruct Verified Qwen: Qwen3 VL 235B A22B Instruct | 40 | 8K | Online | |||
| Llama-3.3-70B-Instruct | 40 | 131K | 10K neurons/day (shared) | Online | ||
| Llama 3.1 70B | 40 | 131K | See provider page | Online | ||
| Llama 3.1 70B | 40 | 131K | Community-powered, no hard cap | Online | ||
| Qwen2.5-7B-Instruct | 40 | 131K | Credit-metered | Online | ||
| GPT OSS 20B Paid | 39 | 131K | Online | |||
| GPT OSS 20B Paid | 39 | 131K | Online | |||
| GPT OSS 20B Paid | 39 | 131K | Online | |||
| GPT OSS 20B Paid | 39 | 131K | Online | |||
| GPT OSS 20B Paid | 39 | 131K | Online | |||
| GPT OSS 20B Paid | 39 | 131K | Online | |||
| 39 | 131K | $25/month free credits, resets monthly | Online | |||
| 39 | 8K | Online | ||||
| @cf/qwen/qwq-32b Verified | 39 | 8K | Online | |||
| 39 | 8K | Online | ||||
| 39 | 8K | Online | ||||
| @cf/google/gemma-7b-it-lora Verified | 39 | 8K | Online | |||
| big-pickle Verified | 39 | N/A | Online | |||
| Qwen/Qwen3-235B-A22B Verified Qwen: Qwen3 235B A22B | 39 | 8K | Online | |||
| mistralai/mistral-large-3-675b-instruct-2512 | 39 | 131K | See provider page | Online | ||
| Llama 3.1 70B | 39 | 131K | Unlimited for free models | Online | ||
| @cf/qwen/qwen2.5-coder-32b-instruct Verified @cf/qwen/qwen2.5-coder-32b-instruct | 39 | 8K | Online | |||
| Codestral (latest) Verified | 38 | 256K | Online | |||
| 38 | 131K | Online | ||||
| allam-2-7b Verified | 38 | 8K | Online | |||
| Qwen/Qwen3-Next-80B-A3B-Instruct Verified Qwen: Qwen3 Next 80B A3B Instruct | 38 | 8K | Online | |||
| meta/llama-guard-4-12b Verified meta/llama-guard-4-12b | 38 | 1.0M | Up to 40 RPM | Online | ||
| Nemotron 3.5 Lightning 30B A3B Verified | 37 | 262K | Online | |||
| Nemotron 3 Nano 30B A3B Verified | 37 | 262K | Online | |||
| MiniMax/MiniMax-M1-80k Verified | 37 | 8K | Online | |||
| Qwen/Qwen3-VL-8B-Thinking Verified Qwen: Qwen3 VL 8B Thinking | 37 | 8K | Online | |||
| Qwen/Qwen3-VL-8B-Instruct Verified Qwen: Qwen3 VL 8B Instruct | 36 | 8K | Online | |||
| Qwen/Qwen3-235B-A22B-Thinking-2507 Verified Qwen: Qwen3 235B A22B Thinking 2507 | 36 | 8K | Online | |||
| GPT OSS 20B Verified gpt-oss:20b | 36 | 131K | Online | |||
| deepseek-v4-flash Verified | 35 | 0 | See provider page | Online | ||
| glm-5.3-flash Verified | 35 | 0 | See provider page | Online | ||
| Qwen/Qwen3-30B-A3B Verified Qwen: Qwen3 30B A3B | 35 | 8K | Online | |||
| mistralai/Mistral-Large-Instruct-2407 | 35 | 8K | Online | |||
| 34 | 131K | Credit-metered | Online | |||
| PaddlePaddle/ERNIE-4.5-0.3B-PT Verified | 34 | 8K | Online | |||
| Qwen/Qwen3-Next-80B-A3B-Thinking Verified Qwen: Qwen3 Next 80B A3B Thinking | 34 | 8K | Online | |||
| @cf/qwen/qwen3-30b-a3b-fp8 Verified Qwen: Qwen3 30B A3B | 34 | 8K | Online | |||
| Meta-Llama-3.1-8B-Instruct Verified Meta-Llama-3.1-8B-Instruct | 34 | 128K | Credit-metered | Online | ||
| 33 | 128K | Online | ||||
| gemma-3-4b-it | 33 | 131K | Credit-metered | Online | ||
| laguna-s-2.1:free Verified | 32 | 0 | See provider page | Online | ||
| LongCat-2.0 Verified | 32 | 0 | See provider page | Online | ||
| Qwen/Qwen3-30B-A3B-Thinking-2507 Verified Qwen: Qwen3 30B A3B Thinking 2507 | 32 | 8K | Online | |||
| Qwen/Qwen3-14B Verified Qwen: Qwen3 14B | 32 | 8K | Online | |||
| 31 | 131K | $25/month free credits, resets monthly | Online | |||
| 31 | 8K | Online | ||||
| 31 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| 30 | 8K | Online | ||||
| Qwen/Qwen3-8B Verified Qwen: Qwen3 8B | 30 | 8K | Online | |||
| @cf/meta/llama-3.2-3b-instruct Verified @cf/meta/llama-3.2-3b-instruct | 29 | 8K | Online | |||
| @cf/meta/llama-3.1-8b-instruct-fp8 Verified Meta-Llama-3.1-8B-Instruct | 28 | 8K | Online | |||
| @cf/meta/llama-guard-3-8b Verified | 27 | 8K | Online | |||
| @cf/qwen/qwen3-embedding-0.6b Verified | 27 | 8K | Online | |||
| @cf/pfnet/plamo-embedding-1b Verified | 27 | 8K | Online | |||
| @cf/google/embeddinggemma-300m Verified | 27 | 8K | Online | |||
| @cf/meta/llama-3.2-1b-instruct Verified @cf/meta/llama-3.2-1b-instruct | 27 | 8K | Online |
How to Get Started with Free LLM APIs
- Pick a free LLM model — Click any model name to see details, rate limits, and API key signup link.
- Get your API key — Sign up on the provider's website (most require no credit card).
- Copy the config — Go to the Config Generator, pick your tool and backend, copy the ready-to-use snippet.
- Test it — Use the Playground to test your API key before integrating.
New to LLM terminology? Check the 📖 Glossary — 22 terms explained in plain English →
FAQ: Common questions about free LLM APIs →About This Free LLM API Directory
Finding reliable free LLM API resources online can be frustrating. Many developers traditionally rely on static GitHub repositories to find endpoints. While those lists are a good starting point, they often become outdated quickly, leaving you with dead links, expired API keys, and unverified rate limits.
That's why we built this dynamic, auto-updating directory. If you are looking for a reliable alternative to GitHub free LLM API lists, this page tracks over 410 free LLM models online in real-time. Whether you need a free API key for text generation, vision, or coding tasks, you can compare context windows, capabilities, and strict rate limit data side-by-side.
Our goal is to be the most accurate and comprehensive list of free AI APIs for developers. Use the filters above to find providers that don't require credit cards or phone verification, and grab your free API keys to start building immediately.