424+ free LLM APIs from 30 providers — find, compare & configure free models in seconds.
🌐 Live at freellm.net — Browse models · Playground · Config generator · API keys
🔄 Data refreshed daily from freellm.net — Last updated: 2026-08-06
🌐 English · 简体中文 · 繁體中文 · 日本語 · 한국어
Finding a free LLM API shouldn't mean hunting through a dozen GitHub READMEs, signing up for five different platforms, or guessing which models still have a free tier.
This repo is a structured, machine-readable directory of every free LLM API — rate limits, context windows, one-click config snippets, and direct API key links. Updated daily.
Why this repo + freellm.net:
- ✅ Always up-to-date — data refreshed daily via automated monitoring, not a 2-year-old static list
- ✅ Credit card transparency — clearly shows which providers require a card, phone verification, or nothing at all
- ✅ One-click configs — ready-to-copy snippets for Claude Code, Cursor, Codex, Aider, and 10+ more tools
- ✅ Side-by-side comparison — compare context windows, rate limits, and modalities across providers instantly
- Pick a provider — see Provider Directory below. Start with Groq (no credit card, 30 RPM free).
- Get your API key — click any Get Key → link below, sign up (most need just an email), and copy your key. Takes < 1 minute.
- Plug it in — copy the base URL + model ID, paste into the Quick Start examples below.
Configuring a specific tool? Claude Code · Cursor · Codex · OpenHuman · OpenCode · OpenClaw — one-click configs at freellm.net/config/.
Never used an API before? Here's the simplest path: go to console.groq.com/keys, sign up with just an email (no credit card), copy your free key, and paste it into any example below. You'll be running in under a minute.
All providers below expose an OpenAI-compatible endpoint. Any tool that accepts baseURL + apiKey works — just swap the base URL and key.
from openai import OpenAI
client = OpenAI(
base_url="https://api.groq.com/openai/v1", # free, no credit card
api_key="GROQ_API_KEY", # get at console.groq.com/keys
)
response = client.chat.completions.create(
model="llama-3.3-70b-versatile", # see Best Models table below
messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
# Groq free tier: 30 RPM, 14,400 RPD — generous for personal useexport OPENAI_BASE_URL="https://api.groq.com/openai/v1"
export OPENAI_API_KEY="your-groq-key" # get at console.groq.com/keys
codex --model "llama-3.3-70b-versatile"Settings → Models → Add Model
Model name: llama-3.3-70b-versatile
Base URL: https://api.groq.com/openai/v1
API key: your-groq-key # get at console.groq.com/keys
# Claude Code needs an Anthropic-compatible API — use OpenRouter
export ANTHROPIC_BASE_URL="https://openrouter.ai/api"
export ANTHROPIC_AUTH_TOKEN="sk-or-v1-your-key" # openrouter.ai/keys
export ANTHROPIC_API_KEY="" # must be empty
# Note: OpenRouter Anthropic models need $10 top-up (one-time)Most AI dev tools accept custom API endpoints — just point them at any provider above. Grab your free key, then:
- Claude Code — set
ANTHROPIC_BASE_URL+ANTHROPIC_AUTH_TOKEN. Step-by-step → - Cursor — Settings → Models → Add Model. Step-by-step →
- Codex CLI — set
OPENAI_BASE_URL+OPENAI_API_KEY. Step-by-step → - OpenHuman — edit
config.toml. Step-by-step → - Aider — edit
.aider.conf.yml. Step-by-step → - Cline (VS Code) — API provider settings. Step-by-step →
- Open WebUI — Settings → Connections. Step-by-step →
More ready-to-copy configs at freellm.net/config/.
All providers, base URLs, and API key links are in the Quick Reference below.
These providers offer a permanently free tier — no credit card required for most.
| Provider | Free Models | Credit Card? | Max Context | Modalities | Get API Key |
|---|---|---|---|---|---|
| NVIDIA NIM | 123 | Phone verification | 1M | audio, embedding, image, reasoning, rerank, text, video, vision | → |
| ModelScope | 53 | Registration | 1M | audio, image, reasoning, text, video, vision | → |
| Cloudflare Workers AI | 39 | No | 10M | code, image, reasoning, text, video | → |
| GitHub Models | 16 | No | 1M | image, pdf, reasoning, text | → |
| Google Gemini | 15 | No | 1M | audio, image, pdf, reasoning, text, video, vision | → |
| Ollama Cloud | 14 | Registration | 1M | audio, code, image, reasoning, text, video | → |
| OVHcloud AI Endpoints | 14 | Registration | 262K | audio, code, image, reasoning, text, video | → |
| LLM7.io | 13 | No | 1M | audio, code, image, pdf, reasoning, text, video, vision | → |
| Groq | 12 | No | 262K | image, reasoning, text | → |
| Mistral AI | 12 | No | 256K | code, image, text | → |
| Cohere | 12 | No | 436K | image, text | → |
| Kilo Code | 12 | No | 1M | audio, code, image, reasoning, text, video | → |
| Cerebras | 8 | No | 131K | image, reasoning, text | → |
| OpenCode Zen | 8 | Registration | 1M | audio, reasoning, vision | → |
| Aion Labs | 7 | Registration | 131K | text | → |
| Hugging Face | 7 | No | 131K | code, text | → |
| Agnes AI | 5 | Registration | 256K | image, text, video, vision | → |
| Alibaba Cloud Model Studio | 5 | Registration | 1M | code, image, text | → |
| Z AI (Zhipu AI) | 4 | No | 200K | image, reasoning, text, video | → |
| SambaNova | 4 | Registration | 128K | image, reasoning, text | → |
| SiliconFlow | 3 | Registration | 131K | text | → |
| xAI | 3 | Registration | 2M | text | → |
| Chutes.ai | 2 | Registration | 131K | reasoning, text | → |
| Glhf.chat | 2 | Registration | 131K | text | → |
| Grok (xAI) | 2 | Registration | 131K | text | → |
| AI21 Labs | 2 | Registration | 256K | text | → |
| DeepSeek | 2 | Registration | 128K | text | → |
| Nscale | 2 | Registration | 128K | text | → |
| Nebius | 1 | Registration | 128K | text | → |
Providers that periodically renew free credits.
| Provider | Free Models | Credit Model | Max Context | Modalities | Get API Key |
|---|---|---|---|---|---|
| OpenRouter | 22 | Free tier + $10 topup → 1K RPD | 1M | audio, code, embeddings, image, reasoning, rerank, speech, text, video | → |
| Provider | Base URL | Get API Key | Credit Card? |
|---|---|---|---|
| NVIDIA NIM | https://integrate.api.nvidia.com/v1 |
Get Key → | Phone verification |
| ModelScope | https://api-inference.modelscope.cn/v1 |
Get Key → | Registration |
| Cloudflare Workers AI | https://api.cloudflare.com/client/v4/accounts/{account_id}/ai/run |
Get Key → | No |
| OpenRouter | https://openrouter.ai/api/v1 |
Get Key → | Registration |
| GitHub Models | https://models.github.ai/inference |
Get Key → | No |
| Google Gemini | https://generativelanguage.googleapis.com/v1beta |
Get Key → | No |
| Ollama Cloud | https://api.ollama.com |
Get Key → | Registration |
| OVHcloud AI Endpoints | https://oai.endpoints.kepler.ai.cloud.ovh.net/v1 |
Get Key → | Registration |
| LLM7.io | https://api.llm7.io/v1 |
Get Key → | No |
| Groq | https://api.groq.com/openai/v1 |
Get Key → | No |
| Mistral AI | https://api.mistral.ai/v1 |
Get Key → | No |
| Cohere | https://api.cohere.com/v2 |
Get Key → | No |
| Kilo Code | https://api.kilo.ai/api/gateway |
Get Key → | No |
| Cerebras | https://api.cerebras.ai/v1 |
Get Key → | No |
| OpenCode Zen | https://opencode.ai/zen/v1 |
Get Key → | Registration |
| Aion Labs | https://api.aionlabs.ai/v1 |
Get Key → | Registration |
| Hugging Face | https://router.huggingface.co/v1 |
Get Key → | No |
| Agnes AI | https://apihub.agnes-ai.com/v1 |
Get Key → | Registration |
| Alibaba Cloud Model Studio | https://dashscope-intl.aliyuncs.com/compatible-mode/v1 |
Get Key → | Registration |
| Z AI (Zhipu AI) | https://open.bigmodel.cn/api/paas/v4 |
Get Key → | No |
| SambaNova | https://api.sambanova.ai/v1 |
Get Key → | Registration |
| SiliconFlow | https://api.siliconflow.cn/v1 |
Get Key → | Registration |
| xAI | https://api.x.ai/v1 |
Get Key → | Registration |
| Chutes.ai | https://api.chutes.ai/v1 |
Get Key → | Registration |
| Glhf.chat | https://glhf.chat/api/openai/v1 |
Get Key → | Registration |
| Grok (xAI) | https://api.x.ai/v1 |
Get Key → | Registration |
| AI21 Labs | https://api.ai21.com/studio/v1 |
Get Key → | Registration |
| DeepSeek | https://api.deepseek.com/v1 |
Get Key → | Registration |
| Nscale | https://inference.api.nscale.com/v1 |
Get Key → | Registration |
| Nebius | https://api.studio.nebius.com/v1 |
Get Key → | Registration |
| Provider | Best Free Model | Model ID | Max Context | Rate Limit |
|---|---|---|---|---|
| NVIDIA NIM | z-ai/glm-5.2 | z-ai/glm-5.2 |
1M | Up to 40 RPM |
| poolside/laguna-xs-2.1 | poolside/laguna-xs-2.1 |
262K | Up to 40 RPM | |
| z-ai/glm-5.1 | z-ai/glm-5.1 |
202K | Up to 40 RPM | |
| ModelScope | MiniMax-M2.5-highspeed | MiniMax/MiniMax-M2.5 |
204K | See provider |
| Qwen/Qwen3.5-35B-A3B | qwen-qwen3-5-35b-a3b |
131K | 2,000 RPD total; <=500 .. | |
| Qwen/Qwen3.5-27B | qwen-qwen3-5-27b |
131K | 2,000 RPD total; <=500 .. | |
| Cloudflare Workers AI | Mistral 7B | @cf/mistral/mistral-7b-instruct-v0.1 |
32K | See provider |
| Qwen 1.5 7B | @cf/qwen/qwen1.5-7b-chat |
32K | See provider | |
| @cf/meta/llama-3.3-70b-instruct-fp8-fast | @cf/meta/llama-3.3-70b-instruct-fp8-fast |
131K | 10K neurons/day (shared) | |
| OpenRouter | NVIDIA: Nemotron 3 Ultra (free) | nvidia/nemotron-3-ultra-550b-a55b:free |
1M | See provider |
| Poolside: Laguna M.1 (free) | poolside/laguna-m.1:free |
262K | See provider | |
| NVIDIA: Nemotron 3 Super (free) | nvidia/nemotron-3-super-120b-a12b:free |
262K | See provider | |
| GitHub Models | Phi-4 | Phi-4 |
131K | See provider |
| Mistral Large (24.11) | Mistral-large-2411 |
131K | See provider | |
| AI21 Jamba 1.5 Large | AI21-Jamba-1.5-Large |
256K | See provider | |
| Google Gemini | Gemini 3.6 Flash | gemini-3.6-flash |
1M | 15 RPM, 1,500 RPD |
| Gemini 3.5 Flash | gemini-3.5-flash |
1M | 15 RPM, 1,500 RPD | |
| Gemini 3.5 Flash-Lite | gemini-3.5-flash-lite |
1M | 30 RPM, 1,500 RPD | |
| Ollama Cloud | deepseek-v4-pro | deepseek-v4-pro |
128K | Session/weekly limits (.. |
| deepseek-v4-flash | deepseek-v4-flash:preview |
1M | Session/weekly limits (.. | |
| minimax-m3 | minimax-m3 |
1M | Session/weekly limits (.. | |
| OVHcloud AI Endpoints | Qwen3.5-397B-A17B | qwen3.5-397b-a17b |
131K | 2 RPM (anonymous) |
| Meta-Llama-3_3-70B-Instruct | meta-llama-3_3-70b-instruct |
131K | 2 RPM (anonymous) | |
| Qwen3.6-27B | qwen3.6-27b |
131K | 2 RPM (anonymous) | |
| LLM7.io | deepseek-r1-0528 | deepseek-r1-0528 |
131K | 30 RPM (120 with token) |
| deepseek-v3-0324 | deepseek-v3-0324 |
131K | 30 RPM (120 with token) | |
| gpt-4o-mini | gpt-4o-mini |
131K | 30 RPM (120 with token) | |
| Groq | Moonshot Kimi K2 | moonshotai/kimi-k2-instruct |
131K | See provider |
| Moonshot Kimi K2 0905 | moonshotai/kimi-k2-instruct-0905 |
131K | See provider | |
| groq/compound | groq/compound |
131K | 30 RPM, 250 RPD | |
| Mistral AI | Mistral 7B | open-mistral-7b |
32K | See provider |
| Mixtral 8x7B | open-mixtral-8x7b |
32K | See provider | |
| Mistral Medium 3.5 (128B) | mistral-medium-3-5-128b |
256K | ~1 RPS, 500K TPM | |
| Cohere | Command A+ (218B) | command-a-218b |
436K | 20 RPM |
| Command A (111B) | command-a-111b |
288K | 20 RPM | |
| Command R+ | command-r |
128K | 20 RPM | |
| Kilo Code | nvidia/nemotron-3-ultra-550b-a55b:free | nvidia/nemotron-3-ultra-550b-a55b:free |
1M | ~200 req/hr |
| stepfun/step-3.7-flash:free | stepfun/step-3.7-flash:free |
262K | ~200 req/hr | |
| nvidia/nemotron-3-super-120b-a12b:free | nvidia/nemotron-3-super-120b-a12b:free |
262K | ~200 req/hr | |
| Cerebras | Llama 3.1 70B | llama3.1-70b |
131K | See provider |
| gpt-oss-120b | gpt-oss-120b |
131K | 5 RPM, 30K TPM, 1M TPD | |
| zai-glm-4.7 (deprecated Aug 2026) | zai-glm-4.7 |
131K | 5 RPM, 30K TPM, 1M TPD | |
| OpenCode Zen | big-pickle | big-pickle |
0 | |
| DeepSeek V4 Flash | deepseek-v4-flash-free |
1M | ||
| MiMo-V2.5 | mimo-v2.5-free |
1M | ||
| Aion Labs | Aion 2.5 | aion-2-5 |
128K | 15 RPM, 20K TPD |
| Aion 2.0 | aion-2-0 |
128K | 15 RPM, 20K TPD | |
| Aion-RP 1.0 (8B) | aion-rp-1-0-8b |
32K | 15 RPM, 20K TPD | |
| Hugging Face | Meta-Llama-3.1-8B-Instruct | meta-llama-3-1-8b-instruct |
128K | Credit-metered |
| gemma-3-4b-it | gemma-3-4b-it |
131K | Credit-metered | |
| Qwen2.5-Coder-7B-Instruct | qwen2-5-coder-7b-instruct |
131K | Credit-metered | |
| Agnes AI | agnes-1.5-flash | agnes-1.5-flash |
256K | 30 RPM |
| agnes-2.0-flash | agnes-2.0-flash |
256K | 30 RPM | |
| agnes-image-2.0-flash | agnes-image-2.0-flash |
4K | 30 RPM (1K) | |
| Alibaba Cloud Model Studio | Qwen3-Max | qwen3-max |
128K | Tiered by region |
| Qwen3-Plus | qwen3-plus |
1M | Tiered by region | |
| Qwen3-VL-Plus | qwen3-vl-plus |
128K | Tiered by region | |
| Z AI (Zhipu AI) | GLM-4.7-Flash | glm-4.7 |
200K | 1 concurrent request |
| GLM-4.5-Flash | glm-4.5 |
128K | 1 concurrent request | |
| GLM-4.6V-Flash | glm-4.6 |
128K | 1 concurrent request | |
| SambaNova | DeepSeek-V3.1 | deepseek-v3-1 |
128K | 20 RPM, 20 RPD, 200K TPD |
| DeepSeek-V3.2 (Preview) | deepseek-v3-2-preview |
128K | 20 RPM, 20 RPD, 200K TPD | |
| MiniMax-M2.7 | minimax-m2-7 |
128K | 20 RPM, 20 RPD, 200K TPD | |
| SiliconFlow | deepseek-ai/DeepSeek-R1-Distill-Qwen-7B | deepseek-ai-deepseek-r1-distill-qwen-7b |
131K | 30 RPM, 60K TPM |
| Abbreviation | abbreviation |
131K | See provider | |
| deepseek-ai/DeepSeek-OCR | deepseek-ai-deepseek-ocr |
131K | 30 RPM, 60K TPM | |
| xAI | grok-4.3 | grok-4-3 |
1M | Credit-based |
| grok-4.1-fast | grok-4-1-fast |
2M | Credit-based | |
| grok-3-mini | grok-3-mini |
131K | Credit-based | |
| Chutes.ai | DeepSeek-R1 | deepseek-ai/DeepSeek-R1 |
131K | Community-powered, no h.. |
| Llama 3.1 70B | meta-llama/Meta-Llama-3.1-70B-Instruct |
131K | Community-powered, no h.. | |
| Glhf.chat | Llama 3.1 70B | meta-llama/Meta-Llama-3.1-70B-Instruct |
131K | Unlimited for free models |
| Mixtral 8x7B | mistralai/Mixtral-8x7B-Instruct-v0.1 |
32K | Unlimited for free models | |
| Grok (xAI) | Grok-2 | grok-2 |
131K | $25/month free credits,.. |
| Grok-2 Mini | grok-2-mini |
131K | $25/month free credits,.. | |
| AI21 Labs | Jamba Large 1.7 | jamba-large-1-7 |
256K | 200 RPM, 10 RPS |
| Jamba Mini 2 | jamba-mini-2 |
256K | 200 RPM, 10 RPS | |
| DeepSeek | deepseek-chat (V3.2) | deepseek-chat-v3-2 |
128K | Dynamic |
| deepseek-reasoner (R1) | deepseek-reasoner-r1 |
128K | Dynamic | |
| Nscale | Llama-3.3-70B-Instruct | llama-3-3-70b-instruct |
128K | Fair-use |
| DeepSeek-R1-Distill-Llama-70B | deepseek-r1-distill-llama-70b |
128K | Fair-use | |
| Nebius | Qwen3-235B-A22B | qwen3-235b-a22b |
128K | Tier-based |
| Tool | Type | Highlights |
|---|---|---|
| Ollama | CLI + API | 100+ models, GPU acceleration, OpenAI-compatible endpoint |
| LM Studio | Desktop GUI | Any GGUF model, built-in model browser, offline |
| llama.cpp | C/C++ engine | Runs any GGUF, minimal dependencies |
| GPT4All | Desktop app | CPU-only, no GPU required, open source |
| Jan.ai | Desktop app | Privacy-focused, 100% offline ChatGPT alternative |
| KoboldCpp | Single executable | Optimized for creative writing, GGUF |
Data from freellm.net, updated daily via API monitoring.
| Model | Provider | Context | Weekly Usage |
|---|---|---|---|
| z-ai/glm-5.2 | NVIDIA NIM | 1M | 2998B tokens |
| NVIDIA: Nemotron 3 Ultra (free) | OpenRouter | 1M | 2326B tokens |
| Poolside: Laguna M.1 (free) | OpenRouter | 262K | 768B tokens |
| NVIDIA: Nemotron 3 Super (free) | OpenRouter | 262K | 315B tokens |
| Cohere: North Mini Code (free) | OpenRouter | 256K | 255B tokens |
| poolside/laguna-xs-2.1 | NVIDIA NIM | 262K | 171B tokens |
| z-ai/glm-5.1 | NVIDIA NIM | 202K | 158B tokens |
| Poolside: Laguna S 2.1 (free) | OpenRouter | 262K | 83B tokens |
| Poolside: Laguna XS 2.1 (free) | OpenRouter | 262K | 81B tokens |
| NVIDIA: Nemotron 3 Nano 30B A3B (free) | OpenRouter | 256K | 45B tokens |
awesome-free-llm-apis/
├── README.md ← Complete provider directory & code examples
├── code-examples/ ← Ready-to-use config snippets
│ ├── claude-code.md
│ ├── cursor.md
│ └── codex.md
└── LICENSE ← MIT
For the full structured dataset with 453 models and daily updates, visit freellm.net.
We welcome contributions!
- Add a missing free model — Open an issue or submit a PR
- Fix inaccurate data — Rate limits change, providers graduate. PRs welcome
- Add a config snippet — Have a working config for a tool we don't cover? Add it to
code-examples/
A model belongs in this list if:
- The provider explicitly offers a free tier (not just a trial credit)
- The API is publicly accessible (no waitlist, closed beta, or reverse-engineering)
- For trial credits: clearly labeled and minimum $1 credit value
- 🌐 Live site: freellm.net — search, compare, playground, config generator
- 🔑 API key directory: freellm.net/free-llm-api-keys/
- ⚙️ Config generator: freellm.net/config/
- 🎮 Playground: freellm.net/playground/
- 📊 Compare models: freellm.net/compare/
MIT © open-free-llm-api
Last updated: 2026-08-06