Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
A list of free LLM inference resources accessible via API.
| Date | Stars |
|---|---|
| 2026-07-31 | 28838 |
| 2026-08-01 | 29035 |
| 2026-08-06 | 29035 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
<!--- WARNING: DO NOT EDIT THIS FILE DIRECTLY. IT IS GENERATED BY src/pull_available_models.py ---> # Free LLM API resources This lists various services that provide free access or credits towards API-based LLM usage. > [!NOTE] > Please don't abuse these services, else we might lose them. > [!WARNING] > This list explicitly excludes any services that are not legitimate (eg reverse engineers an existing chatbot) - [Free Providers](#free-providers) - [OpenRouter](#openrouter) - [Google AI Studio](#google-ai-studio) - [NVIDIA NIM](#nvidia-nim) - [Mistral (La Plateforme)](#mistral-la-plateforme) - [Mistral (Codestral)](#mistral-codestral) - [HuggingFace Inference Providers](#huggingface-inference-providers) - [Vercel AI Gateway](#vercel-ai-gateway) - [Kilo Gateway](#kilo-gateway) - [OpenCode Zen](#opencode-zen) - [Cerebras](#cerebras) - [Groq](#groq) - [Cohere](#cohere) - [Cloudflare Workers AI](#cloudflare-workers-ai) - [Providers with trial credits](#providers-with-trial-credits) - [Fireworks](#fireworks) - [Baseten](#baseten) - [Nebius](#nebius) - [Novita](#novita) - [AI21](#ai21) - [Upstage](#upstage) - [NLP Cloud](#nlp-cloud) - [Alibaba Cloud (International) Model Studio](#alibaba-cloud-international-model-studio) - [Modal](#modal) - [Inference.net](#inferencenet) - [Hyperbolic](#hyperbolic) - [SambaNova Cloud](#sambanova-cloud) - [Scaleway Generative APIs](#scaleway-generative-apis) ## Free Providers ### [OpenRouter](https://openrouter.ai) **Limits:** [20 requests/minute<br>50 requests/day<br>Up to 1000 requests/day with $10 lifetime topup](https://openrouter.ai/docs/api/reference/limits) Models share a common quota. - [Cohere North Mini Code](https://openrouter.ai/cohere/north-mini-code:free) - [Ling 3.0 Flash](https://openrouter.ai/inclusionai/ling-3.0-flash:free) - [NVIDIA Nemotron 3 Nano Omni 30B A3B (Reasoning)](https://openrouter.ai/nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free) - [NVIDIA Nemotron 3 Super 120B A12B](https://openrouter.ai/nvidia/nemotron-3-super-120b-a12b:free) - [NVIDIA Nemotron 3 Ultra 550B A55B](https://openrouter.ai/nvidia/nemotron-3-ultra-550b-a55b:free) - [NVIDIA Nemotron 3.5 Content Safety](https://openrouter.ai/nvidia/nemotron-3.5-content-safety:free) - [Poolside Laguna S 2.1](https://openrouter.ai/poolside/laguna-s-2.1:free) - [Poolside Laguna XS 2.1](https://openrouter.ai/poolside/laguna-xs-2.1:free) - [google/gemma-4-26b-a4b-it:free](https://openrouter.ai/google/gemma-4-26b-a4b-it:free) - [google/gemma-4-31b-it:free](https://openrouter.ai/google/gemma-4-31b-it:free) - [nvidia/nemotron-3-nano-30b-a3b:free](https://openrouter.ai/nvidia/nemotron-3-nano-30b-a3b:free) - [nvidia/nemotron-nano-12b-v2-vl:free](https://openrouter.ai/nvidia/nemotron-nano-12b-v2-vl:free) - [nvidia/nemotron-nano-9b-v2:free](https://openrouter.ai/nvidia/nemotron-nano-9b-v2:free) - [openai/gpt-oss-20b:free](https://openrouter.ai/openai/gpt-oss-20b:free) ### [Google AI Studio](https://aistudio.google.com) Data is used for training when used outside of the UK/CH/EEA/EU. <table><thead><tr><th>Model Name</th><th>Model Limits</th></tr></thead><tbody> <tr><td>Gemini 3.6 Flash</td><td>250,000 tokens/minute<br>20 requests/day<br>5 requests/minute</td></tr> <tr><td>Gemini 3.5 Flash</td><td>250,000 tokens/minute<br>20 requests/day<br>5 requests/minute</td></tr> <tr><td>Gemini 3 Flash</td><td>250,000 tokens/minute<br>20 requests/day<br>5 requests/minute</td></tr> <tr><td>Gemini 3.5 Flash-Lite</td><td>250,000 tokens/minute<br>500 requests/day<br>15 requests/minute</td></tr> <tr><td>Gemini 3.1 Flash-Lite</td><td>250,000 tokens/minute<br>500 requests/day<br>15 requests/minute</td></tr> <tr><td>Gemini 2.5 Flash</td><td>250,000 tokens/minute<br>20 requests/day<br>5 requests/minute</td></tr> <tr><td>Gemini 2.5 Flash-Lite</td><td>250,000 tokens/minute<br>20 requests/day<br>10 requests/minute</td></tr> <tr><td>Gemini 3.1 Flash TTS</td><td>10,000 tokens/minute<br>10 req
Excerpt of 13,280 characters
Read on GitHubJun Siang Cheah · United Kingdom
260
161
Salman Chishti · GitHub, ex-Microsoft · United Kingdom
2
Kevin Markham · Data School · United States
1
1
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:b50e2b7e74641bb8, topic:llm, topic:llama