Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
A Curated Collection of resources for applied AI engineering (work in progress).
| Date | Stars |
|---|---|
| 2026-07-31 | 539 |
| 2026-08-06 | 543 |
Today
+4 stars today
This week
— stars this week
This month
— stars this month
Momentum
16.0
growth rate 0.00%/day
# 🌟 Awesome LLM Resources A Curated Collection of LLM resources. 💡✨ **🌐 Updated: 22nd of June 2025** ### 'Serverless' Hosting of Private/OS Models | Platform/Tool | Rel. | Scale Down | OS 🔓 | GH | Start | One-Click | Dev Exp. | Free-Tier | | --------------------------------| -------- | -------------| -------------- | ----------|----------------|---------------|-------------------------|--------------| | [Baseten](https://www.baseten.com/) | 2019 | > 15 min | 🔴 | [](https://github.com/basetenlabs) | [Guide](https://docs.baseten.co/deploy/guides/private-model) | 🟡 | 👍 | $30 | | [Modal](https://modal.com/) | 2021 | < 1 min | 🔴 | [](https://github.com/modal-labs) | [Helpers](https://github.com/ilsilfverskiold/Awesome-LLM-Resources-List/tree/main/helpers/scripts/modal) | ❌ | 👍 | $30/m | | [HF Endpoints](https://ui.endpoints.huggingface.co/) | 2023 | > 15 min | 🔴 | [](https://github.com/huggingface) | None Needed | ✅ | 😓 | ❌ | | [Replicate](https://replicate.com/) | 2019 | < 1 min | 🔴 | [](https://github.com/replicate) | [Guide](https://replicate.com/docs/guides/push-a-transformers-model) | 🟡 | 🤷 | ❌ | | [Sagemaker (Serverless)](https://aws.amazon.com/sagemaker/) | 2017 | N/A | 🔴 | [](https://github.com/aws/amazon-sagemaker-examples) | N/A | ❌ | ❌ | 300,000s | | [Lambda w/ EFS (AWS)](https://aws.amazon.com/pm/lambda/) | 2014 | < 1 min | 🔴 | [](https://github.com/awsdocs/aws-lambda-developer-guide) | [Guide](https://aws.amazon.com/blogs/compute/hosting-hugging-face-models-on-aws-lambda/) | ❌ | ❌ | ✅ | | [RunPod Serverless](https://www.runpod.io/serverless-gpu) | 2022 | > 30s | 🔴 | [](https://github.com/runpod) | N/A | ❌ | 🤷 | ❌ | | [BentoML](https://www.bentoml.com/) | 2019 | > 5 min | [](https://github.com/bentoml/BentoML) | [](https://github.com/bentoml) | [Gallery](https://www.bentoml.com/gallery) | 🟡 | 👍 | 🆓 $10 | It goes without saying that these platforms can usually do more than LLM serving** ### 🧮 Serverless Compute Pricing & Limits – Lambda vs Modal (on CPU) | Platform | 💵 Compute Unit | 📥 Per-Request Fee | 🆓 Free Tier | ⏱️ Max Timeout | 🚦 Concurrency Limit | |------------------------|----------------------------------------------------|----------------------------------------------------|------------------------------------------------------|------------------------------------------|-------------------------------------------------------------------| | **AWS Lambda + API GW**| GB-sec @ $0.000016667 | $0.20/M Lambda + $1.00/M HTTP API calls | 1M req + 400k GB-s/mo (12 mo) + 1M API calls/mo | 15 min | 1,000 per region (can request more)
Excerpt of 102,642 characters
Read on GitHub149
Gabriel Bianconi · United States
1
1
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:d6fd2ca0b53ad3ab, topic:large-language-models, topic:llm