lyogavin/airllm
quality grade A, 83 out of 100AirLLM 70B inference with single 4GB GPU
- stars
- 35k
- stars gained this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Adapting pretrained models: PEFT/LoRA, instruction tuning, RLHF, DPO and preference alignment.
Signals: fine-tuning, finetuning, lora, peft, qlora, rlhf, dpo, instruction-tuning
283 results
AirLLM 70B inference with single 4GB GPU
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
An MCP adapter that bridges the Abilities API to the Model Context Protocol, enabling MCP clients to discover and invoke WordPress plugin, theme, and core abilities programmatically.
🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
Build, Evaluate, and Optimize AI Systems. Includes evals, RAG, agents, fine-tuning, synthetic data generation, dataset management, MCP, and more.
:zap: The most powerful PowerShell module for fine-tuning Windows 10 & Windows 11 on GitHub
Train Large Language Models on MLX.
Covers pre-training data, Tokenizer, Transformer, MoE,distributed training, Scaling Laws, inference & alignment .6 progressive code assignments for full-stack LLM learning | 涵盖预训练数据、分词器、Transformer、MoE、分布式训练、缩放定律、推理与对齐,6 项渐进代码作业,掌握 LLM 全栈知识
Textbook on reinforcement learning from human feedback
Easily fine-tune, evaluate and deploy Qwen, Gemma, or any open weight LLM!
🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.
Training Sparse Autoencoders on Language Models
A native Mac App for LLM fine-tuning on Apple Silicon — fully on-device, fully open source.
One-click Portable Windows installation of 'AI-Toolkit by Ostris'
The unofficial OpenAI SDK for Dart & Flutter. Full API coverage + OpenAI-compatible providers (Azure, DeepSeek, LM Studio, Ollama): Responses, Chat, Realtime, Videos, Batch, Fine-tuning.
Train and serve LLMs at extreme speed and massive throughput.
An alignment auditing agent capable of quickly exploring alignment hypothesis
:sparkles::sparkles:Latest Advances on Multimodal Large Language Models
SD-Trainer. LoRA & Dreambooth training scripts & GUI use kohya-ss's trainer, for diffusion model.
ACP adapter for pi coding agent
Trinity-RFT is a general-purpose, flexible and scalable framework designed for reinforcement fine-tuning (RFT) of large language models (LLM).
Developer Asset Hub for NVIDIA Nemotron — A one-stop resource for training recipes, usage cookbooks, datasets, and full end-to-end reference examples to build with Nemotron models
250+ Fine-tuning & RL Notebooks for text, vision, audio, embedding, TTS models.
24,535 repositories in the index in total.