transformerlab/transformerlab-app
quality grade A, 86 out of 100The open source research environment for AI researchers to seamlessly train, evaluate, and scale models from local hardware to GPU clusters.
- stars
- 5.2k
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Adapting pretrained models: PEFT/LoRA, instruction tuning, RLHF, DPO and preference alignment.
Signals: fine-tuning, finetuning, lora, peft, qlora, rlhf, dpo, instruction-tuning
284 results
The open source research environment for AI researchers to seamlessly train, evaluate, and scale models from local hardware to GPU clusters.
Research of DeepSeek Engram Architecture based on Qwen-3 and Stable Diffusion series.
Paperclip adapter for Hermes Agent — run Hermes as a managed employee in a Paperclip company
Easily fine-tune, evaluate and deploy Qwen, Gemma, or any open weight LLM!
An MCP adapter that bridges the Abilities API to the Model Context Protocol, enabling MCP clients to discover and invoke WordPress plugin, theme, and core abilities programmatically.
Synthetic data curation for post-training and structured data extraction
streamline the fine-tuning process for multimodal models: PaliGemma 2, Florence-2, and Qwen2.5-VL
Official Repo of "D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models"
The implementation of “Fine-tuning Graph Neural Networks by Preserving Graph Generative Patterns”
Implementation for <Large-Margin Softmax Loss for Convolutional Neural Networks> in ICML'16.
Official implementation of Pref-GRPO: Pairwise Preference Reward-based GRPO for Stable Text-to-Image Reinforcement Learning
[ECCV 2026] Official Code of "Distribution Matching Distillation Meets Reinforcement Learning"
This repo contains the source code for the paper "Evolution Strategies at Scale: LLM Fine-Tuning Beyond Reinforcement Learning"
A library that integrates huggingface transformers with the world of fastai, giving fastai devs everything they need to train, evaluate, and deploy transformer specific models.
Guide: Finetune GPT2-XL (1.5 Billion Parameters) and finetune GPT-NEO (2.7 B) on a single GPU with Huggingface Transformers using DeepSpeed
Pretrain and finetune ELECTRA with fastai and huggingface. (Results of the paper replicated !)
No description
Inference, Fine Tuning and many more recipes with Gemma family of models
Single image to Lora Model for Flux in ComfyUI using Llm and Flux Kontext
Finetuning of Falcon-7B LLM using QLoRA on Mental Health Conversational Dataset
[EMNLP 2024] LongAlign: A Recipe for Long Context Alignment of LLMs
Banishing LLM Hallucinations Requires Rethinking Generalization
输入现代汉语句子,生成古汉语风格的句子。基于荀子基座大模型,采用“文言文(古文)- 现代文平行语料”中的部分数据进行LoRA微调训练而得。
Deepspeed、LLM、Medical_Dialogue、医疗大模型、预训练、微调
24,523 repositories in the index in total.