SalesforceAIResearch/DiffusionDPO
quality grade B, 68 out of 100Code for "Diffusion Model Alignment Using Direct Preference Optimization"
- stars
- 707
- stars gained this week
- —this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Adapting pretrained models: PEFT/LoRA, instruction tuning, RLHF, DPO and preference alignment.
Signals: fine-tuning, finetuning, lora, peft, qlora, rlhf, dpo, instruction-tuning
284 results
Code for "Diffusion Model Alignment Using Direct Preference Optimization"
ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment
Accessible Google Colab notebooks for Stable Diffusion Lora training, based on the work of kohya-ss and Linaqruf
[UNMAINTAINED] Automated machine learning- just give it a data file! Check out the production-ready version of this project at ClimbsRocks/auto_ml
A package that makes it trivial to create and evaluate machine learning pipeline architectures.
TransmogrifAI (pronounced trăns-mŏgˈrə-fī) is an AutoML library for building modular, reusable, strongly typed machine learning workflows on Apache Spark with minimal hand-tuning
PyTorch Library for Active Learning to accompany Human-in-the-Loop Machine Learning book
A PyTorch coding practice platform — covering LLM, Diffusion, PEFT, and more A friendly environment to help you deeply understand deep learning components through hands-on practice. Like LeetCode, but for tensors. Self-hosted. Supports both Jupyter and Web interfaces.
Code repo for "A Simple Baseline for Bayesian Uncertainty in Deep Learning"
(ICLR 2025) TabM: Advancing Tabular Deep Learning With Parameter-Efficient Ensembling
Code for "Discovering Symbolic Models from Deep Learning with Inductive Biases"
Deep learning simplified by transferring prior learning using the Python deep learning ecosystem
Language model alignment-focused deep learning curriculum
Joint Face Detection and Alignment using Multi-task Cascaded Convolutional Neural Networks
[ICCV 2025] Official code of DeepMesh: Auto-Regressive Artist-mesh Creation with Reinforcement Learning
🌾 OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.
Implementation of ChatGPT RLHF (Reinforcement Learning with Human Feedback) on any generation model in huggingface's transformer (blommz-176B/bloom/gpt/bart/T5/MetaICL)
Collection of scripts and notebooks for OpenAI's latest GPT OSS models
No description
Official Code of Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
Official repository of 'Visual-RFT: Visual Reinforcement Fine-Tuning' & 'Visual-ARFT: Visual Agentic Reinforcement Fine-Tuning'’
A Framework for Speech, Language, Audio, Music Processing with Large Language Model
手把手带你实战 Huggingface Transformers 课程视频同步更新在B站与YouTube
Communicate Freely
24,523 repositories in the index in total.