uclaml/SPIN
quality grade C, 54 out of 100The official implementation of Self-Play Fine-Tuning (SPIN)
- stars
- 1.3k
- stars gained this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Adapting pretrained models: PEFT/LoRA, instruction tuning, RLHF, DPO and preference alignment.
Signals: fine-tuning, finetuning, lora, peft, qlora, rlhf, dpo, instruction-tuning
283 results
The official implementation of Self-Play Fine-Tuning (SPIN)
Robust fine-tuning of zero-shot models
Official repository of my book "A Hands-On Guide to Fine-Tuning LLMs with PyTorch and Hugging Face"
Expert Specialized Fine-Tuning
Fine-tune CNN in Keras
Home of StarCoder: fine-tuning & inference!
[NeurIPS 2023] Official implementations of "Cheap and Quick: Efficient Vision-Language Instruction Tuning for Large Language Models"
Papers and Datasets on Instruction Tuning and Following. ✨✨✨
[CVPR2024] The code for "Osprey: Pixel Understanding with Visual Instruction Tuning"
No description
Deita: Data-Efficient Instruction Tuning for Alignment [ICLR2024]
[SIGIR'2024] "GraphGPT: Graph Instruction Tuning for Large Language Models"
Generative Representational Instruction Tuning
Instruction Tuning with GPT-4
An easy way to apply LoRA to CLIP. Implementation of the paper "Low-Rank Few-Shot Adaptation of Vision-Language Models" (CLIP-LoRA) [CVPRW 2024].
(AAAI 2024) BLIVA: A Simple Multimodal LLM for Better Handling of Text-rich Visual Questions
A Benchmarking Study of Embedding-based Entity Alignment for Knowledge Graphs, VLDB 2020
Prompt engineering for developers
Qwen-Image text to image lora trainer
[WACV'25 Oral] Fine-Tuning Image-Conditional Diffusion Models is Easier than You Think
[AAAI 2025] Official codes of "ResAdapter: Domain Consistent Resolution Adapter for Diffusion Models".
[CVPR 2024] X-Adapter: Adding Universal Compatibility of Plugins for Upgraded Diffusion Model
HY-SOAR:Self-Correction for Optimal Alignment and Refinement in Diffusion Models
The image prompt adapter is designed to enable a pretrained text-to-image diffusion model to generate images with image prompt.
24,535 repositories in the index in total.