deepseek-ai/ESFT
quality grade D, 44 out of 100Expert Specialized Fine-Tuning
- stars
- 742
- stars gained this week
- —this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Adapting pretrained models: PEFT/LoRA, instruction tuning, RLHF, DPO and preference alignment.
Signals: fine-tuning, finetuning, lora, peft, qlora, rlhf, dpo, instruction-tuning
284 results
Expert Specialized Fine-Tuning
Fine-tune CNN in Keras
Home of StarCoder: fine-tuning & inference!
[NeurIPS 2023] Official implementations of "Cheap and Quick: Efficient Vision-Language Instruction Tuning for Large Language Models"
Papers and Datasets on Instruction Tuning and Following. ✨✨✨
[CVPR2024] The code for "Osprey: Pixel Understanding with Visual Instruction Tuning"
No description
Deita: Data-Efficient Instruction Tuning for Alignment [ICLR2024]
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
[SIGIR'2024] "GraphGPT: Graph Instruction Tuning for Large Language Models"
Generative Representational Instruction Tuning
Instruction Tuning with GPT-4
An easy way to apply LoRA to CLIP. Implementation of the paper "Low-Rank Few-Shot Adaptation of Vision-Language Models" (CLIP-LoRA) [CVPRW 2024].
(AAAI 2024) BLIVA: A Simple Multimodal LLM for Better Handling of Text-rich Visual Questions
A Benchmarking Study of Embedding-based Entity Alignment for Knowledge Graphs, VLDB 2020
Prompt engineering for developers
Text to speech alignment using CTC forced alignment
Qwen-Image text to image lora trainer
[WACV'25 Oral] Fine-Tuning Image-Conditional Diffusion Models is Easier than You Think
[AAAI 2025] Official codes of "ResAdapter: Domain Consistent Resolution Adapter for Diffusion Models".
[CVPR 2024] X-Adapter: Adding Universal Compatibility of Plugins for Upgraded Diffusion Model
HY-SOAR:Self-Correction for Optimal Alignment and Refinement in Diffusion Models
The image prompt adapter is designed to enable a pretrained text-to-image diffusion model to generate images with image prompt.
DDPO for finetuning diffusion models, implemented in PyTorch with LoRA support
24,523 repositories in the index in total.