Jack000/glid-3-xl
quality grade D, 38 out of 1001.4B latent diffusion model fine tuning
- stars
- 265
- stars gained this week
- —this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Adapting pretrained models: PEFT/LoRA, instruction tuning, RLHF, DPO and preference alignment.
Signals: fine-tuning, finetuning, lora, peft, qlora, rlhf, dpo, instruction-tuning
284 results
1.4B latent diffusion model fine tuning
[ACM Computing Surveys] The collection of awesome papers on alignment of diffusion models.
Automatic1111 WEBUI extension to autofill keyword for custom stable diffusion models and LORA models.
The stable diffusion webui training aid extension helps you quickly and visually train models such as Lora.
This extension replaces the built-in LoRA forward procedure.
Official implementation for "Towards a Unified Self-Distillation Framework for Large Language Models" (https://arxiv.org/abs/2605.06597).
DISC-FinLLM,中文金融大语言模型(LLM),旨在为用户提供金融场景下专业、智能、全面的金融咨询服务。DISC-FinLLM, a Chinese financial large language model (LLM) designed to provide users with professional, intelligent, and comprehensive financial consulting services in financial scenarios.
Finetuning Large Language Models on One Consumer GPU in 2 Bits
Finetuning large language models for GDScript generation.
Best practices for distilling large language models.
Training Large Language Model to Reason in a Continuous Latent Space
BELLE: Be Everyone's Large Language model Engine(开源中文对话大模型)
A project to improve skills of large language models
Aligning pretrained language models with instruction data generated by themselves.
Foundations of Medical Large Language Model Learning
Self-Adapting Language Models
:sparkles::sparkles:Latest Advances on Multimodal Large Language Models
[MICCAI 2019 Young Scientist Award] [MEDIA 2020 Best Paper Award] Models Genesis, one of the first "foundation" models in medical image analysis for multiple downstream tasks
A Unified Parameter-Efficient Transfer Learning Benchmark for Computer Vision Tasks
SCEPTER is an open-source framework used for training, fine-tuning, and inference with generative models.
High Accuracy and efficiency multi-task fine-tuning framework for Code LLMs. This work has been accepted by KDD 2024.
Quick Start for Large Language Models (Theoretical Learning and Practical Fine-tuning) 大语言模型快速入门(理论学习与微调实战)
AgentCPM-GUI: An on-device GUI agent for operating Android apps, enhancing reasoning ability with reinforcement fine-tuning for efficient task execution.
Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference time
24,523 repositories in the index in total.