RL-VIG/LibFewShot
quality grade C, 61 out of 100[TPAMI 2023] LibFewShot: A Comprehensive Library for Few-shot Learning.
- stars
- 1.1k
- stars gained this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Adapting pretrained models: PEFT/LoRA, instruction tuning, RLHF, DPO and preference alignment.
Signals: fine-tuning, finetuning, lora, peft, qlora, rlhf, dpo, instruction-tuning
284 results
[TPAMI 2023] LibFewShot: A Comprehensive Library for Few-shot Learning.
We unified the interfaces of instruction-tuning data (e.g., CoT data), multiple LLMs and parameter-efficient methods (e.g., lora, p-tuning) together for easy use. We welcome open-source enthusiasts to initiate any meaningful PR on this repo and integrate as many LLM related technologies as possible. 我们打造了方便研究人员上手和使用大模型等微调平台,我们欢迎开源爱好者发起任何有意义的pr!
LLM Finetuning with peft
Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"
Mastering Applied AI, One Concept at a Time
Build, personalize and control your own LLMs. From data pre-processing to fine-tuning, xTuring provides an easy way to personalize open-source LLMs. Join our discord community: https://discord.gg/TgHXuSJEk6
Secrets of RLHF in Large Language Models Part I: PPO
Recipes to train reward model for RLHF.
WebGLM: An Efficient Web-enhanced Question Answering System (KDD 2023)
A Doctor for your data
OpenClaw-RL: Train any agent simply by talking
Robust recipes to align language models with human and AI preferences
The official GitHub page for the survey paper "A Survey of Large Language Models".
OpenAssistant is a chat-based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so.
Framework for enhancing LLMs for RAG tasks using fine-tuning.
Official Pytorch Code of the Paper "FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization"
Ethereum Semi Fungible Standard (ERC-1155)
Qwen3 Fine-tuning: Medical R1 Style Chat
聚宝盆(Cornucopia): 中文金融系列开源可商用大模型,并提供一套高效轻量化的垂直领域LLM训练框架(Pretraining、SFT、RLHF、Quantize等)
GraphGen: Enhancing Supervised Fine-Tuning for LLMs with Knowledge-Driven Synthetic Data Generation
chatglm 6b finetuning and alpaca finetuning
Low-rank adaptation for Erasing COncepts from diffusion models.
SD-Trainer. LoRA & Dreambooth training scripts & GUI use kohya-ss's trainer, for diffusion model.
Using Low-rank adaptation to quickly fine-tune diffusion models.
24,523 repositories in the index in total.