allwefantasy/byzer-llm
quality grade D, 43 out of 100Easy, fast, and cheap pretrain,finetune, serving for everyone
- stars
- 314
- stars gained this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Adapting pretrained models: PEFT/LoRA, instruction tuning, RLHF, DPO and preference alignment.
Signals: fine-tuning, finetuning, lora, peft, qlora, rlhf, dpo, instruction-tuning
284 results
Easy, fast, and cheap pretrain,finetune, serving for everyone
Implementation for "Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs"
AutoAudit—— the LLM for Cyber Security 网络安全大语言模型
Finetune ALL LLMs with ALL Adapeters on ALL Platforms!
HuatuoGPT2, One-stage Training for Medical Adaption of LLMs. (An Open Medical GPT)
Library for industrial alignment.
该仓库主要记录 LLMs 算法工程师相关的顶会论文研读笔记(多模态、PEFT、小样本QA问答、RAG、LMMs可解释性、Agents、CoT)
LLM Tuning with PEFT (SFT+RM+PPO+DPO with LoRA)
Dự án bao gồm: 1. Xây dựng bộ dữ Instructions Vietnamese (chất lượng, nhiều, và đa dạng). 2.LLM Training, Finetuning, Evaluating & Testing trên Open-source mô hình ngôn ngữ: Bloomz,T5, UL2, LLaMA (1&2), OpenLLaMA, GPT-J pythia etc. 3. Ứng dụng và Giao diện Người dùng (UI)
Multi-agent Social Simulation + Efficient, Effective, and Stable alternative of RLHF. Code for the paper "Training Socially Aligned Language Models in Simulated Human Society".
Train a Language Model with GRPO to create a schedule from a list of events and priorities
[ICML'24] Data and code for our paper "Training-Free Long-Context Scaling of Large Language Models"
ArcticTraining is a framework designed to simplify and accelerate the post-training process for large language models (LLMs)
PiSSA: Principal Singular Values and Singular Vectors Adaptation of Large Language Models(NeurIPS 2024 Spotlight)
Official PyTorch implementation of DistiLLM: Towards Streamlined Distillation for Large Language Models (ICML 2024)
Code for "From Context to Skills: Can Language Models Learn from Context Skillfully? "
LoRAMoE: Revolutionizing Mixture of Experts for Maintaining World Knowledge in Language Model Alignment
[KDD'2024] "UrbanGPT: Spatio-Temporal Large Language Models"
Official implementation of paper: SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
Let AI think and express like you. This framework provides a complete assembly line: from the noise processing of original chat records, to the seamless switching of multi-model adaptation layers, to local lightweight fine-tuning (LoRA).
A set of solutions is provided, leveraging Openpangu - 7B as the base model for fine - tuning and application of large language models (LLMs) in operations research optimization tasks.
USING BERT FOR Attribute Extraction in KnowledgeGraph. fine-tuning and feature extraction. 使用基于bert的微调和特征提取方法来进行知识图谱百度百科人物词条属性抽取。
A structured reading list on Vision-Language-Action (VLA) models — from diffusion/flow matching foundations through state-of-the-art robot foundation model architectures to data scaling, RL fine-tuning, and world models. Papers in reading order.
A fine-tuned model from Qwen2.5-1.5B-Instruct, capable of handling sensitive topics. / 从 Qwen2.5-1.5B-Instruct 微调,主要擅长处理色情话题
24,523 repositories in the index in total.