AI4Bharat/IndicLLMSuite
quality grade D, 45 out of 100A blueprint for creating Pretraining and Fine-Tuning datasets for Indic languages
- stars
- 412
- stars gained this week
- —this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Adapting pretrained models: PEFT/LoRA, instruction tuning, RLHF, DPO and preference alignment.
Signals: fine-tuning, finetuning, lora, peft, qlora, rlhf, dpo, instruction-tuning
284 results
A blueprint for creating Pretraining and Fine-Tuning datasets for Indic languages
AdaLoRA: Adaptive Budget Allocation for Parameter-Efficient Fine-Tuning (ICLR 2023).
🐢 💨 Speedup your MacOS setup with this fine tuning settings
The Fastest Way to Fine-Tune LLMs Locally
Easy and Efficient dLLM Fine-Tuning
GPT Fine-Tuning using Node.js - an easy to use starter project
Fine-tuning code for CLIP models
LLM fine-tuning and eval
Quick exploration into fine tuning florence 2
Fine-tuning LLMs using QLoRA
Hands-on tutorials on fine-tuning various LLMs using different fine-tuning techniques
[NAACL'24] Self-data filtering of LLM instruction-tuning data using a novel perplexity-based difficulty score, without using any other models
The official codes for "Aurora: Activating chinese chat capability for Mixtral-8x7B sparse Mixture-of-Experts through Instruction-Tuning"
PromptCBLUE: a large-scale instruction-tuning dataset for multi-task and few-shot learning in the medical domain in Chinese
[ACL'24] Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
Shepherd: A foundational framework enabling federated instruction tuning for large language models
🐙 OctoPack: Instruction Tuning Code Large Language Models
MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs
PMetal: high-performance Apple Silicon framework for local LLM inference, LoRA/QLoRA fine-tuning, serving, quantization, and MLX/Metal acceleration.
Easy and Efficient Finetuning LLMs. (Supported LLama, LLama2, LLama3, Qwen, Baichuan, GLM , Falcon) 大模型高效量化训练+部署.
A proof-of-concept project that showcases the potential for using small, locally trainable LLMs to create next-generation documentation tools.
LongCite: Enabling LLMs to Generate Fine-grained Citations in Long-context QA
ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search (NeurIPS 2024)
Simple Python library/structure to ablate features in LLMs which are supported by TransformerLens
24,523 repositories in the index in total.