Open Source LLM Fine-Tuning Projects
Tools and frameworks for fine-tuning large language models (LLMs) on custom datasets. Covers supervised fine-tuning (SFT), instruction tuning, reinforcement learning from human feedback (RLHF), Direct Preference Optimization (DPO), GRPO, and parameter-efficient methods like LoRA and QLoRA. Includes both local training frameworks and managed fine-tuning services.
