The Fine-Tuning Index / RLHF & Preference / #66
Yog-Sotho/LLM-fine-tuner
by Yog-Sotho · RLHF & Preference · updated 17d ago
Powerful no-code LLM fine-tuner: upload data → train → deploy in minutes. Unsloth 2-5× acceleration · QLoRA/DPO/RLHF/PPO/ORPO · Reward Model training · GGUF export · vLLM inference · BLEU/ROUGE/BERTScore · full CLI · Heretic Mode to unlock full model potential
46
momentum
33
stars
4
forks
#66
rank
abliterationaidpofine-tuningggufgradiollmllmslm-studiolocal-ailocal-llmollama
View on GitHub →