The Fine-Tuning Index / RLHF & Preference / #40
Goekdeniz-Guelmez/MLX-LoRA-Studio
by Goekdeniz-Guelmez · RLHF & Preference · updated 17d ago
A native Mac App for LLM fine-tuning on Apple Silicon — fully on-device, fully open source.
59
momentum
265
stars
28
forks
#40
rank
deep-learningllm-trainingllmsmachine-learningmlxmlx-lmmlx-lm-loranlppreference-learningreinforcement-learningrlrlhf
View on GitHub →