The Fine-Tuning Index / RLHF & Preference / #8
oumi-ai/oumi
by oumi-ai · RLHF & Preference · updated today
Easily fine-tune, evaluate and deploy Qwen, Gemma, or any open weight LLM!
77
momentum
9,384
stars
789
forks
#8
rank
dpoevaluationfine-tuninggpt-ossgpt-oss-120bgpt-oss-20binferencellamallmsopen-weightopen-weight-modelsopen-weights
View on GitHub →