The Fine-Tuning Index / RLHF & Preference / #125

TUDB-Labs/mLoRA

by TUDB-Labs · RLHF & Preference · updated 1y ago

An Efficient "Factory" to Build Multiple LoRA Adapters

29
momentum
385
stars
70
forks
#125
rank
baichuanchatglmdpofinetunegpullamallama2llmloramlorapeftrlhf
View on GitHub →