mirror of
https://github.com/hiyouga/LlamaFactory.git
synced 2026-01-31 06:42:05 +00:00
101 B
101 B
Usage:
pretrain.shsft.sh->reward.sh->ppo.shsft.sh->dpo.sh->predict.sh