v0.0.3-alpha29
Add LoRA and Prefix-Tuning as Modeling Options for Improved Memory Efficiency + performance (potentially) (#2840)
fix typos (#3216)
sft8 training preparation (#2988)
Add missing parameters in model_chat.py (#3233)
Rl training (#2206)
readme.md dataset Table Formatting (#3219)
Add GPTNeoXRewardModel (#2182)
Fix typo in check_dataset_appearances.py (#3205)
fix bug
augment rank results for reward model and some other improvements for RM training (#2321)