v0.0.3-alpha35
Add LoRA and Prefix-Tuning as Modeling Options for Improved Memory Efficiency + performance (potentially) (#2840)
Add flash-attention patch for falcon-7b (#3580)
add one .gitignore to resolve conflicts
Simplify message and token format doc (#3324)
update READMEs (#3487)