DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales Paper • 2308.01320 • Published Aug 2, 2023 • 46
TokenBender/llama2-7b-chat-hf-codeCherryPop-qLoRA-merged Text Generation • Updated Aug 8, 2023 • 12 • • 68
enterprise-explorers/Llama-2-7b-chat-coreml Text Generation • Updated Jul 18, 2023 • 1.6k • 136
Stack More Layers Differently: High-Rank Training Through Low-Rank Updates Paper • 2307.05695 • Published Jul 11, 2023 • 24
Becoming self-instruct: introducing early stopping criteria for minimal instruct tuning Paper • 2307.03692 • Published Jul 5, 2023 • 27