Hugging Face daily papers·3d agoRufus-Air: An Open LLM Post-Training Recipe#post-training#sft#reinforcement-learning