🎉 Smithery is now a part of Arcade.dev! Read more in our announcement here
Train and fine-tune LLMs using HuggingFace TRL, Transformers, and cloud GPU infrastructure with SFT, DPO, GRPO methods