huggingface-llm-trainer
Train or fine-tune language and vision models using TRL (Transformer Reinforcement Learning) or Unsloth with Hugging Face Jobs infrastructure. Covers SFT, DPO, GRPO and reward modeling training methods, plus GGUF conversion for local deployment. Includes guidance on the TRL Jobs package, UV scrip...
Install this skill
or
huggingface-llm-trainer19 files
Comments
Sign in to leave a comment.
No comments yet. Be the first to comment!