Skip to main content

huggingface-llm-trainer

Train or fine-tune language and vision models using TRL (Transformer Reinforcement Learning) or Unsloth with Hugging Face Jobs infrastructure. Covers SFT, DPO, GRPO and reward modeling training methods, plus GGUF conversion for local deployment. Includes guidance on the TRL Jobs package, UV scrip...

Install this skill

or
huggingface-llm-trainer19 files

Comments

Sign in to leave a comment.

No comments yet. Be the first to comment!
Installation guide →
Rate this skill
Categorydevelopment
UpdatedOctober 9, 2026
jacobwell/skills