Skills
All skills
Official skills
Leaderboard
Saved
Categories
Coding & Dev Tools
9772
AI & ML
5010
DevOps & Infra
2864
Integrations & APIs
2327
Productivity
2190
Security
1976
All categories →
Packs
Docs
menu-rounded
Skills
Categories
Packs
Docs
My skills
Saved
light-dark-mode
Light
Dark
System
Sign in
Profile
My skills
Saved
Collections
Preferences
Submit a skill
Sign out
Packs
1 pack
curated
Fine-Tune Transformer Model
Fine-tune transformer language models using TRL with support for SFT, DPO, GRPO, and reward model training.
8 skills · pack
Results for “language-model”
1 skill
orchestra-research
openrlhf-training
Train large language models (7B-70B+) with RLHF using PPO, GRPO, DPO, and other algorithms, accelerated by Ray and vLLM for distributed multi-GPU setups.
10.4k
·
bundle