cxcscmu/skilllearnbench-pytorch-preference-optimization
Guide for implementing preference optimization methods (DPO, SimPO, IPO) in PyTorch. Use when implementing loss functions for RLHF-style training.
npx skills add https://github.com/cxcscmu/SkillLearnBench --skill pytorch-preference-optimization
Take cxcscmu/skilllearnbench-pytorch-preference-optimization from the repository into ~/.claude/skills for personal
use, or into .claude/skills inside a project.
The agent identifies a skill by the name field in its header. Two skills with the
same name cannot sit side by side — one of them will be ignored.