preference-optimization
Maintained by wshobson
Align a fine-tuned model with preference data using DPO, ORPO, KTO, or SimPO. Use when preference pairs or thumbs-up/down feedback exist, when choosing between preference-optimization methods, or when a DPO run needs hyperparameters or debugging.
- Current version
- Unknown
- License
- Unknown
- Network access
- Unknown / not assessed
- Review status
- Not verified
Problem it solves
This catalog entry helps users find and evaluate preference-optimization for the task described by its available catalog summary. Confirm the exact scope in the linked original source when one is available.
When to use it
Consider preference-optimization when its available catalog summary matches the task at hand. When available, review the linked original source before use for precise instructions, requirements, and limitations.
Installation and updates
These commands are displayed for copying only and are never executed on RefHub servers. Review the linked upstream source before running them.
npx skills add wshobson/agents --skill preference-optimization -y
Agent compatibility
No compatibility test has been recorded
Do not assume agent compatibility until documented test evidence is available.