Loading

preference-optimization

Maintained by wshobson

Align a fine-tuned model with preference data using DPO, ORPO, KTO, or SimPO. Use when preference pairs or thumbs-up/down feedback exist, when choosing between preference-optimization methods, or when a DPO run needs hyperparameters or debugging.

Current version
Unknown
License
Unknown
Network access
Unknown / not assessed
Review status
Not verified

Problem it solves

This catalog entry helps users find and evaluate preference-optimization for the task described by its available catalog summary. Confirm the exact scope in the linked original source when one is available.

When to use it

Consider preference-optimization when its available catalog summary matches the task at hand. When available, review the linked original source before use for precise instructions, requirements, and limitations.

preference-optimization is listed as an agent skill in RefHub. The listed maintainer is wshobson. The available catalog summary is: Align a fine-tuned model with preference data using DPO, ORPO, KTO, or SimPO. Use when preference pairs or thumbs-up/down feedback exist, when choosing between preference-optimization methods, or when a DPO run needs hyperparameters or debugging. When available, review the linked original source for exact usage instructions, required tools, and limitations.

Installation and updates

These commands are displayed for copying only and are never executed on RefHub servers. Review the linked upstream source before running them.

Install command
npx skills add wshobson/agents --skill preference-optimization -y

Agent compatibility

No compatibility test has been recorded

Do not assume agent compatibility until documented test evidence is available.

Source information and review status

The overview above is structured catalog copy and has no recorded editorial review; technical facts and verification status are shown separately.

Source last reviewed
Not recorded
Catalog source
Open catalog source