Loading

vision-sft

Maintained by wshobson

Fine-tune vision-language models (VLMs) with supervised learning on image+text data. Use when adapting a VLM to a visual domain or task, configuring frozen-vision-tower LoRA, or debugging a VLM fine-tune that trains without learning.

Current version
Unknown
License
Unknown
Network access
Unknown / not assessed
Review status
Not verified

Problem it solves

This catalog entry helps users find and evaluate vision-sft for the task described by its available catalog summary. Confirm the exact scope in the linked original source when one is available.

When to use it

Consider vision-sft when its available catalog summary matches the task at hand. When available, review the linked original source before use for precise instructions, requirements, and limitations.

vision-sft is listed as an agent skill in RefHub. The listed maintainer is wshobson. The available catalog summary is: Fine-tune vision-language models (VLMs) with supervised learning on image+text data. Use when adapting a VLM to a visual domain or task, configuring frozen-vision-tower LoRA, or debugging a VLM fine-tune that trains without learning. When available, review the linked original source for exact usage instructions, required tools, and limitations.

Installation and updates

These commands are displayed for copying only and are never executed on RefHub servers. Review the linked upstream source before running them.

Install command
npx skills add wshobson/agents --skill vision-sft -y

Agent compatibility

No compatibility test has been recorded

Do not assume agent compatibility until documented test evidence is available.

Source information and review status

The overview above is structured catalog copy and has no recorded editorial review; technical facts and verification status are shown separately.

Source last reviewed
Not recorded
Catalog source
Open catalog source