Agent skill · wshobson

vision-sft

Fine-tune vision-language models (VLMs) with supervised learning on image+text data. Use when adapting a VLM to a visual domain or task, configuring frozen

Reference it in any coding agent with:

@skills wshobson/vision-sft

Browse the @skills marketplace