Agent skill · wshobson
vision-sft
Fine-tune vision-language models (VLMs) with supervised learning on image+text data. Use when adapting a VLM to a visual domain or task, configuring frozen
Reference it in any coding agent with:
@skills wshobson/vision-sft