Vision Sft

by wshobson46891e7e60daNo licenseListed Oct 8, 2026Updated Oct 8, 2026

Fine-tune vision-language models (VLMs) with supervised learning on image+text data. Use when adapting a VLM to a visual domain or task, configuring frozen-vision-tower LoRA, or debugging a VLM fine-tune that trains without learning.

Instructions onlyAI & Agents
  1. 46891e7e60daCurrentcommit 46891e7Published Oct 8, 2026

Source and attribution

Source:wshobson/agentsinplugins/llm-finetuning/skills/vision-sftat commit46891e7

License: No license

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal