Vision Sft

by wshobson46891e7e60daNo licenseListed Oct 8, 2026Updated Oct 8, 2026

Fine-tune vision-language models (VLMs) with supervised learning on image+text data. Use when adapting a VLM to a visual domain or task, configuring frozen-vision-tower LoRA, or debugging a VLM fine-tune that trains without learning.

Instructions onlyAI & Agents

Only the file list is public. File contents are available once the skill is installed in a workspace.

PathSizeType
references/collators-and-pitfalls.md5.8 KBtext/markdown
SKILL.md7.7 KBtext/markdown

Source and attribution

Source:wshobson/agentsinplugins/llm-finetuning/skills/vision-sftat commit46891e7

License: No license

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal