Incantory
Sign in
SkillMITNot scanned

Vision Sft

Fine-tune vision-language models (VLMs) with supervised learning on image+text data. Use when adapting a VLM to a visual domain or task, configuring frozen-vision-tower LoRA, or debugging a VLM fine-tune that trains without learning.

wshobson/agentsv10 stars · 0 forks · 0 makes≈1.9K tokens

Family tree

1 prompt · 1 version shown · 0 makes · Full screen

No forks or makes yet. When someone forks this prompt or shares something they made with it, it grows here.

The tree as a list