Vision Sft
Fine-tune vision-language models (VLMs) with supervised learning on image+text data. Use when adapting a VLM to a visual domain or task, configuring frozen-vision-tower LoRA, or debugging a VLM fine-tune that trains without learning.
wshobson/agentsv10 stars · 0 forks · 0 makes≈1.9K tokens
Scores by version
No eval runs reported yet. Incantory never runs prompts: owners report results from their own CI.
Datasets
No datasets.