Introducing vision to the fine-tuning API
OpenAI Blog
OpenAI's fine-tuning API now supports GPT-4o with both images and text inputs, allowing developers to customize the model's vision capabilities. The update enables model training on multimodal datasets containing 1 to 500 training examples per task. This expands the fine-tuning options beyond text-only models to include vision-based applications.
Why it matters
Developers can now fine-tune GPT-4o with images and text to improve vision capabilities