Upgrading the Moderation API with our new multimodal moderation model
OpenAI Blog
Anthropic's Moderation API now includes a new multimodal model based on GPT-4o that detects harmful content in both text and images. The upgraded system improves accuracy compared to previous versions, though specific performance metrics are not disclosed in the announcement. Developers can now use a single API to moderate mixed-media content rather than relying on separate text and image moderation tools.
Why it matters
We’re introducing a new model built on GPT-4o that is more accurate at detecting harmful text and images, enabling developers to build more robust moderation systems.