AI Potluck
Product / UX / Safety & Guardrails

OpenAI Moderation API

OpenAI

OpenAI's hosted moderation endpoint, a free API that classifies text (and images) against OpenAI's content-policy categories (hate, harassment, self-harm, sexual, violence, and more). Closed, API-only; no weights or self-hosting.

OpenAI Moderation API; free hosted endpoint, proprietary (no weights, no self-host). Verified live June 2026.

Openness

1 high confidence
1.0
service
proprietary hosted API(no weights, no self-host)
access
free moderation endpoint
source
closed
license
proprietary

Closed, API-only hosted classifier; no weights or source.

Adoption

4 medium confidence
4.0

Free, default moderation endpoint for the large OpenAI developer base; widely used as a first-line filter.

Capability

3 low confidence
3.0

Convenient broad-category moderation; fixed policy, no customization vs policy-conditioned models.

Unchanged since 2026-08-01 (last edited, not re-checked)