OpenAI Moderation logo

OpenAI Moderation

Freemium
访问

OpenAI Moderation helps applications identify potentially harmful content in text and images. Developers can classify standalone inputs or receive moderation signals alongside generated responses through the OpenAI API.

什么是 OpenAI Moderation?

OpenAI Moderation is an OpenAI API capability for identifying potentially harmful content in text and images. Applications can submit standalone inputs to the moderation endpoint or request moderation results as part of a response-generation request. The returned signals can support application policies such as filtering content, routing items for review, or taking action on accounts that submit flagged material.

OpenAI Moderation 能做什么?

Text and image moderation

The `omni-moderation-latest` model accepts text and image inputs. It does not classify audio, and image files can be up to 20 MB.

Standalone classification

Use the moderation endpoint when an application needs to assess text or images without generating a model response.

Inline moderation for generated content

Add a moderation object to a Responses API or Chat Completions API generation request to receive moderation scores for the model input and generated output without making a separate moderation request.

Interpretable moderation results

Results can include a flagged status, moderation categories, and scores. Applications should review these results before displaying generated output or taking downstream action.

Policy-oriented workflows

Moderation results can be incorporated into application rules for filtering, human review, or intervention on accounts that submit flagged content.

使用场景

“Pre-screen user submissions”

Classify text or images before publishing them in a community, marketplace, or other user-facing experience, then apply the application’s content policy.

“Review generated responses”

Request moderation signals alongside generated text and inspect the input and output results before showing a response to a user.

“Route borderline content for review”

Use categories, scores, and flagged statuses to send selected items to a human-review workflow instead of applying the same action to every submission.

“Moderate multimodal inputs”

Assess supported combinations of text and images when an application accepts both modalities, while handling audio through a separate approach.

常见问题

What inputs does OpenAI Moderation support?

The documented `omni-moderation-latest` model accepts text and image inputs. It does not classify audio. Image files can be up to 20 MB.

Can moderation run without generating a response?

Yes. The moderation endpoint can classify standalone text or images without generating a model response.

Can I moderate generated content in the same request?

Yes. For requests made through the Responses API or Chat Completions API, add a top-level moderation object with a moderation model. The API returns moderation results for the model input and generated output while the model still generates normally.

What should an application do with moderation results?

The results can inform actions such as filtering content, routing a request for review, or intervening with accounts that submit flagged content. OpenAI advises reviewing results before displaying generated output or taking downstream actions.

Is the Moderation API intended for CSAM detection?

No. The documented guidance says not to send known or suspected child sexual abuse material to the Moderation API. It is not designed for CSAM detection or handling and is not a substitute for dedicated child-safety safeguards.

快速信息

Category
Developer Tool
Platform
OpenAI API
Primary inputs
Text and images
Documented model
omni-moderation-latest
Workflow options
Standalone classification or moderation alongside generated responses
Endpoint pricing
Free to use

OpenAI Moderation 替代品

Sightengine logo

Sightengine

sightengine.com

用于审核和分析图像、视频、文本与音频的 API 平台

Arena logo

Arena

arena.im

Arena 是面向网站和应用的社区互动平台,提供评论、实时博客、群聊、投票和 AI 辅助审核功能,帮助出版商、媒体、品牌及电商团队将受众互动引导至自有平台。

CommunityOne logo

CommunityOne

communityone.io

CommunityOne 是面向 Discord 社区的 AI 支持、管理、互动、发现与分析平台,帮助服务器运营者自动化日常工作,并了解成员在 Discord、网站和 YouTube 评论中的互动。

Sendbird logo

Sendbird

sendbird.com

面向企业的 AI 客户体验平台,支持聊天、客服、内容审核和 AI 辅助互动

Hive logo

Hive

thehive.ai

Hive provides APIs for understanding, searching, moderating, and generating text, image, video, and audio content. Developers can use its pre-trained and open-source models for content workflows such as moderation, classification, brand protection, and media generation.

Canopy logo

Canopy

canopy.us

Canopy 是一款面向家庭和成人的内容过滤与家长控制应用,可过滤网站、社交媒体和 AI 聊天机器人中的露骨图片、视频及文字,同时不会屏蔽整个互联网。