Skip to main content

Overview

Moderation API detects harmful or inappropriate content in text, helping you:
  • Content Filtering: Automatically filter inappropriate user submissions
  • Safety Review: Detect potential violations before publishing
  • Compliance Check: Ensure content meets platform guidelines
  • Risk Warning: Identify potentially harmful content types
Moderation API uses OpenAI’s moderation model and is free to use without consuming token quota.

Quick Start

Basic Example

Batch Moderation

Detection Categories

Practical Examples

1. User Input Filter

2. Chatbot Safety Layer

3. Custom Threshold

Best Practices

Multi-layer Protection

Pricing

Moderation API is currently free to use and does not count towards token consumption.

FAQ

Based on OpenAI’s model with high accuracy, but not 100% reliable. Recommend:
  • Combine with human review for critical scenarios
  • Set appropriate thresholds
  • Provide appeal channels
Supports multiple languages including Chinese. English works best.
There are rate limits. Control request frequency for batch processing.

Text Generation

Chat API documentation

Content Safety

Content safety policy

API Reference

API reference

Data Security

Data privacy protection