USE-CASES

Content moderation classification with a review queue

Turn a written text policy into a repeatable triage step. The classifier helps sort the queue; your team remains responsible for enforcement and appeals.

Write the rule before testing the model

Specify the behavior that violates your policy, permitted exceptions, and what counts as insufficient context. Avoid a single vague instruction such as “detect bad content.” Quotation, reporting abuse, and satire can change the meaning of the same phrase.

Use three outcomes

Allow covers clearly permitted text. Review covers ambiguity and missing context. Block should be reserved for clear violations supported by your policy and evaluation results. A confidence threshold adds a second review condition; it does not replace the policy.

This workflow accepts text. Do not use it as an image, audio, or video moderation system.

Audit errors in both directions

False positives silence legitimate users; false negatives expose others to harmful content. Review both, compare performance across supported languages, and provide an appeal path. Do not represent a model score as a legal finding or certainty.

Try a decision ↗

Continue your workflow

Ask about plans & usage

Find a product answer

Answers come from the published product guide. For account-specific questions, contact support.

Contact

Take the next decision into your workspace.

3 anonymous attempts per day · 20 signup credits · No card for the trial

Sign in ↗

Tell us what you need

Describe the workflow you want to classify, the volume you expect, and the outcome you need. We reply by email.

10–2,000 characters. Never include passwords, API keys, or payment details.