Comment moderation

The dataset below contains 30 fictional comments with harmful language or respectful discussion. Below, we’ve run Toxic language, Insults, Threats, and Identity attacks checks on it to check comments for harmful content.

Text

Results

Records checked
30/30
Records flagged
15/30 (50%)

Actions

Categories

Checks

Records

Quick filters:
Text CategoriesActivate to sort ascending. ChecksActivate to sort ascending. DecisionActivate to sort ascending. ActionActivate to sort ascending.
People from that country are disgusting and should not be allowed here Harmful content Identity attacks, Toxic language Review Review
Everyone who follows that religion is inferior and unwelcome Harmful content Identity attacks, Toxic language Review Review
Disabled people are worthless and do not belong in this community Harmful content Identity attacks, Toxic language Review Review
People of that ethnicity should be excluded from this forum Harmful content Identity attacks, Toxic language Review Review

Use this with your own content

Configure similar checks for your own dataset or product

Set up your own checks