AI content moderation that knows when to ask a person

Queue 4
Profiles, Spanish Reports, English Listings, Portuguese

Lucía, 31

Madrid, joined today

Nurse, love the sea and salsa nights.

Not here much, write me on Telegram @lu_31

MD-48211 Blurred by default
Policy

Romance scam pattern

Section 4.2 Moving users off the platform to ask for money

AI score
0.71
Audit log
09:40MD-48210 AI 0.55 → human review → Approved, reviewer 022
09:39MD-48209 AI 0.97 → removed under your threshold
09:38MD-48207 AI 0.71 → human review → Removed, reviewer 014

Where automated content moderation works

  • Spam, duplicates and copy paste abuse at high volume
  • Explicit nudity and clear graphic violence in images
  • Known prohibited terms, links and contact details in listings
  • Language detection and routing to the right reviewer
  • Ranking a queue by risk so the worst items are seen first

Where models still need people

  • Context and intent: a slur quoted in a report, sarcasm, reclaimed words
  • Romance scams and fraud that look polite message by message
  • New slang, coded language and emoji used to evade filters
  • Local law and cultural norms that differ by market
  • Appeals, where the user adds context the model never saw
  • Generated content that imitates real people or brands

Two thresholds and three outcomes

The model is checked too

AI scoring by content type

Content typeWhat AI scoresWhat people decide
Text and commentsspam, toxicity, threats, personal datacontext, intent, quoted speech
Chat and DMsscam patterns, harassment, linksconversation history, grooming signals
Imagesnudity, violence, symbols, text in imageborderline nudity, art, news value
Videoframes and audio flagged for reviewcontext across the whole clip
Profilesstolen photos, contact details, impersonationreal identity, verification disputes
Listingsprohibited items, counterfeit termsprice anomalies, fraud patterns

See your AI and human split

Get your moderation plan in two minutes

Tell us your platform type, content, volume, languages and SLA. The plan shows your team per shift, the AI and human split and a monthly estimate.

Get my moderation plan