Moderation api
Moderation api
Changelog
  • Changelog
  • Support portal

Moderation api changelog

March updates: 5 new models, email reports & shadow flagging

March updates: 5 new models, email reports & shadow flagging

This month we're shipping several features that give you more control over how you monitor and tune your moderation setup - plus five new policy models covering regulated content categories.

Email reports

You can now receive email reports for new and unresolved queue items. Each queue can be configured separately, and each user manages their own notification preferences.

Daily, weekly, and monthly reports are available on all plans. Hourly reports are available on Growth and Enterprise.

Shadow flagging

Shadow flagging lets you test new policies before fully enabling them. Items matched by a shadow-flagged policy appear in a dedicated queue but are not acted on. This gives you a safe way to evaluate coverage and tune thresholds before going live.

To enable it, set a policy to "Do not flag" and configure a queue to show shadow flagged content.

Available on all plans.

Policy thresholds

You can now set per-policy thresholds to control how strictly a policy is enforced. To help you calibrate, we show how many messages would have been flagged at any given threshold over the last 30 days.

Available on all plans.

5 new policy models

We've added five new models covering regulated content categories. Each is trained using our risk-based approach: content that poses a direct risk to the end user scores higher, while a passing mention scores lower. This lets you set a threshold that blocks promotional or facilitative content without flagging incidental references.

  • Adult: Detects content related to adult services, products, and websites.

  • Firearms: Detects content related to the sale, acquisition, or promotion of firearms and related accessories.

  • Gambling: Detects content promoting gambling services, platforms, or solicitations.

  • Crypto: Detects content related to cryptocurrency promotions, investment solicitations, and related financial schemes.

  • Cannabis: Detects content related to cannabis products, dispensaries, and related services.

Secondary features

  • Usage rate limit cadence indicator: You can now see your maximum request cadence directly in Billing → Usage & Limits, making it easier to plan traffic bursts and avoid throttling. Available on all plans.

Improvements

  • API: Switch to a 1-minute rate limit window for clearer, more predictable throttling behavior

  • Dashboard: Change overview percentages to show share of total messages instead of month-over-month deltas

  • Discord: Update Discord plugin to use the new endpoint; upgrade to the latest plugin version

  • General: Fine-tune topic detection for better categorization on edge cases

  • General: Retrain phishing model to reflect the latest tactics and improve precision

  • Improvement: Accept Base64 images in submission requests to simplify file handling

  • Improvement: Update WordPress plugin to the new endpoint; install the latest version to stay compatible

  • Improvement: Improve item detail load time for a faster review experience

  • Improvement: Save request timings so you can inspect end-to-end latency in logs

  • Performance: Speed up review queue queries for snappier filtering and navigation

Fixes

  • Bug: Author histogram now updates correctly after new data ingests

  • Bug: Author trust scores classify legacy accounts correctly instead of marking them as new

  • Bug: Insights now attach when adding policies through the API

  • Bug: Filter by action works as expected across all queue views

  • Bug: Resolve slow queries to stabilize response times under load

  • Bug: Now showing more accurate unique author counts in project overviews