Instagram is rolling out tools to filter out abusive DMs based on keywords and emojis, and to block users along with any new accounts they may create
Facebook and its family of apps have long grappled with the issue of how to better manage — and eradicate — bullying and other harassment …
Context & Ripple Effects
Instagram had already moved moderation upstream with warnings before offensive comments are posted and a quiet Restrict option for abusive users. Its earlier controls to turn off comments and remove followers addressed public or account-level exposure; the new tools extend that approach into direct messages and repeat-account evasion.
The change matters because Instagram is pairing user-defined filtering signals—keywords and emojis—with a block that reaches beyond a single profile, making harassment controls less dependent on victims repeatedly managing new accounts.
First-order effects
- Instagram users gain controls to divert abusive direct messages based on keywords and emojis, reducing the messages that reach their inboxes.
- A blocked user’s newly created accounts can also be blocked, raising the immediate cost of returning to contact the same person.
Second-order effects
- Instagram must operationalize account linkage accurately enough that expanded blocks stop repeat abuse without incorrectly sweeping in unrelated accounts.
- Platforms that have relied on keyword-based anti-harassment controls, including Twitter’s earlier planned post-filtering tool, face a clearer benchmark: message filtering alone does not address account recreation.
Third-order effects
- If account-level enforcement becomes a standard safety control, platform moderation shifts from removing individual posts or profiles toward managing the identity relationships behind repeat abuse.
- The combined rollout points to safety products being designed as layered user controls—prevention, filtering, and repeat-offender blocking—rather than one-off reporting features.
The trend: Social platforms are evolving anti-harassment systems from content moderation and single-account blocks toward layered controls that limit repeat contact.