Twitter introduces Safety Mode, which blocks harassing accounts for seven days, in beta for some English-language accounts
Feeling safe on Twitter looks different for everyone. We've rolled out features and settings that may help you to feel more comfortable and in control of your experience …
Context & Ripple Effects
Safety Mode is the arrival of a feature Twitter first previewed in February 2021, when it said users would be able to auto-block accounts 'acting abusive or spammy.' It extends a moderation lineage going back years: automatic filtering of abusive mentions in 2015, then hiding abusive replies and safe search in February 2017, followed by expanded word- and account-muting in March of that year.
What distinguishes this round is automation with teeth — a seven-day block applied without the user lifting a finger — and it proved durable: by early 2022 Twitter had widened the beta to half of US users and five other markets. The bet is that harassment controls must act before the user notices the abuse.
First-order effects
- Beta participants get automatic seven-day blocks on accounts using harmful language in their replies, shifting enforcement from reactive blocking to pre-emptive filtering for the first time at scale.
- Accounts flagged as harassing or spammy face temporary invisibility in replies they never see coming, with no manual appeal step described in the launch.
Second-order effects
- If the expansion holds, abusers adapt by cycling through new accounts to evade the seven-day block, pressuring Twitter's older problem of keeping banned users from rejoining.
- Rival platforms face user expectations calibrated to automated reply-level protection, pushing mute-and-block tooling from opt-in settings toward default behavior across social products.
Third-order effects
- Moderation is migrating from platform-enforced rules judged after posting to per-user algorithmic shields acting during conversation — a structural shift that makes individual feeds self-policing and reduces reliance on centralized takedowns.
- As automated blocks widen (the beta already covered 50% of US users within months), questions about false-positive silencing and transparency will move from product blogs into regulator territory.
The trend: Social platforms are automating harassment defense at the reply level, turning user safety from a settings page into an always-on algorithmic layer.