/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

← → days · ↑ ↓ browse · Enter similar · o open

Tech companies can do a lot more to protect the identities of people speaking in recordings used to train AI, like shifting the voice or gender of the speaker

It turns out that Siri, Alexa, and whatever it is you call Facebook Messenger have been a little loose-lipped with your conversations.

Slate April Glaser

Context & Ripple Effects

This Slate argument lands mid-spiral in the 2019 voice-assistant privacy run: months earlier, sources described how Amazon Alexa's AI-training pipeline exposed account numbers and private conversations to transcribing workers, and The Verge had flagged how voice-analysis research built on these recordings raises its own accuracy and privacy questions.

The piece's contribution is that the fix is technical, not just policy: shifting a speaker's voice or gender before clips reach reviewers would decouple model improvement from identity exposure. By December, Bloomberg's investigation of Amazon, Apple, Google, and Facebook confirmed the industry was still leaning on contractors with raw audio.

First-order effects

  • Contractors at Amazon, Apple, Google, and Facebook reviewing voice clips would lose access to identifiable speech if voice- or gender-shifting were applied upstream, directly shrinking the exposure documented in Alexa's training process.
  • Siri, Alexa, and Facebook Messenger owners face an immediate choice between adopting anonymization in their labeling pipelines and defending the current practice publicly.

Second-order effects

  • If one assistant maker ships voice-shifted training audio first, it turns reviewer privacy into a competitive claim the others must match, pressuring the shared contractor ecosystem that all four rely on.
  • Anonymized training corpora also cut both ways for users: the same voice-morphing capability that protects speakers is adjacent to the voice-cloning tools later shown defeating bank biometrics, sharpening scrutiny of who controls synthetic-voice technology.

Third-order effects

  • If the pattern holds, 'de-identify the training data' becomes a baseline expectation for any product that records humans — moving voice-assistant privacy from consent disclosures toward engineering requirements regulators can audit.
  • Assistants whose value depends on ambient listening would then compete on pipeline design rather than features alone, since the recordings feeding models are the liability surface.

The trend: Voice-assistant privacy is shifting from disclosure-and-consent debates toward technical anonymization of the recordings themselves as the default input to AI training.

Discussion

  • @aprilaser April Glaser on x
    I wrote about trading convenience for autonomy https://slate.com/...
  • @slate @slate on x
    Tech companies want to improve Siri, Alexa, and Google Assistant. Can they do it without listening to what we say around our devices? https://slate.com/...
  • @josephfcox Joseph Cox on x
    New: after we found Microsoft is hiring contractors to listen to some Skype calls, the company has updated its privacy policy and other pages to explicitly say humans/employees may listen to audio. Wasn't clear before. Original piece based on leaked docs https://www.vice.com/... …
  • @jason_koebler Jason Koebler on x
    NEW: Leaked documents from a Microsoft contractor show that the humans who listen to & transcribe your conversations have horrible, terrible, underpaid jobs http://www.vice.com/... https://twitter.com/...
  • @josephfcox Joseph Cox on x
    These parts of the training materials for contractors listening to Cortana recordings are... a bit on the nose https://www.vice.com/... https://twitter.com/...
  • @j2sheck Jordan Sheckman on x
    There's no way Cortana gets more than 200 audio clips per hour. https://twitter.com/...
  • @derektmead @derektmead on x
    All the big tech co's are hiring real humans to do the work of making sure their voice assistants are useful. How much do these contractors, who are left to listen to random users' voice conversations, get paid? Very very little! https://www.vice.com/...
  • @josephfcox Joseph Cox on x
    New: leaked docs show what it's really like behind scenes of a tech giant's artificial intelligence system. For Microsoft's Cortana, contractors paid $12-14/hour, with extra $1 if they hit a milestone. Training manuals of 100s of pages of monotonous tasks. https://www.vice.com/..…
  • @k_trendacosta Katharine Trendacosta on x
    Increasingly, every tech company is revealed to be the fucking train from Snowpiercer. Rip off the lid and it's not breakthrough technology, it's abusive labor practices. https://twitter.com/...
  • @josephfcox Joseph Cox on x
    Microsoft's privacy statement now says “Our processing of personal data for these purposes includes both automated and *manual (human) methods* of processing” While Facebook, Apple, Google have suspended use of humans though, Microsoft holding on https://www.vice.com/... https://…
  • @sarahfrier Sarah Frier on x
    Facebook confirms it has done this, but says it's not happening anymore: “Much like Apple and Google, we paused human review of audio more than a week ago.”