/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Due to safety concerns like facial recognition abuse, OpenAI hasn't widely shipped GPT-4's “multimodal” capability that can respond to images and text prompts

An advanced version of ChatGPT can analyze images and is already helping the blind.

New York Times Kashmir Hill

Context & Ripple Effects

When GPT-4 launched, OpenAI held back its image-and-text mode rather than shipping it alongside the text model, citing facial-recognition abuse among the risks. The related coverage shows how the gap closed: a hands-on test of ChatGPT's image features found the shipped version simply refuses to discuss faces — the guardrail that made release possible.

The eventual wide rollout came via GPT-4o's text and image capabilities for ChatGPT Plus and Team users, and OpenAI paired the push with a formal Safety and Security Committee, institutionalizing the review process this withholding foreshadowed.

First-order effects

  • Blind users are the immediate beneficiaries: an advanced ChatGPT version that can analyze images is already in use as a visual aid, making accessibility the beachhead market for a capability held back from everyone else.
  • Paying ChatGPT subscribers gain image analysis ahead of the general public, with face-related queries blocked by design rather than left to user discretion.

Second-order effects

  • Rival labs shipping their own vision features now have to match not just the capability but the refusal policy — 'won't discuss faces' becomes a competitive baseline, not a differentiator.
  • Accessibility tools built on ChatGPT's vision get a capability moat competitors can only close by accepting the same facial-recognition restrictions, shaping what assistive products can promise.

Third-order effects

  • If the pattern holds, frontier capabilities ship in gated stages — restricted preview, paid tiers, general release — with safety committees deciding the cadence, turning release timing itself into a governance instrument.
  • Refusal-based guardrails like the face ban become the template regulators and enterprises expect when evaluating whether a multimodal model is safe to deploy.

The trend: Frontier labs are shifting from shipping capabilities all at once to staged, safety-gated rollouts where the guardrail policy — not just the model — defines the product.

Discussion

  • @epro.social Emil Protalinski on bluesky
    “What OpenAI doesn't want ChatGPT to become is a facial recognition machine.”  —  “Google offers an opt-out for well-known people who don't want to be recognized, and OpenAI is considering that approach.”  —  Tech companies love playing the “well, you can always opt-out” game bec…
  • @kashhill.bsky.social @kashhill.bsky.social on bluesky
    Blind users have had access for months to a version of ChatGPT that analyzes images.  One user described it as extraordinary.  But it recently started blurring people's faces.  I talked to OpenAI about why: https://www.nytimes.com/...
  • @glenngabe Glenn Gabe on x
    And in case you are wondering, you can upload images to Bard and use Lens, but if it's a person, that image is removed and Bard explains it cannot provide info about people yet. So Google is in alignment with OpenAI here. [image]
  • @kashhill @kashhill on x
    Blind users have had access for months to a version of ChatGPT that analyzes images. One user described it as extraordinary. But it recently started blurring people's faces so it can't give information about them. I talked to OpenAI about why: https://www.nytimes.com/...
  • @marietjeschaake Marietje Schaake on x
    OpenAI worries a tool intended for blind people would share things incl gender or emotional state. OpenAI is ‘figuring out how to address these and other safety concerns before releasing the image analysis feature’> These shouldn't be internal decisions ↘️ https://www.nytimes.com…
  • @garymarcus Gary Marcus on x
    Yet again confirming my 2001 hypothesis that neural networks would hallucinate until they bridged gap with symbols, new reports show that GPT-4's unreleased visual system hallucinates objects (eg remote control buttons that don't exist; absent faces, etc). https://www.nytimes.com…
  • @katecrawford Kate Crawford on x
    OpenAI says it will follow public opinion on whether GPT4 should include face, emotion and gender recognition. Which public? In which nations? Whose opinions will count? https://www.nytimes.com/...
  • @random_walker Arvind Narayanan on x
    It's surprising that safety concerns have delayed the release by 4+ months, considering that Bing has already rolled it out with what appears to be a simple fix (face blur). A different approach than the much more cavalier release of Code Interpreter. https://twitter.com/...
  • @random_walker Arvind Narayanan on x
    The reason I predicted this is that facial recognition is “merely” extreme multiclass image classification, so if you train a model to describe images there isn't a strong reason it won't also learn faces. Still, facial recognition being emergent behavior feels creepy and sci-fi.
  • @random_walker Arvind Narayanan on x
    When @kashhill asked me a while back why OpenAI hadn't released multimodal GPT-4 yet, I wildly speculated that it's because it can do facial recognition and OpenAI doesn't want that out in the wild. Well... https://www.nytimes.com/...
  • @tomwarren Tom Warren on x
    Bing Chat got a new search by image feature today so I asked it to identify Frank and wow [image]
  • @kashhill @kashhill on x
    In the old blogging days I would have had a “hat tip” at the bottom of this article to @random_walker