/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Anthropic details three infrastructure bugs that intermittently degraded Claude's responses between August and early September, and explains how it fixed them

This is a technical report on three bugs that intermittently degraded responses from Claude.  Below we explain what happened …

Anthropic

Context & Ripple Effects

This disclosure sits in a recurring record of Claude reliability and quality problems, including a later incident in which Anthropic said it had fixed three Claude Code quality causes and a separate period of elevated errors across Claude surfaces followed by a stated fix. The repeated pattern makes root-cause transparency consequential for users deciding whether degraded output is a model limitation, a product-policy change, or an operational fault.

First-order effects

  • Anthropic’s fixes should remove the identified intermittent failure modes for affected Claude users and give customers a more concrete explanation for the degraded responses.
  • The report distinguishes infrastructure-caused quality degradation from intentional model behavior, directly addressing a source of uncertainty for users evaluating Claude output.

Second-order effects

  • Enterprise and developer users may place greater weight on monitoring, validation, and fallback workflows when Claude quality shifts, rather than treating output changes as solely model-level behavior.
  • The disclosure creates a benchmark for competing AI providers: operational quality claims increasingly need incident explanations and remediation details, not just aggregate availability signals.

Third-order effects

  • As AI systems become embedded in production work, reliability management will extend from uptime to output quality: providers may be judged on whether they can detect, diagnose, and communicate silent degradations.
  • If such episodes persist across the sector, deployment governance will likely treat model quality, serving infrastructure, and product configuration as a single operational risk surface.

The trend: Frontier AI providers are moving toward operational accountability for output quality, not merely service availability.

Discussion

  • @_sholtodouglas Sholto Douglas on x
    We're sorry - and we'll do better. We're working hard on making sure we never miss these kind of regressions and rebuilding our trust with you.
  • @ctrlaltdwayne Dwayne on x
    @_sholtodouglas it's nice to see this being discussed sholto, but doesn't make it any less disappointing. many of us paying $200 a month for something that didn't work as well as it did 1 month ago. the fact despite the loud chorus of complaints anthropic said nothing. maybe it w…
  • @markshust Mark Shust on x
    Really appreciate this. You typically don't see this level of transparency and empathy from a company of this size.
  • @claudeai Claude on x
    We've published a detailed postmortem on three infrastructure bugs that affected Claude between August and early September. In the post, we explain what happened, why it took time to fix, and what we're changing:
  • @tszzl Roon on x
    I love Anthropic because they are apologizing for mildly degrading 0.8% of requests which is a normal Tuesday at most software companies
  • @krinetix1234 Krish Maniar on x
    Yet to have seen this detailed of a postmortem across other AI labs and startups. Many companies underestimate how much trust things like this build
  • @vikhyatk Vik on x
    @_sholtodouglas not hearing the word, “refund”
  • @aihsannergiz Ali Ihsan Nergiz on x
    @_sholtodouglas too late :( codex's timing and cursor price changes really made it very easy to break from using claude code and claude on cursor after paying both anthropic and cursor for more than a year. probably not gonna touch claude unless next version is insanely better
  • @rmedranollamas Ramón Medrano Llamas on x
    one of the best postmortem activity of the year. this shows where the SRE job is heading and why will remain relevant. and I have to say that's pretty damn great where we are traveling to. kudos to @AnthropicAI (I'm sure @tmu was here ;)
  • @timkellogg.me Tim Kellogg on bluesky
    Anthropic full postmortem  — long context routing bug  — output corruption bug  — approx top-k bug on TPU  —  plus: why it was so difficult and slow to fix  —  this is going to be a thread  —  www.anthropic.com/engineering/ ...
  • r/ClaudeAI r on reddit
    anthropic published a full postmortem of the recent issues - worth a read!
  • r/artificial r on reddit
    Anthropic outlines three infrastructure bugs that disrupted Claude's responses and how they were resolved