The Compute Reckoning: Anthropic Finally Admits What Customers Suspected for Ten Months

📊 Full opportunity report: The Compute Reckoning: Anthropic Finally Admits What Customers Suspected for Ten Months on ThorstenMeyerAI.com — validation score, market gap, and execution plan.

TL;DR

Anthropic has officially acknowledged that its recent customer issues stem from insufficient compute capacity. A new agreement with SpaceX’s Colossus 1 data center significantly boosts available resources, marking a strategic shift.

Anthropic has confirmed that its recent customer experience problems, including frequent rate limits and outages, were caused by a shortage of compute capacity. The company announced a new agreement with SpaceX to utilize the entire Colossus 1 data center in Memphis, which provides over 300 megawatts and more than 220,000 NVIDIA GPUs, effectively addressing the compute scarcity that had constrained its services for nearly a year.

On May 6, 2026, Anthropic revealed that the persistent degradation in customer experience—such as increased rate limits, outages, and throttling—was primarily due to insufficient compute infrastructure. For months, users faced restrictions that hampered productivity, with some paying subscribers hitting quotas within minutes. The company’s own April statement to Fortune acknowledged that demand had outstripped available capacity, notably during peak hours.

The new deal with SpaceX’s Colossus 1 data center, announced simultaneously, involves over 300 MW of power and more than 220,000 NVIDIA GPUs, making it one of the largest compute commitments in the AI industry. This capacity is expected to come online within the month, marking a significant step in resolving the longstanding supply-demand imbalance. The agreement is part of a broader strategy that includes commitments from Amazon, Google, Microsoft, and Fluidstack, totaling several gigawatts of AI infrastructure investments.

Industry insiders note that this move shifts Anthropic from a ‘compute-constrained challenger’ to a well-resourced frontier lab. The company’s previous struggles, including the rollout of rate limits and outages, were driven by compute scarcity rather than strategic or safety issues, according to sources familiar with the matter. OpenAI internal memos leaked to CNBC also suggest that Anthropic’s failure to secure enough compute was a strategic misstep, now being rectified.

The Compute Reckoning — Anthropic’s SpaceX Deal Closes Ten Months of UX Degradation
DISPATCH / MAY 2026 ANTHROPIC · SPACEX · COMPUTE RECKONING
▲ Breaking · T+0 Announced May 6, 2026
Anthropic + SpaceX · Compute Reckoning

Ten months. One admission.

Anthropic finally got the compute. The customer-experience problem was scarcity all along.

May 6, 2026 — Anthropic announced SpaceX Colossus 1 deal · 300+ MW · 220,000+ NVIDIA GPUs · online within May. Effective immediately: Claude Code 5-hour rate limits doubled. Peak-hour throttling removed. API limits up 1,500% input / 900% output for Opus on Tier 1. Closes ten-month UX degradation arc. Compute risk in IPO disclosure framework materially de-risked.

Announced
May 6yesterday · t+0
SpaceX Colossus 1 · 300+ MW · 220,000+ NVIDIA GPUs · online within May 2026 · all of facility’s compute capacity
Plus orbital ambition
multi-GW exploration
220K+
NVIDIA GPUs · SpaceX Colossus 1
300+ MW · online within May 2026
Claude Code 5-hour rate limits
Pro / Max / Team / Enterprise · effective May 6
+1,500%
API Tier 1 input tokens/min · Opus
+900% output · effective May 6
50/35/15
Next-90-days scenario probability
Bullish · Base · Bearish
MAY 6, 2026 ANTHROPIC + SPACEX COLOSSUS 1 · 300+ MW · 220K NVIDIA GPUS 10-MONTH ARC JULY 2025 WEEKLY LIMITS → MARCH 2026 PEAK THROTTLING → MAY 2026 RESET RATE LIMITS CLAUDE CODE 5HR DOUBLED · PEAK-HOUR THROTTLING REMOVED FOR PRO/MAX API JUMPS +1,500% INPUT / +900% OUTPUT TIER 1 OPUS · EFFECTIVE IMMEDIATELY RIVAL COOPERATION SPACEX/XAI MEMPHIS FACILITY · DIRECT COMPETITOR PROVIDES COMPUTE ORBITAL AMBITION MULTI-GW IN SPACE · SOLVES TERRESTRIAL POWER CONSTRAINT MAY 6, 2026 ANTHROPIC + SPACEX COLOSSUS 1 · 300+ MW · 220K NVIDIA GPUS 10-MONTH ARC JULY 2025 WEEKLY LIMITS → MARCH 2026 PEAK THROTTLING → MAY 2026 RESET
Ten-month UX degradation arc

Nine moments. One constraint.

For ten months, Claude users experienced compute scarcity as broken product. Anthropic experienced it as the binding constraint on growth. May 6 closes the gap — at the announcement level. Verification follows.

UX degradation arc · July 2025 → May 2026
From weekly rate limits to peak-hour throttling to compute reckoning.
Jul 2025
Weekly rate limits introducedPro/Max users running Claude Code in background. Framing: “<5% affected." Reality: power users hit constantly.
Constraint
Oct 9, 2025
Discord mega-thread documents discontentSubscribers paying $100-200/mo report hitting limits faster than expected. Anthropic largely silent through Q4.
Backlash
Dec 25-31, 2025
Holiday usage doublingLimits doubled during Christmas-New-Year. Framing: “holiday gift.” Structural admission: idle enterprise capacity revealed baseline rationing.
Tell
Jan 4, 2026
Post-holiday revert · bug reportsAnthropic dismisses “unfounded” complaints. Discord amplifies — paying customers get worse product in January than December.
Friction
Mar 13-28, 2026
Off-peak doubling promotionLimits doubled during off-peak only. Structural admission: peak-hour compute is binding constraint. Time-of-day rationing as management tool.
Tell
Mar 26, 2026
Peak-hour throttling officially admittedThariq Shihipar on X: “5-hour session limits adjusted during peak hours.” First explicit official acknowledgment compute scarcity drives UX changes.
Admission
Mar-Apr 2026
Max users hit quota in 19 minutes$200/mo Max subscribers exhaust 5-hour quota in ~19 minutes. Anthropic acknowledges “investigating.” Bug + capacity rationing.
Crisis
Apr 24, 2026
Fortune publishes performance-decline analysisFull pattern visible. Anthropic statement: “infrastructure stretched, particularly at peak hours.” OpenAI memo: “strategic misstep” / “smaller curve.”
Public
May 6, 2026
SpaceX deal · the reset300+ MW · 220K+ GPUs · online within May. Rate limits doubled. Peak-hour throttling removed. API limits +900-1,500%. Ten-month arc closes — at announcement level.
Reset
Compute scarcity drove ten months of UX degradation. May 6 is the inflection.
Compute portfolio · five partnerships
NVIDIA 900-2G610-0000-000 Tesla P40 24GB GDDR5 PCIE 3.0 X16 Passive Cooling

NVIDIA 900-2G610-0000-000 Tesla P40 24GB GDDR5 PCIE 3.0 X16 Passive Cooling

Series: Tesla P40, Model: 900-2G610-0000-000

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Five partnerships. One arms race.

Anthropic now operates the second-largest publicly disclosed compute portfolio of any frontier lab — behind only Microsoft-OpenAI. Multi-vendor by design: Trainium + TPU + NVIDIA + custom · five major partners · multi-jurisdictional.

Anthropic compute portfolio · five major partnerships
SpaceX added May 6 to existing Amazon · Google · Microsoft · Fluidstack commitments.
Partner Detail Scale Status
SpaceXColossus 1 · Memphis
All compute capacity at xAI/SpaceX Memphis facility. Direct rival cooperation — unusual.
300+ MW220K+ GPUs
May 2026
Amazon (AWS)Trainium primary
Up to 5 GW agreement. Nearly 1 GW of new capacity by end of 2026. Inference in Asia and Europe.
Up to 5 GW~1 GW in 2026
2026-30
Google + BroadcomTPU + custom silicon
5 GW agreement. Begins coming online 2027. Multi-year capacity commitment.
5 GW2027 start
2027+
Microsoft + NVIDIAAzure capacity
Strategic partnership. $30B Azure capacity commitment. NVIDIA hardware focus.
$30BAzure capacity
2026-28
FluidstackAmerican AI infrastructure
$50B investment in American AI infrastructure. US-resident compute commitment.
$50BUS infrastructure
2026-30
SpaceX orbitalSpeculative · exploration
Multi-gigawatt orbital AI compute capacity. Bypasses terrestrial power constraint.
Multi-GWaspirational
2028+ spec
Three scenarios · next 90 days
NVIDIA Tesla V100 (Volta) 32GB NVLINK 2.0 SXM2 GPU

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Three scenarios. Verification follows.

50/35/15 probability allocation. The May 6 announcement either delivers on customer experience improvements or doesn’t. Setup factors favor bullish: SpaceX execution capability, IPO incentive alignment.

Three scenarios · how May 6 resolves through Q3 2026
Bullish · Base · Bearish. Probability allocation 50/35/15.
▲ Bullish · capacity delivers
50%
Capacity delivers; UX dramatically improves.
  • Online May 2026SpaceX capacity as announced.
  • UX improvements stickDoubled limits, no peak throttle.
  • Trust rebuilds Q3ARR growth continues.
  • IPO Q4 2026 catalyzesPositive market response.
  • Outcome: Compute reckoning is start of positive arc.
▶ Base · partial delivery
35%
Most capacity arrives; gaps remain.
  • Some delayCapacity partial through May.
  • Mostly deliversSome peak-period gaps.
  • Trust rebuild slowerThrough Q3-Q4.
  • IPO early 2027Pushed if needed.
  • Outcome: Continuation trajectory with friction.
▼ Bearish · implementation gap
15%
Implementation gap; trust deficit persists.
  • Capacity lateOr arrives in pieces.
  • Partial improvementsIssues recur in different form.
  • Competitive erosionOpenAI / Google gain share.
  • IPO substantially delayedOr repriced.
  • Outcome: Trust deficit compounds. Multi-quarter rebuild.

The era of “build your own compute” yields to “share compute across rival workloads when economics support it.” SpaceX/xAI’s flagship Memphis facility leases to a direct competitor — that’s how severe compute scarcity has become across the AI lab category.

— The structural read · May 2026
What to do this quarter · through Q2-Q3 2026
The Scaling Era: An Oral History of AI, 2019–2025

The Scaling Era: An Oral History of AI, 2019–2025

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Four assignments. By role.

Claude Users

Verify actual delivery vs announced.

Test the doubled rate limits in your workflow. Monitor performance through May-June. Consider whether to retain, upgrade, or cancel based on demonstrated improvement rather than announced improvement. The trust deficit from ten months of degradation requires sustained performance to repair. Anthropic has incentive to deliver — IPO timing depends on it.

API Developers

Re-architect for new headroom.

1,500% input / 900% output Tier 1 increase is substantial. Scale rate-limit-bottlenecked applications. The structural implication: Anthropic now competitive with OpenAI on API capacity, narrowing what had been meaningful OpenAI advantage. Document delivered vs announced capacity in your monitoring.

IPO Investors

Update models · compute risk de-risked.

The compute risk factor in the Anthropic IPO disclosure framework is materially de-risked. Q3-Q4 2026 IPO window becomes more credible. Valuation case strengthens — $30B ARR, $400-500B precedent from frontier-lab benchmarks, credible compute portfolio. Position based on demonstrated delivery through Q2-Q3 2026.

NVIDIA Demand

Direct demand validation for Q1 FY27 print.

220K+ GPUs from SpaceX deal alone. Aggregate NVIDIA-attributable demand from Anthropic’s compute portfolio plausibly $20-40B over 2026-2028. NVIDIA Q1 FY27 dispatch bull case gets concrete numbers. Hyperscaler capex thesis demand-pull validation gets specific evidence. Watch May 20 print for confirmation.

  • The Anthropic IPO Disclosure Document
  • The $725B Hyperscaler Capex Question
  • The NVIDIA Q1 FY27 Earnings Preview
  • The Bubble Question, Disentangled
  • Anthropic · Higher usage limits + SpaceX deal · May 6, 2026
  • Yahoo Finance · Anthropic SpaceX compute deal · May 6, 2026
  • CNBC · Anthropic-SpaceX compute deal includes space development · May 6
  • Fortune · Anthropic explains Claude Code performance decline · April 2026
  • The Register · Anthropic admits Claude Code quotas running too fast · March 31
  • TechRadar / MacRumors / DevOps · Peak-hour throttling coverage · March 2026
  • OpenAI internal memo (CNBC) · “strategic misstep” framing
  • Anthropic ARR · $30B run rate (Fortune Apr 2026) · 3× growth in 12 months
Colophon

Set in Lora, Plus Jakarta Sans, & JetBrains Mono. Composed for ThorstenMeyerAI.com, May 2026. Free to embed with attribution.

thorstenmeyerai.com

InfiniBand XDR 800G For AI & HPC Clusters: Configure RDMA, GPU Networking OpenSM, NCCL, And Low-Latency Data Center Fabrics

InfiniBand XDR 800G For AI & HPC Clusters: Configure RDMA, GPU Networking OpenSM, NCCL, And Low-Latency Data Center Fabrics

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Implications for Anthropic’s Market Position

This development significantly alters Anthropic’s strategic landscape. By addressing its compute limitations, the company can now focus on product innovation and scaling without the previous constraints. The capacity boost also reduces the risk factors associated with its upcoming IPO, as investors can now view it as less vulnerable to infrastructure shortages. Additionally, the deal with SpaceX signals a potential move towards orbital AI compute, which could redefine how AI models are trained and deployed at scale.

For users, the immediate benefit is a reduction in throttling and outages, leading to more reliable access to Claude models. For the broader AI industry, the announcement underscores the importance of compute infrastructure as a core competitive factor, highlighting the shift from scarcity-driven limitations to resource abundance.

Background of Compute Scarcity and Industry Response

Throughout 2025, Anthropic faced mounting criticism over deteriorating user experiences, including frequent rate limits, outages, and delayed model improvements. Starting in July 2025, weekly rate limits were introduced, followed by peak-hour throttling in March 2026. Subscribers paying up to $200/month reported hitting quotas within minutes, and the discourse around ‘Claude being dumber on Tuesdays’ became widespread.

Sources close to the company revealed that demand for Claude grew exponentially, outpacing infrastructure. In April, Anthropic acknowledged that its infrastructure was stretched during peak hours, a situation that internal memos from OpenAI described as a ‘strategic misstep.’ Prior commitments from Amazon, Google, Microsoft, and Fluidstack indicated industry awareness of the compute bottleneck, but Anthropic’s situation was more acute due to its rapid growth and limited initial capacity.

The May 6 announcement marks a turning point, as it confirms that the root cause was compute scarcity and demonstrates a strategic pivot to secure substantial new capacity.

“We are committed to providing reliable, scalable AI services and have secured significant compute capacity through our partnership with SpaceX to meet growing demand.”

— Anthropic spokesperson

Remaining Questions About Future Capacity and Strategy

It is not yet clear how quickly the new capacity from SpaceX will fully integrate into Anthropic’s infrastructure or how it will impact future model performance and safety measures. The long-term orbital compute ambitions remain speculative, with no confirmed timelines or technical details. Additionally, the extent to which this capacity will prevent future throttling or outages is still to be validated through operational data.

Next Steps for Capacity Deployment and Product Scaling

Anthropic is expected to activate the new capacity within the coming weeks, with user-facing improvements likely to follow shortly thereafter. The company will likely report on operational metrics and user experience enhancements in its upcoming quarterly disclosures. Industry analysts will watch for how this capacity influences Anthropic’s product development, safety protocols, and IPO prospects. Further, the company may explore orbital AI compute initiatives, though details remain unconfirmed.

Key Questions

How does the new SpaceX deal impact Anthropic’s ability to scale?

The deal provides over 300 MW and 220,000 GPUs, dramatically increasing capacity and reducing previous constraints, enabling faster scaling and more reliable service.

Will this capacity expansion eliminate all outages and throttling?

While the capacity increase should significantly improve reliability, it remains to be seen how effectively it addresses peak-hour demand and future growth. Operational data will clarify this in upcoming months.

What does the orbital AI compute ambition involve?

Anthropic has expressed interest in developing multi-gigawatt orbital AI compute capacity, but details, timelines, and technical feasibility are still speculative at this stage.

How does this affect Anthropic’s competitive position?

Securing substantial compute resources narrows the gap with competitors like OpenAI and positions Anthropic as a better-resourced player, potentially influencing market dynamics and IPO timing.

What are the implications for safety and model development?

With increased capacity, Anthropic can focus more on safety and model improvements without the previous infrastructure constraints, though safety protocols remain a priority.

Source: ThorstenMeyerAI.com

You May Also Like

When AI Builds Itself: Inside Anthropic’s Evidence on Recursive Self-Improvement

Anthropic reports measurable acceleration in AI’s ability to build and improve itself, based on internal data and public benchmarks, raising questions about future AI development.