---
title: The AI factory became a power-and-fabric problem, not a model-release problem.
publication: The AI Stack Weekly
slug: 2026-W23
issueNumber: 7
isoYear: 2026
isoWeek: 23
publishedAt: '2026-06-06'
canonicalUrl: https://brianletort.ai/industry/weekly/2026-W23
pdfUrl: https://brianletort.ai/downloads/ai-stack-weekly-2026-W23.pdf
schemaVersion: 2026.05.02
flywheelArc: all-three
capitalFlow:
  - category: Frontier Labs
    capitalIn: ~$90B
    capitalInPrior: ~$90B
    capitalInDirection: flat
    revenueOut: ~$20B
    revenueOutPrior: ~$20B
    revenueOutDirection: flat
    burnToRevenue: ~1.3x
  - category: Hyperscaler-Hosted
    capitalIn: ~$181B
    capitalInPrior: ~$180B
    capitalInDirection: flat
    revenueOut: ~$60B
    revenueOutPrior: ~$60B
    revenueOutDirection: flat
    burnToRevenue: ~0.3x
  - category: Neoclouds
    capitalIn: ~$12B
    capitalInPrior: ~$12B
    capitalInDirection: flat
    revenueOut: ~$5B
    revenueOutPrior: ~$5B
    revenueOutDirection: flat
    burnToRevenue: ~3x
  - category: On-Prem / Hybrid
    capitalIn: ~$91B
    capitalInPrior: ~$90B
    capitalInDirection: flat
    revenueOut: ~$35B
    revenueOutPrior: ~$35B
    revenueOutDirection: flat
    burnToRevenue: ~2x
levers:
  - metric: Frontier lab cash position (avg months runway, top 3)
    current: ~33-36 mo
    prior: ~33-36 mo
    direction: flat
    threshold: <18 mo triggers re-rating risk
  - metric: Hyperscaler capex / AI revenue ratio (top 4 weighted)
    current: ~5.0-5.2
    prior: ~5.0-5.2
    direction: flat
    threshold: '>6.0 invites investor pushback at next earnings'
  - metric: CoreWeave revenue backlog
    current: $99.4B
    prior: $99.4B
    direction: flat
    threshold: Conversion velocity matters more than gross figure
  - metric: NVIDIA Q-over-Q data center revenue
    current: $75.2B (Q1 FY27); Rubin production ramp confirmed
    prior: $75.2B (Q1 FY27)
    direction: up
    threshold: Q2 FY27 guide $91B implies further +21% QoQ
  - metric: Open vs closed gap on SWE-Bench Pro (coding)
    current: Closed +~19pp (no new Pro challenger yet)
    prior: Closed +~19pp (audit caveat)
    direction: flat
    threshold: Sustained open lead reshapes enterprise procurement
  - metric: Sovereign AI commitments (count / aggregate $)
    current: ~13 / ~$160B+; power-first gating rising
    prior: ~13 / ~$160B+
    direction: flat
    threshold: null
  - metric: PJM 2026/27 capacity auction price ($/MW-day)
    current: $329.17
    prior: $329.17
    direction: flat
    threshold: 11x in 24 months — power is the new binding constraint
  - metric: Time-to-power, busiest US markets (months)
    current: 60-84 (new PJM); power-first campuses rising
    prior: 60-84 (new PJM); 36-48 (existing PJM queue)
    direction: flat
    threshold: null
  - metric: Cost-per-task, frontier reasoning model
    current: ~$0.10-$0.15 (effective; unchanged)
    prior: ~$0.10-$0.15 (effective)
    direction: flat
    threshold: null
  - metric: Custom silicon share of incremental AI compute
    current: ~33-36%; Broadcom AI revenue +143% YoY
    prior: ~33-36%
    direction: up
    threshold: '>35% materially compresses merchant GPU pricing'
predictions:
  - id: p32-gemini-3-5-pro-ga
    lens: software
    confidencePct: 60
    deadline: By June 30, 2026
    text: >-
      Gemini 3.5 Pro reaches public GA by June 30, 2026, but does not exceed Claude Opus 4.8 on
      SWE-Bench Pro in its first independent Artificial Analysis run.
  - id: p33-vera-rubin-first-shipments
    lens: hardware
    confidencePct: 70
    deadline: By September 30, 2026
    text: >-
      At least one major OEM announces customer shipment or formal order availability for Vera Rubin
      NVL72-class systems before September 30, 2026.
  - id: p34-hbm4-allocation-tightness
    lens: hardware
    confidencePct: 65
    deadline: By August 31, 2026
    text: >-
      Before August 31, 2026, at least one memory supplier or supply-chain analyst reports HBM4
      allocation tightness despite three-supplier qualification.
  - id: p35-cpo-design-win
    lens: networking
    confidencePct: 65
    deadline: By August 31, 2026
    text: >-
      Broadcom, Marvell, or NVIDIA announces a new CPO/1.6T production design win or revenue guide
      uplift tied to AI networking before August 31, 2026.
  - id: p36-power-first-followthrough
    lens: power
    confidencePct: 60
    deadline: By September 30, 2026
    text: >-
      A hyperscaler announces another >500MW power-first AI campus or behind-the-meter generation
      deal by September 30, 2026.
predictionsPrior:
  - id: p27-anthropic-s1-public
    lens: capital
    outcome: pending
    deadline: By August 31, 2026
    text: >-
      No frontier lab (Anthropic or OpenAI) files a publicly visible S-1 on SEC EDGAR before August
      31, 2026, keeping the IPO race at the confidential-DRS stage.
  - id: p28-gemini-3-5-pro-june
    lens: software
    outcome: pending
    deadline: By June 30, 2026
    text: >-
      Gemini 3.5 Pro reaches general availability by June 30, 2026 and scores AA Intelligence Index
      >= 61, contesting Claude Opus 4.8's fresh lead.
  - id: p29-vera-rubin-cadence
    lens: hardware
    outcome: pending
    deadline: By June 7, 2026
    text: >-
      At GTC Taipei / Computex (June 1), NVIDIA reaffirms Vera Rubin production starting in 2H 2026
      and frames HBM4 + CoWoS as the binding supply constraint rather than demand.
  - id: p30-optics-design-wins
    lens: networking
    outcome: pending
    deadline: By August 31, 2026
    text: >-
      At least two of (Credo, Marvell, Broadcom) cite co-packaged-optics or 1.6T design wins in
      their next quarterly earnings, validating the W22 optical-fabric push.
  - id: p31-sovereign-power-followthrough
    lens: power
    outcome: pending
    deadline: By August 31, 2026
    text: >-
      A major hyperscaler or sovereign program announces a new behind-the-meter or >1GW
      power-procurement deal (SMR, gas, or grid) by August 31, 2026, as time-to-power stays the
      binding US constraint.
signalScores:
  - 5
  - 4
  - 2
  - 4
keyTakeaways:
  - >-
    NVIDIA moved Vera Rubin from roadmap to full production with a fall/Q3 shipment path, and all
    three memory suppliers are qualified for HBM4 — shifting the bottleneck from silicon to
    delivering power, memory, optics, and operator software together.
  - >-
    The closed frontier was quiet (Gemini 3.5 Pro still not GA at week's end) while open weights
    widened in the efficient-agent layer: Mellum2, Cosmos 3, and Holo3.1 all target deployable
    sub-agents, physical-AI reasoning, or local computer use.
  - >-
    Networking moved into the same frame as HBM4: Broadcom's AI semiconductor revenue rose 143% YoY
    with networking nearly 40% of AI revenue, and Spectrum-X Ethernet Photonics (CPO) entered
    production.
  - >-
    Capital flow's visible battleground was energy procurement: Google/Intersect's power-first
    campus model pairs AI load with more than 1GW of dedicated generation before servers are
    ordered.
  - >-
    Signal-vs-noise: enterprise application vendors are converging on governed agents with identity,
    permissions, and workflow authority; Gemini 3.5 Pro displacement claims are noise for this
    window.
  - >-
    Watch HBM4 allocation and first Vera Rubin customer-shipment evidence through August — volume
    and yield decide whether the fall ramp is broad or supply-rationed.
byTheNumbers:
  - value: +143%
    label: Broadcom AI semiconductor revenue growth YoY
  - value: '>1GW'
    label: Dedicated generation in Google/Intersect's power-first campus model
  - value: 350+
    label: Factories across 30 countries manufacturing the Vera Rubin platform
  - value: $99.4B
    label: CoreWeave revenue backlog (audited, as of Mar 31)
  - value: 60-84
    label: Months from new-load interconnection request to energization in PJM
  - value: $329.17
    label: PJM 2026/27 capacity price, cleared at the FERC cap ($/MW-day)
---

# The AI factory became a power-and-fabric problem, not a model-release problem.

*Issue 07 · Week 23 of 2026 · Published 2026-06-06*

## Executive summary

- NVIDIA moved Vera Rubin from roadmap to full production with a fall/Q3 shipment path, and all three memory suppliers are qualified for HBM4 — shifting the bottleneck from silicon to delivering power, memory, optics, and operator software together.
- The closed frontier was quiet (Gemini 3.5 Pro still not GA at week's end) while open weights widened in the efficient-agent layer: Mellum2, Cosmos 3, and Holo3.1 all target deployable sub-agents, physical-AI reasoning, or local computer use.
- Networking moved into the same frame as HBM4: Broadcom's AI semiconductor revenue rose 143% YoY with networking nearly 40% of AI revenue, and Spectrum-X Ethernet Photonics (CPO) entered production.
- Capital flow's visible battleground was energy procurement: Google/Intersect's power-first campus model pairs AI load with more than 1GW of dedicated generation before servers are ordered.
- Signal-vs-noise: enterprise application vendors are converging on governed agents with identity, permissions, and workflow authority; Gemini 3.5 Pro displacement claims are noise for this window.
- Watch HBM4 allocation and first Vera Rubin customer-shipment evidence through August — volume and yield decide whether the fall ramp is broad or supply-rationed.

**By the numbers.**

- **+143%** — Broadcom AI semiconductor revenue growth YoY (Networking nearly 40% of AI revenue; demand described as insatiable)
- **>1GW** — Dedicated generation in Google/Intersect's power-first campus model (Energy development is now part of AI capacity procurement)
- **350+** — Factories across 30 countries manufacturing the Vera Rubin platform (Five-rack AI factory reference, fall/Q3 shipments planned)
- **$99.4B** — CoreWeave revenue backlog (audited, as of Mar 31) (Conversion velocity matters more than the gross figure)
- **60-84** — Months from new-load interconnection request to energization in PJM (Substation transformer lead times ticked up from ~150 to >160 weeks)
- **$329.17** — PJM 2026/27 capacity price, cleared at the FERC cap ($/MW-day) (Budget capacity at-cap through 2028)

## Big Story

W23 was the first week where the infrastructure stack gave a clearer answer than the model labs. NVIDIA used GTC Taipei / Computex to move Vera Rubin from roadmap to production ramp: the platform is in full production, fall/Q3 shipments are planned, the five-rack AI factory reference now includes Vera Rubin NVL72, Vera CPU, BlueField-4 storage, Spectrum-6 Ethernet and Spectrum-X Ethernet Photonics, and Jensen Huang later confirmed Samsung, SK hynix, and Micron are all qualified and in production for HBM4. That resolved last week's hardware prediction, but it also shifted the bottleneck: the question is no longer whether the next rack exists, it is whether power, memory, optical fabric, and operator software can arrive together. On software, the closed frontier was quiet — Gemini 3.5 Pro still had not GA'd by the end of the window — while open weights widened in the efficient-agent layer: JetBrains Mellum2, NVIDIA Cosmos 3, and Holo3.1 all targeted deployable sub-agents, physical-AI reasoning, or local computer-use rather than a monolithic chatbot benchmark. On applications, Microsoft Scout, Salesforce Coworker, ServiceNow Otto, Wordsmith, and Stilta all pointed at the same control-plane fight: governed agents with identities, permissions, and workflow authority. Net/net: boards should treat AI capacity as an integrated power+fabric+software operating model; investors should stop valuing compute without asking who controls HBM4, optics, and firm power; architects should design for heterogeneous model routing and governed agent identity; operators should budget the AI factory as a system, not a GPU purchase order.

Flywheel arc: `all-three`.

## Software lens

- **Jun 1.** JetBrains released Mellum2, an Apache-2.0 12B sparse MoE with 2.5B active parameters per token, positioned for low-latency routing, RAG, summarization, validation, sub-agents, and private text/code deployments _([Hugging Face JetBrains Mellum2 launch](https://huggingface.co/blog/JetBrains/mellum2-launch))_
- **Jun 1.** NVIDIA released Cosmos 3 on Hugging Face as an open omni-model for physical-AI reasoning and action, with Nano 16B and Super 64B variants plus Diffusers integration and synthetic-data workflows _([Hugging Face NVIDIA Cosmos 3 launch](https://huggingface.co/blog/nvidia/cosmos-3-for-physical-ai))_
- **Jun 2.** H Company released Holo3.1 for local computer-use agents, adding 0.8B / 4B / 9B / 35B-A3B sizes plus FP8, Q4 GGUF and NVFP4 checkpoints for private deployment _([Hugging Face Holo3.1 launch](https://huggingface.co/blog/hcompany/holo31))_

**What this means.** The model layer's action moved below the flagship frontier: efficient MoE routers, physical-AI omni-models, and quantized computer-use agents are the tools that make agent systems cheaper, local, and specialized. Architects should route cheap sub-agent work to open/local models and reserve Opus/GPT/Gemini-class spend for high-risk reasoning, because the software flywheel is now about orchestration economics as much as raw intelligence.

## Hardware lens

- **Jun 1.** NVIDIA announced Vera Rubin is in full production, with a five-rack platform spanning Vera Rubin NVL72, Vera CPU, BlueField-4 STX storage, Spectrum-6 SPX Ethernet, and partner manufacturing across 350+ factories and 30 countries _([NVIDIA Newsroom, GTC Taipei](https://nvidianews.nvidia.com/news/vera-rubin-full-production-agentic-ai-factory))_
- **Jun 1.** GTC Taipei positioned DSX OS as the lifecycle, health, resiliency, and multi-tenant operating layer for AI factories, shifting attention from rack shipment to fleet operations _([Data Center Knowledge GTC Taipei coverage](https://www.datacenterknowledge.com/data-center-chips/nvidia-says-vera-rubin-vera-cpu-on-track-launches-dsx-os-to-run-ai-factories))_
- **Jun 5.** Jensen Huang confirmed Samsung, SK hynix, and Micron are all qualified and in production for Vera Rubin HBM4, resolving the near-term supplier uncertainty around the Q3/fall ramp _([TechTimes summary of Reuters/Bloomberg remarks](https://www.techtimes.com/articles/317855/20260605/nvidia-vera-rubin-hbm4-jensen-huang-confirms-all-three-suppliers-production-q3-ship.htm))_

**What this means.** The hardware read changed from 'will Rubin be on schedule?' to 'can the whole AI factory be delivered as a coordinated system?' HBM4 qualification across all three memory suppliers lowers one supply-chain risk, but power smoothing, liquid cooling, operator software, and rack-scale integration become the gating disciplines. Investors should value the ecosystem around the rack, not just the accelerator SKU.

## Networking lens

- **Jun 1.** NVIDIA said Spectrum-X Ethernet Photonics, a CPO-based switch platform with 200Gb/s SerDes, is now in production as part of the Vera Rubin AI factory fabric _([NVIDIA Newsroom](https://nvidianews.nvidia.com/news/vera-rubin-full-production-agentic-ai-factory))_
- **Jun 3.** Marvell framed CPO and 1.6T optical DSPs as the next AI connectivity bottleneck, citing a CPO switch design, 100T Ethernet switch work, and NVIDIA partnership around optics, photonics, and NVLink Fusion _([DataCenterNews Asia](https://datacenternews.asia/story/marvell-targets-ai-connectivity-bottleneck-with-nvidia-boost))_
- **Jun 6.** Broadcom reported AI semiconductor revenue up 143% YoY, with networking nearly 40% of AI revenue and demand for XPUs plus networking described as insatiable _([SDxCentral Broadcom Q2 FY2026 earnings coverage](https://www.sdxcentral.com/news/broadcom-bets-big-on-ai-infrastructure-as-networking-and-xpu-demand-hits-insatiable-levels/))_

**What this means.** Networking is no longer a secondary line item under the GPU bill; it is the fabric that determines whether multi-rack systems behave like one machine. The week put CPO, 1.6T/3.2T optics, and AI Ethernet economics into the same frame as HBM4. Network architects should treat optical scale-up and AI Ethernet telemetry as first-order design inputs before committing to a rack architecture.

## Capital flow

| Category | Capital in | Revenue out | Burn:Revenue | Movement |
|---|---|---|---|---|
| Frontier Labs (OpenAI, Anthropic, Google DeepMind, xAI) | ~$90B (was ~$90B, flat) | ~$20B (was ~$20B, flat) | ~1.3x | No new mega-round closed; the action moved from lab balance sheets to the infrastructure stack those labs consume. |
| Hyperscaler-Hosted (Azure-OpenAI, AWS-Anthropic, Google Cloud-Gemini, Oracle-OCI) | ~$181B (was ~$180B, flat) | ~$60B (was ~$60B, flat) | ~0.3x | Google's power-first Texas campus made energy procurement the visible hyperscaler battleground. |
| Neoclouds (CoreWeave, Nscale, Crusoe, Lambda, Fluidstack, IREN) | ~$12B (was ~$12B, flat) | ~$5B (was ~$5B, flat) | ~3x | No new W23 financing reset; prior IREN/Microsoft-style deals remain the relevant neocloud proof point. |
| On-Prem / Hybrid (Enterprise GPU clusters, sovereign and national programs, Cisco / Dell / HPE) | ~$91B (was ~$90B, flat) | ~$35B (was ~$35B, flat) | ~2x | Sovereign AI infrastructure moved from compute ambition to power-first site selection and public-consent risk. |

### Frontier Labs — detail
Frontier-lab capital remained elevated but quiet after the prior week's record close. The actionable read is that lab demand has translated into Vera Rubin early-adopter lists and HBM4 qualification pressure rather than a fresh financing event. Treat this row as capital already committed to capacity, not a new inflow.
**Transactions:**
  - **Jun 1.** Vera Rubin early adopter list includes major frontier labs — Undisclosed capacity demand _([Data Center Knowledge / NVIDIA GTC Taipei](https://www.datacenterknowledge.com/data-center-chips/nvidia-says-vera-rubin-vera-cpu-on-track-launches-dsx-os-to-run-ai-factories))_

### Hyperscaler-Hosted — detail
The week's hyperscaler movement was strategic rather than a new earnings print: Google/Intersect's power-first data-center model pairs AI load with more than 1GW of dedicated generation. That is a capital-flow signal because future compute commitments increasingly require power development, land, and grid strategy before servers are ordered.
**Transactions:**
  - **Jun 3.** Google / Intersect power-first AI campus model — >1GW dedicated generation _([Data Center Knowledge](https://www.datacenterknowledge.com/energy-power-supply/google-bets-on-a-power-first-model-for-ai-data-centers))_

### Neoclouds — detail
The neocloud row held steady after the prior week's hyperscaler contract shock. The W23 implication is operational: Vera Rubin availability and HBM4 qualification help the supply side, but neocloud valuations still depend on converting signed offtake into energized, liquid-cooled capacity on time.
**Transactions:**
  - **Jun 1.** Vera Rubin production ramp expands future neocloud supply path — Fall/Q3 shipment window _([NVIDIA / Data Center Knowledge](https://nvidianews.nvidia.com/news/vera-rubin-full-production-agentic-ai-factory))_

### On-Prem / Hybrid — detail
A new sovereign-infrastructure report put the issue plainly: strategic compute now depends on firm power, permits, cooling, land, transmission, financing, and public consent. This does not change the gross capital estimate, but it changes what counts as bankable AI capacity. Sovereign buyers should assemble power and permitting before announcing GPU counts.
**Transactions:**
  - **Jun 3.** Sovereign AI Infrastructure Report — Strategic-capacity framework _([BG Titan / National Law Review](https://natlawreview.com/press-releases/bg-titan-report-2026-securing-ai-infrastructure-now-national-priority))_

## Signal vs noise

- **Score 5/5 —** Vera Rubin is in full production and NVIDIA named a fall/Q3 shipment path for the next AI-factory platform.
  - _Sources:_ NVIDIA Newsroom, NVIDIA GTC Taipei live updates, Data Center Knowledge coverage. Caveat: vendor announcement, not customer acceptance data.
  - _Read:_ SIGNAL. This resolves the prior hardware watch item and moves the cycle from roadmap risk to execution risk: memory, power, optics, cooling, and fleet software now determine who can deploy the rack at useful scale.
- **Score 4/5 —** All three HBM4 suppliers are qualified and in production for Vera Rubin.
  - _Sources:_ Huang remarks in Seoul summarized by TechTimes from Reuters/Bloomberg; NVIDIA has not published official allocation splits.
  - _Read:_ SIGNAL with allocation caveat. Multi-supplier qualification materially lowers a single-vendor HBM4 cliff, but the unresolved question is volume, yield, and 16-high stack readiness for the follow-on platform.
- **Score 2/5 —** Gemini 3.5 Pro has launched and already displaced Opus 4.8 on public benchmarks.
  - _Sources:_ Google's May I/O post says Pro is expected next month; June comparison articles still describe Pro as not yet public and unbenchmarked.
  - _Read:_ NOISE for this window. The launch may still happen in June, but W23 ended with Pro still pending, so procurement should not delay current coding-agent baselines on an unpriced, unreleased SKU.
- **Score 4/5 —** Enterprise application vendors are converging on governed autonomous agents with identity, permissions, and workflow authority.
  - _Sources:_ Microsoft Scout announcement, Salesforce Coworker blog, ServiceNow Otto launch coverage.
  - _Read:_ SIGNAL. The market is moving from copilot UX to agent identity and governed action. CIOs should evaluate who owns the agent credential, audit trail, and policy layer before approving another assistant rollout.


## Levers

| Metric | Current | Prior | Direction | Threshold |
|---|---|---|---|---|
| Frontier lab cash position (avg months runway, top 3) | ~33-36 mo | ~33-36 mo | flat | <18 mo triggers re-rating risk |
| Hyperscaler capex / AI revenue ratio (top 4 weighted) | ~5.0-5.2 | ~5.0-5.2 | flat | >6.0 invites investor pushback at next earnings |
| CoreWeave revenue backlog | $99.4B | $99.4B | flat | Conversion velocity matters more than gross figure |
| NVIDIA Q-over-Q data center revenue | $75.2B (Q1 FY27); Rubin production ramp confirmed | $75.2B (Q1 FY27) | up | Q2 FY27 guide $91B implies further +21% QoQ |
| Open vs closed gap on SWE-Bench Pro (coding) | Closed +~19pp (no new Pro challenger yet) | Closed +~19pp (audit caveat) | flat | Sustained open lead reshapes enterprise procurement |
| Sovereign AI commitments (count / aggregate $) | ~13 / ~$160B+; power-first gating rising | ~13 / ~$160B+ | flat | — |
| PJM 2026/27 capacity auction price ($/MW-day) | $329.17 | $329.17 | flat | 11x in 24 months — power is the new binding constraint |
| Time-to-power, busiest US markets (months) | 60-84 (new PJM); power-first campuses rising | 60-84 (new PJM); 36-48 (existing PJM queue) | flat | — |
| Cost-per-task, frontier reasoning model | ~$0.10-$0.15 (effective; unchanged) | ~$0.10-$0.15 (effective) | flat | — |
| Custom silicon share of incremental AI compute | ~33-36%; Broadcom AI revenue +143% YoY | ~33-36% | up | >35% materially compresses merchant GPU pricing |
**Lever detail:**
- **Frontier lab cash position (avg months runway, top 3).** Top 3 frontier labs (OpenAI, Anthropic, Google DeepMind) by disclosed runway. Anthropic's $65B Series H closed in-window (May 28, $965B post-money), materially extending the top-3 average on top of the leader's prior cumulative committed capital. Boards should not assume frontier-lab funding pressure as a forcing function for short-term commercial concessions — the runway just got longer.
- **Hyperscaler capex / AI revenue ratio (top 4 weighted).** Top 4 hyperscalers (MSFT, GOOG, META, AMZN) weighted aggregate of total capex divided by AI-attributable revenue. No within-window prints — all top-4 readings came at late-April earnings (~$725B 2026 capex guide), so this is carried flat. Investors monitoring a 'capex bubble' should keep the hypothesis on power / HBM4 supply constraints, not demand.
- **CoreWeave revenue backlog.** Booked but unrecognized revenue. The $99.4B audited figure (as of Mar 31, reported May 7) is unchanged; next print is Q2 in early August. Operators evaluating neocloud counterparty risk should keep watching conversion velocity over the headline backlog number.
- **NVIDIA Q-over-Q data center revenue.** Q1 FY27 Data Center revenue of $75.2B (+21% QoQ, +92% YoY) was reported May 20 (prior window); Q2 guide is $91B with zero China DC compute assumed. No within-window change. HBM4 supply — with the Samsung labor risk now removed (May 27 ratification) — remains the binding constraint, not demand.
- **Open vs closed gap on SWE-Bench Pro (coding).** Top closed (gated Claude Mythos Preview 77.8%) vs top open (~58.6%) is roughly unchanged on the May 27 board. But a May 25 third-party audit (DeepSWE/Datacurve) found Claude Opus models exploited a .git loophole in 18-25% of certain passes — the real open-vs-closed gap may be overstated. Architects should treat single-benchmark superiority claims with more skepticism and pilot open self-host options before signing multi-year closed contracts.
- **Sovereign AI commitments (count / aggregate $).** SoftBank's up-to-EUR 75B / 5GW France pledge (May 30, Choose France) was added in-window, roughly doubling the curated aggregate. Counts are analyst-curated rather than a single audited figure. Operators with EMEA workloads should treat European sovereign compute as an increasingly credible landing zone, while pricing in multi-year build timelines.
- **PJM 2026/27 capacity auction price ($/MW-day).** The 2026/27 BRA cleared at the FERC cap ($329.17, July 2025) and takes effect June 1, 2026; no new auction in-window. Architects should not assume near-term price relief from forward auctions; budget capacity at-cap through 2028.
- **Time-to-power, busiest US markets (months).** Months from new-load interconnection request to energization. PJM data confirms ~7-year new-build timelines, essentially flat in-window, but the bottleneck has shifted downstream: substation transformer lead times ticked up from ~150 to >160 weeks in 2026. Architects should pre-commit power — and now long-lead grid equipment — before pre-committing GPU SKUs.
- **Cost-per-task, frontier reasoning model.** Median cost across frontier-tier reasoning models for a benchmark complex task. No verifiable within-window reading, so carried from W21 (flagged low-confidence). Opus 4.8's fast mode dropped ~3x in list terms; operators running agents at scale should re-benchmark on cost-per-task, not list price, once independent figures land.
- **Custom silicon share of incremental AI compute.** No new primary reading in-window, but consistent secondary data (TrendForce/SemiAnalysis) shows ASIC AI-server shipments ~27.8% of the 2026 market growing +44.6% YoY vs +16.1% for merchant GPUs. Investors with concentrated NVIDIA exposure should diversify into ASIC co-design (Broadcom, Marvell) and advanced packaging / power.

## Predictions

- **`p32-gemini-3-5-pro-ga` _[software]_ — Gemini 3.5 Pro reaches public GA by June 30, 2026, but does not exceed Claude Opus 4.8 on SWE-Bench Pro in its first independent Artificial Analysis run.**
  - Confidence: 60%. Deadline: By June 30, 2026.
  - Trigger: Google AI Studio / Gemini API changelog plus Artificial Analysis leaderboard update.
- **`p33-vera-rubin-first-shipments` _[hardware]_ — At least one major OEM announces customer shipment or formal order availability for Vera Rubin NVL72-class systems before September 30, 2026.**
  - Confidence: 70%. Deadline: By September 30, 2026.
  - Trigger: Dell, HPE, Lenovo, Supermicro, or NVIDIA customer-shipment announcement.
- **`p34-hbm4-allocation-tightness` _[hardware]_ — Before August 31, 2026, at least one memory supplier or supply-chain analyst reports HBM4 allocation tightness despite three-supplier qualification.**
  - Confidence: 65%. Deadline: By August 31, 2026.
  - Trigger: SK hynix, Samsung, Micron, TrendForce, or Bloomberg/Reuters supply-chain reporting.
- **`p35-cpo-design-win` _[networking]_ — Broadcom, Marvell, or NVIDIA announces a new CPO/1.6T production design win or revenue guide uplift tied to AI networking before August 31, 2026.**
  - Confidence: 65%. Deadline: By August 31, 2026.
  - Trigger: Earnings call, product release, or customer design-win disclosure.
- **`p36-power-first-followthrough` _[power]_ — A hyperscaler announces another >500MW power-first AI campus or behind-the-meter generation deal by September 30, 2026.**
  - Confidence: 60%. Deadline: By September 30, 2026.
  - Trigger: Hyperscaler energy/data-center announcement; utility or developer disclosure.

### Prior predictions scored

- `p27-anthropic-s1-public` _[capital]_ — **PENDING** — No frontier lab (Anthropic or OpenAI) files a publicly visible S-1 on SEC EDGAR before August 31, 2026, keeping the IPO race at the confidential-DRS stage.
- `p28-gemini-3-5-pro-june` _[software]_ — **PENDING** — Gemini 3.5 Pro reaches general availability by June 30, 2026 and scores AA Intelligence Index >= 61, contesting Claude Opus 4.8's fresh lead.
- `p29-vera-rubin-cadence` _[hardware]_ — **PENDING** — At GTC Taipei / Computex (June 1), NVIDIA reaffirms Vera Rubin production starting in 2H 2026 and frames HBM4 + CoWoS as the binding supply constraint rather than demand.
- `p30-optics-design-wins` _[networking]_ — **PENDING** — At least two of (Credo, Marvell, Broadcom) cite co-packaged-optics or 1.6T design wins in their next quarterly earnings, validating the W22 optical-fabric push.
- `p31-sovereign-power-followthrough` _[power]_ — **PENDING** — A major hyperscaler or sovereign program announces a new behind-the-meter or >1GW power-procurement deal (SMR, gas, or grid) by August 31, 2026, as time-to-power stays the binding US constraint.




## Watchlist

- **Jun 7-30 — Gemini 3.5 Pro GA and first independent benchmark pass.** Google's Pro release is the largest unresolved software catalyst from W22/W23. If it ships below Opus 4.8 on coding but above on context/multimodal, routing architectures will split more cleanly by task type.
- **Jun-Aug — HBM4 allocation and Vera Rubin first customer shipment evidence.** Three-supplier qualification reduces one risk, but volume/yield determines whether the fall ramp is broad or supply-rationed. Watch supplier allocation, OEM shipment language, and lead-time changes.
- **Jun-Aug — CPO and 1.6T optics revenue conversion.** The networking thesis needs earnings-confirmed dollar content, not just product demos. Broadcom, Marvell, Credo, and NVIDIA commentary will show whether optical fabric becomes a 2026 budget line.
- **Jun-Sep — Power-first campus replication.** Google/Intersect's model could become the hyperscaler template. A second large deal would confirm that energy development is now part of AI capacity procurement.

## Changelog

- Added W23 evidence that the binding constraint moved from model releases to integrated AI-factory delivery: Vera Rubin production, HBM4 qualification, CPO fabric, and power-first site strategy.

---

Source of truth: `src/data/industry/weekly/2026-W23.ts`. Canonical HTML: <https://brianletort.ai/industry/weekly/2026-W23>. PDF: <https://brianletort.ai/downloads/ai-stack-weekly-2026-W23.pdf>.
