AI Infrastructure
Articles (50)

Gigawatt Supercomputers Span Buildings. Scale-Across Networking Has a Benchmark. Heterogeneity Doesn't.
NVIDIA has priced training across 1,000 km; Google trains synchronously across data centers. Missing: what one job pays to span mixed silicon and fabrics.

As the Public Web Tightens, AI Labs Are Buying Data From Bankruptcies. Spirit Airlines Is the Latest.
Google outbid a training-data startup for Spirit Airlines' internal records. One auction does not prove a data wall, but it shows how far labs will go.
OCP Maps the Co-Packaged Optics Stack. Photonic Die Testing Remains an Open Problem.
Open Compute Project's 294-page vision assigns an owner to each layer of the co-packaged optics stack, except one: photonic die testing, which the paper itself labels unowned.

NVIDIA's $500 Billion Compute-Finance Push Tests a Market Without a Residual-Value History
Rental indices, secondary venues and syndicated credit already exist for GPU compute. What lenders don't have is a residual-value history to underwrite against.

A New Bottleneck for AI Power: A 2031 Gas-Turbine Slot
As data-center developers add on-site generation to bridge utility delays, the grid queue is being joined by queues for turbines, air permits, gas infrastructure, storage, and EPC capacity.

The CUDA Moat Is Becoming a Composability Problem
AMD's inference optimizations work well one at a time. The gap opens when they run together, and closing it needs test hardware more than it needs missing APIs.

Korea Enforced AI Sovereignty Where It Could
Korea relaxed its chip mandate after two tenders selected no operator, then cut Naver over foreign model weights. Enforcement went where state power reached.

Sophia Space's Orbital Data Center Patent Trades Compute Density for Radiator Area
A worked example budgets 259 W of solar output per square meter of tile, implying a similar heat load. SpaceX's draft AI1 makes the opposite supercomputing bet: hotter liquid cooling and pumps.

Advanced Packaging Is Constraining AI Accelerators - While US Funding Targets the Next Bottlenecks
TSMC says packaging capacity is limiting customer growth. The $874 million Commerce proposes funds the bottleneck after this one.

Is Anduril's Menace-I a Tactical Supercomputer? The Public Specs Don't Say
Anduril says two operators can bring the AWS Outposts-equipped shelter online in under ten minutes, but technical details stay undisclosed.

AMD EPYC 9006 Makes the CPU Case for the Agentic Data Center
Venice adds cores, memory bandwidth and I/O for the work around inference, while AMD's rack-scale claims still await independent validation.

Why Pharma Is Building Its Own AI Supercomputers
Biology models now train on proprietary experiments and feed results back to the lab. That closed loop strengthens the case for owning the machine.

FERC's grid fast lane comes with a curtailment clause
FERC is pushing grid operators to create faster paths for large loads. The price may be flexibility when the system runs short.

Two Deployment Companies, One Week, and the Same Private Capital
OpenAI and Anthropic stood up deployment companies a week apart in May, both on outside capital, some of it from the same firms already financing the chips.

Four Listings, One Ledger: China's GPU Buildout in Audited Numbers
China's listed GPU pure-plays now have comparable audited numbers: an industry that is real, unprofitable, customer-concentrated, and scaling anyway.

AI's Power Moat Is Already Held: Contracted Megawatts Are the Buildout's Scarcest Asset
Shells rise in 18 to 24 months; grid power averages four years. The biggest AI builders have already contracted the firm megawatts everyone else is still queuing for.

Delay-Line Memory Returns: A Swedish Team Proposes Trading HBM for 14,000 Kilometers of Glass
Uppsala researchers propose streaming LLM weights through spooled fiber instead of HBM. The design is a sketch; the DRAM market that provoked it is very real.

Time to First Token Is a Real Metric. It Isn't the One That Defines the Era.
Time to first token is the AI industry's favorite candidate for era-defining metric. It is a real whole-stack signal, not a stand-in for business success.

NVIDIA's 45°C Cooling Spec Isn't New Physics. It's a Forcing Function for the Supercomputing Data Center.
NVIDIA specs every MGX rack (GB200, GB300, Vera Rubin) for 45°C inlet water. Warm-water cooling is a decade old; the default on 200 kW-class racks is what's new.

Three Bets Against Nvidia's Inference Margin, One Shared Dependency
OpenAI, Qualcomm, and Etched are betting against Nvidia's inference margin. Escaping it means queuing for the same TSMC packaging and memory. Most ASIC challengers die on software, not silicon.

One GPU in Orbit, a Million Satellites on Paper: Inside the Orbital Data Center Filing Arms Race
Since Starcloud flew the first H100 last November, SpaceX, Blue Origin, Starcloud, and a five-month-old startup have filed with the FCC for constellations totaling more than a million satellites. The physics hasn't moved as fast as the paperwork.

The President Says It's Fine. The Order Says It Isn't. The Models Are Still Dark.
An export-control order blacked out two Anthropic frontier models worldwide. The President has since softened; the order hasn't. A hosted model enforces a nationality rule only by going da

Reconstructing FP64: How Supercomputing's Establishment Is Adapting Science to AI Silicon
Two papers, Matsuoka's FP8-emulation preprints and the Dongarra 'Ride the Wave' paper, point to a field adapting scientific computing to AI silicon it no longer controls.

Project Prometheus Raised $12B to Train an "Artificial General Engineer." The Training Data Doesn't Exist Yet
Jeff Bezos and Vik Bajaj's startup now has $18.2 billion and roughly 150 employees. What it doesn't have is an internet of manufacturing data, so the corpus will have to be manufactured... much of it on supercomputers.

AWS quietly retired the fat tree. Fifty-year-old graph theory took its place.
By April 2026, Amazon's random-graph fabric had become the default for most new AWS datacenters. The efficiency claims behind it are still Amazon's own, with no independent benchmark yet.

Britain's Sovereign-Compute Day: A £2bn AMD Bet Meets a £1.1bn State Plan
Britain is funding homegrown silicon for machines that, for now, run on an American vendor's chips. Whether that buys sovereignty or quietly rebrands dependence is the question the spending leaves open.

The Electrician Bottleneck: Skilled Trades Increasingly Gate the AI Supercomputer Buildout
The GPU supply chain has the industry's attention. But the constraint that increasingly decides when an AI factory energizes is no longer the chip. It is power delivery, and the licensed electricians who commission it.

Co-Packaged Optics Has Two Front Doors. Only One Fixes the Scale-Up Bottleneck.
At Computex 2026, Wiwynn and eight ecosystem partners showed a full optical scale-up rack built around compute-side optics. It is a useful moment to separate two technologies the supercomputing industry keeps filing under one acronym.

Inside Meta's 83,000-GPU AI Supercomputer: Why It Runs the Silicon at 80% Power on Purpose
Meta's first end-to-end account of running a 150 MW, 83,000-GB200 cluster - when power is the ceiling, the cluster, not the chip, is what you optimize.

The Pentagon's FY2027 Budget Asks Congress for $46 Billion in Sovereign AI Infrastructure
Multi-year mandatory funding for a mix of government-owned, contractor-operated, and commercial-surge compute... reversing the July 2025 White House AI Action Plan that told DoW to lean on hyperscalers.

The 800V DC Rack Transition: How Rubin Ultra Is Rewiring the Supercomputing Industry's Last 50 Feet
NVIDIA has published the architecture. OCP and the supplier alliance have published the spec and the timeline. The colocation operators have published, so far, very little.

Thermodynamic Computing's First Silicon Is Back from the Fab. The Power Math Comes Next.
Normal Computing's CN101 is in characterization. Extropic has a prototype platform, an MIT-co-authored arXiv preprint, and an ETH Zurich hackathon in June. After two years as a manifesto, thermodynamic computing is producing the kind of artifacts readers can evaluate.

AI Training Power Demand Is Outrunning Grid Build Times. xAI Bet It Could Outrun Regulators Too.
xAI operates 46 unpermitted gas turbines at its Southaven power plant. A federal court ruling will determine if the turbine-first playbook is viable.

Nebius's $50 Billion Sells Out. Public Science Gets None of It.
Nebius's $50B backlog: 94% to Meta and Microsoft, zero to NAIRR, CloudBank, or DOE Genesis. The largest neocloud sells out before science can access it.

MRC Gives Open Ethernet Its First 75,000-GPU Production Proof Point
The 50-author MRC paper gives Ethernet its first multi-vendor, open-spec, production-trace answer to the one argument InfiniBand had left at frontier-training scale.

Apple's Mac Shortage Signals Memory Supply Chain Has Reorganized Around Data Center AI
Apple cut Mac memory ceilings and delayed M5 Ultra by four months as HBM production for data center AI consumes edge LPDDR5X allocation.

Orbital Compute in 2026: What Has Flown, What Is Slideware, and What the Physics Allows
Hardware has reached orbit and SpaceX has filed for a million-satellite constellation. Thermal physics, launch cadence, and bandwidth still push gigawatt orbital AI to post-2030, at best.

When the Grid Says No: Denmark and the New Shape of the Power Question
Energinet didn't pause grid connections in a contested metro or a zoning fight - it paused them because 60 GW of queued demand met a 7 GW national peak, and the math stopped working.

The training stack is starting to optimize itself
Anthropic’s 2.9× to 51.9× training-optimization curve signals that AI training infrastructure is becoming machine-optimizable, raising rebound demand and control-plane risks for HPC operators.

DeepSeek V4-Pro on Ascend 950PR: The Two-Stack AI Reality
DeepSeek V4-Pro runs on Huawei Ascend 950PR as the State Department pivots export controls from chip access to model IP, describing two parallel AI stacks.

HFAC Clears 16-Bill Chip Export Package on 150-Day Allied Clock
Sixteen bills cleared, silicon-level verification on deck, and an industry that hasn't spoken.

Sweden's Mimer Is a New Supercomputer Sized for Services, Not Scale
EuroHPC's EUR 29.76M Bull contract deploys a new AI-optimised supercomputer in Linköping - a tenth of IT4LIA's budget, with the services model as the differentiator.

VAST Data's $30B mark is a bet on the middle layer of AI, not storage
VAST Data closed a $1B Series F at a $30B post-money, 3.3x its 2023 mark and roughly 1.7x Everpure's public cap. Here's what the math and the customer list actually signal.

Vera Rubin's Memory Stack Is Korean. How Three Vendors Got There Tells You Why It Will Stay That Way.
Samsung, SK hynix, and Micron converged on SOCAMM2 mass production within six weeks for NVIDIA's Vera Rubin. Korean suppliers now control both memory tiers.

Slingshot Held Performance Under AI Traffic Patterns That Collapsed InfiniBand by 5x on Production Exascale
ISC 2026 research on LUMI, Leonardo, CRESCO8: Slingshot held performance; InfiniBand collapsed 5x under Incast, the AI gradient-sync traffic pattern.

HBM Allocation, Not HBM Supply, Is the 2026 AI Infrastructure Story
HBM scarcity has moved beyond semiconductor supply into system planning. Accelerator availability, server bill-of-materials, cluster economics, and 2026 data center buildouts are all being rewritten around memory - not compute.

Copper Kings Buy the Fiber Layer: Credo and Molex Lock Down Silicon Photonics in 48 Hours
Two acquisitions, two days apart, at adjacent layers of the same stack. Credo's $873M cash-and-stock deal for DustPhotonics and Molex's Teramount buy tell the market the copper-era interconnect champions have decided the AI factory's fiber layer isn't something they're willing to source.

DOE's SYNAPS-I Platform Targets Unified AI Analysis Across Seven Beamline Facilities
DOE's SYNAPS-I targets unified AI analysis across seven beamline facilities. Can it coordinate deployment or will it fragment like existing implementations?

Anthropic Locks 3.5 GW of Google TPU Capacity as Commercial AI Pre-Purchases Infrastructure Scientific Computing Will Need
Broadcom will supply Anthropic with 3.5 GW of Google TPU capacity through 2031; ~23-35x the power of DOE's largest planned science supercomputer.

UCCL-EP vs. NCCL EP: Portability or Consolidation for MoE Communication?
Two new expert-parallel efforts point to different futures for MoE systems: one built for heterogeneous fleets, the other folded into NVIDIA’s stack.