Goldman Sachs: AI Capex Pivots Toward Inference and Enterprise Use

Illustration of AI infrastructure capex shifting from training clusters to distributed inference and enterprise adoption

Goldman Sachs published a note dated July 10, 2026 arguing that AI investment is rotating from headline-grabbing training clusters toward inference workloads and broader enterprise adoption. The bank frames the shift as a maturing phase of the AI capital cycle rather than a slowdown.

Executive Summary

The Goldman Sachs view, as summarized in the release, is that the marginal AI dollar is increasingly directed at inference — the runtime serving of trained models to end users and applications — and at enterprise deployments that put those models to work inside businesses. Training remains significant, but the growth vector is moving.

For infrastructure operators, that framing matters because inference and enterprise AI have a different physical and economic profile than training. They favor latency-sensitive placement, steadier utilization curves, and integration with existing corporate data — all of which reshape where capacity is built, how it is cooled and powered, and which vendors capture the spend.

What ‘Shift to Inference’ Actually Means for Infrastructure

Training a large model is a bursty, capital-intensive event: tens of thousands of accelerators wired together, run flat-out for weeks, tolerant of remote siting as long as power and interconnect are cheap. Inference — the act of answering a user’s query with a trained model — is the opposite. It runs continuously, scales with usage, and rewards proximity to users and to enterprise data. If Goldman’s read is right, the next tranche of AI capex will look less like one giant campus in a remote grid pocket and more like distributed capacity closer to demand.

That has second-order consequences the note itself does not spell out. Metro data centers, edge sites, and existing enterprise colocation footprints become more strategically valuable. Networking — low-latency fiber between inference points, users, and data gravity centers — becomes a first-class concern rather than a training-cluster afterthought.

Enterprise Adoption Changes the Buyer

A capex signal tied to enterprise adoption implies a different customer mix than the hyperscaler-and-frontier-lab spending that has dominated headlines. Enterprises buy differently: they care about data residency, regulatory posture, integration with existing systems, and predictable unit economics. They are also more sensitive to total cost of ownership than to raw peak FLOPS.

If that customer base grows as the note suggests, the winners are likely to include vendors and operators that can package AI capacity as a consumable service — with governance, observability, and support — rather than raw GPU hours. It also expands the addressable market for private cloud, sovereign cloud, and hybrid deployments where the model runs near the data.

Reading the Capex Signal With Appropriate Caution

Analyst notes are directional, not deterministic. Goldman is describing a rotation in how AI dollars are spent, not a retreat from AI spending overall, and the release as summarized does not quantify the magnitude, timing, or geographic distribution of that rotation. It is fair to ask what data underpins the call — enterprise deal flow, hyperscaler capex disclosures, chip shipment mix — and how much of the shift is already priced into infrastructure equities.

The same scrutiny applies to the counter-narrative. Claims that training demand is peaking have been made before and repeatedly revised as new model generations arrived. A durable inference-led phase would still coexist with periodic training surges tied to frontier releases. Buyers planning multi-year builds should treat the shift as a change in mix, not a substitution.

Background

AI infrastructure spending accelerated sharply from 2023 onward, dominated by large training clusters built by hyperscalers and frontier model developers. That phase concentrated capital in a small number of very large sites optimized for dense accelerator deployments, cheap power, and high-bandwidth interconnect.

As foundation models have matured and enterprise pilots have moved toward production, industry attention has increasingly turned to inference — the runtime side of AI — and to the operational, data, and governance challenges of deploying models inside businesses. Goldman’s July 2026 note sits within that broader transition, articulating a capex signal that many operators and vendors have been positioning for.

Source: AI Investment Is Shifting as Inference, Enterprise Adoption Accelerate – Goldman Sachs — Goldman Sachs note dated July 10, 2026 describing a rotation in AI capital spending toward inference workloads and enterprise adoption.