Undercurrent.
All articles14 min left
Undercurrent research14 min read
Undercurrent · Deep Dive · Technology

The AI
Stack Map

AI in 2026 is not a product category. It is a stack. The companies that matter will not merely have the best chatbot, the best model, or the best app. They will sit on the layers every other company has to cross.

AIComputeChipsEnergyRobotics2026
$400bn+
Big-tech AI capex 2025
set to rise +75% in 2026 (IEA)
17%
Data-centre power demand
2025 rise; double by 2030
40×/yr
Inference price decline
to GPT-4 level (Epoch AI)
70%
China refining share
of 19/20 strategic minerals (IEA)

Every computing era gets explained, after the fact, as if the winning layers were obvious from the beginning. They were not.

The map matters more than the demo.

01
The Map Matters More Than The Demo

A Six-Layer Stack, Two Forks

The internet did not simply produce websites. It produced a stack. Physical networks, routing, protocols, databases, servers, operating systems, browsers, search, payments and applications each had their own economics. The people who understood where one layer ended and another began could see why Intel, Cisco, Broadcom, Oracle, Microsoft, Google, Amazon and Apple were not all the same kind of company. They were claims on different choke points.

AI is now at the same stage. The public conversation is still trapped at the application layer: which chatbot is better, which assistant writes cleaner code, which image generator looks more realistic. That is the wrong level of abstraction. The more important question is not what the model can do today. It is which layer of the system compounds as the models get cheaper.

The first useful frame is that AI has become a six-layer stack. From the bottom up, the layers are infrastructure, chips, data, models, execution and applications. The stack then forks. One side is software AI, where inference prices are falling fast and intelligence begins to behave like abundant compute. The other side is physical AI, where the bottlenecks are not tokens but batteries, motors, sensors, materials, factories and energy.

AI is not one market. It is a vertical stack with horizontal forks. The mistake is to value every layer as if it had the same scarcity.

Some scarcity sits below the model, in power, cooling, wafers, lithography and memory. Some sits beside the model, in proprietary data and workflow context. Some sits above the model, in execution systems that turn reasoning into action. And some sits outside the software stack entirely, in the physical supply chains required to make machines move through the world.

02
The Six-Layer AI Stack

Not Born In A Cloud Console. Born In A Power Market.

The AI stack starts before a line of code is written. A model is born in a power market, a grid interconnection queue, a cooling design, a semiconductor fab, a lithography machine, a wafer, a package, a network and a data pipeline. By the time the user sees an answer on screen, the stack has already passed through multiple layers of scarcity.

LayerWhat it coversFulcrum asset
InfrastructurePower, cooling, land, grid, data centres, networks, critical mineralsEnergy access and physical capacity
ChipsGPUs, accelerators, memory, packaging, EDA, lithography, wafers, materialsAdvanced compute supply
DataPublic, proprietary, synthetic, telemetry and labelled feedbackUnique context
ModelsFoundation, open-weight, specialist, multimodal and inference enginesCapability per dollar
ExecutionAgents, tool use, orchestration, memory, permissions, evaluationReliable action
ApplicationsVertical software, copilots, autonomous workflows, physical-AI productsDistribution and workflow ownership

Six layers. Each has a fulcrum asset — the point through which disproportionate value must pass.

The stack is useful because it prevents a common error: confusing visibility with value capture. Applications are visible. Models are famous. Chips are talked about. But the highest-quality businesses may sit in places that users never see: lithography, high-purity quartz, grid capacity, inference routing, evaluation infrastructure, proprietary workflow data, robotics actuators and battery supply.

The OECD describes AI infrastructure as a complex, capital-intensive global supply chain that includes chips, data centres, cloud computing, power and cooling systems, broadband and network infrastructure. The International Energy Agency's 2026 update makes the same point from the energy side. It reports that data-centre electricity demand rose 17% in 2025, that consumption is set to double by 2030, and that power use from AI-focused data centres is poised to triple. The AI stack is not a metaphor. It is a physical system.

03
Foundation: Power, Not Prompts

AI Moves At Software Speed Until It Hits The Grid

The first layer of AI is infrastructure. That sounds boring until the constraint binds. Then it becomes the whole market.

The IEA says capex from five large technology companies rose to more than $400 billion in 2025 and is set to rise another 75% in 2026. That number is not just a symbol of corporate ambition. It is the bill for turning intelligence into an industrial utility.

Data-centre electricity demand rose 17% in 2025; consumption is set to double by 2030, and power use from AI-focused data centres is poised to triple.

— International Energy Agency, April 2026

Data centres require land, substations, grid interconnection, power-purchase agreements, backup capacity, cooling systems, transformers, water or liquid-cooling loops, fibre and permitting. The model layer can iterate every week. The power layer cannot.

ConstraintWhy it becomes a fulcrumWho benefits
Power availabilityTraining and inference loads need large, reliable electricity accessData-centre operators, utilities, hyperscalers with secured power
CoolingHigh-density AI racks push beyond conventional air coolingLiquid-cooling vendors, thermal-management specialists
Grid connectionBottleneck is interconnection and transmission, not just generationOwners of powered land, substations, fast permitting paths
Critical mineralsBatteries, transformers, motors and robotics depend on concentrated supplyMiners, refiners, countries controlling processing
Financing capacityAI build-outs require multi-year commitments before revenue is certainHyperscalers, sovereign-backed projects, low-cost-capital firms

Every improvement in model capability increases inference demand. Every inference call consumes electricity.

This is the first inversion. The world thinks AI abundance will make the stack lighter. In reality, intelligence abundance may make the bottom of the stack more valuable.

04
Chips: Where The Stack Becomes Geopolitical

A System Of Chokepoints, Not A Commodity

The second layer is chips, but "chips" is too small a word. The AI chip layer includes accelerators, high-bandwidth memory, advanced packaging, EDA software, photolithography, deposition, etching, inspection, wafer fabrication, substrates, specialty chemicals and high-purity materials. It is the most globally interdependent layer in the stack.

CSET's semiconductor supply-chain work frames the point clearly: the semiconductor supply chain is not a single line but a set of highly specialised production stages, with national strengths and vulnerabilities distributed across design, fabrication, manufacturing equipment, materials, assembly, testing and packaging.

€32.7 billion in net sales, €4.7 billion of R&D, 535 system sales, 5,100 suppliers — and the first full-specification TWINSCAN EXE:5200B High-NA EUV system delivered. The rest of the stack can only be more ambitious if this layer can advance.

— ASML, Annual Report 2025

SublayerChokepointStrategic question
EDA and design IPSoftware and libraries to design advanced chipsWho controls the design tools used before silicon exists?
LithographyEUV and High-NA EUV systemsWho controls the machines that define the minimum printable feature?
MaterialsPhotoresists, gases, wafers, high-purity quartzWhich suppliers can meet purity at scale?
FabricationLeading-edge foundry capacityWho can turn designs into advanced nodes reliably?
Memory and packagingHBM, advanced substrates, chiplet integrationWho removes the bandwidth wall around AI accelerators?
SystemsServers, racks, networking, data-centre integrationWho converts chips into usable compute clusters?

The chip layer is a system of chokepoints. National policy and corporate strategy meet at every one.

The strange part of AI nationalism is that the most visible American AI assets sit on non-American and globally concentrated physical foundations. NVIDIA's CUDA ecosystem may be one of the most important software layers in AI, but it runs on chips whose creation depends on lithography from the Netherlands, manufacturing and packaging capacity across Asia, Japanese specialty materials, and high-purity quartz from Spruce Pine, North Carolina.

Beyond Google

Sibelco states that its IOTA high-purity quartz is mined from two uniquely pure ore bodies at Spruce Pine, used to produce fused-quartz crucibles for the Czochralski process and quartzware for semiconductor wafer production. The precise global dependency is more nuanced than the popular phrase "every wafer comes from one mine," but the direction is right: the chip stack contains obscure physical inputs whose scarcity is wildly underpriced by the application-layer conversation.

AI is sold as a software revolution. It is also a mineral, chemical and lithography regime.

The most important AI company in a given layer may be a company no consumer has ever heard of. That is what a stack looks like before the map is drawn.

05
Models: Improving Into Commoditization

More Important To The World, Less Defensible Alone

The model layer is the loudest layer because it produces the magic. It is also the layer where rent may compress fastest.

Epoch AI's analysis of inference-price trends shows why. It found that the lowest API price required to reach GPT-4-level performance on a PhD-level science benchmark fell around 40× per year, while decline rates across benchmark thresholds ranged from 9× to 900× per year. The exact rate depends on task, threshold and benchmark. The strategic implication is clear: capability-adjusted intelligence is getting cheaper very quickly.

Estimated annual decline in API price to reach a benchmark threshold (Epoch AI, 2025)
GPT-4 / PhD science
40× / yr
Headline figure cited in Epoch's analysis
Slowest task
9× / yr
Lower bound across measured tasks
Fastest task
900× / yr
Upper bound across measured tasks

That creates a paradox. The model layer is becoming more important to the world but less obviously defensible as a standalone profit pool. If every frontier model becomes better, and if open-weight or cheaper closed models approach the same practical performance on many tasks, then the value above the model shifts to distribution, context and execution. The model becomes the engine. The business is the machine that uses it.

TrendWhat it doesWhere value migrates
Rapid capability gainsMore tasks become automatableApplications expand the surface area of use
Inference-price declinesMore intelligence per workflowExecution systems and high-frequency apps gain leverage
Open-weight competitionBaseline capabilities widely availableProprietary data, deployment and trust matter more
MultimodalityText, image, audio, video and sensor data convergePhysical AI and enterprise workflows become addressable
Model routingWorkloads sent to different models by price/latency/qualityOrchestration and evaluation layers become strategic

The deep question is not whether models become smarter. It is who captures the surplus when they do.

06
Data Is Not Oil. Data Is Memory.

Contextual, Permissioned, Tied To Action

The third layer is data, and the old cliché that "data is the new oil" is now actively misleading. Oil is consumed. Data is remembered, reweighted, recombined and embedded into workflows. In AI, the scarce data is not merely large. It is contextual, permissioned, current and tied to action.

Public internet data trained the first wave. That layer is increasingly exhausted, litigated or commoditised. The next wave is enterprise data, user memory, real-time telemetry, codebases, customer histories, transaction graphs, medical records, industrial sensor data, robotics demonstrations and expert feedback.

Data typeExampleDefensibility
Public knowledgeWeb text, open-source code, books, images, public datasetsHigh utility, increasingly commoditised and contested
Private workflow dataCRM, claims files, ERP, code repos, support tickets, legal matter historyDefensible if permissioned, clean and embedded in daily work
Action feedbackHuman corrections, tool outcomes, robotic demos, user decisionsHighly valuable — it teaches the system what worked

A weak app with no workflow ownership is a skin on a model. A strong app is a data acquisition machine.

The winner is not the company with the largest static dataset. It is the company that owns the loop: observe, suggest, act, measure, learn, repeat.

07
Execution: The Missing Layer

From Generating Text To Completing Work

Most AI maps jump from models to applications. That misses the most important emerging layer: execution. It is what sits between a model that can reason and an application that can do work — agents, tool use, workflow orchestration, memory, permissions, identity, sandboxing, evaluation, observability, compliance, exception handling and handoff to humans.

The Federal Reserve's 2026 note on AI adoption shows that diffusion has already moved beyond novelty. By year-end 2025, about 18% of US firms had adopted AI according to Census business survey data; work-related generative-AI adoption among individuals was about 41% in November 2025; and a Survey of Business Uncertainty estimate found that roughly 78% of the labour force worked at firms that had adopted AI. But adoption is not the same as transformation. A company can have employees using AI and still have no execution layer.

ComponentWhat it solvesWhy it is hard
Tool useLets models query systems, write files, call APIs, update recordsTools create real-world side effects and security risk
MemoryLets systems preserve context across sessions and workflowsMemory must be relevant, permissioned and auditable
PlanningBreaks complex work into steps and dependenciesLong-horizon tasks fail silently without checkpoints
EvaluationMeasures whether outputs and actions are correctMany business tasks lack clean benchmark answers
PermissionsControls who or what can act on behalf of a userEnterprise trust requires identity, logging, revocation
Human handoffEscalates ambiguous or high-risk decisionsAutomation must know when not to automate

Pure generation is easy to copy. Reliable action is not.

This is the "machine that makes the machines" layer. The first wave of AI products helped humans create outputs. The next wave builds systems that create, test, deploy and operate other systems. A coding agent is not merely a better IDE — it is a factory for software changes. A claims agent is not merely a summariser — it is a factory for triage, evidence collection and settlement recommendations.

08
Applications: Where Distribution Becomes Data

Operating Systems For Narrow Domains

Applications are where the stack becomes legible to customers. They are also where people overpay for demos and underpay for workflow ownership. The right question is not "does this app use AI?" Every app will. The right question is whether the application owns a repeated workflow with enough frequency, pain, budget and data exhaust to improve faster than substitutes. AI does not make a weak workflow strong. It makes a strong workflow compound.

ArchetypeWeak versionStrong version
CopilotWrites drafts inside an existing workflowCaptures decisions, feedback and edits until it becomes workflow memory
Vertical agentAutomates a narrow task with fragile promptsOwns permissions, integrations and exception handling end-to-end
Consumer assistantAnswers questionsOwns identity, preferences, payments, scheduling, repeated intent
Developer toolAutocompletes codePlans, edits, tests, reviews, deploys, learns from production outcomes
Physical-AI appDemonstrates a robot taskOwns hardware, fleet operations, maintenance, supply chain and real-world data

Consumer AI is distribution-led. Enterprise AI is trust-led. Both are won by owning the loop.

09
The Fork: Software AI vs Physical AI

Tokens And Atoms Don't Scale The Same Way

The most important structural split appears at the chip layer. Above chips, AI forks into two worlds.

The first is software AI. Agents, copilots, model routing, code generation, enterprise automation, search, media, customer support, analytics and personal assistants. It compounds on declining inference prices.

The second is physical AI. Robots, autonomous vehicles, drones, industrial automation, humanoids, warehouse systems, embodied assistants. It compounds more slowly because atoms do not scale like tokens. Physical AI has to solve batteries, motors, actuators, sensors, safety, maintenance, manufacturing yield, repair networks, regulatory exposure and real-world edge cases.

DimensionSoftware AIPhysical AI
Primary bottleneckReliable execution inside digital workflowsEnergy storage, actuation, sensors, manufacturing, safety
Marginal cost curveFalls with inference efficiency and routingFalls with manufacturing scale and field learning
Data loopPrompts, documents, code, enterprise systems, tool outcomesSensor data, demonstrations, fleet telemetry, physical failures
Deployment speedFast — software ships continuouslySlow — hardware must be built, tested, certified, maintained
Dominant scarcityDistribution, trust, context, permissionsMinerals, batteries, motors, factories, reliability, operating data

A chatbot and a robot may share a model architecture. They do not share the same bottleneck.

The IEA's Global Critical Minerals Outlook 2025 reports that the average market share of the top three refining nations for key energy minerals rose from around 82% in 2020 to 86% in 2024, with around 90% of refined material supply growth coming from the top single supplier across several key minerals. China is the dominant refiner for 19 of 20 strategic minerals analysed, with an average market share around 70%.

That concentration matters for robots because physical AI is not just a model problem. A robot needs energy storage, motors, magnets, sensors, manufacturing tolerance, maintenance and a supply chain that can turn demonstrations into fleets. A software AI company can scale by lowering inference cost and adding customers. A physical AI company must master the interface between intelligence and matter.

10
The Fulcrum Assets

Where Disproportionate Value Has To Pass

A fulcrum asset is a point in the stack where control over a scarce input changes the economics of every layer above it. It is not always the largest market by revenue. It is the point through which disproportionate value must pass.

LayerFulcrum assetWhy it may compoundCommoditisation risk
InfrastructureSecured power, cooling, grid interconnectionAI demand outruns physical capacity buildMedium — overshoot or shared regulation
ChipsEUV lithography, HBM, advanced packaging, high-purity materialsAdvanced compute depends on specialised tools and inputsLow to medium — expertise and capex create barriers
DataProprietary workflow data and action feedbackData improves the product as usage risesMedium — if data is portable or untied to outcomes
ModelsFrontier capability and low-cost inferenceBetter models expand use cases and enable new workflowsHigh — convergence and routing erode standalone rent
ExecutionAgent orchestration, evaluation, identity, permissionsReliable action becomes the interface between AI and workMedium — platforms may absorb this layer
ApplicationsDistribution, workflow ownership, repeated user intentUsage creates data and habit which improves the productHigh for shallow wrappers; lower for systems of record

Six layers, six fulcrums, six different risks of being competed away.

The most dangerous sentence in AI investing is "the model will do that." The model may do the reasoning. It does not automatically own the energy contract, the chip supply, the proprietary dataset, the customer relationship, the compliance boundary, the workflow permissions or the robot's actuator. Each of those can be a business.

The second most dangerous sentence is "the app is just a wrapper." A spreadsheet was "just a wrapper" on computation until it became the operating surface for finance. A CRM was "just a database" until it became the system of record for revenue. The test is whether the application owns a loop that improves with use.

11
The Structural Position

Value Migrating Toward Bottleneck Control

AI in 2026 should be understood as a stack whose value is migrating away from novelty and toward bottleneck control. The bottom of the stack is becoming more important because intelligence consumes physical resources. Power, cooling, grid access, chips, lithography, high-purity materials and critical minerals are not background inputs. They are the reason some companies can scale and others cannot.

The chip layer turns the AI boom into a geopolitical supply-chain problem. The model layer will keep improving, but capability-adjusted inference prices are falling fast enough to pressure standalone model margins. The data layer becomes more valuable when it is private, contextual and tied to outcomes. The execution layer becomes the bridge between reasoning and work. The application layer wins when distribution becomes data and data becomes execution.

The fork matters because software AI and physical AI do not compound the same way. Software AI rides the collapse in inference cost. Physical AI rides manufacturing scale and supply-chain control. One waits on permissions. The other waits on batteries, motors, factories and minerals.

The honest version is this. AI is not one race. It is a layered contest over scarce control points. The market will remember the famous model names, but the durable winners may be the companies that own the less glamorous boundaries: powered land, lithography, high-purity quartz, memory bandwidth, proprietary workflow data, agent execution, identity, evaluation and physical supply chains.

You are not looking for the company that says "AI" the loudest. You are looking for the company that sits where the rest of the stack has no choice but to pass.

Source References
  1. OECD, "Competition in Artificial Intelligence Infrastructure", 14 November 2025.
  2. International Energy Agency, "Data centre electricity use surged in 2025, even with tightening bottlenecks driving a scramble for solutions", 16 April 2026.
  3. Saif M. Khan, Dahlia Peterson and Alexander Mann, Center for Security and Emerging Technology, "The Semiconductor Supply Chain: Assessing National Competitiveness", January 2021.
  4. ASML, "Annual Report 2025", 2026.
  5. Sibelco, "High Purity Quartz", accessed May 2026.
  6. Epoch AI, "LLM inference prices have fallen rapidly but unequally across tasks", 12 March 2025.
  7. Jeffrey S. Allen, Board of Governors of the Federal Reserve System, "Monitoring AI Adoption in the US Economy", FEDS Notes, 3 April 2026.
  8. International Energy Agency, "Global Critical Minerals Outlook 2025: Executive Summary", 2025.
Undercurrent
The hidden systems behind the world you live in
Deep Dive · Technology · AI & Compute · May 2026

You’ve looked beneath the surface.

Now follow the connection.

The Energy Illusion

Electricity is a physical constraint connecting national energy systems and computing infrastructure.

The Garden Hose

Move from the physical cables carrying data to the layers of infrastructure behind AI.

Escape Velocity

A connection through “The hidden bottleneck”: Find the physical and institutional constraints beneath apparently limitless systems.

Explore this article’s connections ↗
A suggested reading trail

Who owns the future?

From the foundations of AI to the ownership of information.

Download sharing card ↗Follow the next investigation via RSS ↗