LayerRepresentative playersWhat is soldCore tension
ExperiencesOpenAI · Anthropic · Microsoft · Google · Apple · Cursor · vertical AI
Outcomes, attention, distribution
Trust, workflow fit, agent reliability
ModelsOpenAI · Anthropic · Google DeepMind · Meta · xAI · DeepSeek · Mistral
Capability per dollar
Frontier premium vs. open diffusion
Cloud & inferenceAWS · Azure · Google Cloud · Oracle · CoreWeave · Together · Fireworks · Baseten · Sail
GPU-hours and tokens
Latency, throughput, utilization
Networking & opticsNvidia / Mellanox · Broadcom · Marvell · Arista · Cisco · optical suppliers
Bandwidth, latency, ports, fiber
Scale-up performance vs. open scale-out fabrics
Systems softwareCUDA / NCCL · ROCm · PyTorch · Triton · JAX / XLA · serving engines
Portability and utilization
Ecosystem lock-in vs. abstraction
AcceleratorsNvidia · AMD · Google TPU · AWS Trainium · Microsoft Maia · Meta MTIA · Cerebras · Groq · Etched
FLOPs, bytes, watts, yield
Balanced generalist vs. spiky specialist
EDA & chip IPSynopsys · Cadence · Arm · verification and interface IP suppliers
Design tools, cores, interfaces
Faster design cycles vs. concentrated dependencies
Memory & packagingSK hynix · Samsung · Micron · TSMC · ASE · Amkor · substrate suppliers
HBM stacks and packaged systems
Capacity, thermals, yield, long lead times
ManufacturingTSMC · Samsung · Intel · ASML · Applied Materials · Lam Research · KLA
Wafers, HBM, packaged systems
Concentrated capacity and long lead times
Physical plantData-center developers · utilities · grid operators · Vertiv · Schneider Electric · Eaton
Megawatts, uptime, land, permits
Concentration vs. stranded capacity