Generative AI

The generative AI ecosystem,
explained.

Generative AI is the part of the wider AI ecosystem pulled along by models that create new text, code, images, audio and video. It runs on the same infrastructure as everything else, but the money moves through a distinct value chain. This page traces that chain from accelerators to agents.

Last fact checked 29 August 2026

Search the map

The generative AI value chain

Five stages carry a generated answer from raw silicon to a person. Each one maps onto a layer of the full stack, so you can read the detail wherever you want more depth.

  1. Stage 01 · ComputeFull layer →

    Accelerators and memory

    Generative models are matrix maths at enormous scale. Training a frontier model and then serving it to millions of people both come back to accelerators, high bandwidth memory and the packaging that binds them together.

  2. Stage 02 · CapacityFull layer →

    Data centres and cloud

    Someone has to own the buildings, the cooling and the power contracts. Generative AI shifted the shape of that demand from lots of small servers to dense clusters designed around a single training run.

  3. Stage 03 · FuelFull layer →

    Data and retrieval

    Pre-training corpora, licensed content, vector search and the pipelines that keep a model connected to current company data. Most generative products fail on retrieval quality long before they fail on model quality.

  4. Stage 04 · ModelsFull layer →

    Foundation and open-weight models

    Closed frontier models sold through an API, open-weight models you can host yourself, and the platforms that host, fine-tune and evaluate both. This is the layer people usually mean when they say generative AI.

  5. Stage 05 · ProductsFull layer →

    Applications and agents

    Assistants, copilots, image and video tools, coding agents and the workflow software quietly adding generation into features people already use. Value here depends on distribution, not just capability.

Generative AI is not the whole stack

Plenty of AI infrastructure has nothing to do with generation. Forecasting, fraud scoring, recommendation and industrial vision all sit on the same chips and data centres. Treating the two as identical is the most common mistake in the market, because it makes a specific product wave look like the entire industry.

Security, regulation and buyer demand cut across all of them: Cybersecurity, Governance, Regulation & Sovereign AI, Users, Enterprises & Distribution.

Model and platform companies

  • NVIDIA NVDA

    Designs the GPUs and accelerated computing platforms used for much of modern AI training and inference, and sells the networking that ties them together.

  • Microsoft MSFT

    Runs Azure, distributes AI through Microsoft 365 and Copilot, partners commercially with OpenAI, and designs its own Maia accelerators.

  • Amazon AMZN

    Runs AWS, designs Trainium and Inferentia accelerators, hosts third party models through Bedrock, and invests in Anthropic.

  • Alphabet (Google) GOOGL

    Builds Gemini models through Google DeepMind, designs TPU accelerators, operates Google Cloud, and distributes AI through Search, Workspace and Android.

  • OpenAI private

    Develops the GPT family of models and ships them through ChatGPT, an API and enterprise products.

  • Anthropic private

    Develops the Claude family of models, sold through an API, consumer apps and cloud marketplaces.

The models layer →

Application and agent companies

  • Microsoft MSFT

    Runs Azure, distributes AI through Microsoft 365 and Copilot, partners commercially with OpenAI, and designs its own Maia accelerators.

  • Amazon AMZN

    Runs AWS, designs Trainium and Inferentia accelerators, hosts third party models through Bedrock, and invests in Anthropic.

  • Alphabet (Google) GOOGL

    Builds Gemini models through Google DeepMind, designs TPU accelerators, operates Google Cloud, and distributes AI through Search, Workspace and Android.

  • OpenAI private

    Develops the GPT family of models and ships them through ChatGPT, an API and enterprise products.

  • Palantir PLTR

    Sells data integration and decision platforms to governments and large enterprises, increasingly packaged around AI driven workflows.

  • ServiceNow NOW

    Provides a workflow platform for IT, HR and customer operations, with AI agents embedded in those workflows.

The applications layer →

Generative AI ecosystem questions

What is the generative AI ecosystem?

It is the chain of businesses required to turn electricity and silicon into generated text, code, images, audio and video. Compute, capacity, data, models and applications, with security, regulation and buyer demand cutting across all of it.

How is the generative AI ecosystem different from the wider AI infrastructure stack?

The infrastructure stack describes everything that makes modern AI work, including forecasting, recommendation and vision workloads. The generative slice is narrower: it is the part of that stack pulled along by foundation models that produce new content rather than only classify or predict.

Where does the money actually go?

A large share of spend from generative AI applications flows back down to model providers, cloud capacity and ultimately accelerators and power. Application companies keep the customer relationship, but their cost of goods sits several layers below them.

What is the difference between training and inference?

Training creates the model and is a large, occasional capital-like cost. Inference is running the finished model for every user request, and it is a recurring cost that scales with usage. Products that grow quickly feel inference cost far more than training cost.

Are open-weight models part of the same ecosystem?

Yes. Open-weight models such as Llama shift where value accrues rather than removing the layers beneath. You still need accelerators, hosting, data pipelines and evaluation, you simply run more of it yourself.

All frequently asked questions →