---
title: "The $1.3 Trillion Inference War Is Heating Up. 3 Stocks to Watch."
type: "News"
locale: "en"
url: "https://longbridge.com/en/news/295630424.md"
description: "The AI inference market is projected to reach $1.3 trillion by 2032, driving competition among key players. Nvidia leverages LPUs for low-latency decoding, while Cerebras offers high-speed wafer-scale chips, recently partnering with AMD and securing deals with OpenAI and AWS. AMD pursues the market through chiplet designs, acquisitions of MEXT and Taalas to optimize memory and performance, and its collaboration with Cerebras. These companies are positioning themselves as winners in the rapidly expanding inference infrastructure sector."
datetime: "2026-08-12T07:56:46.000Z"
locales:
  - [zh-CN](https://longbridge.com/zh-CN/news/295630424.md)
  - [en](https://longbridge.com/en/news/295630424.md)
  - [zh-HK](https://longbridge.com/zh-HK/news/295630424.md)
---

# The $1.3 Trillion Inference War Is Heating Up. 3 Stocks to Watch.

Inference has become the fastest-growing part of the artificial intelligence (AI) infrastructure market, and Bloomberg Intelligence projects it will double the size of the AI training market by 2032, reaching $1.3 trillion. With so much at stake, both leading chipmakers and upstarts are jockeying to grab a slice of this huge, fast-growing market.

**Nvidia** (NVDA \-0.02%), **Cerebras** (CBRS +2.06%), and **Advanced Micro Devices** (AMD +1.01%) are all tackling this market in different ways. Let's see how these AI stocks stack up and why they could all be winners, given the size and growth of the inference market.

![Artist rendering of AI chip.](https://imageproxy.pbkrs.com/https://g.foolcdn.com/image//query-b3A9cmVzaXplJnVybD1odHRwczovL2cuZm9vbGNkbi5jb20vZWRpdG9yaWFsL2ltYWdlcy84ODI3MTEvZ2V0dHlpbWFnZXMtMTQ5MjMxOTg4NC5qcGcmdz0zODQw?x-oss-process=image/auto-orient,1/interlace,1/resize,w_1440,h_1440/quality,q_95/format,jpg)

Image source: Getty Images.

## 1\. Nvidia

Already the winner in AI model training, Nvidia now has its sights on the inference market. The company's big move to capture share was its "acquisition" of Groq and its language processing units (LPUs). Inference is more about fast memory access and low latency than raw compute power, and LPUs help address this by having SRAM (static random-access memory) embedded directly on their chips.

Expand

![Nvidia Stock Quote](https://imageproxy.pbkrs.com/https://g.foolcdn.com/image//query-b3A9cmVzaXplJnVybD1odHRwczovL2cuZm9vbGNkbi5jb20vYXJ0L2NvbXBhbnlsb2dvcy9tYXJrL05WREEucG5nJnc9MTI4?x-oss-process=image/auto-orient,1/interlace,1/resize,w_1440,h_1440/quality,q_95/format,jpg)

## NASDAQ: NVDA

Nvidia

Today's Change

(-0.02%) $-0.05

Current Price

$217.50

### Key Data Points

Market Cap

$5.3TMarket cap calculated using publicly traded shares outstanding only. Does not include unlisted, private, or dual-class non-traded shares. Implied market cap may vary.

Day's Range

$216.20 - $222.20

52wk Range

$164.07 - $236.54

Volume

101.3M

Avg Vol

149.8M

Gross Margin

74.15%

Dividend Yield

0.13%

LPUs are particularly useful during the decode phase of inference, which is when large language models (LLMs) answer queries. As such, Nvidia now offers complete systems designed specifically for inference, where its graphics processing units (GPUs) handle the pre-fill phase (reading the prompt) while its LPUs handle the decode phase, thereby speeding up response times.

This is a nice solution and positions Nvidia to remain an AI infrastructure leader, even if it doesn't capture the same market share it does in training.

## 2\. Cerebras

Like Nvidia, Cerebras is tackling inference with SRAM-based chips. However, because SRAM is so bulky, instead of just embedding a small amount onto its chips and stringing them together, Cerebras has created huge wafer-sized chips that are five to six times faster than LPUs.

The physical size of Cerebras' chips comes with some trade-offs. They require specialized cooling and energy management solutions and, as such, are only sold or rented as part of the Cerebras CS-3 systems. They also come at a very premium price tag.

Expand

![Cerebras Systems Stock Quote](https://imageproxy.pbkrs.com/https://g.foolcdn.com/image//query-b3A9cmVzaXplJnVybD1odHRwczovL2cuZm9vbGNkbi5jb20vYXJ0L2NvbXBhbnlsb2dvcy9tYXJrL0NCUlMucG5nJnc9MTI4?x-oss-process=image/auto-orient,1/interlace,1/resize,w_1440,h_1440/quality,q_95/format,jpg)

## NASDAQ: CBRS

Cerebras Systems

Today's Change

(2.06%) $4.75

Current Price

$234.76

### Key Data Points

Market Cap

$66BMarket cap calculated using publicly traded shares outstanding only. Does not include unlisted, private, or dual-class non-traded shares. Implied market cap may vary.

Day's Range

$220.50 - $236.75

52wk Range

$160.81 - $386.34

Volume

3.2M

Avg Vol

7M

Gross Margin

41.98%

However, the company has inked major deals with OpenAI and **Amazon**'s AWS, and it recently announced a partnership with AMD that should help reduce the cost of ownership. The two companies will offer an inference solution in which AMD's Helios rack-scale solution will handle the pre-fill phase of inference, which it can do more cheaply, while Cerebras' Wafer-Scale Engine will perform the decode phase, which it can do more quickly. It's a nice way for companies to better compete with Nvidia's offerings.

Given the high cost of its systems, Cerebras has been more of a premium, niche solution, but it now looks set to become a major player in the humongous inference market.

## 3\. AMD

After losing out on the LLM training market to Nvidia, AMD has been aggressively pursuing the inference market to make sure it doesn't get left behind again. Its chiplet design is better suited for inference, as it allows its GPUs to be packaged with more high-bandwidth memory (HBM) and to act as part of an entire unit to reduce latency. Meanwhile, its partnership with Cerebras looks like a smart move to help it better compete with Nvidia's complete inference system.

However, the company has not stopped there. It recently acquired memory optimization company MEXT and chip start-up Taalas to boost its inference offering. Memory is one of the biggest AI bottlenecks right now, and MEXT's solution can offload seldom-accessed data from DRAM to unused flash and then, using predictive AI, can transfer it back into DRAM before an application even requests it. This can reduce the need for more expensive DRAM and help save costs.

Expand

![Advanced Micro Devices Stock Quote](https://imageproxy.pbkrs.com/https://g.foolcdn.com/image//query-b3A9cmVzaXplJnVybD1odHRwczovL2cuZm9vbGNkbi5jb20vYXJ0L2NvbXBhbnlsb2dvcy9tYXJrL0FNRC5wbmcmdz0xMjg?x-oss-process=image/auto-orient,1/interlace,1/resize,w_1440,h_1440/quality,q_95/format,jpg)

## NASDAQ: AMD

Advanced Micro Devices

Today's Change

(1.01%) $4.76

Current Price

$474.32

### Key Data Points

Market Cap

$774BMarket cap calculated using publicly traded shares outstanding only. Does not include unlisted, private, or dual-class non-traded shares. Implied market cap may vary.

Day's Range

$463.21 - $475.99

52wk Range

$149.22 - $584.73

Volume

18.1M

Avg Vol

30.4M

Gross Margin

50.37%

Meanwhile, Taalas has developed chips in which AI models are hardwired directly to bolster inference performance. Since the chips are model-specific, they aren't as flexible, but they are cheaper and much faster. The company plans to use them as part of a complete system where its GPUs would handle the pre-fill phase and Taalas' chips would handle the decode phase.

AMD is tackling inference from a couple of different angles, which should position it to capture a nice share of this huge market.

### Related Stocks

- [NVDA.US](https://longbridge.com/en/quote/NVDA.US.md)
- [CBRS.US](https://longbridge.com/en/quote/CBRS.US.md)
- [AMDD.US](https://longbridge.com/en/quote/AMDD.US.md)
- [NVDY.US](https://longbridge.com/en/quote/NVDY.US.md)
- [NVDS.US](https://longbridge.com/en/quote/NVDS.US.md)
- [AMUU.US](https://longbridge.com/en/quote/AMUU.US.md)
- [NVD.US](https://longbridge.com/en/quote/NVD.US.md)
- [NVDX.US](https://longbridge.com/en/quote/NVDX.US.md)
- [NVDL.US](https://longbridge.com/en/quote/NVDL.US.md)
- [NVDU.US](https://longbridge.com/en/quote/NVDU.US.md)
- [NVDD.US](https://longbridge.com/en/quote/NVDD.US.md)
- [NVDQ.US](https://longbridge.com/en/quote/NVDQ.US.md)
- [07788.HK](https://longbridge.com/en/quote/07788.HK.md)
- [07388.HK](https://longbridge.com/en/quote/07388.HK.md)
- [AMD.US](https://longbridge.com/en/quote/AMD.US.md)
- [AMDL.US](https://longbridge.com/en/quote/AMDL.US.md)
- [NVDB.US](https://longbridge.com/en/quote/NVDB.US.md)
- [NVDG.US](https://longbridge.com/en/quote/NVDG.US.md)
- [NVDO.US](https://longbridge.com/en/quote/NVDO.US.md)
- [NVDW.US](https://longbridge.com/en/quote/NVDW.US.md)
- [NVYY.US](https://longbridge.com/en/quote/NVYY.US.md)
- [NYYY.US](https://longbridge.com/en/quote/NYYY.US.md)
- [DIPS.US](https://longbridge.com/en/quote/DIPS.US.md)
- [MAGX.US](https://longbridge.com/en/quote/MAGX.US.md)
- [AMDG.US](https://longbridge.com/en/quote/AMDG.US.md)
- [DAMD.US](https://longbridge.com/en/quote/DAMD.US.md)
- [AMDY.US](https://longbridge.com/en/quote/AMDY.US.md)
- [AMDW.US](https://longbridge.com/en/quote/AMDW.US.md)
- [CBRX.US](https://longbridge.com/en/quote/CBRX.US.md)
- [CBRG.US](https://longbridge.com/en/quote/CBRG.US.md)
- [CBRZ.US](https://longbridge.com/en/quote/CBRZ.US.md)
- [SMH.US](https://longbridge.com/en/quote/SMH.US.md)
- [SOXX.US](https://longbridge.com/en/quote/SOXX.US.md)
- [SOXL.US](https://longbridge.com/en/quote/SOXL.US.md)
- [SOXQ.US](https://longbridge.com/en/quote/SOXQ.US.md)
- [XSD.US](https://longbridge.com/en/quote/XSD.US.md)
- [PSI.US](https://longbridge.com/en/quote/PSI.US.md)
- [FTXL.US](https://longbridge.com/en/quote/FTXL.US.md)
- [09388.HK](https://longbridge.com/en/quote/09388.HK.md)
- [OpenAI.NA](https://longbridge.com/en/quote/OpenAI.NA.md)
- [AMZN.US](https://longbridge.com/en/quote/AMZN.US.md)
- [NVD.DE](https://longbridge.com/en/quote/NVD.DE.md)

## Related News & Research

- [Anthropic’s Massive AI Buildout Is Another Bullish Signal for AMD Stock](https://longbridge.com/en/news/295700576.md)
- [AI weekly: Nvidia does deal with Wall Street, China's AI weather forecasts](https://longbridge.com/en/news/295777975.md)
- [Jensen Huang's latest pitch to investors: Nvidia's AI compute is its own asset class](https://longbridge.com/en/news/295562673.md)
- [Nvidia’s Best Customers Have a Reason to Stop Buying So Many Nvidia Chips](https://longbridge.com/en/news/295572288.md)
- [Nvidia just recruited Wall Street to help fund $500 billion in AI infrastructure. Here’s the catch.](https://longbridge.com/en/news/295590827.md)