Intermediate 12 minAnnouncements and updates

Ryzen AI Max PRO 400: what 192 GB enables ?

Direct response

Yes: AMD’s Ryzen AI Max PRO 400 lineup (Zen 5), introduced in 2026, includes up to 192 GB of unified memory, with 160 GB assignable to the GPU—a portion of the 192 GB, not an addition. AMD says that these 160 GB are enough to load a model with more than 300 billion parameters in 4-bit. Speed remains the issue: at 273 GB/s of bandwidth, such a model would run at only a few tokens per second.

AMD has unveiled the Ryzen AI Max PRO 400 lineup for business PCs and small workstations, with one headline feature: up to 192 GB of unified memory. That is enough to excite anyone who wants to load a very large language model without a dedicated $10,000 graphics card. Here is what these figures really mean, what still needs to be verified, and why memory capacity does not tell you everything about speed.

Choosing a machine? Our picks by budget → · Our spec sheet Ryzen AI Max mini-PC with 128 GB →

By Mohamed Meguedmi·Update 2026-09-28·Tested on Windows, macOS, and Linux
Recommended hardware

Buying alternative for this guide: BOSGAME M5 128GB / 2TB (Ryzen AI Max+ 395).

Why this choice? Our complete guide on BOSGAME M5 128GB / 2TB (Ryzen AI Max+ 395) →

Compare all options by budget, from €800 to €3,500 →

On the go: which laptop for local AI →

Affiliate links — commission possible at no extra cost to you. As an Amazon Associate, BestLLMfor earns from qualifying purchases.

#What AMD announces with the PRO 400

In 2026, AMD introduced the Ryzen AI Max PRO 400 series, built on the Zen 5 architecture with RDNA 3.5 graphics and an XDNA 2 NPU. On its official blog, AMD writes that these processors combine “192 GB of system memory and 160 GB of VRAM” for mobile workstations or small desktops aimed at AI, content creation, and simulation. This is a manufacturer announcement: the figures come from AMD, not an independent measurement.

The target audience is professional: AI developers, engineers, and creators working with large datasets. AMD cites HP and Lenovo as OEM partners, and the technical press (ServeTheHome) adds ASUS to the list of confirmed partners.

#192 GB of memory, 160 GB for the GPU: the nuance you shouldn’t miss

This is the point that causes the most confusion: the 160 GB announced for the GPU is not added to the 192 GB of total memory. It is 160 GB TAKEN from the unified 192 GB pool, with the remainder (32 GB) available for the system and CPU. AMD specifies this in its technical note: the Ryzen AI Max+ 495 PRO “supports up to 160 GB of dedicated graphics memory.” So there is no configuration with 192 GB for the CPU AND an additional 160 GB for the GPU.

!
Mistake to avoid
192 GB + 160 GB does not equal 352 GB. The GPU draws from the same unified memory as the CPU; 160 GB is the maximum amount you can allocate to it, not a separate block.

#What 192 GB really enables (Q4 model sizes)

Using the usual memory estimates for Q4 quantization (weights only, excluding context: 70B ≈ 40 GB, 32B ≈ 19–20 GB), 160 GB of GPU memory leaves room for models well beyond 70B. AMD states, in a technical footnote, that the Ryzen AI Max+ 495 PRO “supports up to 160 GB of dedicated graphics memory, capable of running a model with more than 300 billion parameters in 4-bit quantization.”

Indicative memory headroom in Q4 (weights only) on a 160 GB GPU
Model sizeEstimated Q4 sizesRemaining context headroom
70B≈ 40 GB≈ 120 GB
120B (dense)≈ 68 GB≈ 92 GB
235B (MoE, Qwen3-235B type)≈ 130 GB≈ 30 GB
300B (Q4, AMD figure)≈ 160 GBAlmost none

The calculation shows one simple thing: loading a 300B model in Q4 nearly saturates all GPU memory, leaving very little room for the context (KV cache) and system overhead. A long context on a model this size therefore isn’t guaranteed even if the model technically fits in memory.

#Why memory bandwidth will hold back what 192 GB promises

This is the point marketing announcements don’t highlight: memory capacity is not generation speed. In the vast majority of cases, an LLM’s generation throughput is limited by memory bandwidth—the model must be reread from memory for every token. The formula to remember, presented as a theoretical estimate rather than a measurement: tokens/s ≈ bandwidth (GB/s) ÷ active weight size (GB).

The PRO 400 advertises 273 GB/s of bandwidth with LPDDR5X-8533, according to ServeTheHome. Applying the formula to a 300B model in Q4 (≈ 160 GB of weights): 273 ÷ 160 ≈ 1.7 tokens/s. This is a theoretical upper-bound estimate, not an actual measurement published by AMD or an independent tester—but it gives an order of magnitude incompatible with fluid conversation. On a dense 70B model (≈ 40 GB), the estimate rises to about 6.8 tokens/s; on a 32B model (≈ 20 GB), to about 13.6 tokens/s.

i
The gradient we often forget
A manufacturer's “up to 192 GB” claim refers to capacity, never interactive throughput. On a machine with modest bandwidth, MoE models (few active parameters per token) remain the best way to take advantage of large memory capacity without a collapse in speed.

#The three chips in the PRO 400 lineup

Ryzen AI Max PRO 400 series (source: AMD)
ModelCores / threadsBoost clockGraphics processing unitsUnified memory
Ryzen AI Max+ PRO 49516C / 32Tup to 5.2 GHz40 CU (Radeon 8065S)192 GB
Ryzen AI Max PRO 49012C / 24Tup to 5.0 GHz32 CU (Radeon 8050S)192 GB
Ryzen AI Max PRO 4858C / 16Tup to 5.0 GHz32 CU (Radeon 8050S)192 GB

The three chips share the same maximum memory capacity of 192 GB; only the CPU cores, frequency, and number of graphics units differ. For local LLM use, the 495 PRO (40 CU) has the most compute power for processing the prompt and managing the context, even though memory bandwidth — and therefore generation speed — remains identical across all three models.

#“Ryzen AI Halo” Developer Kit and PRO 400 OEM PC: two distinct products

AMD announced two different offerings on the same day, which several press outlets have conflated. The “Ryzen AI Halo” is, according to AMD, “the first compact AI development platform designed by AMD,” intended for local prototyping without depending on the cloud. Its first version is based on a Ryzen AI Max+ 395—the previous-generation chip, not the PRO 400—with up to 128 GB of unified memory, sold exclusively through Micro Center with preorders opening in June 2026. It is not until the third quarter of 2026 that AMD plans to upgrade this same developer kit to a PRO 400 Series chip, with up to 192 GB of memory and 160 GB of VRAM.

The Ryzen AI Max PRO 400 discussed here is a separate product: a component that manufacturers (HP, Lenovo, ASUS) integrate into their own business PCs, at their own prices, with no direct connection to Micro Center’s sales channel for the developer kit. To avoid confusion when purchasing: the “Halo” kit is an AMD machine you order directly; the PRO 400 is a chip found in third-party branded PCs, whose availability and pricing depend on each manufacturer.

Ryzen AI Halo developer kit vs. Ryzen AI Max PRO 400 OEM PC
Ryzen AI Halo (developer kit)PRO 400 (OEM professional PCs)
Embedded chipRyzen AI Max+ 395 (Q3 2026: PRO 400)Ryzen AI Max PRO 400 (485/490/495)
Maximum memory128 GB (Q3 2026: up to 192 GB)192 GB
VendorAMD, Micro Center exclusiveHP, Lenovo, ASUS (third-party manufacturers)
Announced timelinePreorders starting in June 2026Machines starting in Q3 2026

#Availability: what is confirmed and what isn't

AMD announces availability through OEM partners “including HP and Lenovo” in the third quarter of 2026; ServeTheHome adds ASUS to the list of confirmed partners. As of this writing, no public price or specific consumer-machine model has been verified for France: these are professional PCs and mobile workstations, not (at this stage) mini PCs affordable on an individual’s budget like the previous generation.

The first concrete example documented by the tech press: in early September 2026, Lenovo unveiled its ThinkCentre X Ultra, a mini workstation featuring “up to” a Ryzen AI Max+ PRO 495 with an integrated Radeon 8065S GPU—but in the announced configuration, it offers 128 GB of RAM, not the chip's maximum 192 GB. Tom's Hardware reports a planned November 2026 release, with an “expected” starting price of $3,669, while specifying that this price “surely does not” apply to the high-end configuration. The takeaway: AMD's announced 192 GB is a chip limit, not a guarantee of a configuration at an entry-level price—you need to check each machine's exact specifications before expecting the maximum memory at the lowest price.

!
Check before buying
Do not rely on a release date or price found on a forum or unofficial product sheet: verify actual availability in an FR/EU store and the amount of memory actually installed in the target configuration, as the PRO 400 is positioned for professional rather than consumer use.

#Against the Ryzen AI Max+ 395: what really changes

The previous generation, the Ryzen AI Max+ 395 (Strix Halo), tops out at 128 GB of unified LPDDR5X-8000 memory, with theoretical bandwidth of 256 GB/s (256-bit bus)—it is the chip used in mini PCs such as the GMKtec EVO-X2, already available to consumers. The PRO 400 adds 64 GB of memory (192 versus 128) and slightly higher theoretical bandwidth (273 versus 256 GB/s, or about 7%), but at this stage remains positioned in professional machines rather than affordable mini PCs.

#Buy now or wait?

If you need to load a dense 70B to 120B model with a comfortable context window without spending on a multi-GPU workstation, the PRO 400 addresses a real use case: its memory capacity is real and verified by AMD. If your goal is to run a 300B model in a smooth interactive conversation, the bandwidth estimate above (around 1 to 2 tokens/s) suggests waiting for independent measurements before committing a professional budget: the capacity is sufficient, but the speed remains unproven.

You already have a Ryzen AI Max+ 395 (128 GB)
The additional 64 GB of gain does not justify an immediate repurchase for most 30B–70B uses, except when you specifically need a very long context or models above 120B.
You are targeting large MoE models
This profile gets the most out of large memory with modest bandwidth: few active parameters per token, so throughput is less penalized than with a dense model of the same size.
You need a consumer machine today
The PRO 400 is advertised for professional PCs and workstations; make sure a configuration fits your budget before waiting for a more affordable variant.
Frequently asked questions
Does the Ryzen AI Max PRO 400 have 192 GB or 352 GB of memory for AI?+
192 GB total, not 352 GB. The 160 GB AMD calls “assignable to the GPU” is taken from the same 192 GB pool shared by the processor and integrated graphics, not added on top. Once 160 GB is reserved for the GPU, 32 GB remains for the operating system and standard CPU tasks—enough for Windows or Linux, but something to watch if you run several heavy applications alongside the LLM.
Can it really run a 300-billion-parameter model?+
AMD claims that 160 GB is enough to load a 300B-plus model in 4-bit. This is a manufacturer-confirmed loading capacity, not a measured speed: with 273 GB/s of bandwidth, the theoretical throughput would drop to around 1 to 2 tokens/s, too slow for live conversation.
When will the Ryzen AI Max PRO 400 be available?+
AMD announces availability in the third quarter of 2026 through OEM partners “including HP and Lenovo”; ServeTheHome adds ASUS to the list of confirmed partners. First concrete example: the Lenovo ThinkCentre X Ultra, announced in early September 2026 with up to a Ryzen AI Max+ PRO 495 and 128 GB of RAM in the configuration shown, a planned November 2026 release, and an “expected” starting price of $3,669 according to Tom's Hardware—probably not for the maximum 192 GB configuration.
How does it differ from the Ryzen AI Max+ 395 in current mini PCs?+
The 395 (Strix Halo) tops out at 128 GB of unified memory and about 256 GB/s of theoretical bandwidth; the PRO 400 reaches 192 GB and 273 GB/s, a more significant capacity gain than speed gain (+7% bandwidth). The 395 is already used in consumer mini PCs; the PRO 400 is currently aimed at professional PCs and workstations from HP, Lenovo, and ASUS, with no affordable mini PC confirmed to date.
Do you need more VRAM or more bandwidth for a local LLM?+
The two serve complementary roles: memory capacity determines which models can be loaded (weights + context), while bandwidth determines generation speed once the model is in memory, because each token rereads all the active weights. A 300B model loaded into 160 GB fits technically, but at 273 GB/s the theoretical estimate drops to 1-2 tokens/s. For a smooth conversation, a smaller model or an MoE with few active parameters is often the better choice.
Are “Ryzen AI Halo” and “Ryzen AI Max PRO 400” the same product?+
No. The Ryzen AI Halo is a developer kit sold directly by AMD, available for preorder at Micro Center starting in June 2026, powered at launch by a Ryzen AI Max+ 395 (up to 128 GB) — not the PRO 400 yet. AMD plans to upgrade this same kit to a PRO 400 chip in the third quarter of 2026, with up to 192 GB. The “OEM” PRO 400 presented above is a separate component, integrated by HP, Lenovo, and ASUS into their own professional PCs.
Did this guide help you?

Feedback, an error, or a clarification? Let us know—it improves the guide for everyone.

Prices in euros (€) are French market prices including VAT, as checked by BestLLMfor. US prices differ: the Amazon buttons show the current US price.