MobbleOpen in Mobble ⇢
Science · Physics · published 2026-09-29 · via The Register

AMD's high-memory desktop chip targets local AI model deployment

AMD's Gorgon Halo system-on-chip now powers consumer systems like GMKtec's EVO-X5 Pro with up to 192 GB of unified memory, enabling users to run large language models locally without cloud dependency. Starting at $6,799, the Gorgon Halo can execute models with around 345 billion parameters at 4-bit precision on Linux, bringing high-profile models like DeepSeek V4 Flash within reach of privacy-conscious users and small businesses. The chip features 16 Zen 5 cores, a 40-compute-unit GPU, and an XDNA 2 NPU overclocked approximately 100 MHz higher than the standard Strix Halo APU.

Expanded Detail

AMD's Gorgon Halo represents a factory-optimized variant of the Strix Halo platform, distinguished primarily by increased memory capacity and modest clock improvements. The chip maintains 16 Zen 5 cores and a 40-compute-unit GPU, but now supports up to 192 GB of faster memory running at 8,533 MT/s compared to the standard model's 128 GB at 8,000 MT/s. This memory bandwidth advantage of approximately 6.5 percent translates to marginal performance gains in token generation speed for language model inference tasks.

The pricing escalation reflects broader semiconductor market dynamics. Memory costs have risen substantially over the past year, with comparable systems doubling in price from $2,000-$3,000 to $4,000. GMKtec's entry-level configuration at $6,799 suggests manufacturers are passing along these cost increases, potentially limiting adoption among price-sensitive consumers despite the hardware's technical capabilities.

Context

Gorgon Halo systems could enable decentralized AI deployment, potentially benefiting privacy-conscious organizations and small businesses seeking on-premises model execution without cloud dependency. However, the steep entry price may restrict adoption primarily to institutions and users with substantial technology budgets, potentially widening disparities in AI access. The technology's practical impact may depend on whether memory costs stabilize, allowing manufacturers to reduce pricing and broaden the addressable market for local AI inference capabilities.

Expanded detail and Context are AI-generated analysis; the linked article remains the authoritative source.
Read the full article at The Register →
This summary is Al-enhanced to contain extended analysis and broader social context. The original is {NAME); the linked article is the authoritative source. Original headline: “AMD's 192 GB Gorgon Halo prices might leave you petrified.” Browse more stories.