Skip to main content

AMD Ryzen AI Max Pro 495

Overview​

AMD Ryzen AI Max Pro 495 (upgraded Strix Halo) was unveiled on October 5, 2026 as AMD's flagship AI PC / agentic on-device APU, deliberately timed two days ahead of NVIDIA's RTX Spark launch event (October 7).

The headline feature is up to 192 GB of unified memory — via AMD's Variable Graphics Memory (VGM) slider, users can allocate up to 160 GB to the GPU, enough to fit and run 100B-parameter-class models entirely offline with no cloud subscription and no data leaving the device.

Core Specs​

ItemSpec
ArchitectureHeterogeneous chiplet (CPU + GPU + I/O die)
CodenameStrix Halo (upgraded)
CPU16 desktop-class Zen 5 cores
GPU architectureRDNA 3.5 integrated graphics
Unified memoryUp to 192 GB LPDDR5X
Memory speed8.5 GT/s
Memory bandwidth~272 GB/s (derived from 256-bit @ 8.5 GT/s)
Allocatable VRAMUp to 160 GB (VGM)
Model capacityRuns 100B-parameter-class models locally, offline
TDPNot disclosed (PRO series reference: 55-120W cTDP)
UnveiledOctober 5, 2026
AvailabilityNot disclosed

⚠️ Note: NPU TOPS, total AI TOPS, pricing, and availability have not been officially disclosed; this table lists only confirmed information.

Head-to-Head vs RTX Spark (N1X)​

MetricRyzen AI Max Pro 495NVIDIA RTX Spark (N1X)
CPU16 Zen 5 cores (native x86)20 Arm cores
GPURDNA 3.5 integrated6,144 CUDA Blackwell
Unified memoryUp to 192 GBUp to 128 GB LPDDR5X
Allocatable VRAMUp to 160 GB (VGM)Shared unified memory
Local model capacity100B-parameter class120B parameters
AI computeNot disclosed1 PFLOPS FP4 sparse
Windows compatibilityNative x86, no emulation layerArm relies on Prism translation
EcosystemRuns directly on Windows + LinuxNative CUDA on Windows

AMD's core pitch: native x86 with no emulation layer — the RTX Spark Arm approach still depends on Microsoft's Prism translation layer in the Windows ecosystem (with lingering power and compatibility costs), while the Pro 495 runs directly on both Windows and Linux. AMD's marketing even rebrands the acronym as "Agentic Micro Devices since 2025."

Real-World Deployment​

Emmy-winning VR studio LightSail VR (16K, 90 fps stereoscopic 3D video) built a "production coordinator" AI agent on a Ryzen AI Max desktop, managing 6-9 production projects simultaneously — the workload of two people. Fully local means zero cloud cost; the same workload in the cloud would cost thousands of dollars per month in tokens alone.

Use Cases​

  • ✅ Local 100B-class LLM inference (192GB / 160GB VRAM)
  • ✅ Privacy-sensitive workloads (tax documents never leave the device)
  • ✅ Agentic AI workflows (parallel agents, zero cloud cost)
  • ✅ Professional creation (16K video, 3D rendering)
  • ❌ Cloud-scale training (MI455X / Helios territory)