
That’s solid performance, but in both cases, “up to” carries a lot on its shoulders. The story of token throughput is told as the context length increases, showcasing what happens when someone actually runs these models locally, not just boots them up cold. Performance drops at higher context lengths, naturally, which could pose some issues for the larger models. If GLM 5.3 Flash provides up to 20 tokens per second, it could very easily decline into unusable territory as the context length increases.
We largely know what performance to expect out of the Ryzen AI Max+ Pro 495, and Gorgon Halo more broadly. It’s a refresh of Strix Halo, with notable spec changes being the bump up to 192GB of unified memory from 128GB, as well as a 100 MHz jump on boost clocks for the 495. Otherwise, the range is using identical core counts and microarchitectures as previous-gen Strix Halo chips.
With these proxies — Strix Halo for Gorgon Halo, and GB10 for RTX Spark — we can already get a good idea about how these parts will stack up. As you can see in our Ryzen AI Halo review (packing the Ryzen AI Max+ 395), AMD’s part universally underperformed compared to the DGX Spark in both time to first token and tokens per second across three models. More unified memory will allow you to run larger models, but that doesn’t mean those models will run faster .
The first Gorgon Halo devices are available for sale now, such as the Minisforum MS-S1 Max-P495 . For the top-line configuration, prices sit around $7,000 right now, though we expect a broad range of prices once different devices are available, likely driving above that $7,000 mark.
(Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) (Image credit: AMD) Image 1 of 24 View Original
Follow Tom's Hardware on Google News , or add us as a preferred source , to get our latest news, analysis, & reviews in your feeds.
Jake Roach Social Links Navigation Senior Analyst, CPUs Jake Roach is the Senior CPU Analyst at Tom’s Hardware, writing reviews, news, and features about the latest consumer and workstation processors.
Neilbob Meh, whatever. More importantly, the messed up font substitution issue we've seen before on AMD slides makes my bum hurt. Reply
Key considerations
- Investor positioning can change fast
- Volatility remains possible near catalysts
- Macro rates and liquidity can dominate flows
Reference reading
- https://www.tomshardware.com/pc-components/cpus/SPONSORED_LINK_URL
- https://www.tomshardware.com/pc-components/cpus/amd-attempts-to-get-ahead-of-expected-rtx-spark-launch-with-gorgon-halo-ai-benchmarks-company-says-it-has-shipped-over-half-a-million-agentic-pcs-to-date#main
- https://www.tomshardware.com/membership
- Russia's uncrewed robot tank fails during debut military display in front of President Putin
- Google AI data center project investigated after 420 football fields of Finnish forest demolished
- German utility provider introduces 'gaming electricity' plan targeting high-consumption households, like those running multiple high-end gaming PCs
- Nintendo Switch 2 drops to £354.99 all-time low to defy the AI tax — pocket £65 in savings across these retailers
- NVIDIA DGX Spark 64GB Gives Developers More Ways to Build and Scale Local AI
Informational only. No financial advice. Do your own research.