Apple (Hacker News, MacRumors): Mac Studio with M5 Max features an 18-core CPU, an up-to-40-core GPU with Neural Accelerators built into each core, and up to 128GB of unified memory, accelerating complex pro and AI workloads. With the powerful M5 Ultra, Mac Studio scales up to a 36-core CPU, up to an 80-core GPU, and […]
Apple (Hacker News, MacRumors):
Mac Studio with M5 Max features an 18-core CPU, an up-to-40-core GPU with Neural Accelerators built into each core, and up to 128GB of unified memory, accelerating complex pro and AI workloads. With the powerful M5 Ultra, Mac Studio scales up to a 36-core CPU, up to an 80-core GPU, and a staggering 512GB of unified memory, enabling users to run enormous LLMs entirely on device. Wi-Fi 7 and Bluetooth 6 come to Mac Studio for the first time, while Thunderbolt 5 rounds out its extensive connectivity, so users can take advantage of blazing-fast external storage, PCIe expansion chassis, and powerful hub solutions for the most intense workloads. Thunderbolt 5 also enables multiple Mac Studio systems to be clustered, bringing up to 3x faster performance for distributed AI inference when compared to a single system.
[…]
Mac Studio with M5 Max starts at $2,499 (U.S.) and $2,299 (U.S.) for education.
[…]
Mac Studio with M5 Ultra starts at $5,499 (U.S.) and $5,099 (U.S.) for education.
It’s not shipping until September 22, with the 512 GB configuration in late October.
Here’s my attempt to put all of the RAM/SSD configurations into condensed tables, so you can see which storage and memory options are available for each chip, and how much they cost.
[…]
It kind of stinks that there are no RAM options for the Studio between the 96 GB base and the $4,000 256 GB upgrade.
The Mac Studio is effectively the replacement for the Mac Pro, and it’s alone in offering the Ultra-class chip. Apple has positioned the Mac Studio as ideal for local AI workflows, and the new models will still be able to cluster via Thunderbolt 5 to create even larger collections of memory and performance. Apple representatives pointed out that a four-Mac Studio AI cluster, like the one I saw on display at WWDC earlier this summer, is so efficient that it can be powered from a single standard wall outlet. (They even showed us a picture of four Mac Studios plugged into a power strip that was plugged into the wall.) In an era of data center excesses, Apple is clearly leaning into the possibility of small, efficient Macs being used to do local AI work rather than relying on huge, power-hungry cloud models.
If these numbers hold up and scale linearly, a local Mixture-of-Experts model such as Qwen 3.5-35B-A3B, which would run at ~17 tokens/sec on average on a base M4 Mac mini with 16 GB of RAM, could realistically generate output at over 60 tokens/second with the base model M6 Mac mini.
Based on what we’ve seen so far, the one downside of the M6 Mac mini is that it does not have Thunderbolt 5 ports; those are exclusive to the M5 Pro model, also announced today.
[…]
Speaking from personal experience, I know that my M3 Ultra Mac Studio using oMLX can run DeepSeek-V4-Flash locally with generation averaging 35 tokens/second. Assuming a linear 4x increase, that would put the same model at over 120 tokens/second on an M5 Ultra Mac Studio. To put things in perspective, that kind of performance would be faster than any AI chatbot website, it’d be faster than many providers who offer a “fast” mode for their models, and it’d only be second to either dedicated NVIDIA PC clusters at home or specialized inference providers such as Cerebras or Groq…which are running in full-blown data centers. Sure, you would need a computer that is likely going to cost more than $20,000 to make it happen, but it’d still be possible on a single machine that is small, quiet, and that – in theory – any consumer can buy off the shelf.
Previously:
Update (2026-08-26): John Gruber:
That upgrade is labeled “+ $300”, which makes it look as though the starting price for the 18/40-core model is $2,800. That’s the price I put in the original version of my chart.
But if you select that option, you’ll notice that the actual starting price jumps from $2,500 to $3,100 — a $600 difference, not $300. The reason is that the $2500 18/32-core version only comes with one option for RAM: 32 GB. The 18/40-core chip has three tiers for RAM: 48, 64, and 128 GB.
Update (2026-09-02): Joe Rossignol:
Apple’s press release for the new Mac Studio with M5 Max and M5 Ultra chips last week initially stated that the computer had “next-generation SSD architecture built on PCIe Gen 6.” However, as spotted by the French blog MacGeneration, Apple removed the PCIe 6.0 mention from the announcement shortly after it was published.
Update (2026-09-22): Federico Viticci (Hacker News):
In my day-to-day experience with agents running on the M5 Ultra, these improvements to token prefill (or how quickly a prompt can be processed) and token generation are the changes I noticed immediately. When comparing a model running on the M3 Ultra and M5 Ultra side by side with Open Minis on iOS, the M5 Ultra was ~70% faster on average than the M3 Ultra at generating a response.
[…]
In my tests, prompt processing is up 150% on average from the M3 Ultra – a ~2.5× improvement from my previous setup. This change alone makes local models solid choices in apps like Open Minis and Hermes Agent.
[…]
As you’ll see from the visualizations later in this article, NVIDIA’s RTX 5090 is still faster than Apple’s M5 Ultra despite its “meager” 32 GB of VRAM, for two different reasons.
Update (2026-10-08): Andrew Cunningham:
In many of our general-purpose CPU and GPU tests, the Ultra’s CPU outruns the M5 Max by 80 or 90 percent, and the GPU is between 50 and 80 percent faster. That’s essentially in keeping with what we’ve observed in past Ultra chips—the CPU comes closer to 2x scaling than the GPU does. Compared to the outgoing M3 Ultra, M5 Ultra usually posts around 30 percent faster single-core CPU speeds, 50 percent faster multi-core CPU speeds, and GPU performance that’s anywhere from 33 to 66 percent faster, depending on the test. As you’d expect for a two-generation upgrade, it’s a big one.
But we also observed some less-expected behavior. The Ultra’s Geekbench multicore performance is only around 26 percent faster than the Max, and in our CPU-based Handbrake video encoding test, the Max is actually faster to complete the H.264 encode (and barely slower at H.265).
Looking at the power consumption numbers offers a possible explanation.
[…]
It’s genuinely fun and freeing to be able to write hobby-project code at usable speeds without my data ever leaving my control. It’s just too bad I’ve made this discovery as Apple has instituted 25-percent-and-up price increases across the entire Mac Studio line.
Previously:
Apple Hardware Announcement Apple M5 Max Apple M5 Ultra Artificial Intelligence Mac Mac Studio macOS Tahoe 26
16 Comments| # | Наименование новости | Тональность | Информативность | Дата публикации |
|---|---|---|---|---|
| 1 | Apple анонсировала Mac Studio с чипами M5 Max и M5 Ultra | 0 | 7.62 | 26-08-2026 |
| 2 | Apple’s new 2026 M5 Max Mac Studio is now $50 off at Amazon | 0 | 18.15 | 25-09-2026 |
| 3 | Apple’s 2026 Mac Studio is nearly $50 off ahead of release tomorrow | 0 | 21.41 | 21-09-2026 |
| 4 | Apple Mac Studio M5 Ultra and Mac Mini M6 Launched | 0 | 4.89 | 26-08-2026 |
| 5 | Apple prepara una revolución: el chip M7 Ultra triplicará la memoria de sus Mac más potentes | 0 | 18.86 | 13-07-2026 |
| 6 | Launch deals now live on new M6 Mac mini and Mac Studio at up to nearly $50 off | 0 | 14.56 | 23-09-2026 |
| 7 | Deals: $600 Off 16″ M5 Max MacBook Pro 36GB/2TB, & $150 Off M5 MacBook Air 15″ | 0 | 14.6 | 28-09-2026 |
| 8 | Deals: $200 Off M5 Max Mac Studio, $120 Off M6 Mac mini, $200 Off M5 MacBook Air, etc | 0 | 10.27 | 05-10-2026 |
| 9 | M5 Pro Mac mini with AppleCare+ gets $30 off at Amazon | 0 | 19.17 | 09-09-2026 |
| 10 | Apple cambia su estrategia de chips: adiós a los M6 Pro y M6 Max | 0 | 17.67 | 25-06-2026 |