Archive for August 25, 2026

Tuesday, August 25, 2026

Mac Studio 2026

Apple (Hacker News, MacRumors):

Mac Studio with M5 Max features an 18-core CPU, an up-to-40-core GPU with Neural Accelerators built into each core, and up to 128GB of unified memory, accelerating complex pro and AI workloads. With the powerful M5 Ultra, Mac Studio scales up to a 36-core CPU, up to an 80-core GPU, and a staggering 512GB of unified memory, enabling users to run enormous LLMs entirely on device. Wi-Fi 7 and Bluetooth 6 come to Mac Studio for the first time, while Thunderbolt 5 rounds out its extensive connectivity, so users can take advantage of blazing-fast external storage, PCIe expansion chassis, and powerful hub solutions for the most intense workloads. Thunderbolt 5 also enables multiple Mac Studio systems to be clustered, bringing up to 3x faster performance for distributed AI inference when compared to a single system.

[…]

Mac Studio with M5 Max starts at $2,499 (U.S.) and $2,299 (U.S.) for education.

[…]

Mac Studio with M5 Ultra starts at $5,499 (U.S.) and $5,099 (U.S.) for education.

It’s not shipping until September 22, with the 512 GB configuration in late October.

John Gruber:

Here’s my attempt to put all of the RAM/SSD configurations into condensed tables, so you can see which storage and memory options are available for each chip, and how much they cost.

[…]

It kind of stinks that there are no RAM options for the Studio between the 96 GB base and the $4,000 256 GB upgrade.

Jason Snell:

The Mac Studio is effectively the replacement for the Mac Pro, and it’s alone in offering the Ultra-class chip. Apple has positioned the Mac Studio as ideal for local AI workflows, and the new models will still be able to cluster via Thunderbolt 5 to create even larger collections of memory and performance. Apple representatives pointed out that a four-Mac Studio AI cluster, like the one I saw on display at WWDC earlier this summer, is so efficient that it can be powered from a single standard wall outlet. (They even showed us a picture of four Mac Studios plugged into a power strip that was plugged into the wall.) In an era of data center excesses, Apple is clearly leaning into the possibility of small, efficient Macs being used to do local AI work rather than relying on huge, power-hungry cloud models.

Federico Viticci:

If these numbers hold up and scale linearly, a local Mixture-of-Experts model such as Qwen 3.5-35B-A3B, which would run at ~17 tokens/sec on average on a base M4 Mac mini with 16 GB of RAM, could realistically generate output at over 60 tokens/second with the base model M6 Mac mini.

Based on what we’ve seen so far, the one downside of the M6 Mac mini is that it does not have Thunderbolt 5 ports; those are exclusive to the M5 Pro model, also announced today.

[…]

Speaking from personal experience, I know that my M3 Ultra Mac Studio using oMLX can run DeepSeek-V4-Flash locally with generation averaging 35 tokens/second. Assuming a linear 4x increase, that would put the same model at over 120 tokens/second on an M5 Ultra Mac Studio. To put things in perspective, that kind of performance would be faster than any AI chatbot website, it’d be faster than many providers who offer a “fast” mode for their models, and it’d only be second to either dedicated NVIDIA PC clusters at home or specialized inference providers such as Cerebras or Groq…which are running in full-blown data centers. Sure, you would need a computer that is likely going to cost more than $20,000 to make it happen, but it’d still be possible on a single machine that is small, quiet, and that – in theory – any consumer can buy off the shelf.

Previously:

Mac mini 2026

Apple (Hacker News, MacRumors):

With M6, Mac mini now delivers up to 4x faster AI performance, 2x faster storage and graphics, and 40 percent faster CPU performance. Everything on Mac mini with M6 feels incredibly fast, from everyday productivity tasks to agentic AI workflows. Mac mini with M5 Pro delivers even more pro-level performance to breeze through demanding projects, from video production to game development. And this new level of performance takes on business workflows with ease whether Mac mini is being used as a primary desktop or for always-on, deskside agentic computing. Both Mac mini models include Wi-Fi 7 and Bluetooth 6, as well as upgraded 2.5Gb Ethernet, with a 10Gb option available.

[…]

Mac mini with M6 starts at $899 (U.S.) and $799 (U.S.) for education.

[…]

Mac mini with M5 Pro starts at $1,699 (U.S.) and $1,599 (U.S.) for education.

Rui Carmo:

The M6 tops out at 170GB/s and 32GB of unified memory, whereas the M5 Pro offers 307GB/s and up to 64GB–nearly twice the bandwidth, as well as twice the memory ceiling. That is the number to watch for AI inference, not just CPU and GPU benchmark deltas.

John Gruber:

Here’s my attempt to put all of the RAM/SSD configurations into condensed tables, so you can see which storage and memory options are available for each chip, and how much they cost.

[…]

If you configure an M6 Mac Mini with 2 TB of storage, the SSD upgrade ($1,000) costs more than the entire base model computer ($900). So too with the 4 TB SSD upgrade for the M5 Pro Mini ($1,800 upgrade for a $1,700 computer).

Nick:

Almost double (in Canadian dollars) what I paid for my M4 mini 18 months ago.

Previously:

Apple M6 and M5 Ultra

Apple (MacRumors M5 Ultra and M6, Hacker News):

M6 is built using cutting-edge 2 nm process technology, packing greater transistor density into a smaller die for a major leap in performance and power efficiency. M6 also introduces a Dual 16-core Neural Engine, providing up to 2x the peak compute over previous generations to make on-device AI workflows run even faster. System frameworks can automatically utilize both engines simultaneously, enabling applications to see faster model execution.

M6 has a brand-new 12-core CPU complex — two more cores than M5 — that consists of 2 super cores, 4 performance cores, and 6 efficiency cores. It delivers the world’s fastest single-threaded performance and up to 1.2x faster multithreaded performance as compared to M5, and up to 2.4x faster than M1.

[…]

M5 Ultra uses UltraFusion to connect two dual-die M5 Max chips to form the quad-die architecture — a first for Apple silicon. UltraFusion increases the inter-die bandwidth to over 4.4TB/s and the connection density by over 6x. Together, these ultra-low-latency, high-bandwidth interconnects allow the four dies to behave as a single unified processor. M5 Ultra also features a large up-to-36-core CPU consisting of 12 super cores and 24 performance cores, delivering up to 1.25x higher single-threaded performance and up to 1.3x higher multithreaded performance than M3 Ultra.

M5 Ultra features a next-generation GPU with up to 80 cores, incorporating a Neural Accelerator in each core to offer up to 4.5x the peak GPU compute for AI compared to M3 Ultra and over 6x more than M1 Ultra.

ghostly_s:

It was not long ago all the illustrations in an M chip press release were graphs showing it flouncing the competition. Now it’s just random screenshots of desktops. Skimming this I didn’t see a single claim of performance advantages over competitors. The massive thing Apple still needs to solve is making sure customers don’t miss being able to plug the latest greatest GPU into their tower, they should not be resting on their laurels.

Greg Pierce:

I’m so old I remember when Apple PR benchmarks were all like “Photoshop Gaussian Blur in only 48 seconds!”

Previously: