Top
Best
New

Posted by thm 18 hours ago

Apple caught off guard by AI demand for Mac Mini and Mac Studio(www.macrumors.com)
366 points | 415 commentspage 3
bilsbie 7 hours ago|
I get the impression they want AI marketing points but don’t actually want people to use local AI on their products.

I have no idea why. They could be so successful if they leaned into local AI.

m463 7 hours ago||
I don't think they're ready for local ai. They have memory + memory bandwidth, that's it.

I also don't think they do "technology". For example, containers have been around for a long time, and apple didn't show up. (I know they have some support now). Imagine an apple-native docker/podman doing something like FROM macos:10.12

I was actually surprised when they did their own chips. I figure it was about control.

jonplackett 7 hours ago||
I think they’re too scared to ‘own’ it - it would be someone else’s model and potential security issue.

But apple have to own everything they do so they’re in a bind

ghostly_s 9 hours ago||
> Apple's unusually timed announcement of new Mac mini and Mac Studio models this week was driven by unexpectedly strong enterprise appetite for AI hardware, according to The Information.

Obviously; no one else can justify the expense.

skybrian 9 hours ago||
Maybe it's not anything specific to Apple? There's high demand and short supply elsewhere due to AI, so it doesn't seem all that odd that many companies would try to buy gear from Apple too.
jmyeet 12 hours ago||
So for people who don't understand, there are two markets for Apple hardware in this space:

1. Running an agent like OpenClaude. The $599 Mac Mini was an insanely good deal for this. I happened to buy a M5 Pro Mac Mini for $999 last year for other reasons. The equivalent is now almost $2000; and

2. Hardware for running inference on local models. This to me is the far more interesting market because Apple has a real opportunity to disrupt NVidia's stranglehold on the market.

With current architecture, the largest model you can reasonbly run is the amount of memory on the GPU and is a function of the quantization (eg int4, int8, fp8, fp16, etc) available and the number of parameters. NVidia aggressively segments the market. The most VRAM on a "consumer" card is 32GB on the 5090, which allows you to run ~31B parameter models.

In comparison, the RTX 6000 Pro has only slightly more CUDA units than a 5090 but has 80GB of VRAM. A few months ago they were $10-11k. Now they're ~$16k.

Macs use a shared memory architecture. Apple has previously sold Mac Studios with up to 512GB of RAM. Almost all of that memory can be used to hold much larger models without taking a penalty for interconnections between different GPUs or machines. Plus Apple interconnects between computers are actually relatively good by chaining TB5. It's still slow but it's about the best non-enterprise option available.

But the previous Mac Studios just didn't have the raw FLOPS and memory bandwidth. The M5 Ultras are up to 1.2TB/s of memory bandwidth. M3 Ultra had ~900GB/s. RTX 5090s and RTX 6000 Pros are 1.8TB/s. The current best HBM3 NVidia DC GPUs are at 3.2TB/s IIRC. But the M5 Ultra has a claimed ~4.5x the FLOPS of the M3 Ultra.

We don't have our hands on these yet but it probably means they are going to be much closer to a 5090. I expect ~50% of a 5090's inference speed. That may sound bad but it's actually really good because a 256/512GB Mac Studio can probably locally run the best Flash models. With NVidia hardware you'll need to spend many tens of thousands for that.

We'll see what the inference speed is but I expect it to be usable. DeepSeek v4 Flash, for example, will be entirely runnable. We're not at DeepSeek v4 Pro local yet.

jubilanti 10 hours ago||
> 1. Running an agent like OpenClaude. The $599 Mac Mini was an insanely good deal for this.

I still have zero clue how "Buy a $599 Mac Mini to have a sandboxed LLM API caller" became the default. If you're not doing local inference and don't need to inject into iMessage or iCloud, all you need to run openclaw-style harnesses that call external APIs is a Raspberry Pi 4B, an N100, an HTPC, or that 10 year old laptop sitting in your desk.

nullbio 5 hours ago|||
Apple has a huge marketing budget, and evidently, they are not beyond using unethical tactics to sell their products. That's how it started.
MaxikCZ 6 hours ago|||
People bought the mac mini and then spent more money to have someone install openclaw for them, it was wild.
subarctic 10 hours ago||
You have the m4 pro right? I thought the m5 pro mac mini was only just announced
ChrisMarshallNY 11 hours ago||
Sounds like people want those bespoke servers that Apple has been rumored to have developed.
comrade1234 17 hours ago||
I wish they sold something that could go in a colo - redundant power supplies, lights out management, etc. you know they have them internally...
dewey 17 hours ago||
> you know they have them internally...

What makes you think that? There's a lot of data centers that sell you access to colocated Mac Mini's, they have added FileVault unlock via SSH in the boot process which also makes things easier. There's not that many reasons to run a Mac in the cloud unless you have some very specific Mac related workload.

giancarlostoro 17 hours ago|||
Because there's been photos of Apple building server racks with Apple Silicon, but also a recent leak.

https://www.macrumors.com/2026/08/26/leaked-images-of-apple-...

https://www.reuters.com/business/apple-begins-shipping-ai-se...

scrlk 17 hours ago||||
Apple built internal M5 servers for private cloud compute:

https://wccftech.com/apples-private-cloud-compute-server-m5-...

https://security.apple.com/blog/private-cloud-compute/

dewey 16 hours ago||
Thanks, missed that article.
N_A_T_E 17 hours ago||||
At this point it’s a well known secret that Apple has real rack mount servers for their internal processes. They actually have officially released video of their servers in the WSJ report on their chip supply chain.

https://forums.macrumors.com/threads/photos-of-apples-own-ne...

reaperducer 10 hours ago||
it’s a well known secret that Apple has real rack mount servers for their internal processes

Only if by "secret" you mean "announced in multiple press releases and a public event with federal, state, and local officials at its new sever factory in Houston."

https://www.apple.com/newsroom/2026/08/apple-opens-advanced-...

ndiddy 17 hours ago||||
> What makes you think that?

There’s articles about them, Apple uses them internally for AI services. https://forums.macrumors.com/threads/photos-of-apples-own-ne...

unrented7977 16 hours ago||||
Apple has to have significant build infrastructure to support internal iOS development, surely? They can't just be using whatever is at the developers' desk, or a big pile of Mac minis in a closet. That's far too pedestrian for Apple internal works.

Plus they did sell rackmount servers for some time.

shepmaster 17 hours ago||||
> What makes you think that?

Here's a leaked / rumor image of Apple servers themselves.

https://www.macrumors.com/2026/08/26/leaked-images-of-apple-...

abtinf 17 hours ago|||
> [cites a convoluted work around]

> [still claims there is no reason]

detourdog 17 hours ago||
They did. Now think they feel a stack on Minis or Studios fills the reduce needs better. The multiple machines one gets software redundancy in addition to everything else.
moezd 10 hours ago||
Classic monopoly move: Control the user base, then control hardware. Any decent always-on local LLM setup with Apple devices will have to compete with these behemoths now. Great.
GeekyBear 10 hours ago|
> Classic monopoly move: Control the user base, then control hardware

Classic monopoly move by who?

Apple created MLX as an open source framework to allow users to run any open model locally.

bel8 10 hours ago||
just to clarify, models already ran locally without MLX years before it existed, on non-Apple environments.

MLX was just Apple's bridge to what already ran in other hardware.

GeekyBear 9 hours ago|||
MLX is an open source framework that allows you to run open models in a manner optimized for Apple's hardware.

You can take advantage of larger amounts of memory, higher memory bandwidth, and clustering multiple systems.

How is using an open source framework to run open models a monopoly move?

Danox 6 hours ago|||
Apple Pay is a bridge too because at the time the other pay services didn’t care about doing anything for the Mac, and over the years that’s the way it’s been that’s the way it goes when you’re the last vertical computer company left from the 1980s, the Apple Watch and Apple Maps is also a bridge for the same reason…
shevy-java 10 hours ago||
This sounds like advertisement, disguised as an "article".
LoganDark 10 hours ago||
I hope Apple does not gain some exclusive enterprise tier for hardware. Part of what I love about them is that everything is available to consumers. A lowly home user can buy the exact same 256 (or 512) gigabytes of memory in a Mac from Apple, as long as they have a couple dozen thousand dollars to spare. I'd be really sad to lose that.
More comments...