Apple Is Secretly Building an AI Server Rack Full of M-Series Ultra Chips
Key takeaways
- Apple is reportedly building a rack server with multiple M-series Ultra chips, targeting a 2029 release
- This would be Apple's first enterprise server since the Xserve was discontinued in 2011
- The M3 Ultra chip features 192GB unified memory, a 32-core CPU, and an 80-core GPU
- The design could appeal to enterprises with data sovereignty concerns who want on-premises AI inference
Apple is quietly working on something that would mark a significant departure from its usual playbook: a dedicated server designed for enterprise AI workloads, reportedly packed with multiple M-series Ultra chips. According to a report from Ars Technica, the machine is targeting a 2029 launch, which would make it Apple's first proper enterprise server in decades.
This is interesting for several reasons, and not just the obvious one about Apple entering yet another market.
What the Hardware Reportedly Looks Like
The details that have emerged suggest Apple is designing a rack-mounted server that consolidates multiple M-series Ultra processors. The M-series Ultra chips are already formidable on paper. Apple's M3 Ultra, for instance, features 192GB of unified memory and a 32-core CPU alongside an 80-core GPU. The unified memory architecture, where the CPU, GPU, and neural processing units all share the same memory pool, is something Nvidia and AMD simply do not offer at comparable scale. If Apple can pack several of these together effectively, the result could be a server with genuinely unusual memory bandwidth characteristics.
Apple's existing Private Cloud Compute infrastructure, which powers the more demanding Apple Intelligence tasks that cannot run on-device, already uses custom Apple Silicon in data centre configurations. This reported server sounds like the next step: making that hardware available as a product that external enterprises could buy or lease, rather than keeping it exclusively internal.
Why 2029 Is Significant
A 2029 debut is not imminent, but the timeline matters. By that point, the AI server market will look very different from today. Nvidia currently dominates with roughly 70 to 80 percent market share in AI accelerators. AMD's MI-series chips have been making inroads, and Intel has been trying to compete in the inferencing space. Google's TPUs and Amazon's Trainium chips are proving that custom silicon can be cost-effective at hyperscale.
Apple arriving in 2029 with a differentiated architecture, rather than trying to out-CUDA Nvidia on Nvidia's own terms, could actually carve out a niche. The unified memory approach makes particular sense for inference tasks on large language models, where moving data between separate memory pools is a significant bottleneck. There are workloads where Apple's architecture would be genuinely faster, not just competitive on price.
The enterprise angle is what makes this most surprising. Apple has historically been reluctant to pursue enterprise hardware in any serious way. The Xserve, Apple's previous rack server, was discontinued back in 2011 after a relatively short and commercially underwhelming run. The company has spent the intervening 15 years building its business around consumer devices and services. Returning to enterprise hardware would require a very different go-to-market motion, different sales relationships, different support infrastructure.
The Bigger Picture
This move makes most sense when you frame it as an extension of Apple Intelligence rather than a standalone hardware play. Apple has staked a significant part of its device strategy on AI features that are genuinely private and on-device wherever possible. For the tasks that cannot run on a phone or laptop, Private Cloud Compute acts as a trusted extension of the device. Selling that infrastructure to enterprises, particularly those with strict data sovereignty requirements, follows a coherent logic.
Financial services, healthcare, and legal firms all have data that they are reluctant to route through Microsoft Azure or Google Cloud, even with encryption guarantees. An Apple server that can process sensitive queries without data leaving a client's own infrastructure could appeal precisely to that anxiety.
There is plenty that could go sideways. Apple has very little track record in enterprise sales, and the ecosystem integrations that make Apple hardware compelling for consumers, things like iCloud, AirDrop, and Continuity, are largely irrelevant in a data centre context. It would need to build or acquire the enterprise credibility from scratch.
But as a signal about where Apple thinks the next decade of computing is headed, a 2029 AI server is one of the more telling things the company has quietly done.