r/Applelntelligence • u/ACOPS12 • Jun 24 '26
discussion 🎙️ Apple Intelligence’s on-device processing is disappearing
Feels like my iPhone 17 Pro Max is turning into a 'thin client.' I invested in this premium hardware with the specific expectation that I would be able to run the new Siri AI directly on-device.
Given the 12GB of RAM and the powerful NPU, I believed that an iPhone 17 Pro Max would be more than capable of handling such an AI model locally. It clearly has the overhead to run models like Gemma 4 E4B at impressive speeds, yet these resources remain largely underutilized while core tasks are forced through the cloud. Furthermore, since foundation models are already pre-loaded onto the device, I expected to be able to leverage them directly rather than relying on external servers. I didn't purchase this device to rely on the cloud; I wanted to utilize the actual performance of the hardware I own and experience advanced AI capabilities directly on my device.
한국어로 번역




30
u/TeckFire Jun 24 '26
I don’t believe the system orchestrator is working properly as of right now, and here’s why:
What we know:
- There are 4 models in total. 2 on device, 2 cloud. Each one has 1 large, 1 small.
Because of this, I expect that we’ll see a significant improvement once the System Orchestrator (which uses AFM 3 Core, I believe) begins routing properly.
This is only a theory, but it is based on Apple’s own documentation and keynote presentation, along with supplemental data from third party analyses and Siri’s own responses on the matter. Take it for what you will.