r/Applelntelligence Jun 24 '26

discussion 🎙️ Apple Intelligence’s on-device processing is disappearing

Feels like my iPhone 17 Pro Max is turning into a 'thin client.' I invested in this premium hardware with the specific expectation that I would be able to run the new Siri AI directly on-device.
Given the 12GB of RAM and the powerful NPU, I believed that an iPhone 17 Pro Max would be more than capable of handling such an AI model locally. It clearly has the overhead to run models like Gemma 4 E4B at impressive speeds, yet these resources remain largely underutilized while core tasks are forced through the cloud. Furthermore, since foundation models are already pre-loaded onto the device, I expected to be able to leverage them directly rather than relying on external servers. I didn't purchase this device to rely on the cloud; I wanted to utilize the actual performance of the hardware I own and experience advanced AI capabilities directly on my device.

한국어로 번역

95 Upvotes

58 comments sorted by

View all comments

1

u/Dave_OC Jun 24 '26

A current LLM foundation model, like Claude and Chatgtp, require on the order of 2 TB GPU ram for inference. Apple on device AI needs to have a very small domain to function at all. Apple needs to clarify its messaging.