I was responding to a comment that directly said this was a stop gap before they made a real AI centric chip, in that context, the chips will need more ram, until they do, they will be offloading a lot of stuff to the server.
The Ultra may support 192gb but that also comes at an insane cost, if AI is going to become mainstream and Apple want to do this on device, the base models are going to need more Ram, processing power on these systems is not the bottleneck, the RAM is.
The whole context of this article and post is Apple using their own chips on the server for AI tasks.
The “stopgap” is using the existing M2 Ultra for this, versus a chip specifically designed for AI server duty.
I don’t think the person you responded to is talking about on-device processing. Apple definitely wants to do as much of this on-device as they can of course, that’s just kind of tangential to this discussion on which server chips they are using.
The Ultra may support 192gb but that also comes at an insane cost, if AI is going to become mainstream and Apple want to do this on device, the base models are going to need more Ram, processing power on these systems is not the bottleneck, the RAM is.