Apple is in early talks with PrismML, whose on-device AI compression shrinks a 27B model to under 4GB to run on an iPhone. Apple has not commented.

PrismML says its compressed version of Alibaba’s Qwen model uses up to 15 times less memory, potentially advancing Apple’s AI push.

Apple is reportedly negotiating with PrismML to compress massive AI models for on-device iPhone use, with implications for privacy, computing, and crypto