PrismML has compressed a 27-billion-parameter AI model to under 4 GB, small enough to run on an iPhone. In the company's own benchmarks, the smallest version keeps 90 percent of the original performance, with math and coding scores barely affected. Apple is reportedly already testing the compression technology, which could help it close the gap in on-device AI.

PrismML says its compressed version of Alibaba’s Qwen model uses up to 15 times less memory, potentially advancing Apple’s AI push.

Apple is reportedly negotiating with PrismML to compress massive AI models for on-device iPhone use, with implications for privacy, computing, and crypto