A Caltech spinout says it fits a 54 GB model into 3.9 GB on a phone, and Apple started testing it the week its Siri beta went public. The timing is the story.

PrismML says its compressed version of Alibaba’s Qwen model uses up to 15 times less memory, potentially advancing Apple’s AI push.

Apple is reportedly negotiating with PrismML to compress massive AI models for on-device iPhone use, with implications for privacy, computing, and crypto