Apple is advancing its AI capabilities on two fronts: talks with startup PrismML to run large models on-device and regulatory approval in China to integrate Alibaba's Qwen model into Apple Intelligence.
PrismML, a Caltech spinout backed by Khosla Ventures, released compressed versions of Alibaba's open-source Qwen model on Tuesday. The company reduced the model from roughly 54GB to under 4GB, allowing its 27 billion parameters to run on an iPhone 15 or newer Source: CNBC. PrismML CEO Babak Hassibi told CNBC that Apple and other companies are evaluating the technology, calling the discussions "very early" but noting "things are progressing nicely" Source: The Next Web. The compression technique reduces each value from 16 bits to one or three possible values, cutting memory use by 10–15 times and speeding up responses by 6–8 times, while sacrificing about 5–10% performance Source: TechCrunch.
China's Cyberspace Administration approved Apple Intelligence for launch in the country, integrating Alibaba's Qwen model across iOS, iPadOS, macOS, and visionOS Source: TechCrunch. Alibaba confirmed the partnership, stating Qwen will be "integrated into Apple Intelligence experiences" for text and image understanding and generation Source: The Next Web. The approval ends a lengthy regulatory process and follows Apple's earlier failed attempts with Baidu, DeepSeek, and ByteDance Source: TechCrunch. Alibaba's US-listed shares rose nearly 4% on the news Source: CNBC.
Analysts caution that real-world performance across millions of devices remains to be proven, with power consumption as a key question Source: The Next Web.
“PrismML CEO Babak Hassibi told CNBC that Apple and other companies have been evaluating the startup's models and measuring their speed, energy efficiency and performance on devices.”
“Alibaba confirmed the company's news to CNBC in a statement, saying that Qwen would be 'integrated into Apple Intelligence experiences,' but did not provide a timeframe.”
“PrismML ships two versions under a free licence. A ternary build runs on a laptop. A smaller 1-bit build, about 3.9GB, is designed to fit within the memory budget of an iPhone 17 Pro.”
“Alibaba confirmed on Wednesday that its Qwen AI model will be integrated into Apple Intelligence across iOS, iPadOS, macOS, and visionOS for users in China.”
“Prior to working with Alibaba, Apple was reportedly exploring a deal with Baidu, but faced issues adapting its models for Chinese customers.”
“The startup said the compressed models use between 10 and 15 times less memory, generate responses six to eight times faster and consume three to six times less energy than conventional versions.”