Apple reportedly trying to distill Google’s multi-trillion-parameter Gemini AI to run on iPhone

Apple reportedly trying to distill Google’s multi-trillion-parameter Gemini AI to run on iPhone

As the integration of generative AI becomes increasingly prevalent in technology, Apple is making significant strides to enhance its Siri virtual assistant. The company has faced delays in rolling out AI improvements since promising an upgraded version of Siri for 2024. However, a strategic collaboration with Google is set to merge Siri with the advanced Gemini AI system later this year, just in time for the upcoming WWDC event. Despite Apple's long-standing commitment to privacy and its preference for local AI processing, recent reports indicate a shift in strategy. According to The Information, the revamped Siri will operate both on-device and in the cloud, which marks a departure from Apple's traditional focus on local processing. This dual approach could raise concerns among Apple enthusiasts who value the company's privacy-centric ethos. With each new chip release, Apple touts its advancements in AI optimization, particularly through its Neural Engine upgrades. However, the reality is that while smartphones are marketed as capable of handling complex AI models, their hardware often falls short. Most mobile GPUs can process more tokens than AI-specific NPUs, yet they still lack the necessary RAM to support vast AI models. The AI models currently running on smartphones are limited in size, typically containing only a few billion parameters. In stark contrast, Google's latest Gemini models boast trillions of parameters. This disparity can lead to on-device AI that operates with reduced precision, making it faster but less accurate in generating responses. Consequently, these smaller models may not deliver the same level of intelligence as their cloud-based counterparts. Google has developed mobile-optimized versions of Gemini, known as Gemini Nano, which are intended for specific tasks like contextual features and audio summarization. However, the requirements of a conversational assistant like Siri demand a different approach, one that is currently reliant on cloud processing. In fact, on Android devices, Google has opted to direct all interactions with Gemini to the cloud, foregoing local processing entirely.

Sources : Ars Technica

Published On : May 28, 2026, 18:35

Computing
Market Turbulence: Four Key Factors Impacting Stocks This Week

This past week has been challenging for the stock market, driven by several significant forces that have created turbule...

CNBC | Jul 25, 2026, 20:05
Market Turbulence: Four Key Factors Impacting Stocks This Week
AI
Shifting Focus: The Cost-Effectiveness of AI Models Takes Center Stage

In recent years, the AI sector has been intensely focused on identifying the most advanced models. While this pursuit re...

Business Insider | Jul 25, 2026, 13:10
Shifting Focus: The Cost-Effectiveness of AI Models Takes Center Stage
Computing
Crisis Averted: Power Line Failure Highlights Urgent Need for Data Center Resilience

A power line failure near Washington, DC, recently showcased a significant challenge faced by the electrical grid due to...

TechCrunch | Jul 25, 2026, 13:50
Crisis Averted: Power Line Failure Highlights Urgent Need for Data Center Resilience
AI
Navigating the AI Landscape: Insights from a Former OpenAI Intern

As the demand for expertise in artificial intelligence surges, many are seeking ways to break into this dynamic field. H...

Business Insider | Jul 26, 2026, 10:10
Navigating the AI Landscape: Insights from a Former OpenAI Intern
AI
Nvidia Strikes Major Deal with SK Hynix to Secure AI Memory Supply

Nvidia has successfully forged a significant partnership with South Korea's SK Hynix to secure memory supplies essential...

CNBC | Jul 25, 2026, 05:15
Nvidia Strikes Major Deal with SK Hynix to Secure AI Memory Supply
View All News