Apple Using Google Gemini Access for On-Device AI Model Distillation

Apple is leveraging Google's Gemini AI models to create smaller, on-device versions through distillation. According to The Information, Google gave Apple "complete access" to Gemini in Google's own data centers, allowing Apple to customize the model for Siri and other AI features.
How the Distillation Process Works
Apple can ask the main Gemini model to perform tasks that provide high-quality results, including a rundown of the reasoning process. Apple then feeds the answers and reasoning information from Gemini to train smaller, cheaper models. This enables the smaller models to learn the internal computations used by Gemini, producing efficient models with Gemini-like performance but requiring less computing power.
Technical Details and Challenges
- Apple can design models built to run on Apple devices without internet connectivity
- Apple can edit Gemini as needed to ensure responses align with Apple's requirements
- Apple has encountered issues because Gemini was tuned for chatbot and coding applications, which doesn't always meet Apple's needs
- The smarter, chatbot version of Siri planned for iOS 27 will rely on Google's Gemini models
Capabilities and Development
Siri will be able to perform many of the same functions as Gemini and other chatbots, including:
- Answering questions
- Summarizing information
- Scanning and understanding uploaded documents
- Telling stories
- Providing emotional support
- Completing tasks like booking travel
The Apple Foundation Models team continues to work on Apple AI models distinct from Gemini models, indicating this is a transitional approach while Apple develops its own AI capabilities.
📖 Read the full source: HN AI Agents
👀 See Also

Claude Service Incident: Elevated Errors Across Platforms
Claude experienced elevated errors across claude.ai, console, and Claude Code platforms on March 2, 2026, with issues affecting login/logout paths and some API methods. The incident was resolved after approximately 4 hours.

Developer Replaces $25/hr Virtual Assistant with AI Agents, Confronts Ethical Implications
A developer replaced a $25/hour virtual assistant with AI agents that handle follow-ups, scheduling, lead tracking, and CRM updates. The AI setup costs about $1,000/month and performs tasks faster and more consistently than the human assistant.

Running OpenClawd for Free: Successes and Challenges
In a recent post on r/clawdbot, a member shares their experience running OpenClawd without API keys, discussing their successes and the challenges faced.
Claude Code v2.1.273 Adds Gateway Hint Headers, MCP Reconnect Alerts, and Remote Control Session Forking
Claude Code v2.1.273 adds opt-in LLM gateway request headers, an MCP disconnect-and-give-up notification, and forking for Remote Control sessions. It also fixes auto-compact firing at roughly half the real context window.