Categories: FAANG

Apple Intelligence Foundation Language Models Tech Report 2025

We introduce two multilingual, multimodal foundation language models that power Apple Intelligence features across Apple devices and services: (i) a ∼3B-parameter on-device model optimized for Apple silicon through architectural innovations such as KV-cache sharing and 2-bit quantization-aware training; and (ii) a scalable server model built on a novel Parallel-Track Mixture-of-Experts (PT-MoE) transformer that combines track parallelism, mixture-of-experts sparse computation, and interleaved global–local attention to deliver high quality with competitive cost on Apple’s Private Cloud Compute…
AI Generated Robotic Content

Recent Posts

qwen 2.1 is very good upscaler

This test used frames taken from H3 generated videos on 0.3MP on my 6GB VRAM.…

18 hours ago

The Roadmap to Mastering LLM Inference Optimization

In this article, you will learn how LLM inference optimization works and which techniques to…

18 hours ago

xAI’s Grok 4.6 is now available in Amazon Bedrock

Today, we are announcing that xAI’s Grok 4.6 is available in Amazon Bedrock, adding a…

18 hours ago

AI, Tariffs, Rare Minerals: What to Expect From Trump’s Upcoming Summit With Xi Jinping

Washington and Beijing have grown ever more linked in the AI boom, making hardware exports…

19 hours ago

Toward physical AI: When the hardware becomes the neural network

Digital computing using silicon chips has transformed nearly every aspect of modern life and enabled…

19 hours ago

Still wishing on a local editing model that can compete with NBP

In exactly two months, it will be 1 year since the release of Nano Banana…

2 days ago