Apple launches Core AI framework for on-device model inference across iOS and macOS.
edge-ai
25 tracked signals on edge-ai.
Google's Gemma 4 brings frontier-level, multimodal, agentic AI to offline edge and mobile devices.
“The North Star of our smaller models is intelligence per byte of memory footprint.”
Apple's new Core AI framework lets developers run advanced AI models entirely on-device in their apps.
“With Core AI, you can build app experiences where user's data never leaves their device.”
NVIDIA's RTX Spark plus Windows software enables autonomous AI agents to run directly on personal PCs.
“My PC became an assistant. While I'm sitting there, of course, this PC would be my great assistant as well.”
Hugging Face launches 200+ WebGPU kernels enabling local AI inference in browsers
NVIDIA Cosmos 3 Edge brings 4B world model robot control to on-device hardware
Google's on-device Gemma 4 models hit 5 million AI Edge Gallery downloads within one month of launch.
“Within one month of the launch, we see more than 5 million downloads”
Google's Gemma models run entirely on-device on phones, enabling multimodal AI, agent skills, and offline use.
“And what's also important to remember is that this is running entirely on the device. So, it will work offline or in areas of low connectivity.”
Microsoft announces Foundry Local 1.2.0 at Build 2026, expanding on-device edge AI development across platforms.
“These updates expand platform support, improve control over inference and acceleration, add new on-device APIs, and simplify deployment across disconnected, regulated, and sovereign environments.”
Microsoft's local AI stack for Windows is now generally available, enabling on-device inference across 1 billion PCs with no cloud, tokens, or network.
“shifting from always running your AI workloads in the cloud to running them locally using just the hardware on everyday PCs and only going to the cloud when your workload truly needs it”
Hugging Face's Reachy Mini robot now runs AI inference fully on-device locally
LFM2.5-VL-3B delivers improved vision-language capabilities optimized for edge deployment
Hugging Face releases LFM2.5-Encoders optimized for fast long-context inference on CPU hardware
Tiny LLMs are essential to scale AI intelligence beyond expensive robots to mass-market devices
“If we want intelligence to get into lots and lots and lots of devices and not just really expensive robots, we are going to need tiny models.”
Google demos Gemma running on-device on Hugging Face's open-source Reachy Mini robot for local AI interaction.
“Faster processing means quicker reactions, which is great for robots navigating the real world.”
NVIDIA introduces the 'AI grid' reference design, positioning telcos to become intelligence providers in the token economy.
“There has truly never been a better time to be in Telecom.”
NVIDIA JetPack 7.2 enables one-command edge deployment of agentic AI with optimized memory and performance.
Frontier models handle big problems but local models offer control and privacy
“you just want some control, you want some privacy, you just want to keep things local and make my stuff my stuff”
Qualcomm VP argues image generation needs architectural innovation, not just scale
Microsoft demos building agentic robots that run generative AI at the edge via Foundry Local and adaptive cloud.
“But the benefits from those advancements have not been evenly distributed.”
Tata Elxsi built IRIS, an AWS edge-based computer vision platform detecting industrial safety risks in seconds.
“Detection of unsafe conditions typically takes 15–45 minutes, depending on operator availability.”
An engineer built a physical, AI-native terminal device with dual displays to interact with an LLM.
“I just wanted to build a device which is physical and AI native like the device which comes from the future.”
OpenBMB promotes its compact MiniCPM models for cheaper, edge-friendly AI deployment at a 'Build Small' hackathon.
“instead of asking how can I use the largest model possible, a better question is can I build a useful project with a compact model that is better, cheaper and easier to deploy”
Qualcomm and Microsoft tout partnership for context-aware AI agents across devices from ring to data center.
“this is the age of AI agents”
Hugging Face announced Cosmos 3 Edge but content body is empty.