Articles Tagged “On Device AI”
19 articles found

Mac mini M6 Speeds Local LLM Prompts by 4.8x for $899
Apple's new Mac mini starts at $899 with the M6 chip, 16GB of unified memory and a claimed 4.8x faster LLM prompt processing in LM Studio than M4.

LFM2.5-DSpark Speeds Local AI Inference Up to 3.2x
Liquid AI released ~300M-parameter draft models for the LFM2.5 family, delivering up to 3.18x GPU throughput and 57% lower function-calling latency.

On-Device Piano Model Autocompletes Music on iPhone
A 125M-parameter transformer generates about 108 piano notes per second on an iPhone 15, trained on 300 million note events with no cloud call.

North Micro Vision Packs Document AI Into 2.4B Params
Cohere Labs released North Micro Vision, a 2.4B Apache 2.0 vision model that reads full-resolution A4 pages and scores 0.921 on DocVQA on local hardware.

On-Device Vision AI Reads Screens in 3GB of Memory
Liquid AI's LFM2.5-VL-3B scores 69.4% average across vision benchmarks and decodes 228 tokens/second on an M5 Max, all inside roughly 3GB of memory.

Snapdragon C Runs 67% Faster on Battery Than Intel N250
Qualcomm's Snapdragon C beat Intel's N250 by up to 67% in unplugged Cinebench multi-core tests, with up to 2.1x better battery power efficiency.
Google Sign Language AI Ships in Gboard on Pixel 11
Google DeepMind's SL2T model brings sign-language-to-text to Gboard and Live Transcribe on Pixel 11, trained on 100,000+ hours across 50 sign languages.

LFM2.5-2.6B Runs Tool-Calling AI Agents in 2.5GB of RAM
Liquid AI's LFM2.5-2.6B runs full tool-calling AI agents on a phone or Raspberry Pi, hitting 220 tokens per second in under 2.5GB of memory.

Liquid AI Encoders Hit 8K Context on CPU 3.7x Faster
Liquid AI's new LFM2.5-Encoders run 8,192-token inputs on a plain CPU roughly 3.7x faster than ModernBERT-base, from just 230M parameters and open weights.

Moonshine Puts Offline Voice AI on a Pico 2 Chip
Moonshine AI fit a full offline voice pipeline — detection, speech-to-text, and neural TTS — onto a Raspberry Pi Pico 2, using just 3.6 MiB of flash.

Raspberry Pi AI Projects Book Covers Local LLMs for £9
Raspberry Pi Press's new AI Projects book covers local LLMs, vision and speech across Pi 5, Pi Zero 2 W and Pico, at an intro price of £8.99.

NVIDIA Cosmos 3 Edge Puts World Models Inside Robots
NVIDIA Cosmos 3 Edge is a 4-billion-parameter world model doing spatial reasoning on Jetson and RTX hardware, adaptable to a robot in about a day.

Edge AI Dev Boards With NPUs: A 2026 Buyer's Guide
Compare five edge AI dev boards from 4 to 67 TOPS and $70 to $249, and see why NPU toolchain maturity, not the TOPS number, decides what runs.

Google's Gemma 4 12B Brings Multimodal AI to a 16GB Laptop
Google DeepMind released Gemma 4 12B on June 3, 2026 — an open multimodal model that reads images and audio and runs on a 16GB laptop, free under Apache 2.0.

ASUS Ascent QN10 — First Snapdragon X2 Elite Mini PC Hits 80 TOPS
ASUS unveiled the Ascent QN10, the first mini PC built on Qualcomm's 18-core Snapdragon X2 Elite — 80 TOPS of on-device AI, up to 32GB LPDDR5, and four-display output.

Gemma 4 12B Brings Full Multimodal AI to a 16GB Laptop — Free Under Apache 2.0
Google DeepMind released Gemma 4 12B on June 3, 2026 — an open-weight, encoder-free multimodal model with native audio that runs locally on a 16GB consumer laptop.

Google Gemma 4 Comes to Android: On-Device AI in 140+ Languages, No Cloud Required
Google's AICore Developer Preview brings Gemma 4 natively to Android devices — offline, privacy-preserving AI inference in over 140 languages that upgrades automatically to Gemini Nano 4.

Apple's M5 Pro and Max Chips Fuse Neural Accelerators Into Every GPU Core — Delivering 4x the AI Compute
Apple's new Fusion Architecture bonds two 3nm dies into a single SoC, embedding dedicated neural accelerators in every GPU core for massive on-device AI gains.

Apple Is Rebuilding Siri From the Ground Up With LLM-Powered Conversational Intelligence in iOS 26.4
Apple's Siri overhaul replaces the command-based architecture with on-device large language models, bringing natural conversation and app-aware context to one billion iPhones.
